Back to Glossary
AI Tech ยท Intermediate
Context Caching
Efficiency
A feature in modern LLM APIs that allows you to pay once to upload a large codebase, and then query it cheaply multiple times without re-uploading tokens.
A feature in modern LLM APIs that allows you to pay once to upload a large codebase, and then query it cheaply multiple times without re-uploading tokens.