Back to Glossary

AI Tech ยท Intermediate

Context Caching

Efficiency

A feature in modern LLM APIs that allows you to pay once to upload a large codebase, and then query it cheaply multiple times without re-uploading tokens.

A feature in modern LLM APIs that allows you to pay once to upload a large codebase, and then query it cheaply multiple times without re-uploading tokens.