topic · #token-management
Everything about token-management
2 connected nodes across dictionary, workflows, comparisons, prompts, tool stacks and use cases.
/ Dictionary · 2
DEFDictionaryNODE·60620B
Contextual Compression
Contextual compression is a technique used to reduce the size of the input context for a Large Language Model (LLM) while retaining its most relevant information, typically by summarizing or filtering.
#context-engineering#llm#summarization
/contextual-compressionopen →
DEFDictionaryNODE·3A3CEE
Token Budget
A token budget is a predefined limit on the number of tokens an AI application or specific request can consume within a given period or for a single interaction. It is a critical mechanism for controlling costs and managing resource allocation for Large Language Model (LLM) usage.
#ai-cost-control#token-management#llm-governance
/token-budgetopen →