AI Dictionary › Prompting
Compressione del Prompt
Prompt compression means cutting the token count of a prompt, or its surrounding context, while preserving essential information, to save cost, reduce latency and fit the context window. It is done by hand (removing redundancy, pleasantries, superfluous examples) or with automatic techniques that filter low-information words. Example: a 3,000-token context of meeting minutes is condensed into 400 tokens that keep decisions, owners and deadlines, dropping chit-chat and repetition, before being passed to the model to draft the follow-up.
Use it when working with long contexts, high volumes or tight token budgets: RAG, document processing, large-scale apps. The risk is cutting useful information: compress deliberately and check that output quality does not drop.
From our network
AGORÀ Intelligence: Enterprise AI Governance Platform
Govern AI at scale: policies, adoption and measurable results on your data. Built for boards and C-suite.
Visit agora-intelligence.com →From the Agora Intelligence blog
📱 Download the Android app (beta) iOS coming soon
Say what you mean. Get what you need.
Grace Certified, the AI coach that trains and certifies your prompt engineering, by Agora Intelligence.