newsReddit r/MachineLearningTrust 52 · CommunityPublished 8d agoLive · 6d ago
Does telling an LLM to "be concise" actually save you money? We measured it across 9 models. Compressing the output can save you money and keep accuracy, compressing the input prompt does not. [R]
LLMs are too verbose and with a black box model the only things you control are what goes in and how you tell it to write back. Yesterday Claude Code shipped a "concise output style" where Claude keeps things short. We already have a paper out about this! We tested both channels, shortening the input prompt versus telling the model to output answer shorter, on the same questions across five reduction levels, and scored cost, accuracy, and whethe
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 53%Evaluate a model properly →
- PossiblePossibly related (embedding) · 51%When Summaries Distort Decisions: Information Fidelity in LLM-Compressed Financial Analysis →
- PossiblePossibly related (embedding) · 48%Two Axes of LLM Abstention: Answer Correctness and Question Answerability →
- PossiblePossibly related (embedding) · 48%Prompt Compression via Activation Aggregation →
- PossiblePossibly related (embedding) · 48%LLM-as-a-Verifier: A General-Purpose Verification Framework →
- PossiblePossibly related (embedding) · 47%jia-gao/leanctx →
Covers
tutorialEvaluate a model properlypaperWhen Summaries Distort Decisions: Information Fidelity in LLM-Compressed Financial AnalysispaperTwo Axes of LLM Abstention: Answer Correctness and Question AnswerabilitypaperPrompt Compression via Activation AggregationpaperLLM-as-a-Verifier: A General-Purpose Verification Framework
Covers (incoming)
Related across the graph
paperWhen Summaries Distort Decisions: Information Fidelity in LLM-Compressed Financial AnalysispaperLLM-as-a-Verifier: A General-Purpose Verification FrameworkpaperTwo Axes of LLM Abstention: Answer Correctness and Question AnswerabilitypaperPrompt Compression via Activation Aggregationrepojia-gao/leanctxtutorialEvaluate a model properly
