Tech News
← Home  ·  All topics

Telegraph Test

1 GoKawiil brief on this topic

Benchmark finds telegram-style prompts cut LLM output tokens up to 49%

A new open-source benchmark called the Telegraph Test shows that instructing large language models to answer in 'cablese' — the clipped, article-free style once used by telegraph operators — reduces billed output tokens by 40-49% on several models' own API meters, while preserving or even slightly improving factual recall. The effect held across four model families tested with roughly 1,300 questions over 50 passages, with one exception: gpt-5-mini's reasoning overhead made cablese prompting roughly double its cost instead of cutting it.