Zum Hauptinhalt springen
tsecurity.de LIVE
Echtzeit-Radar & Feeds
Alle RSS Feeds
👥 Community & Social
Sichere ProgrammierungI audited my own ML linter and had to withdraw its best evidence(21.09.2026 um 22:54 Uhr)
Sichere ProgrammierungQuantum Result Validation for Distributed Computing Systems(21.09.2026 um 22:54 Uhr)
Sichere ProgrammierungJWT Authentication and Role-Based Access Control in LocalHands(21.09.2026 um 22:56 Uhr)
Sichere ProgrammierungStochastic Parrot or Alien Mind?(21.09.2026 um 22:56 Uhr)
Sichere ProgrammierungBuilding AI for the Physical World Is a Different Engineering Problem(21.09.2026 um 22:58 Uhr)
Sichere ProgrammierungI audited my own ML linter and had to withdraw its best evidence(21.09.2026 um 22:54 Uhr)
Sichere ProgrammierungQuantum Result Validation for Distributed Computing Systems(21.09.2026 um 22:54 Uhr)
Sichere ProgrammierungJWT Authentication and Role-Based Access Control in LocalHands(21.09.2026 um 22:56 Uhr)
Sichere ProgrammierungStochastic Parrot or Alien Mind?(21.09.2026 um 22:56 Uhr)
Sichere ProgrammierungBuilding AI for the Physical World Is a Different Engineering Problem(21.09.2026 um 22:58 Uhr)
Intelligence View
⚡ tsecurity.de Intelligence

A GitHub tinkerer teaches Claude to talk less, and that may matter more than it seems

In a quiet corner of GitHub better known for weekend experiments than paradigm shifts, Drona Reddy, a data analyst at Amazon US, has published a single markdown file that promises to cut Claude’s output token usage by more than half, not b…

0
↗ Quelle (infoworld.com)
Reagiere als Erste:r — dein Feedback zählt!








In a quiet corner of GitHub better known for weekend experiments than paradigm shifts, Drona Reddy, a data analyst at Amazon US, has published a single markdown file that promises to cut Claude’s output token usage by more than half, not by changing code, but by reshaping the model’s behavior.





The file, called Claude.md and available under an MIT license, outlines a set of structured instructions that claim to reduce Claude’s output verbosity by about 63% without any code modifications.





These instructions impose strict behavioral constraints on the model, including limits on output length, emphasis on token efficiency and accuracy, controls on speculation, rules for typography, and a zero‑tolerance policy on sycophantic responses. They also simplify code generation and define clear override policies, effectively training the model to respond more concisely and deliberately.





Reducing output tokens





The rationale is straightforward: eliminate what Reddy describes as Claude’s “frivolous” habits, stripping out everything that isn’t strictly necessary. That means no automatic pleasantries like “Sure!” or “Great question!”, no boilerplate sign-offs such as “I hope this helps,” no restating the prompt, and no unsolicited suggestions or over-engineered abstractions.





It also curbs stylistic quirks like “em” dashes, smart quotes, and other Unicode characters that can break parsers, while preventing the model from reflexively agreeing with flawed assumptions.





At scale, that kind of austerity, according to Reddy, could translate into meaningful savings, turning small stylistic trims into outsized efficiency gains.





The data analyst also outlined three distinct use cases where the markdown file could be most effective. First, high-volume automation pipelines, such as resume bots, agent loops, and code generation, where verbosity compounds across repeated calls.
Second, repeated structured tasks, where Claude’s default expansiveness can add up over hundreds of interactions. Third, team environments that require consistent, parseable output formats across sessions, where tighter control over responses improves reliability and downstream usability.





In his own simulations on Claude Sonnet, Reddy said the file could save close to 9,600 tokens a day at 100 prompts, translating to roughly $0.86 in monthly savings. At 1,000 prompts a day, the savings rise to about 96,000 tokens, or $8.64 a month, while across three projects combined, he estimates reductions of nearly 288,000 tokens, equivalent to around $25.92 monthly.





However, the data analyst also warned that the file might be really ineffective, even counterproductive, in certain use cases, such as single one-off queries, fixing deep failures, or exploratory work where feedback is required, as the file itself consumes input tokens on every message.





“The CLAUDE.md file itself consumes input tokens on every message. The savings come from reduced output tokens. The net is only positive when output volume is high enough to offset the persistent input cost. At low usage it costs more than it saves,” Reddy wrote in the repository’s documentation.





Modest enterprise gains





Analysts do see enterprises and their CIOs benefitting from the markdown file, at least to a certain degree, especially as they struggle to balance spiraling inference bills and moving agentic or other AI pilots into production.





“A 63% token reduction can meaningfully lower inference costs and latency for enterprises running high-volume Claude workloads,” said Charlie Dai, principal analyst at Forrester.





The gains, however, may be more operational than transformative.





“For CIOs, this method offers some operational benefits as it improves output consistency, improves latency, and enforces basic token discipline, which can help in scaling automation,” said Pareekh Jain, principal analyst at Pareekh Consulting.





However, Jain pointed out that though this is a “useful tactical optimization”, it does not fundamentally change enterprise AI economics.





“In enterprise settings, the tactic is likely to translate into more modest savings because output tokens are only a portion of total usage as input context, retrieval, and agent orchestration typically dominate costs,” Jain said. “As a result, most enterprises would likely see single-digit savings rather than the headline number,” he added. The markdown file is designed to be model-agnostic and should work across large language models that can follow structured instructions, though Reddy noted he has not tested its effectiveness on local models such as those running on llama.cpp or Mistral.


Ähnliche Beiträge
🔍 Verwandte News

Auch interessante Nachrichten A GitHub tinkerer teaches Claude to talk less, and that may matter more than it seems

Thematisch verwandte Begriffe: GitHub, tinkerer, teaches, Claude · 6 Treffer

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Zum Aktualisieren ziehen
ZERO-DAY CVE-2026-79918 | MaxKB is an open-source AI assistant for enterprise. Prior to version 2.…
Advisory →
TTS Reader • tsecurity.de Voice
tsecurity.de Icon
tsecurity.de App
Offline-Lesen, Eilmeldungen & 0ms Ladezeit

Installiere tsecurity.de direkt auf deinen Home-Bildschirm für das ultimative Vollbild-Magazinerlebnis ohne Browser-Leisten.

Nächster Beitrag
Themen-Radar & Intelligence Matrix
Echtzeit-Taxonomie nach Angriffsvektoren & Plattformen

tsecurity.de Live Threat Radar

🔴 LIVE RADAR
MONITORING
AKTIV
CVE-DATENBANK
LIVE
🔍
Community Radar & Live Chat
Sentinel Bot online • Live-Stream
Dein Cluster: Security Explorer
Match:
lädt…
Verbindung zum Community-Stream wird aufgebaut...
Bearbeitungsmodus — Senden überschreibt deine Nachricht
Community-Puls — was gerade passiert
lädt…
Aktivitäten deiner Analysten
lädt…
Neues Thema oder Eilmeldung einreichen

Reiche interessante Links, Zero-Days oder Debatten ein. Die Community entscheidet per Upvote über die Veröffentlichung.

Heiß diskutierte Einreichungen
🔖 Gespeicherte Artikel
📂 Keine gespeicherten Artikel vorhanden.
Zurück Ziehen Vor
Links: vorheriger Artikel Rechts: nächster Artikel unten: schließen
News NIS-2 Frühwarnung Tier-1 Intel ⏱️ 3 Min vor 10 Min
Artikeldaten werden geladen...

Zurück: vorheriger Vor: nächster
↗ Original-Quelle
Social Reaktionen Deine Reaktion zählt
Einstufung & Relevanz-Poll 0 Stimmen
In sozialen Netzwerken teilen 1-Klick