Odhadované mesačné plytvanie
13 418 USD
Podiel sledovaných nákladov
31,5 %
Zistenia
7
Premium model used for binary classification
Drahý model na jednoduchú úlohu
68% of SupportClassifier tasks are single-label routing decisions under 400 tokens, but run on a premium reasoning model.
Retry storms on malformed JSON output
Nadmerné opakovania
LeadEnricher retries up to 4x when schema validation fails. Retries account for 31% of its spend.
Identical system prompts re-sent without prompt caching
Chýbajúca cache
A 6.1k-token system preamble is re-billed on every call across three agents. Prompt caching is not enabled.
Full document dumps instead of retrieval
Príliš veľký kontext
IncidentSummarizer sends whole log files (avg 11.2k tokens) where top-k retrieval of 1.5k tokens matches quality.
Spend on tasks that never produced an outcome
Zlyhané úlohy
Failed tasks still consume premium model calls before aborting. No early-exit guard is configured.
Same request fingerprint within 60 seconds
Duplicitné požiadavky
Client-side double submits produce identical completions billed twice.
Agent loops between search and reasoning steps
Slučky volaní nástrojov
PipelineTriager exceeds 8 tool calls in 12% of runs without converging; no loop breaker configured.