prompt cachingOne PR Billed 156M Tokens. Cap the Reads, Not the Rate.
A published trace of one 800-line pull request shows a coding agent billed roughly 156 million tokens to produce 289,000 — 98% of the volume was cache re-reads of context re-sent on every one of 512 turns. Coding-agent cost is set by turns per task and the width of each read, not by the model's rate card.
September 3, 2026 · 12 min readAI observability pricingAI Observability Pricing: Same 10 GB, $49 or $930
Eleven AI observability options priced against one workload — 500,000 agent runs a month, 5 million spans, 10 GB of traces. The same telemetry costs $49 at SigNoz Cloud and about $930 at W&B Weave, and retention is a bigger multiplier than traffic.
August 19, 2026 · 17 min readAI observabilityDynatrace Bought Arize. Fix Your Spans Before It Closes.
Dynatrace is paying $915 million for Arize, and the change that matters is the meter: Arize bills trace spans with unlimited users, Dynatrace bills $0.20 per GiB ingested and $0.0035 per GiB scanned. Normalize your OpenInference attributes to OpenTelemetry gen_ai.* in your own collector before the deal closes.
August 16, 2026 · 12 min readAI agent observabilityBest AI Agent Monitoring: Langfuse, Then a Real Kill Switch
Seven agent monitoring tools priced against the same workload: 50,000 runs a month at 12 steps each. Self-hosted Langfuse wins on cost and audit depth — but only one of these tools sits in the request path where it can actually stop a runaway agent, and that is the part you have to build yourself.
August 11, 2026 · 16 min read