LiteLLM Budget Limits: max_budget Needs Postgres
A LiteLLM budget cap does nothing without a database. I tested v1.104.0, then set up per-agent keys and a team spend limit in ten minutes.
Read the full story →Latest
Open-weight decision models: Clef vs Strands Decider
Clef and Strands Decider 2B are open-weight decision models that return a calibrated score, not text. Here's what shipped and why it matters.
The Hidden Risk of Letting AI Manage Ad Campaigns
MCP's own specs disagree on whether it is stateful. Neither defines a record of what changed in your ad account — and…
Your Agent A/B Test Was Broken Before You Ran It
In parimutuel betting your own bet moves the odds you get, so you can't test a strategy against half your bankroll. Agents…
Half My Agent’s Budget Goes to Checking Its Own Work
I built the quality gate last, in an afternoon, as the cheap safety rail. Then I added up the API bills by…
I Rewrote the Test Until My Writing Passed It
Five commits in four days, every one adjusting the grader that judges this blog's drafts, every one triggered by a verdict I…
The Bug Was Reproducible. My Logs Weren’t.
A post on this site ended in a raw API payload for 151 days. Fixing it took two minutes. Explaining it took…
