Writing

2.61Btokens through the router in 14 days$23.85total model spend at list price29.9%of requests failed

Fourteen days of router logs from a research workload running on nine free or near-free LLM providers: 83,497 requests, 2.6 billion tokens, a 29.9% failure rate, and the eleven changes that made it usable.

ai-infrastructurellm-routingfree-modelsresearch-automation
  1. 01What the router runs
  2. 02Nine providers and what each one did
  3. 03Where the failures came from
  4. 04What changed and when
  5. + 4 more sections