Discussion about this post

User's avatar
Marius Laurusevicius's avatar

The "smart decides, fast implements" split now has a public number next to it. The Information reported on 20 August that AT&T cut costs on coding and some other advanced AI tasks by as much as 56% using LiteLLM routers, with quality down 2% — though that is one company's internal measurement and the benchmark set was not published. Do you draw the smart/fast line by hand per task, or has automated routing held up for you?

No posts

Ready for more?