AI21 argues that LLM token spend keeps climbing and that naive model routing alone cannot control it. The post lays out why managing cost requires deeper strategies beyond simply routing requests to cheaper models.
Token spend isn't going down. You need more than naive routing to manage it
calendar_today
June 25, 2026
domain
ai21-labs