The Model Menu Is Getting Out of Hand
Claude, DeepSeek, Kimi, Grok, Nemotron—Vercel's AI Gateway added four new models in a single week. The real story isn't model proliferation. It's what developers are supposed to do with a menu this long.
Vercel's AI Gateway added Kimi K2.7 Code, Nemotron 3 Ultra, Grok Imagine Video 1.5, and DeepSeek via Azure inside the same news cycle. Claude Fable 5 was suspended, then reinstated. The message is clear: the model layer is moving too fast for anyone to track, including the platforms distributing them.
This creates a real problem for product teams. Choosing a model used to be a quarterly decision. Now it's a weekly one, and the switching costs—prompt tuning, evals, cost recalibration—are non-trivial. Anthropic still dominates spend according to Vercel's own data, but DeepSeek is entering the fight for token volume. That gap between spend and volume is worth watching. It suggests teams are running DeepSeek for high-frequency, lower-stakes tasks while keeping Anthropic for work that matters more.
The practical move for teams right now: stop optimizing for the best model and start building model-agnostic evals. If your product's quality depends on a specific model staying available—Claude Fable 5 was suspended and back within days—you're one vendor decision away from a bad week.
The AI SDK update that now supports Claude Code, Codex, and Pi as agent harnesses is relevant here too. Abstraction layers exist precisely because the underlying models are unstable. Use them.
Sources
- Claude Fable 5 access suspended on AI Gateway Vercel blog
- Kimi K2.7 Code now available on AI Gateway Vercel blog
- Program Claude Code, Codex, Pi and other agent harnesses with AI SDK Vercel blog
- Claude Fable 5 now available on AI Gateway Vercel blog
- DeepSeek enters the fight for token volume, Anthropic continues to dominate spend Vercel blog
- Nemotron 3 Ultra now available on AI Gateway Vercel blog
- Grok Imagine Video 1.5 on AI Gateway Vercel blog