DeepSeek V4 Flash now runs updated weights on AI Gateway
DeepSeek V4 Flash now runs updated weights by default on AI Gateway. The change is automatic for traffic to deepseek/deepseek-v4-flash: the same model ID and…
By Dillip Chowdary • Aug 05, 2026 • Source: Vercel Blog
DeepSeek V4 Flash now runs updated weights by default on AI Gateway. The change is automatic for traffic to deepseek/deepseek-v4-flash: the same model ID and existing request code pick up the new weights with no migration step. The update is framed around stronger agentic capabilities rather than a rename or new product surface.
On Terminal-Bench, the updated model scores 82.7. That is 25.8 points above the April preview result of 56.9. The benchmark jump is the concrete performance signal tied to the weight update; AI Gateway surfaces it by routing default deepseek/deepseek-v4-flash calls to those weights without a separate endpoint or version string change.
Advertisement
Tech Pulse Daily
Get tomorrow's pulse first
Join engineers who read Tech Pulse before stand-up. Free, weekday mornings.
For engineers already calling DeepSeek V4 Flash through AI Gateway, the practical effect is an in-place capability lift on agent-style workloads. Terminal-Bench-oriented loops, tool use, and multi-step task completion can improve without client changes, redeploys, or ID swaps. That reduces the usual cost of chasing a better checkpoint when the provider controls default weights behind a stable model string.
On the supply side, DeepSeek is currently the only provider serving the updated weights. Other providers have not yet matched that serving path, so AI Gateway traffic that lands on DeepSeek is the channel that receives the Terminal-Bench gains today. That is a temporary exclusivity window: builders comparing multi-provider routes for the same model ID may see uneven quality until others ship the same weights.
What to watch next is when additional providers start serving the updated DeepSeek V4 Flash weights and whether Terminal-Bench and agentic quality stay aligned across them. Until then, treat deepseek/deepseek-v4-flash on AI Gateway as the path that already defaults to the stronger weights, and re-check provider routing if you pin or load-balance across non-DeepSeek backends that may still be on the older checkpoint.
Advertisement
🔎 More interesting news
- Degraded performance for Claude Mythos 5, Claude Fable 5, and Claude Opus 5
- SkiaSharp 4.0 Establishes Milestone-Aligned Release Cadence
- Show HN: Memcode launches a new terminal coding agent
- Ponytail Agent Skill Corrects Its Own Benchmark After Contributor Challenge
- Today's full Tech Pulse briefing →