OpenRouter reported DeepSeek V4.1 Flash processed ~1T tokens in 24 hours and is on pace for ~2.8T in 48 hours—potentially its biggest paid-model launch window. About 90% of those tokens were cache reads, priced by the market at ~$0.006/M, roughly 5x cheaper than comparable models like GLM-5.3 Flash.
Key Takeaways
- ✓~1T tokens in 24h on OpenRouter; ~2.8T pace at 48h.
- ✓~90% cache reads at ~$0.006/M market price.
- ✓About 5x cheaper cache pricing vs peers like GLM-5.3 Flash.
Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.