Fireworks said DeepSeek V4 Flash now includes native vision on serverless and priority tiers. It reports better benchmarks than the text-only 0731 checkpoint at the same $0.22 input / $0.66 output per 1M tokens. The model page lists a 305B MoE with about 1040K context, function calling, image input, and a speculative-decoding module.
Key Takeaways
- βDeepSeek-V4-Flash-Vision-Exp is live on Fireworks serverless and priority tiers with native image input.
- βSame price as text-only 0731: $0.22 input / $0.66 output per 1M tokens (cached input $0.007).
- β305B MoE, ~1040K context, function calling, and speculative decoding for screenshot-aware coding agents.
Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.