DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Tokens processed for this model through Reroute, per day, over the last 30 days.
Daily token volume shows up here as soon as the first request for this model completes.