This model is no longer available on Elosia. This page is kept for informational purposes.
ThinkingTool UseStructured Output
About this model
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Performance Tier
Balanced
DeepSeek v4 Flash 0423 is a balanced model from DeepSeek : strong performance at a reasonable price.
Strong cost-performance ratio. Reliable for most professional use cases without premium pricing.
Pricing
This model is included in Elosia plans
Eco
Minimal cost. Ideal for very high volume or simple tasks.
Reasoning and code generation are the declared focus, with a reported suite weighted toward science questions, competitive programming and repository patches
MoE 284B total / 13B active, a ratio that keeps throughput high for agents and coding assistants
1M-token context served by hybrid sparse attention
Open-weight MIT, deployable on-prem
Limitations
Limited factual knowledge density, which shows on recall-heavy tasks
Max reasoning mode adds significant latency, so it does not suit real-time UX
Creative writing is the weak axis of a model tuned for reasoning and code