Fireworks AI only route
accounts/fireworks/models/deepseek-v4p1-flash
Uptime
100.0%
- Input / 1M
- $0.22
- Output / 1M
- $0.66
- TTFT
- 1.13s fixed prompt · n=5
- Output rate
- 120.8 tok/s fixed prompt · n=5
deepseek-v4p1-flash
Multimodal open-weight MoE model for long-context reasoning and agentic work
One route serves DeepSeek V4.1 Flash, and every figure below is measured on it. Where the selected window has too little traffic, the figure names what it is based on instead.
accounts/fireworks/models/deepseek-v4p1-flash
TTFT is the median time to the first visible answer token over real ModelRelay requests, with a fixed prompt standing in for models that have no traffic yet. Speed ranks by it. Intelligence is the Artificial Analysis Intelligence Index.