DeepSeek: DeepSeek V4.1 Flash

Text input Image input Text output
Author's Description

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the cost-efficient tier of the V4.1 family. DeepSeek reports that it exceeds V4 Pro on performance, speed, and task...

Key Specifications
Context
1M
Released
Sep 09, 2026
Supported Parameters

This model supports the following parameters:

Temperature Include Reasoning Reasoning Presence Penalty Top Logprobs Stop Tool Choice Logprobs Response Format Top P Frequency Penalty Tools Max Tokens
Features

This model supports the following features:

Reasoning Tools Response Format
Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.3
Completion $1.2
Input Cache Read $0.006

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
DeepSeek
DeepSeek | deepseek/deepseek-v4.1-flash-20260910 1M $0.3 / 1M tokens $1.2 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by deepseek