DeepSeek V4.1 Flash

Chat
deepseek/deepseek-v4.1-flash

DeepSeek V4.1 Flash is DeepSeek's latest efficiency-optimized Mixture-of-Experts model with a 1M-token context window and up to 384K output tokens. It adds native multimodal vision understanding, supports thinking mode (on by default), tool calls, JSON output and prompt caching, and per DeepSeek surpasses V4 Pro on quality, cost and speed. Served upstream under the official model id deepseek-flash.

Kontextfenster
1M
Max. Ausgabe-Tokens
384K
Veröffentlicht
2026-09-10
Fähigkeiten
VisionFunction CallingReasoningPrompt Caching
Verfügbare Anbieter
DeepSeek
Unterstützte Protokolle
openaianthropic

Providers

DeepSeek
Eingabe-Tokens
$0.3/M
Ausgabe-Tokens
$1.2/M
Cache-Lesen
$0.006/M
Protocols
openai/v1/chat/completions/v1/responses
anthropic

Code-Beispiele

from openai import OpenAI
client = OpenAI(
base_url="https://api.ofox.io/v1",
api_key="YOUR_OFOX_API_KEY",
)
response = client.chat.completions.create(
model="deepseek/deepseek-v4.1-flash",
messages=[
{"role": "user", "content": "Hello!"}
],
)
print(response.choices[0].message.content)

Häufig gestellte Fragen

DeepSeek V4.1 Flash auf Ofox.ai kostet $0.3/M pro Million Eingabe-Tokens und $1.2/M pro Million Ausgabe-Tokens. Pay-as-you-go, keine monatlichen Gebühren.