DeepSeek V4.1 Flash
Chatdeepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash is DeepSeek's latest efficiency-optimized Mixture-of-Experts model with a 1M-token context window and up to 384K output tokens. It adds native multimodal vision understanding, supports thinking mode (on by default), tool calls, JSON output and prompt caching, and per DeepSeek surpasses V4 Pro on quality, cost and speed. Served upstream under the official model id deepseek-flash.
Kontextfenster
1M
Max. Ausgabe-Tokens
384K
Veröffentlicht
2026-09-10
Fähigkeiten
VisionFunction CallingReasoningPrompt Caching
Verfügbare Anbieter
DeepSeek
Unterstützte Protokolle
openaianthropic
Providers
DeepSeek
Eingabe-Tokens
$0.3/M
Ausgabe-Tokens
$1.2/M
Cache-Lesen
$0.006/M
Protocols
openai
/v1/chat/completions/v1/responsesanthropic
Code-Beispiele
from openai import OpenAIclient = OpenAI(base_url="https://api.ofox.io/v1",api_key="YOUR_OFOX_API_KEY",)response = client.chat.completions.create(model="deepseek/deepseek-v4.1-flash",messages=[{"role": "user", "content": "Hello!"}],)print(response.choices[0].message.content)
Verwandte Modelle
Häufig gestellte Fragen
DeepSeek V4.1 Flash auf Ofox.ai kostet $0.3/M pro Million Eingabe-Tokens und $1.2/M pro Million Ausgabe-Tokens. Pay-as-you-go, keine monatlichen Gebühren.
DeepSeek V4.1 Flash unterstützt ein Kontextfenster von 1M Tokens mit max. Ausgabe von 384K Tokens, was die Verarbeitung großer Dokumente und lange Konversationen ermöglicht.
Einfach Ihre Base-URL auf https://api.ofox.io/v1 setzen und Ihren Ofox API Key verwenden. Die API ist OpenAI-kompatibel — einfach Base-URL und API Key in Ihrem bestehenden Code ändern.
DeepSeek V4.1 Flash unterstützt folgende Fähigkeiten: Vision, Function Calling, Reasoning, Prompt Caching. Zugriff auf alle Features über die einheitliche Ofox.ai API.