DeepSeek V4.1 Flash
Chatdeepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash is DeepSeek's latest efficiency-optimized Mixture-of-Experts model with a 1M-token context window and up to 384K output tokens. It adds native multimodal vision understanding, supports thinking mode (on by default), tool calls, JSON output and prompt caching, and per DeepSeek surpasses V4 Pro on quality, cost and speed. Served upstream under the official model id deepseek-flash.
コンテキストウィンドウ
1M
最大出力トークン
384K
リリース日
2026-09-10
機能
VisionFunction Calling推論プロンプトキャッシュ
利用可能なプロバイダー
DeepSeek
対応プロトコル
openaianthropic
Providers
DeepSeek
入力トークン
$0.3/M
出力トークン
$1.2/M
キャッシュ読込
$0.006/M
Protocols
openai
/v1/chat/completions/v1/responsesanthropic
コード例
from openai import OpenAIclient = OpenAI(base_url="https://api.ofox.ai/v1",api_key="YOUR_OFOX_API_KEY",)response = client.chat.completions.create(model="deepseek/deepseek-v4.1-flash",messages=[{"role": "user", "content": "Hello!"}],)print(response.choices[0].message.content)
関連モデル
よくある質問
Ofox.aiでのDeepSeek V4.1 Flashの料金は、入力100万トークンあたり$0.3/M、出力100万トークンあたり$1.2/Mです。従量課金制、月額料金なし。
DeepSeek V4.1 Flashは1Mトークンのコンテキストウィンドウと最大384Kトークンの出力に対応。大規模ドキュメントの処理や長い会話の維持が可能です。
ベースURLをhttps://api.ofox.ai/v1に設定し、Ofox APIキーを使用するだけ。OpenAI互換APIなので、既存コードのベースURLとAPIキーを変更するだけです。
DeepSeek V4.1 Flashは以下の機能に対応:Vision, Function Calling, 推論, プロンプトキャッシュ。Ofox.ai統合APIですべての機能にアクセスできます。