DeepSeek V4.1 Flash

Chat
deepseek/deepseek-v4.1-flash

DeepSeek V4.1 Flash is DeepSeek's latest efficiency-optimized Mixture-of-Experts model with a 1M-token context window and up to 384K output tokens. It adds native multimodal vision understanding, supports thinking mode (on by default), tool calls, JSON output and prompt caching, and per DeepSeek surpasses V4 Pro on quality, cost and speed. Served upstream under the official model id deepseek-flash.

コンテキストウィンドウ
1M
最大出力トークン
384K
リリース日
2026-09-10
機能
VisionFunction Calling推論プロンプトキャッシュ
利用可能なプロバイダー
DeepSeek
対応プロトコル
openaianthropic

Providers

DeepSeek
入力トークン
$0.3/M
出力トークン
$1.2/M
キャッシュ読込
$0.006/M
Protocols
openai/v1/chat/completions/v1/responses
anthropic

コード例

from openai import OpenAI
client = OpenAI(
base_url="https://api.ofox.ai/v1",
api_key="YOUR_OFOX_API_KEY",
)
response = client.chat.completions.create(
model="deepseek/deepseek-v4.1-flash",
messages=[
{"role": "user", "content": "Hello!"}
],
)
print(response.choices[0].message.content)

よくある質問

Ofox.aiでのDeepSeek V4.1 Flashの料金は、入力100万トークンあたり$0.3/M、出力100万トークンあたり$1.2/Mです。従量課金制、月額料金なし。