DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash

deepseek-ai/DeepSeek-V4.1-Flash

約DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash is DeepSeek's latest multimodal MoE model with 552B backbone parameters. Built on a Causal Encoder-Decoder architecture, it activates only ~8B parameters during prefill and ~16B during decode, delivering flagship-level intelligence at significantly lower inference cost. The model natively understands images, supports a 1M-token context window. It outperforms DeepSeek-V4-Pro across coding, agentic and automation benchmarks, while reducing the KV cache footprint by roughly 4x versus its predecessor.

利用可能な Serverless

クエリをすぐに実行し、使用量のみを支払います

100万トークン(Input/Output)ごとに

$

0.15

/ M Tokens

100万トークン(Input/Output)ごとに

$

0.003

/ M Tokens

100万トークン(Input/Output)ごとに

$

0.6

/ M Tokens

メタデータ

作成する

ライセンス

MIT

プロバイダー

DeepSeek

ハギングフェイス

仕様

Available

建築

Causal Encoder-Decoder MoE

キャリブレートされた

いいえ

専門家の混合

はい

合計パラメータ

552B

アクティブ化されたパラメータ

8B / 16B

推論

いいえ

Precision

FP8

コンテキスト長

1049K

Max Tokens

393K

対応機能

Serverless

対応

Serverless LoRA

サポートされていません

Fine-tuning

サポートされていません

Embeddings

サポートされていません

Rerankers

サポートされていません

Image入力をサポートする

対応

JSON Mode

対応

構造化されたOutputs

サポートされていません

ツール

対応

Fim Completion

サポートされていません

Chat Prefix Completion

対応

AI開発を 加速する準備はできていますか?

AI開発を 加速する準備はできていますか?

AI開発を 加速する準備はできていますか?