DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash

deepseek-ai/DeepSeek-V4.1-Flash

Tentang DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash is DeepSeek's latest multimodal MoE model with 552B backbone parameters. Built on a Causal Encoder-Decoder architecture, it activates only ~8B parameters during prefill and ~16B during decode, delivering flagship-level intelligence at significantly lower inference cost. The model natively understands images, supports a 1M-token context window. It outperforms DeepSeek-V4-Pro across coding, agentic and automation benchmarks, while reducing the KV cache footprint by roughly 4x versus its predecessor.

Tersedia Serverless

Jalankan kueri segera, bayar hanya untuk penggunaan

Per 1M Token (Input/Output)

$

0.15

/ M Tokens

Per 1M Token (Input/Output)

$

0.003

/ M Tokens

Per 1M Token (Input/Output)

$

0.6

/ M Tokens

Metadata

Buat di

Lisensi

MIT

Penyedia

DeepSeek

Spesifikasi

Negara

Available

Arsitektur

Causal Encoder-Decoder MoE

Terkalibrasi

Tidak

Campuran Ahli

Ya

Total Parameter

552B

Parameter yang Diaktifkan

8B / 16B

Penalaran

Tidak

Precision

FP8

Text panjang konteks

1049K

Max Tokens

393K

Didukung Keberfungsian

Serverless

didukung

Serverless LoRA

Tidak didukung

Fine-tuning

Tidak didukung

Embeddings

Tidak didukung

Rerankers

Tidak didukung

Dukung Image Input

didukung

JSON Mode

didukung

Output Terstruktur

Tidak didukung

Alat

didukung

Fim Completion

Tidak didukung

Chat Prefix Completion

didukung

Siap untuk mempercepat pengembangan AI Anda?

Siap untuk mempercepat pengembangan AI Anda?

Siap untuk mempercepat pengembangan AI Anda?