DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash

deepseek-ai/DeepSeek-V4.1-Flash

정보에 대해서DeepSeek-V4.1-Flash

DeepSeek-V4.1-Flash is DeepSeek's latest multimodal MoE model with 552B backbone parameters. Built on a Causal Encoder-Decoder architecture, it activates only ~8B parameters during prefill and ~16B during decode, delivering flagship-level intelligence at significantly lower inference cost. The model natively understands images, supports a 1M-token context window. It outperforms DeepSeek-V4-Pro across coding, agentic and automation benchmarks, while reducing the KV cache footprint by roughly 4x versus its predecessor.

사용 가능한 Serverless

쿼리를 즉시 실행하고 사용한 만큼만 지불하세요.

1M 토큰당 (Input/Output)

$

0.15

/ M Tokens

1M 토큰당 (Input/Output)

$

0.003

/ M Tokens

1M 토큰당 (Input/Output)

$

0.6

/ M Tokens

메타데이터

생성하다

라이센스

MIT

공급자

DeepSeek

허깅페이스

사양

Available

건축

Causal Encoder-Decoder MoE

교정된

아니요

전문가의 혼합

총 매개변수

552B

활성화된 매개변수

8B / 16B

추론

아니요

Precision

FP8

콘텍스트 길이

1049K

Max Tokens

393K

지원됨 기능

Serverless

지원됨

Serverless LoRA

지원하지 않음

Fine-tuning

지원하지 않음

Embedding

지원하지 않음

Rerankers

지원하지 않음

지원 Image Input

지원됨

JSON Mode

지원됨

구조화된 Outputs

지원하지 않음

도구

지원됨

Fim Completion

지원하지 않음

Chat Prefix Completion

지원됨

AI 개발을 가속화할 준비가 되셨나요?

AI 개발을 가속화할 준비가 되셨나요?

AI 개발을 가속화할 준비가 되셨나요?