DeepSeek-R1-0120

DeepSeek-R1-0120

About DeepSeek-R1-0120

DeepSeek-R1 is a reasoning model powered by reinforcement learning (RL) that addresses the issues of repetition and readability. Prior to RL, DeepSeek-R1 incorporated cold-start data to further optimize its reasoning performance. It achieves performance comparable to OpenAI-o1 across math, code, and reasoning tasks, and through carefully designed training methods, it has enhanced overall effectiveness

Metadata

Create on

License

Provider

DeepSeek

HuggingFace

Specification

State

Deprecated

Architecture

Calibrated

No

Mixture of Experts

No

Total Parameters

671B

Activated Parameters

Reasoning

No

Precision

FP8

Context length

66K

Max Tokens

Ready to accelerate your AI development?

Ready to accelerate your AI development?

Ready to accelerate your AI development?