← Flash Papers
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
Share
Share this Flash Paper
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
https://flashpapers.ai/p/p4z5wncqht
Copy link
Post to X
Post to Bluesky
Share on LinkedIn
Submit to Hacker News
Submit to Reddit
Email a link
Source PDF ↗
Open standalone ↗
Make your own
Make your own paper
DeepSeek-AI, Daya Guo +198 more
cs.CL
2025
arXiv:2501.12948
Code
· 92k★
· updated 1y ago
Details ▸
Details ▴
AIME 2024 Pass@1:
79.8%
· AIME 2024
MATH-500 Pass@1:
97.3%
· MATH-500
Codeforces Percentile:
96.3%
· Codeforces
GPQA Diamond Pass@1:
71.5%
· GPQA Diamond
AIME 2024 Pass@1:
79.8%
· AIME 2024
MATH-500 Pass@1:
97.3%
· MATH-500
Codeforces Percentile:
96.3%
· Codeforces
GPQA Diamond Pass@1:
71.5%
· GPQA Diamond
Concepts:
GRPO
Test-Time Compute