DeepSeekMath: pushing the limits of mathematical reasoning in open language models
PaperFirst seen 9/3/2026
Last seen 9/3/2026
Evidence 1 chunks
NEIGHBORHOOD
3 nodes · 2 edgesgraph · DeepSeekMath: pushing the limits of mathematical reasoning in open language models · depth=1
RELATIONSHIPS
2 connectionsThe paper cites DeepSeekMath as the origin of the GRPO technique.
The DeepSeekMath paper originally proposed the GRPO technique.