articleHuggingFace Blog

DeepMath: A lightweight math reasoning Agent with smolagents

DeepMath est un agent de raisonnement mathématique léger basé sur Qwen3-4B Thinking, affiné avec GRPO pour préférer des traces courtes et axées sur le code. Il exécute des snippets Python dans un bac à sable, réduit drastiquement la longueur des réponses et améliore la précision. L’approche est mise en œuvre via smolagents et évaluée sur MATH500, AIME, HMMT et HLE.

published DEC 04, 2025★★★★★

Read the sourcehuggingface.co/blog/intel-deepmath

[*] Opens in a new tab · no tracking on Lantern's side

Source: HuggingFace Blog
Ingested: DEC 04, 2025 · 19:10
Editorial score: 4.0 / 5