#
aime
Here are 4 public repositories matching this topic...
Frontier-level mathematical reasoning from a 1.5B model: 83.3% on AIME 2024 via inference-time compute and learned verification. Technical report TR-2026-01.
machine-learningmctsreasoningverifierbest-of-naimelarge-language-modelsllmtest-time-computeinference-time-compute
-
Updated
Aug 21, 2026 - Python
Official code for "Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration" (SOE, ACL 2026 Oral). A training-free, test-time scaling method that fixes reasoning collapse in LLM mathematical reasoning via orthogonal probing.
nlpagentmathgemmareasoningself-consistencyaimellmchain-of-thoughtvllmqwenweak-to-strongtest-time-computetest-time-scalinginference-time-scalingacl2026
-
Updated
Jul 7, 2026 - Python
Improve this page
Add a description, image, and links to the aime topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the aime topic, visit your repo's landing page and select "manage topics."