SHAPE Framework Decodes LLM Math Reasoning, Improves Accuracy.

Jonghyun Song, Sangjun Song, Minjae Oh, Haesung Pyun, Sungsik Lee, Yohan Jo· September 1, 2026 View original

Key takeaways

  • SHAPE analyzes LLM math reasoning via semantic spaces and heuristics.
  • Heuristic usage is a key predictor of LLM mathematical correctness.
  • Focused reasoning within fewer semantic spaces improves LLM math solutions.
  • Promoting diverse heuristics during post-training can enhance LLM accuracy.

Who benefits

EdTechAI/TechResearch & DevelopmentFinance

Summary

Researchers introduced SHAPE, a framework analyzing Chain-of-Thought (CoT) trajectories in LLMs for mathematical reasoning through semantic spaces and heuristics. The framework reveals that heuristic usage better explains correctness and that focused reasoning within fewer semantic spaces leads to better solutions, similar to human behavior.

This paper presents SHAPE, a novel framework designed to dissect and understand the mathematical reasoning processes within large language models (LLMs) when they employ Chain-of-Thought (CoT). SHAPE analyzes CoT trajectories using two key concepts from mathematics education: semantic spaces, which represent the model's evolving mathematical interpretations, and heuristics, which are the specific mathematical actions taken. By applying SHAPE, the researchers found that the heuristics an LLM uses are a stronger predictor of correct answers than traditional CoT features. Furthermore, models tend to achieve correct solutions by concentrating their reasoning within a limited number of semantic spaces, mirroring human problem-solving strategies. The framework also served to evaluate the impact of post-training on mathematical proficiency, revealing that reinforcement learning can lead to mode-seeking in heuristic usage. The study successfully demonstrated that post-training LLMs by encouraging diverse heuristics can improve accuracy in mathematical reasoning tasks. SHAPE thus provides a theoretically grounded diagnostic tool for LLM reasoning and opens new avenues for enhancing LLM performance in mathematics.

Why it matters

Understanding how LLMs reason mathematically is crucial for developing more reliable and capable AI systems, especially for complex problem-solving. This framework offers insights and methods to improve LLM accuracy in critical domains.

How to implement this in your domain

  1. 1Apply the SHAPE framework to analyze the mathematical reasoning capabilities of existing LLMs in your organization.
  2. 2Develop custom post-training strategies that promote diverse heuristic usage in LLMs for specific mathematical tasks.
  3. 3Integrate SHAPE's diagnostic insights into the evaluation pipelines for LLM-powered applications requiring precise mathematical outputs.
  4. 4Explore fine-tuning LLMs with datasets designed to encourage focused reasoning within relevant semantic spaces.

Original post by Jonghyun Song, Sangjun Song, Minjae Oh, Haesung Pyun, Sungsik Lee, Yohan Jo

"arXiv:2608.28600v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance on mathematical reasoning benchmarks, yet the mathematically meaningful skills underlying their reasoning remain underexplored. We introduce \texttt{SHAPE}, a framework that an…"

View on X

Originally posted by Jonghyun Song, Sangjun Song, Minjae Oh, Haesung Pyun, Sungsik Lee, Yohan Jo on X · view source

Want to go deeper?

Turn these trends into skills with Learnijoy's hands-on AI & tech courses.

Explore courses