Hugging Face Daily Papers · · 3 min read

SHAPE of Chain-of-Thought in Math Reasoning

Mirrored from Hugging Face Daily Papers for archival readability. Support the source by reading on the original site.

Impressive work</p>\n","updatedAt":"2026-09-01T05:30:35.739Z","author":{"_id":"64a3b603fbd994e0767b52e9","avatarUrl":"/avatars/0eecc4db5b4da27703204b9301440a4b.svg","fullname":"Minjae Oh","name":"Riasok","type":"user","isPro":false,"isHf":false,"isHfAdmin":false,"isMod":false,"isUserFollowing":false}},"numEdits":0,"identifiedLanguage":{"language":"en","probability":0.9043188691139221},"editors":["Riasok"],"editorAvatarUrls":["/avatars/0eecc4db5b4da27703204b9301440a4b.svg"],"reactions":[{"reaction":"🔥","users":["Aiant56"],"count":1}],"isReport":false}}],"primaryEmailConfirmed":false,"paper":{"id":"2608.28600","authors":[{"_id":"6a962d1ecd6ebc484732ebe1","name":"Jonghyun Song","hidden":false},{"_id":"6a962d1ecd6ebc484732ebe2","name":"Sangjun Song","hidden":false},{"_id":"6a962d1ecd6ebc484732ebe3","name":"Minjae Oh","hidden":false},{"_id":"6a962d1ecd6ebc484732ebe4","name":"Haesung Pyun","hidden":false},{"_id":"6a962d1ecd6ebc484732ebe5","name":"Sungsik Lee","hidden":false},{"_id":"6a962d1ecd6ebc484732ebe6","name":"Yohan Jo","hidden":false}],"publishedAt":"2026-06-28T00:00:00.000Z","submittedOnDailyAt":"2026-09-01T00:00:00.000Z","title":"SHAPE of Chain-of-Thought in Math Reasoning","submittedOnDailyBy":{"_id":"64a3b603fbd994e0767b52e9","avatarUrl":"/avatars/0eecc4db5b4da27703204b9301440a4b.svg","isPro":false,"fullname":"Minjae Oh","user":"Riasok","type":"user","name":"Riasok"},"summary":"Large language models (LLMs) achieve strong performance on mathematical reasoning benchmarks, yet the mathematically meaningful skills underlying their reasoning remain underexplored. We introduce SHAPE, a framework that analyzes Chain-of-Thought (CoT) trajectories through two lenses developed in mathematics education: (1) semantic spaces: the model's evolving mathematical interpretations of a problem (e.g., algebraic, geometric), and (2) heuristics: the specific mathematical actions taken within those spaces (e.g., simplifying the problem, working backward). We first use SHAPE to analyze the reasoning patterns of various models. Our findings reveal that the mathematical heuristics employed by a model better explain final answer correctness than traditional CoT features. Furthermore, models are likely to reach correct solutions by concentrating their reasoning effort within a few semantic spaces rather than exploring many disparate ones -- a pattern consistent with human behavior. Next, we utilize the SHAPE lens to evaluate whether post-training truly enhances mathematical proficiency. We find that reinforcement learning induces mode-seeking in heuristic usage. Lastly, we post-train LLMs by promoting diverse heuristics and demonstrate its effectiveness in improving accuracy. Overall, SHAPE provides a theoretically-grounded diagnostic framework for decoding LLM reasoning and offers a new path toward post-training LLMs for math reasoning. The code for our model is available at https://github.com/holi-lab/SHAPE-of-CoT","upvotes":14,"discussionId":"6a962d1ecd6ebc484732ebe7","ai_summary":"SHAPE analyzes chain-of-thought reasoning via semantic spaces and heuristics to diagnose LLM mathematical reasoning and improve post-training.","ai_keywords":["Chain-of-Thought","semantic spaces","heuristics","reinforcement learning","mode-seeking","post-training"],"ai_summary_model":"thinkingmachines/Inkling-Small","organization":{"_id":"66d54dc8033492801db2bf5a","name":"SeoulNatlUniv","fullname":"Seoul National University","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/659ccc9d18897eb6594e897f/_-0BM-1UyM-d-lRiahFnf.png"}},"canReadDatabase":false,"canManagePapers":false,"canSubmit":false,"hasHfLevelAccess":false,"upvoted":false,"upvoters":[{"_id":"66ac7b0997a8c9192bc551df","avatarUrl":"/avatars/41e9d93cde502e8235f9c8bd20be89cc.svg","isPro":false,"fullname":"Sangjun Song","user":"ssangjun706","type":"user"},{"_id":"6a22c78479a2afc4ecb81e7e","avatarUrl":"/avatars/2a56eedd4da50982bd35a71418f27a40.svg","isPro":false,"fullname":"Rafael Mendoza","user":"rfaelmdz","type":"user"},{"_id":"64a3b603fbd994e0767b52e9","avatarUrl":"/avatars/0eecc4db5b4da27703204b9301440a4b.svg","isPro":false,"fullname":"Minjae Oh","user":"Riasok","type":"user"},{"_id":"669f80b63d38b52c79bdf8fc","avatarUrl":"/avatars/1a88d11c1408c1373ba148e186e3a0f1.svg","isPro":false,"fullname":"sungjiblim","user":"sungzip","type":"user"},{"_id":"650fcd442a45730c3ffcbdb6","avatarUrl":"/avatars/17d7de506dcf870d35fdcd0ddd5cc2ee.svg","isPro":false,"fullname":"Heejae Suh","user":"boribori","type":"user"},{"_id":"6552f9e2ab7c20ac6fe7e556","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/6552f9e2ab7c20ac6fe7e556/WYErK8nUyPXn4QmfhNlze.jpeg","isPro":false,"fullname":"John","user":"johnhan00","type":"user"},{"_id":"65e9343d063e16f1c3eabe5b","avatarUrl":"/avatars/49700b15eb7b31769930798fb1d85112.svg","isPro":false,"fullname":"Woojung Song","user":"Opusdei","type":"user"},{"_id":"686605c5eff672038883bad1","avatarUrl":"/avatars/ef688cc260afa6f1712b548a89f0e0a4.svg","isPro":false,"fullname":"Hoyeol Yang","user":"hoyeolyang","type":"user"},{"_id":"662219a6a46ff7ee8823ebb5","avatarUrl":"/avatars/7e5e1288e15ba7bbcd9a645b12199724.svg","isPro":false,"fullname":"Injin Kong","user":"youuor7r","type":"user"},{"_id":"68307ba77a61df79768f5372","avatarUrl":"/avatars/bab03d64360c72d3124f178d871ddf5e.svg","isPro":false,"fullname":"kimnalim","user":"nalim0230","type":"user"},{"_id":"6359cc256a6195408089393d","avatarUrl":"/avatars/79efa258d2e039cacfa4e5ffd1a0089f.svg","isPro":false,"fullname":"Jong Song","user":"hyung22","type":"user"},{"_id":"67e62e2e85286d639823ee15","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/no-auth/hMXbFXaG4bHNLo0QuEvC1.png","isPro":false,"fullname":"SeungWon Kook","user":"Aiant56","type":"user"}],"acceptLanguages":["en"],"dailyPaperRank":0,"organization":{"_id":"66d54dc8033492801db2bf5a","name":"SeoulNatlUniv","fullname":"Seoul National University","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/659ccc9d18897eb6594e897f/_-0BM-1UyM-d-lRiahFnf.png"},"markdownContentUrl":"https://huggingface.co/buckets/huggingchat/papers-content/resolve/2608/2608.28600.md","query":{}}">
Papers
arxiv:2608.28600

SHAPE of Chain-of-Thought in Math Reasoning

Published on Jun 28
· Submitted by
Minjae Oh
on Sep 1
Authors:
,

Abstract

SHAPE analyzes chain-of-thought reasoning via semantic spaces and heuristics to diagnose LLM mathematical reasoning and improve post-training.

Large language models (LLMs) achieve strong performance on mathematical reasoning benchmarks, yet the mathematically meaningful skills underlying their reasoning remain underexplored. We introduce SHAPE, a framework that analyzes Chain-of-Thought (CoT) trajectories through two lenses developed in mathematics education: (1) semantic spaces: the model's evolving mathematical interpretations of a problem (e.g., algebraic, geometric), and (2) heuristics: the specific mathematical actions taken within those spaces (e.g., simplifying the problem, working backward). We first use SHAPE to analyze the reasoning patterns of various models. Our findings reveal that the mathematical heuristics employed by a model better explain final answer correctness than traditional CoT features. Furthermore, models are likely to reach correct solutions by concentrating their reasoning effort within a few semantic spaces rather than exploring many disparate ones -- a pattern consistent with human behavior. Next, we utilize the SHAPE lens to evaluate whether post-training truly enhances mathematical proficiency. We find that reinforcement learning induces mode-seeking in heuristic usage. Lastly, we post-train LLMs by promoting diverse heuristics and demonstrate its effectiveness in improving accuracy. Overall, SHAPE provides a theoretically-grounded diagnostic framework for decoding LLM reasoning and offers a new path toward post-training LLMs for math reasoning. The code for our model is available at https://github.com/holi-lab/SHAPE-of-CoT

Community

Paper submitter about 3 hours ago

Impressive work

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.
Tap or paste here to upload images

· Sign up or log in to comment

Get this paper in your agent:

hf papers read 2608.28600
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper

No model linking this paper

Cite arxiv.org/abs/2608.28600 in a model README.md to link it from this page.

Datasets citing this paper

No dataset linking this paper

Cite arxiv.org/abs/2608.28600 in a dataset README.md to link it from this page.

Spaces citing this paper

No Space linking this paper

Cite arxiv.org/abs/2608.28600 in a Space README.md to link it from this page.

Collections including this paper

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from Hugging Face Daily Papers