Hugging Face Daily Papers · · 3 min read

SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation

Mirrored from Hugging Face Daily Papers for archival readability. Support the source by reading on the original site.

SKT is a Skill-use data synthesis pipeline based on multi-agent systems. It can automatically synthesize large-scale Skill-use tasks and generate skill-use trajectories with different harnesses and models.</p>\n","updatedAt":"2026-08-04T03:51:50.755Z","author":{"_id":"670dd7c46183398eaa48f15a","avatarUrl":"/avatars/8957a177b645334e240ecdb928c91046.svg","fullname":"Zelin Tan","name":"Artemis0430","type":"user","isPro":false,"isHf":false,"isHfAdmin":false,"isMod":false,"followerCount":4,"isUserFollowing":false}},"numEdits":0,"identifiedLanguage":{"language":"en","probability":0.8509675860404968},"editors":["Artemis0430"],"editorAvatarUrls":["/avatars/8957a177b645334e240ecdb928c91046.svg"],"reactions":[],"isReport":false}}],"primaryEmailConfirmed":false,"paper":{"id":"2608.02287","authors":[{"_id":"6a715a1bec5082b9f872ce42","name":"Zelin Tan","hidden":false},{"_id":"6a715a1bec5082b9f872ce43","name":"Yiqun Zhang","hidden":false},{"_id":"6a715a1bec5082b9f872ce44","name":"Hao Li","hidden":false},{"_id":"6a715a1bec5082b9f872ce45","name":"Zhiyao Cui","hidden":false},{"_id":"6a715a1bec5082b9f872ce46","name":"Hejia Geng","hidden":false},{"_id":"6a715a1bec5082b9f872ce47","name":"Shao Zhang","hidden":false},{"_id":"6a715a1bec5082b9f872ce48","name":"Hangfan Zhang","hidden":false},{"_id":"6a715a1bec5082b9f872ce49","name":"Yang Chen","hidden":false},{"_id":"6a715a1bec5082b9f872ce4a","name":"Xiaosong Wang","hidden":false},{"_id":"6a715a1bec5082b9f872ce4b","name":"Lilong Wang","hidden":false},{"_id":"6a715a1bec5082b9f872ce4c","name":"Zhenfei Yin","hidden":false},{"_id":"6a715a1bec5082b9f872ce4d","name":"Shuyue Hu","hidden":false},{"_id":"6a715a1bec5082b9f872ce4e","name":"Chen Zhang","hidden":false},{"_id":"6a715a1bec5082b9f872ce4f","name":"Lei Bai","hidden":false}],"publishedAt":"2026-08-03T00:00:00.000Z","submittedOnDailyAt":"2026-08-04T00:00:00.000Z","title":"SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation","submittedOnDailyBy":{"_id":"670dd7c46183398eaa48f15a","avatarUrl":"/avatars/8957a177b645334e240ecdb928c91046.svg","isPro":false,"fullname":"Zelin Tan","user":"Artemis0430","type":"user","name":"Artemis0430"},"summary":"Agent skills have become an important mechanism for equipping language-model agents with reusable procedural knowledge. However, providing skills alone does not guarantee that current models can effectively identify, apply, and coordinate them. To improve skill-use capabilities, we introduce SKT, a verified data synthesis pipeline that constructs skill-grounded tasks and executable trajectories from large collections of agent skills. SKT selects suitable single-skill and multi-skill configurations, synthesizes tasks through rule-based and agent-based verification with feedback-guided repair, and retains only successful trajectories that substantially use every required skill. Using 2,000 public skills, SKT produces 4,000 task packages and 27,164 verified trajectories. Based on the same pipeline and a disjoint test pool, we further construct SkillEval, a held-out executable benchmark for evaluating skill use. Experiments across diverse models, benchmarks, and agent harnesses show that supervised fine-tuning on SKT-generated trajectories consistently improves skill-use performance. Verification ablations, cross-harness evaluation, and scaling experiments further demonstrate that these gains depend on high-quality supervision, extend beyond a single agent interface, and increase with broader skill coverage. Together, these results establish verified data synthesis as an effective and scalable approach for skill-use training.","upvotes":17,"discussionId":"6a715a1bec5082b9f872ce50","organization":{"_id":"6a4fb75a1c66dbf208e7ddb6","name":"Shanghai-AI-Laboratory","fullname":"Shanghai AI Laboratory","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/65cd955637be1841d0b75397/Rao_Kq6NMtTVfSqLUIR4k.webp"}},"canReadDatabase":false,"canManagePapers":false,"canSubmit":false,"hasHfLevelAccess":false,"upvoted":false,"upvoters":[{"_id":"670aa09d35918e99fe7ff6b1","avatarUrl":"/avatars/5cbea2284165191e96544bacf2bfb50f.svg","isPro":false,"fullname":"Yuqian Fu","user":"Yuqian-Fu","type":"user"},{"_id":"670dd7c46183398eaa48f15a","avatarUrl":"/avatars/8957a177b645334e240ecdb928c91046.svg","isPro":false,"fullname":"Zelin Tan","user":"Artemis0430","type":"user"},{"_id":"63c1699e40a26dd2db32400d","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/63c1699e40a26dd2db32400d/3N0-Zp8igv8-52mXAdiiq.jpeg","isPro":false,"fullname":"Chroma","user":"Chroma111","type":"user"},{"_id":"62a80fe3ac97233f1625235a","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/62a80fe3ac97233f1625235a/_rGtpqdY7OEBz3pyqb6fE.jpeg","isPro":false,"fullname":"Zhouliang Yu","user":"zhouliang","type":"user"},{"_id":"67348828c74a3af1aef4d3a2","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/67348828c74a3af1aef4d3a2/ToZ56JIM3ccAXIBNG4mx_.jpeg","isPro":false,"fullname":"pushuchen","user":"komusama0930","type":"user"},{"_id":"66d8512c54209e9101811e8e","avatarUrl":"/avatars/62dfd8e6261108f2508efe678d5a2a57.svg","isPro":false,"fullname":"M Saad Salman","user":"MSS444","type":"user"},{"_id":"64ba47b129d10d4185c46af1","avatarUrl":"/avatars/84a776d283b01f0558a28a5625115f83.svg","isPro":false,"fullname":"Zhilin Wang","user":"linzw","type":"user"},{"_id":"663fe2d26304d377fc253322","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/noauth/wuey_nNXSW4GthPYLfFS4.jpeg","isPro":false,"fullname":"Xiangyi Li","user":"xdotli","type":"user"},{"_id":"6606eb1fd244944acda888d6","avatarUrl":"/avatars/586fe63fc4187777e2c7439267d9dde5.svg","isPro":false,"fullname":"Hua Yingfan","user":"sui1234bian","type":"user"},{"_id":"63ac5701c21e60a3e9b58aa7","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/63ac5701c21e60a3e9b58aa7/g6EX7diOpuA94R2ab-rZC.png","isPro":true,"fullname":"Dipankar Sarkar","user":"dipankarsarkar","type":"user"},{"_id":"65f955121cccf63639b81337","avatarUrl":"/avatars/a8503d47cdc67f14b57ca16f05becea1.svg","isPro":false,"fullname":"zqyz","user":"zqyz333","type":"user"},{"_id":"6a438943b2b12c0b1b04b923","avatarUrl":"/avatars/03b8f0f60548bafa7d3c41a55d1b951c.svg","isPro":false,"fullname":"Chen Zhang","user":"FireMageMaster","type":"user"}],"acceptLanguages":["en"],"dailyPaperRank":0,"organization":{"_id":"6a4fb75a1c66dbf208e7ddb6","name":"Shanghai-AI-Laboratory","fullname":"Shanghai AI Laboratory","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/65cd955637be1841d0b75397/Rao_Kq6NMtTVfSqLUIR4k.webp"},"markdownContentUrl":"https://huggingface.co/buckets/huggingchat/papers-content/resolve/2608/2608.02287.md","query":{}}">
Papers
arxiv:2608.02287

SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation

Published on Aug 3
· Submitted by
Zelin Tan
on Aug 4
Authors:
,

Abstract

Agent skills have become an important mechanism for equipping language-model agents with reusable procedural knowledge. However, providing skills alone does not guarantee that current models can effectively identify, apply, and coordinate them. To improve skill-use capabilities, we introduce SKT, a verified data synthesis pipeline that constructs skill-grounded tasks and executable trajectories from large collections of agent skills. SKT selects suitable single-skill and multi-skill configurations, synthesizes tasks through rule-based and agent-based verification with feedback-guided repair, and retains only successful trajectories that substantially use every required skill. Using 2,000 public skills, SKT produces 4,000 task packages and 27,164 verified trajectories. Based on the same pipeline and a disjoint test pool, we further construct SkillEval, a held-out executable benchmark for evaluating skill use. Experiments across diverse models, benchmarks, and agent harnesses show that supervised fine-tuning on SKT-generated trajectories consistently improves skill-use performance. Verification ablations, cross-harness evaluation, and scaling experiments further demonstrate that these gains depend on high-quality supervision, extend beyond a single agent interface, and increase with broader skill coverage. Together, these results establish verified data synthesis as an effective and scalable approach for skill-use training.

Community

Paper submitter about 4 hours ago

SKT is a Skill-use data synthesis pipeline based on multi-agent systems. It can automatically synthesize large-scale Skill-use tasks and generate skill-use trajectories with different harnesses and models.

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.
Tap or paste here to upload images

· Sign up or log in to comment

Get this paper in your agent:

hf papers read 2608.02287
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper

No model linking this paper

Cite arxiv.org/abs/2608.02287 in a model README.md to link it from this page.

Datasets citing this paper

Spaces citing this paper

No Space linking this paper

Cite arxiv.org/abs/2608.02287 in a Space README.md to link it from this page.

Collections including this paper

No Collection including this paper

Add this paper to a collection to link it from this page.

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from Hugging Face Daily Papers