The idea is that for supported factual analysis whose approved methods is already known, asking agents to rediscover the method at request time adds a failure surface that is not needed for expressiveness.<br>For supported enterprise queries, separating language interpretation from policy-selected, reviewed analytical programs is compelling alternative to open-ended tool planning. The model/agent helps determining what the user means without reinventing the measuring procedure.<br>This is not an argument that agents cannot work, it asks why to introduce runtime method invention when approved method already exists. That is a useful question for anyone building analytical systems people must be able to trust and audit.</p>\n","updatedAt":"2026-09-09T14:22:00.551Z","author":{"_id":"62c1b81e8b647bdc24f78027","avatarUrl":"/avatars/b994d31a0c4161f715bb2153c0f0a83f.svg","fullname":"Viktoria Rojkova","name":"vrojkova","type":"user","isPro":true,"isHf":false,"isHfAdmin":false,"isMod":false,"followerCount":1,"isUserFollowing":false}},"numEdits":0,"identifiedLanguage":{"language":"en","probability":0.9332124590873718},"editors":["vrojkova"],"editorAvatarUrls":["/avatars/b994d31a0c4161f715bb2153c0f0a83f.svg"],"reactions":[],"isReport":false}}],"primaryEmailConfirmed":false,"paper":{"id":"2609.03209","authors":[{"_id":"6aa16b29d8c54e38c0a369e5","name":"MasterControl AI Lab","hidden":false}],"publishedAt":"2026-09-02T00:00:00.000Z","submittedOnDailyAt":"2026-09-09T00:00:00.000Z","title":"MasterControl Seventeen Every Time","submittedOnDailyBy":{"_id":"62c1b81e8b647bdc24f78027","avatarUrl":"/avatars/b994d31a0c4161f715bb2153c0f0a83f.svg","isPro":true,"fullname":"Viktoria Rojkova","user":"vrojkova","type":"user","name":"vrojkova"},"summary":"We study a governed approach to enterprise analytics: a language model interprets the question, while deterministic policy selects and runs a pre-approved analytical program that returns both results and evidence. We show that this restriction can remain expressive within a defined analytical class, using relational operations plus aggregation, comparison, windows, ranking, and similarity. Fixed meaning, policy, data, and execution rules also make results replayable. Across 440 runs, three 8B models generated SQL and selected tools at runtime, while Qwen3-8B interpreted intent only and policy executed the approved program. None of 330 runtime-planning episodes matched the full answer-and-evidence contract across all test datasets; the policy-executed analyzer matched 110 of 110. This is a configuration-specific result, not evidence that runtime agents cannot succeed under other designs.","upvotes":1,"discussionId":"6aa16b2ad8c54e38c0a369e6","ai_summary":"A governed analytics framework pairs language models for intent interpretation with deterministic policy execution of pre-approved programs, achieving full answer-and-evidence reliability where runtime-planning agents failed.","ai_keywords":["language model","Qwen3-8B","runtime-planning","agents"],"ai_summary_model":"thinkingmachines/Inkling-Small","organization":{"_id":"62c1b844cc51de96358c7b4f","name":"MasterControlAIML","fullname":"MasterControl","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/noauth/gVXmuxUCuZ6xh1fdgvxWm.png"}},"canReadDatabase":false,"canManagePapers":false,"canSubmit":false,"hasHfLevelAccess":false,"upvoted":false,"upvoters":[{"_id":"677d9f52a1902bef49a962ef","avatarUrl":"https://cdn-avatars.huggingface.co/v1/production/uploads/no-auth/eBbbcP5jH987Ef9jp2d5W.png","isPro":false,"fullname":"Manoj Dobbali","user":"mdobbali","type":"user"}],"acceptLanguages":["en"],"dailyPaperRank":0,"organization":{"_id":"62c1b844cc51de96358c7b4f","name":"MasterControlAIML","fullname":"MasterControl","avatar":"https://cdn-avatars.huggingface.co/v1/production/uploads/noauth/gVXmuxUCuZ6xh1fdgvxWm.png"},"markdownContentUrl":"https://huggingface.co/buckets/huggingchat/papers-content/resolve/2609/2609.03209.md","query":{}}">
MasterControl Seventeen Every Time
Abstract
A governed analytics framework pairs language models for intent interpretation with deterministic policy execution of pre-approved programs, achieving full answer-and-evidence reliability where runtime-planning agents failed.
We study a governed approach to enterprise analytics: a language model interprets the question, while deterministic policy selects and runs a pre-approved analytical program that returns both results and evidence. We show that this restriction can remain expressive within a defined analytical class, using relational operations plus aggregation, comparison, windows, ranking, and similarity. Fixed meaning, policy, data, and execution rules also make results replayable. Across 440 runs, three 8B models generated SQL and selected tools at runtime, while Qwen3-8B interpreted intent only and policy executed the approved program. None of 330 runtime-planning episodes matched the full answer-and-evidence contract across all test datasets; the policy-executed analyzer matched 110 of 110. This is a configuration-specific result, not evidence that runtime agents cannot succeed under other designs.
Community
The idea is that for supported factual analysis whose approved methods is already known, asking agents to rediscover the method at request time adds a failure surface that is not needed for expressiveness.
For supported enterprise queries, separating language interpretation from policy-selected, reviewed analytical programs is compelling alternative to open-ended tool planning. The model/agent helps determining what the user means without reinventing the measuring procedure.
This is not an argument that agents cannot work, it asks why to introduce runtime method invention when approved method already exists. That is a useful question for anyone building analytical systems people must be able to trust and audit.
Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.
Tap or paste here to upload images
Cite arxiv.org/abs/2609.03209 in a model README.md to link it from this page.
Cite arxiv.org/abs/2609.03209 in a dataset README.md to link it from this page.
Cite arxiv.org/abs/2609.03209 in a Space README.md to link it from this page.
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.