{"version":1,"type":"story","url":"https://digestai.news/story/study-proposes-skill-habits-to-fix-ai-agent-inconsistency","json":"https://digestai.news/story/study-proposes-skill-habits-to-fix-ai-agent-inconsistency.json","markdown":"https://digestai.news/story/study-proposes-skill-habits-to-fix-ai-agent-inconsistency.md","slug":"study-proposes-skill-habits-to-fix-ai-agent-inconsistency","headline":"Study proposes skill habits to fix AI agent inconsistency","summary":"A new arXiv paper addresses the lack of consistency in AI agents, which often produce varying results for identical tasks. The authors tested 42 tasks three times each and found that 38% to 74% of responses were inconsistent depending on the model. They also noted that 95.3% to 97.2% of generated tokens are wasted on re-deriving known plans.\n\nThe researchers propose \"skill habit formation,\" where agents mine their own execution history to create deterministic scripts for common tasks. These scripts compete with standard reasoning, allowing the common case to run as a script while complex cases fall through to reasoning. On text-to-SQL tasks, a habit-formed variant reproduced its output on all 456 repeated dispatches, compared to 11 to 26 out of 42 for standard reasoning arms. This approach also reduced token usage by 14% to 56%.\n\nHowever, the method has trade-offs. The system incorrectly admitted work it should have deferred on 2.6% of natural paraphrases and 26% of boundary inputs. Most of these failures were invisible to the system's safety checks. The authors conclude that while deterministic errors repeat exactly, separating routing from parameter extraction improved end-to-end accuracy from 0.888 to 0.952 at 43% of the cost.","keyPoints":["Agents showed 38% to 74% inconsistency across 42 repeated tasks in the study.","Skill habit formation reduced token usage by 14% to 56% on text-to-SQL tasks.","The method improved end-to-end accuracy from 0.888 to 0.952 at 43% of the cost."],"whyItMatters":"This research offers a path to making AI agents reliable enough for regulated industries like auditing and finance, where consistent, auditable outputs are required. It also significantly reduces computational costs by minimizing redundant reasoning.","category":{"slug":"research","name":"Research","url":"https://digestai.news/category/research"},"entities":{"companies":[],"models":[],"people":[]},"firstPublishedAt":"2026-09-23T04:00:00Z","updatedAt":"2026-09-23T04:00:00Z","sourceCount":1,"hasPrimarySource":true,"sources":[{"outlet":"arXiv cs.AI","title":"Making Agents More Consistent: Skills Should Form Habits for Repeat Tasks","url":"https://arxiv.org/abs/2609.25299","publishedAt":"2026-09-23T04:00:00Z","type":"primary","primary":true,"lead":true}],"sourceNotes":null,"discussions":[],"thread":null,"cite":{"text":"Digest AI, \"Study proposes skill habits to fix AI agent inconsistency\", 23 September 2026, https://digestai.news/story/study-proposes-skill-habits-to-fix-ai-agent-inconsistency","publisher":"Digest AI","title":"Study proposes skill habits to fix AI agent inconsistency","datePublished":"2026-09-23T04:00:00Z","url":"https://digestai.news/story/study-proposes-skill-habits-to-fix-ai-agent-inconsistency"},"generatedBy":"Written by Digest AI's editorial model from the linked sources; the sources are the record.","license":"Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse"}