# Microsoft study finds AI models score 27% on long-term decision tasks

Digest AI · Research · published 2026-09-30T17:00:00Z

Canonical: https://digestai.news/story/microsoft-study-finds-ai-models-score-27-on-long-term-decision-tasks

## Summary

Microsoft researchers published a paper showing AI systems struggle with long-term decision-making. In a simulated year-long task requiring interconnected choices, models reached only 27% of human performance. The best-performing setup, Qwen3.7-Max with Hermes, still lagged far behind human benchmarks, highlighting persistent reliability gaps in extended planning.

The findings, shared by Rohan Paul on social media, raise questions about industry progress. While unnamed market observers link the results to Anthropic’s November 2026 timeline for a top-tier model, the study itself does not name competitors or speculate on pricing. Google and Meta’s upcoming model releases could further shape perceptions of AI capabilities in this area.

## Key points

- Microsoft’s study shows AI models achieve 27% of human performance in long-term decision tasks over a year
- Qwen3.7-Max with Hermes was the top-performing setup but still underperformed human benchmarks significantly
- Market observers speculate findings may impact Anthropic’s competitive position ahead of November 2026

## Why it matters

The study underscores a critical weakness in AI’s ability to handle sustained, self-correcting decision-making, a core requirement for real-world applications like finance, healthcare, and logistics.

## Sources

1. [Microsoft study reveals AI struggles with long-term decision-making tasks](https://cryptobriefing.com/microsoft-study-reveals-ai-struggles-with-long-term-decision-making-tasks) (cryptobriefing.com, 2026-09-30)

## Cite

Digest AI, "Microsoft study finds AI models score 27% on long-term decision tasks", 30 September 2026, https://digestai.news/story/microsoft-study-finds-ai-models-score-27-on-long-term-decision-tasks

---

Written by Digest AI's editorial model from the linked sources; the sources are the record. Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse
JSON: https://digestai.news/story/microsoft-study-finds-ai-models-score-27-on-long-term-decision-tasks.json
