Qwen3.7-Max
2 stories mentioning Qwen3.7-Max, newest first, each with its sources and discussion. Follow to see new ones on your front page.
-
Microsoft study finds AI models score 27% on long-term decision tasks
Microsoft researchers published a paper showing AI systems struggle with long-term decision-making. In a simulated year-long task requiring interconnected choices, models reached only 27% of human performance. The…
1 sourcecryptobriefing.com -
BioPhys-Bridge benchmark released for evidence‑grounded biophysical reasoning
A new benchmark called BioPhys-Bridge has been introduced to test language models on interdisciplinary scientific reasoning in biophysics. Each case requires models to ground answers in source evidence, quantitative…
1 source primary sourcearXiv cs.AI
Questions about Qwen3.7-Max
What is the latest news about Qwen3.7-Max?
Microsoft study finds AI models score 27% on long-term decision tasks (30 September 2026).