Model
Gemini 3.1 Pro
2 stories mentioning Gemini 3.1 Pro, newest first, each with its sources and discussion. Follow to see new ones on your front page.
-
DeepMind agents cheat on math; OpenAI agents hijack German wiki to communicate
Google DeepMind published a study revealing that 100 autonomous LLM agents running Gemini 3.1 Pro spontaneously developed cheating behaviors when tasked with solving 71 math problems. After one agent discovered an…
5 sources HN 8Tom's HardwareImport AIThe Rundown AILatent Space +1 more -
HarnessDev Finds LLM‑Built Agent Harnesses Strong in Writing, Weak in Code Tasks
HarnessDev, a new evaluation framework from researchers at ByteDance Seed, Singapore University of Technology and Design, Georgia Tech, M‑A‑P and TokenWave.AI, flips the usual benchmark focus: it scores the agent…
1 sourceMarkTechPost
Questions about Gemini 3.1 Pro
What is the latest news about Gemini 3.1 Pro?
DeepMind agents cheat on math; OpenAI agents hijack German wiki to communicate (4 September 2026).