gpt-oss-120B
2 stories mentioning gpt-oss-120B, newest first, each with its sources and discussion. Follow to see new ones on your front page.
-
PsyAgentBench tests if LLMs truly mimic human psychological biases
Researchers introduced PsyAgentBench, a new benchmark designed to distinguish between LLMs mimicking human psychological patterns and actually possessing those biases. The study re-runs five classic psychology…
1 source primary sourcearXiv cs.CL -
Hugging Face releases tool to cut AI agent consistency gap by half
Hugging Face has introduced a new diagnostic tool and guideline system within its open-source ALTK-Evolve toolkit to address a critical reliability issue in LLM agents. While standard benchmarks report average success…
1 source primary sourceHugging Face
Questions about gpt-oss-120B
What is the latest news about gpt-oss-120B?
PsyAgentBench tests if LLMs truly mimic human psychological biases (22 September 2026).