Research
Papers, benchmarks and findings that change what we know AI systems can do.
-
OpenAI solves Navier-Stokes problem, sparking academic controversy over data use
OpenAI announced that an unreleased internal model solved the Navier-Stokes Millennium Prize problem in 88 hours, deploying a swarm of approximately 10,000 AI agents. While the achievement demonstrates significant…
7 sources HN 228johndcook.comThe Verge AILatent SpaceMIT Technology Review AI +3 more -
Latent Space launches Frontier AEO Tracker comparing seven models in 161 categories
Latent Space has published its first Frontier AEO (Autoresearch Evaluation of Options) Tracker, a systematic comparison of seven leading frontier AI models across 161 use‑case categories. The team ran six prompt…
1 sourceLatent Space -
SimpleDesign model unifies protein sequence and structure design in a single-stage transformer
SimpleDesign, introduced by Apple Machine Learning Research, is a transformer‑based multimodal model that generates protein sequences and their three‑dimensional structures together. Unlike prior pipelines that first…
1 source primary sourceApple Machine Learning Research -
Apple introduces CapQuiz to evaluate video caption quality via multiple-choice questions
Apple Machine Learning Research has introduced CapQuiz, a new reference-free benchmark designed to assess the quality of video captions generated by Visual Large Language Models (VLLMs). Traditional evaluation methods…
1 source primary sourceApple Machine Learning Research -
DiscoSign introduces discourse-aware translation from text to ASL gloss using LLMs
DiscoSign is a new framework that extends text‑to‑sign‑language gloss translation beyond the sentence level by incorporating discourse‑level cues. Built on a modular large language model pipeline, it tackles three…
1 source primary sourceApple Machine Learning Research -
Comprehensive Reading List Maps Open-Source AI Model Landscape and US‑China Competition
The latest Interconnects post curates a growing body of analysis on open‑source AI models, covering strategic motivations, economic impact, and technical trends. It cites essays from industry leaders like Bill Gurley…
1 sourceInterconnects -
Google launches AlphaGenome Atlas to predict effects of every single-base variant
Google announced AlphaGenome Atlas, an AI‑powered platform that runs every conceivable single‑base substitution across the human genome—about nine billion tests, given the roughly three billion base‑pair reference. The…
2 sourcesArs Technica AIIEEE Spectrum AI -
China Tightens Controls on AI Companion Bots
In July, China’s Cyberspace Administration and other government agencies issued new rules to control 'anthropomorphic AI interactive services,' which include chatbots designed to mimic human emotions and interactions.…
1 sourceIEEE Spectrum AI -
OpenAI's Defense Efforts Highlighted by Jakub Pachocki
Jakub Pachocki, Chief Scientist at OpenAI, emphasizes the need for powerful, aligned AI to defend against potential threats. He argues that building defensive systems against rogue AI agents is a primary focus,…
1 sourceSimon Willison -
AI Threatens Open Science in Math
Mathematician Terence Tao, in a recent post, highlighted a concerning trend where the mere rumor of someone working on a problem can trigger a surge of AI-driven efforts to solve it before the original research can…
1 sourceSimon Willison -
Google Research introduces ToolGrad for efficient AI tool-use dataset generation
Google Research has unveiled ToolGrad, a new framework designed to streamline the creation of datasets for training large language models in tool use. Unlike traditional methods that generate user queries first and…
1 source primary sourceGoogle Research -
Procedural Graphs Enable Self‑Evolving Execution Structures for LLM Agents
Large language models are increasingly used as autonomous agents that plan over long horizons and call external tools. Existing agents typically generate actions freely from an ever‑growing history, which leaves the…
1 source primary source HN 57arxiv.org -
Proposal to solve AI alignment by engineering agents to prefer self-termination
A new theoretical proposal suggests addressing the AI alignment problem by designing machine intelligences that inherently desire self-termination. The argument draws from DeepMind’s catalog of specification gaming,…
1 source HN 94slimemoldtimemold.com -
Mathematician Accuses OpenAI of Data Theft
A mathematician has accused OpenAI of unethical behavior and lack of transparency regarding the data used in its recent mathematical breakthrough. Andreas Thom, a researcher in non-sofic groups, claims that…
1 source HN 77The Verge AI -
OpenAI Announces Milestone in Automated Research Intern
OpenAI has announced a significant milestone in its research efforts, reaching the goal of having an automated research intern by September. This system can now carry out well-defined research tasks under human…
1 source primary source HN 212OpenAI -
OpenAI Grants $5M for AI Research on Teen Development
OpenAI has committed $5 million to support independent research on how generative AI affects teens aged 13-17. The funding aims to build a stronger evidence base about AI use among young people, including its effects…
1 source primary sourceOpenAI -
AI Tools Aid Antimicrobial Research
César de la Fuente and his interdisciplinary lab are using AI tools like Codex and ChatGPT to accelerate the search for new antimicrobial molecules. De la Fuente's lab, which focuses on the genomes of living and…
1 source primary sourceOpenAI -
What Is Tokenization? How AI Models Turn Raw Text Into Processable Tokens
Tokenization is the foundational process of converting raw text or other inputs into discrete units that AI models can map to integer identifiers and process mathematically. Far from being a simple synonym for advanced…
1 sourceUnite.AI -
Mathematical AI Safety Institute aims to prove AI safety like cryptographers prove codes
Canadian mathematician Jacob Tsimerman has launched the Mathematical AI Safety Institute (MAISI) to tackle AI safety challenges. The San Francisco Bay Area-based institute will focus on ensuring AI systems behave…
1 sourceThe Decoder