Hallo, Deutschland!
Mistral opens a Munich hub for Physics AI and Industrial AI research, partnering with German industry.
Read at mistral.ai ↗
Papers, evaluations, benchmarks and lab write-ups.
Mistral opens a Munich hub for Physics AI and Industrial AI research, partnering with German industry.
Read at mistral.ai ↗
In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.
Read at anthropic.com ↗
MentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations.
Read at openai.com ↗
OpenAI is working with an independent Advisory Group on Mathematics and Artificial Intelligence to guide the review and communication of emerging AI results.
Read at openai.com ↗
We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic.
Read at anthropic.com ↗
New OpenAI Economic Research shows how workers use AI beyond traditional roles and which new activities become recurring parts of their work.
Read at openai.com ↗
The true measure of AI is who it helps. Here’s how it’s impacting lives today.
Read at blog.google ↗
Christina Koch sits down with James Manyika, Google’s Senior Vice President of Research, Labs, Technology & Society.
Read at blog.google ↗
6 Sol with Codex to autonomously run quantum computing experiments, analyze results, and calibrate qubits.
Read at openai.com ↗AlphaGenome Atlas maps the molecular effects of 9 billion single-letter DNA variants across the human genome.
Read at deepmind.google ↗We’re sharing an AI-generated solution to the Navier–Stokes Millennium Prize Problem, including a writeup and a formal proof in Lean.
Read at openai.com ↗Apply now for OpenAI’s $5 million grant program supporting independent research on how generative AI affects teen development, well-being, and safety.
Read at openai.com ↗Inside OpenAI, coding agents are reshaping AI research. Explore early data on agent usage, experiment velocity, task complexity, and research acceleration.
Read at openai.com ↗
ChatGPT can now connect to trusted healthcare data, helping clinicians securely access patient context, medical research, and more.
Read at openai.com ↗
Building trust in proprietary model benchmarks using cryptographically secure environments.
Read at deepmind.google ↗A randomized study of more than 1,000 students examines ChatGPT, critical thinking, originality, and student performance on a real-world university assignment.
Read at openai.com ↗
A Blog post by IBM Granite on Hugging Face.
Read at huggingface.co ↗
Google DeepMind partners with game studios to prototype breakthrough AI gameplay.
Read at deepmind.google ↗
OpenAI research reveals how enterprises are adopting agentic AI, using ChatGPT and Codex, and how frontier firms are pulling ahead in AI adoption.
Read at openai.com ↗
OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
Read at openai.com ↗
6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.
Read at openai.com ↗A new field report shows how scientists use AI coding agents to modernize scientific computing, accelerating software development and discovery in genomics and beyond.
Read at openai.com ↗
S. Department of Energy and national labs to use frontier AI to accelerate discovery.
Read at openai.com ↗
A new analysis from OpenAI reveals issues in SWE-Bench Pro, a popular coding benchmark, raising concerns about reliability and accuracy in evaluating AI models.
Read at openai.com ↗
A Blog post by Microsoft on Hugging Face.
Read at huggingface.co ↗
A Blog post by Photoroom on Hugging Face.
Read at huggingface.co ↗
Today, Google DeepMind and A24 are announcing a first-of-its-kind partnership focused on research.
Read at deepmind.google ↗
Introducing GeneBench-Pro, a new benchmark testing AI performance in genomics, biology, and scientific research using complex, real-world datasets.
Read at openai.com ↗
GPT-5 Pro helped solve a 3-year-old immunology mystery, offering insights into T cell behavior. The breakthrough could support cancer and autoimmune research.
Read at openai.com ↗5 Instant improves ChatGPT’s health and wellness responses with stronger reasoning, better context, clearer communication, and physician-informed evaluations.
Read at openai.com ↗Researchers used an OpenAI reasoning model to help diagnose rare diseases, identifying 18 new diagnoses in previously unsolved cases.
Read at openai.com ↗4 improved a key drug-making reaction, advancing medicinal chemistry research.
Read at openai.com ↗Introducing LifeSciBench, an expert-authored, expert-reviewed benchmark for evaluating how AI systems handle real-world life science research tasks and decisions.
Read at openai.com ↗
Discover how astrophysicist Chi-kwan Chan uses Codex to build black hole simulations, helping scientists study extreme physics and test Einstein’s theory of general relativity.
Read at openai.com ↗