AI Research & Science
Peer-reviewed breakthroughs, university studies, and lab discoveries — explained in plain English. AI Maestro tracks the frontiers of machine learning research, neuroscience meets AI, and the science driving the next wave of intelligent systems.
Mistral Large 4
Mistral has released Mistral Large 4, a new model from the French firm that competes directly with the latest offerings from American…
Top stories

Reflection debuts Beam, an open-weight AI model to rival Chinese models at lower compute cost
yesterday
ReviewBench: An open benchmark for AI code review
yesterdayEmTech Future 2026: When AI Meets Everything
yesterday
NASA and IBM’s open source lunar model turns 17 years of orbiter data into a foundation for lunar science
2d agoMore ai research & science

More than 20 leading AI researchers warn that automated AI research poses extreme risks
More than 20 AI researchers, including Geoffrey Hinton, Yoshua Bengio, and Jakub Pachocki, have…
28 Sep 2026
FBI Hackers Say They Won’t Publish Massive Trove of FBI Employee Data
The hackers behind the massive FBI breach told 404 Media on Monday they do…
28 Sep 2026
Anthropic’s Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks while costing up to 30 percent less per task
Anthropic has launched Claude Sonnet 5.5, the second iteration of its Claude 5.5 family.…
28 Sep 2026
When can we say AI made a scientific discovery?
On Wednesday, Anthropic announced that its molecular biology lab had made a discovery after…
28 Sep 2026
Holo4: powering generalist computer-use agents
Holo4: powering generalist computer-use agents The new Holo4 model series arrives in two sizes:…
28 Sep 2026
The Next Evolution of AI Is Learning From Your Dodgy Gaming Skills
A British startup is betting that the thumbstick twirls and trigger squeezes of casual…
28 Sep 2026
A Coding Guide to Google Research’s MSEB: Writing Sound Encoders to the Benchmark Contract and Scoring Them Across Classification, Clustering, Retrieval and Segmentation
Google Research has released MSEB, a Massive Sound Embedding Benchmark designed to test how…
27 Sep 2026
OpenAI’s GPT-6 Astra can now tell you exactly where you screwed up your IKEA shelf
OpenAI’s GPT-6 Astra has achieved an 80% accuracy rate in identifying assembly errors on…
26 Sep 2026
Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for Exhaustive List Building
Exa has released Agent Ultra, the highest effort level of its Exa Agent API.…
26 Sep 2026
