AI Research & Science
Peer-reviewed breakthroughs, university studies, and lab discoveries — explained in plain English. AI Maestro tracks the frontiers of machine learning research, neuroscience meets AI, and the science driving the next wave of intelligent systems.
JetBrains Releases Mellum2.1: A 12B MoE Open Model for Coding Agents
JetBrains has released Mellum2.1, an open model built for coding agents and fast sub-agents. Mellum2.1 is a 12B mixture-of-experts thinking model from…
Top stories

Claude Haiku 5.5 arrives with massive price cuts proving the AI pricing arms race is far from over
yesterday
Google bets Gemini can turn casual players into game developers with new Playground feature
yesterday
Google invests millions in Mark Zuckerberg’s efforts to create a ‘virtual cell’
yesterdayMistral Large 4
2d agoMore ai research & science

Strix Halo Llama.cpp MTP Benchmarks: 27B Gets Much Faster, 35B Is Mixed
“`html Strix Halo LLaMA.cpp MTP Benchmarks: 27B Gets Much Faster, 35B Is Mixed Strix…
16 May 2026Sam Altman’s ego was OpenAI’s downfall
Sam Altman’s Ego Was OpenAI’s Downfall The more I watch OpenAI, the more convinced…
16 May 2026
My professor uses an AI detector that flags literally everything. Here is how I finally beat it.
**Editorial Brief** The news item highlights a professor’s struggle with an AI detector that…
16 May 2026
Qwen3.6-35B-A3B and 9B are officially on the public Terminal-Bench 2.0 leaderboard!
“`html British AI research has officially added Qwen3.6-35B-A3B and Qwen9B to the public Terminal-Bench…
16 May 2026
New benchmark shows Claude Mythos and GPT-5.5 can develop real browser exploits autonomously
New benchmark shows Claude Mythos and GPT-5.5 can develop real browser exploits autonomously Key…
16 May 2026
For $1.3 million a month, OpenClaw founder Peter Steinberger runs 100 AI agents that code, review PRs, and find bugs
Peter Steinberger, founder of OpenClaw, operates an AI-driven development team where 100 codex instances…
16 May 2026
Poetiq’s Meta-System Automatically Builds a Model-Agnostic Harness That Improved Every LLM Tested on LiveCodeBench Pro Without Fine-Tuning
Poetiq has just published some very interesting results showing its Meta-System reached a new…
16 May 2026
Best AI Agents for Software Development Ranked: A Benchmark-Driven Look at the Current Field
The AI coding agent market looks almost unrecognizable compared to 2024 or even early…
16 May 2026
Arxiv cracks down on unchecked AI-generated content in research papers
Source Read original →
16 May 2026

