AI & Machine Learning
177 articles RSS
Liquid AI's 230M-Parameter LFM2.5 Beats Models Four Times Its Size at Data Extraction and Runs on a Raspberry Pi
The MIT spinout's smallest model yet pairs convolution and attention blocks to outscore billion-parameter rivals on data extraction while decoding at 42 tokens per second on a Raspberry Pi 5.
xAI's Grok 4.3 Arrives on Amazon Bedrock, Running on a New 'Mantle' Inference Engine Reached Through an OpenAI-Compatible Endpoint
AWS added Grok 4.3 to Bedrock on June 15, making xAI a model provider on the platform and debuting a separate Mantle inference path.
Mistral Renames Le Chat to Vibe, Folding Chat, Office Automation, and Cloud Coding Into One Agent
Mistral rebranded its Le Chat assistant as Vibe, a single agent spanning Work Mode for office tasks and Code Mode for remote coding.
OpenAI and Molecule.one Report a Near-Autonomous AI Chemist That Improved a Stubborn Drug-Making Reaction Across 10,080 Experiments
GPT-5.4 paired with Molecule.one's Maria platform proposed, ran, and analyzed a 10,080-reaction campaign that lifted yields of a hard sulfonamide coupling using the additive TEMPO.
OpenAI's 'Deployment Simulation' Replays 1.3 Million Past Conversations Through Candidate Models to Predict Misbehavior Before Release
OpenAI's pre-release method regenerates real past conversations with an unreleased model to estimate undesired-behavior rates, with a median multiplicative error of 1.5x.
Zhipu Releases GLM-5.2, a 744-Billion-Parameter Open Coding Model, Days After Washington Cut Off Foreign Access to Anthropic's Claude
Zhipu's GLM-5.2 ships under an MIT license as a frontier coding alternative just after the US barred foreign nationals from Anthropic's Fable 5 and Mythos 5.
MIT's Recursive Language Models Let an LLM Read Its Own Prompt as Code, Beating Frontier Long-Context Scaffolds
A new MIT CSAIL inference method has a model inspect its prompt in a Python REPL and recursively call itself over snippets, processing inputs beyond its context window.
DeepMind's AlphaProof Nexus Solves 9 Open Erdős Problems by Pairing Gemini With the Lean Proof Checker
A DeepMind preprint reports an LLM-and-Lean agent that autonomously solved 9 of 353 open Erdős problems and proved 44 of 492 OEIS conjectures for a few hundred dollars each.
Moonshot AI Open-Sources Kimi K2.7-Code, a Trillion-Parameter Coding Model That Cuts Reasoning Tokens by 30%
Moonshot AI released Kimi K2.7-Code, an open-weight 1-trillion-parameter MoE coding model under a Modified MIT License, claiming roughly 30% lower thinking-token usage than K2.6 on self-run benchmarks.
Sanofi Deepens Owkin Partnership With Five-Year Deal to Build Agentic 'Biopharma Agents' for Drug Development
Sanofi will license Owkin's K Pro platform for five years and co-develop autonomous AI agents for drug R&D, extending a partnership that began in 2021 and made Owkin a unicorn.
Anthropic Releases Claude Fable 5, Its First Public Mythos-Class Model, With Sensitive Queries Routed to Opus 4.8
Anthropic launched Claude Fable 5 on June 9, a public version of its restricted Mythos model that falls back to Opus 4.8 on high-risk topics in under 5% of sessions.
Microsoft Launches Seven In-House MAI Models, Built From Scratch Without Distillation to Cut OpenAI Reliance
Microsoft unveiled seven proprietary MAI models led by the 35B-active reasoning model MAI-Thinking-1, all trained from scratch without distillation as it reduces its dependence on OpenAI.