AI & Machine Learning
177 articles RSS
CDT Study Catalogs 37 'Dark Patterns' Across AI Chatbots, From ChatGPT and Claude to Replika and Character.AI
A Center for Democracy & Technology study built a taxonomy of 37 manipulative design patterns it found across major AI chatbots and companion apps.
Linux Foundation Launches Tokenomics Foundation to Standardize AI Cost Management, Extending FinOps' FOCUS Spec to Token-Based Spending
The vendor-neutral foundation will partner with FinOps to extend the FOCUS billing spec to AI token spend, backed by 12 enterprises.
MIT and Harvard Teach Language Models to Ask Better Questions, Lifting a Small Model's Battleship Win Rate From 8% to 82%
An ICLR paper from MIT CSAIL and Harvard shows Monte Carlo inference helps Llama 4 Scout outpace GPT-5 at a Battleship test bed for around 1% of its cost.
NVIDIA Open-Sources Nemotron 3 Ultra, a 550B Mamba-Transformer Mixture-of-Experts Built for Long-Running Agents
The 550-billion-parameter model activates 55 billion parameters per token, ships under the Linux Foundation's OpenMDW-1.1 license, and trades blows with China's Kimi-K2.6 on coding benchmarks.
OpenAI Pushes Codex Beyond Code Into Finance and Legal, Squaring Off Against Anthropic's Claude for Legal
OpenAI is extending Codex into finance and legal work, weeks after Anthropic expanded Claude for Legal with 12 plugins and 20-plus integrations.
Google Releases Gemma 4 12B, an Encoder-Free Multimodal Model With Native Audio That Runs on a 16GB Laptop
Google's new 12-billion-parameter open model drops separate vision and audio encoders, projecting raw image patches and audio waveforms straight into the LLM, and ships under Apache 2.0.
Microsoft Unveils Project Solara, an Android-Based Platform for 'Agent-First' Devices, With a Wearable AI Badge
At Build 2026, Microsoft revealed Project Solara, a chip-to-cloud platform built on AOSP for devices that run AI agents instead of apps, including a reference-design wearable badge.
Raindrop Open-Sources Workshop, a Local MIT-Licensed Debugger That Lets Coding Agents Write and Run Their Own Agent Evals
Raindrop released Workshop, a free local debugger that streams an AI agent's tokens, tool calls, and spans to a browser and lets Claude Code write and fix evals against the trace.
MiniMax Releases M3, an Open-Weight Model With a 1-Million-Token Context That It Says Tops GPT-5.5 on SWE-Bench Pro
Shanghai-based MiniMax launched M3 on June 1, pairing a 1-million-token context with a new sparse-attention design and company benchmarks that top GPT-5.5, with weights promised within 10 days.
Microsoft Build 2026 Bets on Windows as an Agent Platform, Unveils Project Polaris and Azure Agent Mesh
At Build 2026 in San Francisco, Microsoft unveiled Project Polaris to replace GPT-4 Turbo in GitHub Copilot, open-sourced the Windows Agent Framework, and previewed the Windows Agent Runtime and Azure Agent Mesh.
Cognition Raises $1 Billion at $26 Billion Valuation as Devin AI Engineer Hits $492 Million in Annualized Revenue
The maker of Devin, an autonomous AI software engineer, closed a Series D round led by Lux Capital, General Catalyst, and 8VC, more than doubling its valuation from $10.2 billion eight months earlier.
AWS Rebuilds Amazon OpenSearch Serverless From the Ground Up for Agentic AI, Reaching GA With 20x Faster Scaling and Up to 60% Lower Cost
AWS launched the next-generation OpenSearch Serverless on May 28, decoupling compute from storage to hit true scale-to-zero and 20x faster autoscaling for bursty agentic workloads.