Chips
40 articles RSS
Cerebras to Supply About 100 Megawatts of CS-4 Systems to Gimlet Labs for a Mixed-Silicon Inference Cloud
Cerebras will supply roughly 100 megawatts of CS-4 systems to inference-cloud startup Gimlet Labs, which pairs wafer-scale chips with GPUs and targets up to 3,000 tokens per second.
Microsoft Details Maia 200 AI Accelerator Architecture at Hot Chips 2026
At Hot Chips 2026, Microsoft and an accompanying arXiv paper detailed Maia 200's software-defined dataflow architecture, a 750-watt inference chip delivering 10,145 Tflop/s of FP4 compute.
Anthropic Held, Then Abandoned, a $7 Billion Bid for AI Chip Startup MatX, Reuters Reports
Anthropic discussed buying AI chip startup MatX for roughly $7 billion, then walked away; the two sides are now said to be exploring a supply partnership instead.
Apple's M6 and M5 Ultra Chips Debut a New Core AI Framework for On-Device Model Training
Apple's new M6 and M5 Ultra chips ship alongside Core AI, a new framework for building and deploying AI models on Apple silicon, with M5 Ultra supporting 512GB of unified memory.
NVIDIA Groq 3 LPX Inference Chip Enters Full Production, Claiming 4x Faster Response for Coding Agents
NVIDIA's Groq-derived Groq 3 LPX inference accelerator is now shipping, promising ultrafast token generation for agentic coding workloads, with Nebius first to deploy it.
Fractile Seeks $6.5 Billion Valuation After $250 Million Anthropic Chip Deal, Six Times Its May Price
UK inference-chip startup Fractile is in talks to raise about $600M at a $6.5B pre-money valuation, driven by an initial $250M chip deal with Anthropic.
Cerebras Launches CS-4 AI Accelerator, Claiming 30x Faster Inference Than GPUs on an Overclocked WSE-3
Cerebras unveiled its CS-4 rack-scale inference system, claiming 30x faster performance than GPUs, though independent analysis finds the chip inside is an overclocked WSE-3, not a new design.
AMD Acquires Taalas, Betting Model-Specific Chips That Etch Weights Into Silicon Can Boost AI Inference Speed
AMD is buying Toronto-based Taalas, whose chips etch model weights directly into silicon instead of storing them in memory, to strengthen its AI inference lineup.
Etched Exits Stealth With Working Sohu Transformer ASIC, $800 Million Raised at a $5 Billion Valuation and Over $1 Billion in Chip Orders
Etched unveiled working A0 silicon for Sohu, its transformer-only inference ASIC, disclosing $800 million raised at a $5 billion valuation and more than $1 billion in signed customer contracts.
OpenAI and Broadcom Unveil 'Jalapeño,' a Custom LLM-Inference Chip Designed in Nine Months and Targeting Deployment by the End of 2026
OpenAI's first custom silicon, the Broadcom-built Jalapeño inference accelerator, went from design to tape-out in nine months and targets deployment by the end of 2026.
Tensordyne Sends Its Logarithmic-Math AI Chip Napier to TSMC, Claiming Four Times the Speed at a Fifth the Power of an Nvidia GB300 System
The startup says its 3nm Napier inference chip turns matrix multipliers into adders, with commercial sales of a 72-chip system slated for the second half of 2027.
Qualcomm Announces Snapdragon C, Pushing Windows on Arm and an On-Chip NPU Into $300 Budget Laptops
Qualcomm unveiled the Snapdragon C Platform for entry-tier Windows laptops starting at $300, with Acer, HP, and Lenovo as launch partners and devices expected later this year.