Topic
Hardware — AI news & analysis
AI-generated content. Everything on this page was written by an automated AI editorial system and published without prior human review. How this works ›
Every Pick Right story tagged hardware — 4 articles, newest first. All news →
OpenAI is serving its biggest model at 750 tokens a second on Cerebras — speed is now the third axis of the AI race
On 13 August 2026 Cerebras announced it powers a new OpenAI 'Ultrafast' tier that runs GPT-5.6 Sol at up to 750 output tokens per second — about 14× faster than Standard. It ships as a limited preview with no price, no SLA and no region list. Here is what wafer-scale inference actually changes for agents and coding, and why speed — not just intelligence or price — is becoming the axis buyers optimise next.
Read story →Anthropic's 2-gigawatt AMD deal is the biggest crack yet in NVIDIA's monopoly — and AMD is paying to make it happen
AMD and Anthropic announced a strategic partnership on July 22: Anthropic will deploy up to 2 gigawatts of AMD Instinct MI450-series GPUs in Helios rack-scale systems, with the first gigawatt landing in H1 2027 — and AMD will make a strategic equity investment of up to $5 billion in Anthropic. There's also a reciprocal engineering deal where Claude tunes workloads for AMD's own GPUs. Here's why it matters, the circular-financing pattern worth noticing, and what it means for the prices and rate limits you actually pay.
Read story →Google's 'Frozen v2' would etch Gemini's architecture into silicon — a 6–10× efficiency bet that model design has stopped moving
Google is reportedly developing an AI inference chip codenamed Frozen v2 that hardwires Gemini's architecture — the blueprint, not the weights — directly into silicon. Engineers project 6 to 10 times more tokens per watt than Google's latest TPUs, with deployment targeted for 2028. Alphabet stock moved on the report. Here's what's reported versus confirmed, why the design is a bet that model architecture has stopped changing, and what it actually means for what you'll pay for inference.
Read story →NVIDIA hand-delivers first Vera CPUs to Anthropic, OpenAI, SpaceX, Oracle — first standalone NVIDIA CPU lands at frontier labs
NVIDIA's Ian Buck hand-delivered the first Vera CPU systems to Anthropic (San Francisco), OpenAI (Mission Bay), SpaceX (Palo Alto), and Oracle Cloud Infrastructure (Santa Clara) at the end of May 2026. Vera is NVIDIA's first standalone data-center CPU, purpose-built for agentic AI workloads, competing directly with Intel Xeon and AMD EPYC. In full production since March 2026. The frontier-lab customer roster maps NVIDIA's $100B-class strategic compute alliances for the agentic-AI era.
Read story →