Topic

Infrastructure — AI news & analysis

AI-generated content. Everything on this page was written by an automated AI editorial system and published without prior human review. How this works ›

Every Pick Right story tagged infrastructure — 8 articles, newest first. All news →

nvidia hugging-face

Nvidia is reportedly buying the shelf your open-weight models sit on — eight days after Stripe bought the router

The Information reported on 26 August 2026 that Nvidia agreed to acquire Hugging Face for $12.9B. Business Insider says the talks are unresolved. Neither company has confirmed anything. What is not in dispute is the shape: in eight days, both layers buyers picked because they were vendor-neutral — the router and the hub — got bids from parties with a direct stake in what runs through them. Here is what to check in your build pipeline this week, and why the deal being unsigned is the reason to act now rather than later.

Read story →
openai nvidia

OpenAI's Jalapeño chip posts its first numbers — and it still can't escape the memory squeeze

OpenAI published Jalapeño's first benchmark results on 25 August 2026: 1.5-1.9x more throughput per kilowatt and up to 3.6x lower latency than Nvidia's GB200 and GB300, from a 700W package against their 1,200-1,400W. SemiAnalysis ran its InferenceX suite at OpenAI's labs, but OpenAI supplied every number. The chip is not for sale, deploys in volume only in 2027, and carries six stacks of exactly the HBM4 that is driving the industry's cost increase. Here is what it changes for a buyer, and what it does not.

Read story →
nvidia pricing

Nvidia's 15% server price rise puts a floor under the AI price war — and the discounts expire first

Bloomberg reported on 22 August 2026 that contract server builders have told Nvidia's largest customers that Grace Blackwell and Vera Rubin systems shipping in early 2027 will cost more than 15% extra, driven by DRAM, LPDDR and HBM4 shortages. Every headline AI price cut of the last month was set when memory was cheap, and the promotional clocks — Sol to 21 November, Gemini 3.7 Flash to 31 December — run out right where the hardware cost increase begins. Here is what that does and does not mean for your 2027 budget.

Read story →
anthropic reliability

Claude broke again on Monday — and Anthropic no longer sells a tier that promises it won't

Anthropic's status page logged 21 incidents between 1 and 24 August 2026, including one critical and six major. Monday's three-hour outage hit Mythos 5, Fable 5, Opus 5 and Opus 4.8. The buyer problem isn't the incident count — OpenAI's is comparable. It's that Priority Tier, the only Anthropic tier with a published uptime target, is now marked 'no longer available for purchase' — and it never covered Mythos 5, Opus 5 or Sonnet 5 anyway.

Read story →
openrouter stripe

Stripe bought the neutral layer for $7.5B — and Ramp gave it away free the same day

On 19 August 2026 Stripe confirmed it is acquiring OpenRouter, the model gateway that routes across 400+ models from 80+ providers. Hours later Ramp launched a competing router, free through 2026. Two events, one lesson: the 'neutral' routing layer was never neutral infrastructure. It is a seat next to your token spend, and it just became the most contested seat in the stack.

Read story →
anthropic google

Anthropic's compute is being bought with somebody else's money — and kept off its balance sheet before an IPO

Google has assembled a chip-financing programme worth more than $150bn to supply Anthropic with TPUs. The hardware sits in a special-purpose vehicle, Broadcom guarantees roughly $31bn of the senior debt, Morgan Stanley occupies three roles at once, and bitcoin miners provide the buildings. S&P has already downgraded Broadcom over it.

Read story →
openai infrastructure

OpenAI unveils Jalapeño, its first custom AI chip — built with Broadcom for LLM inference, deploying by end of 2026

OpenAI and Broadcom announced Jalapeño on June 24, 2026 — OpenAI's first custom 'Intelligence Processor,' an accelerator designed from the ground up for LLM inference. Co-developed with Broadcom and Celestica, it went from design to tape-out in nine months (reportedly the fastest ASIC cycle ever), with OpenAI's own models used to speed development. OpenAI claims performance-per-watt 'substantially better' than current state of the art. Initial deployment is end of 2026 at gigawatt scale with Microsoft and other partners. OpenAI joins Google, Amazon, Microsoft, and Anthropic in building its own silicon — here's what it means for the AI tools you use.

Read story →
anthropic claude

Anthropic locks up SpaceX's Colossus 1, doubles Claude Code limits, and ships a 10-agent Wall Street pack

On May 6, 2026, Anthropic announced a SpaceX compute deal taking all of the Colossus 1 data center (300+ MW, 220,000+ NVIDIA GPUs), doubled Claude Code rate limits across Pro/Max/Team/Enterprise, removed peak-hour throttling for Pro and Max, and the same week shipped 10 finance agent templates with Microsoft 365 GA integration and named JPMorgan, Goldman Sachs, and Citi as live customers. The capacity story, the user story, and the enterprise story landed together — and they're connected.

Read story →