Topic

Model-launch — AI news & analysis

AI-generated content. Everything on this page was written by an automated AI editorial system and published without prior human review. How this works ›

Every Pick Right story tagged model-launch — 7 articles, newest first. All news →

open-weights china

Two open-weight 'Flash' models landed on the same day — and the better one spent the previous week in your router with no name on it

On 26 August 2026 Alibaba shipped Qwen3.8-Flash-Next (125B total, 6B active, Apache 2.0, $0.16/$0.47) and Z.ai shipped GLM-5.3-Flash (320B total, 18B active, MIT, $0.15/$0.50). Both claim coding scores at or above 2026 flagships at roughly a tenth of flagship price. The pricing is the smaller story. GLM-5.3-Flash is 'ox-alpha' — the unnamed free model that became OpenRouter's most popular listing of the week, retained every prompt, and ran entirely on Chinese domestic silicon. Here is what actually changed for a buyer, and what to check in your router config today.

Read story →
google gemini

Google shipped Gemini 3.7 Flash — a cheaper, faster coding workhorse — while the flagship it actually promised is still missing

On 13 August 2026 Google launched Gemini 3.7 Flash, its third Flash model in seven weeks, with big coding gains (DeepSWE 49%→65.3%) and an introductory price of $0.75/$3.75 per million tokens — half of 3.6 Flash. It is a genuinely strong mid-tier coding-agent model and a clear price-war move. It is also, conspicuously, not the Gemini 3.5 Pro flagship Google promised in May and has now failed to ship for three months, in the middle of a leadership reshuffle. Here is what 3.7 Flash actually delivers, what it costs after the intro window, and why the model Google keeps shipping is not the one that matters most.

Read story →
openai security

OpenAI built a model that writes exploits — and the interesting part is who is allowed to use it

GPT-5.6-Cyber completes 95% of exploit-chain, privilege-escalation and authentication-bypass requests, against 1.5% for the standard GPT-5.6 Sol it is built on. OpenAI is not selling it. Access runs through a vetted partner tier, Daybreak Red, and the model has already found a chainable zero-day in Chrome's V8 engine. The refusal rate was not a bug being fixed — it was a product decision.

Read story →
google gemini

Google I/O 2026 recap — Gemini 3.5 Flash ships; Omni, Spark, and a $100 AI Ultra tier

Google I/O 2026 keynote (May 19) shipped Gemini 3.5 Flash to production, previewed Gemini 3.5 Pro for next month, launched the Gemini Omni video model (image + audio + video + text input → editable video output), introduced the Gemini Spark personal agent, restructured Google AI Ultra to $100/month (with the old $250 tier dropping to $200), and confirmed Android XR audio glasses for fall 2026. Here's the verified recap for Pick Right readers.

Read story →
xai coding

xAI enters the coding agent race — Grok Build ships in early beta with 8 parallel agents and Arena Mode

xAI dropped Grok Build, its first CLI coding agent, in early beta during early May 2026. The product runs up to 8 parallel sub-agents simultaneously, ships an automated 'Arena Mode' that scores competing outputs, and runs local-first (code never leaves the developer's machine). The underlying model — grok-code-fast-1 — scores 70.8% on SWE-Bench Verified at $0.20 input / $1.50 output per million tokens. Here's where Grok Build fits in the Claude Code / Codex / Cursor landscape, and where it falls short.

Read story →
openai voice

OpenAI ships three new realtime voice models — GA, GPT-5-class reasoning, 70-language translation

On May 7, 2026, OpenAI took the Realtime API out of beta and launched three new voice models: GPT-Realtime-2 (the first voice model with GPT-5-class reasoning), GPT-Realtime-Translate (70 input languages → 13 output languages, live), and GPT-Realtime-Whisper (streaming speech-to-text). The release puts voice-AI on a 'listen-reason-translate-act' arc that affects ElevenLabs, the call-center category, and any product that handles spoken-language input.

Read story →
openai chatgpt

GPT-5.5 Instant becomes ChatGPT's default model — fewer hallucinations, smarter web routing

On May 5, 2026, OpenAI made GPT-5.5 Instant the default model for ChatGPT, retiring GPT-5.3 Instant. The new Instant model lifts AIME 2025 math from 65.4 to 81.2, MMMU-Pro multimodal from 69.2 to 76, and reduces hallucination on sensitive prompts (medicine, law, finance) — without changing the price for end users. The rollout is web-Plus/Pro first, with Free, Go Business, and Enterprise following over the coming weeks.

Read story →