OpenRouter

workbuddy hy4 preview work and coding tasks a solid briefcase

Hy4 preview Is Now Available on WorkBuddy for Work and Coding Tasks

Tencent’s 770B open-weight Hy4 preview is now selectable inside WorkBuddy, the company’s desktop AI agent for office and coding work, and it is free there until 10 September 2026. This piece covers what the integration changes for daily work, how the model switch works and what it is recommended for, the 163-expert blind test that Tencent ran inside WorkBuddy against GLM 5.3 and Kimi K3, the benchmarks that matter for work and coding tasks, the quota and reasoning-time caveats, the API and local GGUF routes for teams that would rather not use the app, and how WorkBuddy compares with Claude Cowork.

Read more
GLM-5.3-Flash - glm 5 3 flash open weight 320b model a solid lightning bolt

GLM-5.3-Flash: The 320B Open-Weight Model That Ran on Chinese Chips

Z.ai spent six days serving an anonymous model called Ox Alpha on OpenRouter, took nearly 20% of the platform’s weekly token share, and only then revealed it was GLM-5.3-Flash — a 320B mixture-of-experts model with 18B active parameters, a one-million-token context window and MIT-licensed weights. This piece works through the hybrid attention architecture, what the benchmark table supports and what it does not, what the API actually costs once the launch promotion ends, how credible the domestic-silicon claim is, and what any of it changes for a business choosing a model this quarter.

Read more
hy4 preview tencent open weight moe 1m context a solid dodecahedron

Tencent Releases Hy4 preview: An Open-Weight MoE Model With a 1M Context Window

Tencent open-sourced Hy4 preview on 28 August 2026 under Apache 2.0: a 770B Mixture-of-Experts model that activates just 49B parameters per token and reads a one-million-token context window. This breakdown covers the full architecture, from 256 routed experts per layer to Gated DeepSeek Sparse Attention and the built-in speculative decoding layer; every benchmark figure Tencent published, including the 163-expert blind evaluation against GLM 5.3 and Kimi K3; the API price list against GPT-5.6 Sol; the eight-GPU serving recipes; and the four caveats worth naming before any of it reaches production.

Read more
Z.ai - z ai lab behind ox alpha model a treasure chest closed lid

Surprise: Z.ai Is the AI Lab Behind the Mysterious Ox Alpha Model

On 26 August 2026 the mystery ended: Z.ai, the Beijing lab formerly known as Zhipu AI, confirmed that the anonymous Ox Alpha model topping OpenRouter and OpenCode was the newest iteration of its GLM series, and published the weights the same evening as GLM-5.3-Flash under an MIT licence. This breakdown covers what was confirmed and when, the architecture the model card revealed — 320 billion total parameters with just 18 billion active, hybrid sparse and linear attention, a 1,048,576-token context and forced reasoning that cannot be disabled — the published benchmark table showing 84.3 on Terminal Bench 2.1 against 85.0 for Claude Opus 4.8 and 87.4 for GPT-5.6 Terra, the viral 80 per cent DeepSWE claim that came from a 10-task subset and collapsed to 63.4 on the full 113-task run, the 44 trillion tokens and 503,000 users the stealth week generated, the tokenizer and error-code forensics that unmasked the lab before it spoke, the $0.15 and $0.50 per million token pricing, the company’s Hong Kong listing and US entity-list status, and a buyer’s checklist for deciding between the hosted API and self-hosted weights.

Read more
ox alpha mystery ai model a solid theatre mask

0x Alpha: The Mystery AI Model Everyone Is Testing for Free

Ox Alpha — often searched as 0x Alpha — appeared anonymously on OpenRouter on 20 August 2026: a free frontier-class reasoning model with a million-token context window, video input and a viral claim to beat paid rivals at coding. This guide covers the verified specification sheet, the 10-task benchmark caveats, the serving-layer forensics pointing to Zhipu AI, the four previous stealth models that were all claimed by Chinese labs, what the free week really costs in data terms on each of the two access routes, and a defensive pilot plan for businesses that want the free market intelligence without handing an anonymous operator their company data.

Read more
stripe openrouter acquisition ai gateway a funnel wide top narrow spout

Stripe Will Reportedly Acquire AI Gateway Startup OpenRouter for $7B+

Stripe OpenRouter reports moved from rumour to near-certainty on 16 August 2026, when Bloomberg said the payments company had finalised an agreement to buy the AI model gateway for more than $7 billion. If the figure holds, it is the largest acquisition Stripe has ever made, and it values a company that raised money at […]

Read more
CHAT