AI models

Qwen3.8-Omni-Flash - qwen3 8 omni flash 1m token context window a satellite dish standing angled on one short post

Alibaba Releases Qwen3.8-Omni-Flash with 1M-Token Context Window

Alibaba released Qwen3.8-Omni-Flash on 18 September 2026: an omni-modal model built on the Qwen3.8-Flash-Next architecture that accepts text, images, audio and video into a 1 million token context window and returns text. The headline is price — more than 98% off an hour of audio input against the previous generation — but the mechanism is more interesting. Agentic perception lets the model decide what to watch and listen to, raising OmniVideoBench accuracy from 63.4 to 67.8 while cutting tokens per query by about 45.7%. This article covers the specification, the benchmark claims, the pricing arithmetic, the Apache-2.0 plugin suite and the limits.

Read more
grok 4 7 gcp quotas page spotted ahead of release a grok 4.7 tank cylinder with one tap spout v2

Grok 4.7 Spotted on GCP Quotas Page Ahead of Potential Release

A Google Cloud quota entry for grok-4.7 was reported on 17 September 2026, ten days after a dated identifier leaked through Cursor’s configuration. Neither xAI nor Google has announced anything, and both catalogues still stop at Grok 4.6. We set out what the two artifacts actually prove, why a quota row is the strongest pre-launch signal short of a model card, what the Kalshi and Polymarket books price it at, and how it sits against four public release windows that have already passed.

Read more
gemini 3 8 flash agent studio google cloud a gemini 3.8 flash console slider tracks

Gemini 3.8 Flash Available in Agent Studio on Google Cloud Platform

Gemini 3.8 Flash landed in Agent Studio on Google Cloud Platform on 2 September 2026 — Google’s third Flash release in six weeks. It scores 73.7% on DeepSWE v1.1, within 0.3 points of Claude Opus 5, at $0.75 per million input tokens until January. This article covers what the model changes inside Agent Studio, the benchmarks and pricing worth trusting, the restricted Flash Cyber security variant, and what a three-weekly release cadence means for teams building production agents.

Read more
GPT-5.6 Sol - gpt 5 6 sol a three ascending rounded pillars

GPT-5.6 Sol: Complete Guide to OpenAI’s Best Model Yet

GPT-5.6 Sol is OpenAI’s flagship model, launched publicly on 9 July 2026 alongside Terra and Luna. This guide maps the tier scheme, the full API price list including the 272K long-context surcharge, honest benchmarks against Anthropic’s Claude, and availability across ChatGPT plans and cloud platforms. It closes with what the Doug pre-training project and Astra signal about GPT-6.

Read more
forgetting featured

Powerful Forgetting May Be the Secret to Better AI Language Learning

The concept of forgetting in AI language learning is transforming how researchers design artificial intelligence systems. Forgetting in AI language learning refers to the deliberate removal or suppression of certain learned patterns, memories, or associations during the training process. While this may seem counterintuitive, recent research demonstrates that strategic forgetting can significantly improve how AI […]

Read more
CHAT