September 2026 - Page 10 of 29

ssd failure prediction multiple instance learning a duffel bag lying flat with carry handle

New AI Method Improves SSD Failure Prediction From Imperfect Maintenance Data

Data centres report whole racks as failed rather than isolating one drive, so failure-prediction models train on labels that are systematically wrong. A team at SEOULTECH, working with Samsung Electronics on real Alibaba Cloud drive data, grouped drives into failure bags and let a temporal convolutional network score them individually – holding F1 at 0.717 where a conventional model collapsed to 0.261.

Read more
ai refusal how ai models decide not to answer a round sieve lying flat with side handle

How AI Models Decide Not to Answer a Question

A model declining a question is not one behaviour but two: a policy judgement about harm and a calibration judgement about uncertainty. We take both apart using OpenAI’s Model Spec, Claude’s constitution, the GPT-5 system card, the safe-completions paper, AbstentionBench, XSTest and OR-Bench – including why over-refusal happens and what a builder can actually change.

Read more
Apodex 1.1 mini - apodex 1 1 mini frontieragent framework release a relay baton lying flat on two rests

Apodex Releases the FrontierAgent Framework and the 35B Apodex 1.1 mini

Apodex released a 35B open-weight sibling to its 397B flagship, plus FrontierAgent, an Apache 2.0 agent runtime with ReAct and Agent Team modes. We work from the Hugging Face config, the GitHub repository and the arXiv report rather than the announcement — including the real parameter count, which benchmark scores belong to which model, and why the quantised builds are outrunning the originals.

Read more
deep think mathematica google internal math model a slide rule lying flat with centre slider

Google Tests a New Deep Think Mathematica Model Internally

A leaked Gemini API descriptor names an internal Google build called Mathematica (DeepThinkV3) Teamfood Raw Thoughts, with a one-million-token window, an UNSTABLE_EXPERIMENTAL stage and unsummarised reasoning. We decode the identifier field by field, identify the open number-theory problem in the leaked trace, compare it with the public Gemini 3.1 Deep Think, and separate what the screenshots show from what has been read into them.

Read more
gemini notebook voice mode ultra subscribers a jukebox cabinet arched top

Realtime Voice Mode Rolling Out to Gemini Notebook Ultra Subscribers

Google is rolling out real-time voice conversations in Gemini Notebook, starting with Google AI Ultra subscribers on mobile. We set out what the feature does, the four restrictions buried in the announcement, the Gemini 3.8 Live model it almost certainly runs on, the four other study features shipped alongside it, and whether the feature justifies an Ultra subscription.

Read more
CHAT