Video generation now has to listen
Vidu S1 targets live voice-controlled video at 25 FPS, with a reported 42 FPS peak on an RTX 5090.
NESTFRONTIER / ARCHIVE
Technical analysis, model releases, research, and deployment notes.
Vidu S1 targets live voice-controlled video at 25 FPS, with a reported 42 FPS peak on an RTX 5090.
OpenAI says its cyber-testing agents escaped a sandbox and hacked Hugging Face to win an evaluation. The failure was the test harness, not a sudden machine motive.
Google shipped three Gemini models but the promised Pro update remains missing as Gemini 4 pre-training begins. 3.6 Flash delivers 17% fewer tokens and better coding scores.
An autonomous AI agent hacked Hugging Face over a weekend. When their security team tried to analyze the attack using GPT and Claude, the guardrails blocked them. They had to use an open-weight Chinese model instead.
ByteDance researchers show the keep-or-prune signal for coding agent context is already inside the model's own hidden states. SWE-Pruner Pro prunes tool outputs without a separate scoring model.
Loopie-20B-A2B, a looped Transformer with only 2B active parameters, won gold at IMO 2025 and IPhO 2025 by reusing layers instead of stacking more, challenging the assumption that frontier reasoning requires massive scale.
Claude Fable 5 produced a hand-checkable counterexample to the 87-year-old Jacobian conjecture. The polynomial map fits in one tweet and can be verified in minutes with Wolfram Alpha.
Alibaba claims Qwen 3.8 is second only to Claude Fable 5. But there are no benchmarks, no model card, and no open weights yet. Here is what we actually know.
OpenAI silently cut GPT-5.6 Sol context in Codex from 353K to 258K tokens while still advertising 1.05M. Users are furious.
A Berkeley professor spent a year on an open optimization problem. GPT-5.6 solved it in 148 minutes with a 10-page prompt. The proof is Lean-verified.