❌

Reading view

OpenAI expands Codex and its API at DevDay with security scans, a Decisions API, and Ultrafast

Smartphone view of Codex showing the "All," "Cloud," and "Jay's MacBook Pro" filters above a project list that includes "Landing page" and "Sales dashboard," with a purple cloud icon featuring a terminal symbol next to it.

At DevDay 2026, OpenAI gave Codex reusable cloud environments, automatic security scans for GitHub repositories, and a code review view in the desktop app. The Agents API now supports Computer Use, and a new Decisions API handles fast, single decisions. The Ultrafast premium tier promises up to eight times the speed at six times the price.

The article OpenAI expands Codex and its API at DevDay with security scans, a Decisions API, and Ultrafast appeared first on The Decoder.

  •  

ElevenLabs' new v4 speech model makes AI voices more expressive and consistent

"Eleven V4" lettering in white against a blurred, diagonal color gradient of orange, pink, and green.

Elevenlabs' new speech model, Eleven v4, follows cues for laughter and whispering more accurately and keeps voices consistent across long productions like audiobooks. Its Turbo variant starts speaking in 150 milliseconds and is built for real-time voice agents. On Artificial Analysis' Voice Arena leaderboard, v4 ranks ahead of Cartesia and Google's Gemini.

The article ElevenLabs' new v4 speech model makes AI voices more expressive and consistent appeared first on The Decoder.

  •  

Manus 2.0 lets users edit videos, host multiplayer games, and run agents remotely from their phone

Manus is turning its AI agent into a platform with version 2.0, letting users edit videos, host multiplayer games, and run personal agents with their own phone numbers.

The article Manus 2.0 lets users edit videos, host multiplayer games, and run agents remotely from their phone appeared first on The Decoder.

  •  

AI agents do more of the work in model development, but humans still make the decisions

Collage: Three people interact with colorful tree branches as a symbol of collaborative research and idea development.

A research team analyzed 769 task logs from building its own AI model. AI agents supplied up to 55 percent of method proposals, but humans made more than 85 percent of final decisions. A third of the tasks wouldn't have been attempted without AI. The authors warn that more agent activity doesn't mean more autonomy.

The article AI agents do more of the work in model development, but humans still make the decisions appeared first on The Decoder.

  •  

Nvidia drops a free 100M-parameter model that identifies up to eight speakers in real time

Nvidia released Nemotron 3 Diarization, an AI model that identifies which speaker is talking at any given moment in a conversation.

The article Nvidia drops a free 100M-parameter model that identifies up to eight speakers in real time appeared first on The Decoder.

  •  

Nvidia's SoL-Pi system cuts coding agent token usage nearly in half by optimizing the harness

Stylized illustration of modular AI modules with intertwined cables and a glowing data stream

SoL-Pi cuts coding agents' token usage by up to 49 percent with little change in performance by optimizing the control layer between the model and its environment. A research agent tested 152 approaches across more than 3,000 runs to develop the system, though the gains were smaller on other benchmarks.

The article Nvidia's SoL-Pi system cuts coding agent token usage nearly in half by optimizing the harness appeared first on The Decoder.

  •  

Tencent's Gander aims to keep talking while it works in the background

Tencent's Gander processes speech, images, and text while handling tasks in the background. A "cerebellum" keeps the conversation going, while a swappable "brain" searches files, writes code, or tackles other complex work. Users can interrupt or change the task mid-conversation. In benchmarks, Gander interrupted users in just 8 percent of cases, less often than its rivals, but trailed on task accuracy.

The article Tencent's Gander aims to keep talking while it works in the background appeared first on The Decoder.

  •  

Runway wants to turn AI video generation into a live stream you control in real time

Runway wants to stream AI video as users prompt it, rather than make them wait for finished clips. The approach builds on GWM-1, its world model that generates video frame by frame. Beyond creative tools, Runway sees uses in robotics and autonomous driving.

The article Runway wants to turn AI video generation into a live stream you control in real time appeared first on The Decoder.

πŸ’Ύ

πŸ’Ύ

πŸ’Ύ

  •  

Simulated students that make realistic mistakes help AI tutors learn faster

Illustration: Four people are solving interactive tasks such as puzzles, mazes, and coding challenges as part of an AI-powered tutoring workflow.

Microsoft and the University of Illinois built StudentSim to replicate individual students from limited data and give AI tutors fast, low-cost feedback. In tests covering 60 students across chess, English, and math, it outperformed GPT-5.4. A chess tutor trained with StudentSim also earned the highest expert ratings among three versions tested.

The article Simulated students that make realistic mistakes help AI tutors learn faster appeared first on The Decoder.

  •  

Google Deepmind's Dream-RSI helps AI agents improve by β€œdreaming” about past attempts

An abstract digital collage of overlapping purple, green, and orange geometric shapes against a dark background.

Google and Deepmind's Dream-RSI lets AI agents "dream" through past search runs to test new strategies without costly recalculations. In tests, it matched or beat existing results, cutting iterations by a factor of up to 2.43. Only the search strategy adapts, while the underlying AI model stays unchanged.

The article Google Deepmind's Dream-RSI helps AI agents improve by β€œdreaming” about past attempts appeared first on The Decoder.

  •  

Iris-mini and Iris-pro are the strongest open-weight search agents in their class

Colorful browser windows and speech bubbles, connected by arrows to a glowing network, symbolize AI search agents.

The AllSpark team has released Iris-mini and Iris-pro, two open-source search agents built on Qwen models that lead benchmarks among open-weight models in their size classes. According to the paper, the training data and models also improved performance on tasks they were never trained for, including general tool use and office work.

The article Iris-mini and Iris-pro are the strongest open-weight search agents in their class appeared first on The Decoder.

  •  

Two-year university study finds banning AI from classrooms leaves students worse off

Law students with stacks of files on their desks use a holographic AI dashboard to analyze legal data.

A law professor spent two years testing how an AI ban, unguided AI use, and structured training affect student performance. The group without AI finished last both years. "I was wrong," the researcher writes, who had assumed that AI without guidance would do more harm than good.

The article Two-year university study finds banning AI from classrooms leaves students worse off appeared first on The Decoder.

  •  

AI models' written reasoning steps correspond to distinct internal patterns, a new study finds

A multi-stage pipeline made up of colorful geometric shapes and arrows symbolizes structured information processing.

Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model's internal states, especially in the middle layers. That matters for AI safety, because models process more than their visible chain of thought reveals.

The article AI models' written reasoning steps correspond to distinct internal patterns, a new study finds appeared first on The Decoder.

  •  

Google's new AI model predicts the future from sales data, weather, and discount schedules

Melting ice flows toward shops, calendars, and discount icons, symbolizing rising sales during the summer season.

Google Research has released TimesFM-3, a forecasting model that analyzes time series alongside related data and known future events like sales promotions or weather forecasts. Instead of predicting the future step by step, the 330-million-parameter model fills in all future time points in a single pass, which cuts compute time and reduces compounding errors.

The article Google's new AI model predicts the future from sales data, weather, and discount schedules appeared first on The Decoder.

  •  

New Deepseek model V4.1-Flash cuts memory needs for AI agents

Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents.

The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder.

  •  

Muse can shop, write emails, and negotiate prices for users, all through WhatsApp

A blue Muse logo with a handwritten look, surrounded by speech bubbles containing tasks such as "Book it please!" and "I analyzed your new expenses and adjusted your budget"; on the right is the beige Muse avatar.

Meta unveils Muse, an AI agent that books travel, handles purchases, and sends emails through WhatsApp, complete with a payment feature that runs through Stripe's Link. That puts Meta ahead of OpenAI, which stopped its direct checkout feature in ChatGPT. A separate security agent called Sentinel monitors every action before it reaches the internet.

The article Muse can shop, write emails, and negotiate prices for users, all through WhatsApp appeared first on The Decoder.

πŸ’Ύ

πŸ’Ύ

  •  

Suno launches v6 music models built with Warner, BMG, and Believe

Feature image of the Suno v6 model with the "v6" lettering in bright orange and pink on a dark background; the Suno logo is in the upper left corner.

Suno has unveiled a new AI music model generation, v6, in three versions, built together with Warner Music Group, BMG, and Believe. All older models are being shut down. Songs can now be partially changed through text commands or generated multimodally from text, audio, and images. The company won't say which catalogs went into training, while Universal and Sony keep suing.

The article Suno launches v6 music models built with Warner, BMG, and Believe appeared first on The Decoder.

  •  
❌