AI — Page 13
Models, agents, infra, applied AI.
- Hidden ChatGPT modal previews new ImageGen 2.5 model
The unreleased interface promises higher-quality results and smarter tools, while OpenAI's public documentation still lists Images 2.0.
- OpenAI prices GPT-6 Astra at Claude Fable 5.1's $10/$50 rate
The new model costs 2.5x more per token than GPT-5.6 Sol's current promotional rate, putting the premium on autonomous work.
- Codex leak labels OpenAI's Astra as GPT-6 ahead of release
Two model IDs appeared in a Codex Desktop feature flag two days after OpenAI said Astra would be available soon.
- Osaurus ships a 6.3 GB agent model for 8GB and 16GB Macs
Raptor 0.5 activates about 1B of 7.9B parameters and targets file, spreadsheet and email tasks on Apple Silicon.
- Browser Use gives web agents single-use cards for checkout
The Link integration lets users approve purchases without exposing their underlying card details to the browser agent.
- Google DeepMind plans to bring WeatherNext 3 forecasts to Search, Maps and Gemini
The joint Google DeepMind and Google Research project is moving into consumer products, developer APIs and cloud datasets, where company-reported benchmark gains will meet real weather.
- Google rolls Gemini voice controls into Gmail, Docs and Keep
Gmail Live and Keep Live reach Plus, Pro and Ultra plans; Docs Live requires Pro or Ultra, with business access promised later.
- IFM releases six K2 Horizon models with training records for outsiders to inspect
Eric Xing's MBZUAI institute published models from 0.9B to 375B parameters, along with code, data recipes, checkpoints and benchmark caveats.
- Ansh Chokshi and Shashwat Kapoor build Mireye's location layer for AI agents
Mireye's API and MCP server return cited parcel, hazard, utility and site-selection data across the US.
- OpenAI teases Astra with MIT's 1979 voice-and-gesture interface
The clip points toward a multimodal launch two days after OpenAI said its delayed Astra model would arrive soon.
- Anthropic cuts Fable 5.1 cache-read prices 75% for long-running agents
Input remains $10 and output $50 per million tokens, while Anthropic estimates the cache discount can lower highly agentic workload costs by roughly 45%.
- OpenAI investigates ChatGPT and Codex outage across nearly every major feature
Errors hit conversations, login, Search, uploads, Voice, Agent and Codex; OpenAI had not published a cause.
- MiniMax M3 powers HUMAIN's 428B-parameter Arabic model
HUMAIN further trained MiniMax M3 on more than 1 trillion Arabic tokens and put the 428-billion-parameter result into a limited API preview.
- Flock Safety's AI watchlists let police search for people by description
A reconstructed interface shows map-based person searches, image-ranking feedback, and moderation controls that can block or warn on sensitive queries.
- Botika runs a 100-terabyte fashion AI stack on Modal, minus Kubernetes
Eran Dagan says moving data, training and 15 production models onto Modal let Botika scale research without doubling its infrastructure staff.
- NTU researchers release Puffin-World with camera-grounded 3D states
According to the [Hugging Face model card](https://huggingface.co/ACERobotics/Puffin-World?ref=runtimewire), the NTU-led model generates camera-controlled RGB and depth views that can be consolidated into 3D point clouds.
- Wayve launches London's first robotaxis, with a licensed driver still aboard
Alex Kendall's driving AI is entering public service through Uber with fewer than 20 Mustang Mach-Es and regulatory limits on driverless operation.
- Nous Research adds persistent multi-gateway control to Hermes Desktop
The v0.21.0 registry keeps local, cloud, remote and SSH agents visible while background jobs continue streaming.
- Meta launches Muse Voice Transcribe with diarization at $0.18 an hour
The model handles 20-plus speakers and code-switching, while Speechmatics and AWS publish higher speaker ceilings.
- Alexandr Wang's Meta team ships Muse Spark 1.3 for coding agents
The fourth Muse Spark release since April targets coding and agentic tasks as Meta expands its model distribution.
- Trump DOJ backs OpenAI in New York Times copyright case
The government says licensing requirements would protect large publishers, entrench rich AI labs and weaken U.S. competition.
- Gemini 3.8 Flash goes live as RuntimeWire begins head-to-head testing
RuntimeWire is testing Google's callable model against the Fable 5.1 model, while Google's public catalog still stops at Gemini 3.7 Flash.
- Charlie Ruan leads WebLLM, which runs open models locally in your browser
Charlie F. Ruan leads an open-source runtime that uses WebGPU, WebAssembly and compiler-generated artifacts to run language models on users' devices.
- OpenAI's policy chief says frontier lab insiders are worried about AI, too
Dean Ball, hired in July to shape frontier policy, says he will keep issuing candid warnings even when they complicate OpenAI's message.