AI — Page 12
Models, agents, infra, applied AI.
- OpenAI chief scientist seeks safety pact as lab scales agent research
Jakub Pachocki wants frontier labs to set shared triggers for slowing development, three days after GPT-6 Astra and following an OpenAI agent containment failure.
- Fal brings H3 Max reference-to-video to real time with four references
The GA endpoint accepts images, video and audio; Fal's API allows 12 files, though the company ties its real-time performance claim to four references.
- OpenAI launches GPT-6 Astra with a Blender-to-Unreal computer-use demo
The controlled demonstration shows the model building an editable 3D scene and moving it into Unreal Engine 5, while artist Stefan Vaskevich warns about the pace of improvement.
- Rootly's Sylvain Kalache warns AI could dull on-call skills
Sylvain Kalache's warning accompanies Rootly's partnership with Uptime Labs, which uses simulated outages and AI role-play to train incident responders.
- Reddit says Anthropic is calling it an AI rival to restrict discovery
Ahead of a September 10th hearing, Reddit says Anthropic wants tighter controls on its lawyers, experts, technical records and electronic evidence.
- Perplexity publishes its three-layer stack for search embeddings and ranking
Ivy prepares requests, Tulip batches them and ROSE runs models; Perplexity says the custom stack makes online and batch embeddings faster and cheaper, without disclosing production benchmarks.
- Ant Ling releases Ling-3.0-flash-VL for visual agents and GUI control
Ling-3.0-flash-VL reads images and short videos, writes websites from screenshots, and operates interfaces through Ant Ling's hosted API.
- NVIDIA says new optimizations make local agents up to 1.9x faster
The llama.cpp and vLLM updates target RTX 5090, RTX PRO 6000 and dual DGX Spark systems, with support through LM Studio and Ollama.
- Replit adds GPT-6 Astra one day after OpenAI's release
CEO Amjad Masad confirmed the rollout as Replit turns rapid access to frontier models into a core product advantage.
- Alexandr Wang releases Meta's Muse Spark 1.3 Max after safety testing
Meta's Max reasoning setting, which powered its strongest Muse Spark 1.3 benchmark results, is now available in Muse Code and the Meta Model API.
- Adaption launches a dataset generator that starts with instructions, not data
Sara Hooker and Sudip Roy's new tool turns a behavior description into synthetic training examples, then feeds them into AutoScientist.
- MiniMax and Together AI will bring open-model economics to London
The two AI companies will host Rio Shen, Max Ryabinin and Sarung Tripathi at a September 16th event on the production economics of open models.
- Google puts its AI teacher training on a monthly release schedule
New modules will arrive each month, while a September 19th virtual event will offer K-12 teachers live training and digital badges.
- Microsoft ships MAI-Image-2.6 to Foundry at about 4 cents an image
Mustafa Suleyman's lab paired Arena's No. 2 text-to-image model with a half-price Flash version for high-volume workloads.
- Google brings Gemini's Daily Brief to more free U.S. users
The morning agent reads across Gmail, Calendar and Gemini chats, provided users enable Memory and connect their Google Workspace data.
- GitHub launches HydraFusion to make several AI models do one coding job
The Copilot preview routes tasks through single-model, cascade, or critique workflows, trading extra inference for lower estimated costs.
- COPA releases 9B Armenian model with full training data and recipe
Erik Arakelyan's team released a 9B-parameter base model, a 4.37M-document corpus and verified STEM data for Armenian developers.
- Ozz Health pilots an AI guide for the hours hospitals leave unexplained
Galit Desheh, a trauma and gender scholar, founded Ozz Health to give patients context clinicians often lack time to provide.
- TrustKernel ships a $149 AI computer designed to keep agents off your phone
Li Wenhao's PlugClaw pairs a USB-C Android computer with cloud inference inside a hardware-protected environment that TrustKernel says keeps data hidden from both itself and the cloud provider.
- AgentConnect study finds coding agents often choose grep over LSP
Pengcheng Xu's preliminary study found that tool choice followed the task, while location-only LSP results increased follow-up reads and token use.
- Figure commits $3.5B in compute to Nscale for up to 100,000 Rubin GPUs
Figure targets initial deployment in Barstow, Texas, in the second half of 2027; Nscale will take an undisclosed stake and serve as Figure's preferred compute provider.
- Cloudflare launches OpenAI-powered vulnerability defense tied to live traffic
Matthew Prince's latest security bet uses GPT-5.6 Cyber to rank findings, propose patches and prepare WAF rules, with humans approving every change.
- Crusoe signs reported $13B Jane Street cloud deal amid $3B fundraising talks
The five-year GPU contract would turn a former quant trader's energy startup into Jane Street's second multibillion-dollar AI cloud supplier.
- NVIDIA PAIR turns spare PCs into a local AI inference cluster
The open-source beta routes parallel agent requests across Windows, Linux and macOS devices without pooling their GPU memory.