AI — Page 14
Models, agents, infra, applied AI.
- OpenAI's policy chief says frontier lab insiders are worried about AI, too
Dean Ball, hired in July to shape frontier policy, says he will keep issuing candid warnings even when they complicate OpenAI's message.
- IBM launches K-12 AI fellowship as classroom use outruns teacher training
The first cohort will include up to 100 New York-area superintendents, education leaders and educators, with SkillsBuild learning, credentials and responsible-AI capstone projects this fall.
- Grok 4.6 kept its name and changed its biosecurity behavior, LatchBio says
LatchBio found 59.2% red-team refusal and 64.8% routine completion, showing how far behavior can shift under one model name.
- Alibaba updates Qwen3.8-Max, leaves the -0902 checkpoint mapping unclear
Alibaba says the September 2nd update improves coding and Cowork performance; its Singapore service lists $2 per million input tokens and $6 per million output tokens.
- Google readies Gemini 3.8 Flash as its engineers reportedly prefer it to Claude Opus
The reported coding model could arrive 20 days after Gemini 3.7 Flash, though Google has yet to publish benchmarks, pricing or a model card.
- Baseten's inference essay separates real efficiency gains from cost shifts
Baseten's technical essay turns latency, throughput and model quality into a framework for deciding where expensive inference gains come from.
- Basis, Clay and Exa turn onboarding, sales and integrations into agent workflows
Basis, Clay and Exa Labs show three ways founders are turning bounded internal processes into agent workflows, from onboarding and account triage to tested developer integrations.
- MiniMax H3 generates a 10.125-second audiovisual file in under nine seconds
MiniMax's open-weight H3 now runs through vLLM-Omni, while its community license excludes the US, EU, UK and Republic of Korea.
- World Labs launches Atlas for video, 3D reconstruction and robot simulation
Fei-Fei Li's company is putting camera control, scene reconstruction and synthetic robot views into one model, with access initially limited to selected partners.
- Altman says faster AI self-improvement would push OpenAI's IPO further out
The CEO linked OpenAI's path to a public offering to the pace of recursive self-improvement as the company slows frontier research and tightens model safeguards.
- Z.ai gives GLM Coding Plan subscribers a quota refill after one year
Every current subscriber gets a Reset Card for weekly and five-hour limits, with redemption tied to a signed-in ZCode account.
- Flower Labs launches Endeavor 1.0 for private frontier AI deployments
The Cambridge-linked founders are extending Flower's federated-learning stack into a licensed model preview for sensitive enterprise workloads.
- Alibaba says Qwen3.8-Max leads open-weight models on commerce agent benchmark
The Accio team's state-based benchmark reports Qwen3.8-Max completing 56 tasks through Accio, 47 through OpenClaw and 53 through Pi; Alibaba calls it the strongest overall open-weight result.
- Manus resumes independent operations after Beijing forced Meta to unwind its $2B deal
Founders Red Xiao, Tao Zhang and Yichao "Peak" Ji remain in charge after a separation that forced account and task-data deletions.
- DHH says agents wrote Omarchy Quattro, while Basecamp exposed their limits
DHH says agents handled Quattro's implementation, while 37signals' Basecamp work showed how quickly mature architecture can become a cleanup job.
- Prodigy Research says its AI returned 140%, with few details attached
The two-person YC lab says delta-neutral strategies gained 140%, but it has not published the period, capital base, leverage or methodology.
- Apple says ex-engineer used its schematic for OpenAI hardware work
A forensic review of Chang Liu's Apple-issued MacBook has become the iPhone maker's main argument for accelerating discovery.
- Nous Research ships Hermes Agent Bot Mode for multi-agent group chats
Version 0.21.0 adds persistent bot-to-bot messaging, steerable subagents, browser control and scheduled jobs that remember prior runs.
- Meta takes Muse Code out of beta and adds $5 monthly plans
Meta's terminal coding agent adds cross-session messaging, multi-agent workflows and an SDK preview after a 26-day beta.
- Pennant joins YC with AI software for institutional proxy voting
The two-person startup says its early customers include a large US public pension plan, a top law firm and governance advisors.
- Runway says Solaris generates interfaces frame by frame in real time
The early-access research model turns clicks and drags into synthesized 720p frames, while text, trust and accessibility remain open problems.
- DeepSeek publishes 168GB vision model weights under an MIT license
The 305B-parameter checkpoint lands 10 days after its API debut, while DeepSeek's headline benchmark uses Anthropic's superseded Opus 4.8.
- Together AI signs 250 MW Saudi deal, projects $5B annual revenue
HUMAIN will supply power and 120,000 AI chips under a revenue-sharing arrangement for open-model training and inference.
- Grok 4.7 rumor revives Musk's September promise and SpaceX data bet
A viral post adds an unverified training-complete claim to Musk's August promise of a September model trained on SpaceX engineering data.