AI — Page 45
Models, agents, infra, applied AI.
- Mollick's Claude Fable 5 test highlights hours-long agent work, not another launch demo
Ethan Mollick says Claude Fable 5 worked for hours across research and coding tasks, offering a long-horizon outside read on the Amodeis' agent bet.
- Guan Wang's Sapient Says It Trained a 1B-Parameter Model for About $1,500
Sapient Intelligence's HRM-Text claim targets the enterprise fear that custom AI means frontier-model budgets and vendor lock-in.
- Ari Jacoby's Concentrate AI enters the AI routing fight as token bills bite
Concentrate emerged from stealth with more than $5 million, while OpenRouter's recent $113 million round shows how fast the gateway layer is heating up.
- Tesla hacker Yoni Ramon brings Pi out of stealth with $35M for AI security
Pi is valued at $100 million and counts Navan as an early customer, while Forbes reports xAI is also using the system.
- Mike Krieger turns Anthropic's Fable 5 launch into a product test
Anthropic says Fable 5 routes under 5% of sessions to Opus 4.8, while Mythos 5 keeps higher-risk capability behind trusted access.
- Instawork turns its gig marketplace into a robot-training data line
Instacore puts five cameras and a compute backpack on workers to capture commercial tasks for AI labs, with customers still unnamed.
- Anthropic launches Claude Fable 5 with a gated Mythos 5 for cyber use
The new model is priced at $10 per million input tokens and $50 per million output tokens, with some requests routed to Opus 4.8.
- Google ships Gemini 3.5 Live Translate across consumer, enterprise and developer tools
The audio model streams speech-to-speech translation across 70-plus languages, with Meet access limited to private preview this month.
- Harness-1 researchers say a 20B open search agent beat GPT-5.4 on recall
The UIUC, UC Berkeley and Chroma project shifts search memory from the model context window into a structured software environment.
- Leopold Aschenbrenner turns an AI thesis into a $20 billion hedge fund
WSJ reports Jane Street is now an investor in Situational Awareness, whose biggest disclosed win is tied to Anthropic.
- MMAE benchmark tests whether AI can edit audio without collateral damage
Tencent Hunyuan and university collaborators say current models post an Exact Match Rate below 5% on the new speech and audio editing benchmark.
- echohive turns Codex into a creative-coding assistant
The Three.js visual demo points to a smaller but important market for coding agents: creators selling workflows, not software seats.
- Hugging Face turns its community toward small-model efficiency
The Build Small Hackathon track asks developers to build with smaller models as open-weight systems move closer to local production use.
- 0G Labs says its coding-agent model fits locally in 18GB
The company says the Apache 2.0 model runs at 4-bit quantization, but the source material does not include a model card, repo or benchmarks.
- Vercel Sandbox persistence GA pushes agent state into managed infrastructure
The reported GA separates storage from compute, but the public item leaves pricing, limits, release date and official Vercel docs unverified.
- Infini-AI-Lab says Vortex hits 3.46x throughput with agent-generated attention
The research framework lets agents write attention flows in Python, compile them into serving kernels, and benchmark end-to-end LLM throughput.
- Startup Spotlight: MagicPath, Pietro Schirano's shared AI-native canvas for human and agent designers
Village Global has publicly tied MagicPath to investment activity, while Schirano's profiles identify him as founder and CEO of the AI design workspace; funding terms, customers and rollout details remain undisclosed.
- Claimed DeepSeek GUI leak mirrors OpenAI's Codex agent workspace
The screenshot remains unverified, but its project rail, agent canvas and bottom command composer suggest DeepSeek may be following the product philosophy OpenAI is pushing with Codex.
- CodeGuide teases Mac-1, a local 6.6B model built for macOS tools
Zafir says Mac-1 runs on Macs with 8GB-plus RAM and can chain tasks across 487 native macOS tools at about 65 tokens per second.
- Lockheed Martin Tests Combat AI Agents in Simulated Fight Club
The defense contractor says its synthetic environment ran virtual 4-on-4 air combat scenarios with Ansys Government Initiatives and ATG, compressing what it called 114 years of testing into one month.
- Patrick Jiang's Harness-1 externalizes memory for a 20B search agent
The paper reports 0.730 average curated recall across eight retrieval benchmarks, with code and model weights now public.
- Kilo Code AI says MiniMax M3 matched Claude Opus 4.8 on a code audit for $0.07
The open-source AI coding assistant company, founded by Scott Breitenother and Sytse Sijbrandij, says its self-run test found 13 of 17 planted bugs with MiniMax M3, tying the cheapest Claude run it priced at $1.30.
- Markus Buehler frames AI discovery as a verified regime shift
The arXiv preprint uses category theory and materials-science examples to define discovery as auditable schema change, not search inside a fixed problem space.
- Higgsfield AI's $500K feature film turns AI actors into a Hollywood test case
The 95-minute "Hell Grind" cost about $500,000 and turns generative video from short-clip promise into a labor and distribution test.