All Stories — Page 26
Every story published on RuntimeWire, newest first.
- Duolingo cuts AI video-call cost below 1 cent with open models
The cost decline is helping Luis von Ahn move conversation practice into Super and reconsider Duolingo's premium Max tier.
- Insilico previews virtual-cell research that models biological age
Insilico previews a Virtual Aging Cell concept that adds biological age to cell simulations, though the posting date is unclear.
- Waku combines coding-agent CLIs in a native local desktop app
Version 0.1.0 uses Rust, GPUI and Git checkpoints to combine agent activity without a Waku account or cloud intermediary.
- Startup Spotlight: Sellara is building AI agents for the work after a Wall Street trade
Sellara's former Citi bankers are building specialized AI agents for trade capture and reconciliation, backed by a16z Speedrun.
- OpenAI benchmarks a 16x load-time gain for giant ChatGPT and Codex threads
A 741-turn, 231 MB thread loaded in 1.66 seconds in a test, with lower memory growth and 98% fewer requests.
- Printytron turns plain-English requests into dimensioned STL parts without CAD
Printytron's independent alpha focuses on brackets, adapters and replacement parts, then uses a second AI to inspect each generated model.
- Grok 4.6 tops Newsroom Reliability v0.2 benchmark at 0.78
In a 50-task run of Newsroom Reliability v0.2, SpaceXAI’s Grok 4.6 ranked first with a score of 0.78. OpenAI’s GPT-5.6 Luna Pro followed at 0.77 with the lowest reported cost among the top three, at $0.0005 per task.
- Netflix tests GenRec on 10% of traffic, reports 0.006% relative lift
Netflix engineers Ying Li, Arjun Rao and Shradha Sehgal tested an LLM-backed ranker using about 40 times fewer Phase 2 labels than the production model, reporting a 0.006% relative lift in an undisclosed core metric.
- Stephen Cresswell ships Yadda 3 after a one-day Claude-assisted rebuild
The longtime JavaScript maintainer says Claude handled most of the work, while an existing test suite kept the agent from redefining correctness.
- Sankalp uses Codex loop to place 12th in GPU Mode's B200 QR contest
Sankalp placed 12th among 183 entrants after using Codex, profiling tools and more than 1,500 submissions to optimize a B200 QR kernel.
- Andon Labs carries out first firing recommended by its AI manager
Luna needed a human prompt to recover its attendance policy, and people still reviewed and executed the termination.
- Xia Chen ships ThoughtDAG v0.3.13 to expose LLM context
The open-source desktop app lets users branch, prune, merge and preview the exact conversation history sent to a model.
- Amp holds SOC 2 Type II without requiring pull requests
Amp, which [LinkedIn lists at 11-50 employees](https://www.linkedin.com/company/amp-code/), uses signed commits, automated checks and agent threads while skipping mandatory pull requests.
- Deltix pairs AI app exploration with replayable iOS regression tests
Its Mac agent keeps app binaries local, while screenshots and interface data still travel to Deltix's cloud and Anthropic.
- Owen Van Natta, Facebook's early COO and startup investor, dies at 56
Dan Rose credited his longtime mentor with negotiating pivotal Amazon and Facebook deals before Van Natta became an early-stage investor.
- DeepSeek open-sources coding-agent harness before V4-Pro API price hike
Harness v0.1 lets developers swap models, tools and runtimes, while higher V4-Pro rates take effect August 16th.
- California DMV permits Aurora and Kodiak to test autonomous heavy trucks
Kodiak has begun supervised runs near Mountain View, while Aurora and Kodiak must each complete 500,000 autonomous miles before driverless testing and another 500,000 before deployment.
- Maple said it made DeepSeek V4 Flash available to paid customers
An undated Maple post announced access for paid customers, while public materials leave the qualifying plans, Proxy availability and model-specific pricing unclear.
- Community attributes animated SVG to Qwen3.8-27B without reproducible run details
A pelican-on-a-bicycle animation attributed to Alibaba's open-weight model lacks the prompt, raw SVG and execution details needed for independent reproduction.
- GitHub adds xAI's Grok 4.6 to Copilot for long-running coding agents
xAI's Grok 4.6 is entering GitHub's IDE, command-line and enterprise workflows two days after launch, under usage-based billing.
- Anthropic stole ChatGPT's aura. Then it gave it back.
Claude Code made Anthropic Silicon Valley's chosen winner. Its models and business remain formidable. Self-inflicted errors and a contested warning about collapsing morale have punctured the sense that the company could do no wrong.
- Joshua Kushner warns AI excitement could weaken venture discipline
Thrive's first investor letter defends concentrated investing while its deep OpenAI ties put that discipline under a brighter light.
- sandbox.bio ships an embeddable Linux terminal for runnable technical tutorials
Robert Aboukhalil and Maria Nattestad are turning a five-year bioinformatics teaching tool into infrastructure publishers can add to their own sites.
- 224 Ventures outlines seed strategy for robotics, work software and AI infrastructure
Shaun Johnson says the $100M-plus firm will seek nonconsensus seed investments with Oriol Vinyals and Yann LeCun.