RuntimeWire
Real-time startup intelligence for the AI economy. Funding, model launches, infra shifts, and founder moves — the signals that matter.
Latest stories — page 85
- Exclusive: RoboRank, the 'LeetCode for roboticists,' will open evaluation environments and run a public scoreboard
RoboRank, a LeetCode-like benchmarking project for robotics at roborank.dev, will separate and open-source its evaluation environments, target VLAs and world models, and harden sandboxed code execution so researchers can contribute tasks and models in the open.
- OpenClaw momentum builds around a local, open agent as a Google 'Spark' rumor circulates
Aligned News cited a 300,000-star moment and a Google 'Spark' entrant; while unverified, the buzz spotlights OpenClaw's local, open, self-hosted agent thesis.
- Zeb Evans cuts 22% at ClickUp and bets on 3,000 AI agents to build a 100x org
The ClickUp CEO says savings will fund million-dollar salary bands for AI-leveraged top performers, even as Gartner warns automation cuts do not guarantee returns.
- AlphaProof Nexus teaser hints at agentic math push, but the builders stay unnamed
A brief X post teased an agentic framework for research-level math, but shared no docs or team names identifying what AlphaProof Nexus is.
- Replit backs Musixmatch Pro Musicathon, a global remote hackathon with $25k+ in prizes
The June 15th-21st event is fully remote and supported by partners including ElevenLabs, Songstats, LALAL.AI, and Cyanite, according to Replit's post on X.
- Pushmeet Kohli shares Google DeepMind's AlphaProof Nexus results: agentic proof search in Lean
VP of Research Pushmeet Kohli points to a GitHub trove of Lean-formalized proofs and prose by AlphaProof Nexus, signaling progress while holding back the framework code.
- Season with volts: Kirin's Electric Salt Cup and Spoon aim to keep low-sodium joyful
A Facebook friend got one as a gift and I had to know more. Yes, there is a battery at the table. Kirin's new cup and renewed spoon use weak current to boost perceived saltiness and umami, with Japan online sales starting Sep 9th and retail in November.
- Anthropic's Claude Code Auto Mode rolls out to Pro and adds Sonnet 4.6, per community post
A widely reshared ClaudeDevs note, surfaced by Aligned News in a post on X, says Auto Mode now runs on Pro with Sonnet 4.6 and Opus 4.7.
- LimX Dynamics unveils Luna, a full-size female humanoid built for malls and theme parks
LimX Dynamics is positioning Luna for public-facing roles, but the X post offered no specs, pricing, or timeline to back the mass-deliverable claim.
- Seth Howes says he sequenced a full human genome at home to 30x coverage
In a thread on X, Howes details a one-room setup, cites a $28k sequencer and ~$1.2k per run, and calls out a discontinued P2 Solo unit.
- Chert launches Twilio-for-iMessage API and GTM service
YC-backed Chert is pitching a single API for blue-bubble threads with SMS/RCS fallback and CRM integrations, and says it will run outbound as a service for teams.
- Aligned News flags new paper on evaluation awareness in frontier LLMs
Haritz Puerto says a paper on decomposing and measuring evaluation awareness just dropped, plus a resource called EvalAwa..., but links and authorship were not shared.
- Changling Li leads EvalAwareBench to measure when LLMs know they are being tested
In a new paper and open releases, Li and collaborators decompose evaluation awareness, test nine models across four benchmarks, and publish a factor-controlled dataset and code.
- Hugging Face leader says Gemma 4 tops 120M downloads in weeks, counting Hugging Face and Ollama only
The tally counts only Hugging Face and Ollama pulls, hinting at on-device demand but leaving methodology and release timing unclear.
- Zhao Tongyang's EngineAI starts 10,000-unit humanoid line; first T800s roll off Shenzhen base
XRoboHub reports the line is live at an integrated Shenzhen facility, but capacity timing and T800 specs were not disclosed.
- Asimov plans Palo Alto, SF, and Austin meetups to talk humanoids and AI
Coffee meetups land June 2nd, 4, and 6, with RSVPs running through Luma as Asimov convenes folks interested in humanoids and AI.
- Rumor: Sam Altman backs Pharia Health, transcranial magnetic stimulation (TMS) protocol startup
The company targets high-performing professionals, pairing FDA-cleared TMS with d-cycloserine and optional $200 per month maintenance sessions.
- Ascii.dev founders Anicet Nougaret and Kirill Makarov push Box, claiming 5x cheaper agent sandboxes
Ex-ESA data science intern and Volvo Group engineer turn to agent infrastructure; Box targets agent developers with one-command sandboxes, a simple API, and pricing claims 5x below popular offers at equal specs.
- Freu AI launches Mac agent that compiles your cross-app workflow once, then runs it locally with zero recurring token cost
Demos show a Mac agent that records your cross-app workflow once, compiles it into a deterministic DSL, and replays it locally with zero recurring tokens. With freu-cli open-sourced and a local vision execution model coming, can ahead-of-time semantic compilation beat screenshot agents on cost and latency?
- Socket uncovers TrapDoor campaign stealing keys and wallets via open source packages
Socket researchers say the TrapDoor campaign planted credential-stealing payloads in more than 34 packages and 384+ versions, targeting crypto and AI developers.
- DeepMind preprint: LLM+Lean agent resolves 9 Erdos problems and 44 OEIS conjectures
Google DeepMind reports a full-featured AlphaProof Nexus agent solved 9 of 353 open Erdos problems at a few hundred dollars per problem and proved 44 of 492 OEIS conjectures; code and Lean proofs are on GitHub. This is an arXiv preprint and community validation is pending.
- Yuyin Zhou releases ClinSeekAgent, an open-source clinical AI agent that seeks its own evidence
Zhou's team links raw EHR, web, and chest X-ray tools, and reports open-source SOTA after distilling Claude Opus 4.6 into a 35B model.
- Greg Brockman walks through the 72 hours that almost killed OpenAI
In a rare interview on The Knowledge Project, the OpenAI co-founder recounts quitting within hours, sketching a backup company, and why they stopped showing reasoning traces.
- Draw Things adds on-device Qwen 3.5 4B and a clearer mode selector on iOS and macOS
The app says it now runs a lightweight local LLM as an interrogator model and splits generation and editing into distinct modes for a cleaner workflow.