AI
Models, agents, infra, applied AI.
- Supertake turns investing theses into portfolios, starting with Robinhood
Michael Mignano's private-beta startup lets users share and fork portfolios; its terms identify Robinhood as the brokerage link, while his launch thread also named Coinbase.
- ElevenLabs ships v4 voice models, ranked No. 1 by Artificial Analysis
The new models add performance controls and a real-time version; a two-week introductory API offer starts at $22 per million characters.
- Meta puts Muse models, agents and coding tools under a new enterprise platform
Mark Zuckerberg named former MongoDB CEO CJ Desai to lead the effort, which brings Meta's AI models, business agents and developer tools to businesses.
- Four AI models designed racing creatures from scratch. Opus 5.5's six-legged creature won
Tomoyuki Muranaka set the same 18-second obstacle-course task for four models, then asked them to explain their results.
- David Zhang's jevgrep saves 29% on agent costs, excluding its own bill
The CLI searches code repositories for coding agents; its latest 10-task test kept the baseline's solve rate, while the launch post cited a different 40% result.
- Sakana AI uses simulation search to refine robot plans without retraining models
Its SAIL method raised simulated success from 25% to 73% with a 45-node search, but physical testing covered one task and six trials.
- Claude Opus 5.5 builds a browser Spider-Man game, with Shikhar steering
The non-commercial Spiderbench demo combines AI-directed code, Blender-built assets and Three.js; its creator says the goal is to test the model, not ship a game.
- Sidewinder brings arcade spinning-wheel controls to a Super Off Road-inspired game
Figma AI engineering VP Josh Clemm says he built the personal project with Codex, GPT-6 Astra and other tools; a full-game link has not been posted.
- Respan lists Span-01 on OpenRouter after claiming 4.5B tokens processed
The behavior-scoring model targets AI agent monitoring; Respan says its family has processed the volume since launch, while direct API access remains in early access.
- Panda takes $100 deposits for a $2,997 home AI computer due in 2027
The preorder promises private, subscription-free AI and December 2027 shipping, but Panda has not published the hardware specifications needed to assess its most demanding claims.
- K-Dense releases 32 lab-instrument connectors, all awaiting physical tests
The open-source LabMCP catalog exposes 411 actions, but K-Dense says none of its simulated connectors has been confirmed on a physical instrument.
- Claude Opus 5.5 wins Agent Wars bridge test with an estimated 130 lb load
Nickson reports that Claude's bridge held an estimated 130 lb, ahead of Meta's 26.5 lb and OpenAI's estimated 17.5 lb. Grok and Kimi did not complete the challenge.
- Fireworks says its Kimi K3 variant cuts reasoning tokens by 40%
Co-founder Dmytro Dzhulgakov promoted Ember-1 on September 27th; Fireworks introduced the model four days earlier with benchmark and customer-test results.
- TypeSafe's Jev makes Minecraft decisions in 24 milliseconds, Hao AI Lab says
The demo puts the model's fast, structured-action pitch in a game; TypeSafe raised a $40M seed round led by DCVC on September 15th.
- NaiveAI releases a 309B model built with AI-assisted research
The open-weight model targets coding and AI research, pairing a 1-million-token context with a custom inference runtime and company-reported speed claims.
- Earn an Honest Dollar tests whether extraction tools can admit what they don't know
Its synthetic-page benchmark found a one-line instruction sharply reduced invented fields, while showing how far agent-service verification still has to go.
- Synthetic Sciences takes OpenScience out of beta with 30-plus models
The YC W26 startup is adding a redesigned research workspace and paid model access to its open-source science agent.
- TinyAIArena turns AI-agent battles into replays; the benchmark still has to earn its score
Developer hp6's browser project pairs match playback with an Elo leaderboard, but its public interface leaves the rules and rating method unexplained.
- Likely OpenAI-linked agents used relays to retrieve UNCTAD data, researcher finds
Rowan Howard-Jones traces URLQuery, relay and browser-based routes to UNCTADstat. Transluce separately counted more than 1,000 URLQuery reports over two weeks, mostly involving UNCTAD statistics.
- Gen's AI privacy checker wants your email before explaining where data goes
Chief AI officer Howie Xu, a co-founder of acquired security startup TrustPath, is expanding Gen's Agent Trust Hub with a consumer-facing beta tool.
- Anthropic commits $11.6B to Akamai for seven years of CPU capacity
Anthropic is securing seven years of CPU capacity; Akamai says the deal could grow by $9B, subject to further purchases and service conditions.
- Supersonic Labs releases a 144M-parameter classifier built for CPU inference
Julia 1 is open under Apache 2.0, but Supersonic's own small Banking77 pilot scored below its comparison benchmark.
- Meta's Muse may need 2,000 server trays at reported user scale
Tom's Hardware's estimate combines Meta's Muse's reported 500,000-plus daily active users with 256 sandboxes per hypothetical tray. The figure is napkin math, not a disclosed Meta deployment.
- Benjamin Breen asks AI labs to fund historians' frontier-model research
The UC Santa Cruz historian argues for collaborations around tractable archival questions, drawing on early experiments in alchemy, cryptography and historical texts.