AI — Page 10
Models, agents, infra, applied AI.
- CostPerPrompt publishes calculators for estimating monthly AI application costs
The site packages model-price comparisons into calculators for APIs, chatbots, agents, retrieval systems and other AI workloads.
- Meta ships competitive AI models as capital spending tests returns
Independent benchmarks and access to 3.60 billion daily users undercut claims that Meta has little to show for AI, but a capital budget of up to $145 billion leaves the business case unproven.
- Head to head: Bernini-R Edit Video vs Marey Realism V1.5
This matchup wasn’t especially close on the scoreboard, even if a couple of individual prompts were more debatable on second pass. Bernini-R Edit Video separates itself where these tests usually matter most: prompt adherence through motion, occlusion, and multi-step action continuity.
- ByteDance ships Seedance 2.5 with 30-second video and timeline editing
The model doubles single-pass length, adds dense reference inputs and rolls out through ByteDance's own AI products, with an API planned.
- Truffle Security said its Hugging Face scan found 221,303 live credentials
Dylan Ayrey's couch-born TruffleHog found reused keys across public AI corpora, where deleting one source cannot erase its copies.
- Head to head: Bagel vs Fibo
Bagel shows flashes of polish, but this matchup wasn’t close. Fibo swept the board on prompt discipline, spatial reasoning, style control, and reflection-heavy realism, leaving Bagel with no task wins and only one split decision.
- Granola sued over bot-free meeting capture and AI training defaults
The proposed class action says non-users cannot consent or opt out when Granola transcribes calls and uses anonymized data for model improvement.
- OpenAI says unreleased Astra model solved 10 open math problems
The lab released a 249-page paper, reasoning transcripts and Lean certificates for claims spanning geometry, cryptography and quantum complexity.
- DeepSeek ships V4 Flash with strong agent scores and low API prices
DeepSeek's open-weight model undercuts rivals on API price, while third-party tests put its max-reasoning version near the top of its class.
- Claude Code reverse engineering alleges hidden Fable 5 model switches
Lon says a Claude Code path can route requests to Opus 4.8 without the notice and opt-out Anthropic's own help page promises.
- OpenAI previews Astra model built to coordinate long-running agents
Sam Altman demonstrated the unreleased model to policymakers as a federal review process approaches its August 1st deadline.
- Head to head: Bernini-R Edit Video vs Luma Ray 3.2 Image to Video
This one wasn’t subtle. Across four judged prompts, Bernini-R Edit Video consistently delivered tighter prompt adherence, cleaner motion control, and more reliable scene logic than Luma Ray 3.2 Image to Video.
- Plasma AI released a deterministic wiki CLI for coding agents
The open-source tool lets agents map, search and read slices of Markdown knowledge before modifying a codebase.
- MemTensor releases Metis models that write memory inside the backbone
The Qwen3.5-based research preview includes 4B, 9B and 27B checkpoints, while withholding training data and restricting the code to noncommercial use.
- Google backs away from an AI Studio app after claiming 800,000 preorders
The app was due August 1st; Google's existing product already lets builders create and distribute software through the browser.
- Explorative Modeling adds best-of-K search to generative model pretraining
Alexi Gladstone's open-source method spends extra compute during training and reports gains across image, video and masked language models.
- Edison Scientific launches research hub with open models for reading chemical patents
The FutureHouse spinout is publishing weights, code and benchmarks to support its push to sell Kosmos into biopharma R&D.
- Tracer launches Echo with near-Claude Fable scores at one-third the cost
Solo founder Adam Rida built the YC S26 system to allocate compute across open-weight models through one OpenAI-compatible endpoint.
- Supabase launches open benchmark for AI coding agents building backends
The framework tests Claude Code, Codex and OpenCode on database, auth and Edge Function tasks, then publishes the scores.
- Head to head: AuraFlow vs Cosmos 3 Super
One model showed flashes of style; the other actually did the assignment. Across eight image prompts, Cosmos 3 Super separated itself on prompt adherence, spatial reasoning, realism, and text handling—and the result wasn’t remotely close.
- Argo IQ says revenue grew 10x as it shifted from AI QA to growth services
Founder Trenton Hughes is selling recurring content and growth operations to B2B companies, with human experts supervising AI output.
- Argil adds MiniMax H3 to its AI-avatar social video workflow
Founders Laodis Menard and Brivael Le Pogam are placing MiniMax's video model inside Argil's production system for avatars, scripts, captions, B-roll and finished social clips.
- Huawei releases 505B-parameter openPangu model trained on Ascend chips
The 512K-context release tests whether Huawei can turn its chip, model and inference stack into a credible alternative to Nvidia-centered AI.
- Altman claims AI cost declines are 20x faster than Moore's law
The OpenAI CEO made the claim after the company reported a 20% serving-cost reduction and greater than 15% token-efficiency gain from GPT-5.6 Sol, while passing some savings to Luna and Terra users.