AI — Page 17
Models, agents, infra, applied AI.
- Hugging Face explains why DeepSeek wants lookup tables inside LLMs
Blackroot's new guide translates a January architecture that trades some MoE experts for hashed N-gram embeddings.
- Alibaba switches on Qwen3.8-Flash API at $0.16 per million input tokens
The same-day follow-up puts Qwen's new model on QwenCloud with a default 1-million-token context and OpenAI and Anthropic API compatibility.
- Perceptron ships Isaac 0.5 as an open-weight model for factory-floor robots
The former Meta researchers behind Isaac 0.5 are betting one model can handle perception, reasoning and control without a stack of specialized systems.
- Reuters puts its news archive in Snowflake for enterprise AI
Customers get Reuters journalism dating to 1987 in five languages, while pricing and permitted model-training uses remain undisclosed.
- Z.AI confirms Ox Alpha is a GLM model, plans to release its weights
The Beijing lab told Bloomberg it built the anonymous coding model and planned to release its weights on August 26th.
- Alibaba's Qwen3.8-27B ranks first among open models in Arena's Image-to-WebDev test
The 27-billion-parameter open-weight model ranks seventh overall in a human-preference test of visual web development, a useful result that says little about production code quality.
- ModelBest puts MiniCPM in a $2,999 desktop edge AI node
ModelBest's Pinea Pi pairs local MiniCPM models with Intel or Nvidia compute and integrated sensors ahead of a planned Kickstarter campaign.
- Thinking Machines offers up to $50,000 in credits for Inkling safety research
Thinking Machines Lab will subsidize outside safety work on Inkling, although it has not disclosed the program's budget or number of awards.
- MiniMax says M3 sent a business email for $0.018
Founder Yan Junjie's efficiency thesis gets a tidy agent demo, though the visible test does not establish a repeatable cost.
- OpenAI bans Russia-linked ChatGPT accounts promoting a think tank built on copied papers
The campaign used ChatGPT for posts and channel logos while an Israel-branded website republished academic work and praised Russia.
- OpenAI launches $100 ChatGPT Business seats for heavier AI workloads
Premium costs five times the Standard rate and includes 5x usage, while letting workspace owners mix both seat types across a team.
- Stardock lets teams share one AI session without moving it to the cloud
Brad Wardell's Clairvoyance 0.85 lets colleagues work with the same agents while the host computer retains the session, files and local execution.
- OpenAI reports up to 3.6x lower Jalapeno latency in engineering tests
OpenAI reports gains over Nvidia GB200 and GB300 systems in company-run tests, with production deployment planned by the end of 2026.
- Figure launches Index, a paid human-video pipeline for training humanoid robots
Figure says 44,000 weekly creators have uploaded 16M videos, and it plans to spend more than $1B on data and compute over the next year.
- OpenAI says Jalapeno beats Nvidia Blackwell on inference speed per watt
The first-party accelerator beat GB200 and GB300 systems in OpenAI-run tests, with deployment planned inside OpenAI's infrastructure by December.
- Skild AI ships S1, a robot model prompted by one video
The model learns unseen tasks lasting up to 10 minutes without fine-tuning, according to tests published by the robotics developer.
- Ox Alpha filters domestic Chinese political risks, CTGT audit finds
The anonymous coding model matched GLM-5.2's tokenizer and heavily suppressed seven topics tied to Xi, party legitimacy and domestic protest.
- IBM cuts the decoder from Granite Speech, says transcription runs 20x faster
IBM's 470M-parameter English models trade translation and keyword biasing for lower latency, smaller memory needs and edge deployment.
- Anima Anandkumar launches Accelerated Understanding to model physical systems
Anandkumar and Benedikt Jenik kept building Accelerated Understanding after declining a proposed Prometheus package; the company says its physics-focused model handled 5 trillion pieces of data in one prompt.
- Laude ships an agent harness that keeps thinking after you stop asking
The Laude-MIT collaboration runs a self-directed thought loop in less than 10,000 lines of Bash, with continuous token costs and familiar security problems.
- Plicara finds non-English agent skills reached 16.3% in Q2
Its 255,068-skill comparison shows multilingual growth, while lexical discovery remains a likely constraint on cross-language reuse.
- Microsoft Paint embeds server-issued IDs in locally generated AI images
Reverse engineer Xusheng Li found that Paint and Photos alter output pixels with a GUID untouched by the visible-watermark setting.
- MiniMax catalogs H3 integrations while keeping Context-IR hosted
The new index covers 24GB local deployment, ComfyUI and multi-GPU serving, three weeks after MiniMax released H3's weights.
- Meta plans Hatch agent launch while considering a $199.99 monthly tier
Internal documents point to a late August or early September rollout, followed by Meta's Watermelon model in October.