Skip to content

Product updates

New features, redesigns and pricing in AI products — whose product got better, pricier or finally usable.

842 picksRelated topicsModel releasesIndustryAI coding

Latest picks

161–180 of 842

May 31Sunday

AI HOT (Curated Pool)

Apple WWDC AI Upgrade: Gemini-Distilled Model Runs Locally, With Heavy External Dependencies

Apple will present Siri and on-device AI upgrades at next month’s WWDC, with iPhones running a smaller Gemini-distilled model locally while complex queries route to Google Cloud using Nvidia confidential computing.

Why it matters: HKR-H/K/R all pass: the Apple-Google-Nvidia stack is a strong WWDC AI hook with a concrete routing mechanism and clear industry tension. Capped at 82 because this is a single X-sourced claim with no model size, latency, pricing, or contract terms disclosed.

r/LocalLLaMA

Use any model and provider with the official OpenAI Codex Desktop App without modifying its code

Reddit user thibautrey describes a 3-step setup: edit Codex Desktop config.toml, store an API key, and use a multicodex proxy alias to map gpt-5.3-codex to MiniMax-Latest. The post lists a local base_url of 127.0.0.1:1455 and says the proxy disguises returned model names as gpt-5.3-codex.

Why it matters: This is a reproducible developer workflow trick, not an official release. HKR-H comes from the lock-in workaround, HKR-K has concrete config details, and HKR-R hits cost and model-choice pressure, placing it at the tutorial featured threshold.

QbitAI · WeChat

NVIDIA’s MacBook Pro-like laptop reportedly uses an in-house CPU

NVIDIA, Microsoft, and Arm posted the same “new era of PC” teaser, and the article says the rumored N1X laptop may use a 20-core Arm CPU, a Blackwell GPU, 6,144 CUDA cores, and 128GB of LPDDR5X unified memory, while bandwidth and x86 translation remain the stated constraints.

Why it matters: HKR-H/K/R all pass, but the story rests on hints and rumored specs; launch date, price, and production plan are not confirmed. Treat it as a strong hardware rumor, not a same-day must-write release.

AI HOT (Curated Pool)

Tesla FSD completes a 6,000 km zero-intervention autonomous drive across Canada

Tesla FSD V14.3.3 completed a 6,051 km zero-intervention drive from Vancouver to Halifax in 4 days and 21 hours, with the system handling lane changes, complex road conditions, and parking without disengagements or human corrections.

Why it matters: HKR-H/K/R all pass: Tesla FSD V14.3.3 has a concrete 6,051 km zero-intervention claim. It stays below 85 because the item gives the result but lacks independent validation, route detail, and failure boundaries.

AI HOT (Curated Pool)

Run Python ASGI Apps in the Browser with Pyodide and Service Workers

Simon Willison demonstrated running Python ASGI apps in the browser with Pyodide and Service Workers, with Claude Opus 4.8 assisting development, and showed two working demos: a basic ASGI FastCGI demo and Datasette 1.0a31.

Why it matters: HKR-H/K/R all pass: the post has a surprising browser-runtime hook, concrete mechanisms, and developer resonance. Impact stays in the 72–77 band because this is a developer experiment, not a model or platform launch.

AI HOT (Curated Pool)

DynoSim: Simulation-Driven Inference Stack Optimization

NVIDIA released DynoSim for optimizing its Dynamo inference serving stack; the Rust-based tool models thousands of deployment configurations on a single virtual timeline and reached 1,500x real-time speed in tests.

Why it matters: HKR-H/K/R all pass: the hook is 1500x real-time simulation, with a concrete virtual-timeline mechanism and infra cost resonance. Single-source NVIDIA product update keeps it in the lower featured band.

AI HOT (Curated Pool)

“What a joke”: GitHub Copilot’s new token-based billing draws developer backlash

GitHub Copilot changed billing to token-based metering, and the RSS snippet says developers are unhappy; the post does not disclose pricing, per-token rates, or the rollout date.

Why it matters: HKR-H/K/R all pass: Copilot’s token billing creates conflict, a concrete mechanism, and a cost nerve for developers. Missing price, unit economics, and start date keep it in the lower featured band.

May 30Saturday

TechCrunch · AI

I put Google’s 24/7 AI assistant Gemini Spark to work, and it’s actually pretty useful

TechCrunch tested Google’s Gemini Spark as a 24/7 AI assistant for inbox summaries and local event planning; the RSS snippet does not disclose pricing, release timing, or why Google made it a separate product.

Why it matters: HKR-H/K/R pass: the hands-on angle is clickable, and inbox plus local-planning automation gives concrete substance. The score stays in the low featured band because price, launch timing, and product positioning are not disclosed.

AI HOT (Curated Pool)

Nano Banana Pro and Nano Banana 2 officially released

Google AI Developers released Nano Banana Pro and Nano Banana 2, mapped to gemini-3-pro-image and gemini-3.1-flash-image. The post says both are production-ready through the Gemini API, but does not disclose pricing, benchmarks, or runtime limits.

Why it matters: HKR-H/K/R all pass: Google names two image models and production Gemini API access. Missing pricing, benchmarks, and invocation limits keep it in the mid product-update band rather than a must-write release.

Xinzhiyuan · WeChat

Claude AI fluency scorecard surfaces, with strong users scoring 7.5

Anthropic is testing a Claude AI Fluency scorecard that analyzes Chat, Cowork, and Claude Code history against 11 observable behaviors, with an 11-point maximum score. The underlying study used 9,830 anonymized multi-turn conversations, and iteration appeared in 85.7% of high-quality conversations.

Why it matters: HKR-H/K/R all land: the angle is clickable, the scorecard has concrete numbers, and Claude users will debate being graded. This is not a model launch or major capability release, so it stays in the 78–84 featured band.

Xinzhiyuan · WeChat

Opus 4.8 Builds a Historical Rebirth Simulator for 117 Billion Humans

Ethan Mollick used Claude Opus 4.8 to generate The Veil of History, a website that weights a random human life by 117 billion historical births and, according to the article, uses 4,000 Monte Carlo runs to estimate regional and era distributions.

Why it matters: HKR-H/K/R all pass: Mollick’s Claude Opus 4.8 demo has a strange hook, concrete numbers, and a builder-relevant prototyping angle. It is not an Anthropic release, so it stays in the lower featured band.

AI HOT (Curated Pool)

xAI drops JAX GPU for an in-house training framework

SemiAnalysis says xAI dropped JAX GPU and moved to a C training framework written with Grok Build; the snippet claims xAI’s JAX stack had MFU below 10%, but the post does not disclose reproducible benchmark conditions.

Why it matters: HKR-H/K/R all pass: xAI changing its training stack is a strong hook, MFU <10% is a concrete claim, and infra cost will spark debate. Single-source tweet format and no reproducible setup keep it at 80, not P1.

AI HOT (Curated Pool)

Codex Can Manage Conversation Threads and Parallel Tasks

Codex can now create, search, organize, and pin conversation threads inside the Codex interface, and start worktrees for parallel tasks.

Why it matters: HKR-H/K/R pass: Codex gets concrete thread-management and parallel-worktree mechanics that matter to coding-agent users. Scope, pricing, and performance data are not disclosed, so this stays in the lower featured band.

AI HOT (Curated Pool)

Codex now supports computer use on Windows

OpenAI added Windows computer-use support for Codex, letting users start, review, and guide tasks on a Windows PC through the ChatGPT mobile app; the post states this is an early experience and does not disclose pricing or rollout scope.

Why it matters: HKR-H/K/R all pass: OpenAI adds Windows computer use to Codex, controlled through ChatGPT mobile. The post gives the workflow and early-stage condition, but not permissions, pricing, or rollout scope, so this stays at the featured threshold.

The Verge · AI

Tech companies desperately want to film you doing chores

Shift said it would clean New Yorkers’ homes for free if it can film cleaners doing chores such as washing dishes, wiping counters, dusting tables, and mopping floors, creating domestic robot training data; the snippet does not disclose consent terms, data retention, pricing, or expansion timing for cities such as London.

Why it matters: HKR-H/K/R all pass: the odd trade is clickable, the post gives a concrete chore-video collection mechanism, and it hits robotics data plus privacy nerves. This is a strong industry feature, not a major model or platform release, so it sits in 72–77.

AI HOT (Curated Pool)

OpenRouter supports model-generated file patches

OpenRouter now supports apply_patch, a server-side tool that lets any model propose file edits through the Responses API using V4A diffs, covering file creation, updates, and deletion, with OpenRouter validating diff syntax on the server.

Why it matters: HKR-H/K/R pass: the OpenRouter update gives coding agents a concrete cross-model patch path with V4A diffs and server validation. It is useful infra, not a model-level release, so it sits low in the 72–77 band.

AI HOT (Curated Pool)

xAI Releases Grok Build 0.1 Public Beta

xAI released grok-build-0.1 as a public beta through its API; the same model powers the Grok Build CLI, targets agentic coding, and is priced at $1 per million input tokens and $2 per million output tokens.

Why it matters: HKR-H/K/R all pass, but the post is thin: beta, CLI, pricing, and agent-coding positioning only; no benchmarks, context window, or hands-on results. This fits a mid-weight product update.

May 29Friday

The Verge · AI

This AI startup will clean your home for free to train future robots

Shift offers free home cleaning and records cleaners scrubbing, vacuuming, dusting, tidying, and washing to collect robot training footage; the RSS snippet does not disclose service cities, privacy terms, consent mechanics, or dataset scale.

Why it matters: HKR-H and HKR-R are strong: Shift turns home cleaning into robot-training data collection. HKR-K has a clear mechanism, but city scope, privacy terms, and dataset scale are not disclosed, so this stays low-featured.

The Verge · AI

Adobe’s Conversational AI Agent Is a Mediocre Design Intern

The Verge tested Adobe Firefly AI Assistant in beta. It can operate Adobe design apps as a conversational middleman, rather than only generating images or video. The post says it explains edit steps clearly, but the results were not impressive. The RSS snippet does not disclose pricing, release timing, or the full list of supported apps.

Why it matters: HKR-H/K/R all pass because this is a Verge hands-on of Adobe’s Firefly AI Assistant beta with a clear negative usability hook. Missing pricing, launch timing, and supported-app details keep it in the 72–77 featured-threshold band.

QbitAI · WeChat

Tencent unveils Code Craft, an AI game creation platform for beginners and developers

Tencent Games unveiled Code Craft, an AI game creation platform that turns natural-language prompts into runnable 2D or 3D games, with a planning knowledge base, Skill system, visual tuning panels, and more than 20,000 free cloud assets; the post does not disclose release timing, pricing, model details, or supported engines.

Why it matters: HKR-H/K/R pass on the Tencent game-creation hook, runnable 2D/3D output, and 20,000+ assets. Pricing, access scope, and model limits are not disclosed, so it stays in the lower featured band.