Skip to content

Product updates

New features, redesigns and pricing in AI products — whose product got better, pricier or finally usable.

842 picksRelated topicsModel releasesIndustryAI coding

Latest picks

181–200 of 842

May 29Friday

Xinzhiyuan · WeChat

Claude Opus 4.8 tests split users: strong at high effort, costly under rate limits

The article says Claude Opus 4.8 scores 63 on an Extra-High senior engineering benchmark, 30 points above Opus 4.7, but drops to 42 at High effort, while $200/month Max users report hitting rate limits within hours on complex agent tasks.

Why it matters: Anthropic/Claude relevance plus concrete test numbers clears HKR-H/K/R: the hook is strength versus cost, K has benchmark and quota details, and R hits agent-budget anxiety. Source is a media test rather than an official release, so this lands at low P1.

r/LocalLLaMA

Liquid AI releases LFM2.5-8B-A1B

Liquid AI released LFM2.5-8B-A1B with a 128K context window, 38T pre-training tokens, large-scale reinforcement learning, doubled vocabulary for non-Latin tokenization, and availability on Hugging Face.

Why it matters: HKR-H/K/R pass: 8B/A1B, 128K context, and 38T tokens are concrete hooks for local inference. No benchmarks, license, or deployment limits are disclosed, so it stays in the mid featured band.

AI HOT (Curated Pool)

Strengthening Societal Resilience with Rosalind Biodefense

OpenAI launched Rosalind Biodefense and provides trusted GPT-Rosalind access to vetted developers and U.S. government partners; the post does not disclose model parameters or pricing.

Why it matters: HKR-H/K/R all pass: OpenAI launched GPT-Rosalind access for vetted developers and US government partners. Missing parameters, pricing, and eval results keep it below a major capability release.

r/LocalLLaMA

StepFun 3.7 Flash

StepFun released Step 3.7 Flash with 196B total parameters, 11B active MoE, a built-in 1.8B ViT, and local execution on 128GB RAM.

Why it matters: HKR-H/K/R pass via the 196B/11B MoE specs and 128GB local-run claim. Sparse Reddit sourcing leaves license, eval method, and access conditions undisclosed, so it stays in the lower featured band.

AI HOT (Curated Pool)

StepFun Releases Step 3.7 Flash, Focused on Agent Efficiency

StepFun released the open-source Step 3.7 Flash model with a 198B-parameter MoE architecture, about 11B active parameters, a 256K context window, and a 67.1 score on ClawEval-1.1.

Why it matters: HKR-H/K/R all pass: the release has a clear sparse-model hook, concrete context and benchmark numbers, and practitioner resonance around open agent efficiency. Official-post sourcing and no independent eval keep it in the 78–84 band.

Bloomberg Technology

Samsung Takes Lead in Shipping Top-End AI Memory Chip Samples

Samsung Electronics has begun shipping samples of its most advanced memory to customers for AI accelerators from companies including Nvidia; the RSS snippet does not disclose the chip model, customer list, sample volume, pricing, or mass-production timeline.

Why it matters: HKR-H/K/R pass on the Samsung AI-memory supply-chain angle, but the facts stop at sample shipments; no model, customer list, or production window keeps it at the featured floor.

AI HOT (Curated Pool)

Apple reportedly tries to fit Google's large Gemini model into iPhone for new Siri

Apple is trying to integrate a large Gemini model into the iPhone for new Siri features. The RSS snippet says full local processing is unlikely because of model size, and a cloud component is likely required; the post does not disclose parameters, latency targets, or a release timeline.

Why it matters: HKR-H/K/R all pass, but this is a reported Apple-Google Siri effort, not a shipped product. The post gives distillation and likely cloud dependency, but no timeline, size, or tests, so it stays at the featured threshold.

TechCrunch · AI

Anthropic releases Opus 4.8 with new Dynamic Workflows tool

Anthropic released Opus 4.8 with a Dynamic Workflows tool for coordinating swarms of subagents. The RSS snippet does not disclose pricing, context window size, benchmarks, or a rollout schedule.

Why it matters: HKR-H/K/R all pass: an Anthropic model release plus an agent orchestration tool fits the 85–94 same-day band. Missing price, context window, and rollout detail keep it below the top of the band.

Hacker News front page

Dynamic Workflows in Claude Code

The title says Claude Code introduces dynamic workflows; the RSS body only provides the article URL, the Hacker News comments URL, 70 points, and 62 comments, and the post does not disclose the workflow mechanism, supported conditions, pricing, or release timing.

Why it matters: HKR-H and HKR-R pass: an official Claude Code update with HN discussion. HKR-K fails because the feed lacks mechanism, conditions, or limits, so this stays at the low featured threshold rather than 78+.

Hacker News front page

Claude Opus 4.8

The title names Claude Opus 4.8, while the RSS body only discloses the Anthropic URL, a Hacker News thread with 139 points, and 49 comments; the post does not disclose model parameters, capability changes, pricing, or release timing.

Why it matters: HKR-H/R pass: an official Claude Opus 4.8 page plus HN traction is a strong Claude-audience hook. HKR-K fails because capabilities, pricing, and context window are not disclosed, so it stays in 78–84, not P1.

The Verge · AI

A $2,000 AI-generated film will debut at Tribeca

Tribeca Festival will premiere the 75-minute AI-generated film Dreams of Violets, which cost $2,000 to make and uses people and images fully created by AI.

Why it matters: HKR-H/K/R all pass: a $2,000, 75-minute AI film entering Tribeca has novelty, concrete numbers, and labor resonance. It is a film-industry application, not a model or platform launch, so it stays in the low featured band.

May 28Thursday

AI HOT (Curated Pool)

Perplexity Computer Now Integrates with Microsoft Office

Perplexity Computer is now available in Microsoft Excel, Word, PowerPoint, and Outlook, letting users access Computer from the app sidebar to coordinate work, draft documents, model data, create presentations, and handle email.

Why it matters: HKR-H/K/R pass: Perplexity brings Computer into four Office apps with sidebar workflows, a useful product fact and competitive hook. Price, permission model, enterprise rollout, and measured results are not disclosed, so it stays at the featured threshold.

TechCrunch · AI

Sneak peek at new Siri app reveals Apple’s plans to take on ChatGPT and more

The title says a new Siri app preview shows Apple’s plan to compete with ChatGPT. The RSS snippet only discloses an iOS 27 AI overhaul, a redesigned Siri experience, and a standalone Siri app; it does not disclose launch timing or feature parameters.

Why it matters: HKR-H/R are strong because Apple is testing a standalone Siri app against ChatGPT. HKR-K is thin: iOS 27 AI revamp and redesign are disclosed, timing and specs are not, so it sits just above featured.

QbitAI · WeChat

Behind DeepSeek V4's Chip-Model Co-Design, China's Compute Ecosystem Gains Speed

QbitAI says DeepSeek V4 validated Ascend chip-model co-design, with CANN open-sourcing 65 repositories and supporting day-zero adaptation for more than 70 mainstream models, while AIGCode reported 65% MFU in MoE pretraining on Ascend.

Why it matters: HKR-H/K/R all pass, but this is mainly a compute-ecosystem progress story, not a DeepSeek V4 capability release. Concrete repo, adaptation, and MFU numbers lift it into featured, below must-write.

r/LocalLLaMA

Zai replaced the network architecture for GLM-5.1 inference, lifting throughput 15%

Zai replaced the ROFT network topology with ZCube on a thousand-GPU GLM-5.1 coding inference cluster, keeping the same GPUs, software stack, and model; the Reddit post cites 33% lower switch and optical module costs, 15% higher GPU inference throughput, and a 40.6% drop in first-token P99 tail latency under prefill-decode disaggregated inference.

Why it matters: HKR-H/K/R all pass: the GLM-5.1 inference cluster has concrete cost, throughput, and P99 latency numbers. Reddit single-source sourcing and infra-niche scope keep it at 78.

AI HOT (Curated Pool)

Mistral AI Releases Search Toolkit

Mistral AI released the public preview of Search Toolkit, an open source framework that combines data ingestion, retrieval, and evaluation behind shared interfaces for cloud, on-premises, or edge deployment.

Why it matters: HKR-K and HKR-R pass: Mistral combines RAG ingestion, retrieval, and evaluation in an open-source Search Toolkit with cloud, local, and edge deployment. HKR-H is weak, so this sits at the featured threshold for a mid-weight product update.

Mistral AI

Mistral upgrades Le Chat into unified agent Vibe, covering office work and coding

Mistral upgraded Le Chat into a unified AI agent called Vibe, with one license covering both office work and coding. Existing chats, settings and plans all carry over. Work Mode supports enterprise knowledge search, structured data analysis, document and report generation, scheduled multi-step tasks and reusable skills, and connects to Google Workspace, Outlook, SharePoint, Slack, GitHub and more.

Why it matters: It discloses Vibe's Work Mode, coding mode and CLI updates in full, so readers can judge how it plugs into existing workflows.

AI HOT (Curated Pool)

AI Now Summit 2026

Mistral AI announced industrial AI work, a Vibe upgrade, and a 10 MW inference data center in Les Ulis at AI Now Summit 2026; it is working with Airbus, BMW Group, and ASML, and the data center is scheduled to start operating in Q3 2026.

Why it matters: HKR-H/K/R pass: Mistral gives a concrete 10 MW inference site, Q3 2026 timing, and major industrial partners. No new model capability or pricing is disclosed, so it stays just above the featured threshold.

The Verge · AI

YouTube Will Let You Ask AI to Make a Custom Video Feed

YouTube is rolling out an AI custom video feed that uses a user-entered prompt or suggested options to build a personalized homepage feed around interests, moods, or topics; the feature supports English first and is available to signed-in users in the US on YouTube’s mobile app and desktop site.

Why it matters: HKR-H/K/R pass: promptable YouTube feeds are a concrete consumer-AI UX shift with rollout conditions. No model, ranking mechanism, or performance data is disclosed, so this stays in the mid-weight product-update band.

Financial Times · Technology

Kirkland & Ellis to Spend $500mn Building Its Own AI Technology

Kirkland & Ellis plans to spend $500mn building its own AI technology platform for the “collective intelligence” of its lawyers; the post does not disclose model architecture, vendors, or launch timing.

Why it matters: FT source authority and a $500mn in-house AI budget give strong HKR-H/K/R for legal AI adoption. Missing model architecture, vendors, and launch timing keep it in the 72–77 enterprise-adoption band.