Skip to content

#产品更新

25 today

May 19Tuesday

AI HOT (Curated Pool)

Cursor releases Composer 2.5 coding model

Cursor released Composer 2.5, claiming up to 10x higher efficiency on long coding tasks; the model is further trained on Moonshot’s Kimi K2.5 and uses text feedback for 100k-token-scale trajectories.

Why it matters: HKR-H/K/R all pass: Cursor is a core AI coding tool, and Composer 2.5 adds concrete claims around 10x long-task gains and Kimi K2.5 tuning. Limited sourcing and no independent eval keep it in the 78–84 band.

AI HOT (Curated Pool)

Take your local GitHub sessions anywhere

GitHub launched remote control sessions for Copilot, letting users start tasks in VS Code or the command line and continue them through github.com or GitHub Mobile.

Why it matters: GitHub Copilot session handoff from VS Code/CLI to web and mobile clears HKR-H/K/R, but the post only gives entry points and use case; permissions, pricing, and supported task scope are not disclosed.

AI HOT (Curated Pool)

NVIDIA fine-tunes Cosmos Predict 2.5 with LoRA/DoRA for robot video generation

NVIDIA published a Hugging Face post on fine-tuning Cosmos Predict 2.5 with LoRA and DoRA to generate robot first-person videos from text prompts; the post does not disclose dataset size, training cost, or evaluation results.

Why it matters: HKR-H/K/R pass: the robot POV video angle is clickable, and LoRA/DoRA on Cosmos Predict 2.5 is a concrete mechanism. Missing dataset scale and metrics keep it in the low featured band.

May 18Monday

AI HOT (Curated Pool)

Baidu Core AI Business Revenue Exceeded RMB 13.6B in Q1

Baidu reported that its core AI-driven business generated more than RMB 13.6 billion in Q1 2026, up 49% year over year, and accounted for more than half of Baidu’s general business revenue for the first time.

Why it matters: HKR-H/K/R all pass, but the source is a Baidu post with revenue framing only; segment breakdown and margin quality are not disclosed. This fits the low featured band as an AI commercialization signal.

Bloomberg Technology

Baidu AI Sales Eclipse Waning Legacy Ads for the First Time

Baidu reported a 1% revenue decline as growth in nascent AI businesses offset shrinking traditional internet revenue; the post does not disclose AI sales, advertising revenue, or details of the agentic AI pivot.

Why it matters: Baidu revenue fell 1% while AI sales topped legacy ads for the first time, so HKR-H/K/R pass. Missing AI/ad dollar splits and agentic-AI mechanics keep it in the low featured band, not p1.

AI HOT (Curated Pool)

Grok Now Supports Video Understanding and Analysis

Grok now supports full-video uploads for real-time analysis, summarization, translation, scene explanation, and context extraction; the post does not disclose duration limits, supported formats, or rollout scope.

Why it matters: HKR-H/K/R all pass, but duration limits, formats, and rollout scope are not disclosed, so this stays at the featured threshold for a mid-weight product update.

Synced · WeChat

openJiuwen releases JiuwenSwarm, an open-source multi-agent swarm framework

openJiuwen released and open-sourced JiuwenSwarm with four components: Agent Swarm, Swarm Skills, Swarm Skills Hub, and self-evolving Swarm Skills, and reports a 94.2% PinchBench score versus 91.6% for OpenClaw.

Why it matters: HKR-H/K/R all pass: an open-source agent-swarm framework with named components and a PinchBench 94.2% claim. It stays at 78 because openJiuwen is not a top lab and the summary lacks license, reproduction setup, and baselines.

AI HOT (Curated Pool)

Tencent AI Design Agent Ardot Enters Public Beta: Generates Editable Designs and Converts Them to Code

Tencent Cloud opened public beta for Ardot, an AI design agent that generates editable app pages, websites, and posters from one-sentence prompts, then converts designs to code.

Why it matters: HKR-H/K/R pass on a concrete Tencent product beta for editable design-to-code workflows. Missing pricing, model details, benchmarks, and field results keep it at the lower featured threshold.

AI HOT (Curated Pool)

Alibaba Cloud launches HappyHorse video generation model

Alibaba Cloud launched HappyHorse on Model Studio, with prompt-to-1080p multi-shot video generation in one workflow; the post lists a limited-time 20% discount but does not disclose pricing, model parameters, or availability terms.

Why it matters: HKR-H/K/R pass on the named model, 1080p multi-shot capability, and cost/competition angle. Thin disclosure on price, parameters, and benchmarks keeps it near the featured threshold.

AI HOT (Curated Pool)

Grok launches Skills feature

xAI launched Grok Skills on May 18, 2026, letting users set preferences, formatting rules, or workflows once and keep them active across all conversations on web, iOS, and Android.

Why it matters: HKR-H/K/R all pass: Grok Skills adds persistent preferences and workflows across web, iOS, and Android. This is a mid-weight xAI product update; rollout scope, limits, and pricing are not disclosed.

AI HOT (Curated Pool)

Composer 2.5 release and technical analysis

Cursor released Composer 2.5, built on a Moonshot open-source checkpoint, trained with synthetic data from real codebases at 25 times the previous scale, and updated with text-feedback reinforcement learning and a sharded Muon optimizer.

Why it matters: HKR-H/K/R all pass: Cursor is a core coding-agent surface, and the post gives concrete training details around Moonshot, 25x data, RL, and Muon. It lacks benchmarks, pricing, or user-facing capability limits, so it stays in the 78–84 band.

Google DeepMind

Google DeepMind adds Street View grounding to Project Genie

Google DeepMind has added Street View real-scene grounding to its experimental prototype Project Genie. Users can pick a US location, then pair it with a style and characters to generate a world.

Why it matters: With Street View imagery wired in, agents and robots can train and navigate in virtual environments that track real places.

Google DeepMind

Introducing Google Antigravity 2.0

Google 发布智能体开发平台 Google Antigravity 2.0。该平台在 Google DeepMind 官网被列为面向开发者的 agentic development platform,与 Gemini 应用、Google AI Studio 并列。原文未披露版本功能、参数或可用性细节。

May 17Sunday

Bloomberg Technology

Apple’s New ChatGPT-Like Siri App Will Have Auto-Deleting Chats

The title says Apple’s ChatGPT-like Siri app will support auto-deleting chats; the RSS snippet only adds that iOS 27 will include a Genmoji upgrade, and the post does not disclose retention periods, release timing, or feature details.

Why it matters: HKR-H and HKR-R pass because Bloomberg frames a specific Apple Siri privacy angle; HKR-K fails since retention and feature mechanics are missing, so this stays at the low featured threshold.

Google DeepMind

Google DeepMind launches Gemini for Science toolset

Google DeepMind released Gemini for Science, which includes three experimental tools on Google Labs: Hypothesis Generation, built on Co-Scientist.

Why it matters: Google is packaging research prototypes like Co-Scientist and AlphaEvolve into apply-to-use science tools, showing what agentic research looks like in practice.

Google DeepMind

Google expands content provenance and verification tools across Search, Gemini, Chrome and Pixel

Google is widening its content transparency and verification tools across Search, Gemini, Chrome, Pixel and Cloud, and deepening industry partnerships. SynthID has watermarked over 100 billion images and videos plus 60,000 years of audio. SynthID verification in the Gemini app has been used 50 million times, and the capability reaches Search today, with Chrome in the coming weeks.

Why it matters: The post lays out where SynthID and C2PA land across Search, Gemini, Chrome and Pixel, which shows the current limits of content provenance tools.

QbitAI · WeChat

A Robot Dog Challenges Nvidia's Compute Lead

Weilan Technology unveiled BabyAlpha A3, a consumer quadruped robot using a six-chip heterogeneous cluster that runs a 7B-parameter model on-device at 280 TPS; the article says it has 66MP vision, 2.232 million point-cloud samples per second, and a planned Q3 launch.

Why it matters: HKR-H/K/R pass: the robot-dog-versus-Nvidia angle is clickable, and 280 TPS on a local 7B model is concrete. Single-source summary lacks price, power draw, and benchmark setup, so it stays near the featured floor.

AI HOT (Curated Pool)

Grok Imagine image generation is officially released

Grok Imagine is now available on X for all users, with text-to-image generation for realistic images and multiple aspect ratios; the post does not disclose model parameters, pricing, or regional limits.

Why it matters: HKR-H/K/R pass, but the post only discloses availability and basic image features; model details, pricing, and regions are absent, so this lands at the featured threshold.

AI HOT (Curated Pool)

MagicPath Integrates with Codex to Combine Design and Development

MagicPath AI CEO @skirano demonstrated MagicPath running inside Codex as a native canvas, with users configuring it through one command, dragging UI elements, and letting Codex generate and edit code in real time.

Why it matters: HKR-H/K/R pass: MagicPath puts a draggable design canvas inside Codex with one-command setup and live code edits. Single-demo sourcing and missing framework support, permissions, and reproducible cases keep it at the lower featured band.

May 16Saturday

AI HOT (Curated Pool)

Codex adds multi-device remote control and shared context

Codex controls multiple devices through ChatGPT, switches by project to access each device’s context and files, and supports remote SSH setup for other VMs.

Why it matters: HKR-H/K/R all pass, but the item is a thin X-post summary with no official release note, pricing, permission model, or reproducible demo. Treat it as a mid-weight coding-agent product update at the featured threshold.

AI HOT (Curated Pool)

OpenAI Restructures as Brockman Takes Over Product Strategy

OpenAI merged ChatGPT, Codex, and API into one product organization, with Greg Brockman taking over product strategy; the post says Anthropic’s valuation reached $900 billion, but it does not disclose the restructuring timeline.

Why it matters: HKR-H/K/R all pass: this is an OpenAI top-level product reorg covering ChatGPT, Codex, and API. Single-source summary keeps it below the highest band, but it is same-day must-write news.

QbitAI · WeChat

Codex Integrates HeyGen for Prompt-Based Video Generation and Editing

Codex integrates the HeyGen plugin to run image generation, talking-avatar video, subtitles, and edits from natural-language prompts; the article tests roughly one-minute avatar generation, trimming content after 10 seconds, and deleting a blink at the eighth second.

Why it matters: HKR-H/K/R all pass, backed by a numbered hands-on test. The scope is still one Codex-to-HeyGen plugin workflow, not a model or platform release, so it lands in the 72-77 featured band.

QbitAI · WeChat

A new AI for 5 million doctors in China: exclusive journal partnership focuses on evidence sources

Alibaba Health launched the medical AI product Qinglizi for China’s 5 million doctors, with access to ten years of content from 70 BMJ Group journals and an evidence workflow constrained by PICO, GRADE, and review from more than 300 clinical experts.

Why it matters: HKR-H/K/R all pass: Alibaba Health and BMJ add concrete evidence sources and review mechanisms to a medical AI product. It remains a vertical product/partnership update, not a foundation-model or platform release.

Computing Life · Share · Yage

OpenAI Reaches Into Your Bank Account

OpenAI uses Plaid to let ChatGPT connect to bank accounts; the post does not disclose launch timing, authorization flow, or the exact data scope ChatGPT can access.

Why it matters: HKR-H/R are strong and HKR-K passes via the Plaid integration mechanism. Missing launch timing, authorization flow, and data scope keep it at the featured threshold rather than a higher OpenAI product-update score.

AI HOT (Curated Pool)

Runway Agent Generates Complete Ads in One Session

Runway Agent turns product photos and ideas into fully produced ads in one session; the post does not disclose the model, pricing, generation length, or regional availability.

Why it matters: Runway’s ad-generation Agent clears HKR-H/K/R as a mid-weight product update. Missing model, pricing, duration, and region details keep it at the featured threshold, not a must-write release.

The Verge · AI

OpenAI now wants ChatGPT to access your bank accounts

OpenAI previewed a ChatGPT feature that lets users connect financial accounts through Plaid, which links to 12,000 institutions. OpenAI says more than 200 million people ask ChatGPT finance questions each month; the post does not disclose a general release date.

Why it matters: HKR-H/K/R all pass: OpenAI is moving ChatGPT toward real financial-account access, with Plaid’s 12,000 institutions and 200M monthly finance askers as concrete facts. It stays below 85 because launch timing is not disclosed.

TechCrunch · AI

OpenAI launches ChatGPT for personal finance, will let users connect bank accounts

OpenAI launched ChatGPT for personal finance, and connected users can view portfolio performance, spending, subscriptions, and upcoming payments; the RSS snippet does not disclose supported banks, launch regions, pricing, or account-security terms.

Why it matters: HKR-H is strong because ChatGPT connects to bank accounts; HKR-K has concrete finance features; HKR-R hits privacy and fintech competition. Banks, regions, and pricing are undisclosed, so this stays in the low P1 band.

May 15Friday

r/LocalLLaMA

Fully Offline Suitcase Robot Built Around Jetson Orin NX SUPER 16GB

CreativelyBankrupt built Sparky as a fully offline suitcase robot on Jetson Orin NX SUPER 16GB, running Gemma 4 E4B Q4_K_M via llama.cpp with q8_0 KV cache, about 200 ms cached TTFT, 14-15 tok/s sustained output, 12K context, 30+ sensors, and no WiFi, Bluetooth, or cellular interface.

Why it matters: HKR-H/K/R all pass, with a named hands-on build and concrete latency/sensor numbers. It stays in low featured because this is a Reddit project post, not a product launch or research release.

MIT Technology Review · AI

The Download: China’s AI Drama Factory and the WHO’s Missing Health Targets

China’s short-drama industry released an average of 470 AI-generated short dramas per day in January, while production timelines fell from months to weeks and costs dropped by up to 90%.

Why it matters: MIT Technology Review provides concrete output, cycle-time, and cost figures for China’s AI short-drama pipeline, clearing HKR-H/K/R. The story is application-layer, not a core model or product release, so it sits at the featured threshold.

MIT Technology Review · AI

How Chinese Short Dramas Became AI Content Machines

Chinese short-drama companies are using AI for full-series production, with DataEye counting an average of 470 AI-generated short dramas released per day in January 2026, while FlexTV says production time fell from three to four months to under one month and North American per-series costs can drop by 80% to 90%.

Why it matters: HKR-H/K/R all pass: the story has a strong content-factory hook, concrete production metrics, and clear labor/cost resonance. It is a quality industry feature, not a model or platform release, so 80 fits the 78-84 band.

Alibaba Technology · WeChat

Qoder 1.0 launches as an agentic development workspace beyond AI IDE

Alibaba released Qoder 1.0 with downloads for Windows, macOS, and Linux, adding a standalone Quest workspace, cross-project parallel agent tasks, a team knowledge engine, and Experts mode with five roles for planning, research, coding, review, and testing.

Why it matters: Alibaba’s Qoder 1.0 is a mid-weight AI coding product release with concrete agent-workflow features and developer resonance. No pricing, benchmark, or task-success data is disclosed, so it stays near the featured threshold.

AI HOT (Curated Pool)

Codex lands on mobile with preview in the ChatGPT app

OpenAI brought Codex to a preview inside the ChatGPT mobile app; the post does not disclose supported platforms, feature scope, pricing, or rollout schedule.

Why it matters: Official OpenAI product update with HKR-H/K/R, but detail is thin. The post does not disclose platform support, feature scope, pricing, or rollout timing, so it sits at the featured threshold.

AI HOT (Curated Pool)

Databricks brings GPT-5.5 to enterprise agent workflows

Databricks made GPT-5.5 available through AI Unity Gateway for AgentBricks and Agent Supervisor API workflows; on OfficeQA Pro, it became the first model above 50% accuracy and reduced errors by 46% versus GPT-5.4.

Why it matters: HKR-H/K/R all pass: GPT-5.5 enters Databricks workflows with 50% OfficeQA Pro accuracy and 46% fewer errors than GPT-5.4. It stays below a full model-release score because the page is a sales-led OpenAI customer story using Databricks’ own benchmark.

AI HOT (Curated Pool)

Connect Grok to the Hermes Agent

xAI connects Grok subscription accounts to Nous Research’s open-source Hermes Agent across all subscription tiers, letting users run Grok 4.3 text chat and reasoning, generate spoken replies with text-to-speech, create images and videos with Grok Imagine, and connect the agent to WhatsApp or Discord.

Why it matters: HKR-H/K/R all pass, but this is a mid-weight xAI product integration with an open-source agent, not a flagship model release. Featured fits; it does not clear the 85+ same-day bar.

AI HOT (Curated Pool)

ChatGPT launches personal finance experience

OpenAI launched a personal finance preview for ChatGPT Pro users in the US, letting users connect financial accounts and receive analysis based on their finances, goals, and priorities.

Why it matters: HKR-H/K/R all pass: OpenAI is moving ChatGPT into personal finance with linked financial accounts. The score stays at 76 because it is a US Pro preview, with partners, controls, and pricing not disclosed.

AI HOT (Curated Pool)

Claude Agent Tool v2.1.142 Release

Claude Agent Tool v2.1.142 adds eight command-line flags for configuring background sessions, upgrades Fast mode’s default model to Opus 4.7, and fixes more than 15 issues including MCP tool timeouts and Windows network-drive deadlocks.

Why it matters: HKR-H/K/R all pass: this is a small Claude Code release, but the Opus 4.7 Fast-mode default, 8 session flags, and 15+ fixes affect daily dev workflows. Anthropic tool-chain relevance keeps it at the featured floor.

Latent Space

AI-Native Healthcare: 100M Doctor Visits, 10–20 Hours Saved, Prior Auth in Minutes

Abridge says it is projected to support 80M+ patient-clinician conversations this year across 250 large U.S. health systems, 28+ languages, and 50+ specialties, while its clinical documentation workflow reduces clinicians’ documentation burden by 10–20 hours per week.

Why it matters: HKR-H/K/R all pass: the story has a strong scale hook, concrete adoption metrics, and workflow ROI. Claims are company-interview sourced, not an independent benchmark or major platform release, so it sits in low featured.

AI HOT (Curated Pool)

Codex adds automation hooks and programmatic tokens

Codex added hooks and programmatic access tokens: hooks run scripts at key task stages for validation, secret scanning, logging, or repo-specific behavior, while scoped tokens for Business and Enterprise teams support CI/CD, release workflows, and internal automation with expiration or revocation.

Why it matters: HKR-H/K/R all pass: Codex gains concrete automation hooks and programmatic tokens for CI/CD. Score stays in the 72–77 band because the post discloses workflow fit, not pricing, permission detail, or impact data.

Bloomberg Technology

Musk’s xAI Unveils First Coding Agent in Bid to Rival Anthropic

xAI is rolling out its first AI coding agent, Grok Build, for software development workflows; the RSS snippet names Anthropic’s Claude as the rival but does not disclose pricing, availability, benchmarks, or supported IDEs.

Why it matters: HKR-H and HKR-R pass: xAI entering coding agents is a strong competitive hook for developers. HKR-K fails because pricing, availability, and benchmarks are not disclosed, so this stays at the low end of a mid-weight product update.

The Verge · AI

OpenAI’s Codex is now in the ChatGPT mobile app

OpenAI will let users access Codex from the ChatGPT mobile app; the RSS snippet says Codex can write code and use apps on a computer, but the post does not disclose launch timing, pricing, or the full mobile feature scope.

Why it matters: OpenAI added Codex access to ChatGPT mobile, a mid-weight product update. HKR-H/K/R pass through the mobile coding-agent hook, concrete app-control claim, and developer workflow nerve; missing timing, pricing, and support scope keep it at the featured floor.