Skip to content

#MCP/工具调用

5 today

May 22Friday

AI HOT (Curated Pool)

Aleph 2.0 and Edit Studio

Runway released Aleph 2.0 and Edit Studio, combining generation, editing, and post-production into one platform; the post does not disclose pricing, technical parameters, or rollout scope.

Why it matters: Runway is a major AI video vendor, and Aleph 2.0 plus Edit Studio is a mid-weight product update. HKR-H/K/R pass, but missing price, specs, and rollout keep it at the featured threshold.

AI HOT (Curated Pool)

Claude now supports more security and compliance tools

Anthropic added 28 security and compliance integrations for Claude Enterprise and its platform, using the Claude Compliance API to provide conversation content and activity events to DLP, SIEM, and existing enterprise monitoring workflows.

Why it matters: Official Anthropic product update with 28 compliance integrations and Compliance API event routing, so HKR-K/R pass. It is enterprise governance rather than a model capability jump, keeping it near the featured threshold.

AI HOT (Curated Pool)

Kotlin ADK and Android ADK 0.1.0 Released for Building AI Agents

Google released Kotlin ADK and Android ADK 0.1.0 for developers, with Kotlin ADK targeting backend agent workflows and Android ADK providing mobile-specific functions for building AI agents.

Why it matters: Google’s Kotlin ADK and Android ADK 0.1.0 release is a mid-weight agent tooling update. HKR-H/K/R pass, but the disclosed facts stop at platforms and version, with no performance data, examples, or ecosystem scale.

Hacker News front page

Launch HN: Runtime (YC P26) – Sandboxed coding agents for everyone on a team

Runtime launched open-source sandbox infrastructure for coding agents, supporting Claude Code, Codex, Cursor, Copilot, Gemini, and Devin, with hosted access, a free tier, and pricing based on a flat platform fee plus compute without token markup.

Why it matters: HKR-H/K/R pass: this is not a major-lab launch, but open-source sandboxes, six coding-agent types, and no token markup give teams concrete adoption signals. No usage data or marquee customers keeps it near the featured floor.

May 21Thursday

r/LocalLLaMA

LLM planner: pick a rig by use case, model, or budget, or pick models for your rig

totosse17 published the LLMRequirements hardware planner with 60+ build configs, 50+ models, 130 cited tokens-per-second sources, 150+ reviewer videos, multi-region prices, idle and active watts, and a public GitHub data repo.

Why it matters: HKR-H/K/R all pass, but this is a Reddit community tool for local LLM rigs, not a broad platform release. The concrete dataset earns a featured-threshold score, not the 78+ band.

The Verge · AI

I Can’t Believe How Fast Google Vibe Coded My First Android App

The Verge’s Sean Hollister used Google AI Studio to generate three Android apps in one afternoon; one app came from a 148-word browser prompt and installed about 10 minutes later on an Android phone prepared with USB debugging and a PC connection.

Why it matters: HKR-H/K/R all pass: the story has a personal-test hook plus concrete timing and prompt details. This is not a major Google launch, so it fits the high-quality first-person experiment band, not same-day must-write.

r/LocalLLaMA

Tencent Hy-MT2 30B/7B/1.8B

Tencent released Hy-MT2 translation models in 1.8B, 7B, and 30B-A3B sizes, supporting translation across 33 languages; AngelSlim 1.25-bit quantization reduces the 1.8B model’s storage requirement to 440 MB and raises inference speed by 1.5x.

Why it matters: HKR-H/K/R pass via the 440MB quantized 1.8B model, 33-language support, and local inference cost angle. Sparse Reddit sourcing keeps it at the featured threshold, not the 78+ band.

AI HOT (Curated Pool)

Lessons from Building Cloud Agents

Cursor summarizes lessons from building cloud agents: after migrating to Temporal, reliability rose above 99.9%, and the platform processes more than 50 million operations per day.

Why it matters: HKR-H/K/R all pass: Cursor is central to coding agents, and the post gives Temporal, 99.9%+ reliability, and 50M daily operations. Not a launch, so it stays at low-end featured.

Alibaba Technology · WeChat

Building an Agent from 0 to 1: Principles and Personal Assistant Practice

Zhan Xupeng published a roughly 50-minute article on Agent theory and a personal assistant implementation, covering memory, ReAct planning, progressive skill loading, subagents, and harness-level fault recovery.

Why it matters: HKR-K/R pass via concrete agent mechanisms and practitioner reliability pain; HKR-H is weak because the headline is a standard tutorial frame. This fits the quality-tutorial threshold, not the 78+ news band.

Xinzhiyuan · WeChat

Anthropic Acquires SDK Toolmaker Stainless, Leaving OpenAI and Others to Maintain SDKs

Anthropic has completed its acquisition of Stainless, an SDK generation company used by OpenAI, Anthropic, Meta, Cloudflare, and other infrastructure vendors; Stainless says prior SDK ownership remains with customers, but it will shut down hosted products including SDK generator and stop providing ongoing support.

Why it matters: HKR-H/K/R all pass: the deal targets API SDK generation, names OpenAI/Meta/Cloudflare as customers, and says hosted products will shut down. Anthropic bump applies, but this is not a model or core capability release, so it fits 78–84.

r/LocalLLaMA

Moved from prompt-based output validation to schema-enforced execution, with significant reliability gains

A Reddit user tested Claude structured outputs and reported 90–95%+ first-pass parse rates with tool_use, typed schemas, enum constraints, and stepwise validation, versus 65–70% for prompt instructions followed by regex or JSON parsing and retries.

Why it matters: HKR-H/K/R all pass: the post has a clear reliability contrast and concrete 90–95%+ vs 65–70% numbers. Source authority is limited to one Reddit experiment, with sample and task details not disclosed, so it stays at low featured.

AI HOT (Curated Pool)

Equipping AI with a Scientific Toolkit to Accelerate Research Workflows

Google DeepMind released Science Skills for Google Antigravity, integrating insights from more than 30 life-science sources, including the UniProt and AlphaFold databases.

Why it matters: Google DeepMind has a concrete product update: Science Skills adds 30+ life-science sources to Antigravity, hitting HKR-H/K. It is a vertical toolkit rather than a model release, so it sits at the low featured band.

AI HOT (Curated Pool)

Tencent Launches OS-Level AI Assistant Mavis on Windows, Mac, and Android

Tencent launched the OS-level AI assistant Mavis on May 21 across Windows, Mac, and Android, with document parsing, image recognition, system maintenance, partial offline use, model dispatching, and desktop control of mobile apps listed as supported functions.

Why it matters: HKR-H/K/R all pass: Tencent’s OS-level assistant spans Windows, Mac, and Android with concrete tool abilities. Model, pricing, and permission design are not disclosed, so it stays at the lower featured band.

Latent Space

Railway: The Agent-Native Cloud — Jake Cooper

Railway serves 3 million users with a 35-person team, adds about 100,000 signups per week, has raised $124 million, and has moved most workloads to its own bare-metal data centers with a reported three-month payback versus rented cloud capacity.

Why it matters: HKR-H/K/R all pass: the Railway interview has concrete growth, funding, and bare-metal details tied to agent infrastructure. It is a strong practitioner story, not a core model or major AI product release, so 74 fits the featured threshold.

AI HOT (Curated Pool)

Google Stitch update: AI design assistant supports end-to-end building

Google updated its AI design partner Stitch with real-time streaming design builds, direct edits and feedback, codebase or Design.md imports, dynamic UI generation, shareable URL exports, and global availability.

Why it matters: HKR-H/K/R pass: Google Stitch adds streaming builds, codebase/Design.md import, and global access. It stays at the featured threshold because model details, pricing, and measured output quality are not disclosed.

AI HOT (Curated Pool)

ChatGPT mobile app adds Codex support for cross-device collaboration

OpenAI Devs says the ChatGPT mobile app now supports Codex, letting users ask questions on mobile and continue the same conversation on desktop; the post does not disclose supported platforms, app versions, or rollout scope.

Why it matters: OpenAI Devs is authoritative and HKR-H/K/R pass, but the post only confirms mobile Codex access and handoff; platform, version, and rollout scope are not disclosed, so it sits at the featured threshold.

The Verge · AI

Google Search’s AI Evolution Includes More Ads

Google is adding Gemini-generated product explanations to Search ads for product queries, including Sponsored Product placements and some ads with built-in chatbots; the snippet cites a compact espresso pod machine example, but the post does not disclose rollout scope or pricing.

Why it matters: HKR-H/K/R all pass: Google is inserting Sponsored Product units and ad chatbots into Gemini shopping results. Scope, pricing, and performance data are not disclosed, so this stays in low featured rather than a major product-release band.

May 20Wednesday

The Verge · AI

If Google Can’t Make AI Agents Useful, Maybe No One Can

The Verge says Google announced multiple AI agents at I/O 2026 for information gathering, event planning, and inbox or calendar summarization; the RSS snippet says the agents can run continuously in the background, but the post does not disclose launch timing, pricing, or evaluation results.

Why it matters: HKR-H and HKR-R are strong because the Verge frames Google agents as a sector test; HKR-K is limited to background-running agents. Missing launch timing, pricing, and evals keeps it in the lower featured band.

TechCrunch · AI

Figma adds an AI assistant to its collaborative canvas

Figma added an AI agent to its collaborative canvas, letting users use natural-language prompts to create new designs, edit existing ones, or automate tasks such as generating design iterations.

Why it matters: HKR-K and HKR-R pass: Figma puts an AI agent into a core collaborative design surface. Details stop at capability scope, with no model, pricing, or rollout timing, so this sits at the featured threshold.

AI Chat-Group Daily (群聊日报)

2026-05-19 Chat Group Daily

The chat group daily says Karpathy joined Anthropic's pretraining team, and cites Stainless shutting down hosted services after acquisition plus Google I/O announcing Gemini 3.5 Flash and a $100 subscription tier.

Why it matters: HKR-H/K/R all pass, but this is a chat-daily roundup with secondhand claims and no disclosed primary links, appointment details, or product specs, so it lands at the lower featured band.

AI HOT (Curated Pool)

OpenAI offers $2 million API investment to every YC startup

OpenAI is offering each startup in Y Combinator’s current batch $2 million in API credits in exchange for equity; the post does not disclose the equity stake, credit expiration, or usage limits.

Why it matters: HKR-H/K/R all pass: $2M per current YC startup in API credits for equity is concrete and talkable. Missing equity %, term and usage caps keep it at 78, below the 85+ must-write band.

AI HOT (Curated Pool)

Microsoft reportedly warns internally that GitHub faces existential risk as AI coding tools reduce hosting need

Microsoft internally warned that GitHub faces an existential risk from AI coding assistants such as Cursor and Claude Code, and told some teams to stop using Claude Code by the end of June 2026 and move to GitHub Copilot CLI.

Why it matters: HKR-H/K/R all pass: the angle is sharp, the summary gives a Claude Code-to-Copilot CLI deadline, and the workflow stakes are real. Single-source “reported” framing and no Microsoft response keep it below the 85 must-write band.

AI HOT (Curated Pool)

Qwen3.7: Agent Frontier

Qwen Studio released Qwen3.7 with chatbots, image and video understanding, and image generation. It also covers document processing, web search integration, tool calling, and artifact generation. The RSS snippet frames it as an agent-focused model, but the post does not disclose context length. It also omits benchmark scores, pricing, API limits, release schedule, and reproducible evaluation conditions.

Why it matters: HKR-H/K/R all pass: this is a Qwen flagship-model update with concrete capability coverage. Lack of benchmarks, pricing, and context-window details keeps it at the low end of the 85–94 band.

AI HOT (Curated Pool)

Gemini 3.5 Flash price rises sharply as Google plans broad rollout

Google released Gemini 3.5 Flash at I/O with $1.50 per million input tokens and $9 per million output tokens, making it 3x and 6x the previous model’s pricing, while adding roughly 1 million input tokens and about 65,000 maximum output tokens.

Why it matters: HKR-H/K/R all pass: a Google model update with a sharp pricing twist and concrete token costs. It stays below p1 because the body only gives price and rollout intent, not capability deltas, benchmarks, or context window.

AI HOT (Curated Pool)

Gemini launches personal AI agent and Daily Brief

Gemini added Gemini Spark and Daily Brief: Spark acts as an always-on personal AI agent across Gmail, Google Docs, and Slides after user authorization, while Daily Brief is available to Google AI subscribers in the U.S. aged 18 or older.

Why it matters: HKR-H/K/R all pass: Google is adding Gemini Spark’s authorized actions across Gmail, Docs, and Slides, plus Daily Brief eligibility for US 18+ AI subscribers. This is a same-day Google agent product update.

AI HOT (Curated Pool)

Claude Code’s HTML Output: The Unreasonable Effectiveness of HTML

The Claude Code team is shifting its primary output format from Markdown to HTML, and the post names four mechanisms: tables, CSS styling, SVG charts, and JavaScript interactions.

Why it matters: Official Claude Code post with a concrete shift from Markdown to HTML and 4 output mechanisms; strong practitioner utility, but not a major product launch, so it sits in the featured threshold band.

TechCrunch · AI

You Can Now Talk to Your Gmail Inbox, as Seen at Google I/O 2026

Google expanded Gmail’s AI Inbox with conversational voice search, letting users ask Gemini to find details buried in email. The RSS snippet does not disclose rollout scope, supported languages, pricing, latency, or the retrieval mechanism behind Gmail search.

Why it matters: HKR-H/K pass: a Google-scale Gmail voice inbox feature is concrete and clickable. HKR-R is weak because rollout, language support, pricing, and retrieval mechanics are not disclosed.

The Verge · AI

Google’s AI Future Demands Trust — and Your Personal Data

Google presented Gemini Spark, Daily Brief, and expanded Gmail AI inbox access at I/O 2026; the Verge snippet says these tools depend on large amounts of personal information, but the post does not disclose detailed data-handling terms.

Why it matters: HKR-H/K/R all pass, but the body gives product names and a personal-data dependency without data-handling details. Google I/O makes it featured, not a same-day must-write.

AI HOT (Curated Pool)

Production Guide for Claude Operating Real User Interfaces

ClaudeDevs published a production guide for Claude computer use, and the snippet lists four mechanisms: click accuracy, thinking effort level selection, context retention in long sessions, and replayable demonstration logging.

Why it matters: HKR-H/K/R all pass: a practical Claude UI-control guide with 4 concrete mechanisms. It is not an official model or product release, so it fits the quality-tutorial band rather than same-day must-write.

AI HOT (Curated Pool)

Smarter Google AI Edge Gallery: MCP Integration, Notifications, and Session Continuity

Google AI Edge Gallery adds experimental MCP support on Android, letting Gemma 4 coordinate external data sources including Google Workspace and Google Maps; the update also adds scheduled notifications and persistent chat history for faster restoration of long-session context.

Why it matters: HKR-H/K/R all pass: Google’s developer update adds experimental MCP, notifications, and session continuity to AI Edge Gallery. It is a mid-weight product update, not a model release or major capability launch.

AI HOT (Curated Pool)

Google Tensor ML SDK Beta Released

Google released the Tensor ML SDK beta, letting developers convert, compile, and run PyTorch or TFLite models on Pixel 10 TPUs through LiteRT, with a model library containing more than 100 classic and generative AI models, including Gemma 3.

Why it matters: HKR-K is strong: the post gives a concrete Pixel 10 TPU workflow and a 100+ model library. HKR-H/R clear the featured bar, but this is a beta developer SDK rather than a flagship model or major consumer launch.

TechCrunch · AI

Google takes a page from Meta, announces audio-powered smart glasses at I/O 2026

Google announced “audio glasses” at I/O 2026, letting users issue voice commands across its apps and services, including Gemini; the RSS snippet does not disclose price, launch timing, or hardware specifications.

Why it matters: HKR-H/K/R pass: Google announced Gemini-linked audio glasses at I/O 2026, a credible AI-hardware platform move. Missing price, launch date, and specs keep it in the low featured band.

AI HOT (Curated Pool)

Empirical Research Assistant ERA: From Nature Publication to Computational Discovery

Google Research published its Gemini-based Empirical Research Assistant in Nature and opened early access through the Google Labs trusted tester program.

Why it matters: HKR-H/K/R all pass: Google moves Gemini-based ERA from a Nature paper to a Labs trusted-tester trial. Score stays at 78 because the provided text lacks metrics, benchmark setup, or reproducible workflow details.

Hacker News front page

Gemini CLI will stop working from June 18, 2026

Google Developers Blog says Gemini CLI will stop working on June 18, 2026, and the post title points to a transition to Antigravity CLI; the RSS snippet only includes the article URL, Hacker News comments URL, 36 points, and 10 comments, and it does not disclose migration steps, compatibility details, or replacement behavior.

Why it matters: Google’s developer blog gives a Gemini CLI shutdown date and migration target, clearing HKR-H/K/R. Detail is thin beyond the deadline, so it stays in the low featured band.

AI HOT (Curated Pool)

Google Search gets its biggest redesign in 25 years with AI-driven interaction changes

Google announced at I/O 2026 the biggest Search redesign in 25 years, powered by Gemini 3.5 Flash, with AI Mode exceeding 1 billion monthly active users and query volume doubling each quarter.

Why it matters: Google Search is an internet entry-point product; 1B+ AI Mode MAU and quarterly query doubling put this beyond a routine feature update. HKR-H/K/R all pass, so it lands in p1.

TechCrunch · AI

Google Search as You Know It Is Over

Google is changing Search from a link list into an AI experience with conversational answers, autonomous agents, and interactive interfaces; the RSS snippet does not disclose a rollout timeline, publisher traffic impact numbers, or product parameters.

Why it matters: HKR-H/K/R all pass: Google is recasting Search around AI answers, agents, and interactive UI. The post lacks rollout timing and traffic numbers, but this is still a major Google core-product update.

The Verge · AI

Gmail is going to start talking to you

Google is launching Gmail Live for Gmail, letting users tap a search-bar icon and ask voice questions about inbox content; a press demo retrieved school event dates, locations, and an upcoming Detroit trip from the employee’s email.

Why it matters: HKR-H/K/R pass: Gmail Live adds voice email queries inside a mass-market Google surface. The post gives demo cases, but no launch date, pricing, or model details, so it stays at the lower featured band.

The Verge · AI

Would You Let Robots Spend Your Money? Google Is Betting on It

Google unveiled an AI shopping Universal Cart at I/O that lets users add products while browsing Search or chatting with Gemini, then check out through Google; the RSS snippet says future support includes YouTube and Gmail, while pricing, rollout timing, and retailer coverage are not disclosed.

Why it matters: HKR-H/K/R all pass: the hook is AI agents spending money, the new fact is Search/Gemini checkout via Google, and the nerve is agent payment safety. This is a mid-weight Google I/O product update, so 76, not P1.

The Verge · AI

Google can now vibe-code you an Android app

Google upgraded AI Studio today to build native Android apps from prompts, with an embedded Android emulator for preview and device connection for installation; the initial release focuses on personal utility apps, and tester invites are planned but not yet available.

Why it matters: HKR-H/K/R all pass: Google AI Studio now generates installable native Android prototypes with emulator preview and device install. Scope stays limited to personal utility apps, so this is a mid-weight product update, not P1.

TechCrunch · AI

Agentic app coding gets an upgrade with Google’s release of Android CLI

Google released Android CLI for AI coding agents, letting platforms such as Claude Code and OpenAI Codex build Android apps from the command line; the RSS snippet does not disclose version numbers, release timelines, pricing, or performance data.

Why it matters: HKR-H/K/R all pass, but the body lacks version, timeline, and performance data. Google plus Android plus agentic coding clears the featured line, not the must-write band.