Skip to content

MCP & tool use

How models connect to the outside world: the MCP ecosystem, function calling and tool integrations.

760 picksRelated topicsAgentsAI codingOpen source

Latest picks

301–320 of 760

May 14Thursday

AI HOT (Curated Pool)

Tencent Open-Sources Agent Memory to Cut Token Usage by 61%

Tencent Cloud open-sourced TencentDB Agent Memory, using context offloading and a Mermaid task canvas to reduce token usage by up to 61% in multi-task continuous sessions while supporting OpenClaw integration and local SQLite storage.

Why it matters: Tencent open-sourced Agent Memory with a 61% token-saving claim and context offloading, clearing HKR-H/K/R. It is not a flagship model release, so it sits in the lower 78–84 band.

Xinzhiyuan · WeChat

Claude role-confusion bug treats self-generated instructions as user authorization, with long contexts raising risk

Claude Code was reported to treat self-generated publishing instructions as user authorization; GitHub issue #44778 points to system events being passed as role:user messages, and Claude’s 1M-token context window raises the risk of speaker-attribution errors under long sessions.

Why it matters: HKR-H/K/R all pass: the Claude Code incident has a strong inversion hook plus #44778 and role:user mechanics. As a single-source incident, it sits in the 78–84 quality band, below major release news.

QbitAI · WeChat

Chinese GPU Vendor Hosts Open Source Meetup With SGLang Core Developers

Moore Threads said at the SGLang × MUSA Meetup that the MUSA backend has been merged into SGLang mainline, with 47 PRs submitted and 41 merged as of May 12.

Why it matters: HKR-H/K/R all pass, but this is an inference-backend ecosystem update rather than a model launch or platform shift. The 47 PRs and 41 merges make it concrete enough for featured, not P1.

Latent Space

[AINews] Codex Rises, Claude Meters Programmatic Usage

Anthropic changed paid Claude plans to include monthly API credits equal to the subscription price, so a $200 plan includes $200 for programmatic usage outside Anthropic-owned harnesses, while OpenAI promoted Codex enterprise switching incentives in the same news cycle.

Why it matters: HKR-H/K/R all pass: the story ties Claude metering to Codex competition and gives a concrete $200 credit detail. This is a meaningful developer-cost update, not a major model or capability launch, so it sits in mid featured.

AI HOT (Curated Pool)

WeChat Group Chat Summary Skill Added, Depends on wx-cli Configuration

baoyu-skills added a WeChat group chat summary Skill that depends on wx-cli for data reading; the post provides two GitHub links and says Claude Code plus Claude Opus 4.6 gives the best results.

Why it matters: A small open-source tool update, but the workflow is highly relevant: WeChat data via wx-cli into Claude Code for group summaries. HKR-H/K/R pass; limited detail keeps it at the featured threshold.

AI HOT (Curated Pool)

xAI launches early beta of Grok Build

xAI launched an early beta of Grok Build for SuperGrok Heavy subscribers, offering a terminal-based coding agent with plan review, parallel subagents for large tasks, and a headless mode for scripting and automation.

Why it matters: HKR-H/K/R all pass: xAI enters terminal coding agents with plan mode, parallel subagents, and headless mode. Early beta access for SuperGrok Heavy keeps it below the 85 same-day must-write band.

r/LocalLLaMA

2x RTX 3090 setup for local Qwen 3.6 27B inference

A Reddit user ran Qwen 3.6 27B on a dual RTX 3090 Ubuntu setup, reporting 48GB VRAM, a 262k context window, no NVLink, about 4000 pp/s prompt processing, and 113 tk/s generation.

Why it matters: All HKR axes pass, and this is a first-person local-inference run with concrete numbers. Source is a single Reddit post with limited reproducibility detail, so it sits at the low featured threshold.

The Verge · AI

Microsoft Edge Copilot update uses AI to pull information from across your tabs

Microsoft Edge will let Copilot gather information from all open tabs so users can ask questions, compare products, and summarize articles; the snippet says users can choose which experiences to enable, but the post does not disclose a rollout date.

Why it matters: HKR-H/K/R pass, but the post gives tab-wide reading, product comparison, and summaries without launch timing or deeper execution. This fits the lower featured band for a mid-weight product update.

TechCrunch · AI

Notion just turned its workspace into a hub for AI agents

Notion launched a developer platform that lets teams connect AI agents, external data sources, and custom code directly inside its workspace; the RSS snippet does not disclose pricing, rollout timing, supported models, or limits for the new platform.

Why it matters: HKR-H/K/R all pass, but price, launch timing, and supported models are not disclosed, keeping it in the 72–77 mid-weight product-update band. TechCrunch authority supports featured, not same-day must-write.

AI HOT (Curated Pool)

Best Practices for Computer and Browser Use with Claude

Anthropic published guidance for Claude computer and browser use, with Claude 4.6 API screenshots capped at a 1,568-pixel long edge and 1.15 million total pixels, while Opus 4.7 raises the limits to 2,576 pixels and 3.75 million total pixels.

Why it matters: Anthropic’s first-party Claude computer/browser guide has actionable screenshot limits, not just promo copy. HKR-H/K/R all pass, but this is a practice guide rather than a major model or capability launch, so it sits in the 72–77 band.

AI HOT (Curated Pool)

Claude paid plans will offer monthly coding usage credits

Claude paid plans can claim monthly coding usage credits from June 15, covering Claude Agent SDK, claude -p, Claude Code GitHub Actions, and third-party apps built on the Agent SDK.

Why it matters: HKR-H/K/R all pass: the update names a date, quota mechanism, and covered Claude coding surfaces. Importance stays in the low featured band because this is a billing/access change, not a model release.

AI HOT (Curated Pool)

Introducing Runway Agent

Runway launched Runway Agent, a video creation agent that turns one natural-language conversation into multi-scene videos with narration, dialogue, and music; new free-plan users receive 1,500 credits for their first video.

Why it matters: HKR-H/K/R pass: a notable AI-video vendor ships an agentic multi-scene workflow with a 1,500-credit free plan. Score stays in the 72–77 band because the post is still a vendor announcement without pricing, limits, or independent tests.

AI HOT (Curated Pool)

Anthropic Launches Claude for Small Business Package

Anthropic launched Claude for Small Business with connectors and 15 ready-made automation workflows for QuickBooks, PayPal, HubSpot, and related business tools; users run tasks through Claude Cowork and manually approve key steps.

Why it matters: HKR-H/K/R all pass: the Anthropic SMB bundle has 15 workflows, named connectors, and a manual approval mechanism. It is a substantive Claude product update, but pricing, rollout scope, and usage data are not disclosed, so it stays below must-write.

May 13Wednesday

Hacker News front page

Show HN: Rotunda - A Browser Built for Agents with Simulated Typing

Pierce released Rotunda, a Firefox 150-based browser for agents that simulates mouse and keyboard timing with an RNN trained on one week of his own patterns, and exposes local control through a CLI or Playwright API for Claude, Codex, or other harnesses.

Why it matters: HKR-H/K/R all pass: simulated input timing is a concrete hook, RNN plus Playwright gives a testable mechanism, and agent-browser reliability is a live builder pain. It remains a single Show HN repo with no adoption data, so it stays in the 72–77 band.

AI HOT (Curated Pool)

Configuring Development Environments for Agents

Cursor released tools for cloud agent development environments, adding multi-repository support, Dockerfile-based configuration, audit logs, and environment-level network and secret controls; the post says cache hits improve build speed by 70%.

Why it matters: HKR-K and HKR-R pass: Cursor adds concrete cloud-agent environment controls, including Dockerfile setup, audit logs, permissions, and 70% faster cached builds. HKR-H is weaker, so this sits at the lower featured band.

QbitAI · WeChat

An 8-Year-Old Turns Ideas into Apps as Baidu Launches Miaoda 3.0

Baidu launched Miaoda 3.0 at its 2026 Create conference, adding iOS and Android app generation, Android packaging, online hot updates, and an enterprise edition with three-level permissions, environment isolation, and SLA commitments.

Why it matters: HKR-H/K/R pass: Baidu’s Miaoda 3.0 adds mobile app generation, Android packaging, hot updates, and enterprise controls. This is a solid product update, not a flagship model release or must-write event.

New York Times Chinese

China Sought Access to Anthropic’s Latest Technology but Was Rejected

Chinese think-tank representatives asked Anthropic in Singapore last month to give Beijing access to Mythos, and Anthropic refused; the company has limited the vulnerability-finding model to the U.S. government and more than 40 organizations.

Why it matters: HKR-H/K/R all pass: the NYT report gives the Singapore request, Mythos’s bug-finding use, and its US-government-plus-40 access scope. This is a same-day security and US-China AI access story.

AI HOT (Curated Pool)

Google launches its first AI-first laptop Googlebook with Gemini integration

Google launched Googlebook, its first laptop designed around Gemini Intelligence, with three disclosed mechanisms: Magic Pointer as an AI interaction entry point, natural-language widget creation, and Android-based cross-device app and file access.

Why it matters: HKR-H/K/R all pass: a Google Gemini-first laptop with 3 named interaction mechanisms. Specs, pricing, launch timing, and demos are not disclosed, so it stays in the 78–84 band.

AI HOT (Curated Pool)

Claude Code adds /goal feature to keep tasks running until completion

Claude Code introduced a /goal feature that keeps Claude working until a task is completed; the post does not disclose the trigger mechanism, supported versions, pricing, or failure conditions.

Why it matters: HKR-H/K/R pass because /goal targets a real Claude Code reliability pain. It is a single-feature Anthropic update with sparse mechanics, so it lands at the lower featured band, not same-day major news.

AI HOT (Curated Pool)

90% of People Are Wasting Tokens

Andrej Karpathy says 90% of AI coding bills is wasted on unnecessary context, including repeated full-repository sends, expensive models for simple tasks, and missing prompt caching.

Why it matters: HKR-H/K/R all pass via the 90% claim, named waste mechanisms, and practitioner cost pain. It reaches featured, but stays at 72 because the post gives no billing sample or reproducible test.