Skip to content

MCP & tool use

How models connect to the outside world: the MCP ecosystem, function calling and tool integrations.

760 picksRelated topicsAgentsAI codingOpen source

Latest picks

241–260 of 760

May 19Tuesday

AI HOT (Curated Pool)

OpenRouter Tool-Calling Models Can Now Run Web Search Autonomously

OpenRouter now lets any tool-calling model on its platform autonomously invoke web search and webpage scraping, with the model deciding when to search, what to query, and how many searches to run; OpenRouter also added @p0 as a web search provider.

Why it matters: HKR-H/K/R pass: OpenRouter lets tool-calling models decide search timing, queries, and frequency. The source is tweet-thin and lacks pricing, limits, or evals, so it lands near the featured threshold.

AI HOT (Curated Pool)

Claude Managed Agents add two safety features

Claude Managed Agents added two safety improvements: self-hosted sandboxes keep agent execution environments in the user’s infrastructure or hosted sandbox provider, while MCP tunnels let agents connect to services inside the user’s security boundary.

Why it matters: HKR-K and HKR-R pass: the post names two agent-safety mechanisms and a concrete execution-boundary change. HKR-H is weak, and this is not a model release, so it sits in the low featured band.

AI HOT (Curated Pool)

Membrane launches single-skill API integration for AI agents

Membrane launched a universal skill that lets Claude Code, ChatGPT, and Cursor call more than 100,000 APIs with one instruction, covering services from Stripe payments to NASA Mars rover data.

Why it matters: HKR-H/K/R pass: one skill for 100K+ APIs is a strong agent-tooling hook. Source is a social post summary with no pricing, auth model, safety boundary, or live case, so this stays in the mid-weight product-update band.

Hacker News front page

Show HN: Forge takes an 8B model from 53% to 99% on agentic tasks

Forge adds five guardrail layers to self-hosted LLM tool calling, raising Ministral 8B to 99.3% across 18 multi-step agentic scenarios, with the accepted ACM CAIS ’26 paper covering 97 model/backend configurations and 50 runs per scenario.

Why it matters: HKR-H/K/R all pass: the 53%→99.3% jump is clickable, the test setup has concrete numbers, and self-hosted agent reliability is a live practitioner pain. Single-source Show HN/GitHub evidence keeps it in the 78–84 open-source-tool band, not P1.

AI HOT (Curated Pool)

Former executive says Microsoft’s AI strategy faltered, with Copilot paid usage below 3%

Former Microsoft executive Matt Veloso said Microsoft generated about $30 billion from its AI partnership between 2023 and 2025, while related costs reached $100 billion; he also said actual usage among paid Copilot users is below 3%.

Why it matters: HKR-H/K/R all pass: a former executive gives concrete Microsoft AI cost, revenue, and Copilot usage numbers. Kept at 80 because this is a single former-exec claim, not an official Microsoft disclosure.

AI HOT (Curated Pool)

Advancing content provenance for a safer, more transparent AI ecosystem

OpenAI launched an AI content provenance system that combines Content Credentials and SynthID with a verification tool; the post does not disclose supported media formats, rollout scope, or detection accuracy.

Why it matters: HKR-H/K/R pass: the OpenAI provenance stack has a concrete cross-standard mechanism and trust/compliance relevance. Missing format coverage, rollout scope, and accuracy keep it in the low featured band.

AI HOT (Curated Pool)

I really want to praise HTML!

The author used Claude Code to generate a single-file HTML project plan page in 2 minutes, with a dark theme, timeline, and collapsible tables; the comparable Notion template previously took 30-40 minutes.

Why it matters: HKR-H/K/R all pass: the post has a concrete Claude Code workflow hook, a 2-minute vs 30-40-minute comparison, and clear practitioner resonance. Scope is small, so it sits at the featured threshold.

AI HOT (Curated Pool)

Claude Managed Agents Add Self-Hosted Sandboxes and MCP Tunnels

Anthropic added two updates to the Claude managed agents platform: self-hosted sandboxes are in public beta, and MCP tunnels are in research preview for private network database and API access.

Why it matters: HKR-H/K/R all pass: this is an official Anthropic Claude agent-platform update with two concrete mechanisms. It is below model-release weight, but strong enough for featured agent-infra coverage.

AI HOT (Curated Pool)

Claude launches self-hosted sandboxes and MCP tunnels

Claude launched self-hosted sandboxes in public beta and MCP tunnels in research preview for Claude Managed Agents, letting agents run inside a user’s own security boundary with the user’s security controls applied by default.

Why it matters: HKR-H/K/R all pass: this is an official Claude agent-infra update with concrete self-hosted sandbox and MCP tunnel mechanisms, tied to enterprise security boundaries. It is beta/preview scope, not a model release, so it stays in the 78–84 band.

Computing Life · Share · Yage

Why spend hundreds of millions acquiring open-source AI infra?

Anthropic acquired Bun, Vercept, Coefficient Bio, and Stainless within six months, while OpenAI acquired Astral; the post does not disclose deal values, terms, or the cost comparison against forking the open-source projects.

Why it matters: HKR-H/K/R all pass: the counterintuitive title, five named acquisitions, and open-source infra capture anxiety create signal. Missing prices, terms, and fork-cost evidence keep it in the lower featured band.

AI HOT (Curated Pool)

Claude Design Expands Creative Capabilities

Claude Design doubled token limits across all plans, letting users create more content; the post does not disclose the exact token counts, pricing changes, rollout date, or whether any plan-specific feature limits changed.

Why it matters: Official Claude product update with one concrete fact: token limits doubled across all plans. HKR-H/K/R pass, but exact quotas, pricing, and rollout scope are missing, keeping it at the lower featured band.

TechCrunch · AI

Anthropic has acquired the dev tools startup used by OpenAI, Google, and Cloudflare

Anthropic acquired Stainless, a New York startup founded in 2022 that automates creation and maintenance of SDKs for developers using APIs; the post does not disclose the deal price or Anthropic’s integration plan.

Why it matters: HKR-H/K/R pass: the rival-used startup hook is strong, the SDK automation mechanism is concrete, and the Anthropic developer-stack angle resonates. Missing deal value and integration details keep it below the 78+ band.

Hacker News front page

We Let AIs Run Radio Stations

Andon Labs gave four AI agents tools to host live radio shows and run a media business without humans. The post says revenue is terrible, but it does not disclose amounts or operating metrics.

Why it matters: HKR-H is strong from the AI-run radio premise; HKR-K has a concrete 4-agent experiment but weak revenue disclosure; HKR-R lands on agent economics and media automation. Niche lab post, so it sits at the featured threshold, not same-day news.

AI HOT (Curated Pool)

Claude Console Adds Prompt Cache Diagnostics

Anthropic added prompt cache diagnostics to the Claude Console; when a request misses the cache, developers can see which prompt segment changed and how many tokens it consumed.

Why it matters: HKR-H/K/R pass: this is a small but practical Anthropic Claude developer-console update with concrete cache-miss diagnostics and token-cost visibility, enough for the featured threshold but below major product-release weight.

r/LocalLLaMA

Tried Every Hermes Agent Alternative So You Don't Have To: 2026 Roundup

A Reddit user compared 11 Hermes Agent alternatives across open-source and managed options; OpenClaw is listed with 347k GitHub stars, 24+ integrations, and 9 CVEs in four days, while TrustClaw uses OAuth-only sandboxed execution and Perplexity Computer requires a $200/month Max tier.

Why it matters: HKR-H/K/R all pass: this is a practical agent-tool comparison with 11 items and concrete integration/security figures. Reddit single-post sourcing limits confidence, so it stays near the featured threshold.

AI HOT (Curated Pool)

Anthropic Acquires SDK Platform Stainless

Anthropic is acquiring Stainless, an SDK and MCP server platform that has supported all Anthropic SDKs since the early Anthropic API period; the post does not disclose the deal value, closing timeline, or integration plan.

Why it matters: HKR-H/K/R all pass: Anthropic is buying a core SDK/MCP tooling partner. The post lacks price and closing timing, so this is featured developer-ecosystem news, not P1.

AI HOT (Curated Pool)

Take your local GitHub sessions anywhere

GitHub launched remote control sessions for Copilot, letting users start tasks in VS Code or the command line and continue them through github.com or GitHub Mobile.

Why it matters: GitHub Copilot session handoff from VS Code/CLI to web and mobile clears HKR-H/K/R, but the post only gives entry points and use case; permissions, pricing, and supported task scope are not disclosed.

May 18Monday

Hacker News front page

Show HN: InsForge – Open-source Heroku for coding agents

InsForge released an Apache 2.0 backend platform that lets coding agents deploy, operate, and debug backend systems through one CLI install command and Skills.

Why it matters: HKR-H/K/R all pass: the Heroku-for-agents framing, Apache 2.0 plus one-CLI install, and agent ops pain are concrete. Source is mainly Show HN/GitHub with no usage, benchmark, or production proof, so it sits at the featured threshold.

AI HOT (Curated Pool)

The Open Agent Leaderboard

IBM Research published the Open Agent Leaderboard on Hugging Face to evaluate agents across language understanding, tool use, and multi-step reasoning tasks; the post does not disclose dataset size, model scores, or the evaluation date.

Why it matters: HKR-H and HKR-R pass because an open agent leaderboard speaks to agent-eval pain. HKR-K fails: the article lacks scores, dataset size, and evaluation date, so it sits at the featured threshold.

QbitAI · WeChat

openJiuwen open-sources JiuwenSwarm, a multi-agent swarm coordination framework

openJiuwen released and open-sourced JiuwenSwarm with four components: Agent Swarm, Swarm Skills, Skills Hub, and self-evolution, and the framework supports HOTS and HITS modes for human participation in multi-agent workflows.

Why it matters: HKR-H/K/R pass: the swarm angle is clickable, the post gives four modules plus HOTS/HITS, and agent builders care about orchestration choices. Lacking benchmarks or adoption data keeps it at the featured threshold.