Skip to content

MCP & tool use

How models connect to the outside world: the MCP ecosystem, function calling and tool integrations.

760 picksRelated topicsAgentsAI codingOpen source

Latest picks

281–300 of 760

May 15Friday

AI HOT (Curated Pool)

Feishu Open-Source CLI Tool Gets 10,000 Stars in 45 Days with Visible AI Operations

Feishu’s open-source lark-cli gained over 10,000 GitHub stars in 45 days, letting AI create groups and documents through the command line with each operation previewable and reviewable.

Why it matters: HKR-H/K/R all pass, but the source is a single social post and lacks usage, contributor, or adoption data. This fits the lower featured band for an open-source agent tool update.

r/LocalLLaMA

Used over a million tokens in three sessions to test Qwen 3.6 35B MTP

A Reddit user tested Qwen3.6-35B-A3B MTP across three million-token-scale sessions, using 300k context and KV Q8_0, and reported about 1.5x the tok/sec of earlier tests.

Why it matters: HKR-H/K/R all pass: the million-token test is clickable, 300k context and KV Q8_0 add testable detail, and local speed maps to cost. Source is one Reddit post, so it stays below the high-importance band.

Synced · WeChat

Amazon employees reportedly tokenmaxx to meet AI usage KPIs

Amazon required more than 80% of developers to use AI tools each week and created an internal token-consumption leaderboard. Employees reportedly used the internal MeshClaw agent to inflate usage, while Amazon has limited visibility of the statistics to each employee and their direct manager.

Why it matters: HKR-H/K/R all pass: Amazon’s AI-use KPI became token-gaming, with >80% target, leaderboard, MeshClaw, and visibility changes. Impact is workplace-significant, not major-release level, so featured not p1.

Xinzhiyuan · WeChat

Hassabis Praises Google DeepMind's AI-enabled Pointer Powered by Gemini

Google DeepMind released a Gemini-powered AI-enabled pointer and opened two demos in Google AI Studio: image editing and place finding on maps, while the post says Chrome pointer selection and a Googlebook Magic Pointer are planned product paths.

Why it matters: HKR-H/K/R all pass: the prompt-free pointer is clickable, the two AI Studio demos add concrete facts, and UI replacement resonates. Scope is still demo-level, with no metrics or API details, so 78 not 85+.

AI HOT (Curated Pool)

ChatGPT launches personal finance experience

OpenAI launched a personal finance preview for ChatGPT Pro users in the US, letting users connect financial accounts and receive analysis based on their finances, goals, and priorities.

Why it matters: HKR-H/K/R all pass: OpenAI is moving ChatGPT into personal finance with linked financial accounts. The score stays at 76 because it is a US Pro preview, with partners, controls, and pricing not disclosed.

AI HOT (Curated Pool)

API prompt precaching speeds up first-token generation

Claude API prewarms prompt cache with the system prompt, skips output, then hits cache on the real request.

Why it matters: HKR-H/K/R all pass: this is a concrete Claude API latency mechanism, not a vague product tease. It clears featured, but it is a mid-weight inference update rather than a major model or capability release.

AI HOT (Curated Pool)

Claude Agent Tool v2.1.142 Release

Claude Agent Tool v2.1.142 adds eight command-line flags for configuring background sessions, upgrades Fast mode’s default model to Opus 4.7, and fixes more than 15 issues including MCP tool timeouts and Windows network-drive deadlocks.

Why it matters: HKR-H/K/R all pass: this is a small Claude Code release, but the Opus 4.7 Fast-mode default, 8 session flags, and 15+ fixes affect daily dev workflows. Anthropic tool-chain relevance keeps it at the featured floor.

AI HOT (Curated Pool)

Codex adds automation hooks and programmatic tokens

Codex added hooks and programmatic access tokens: hooks run scripts at key task stages for validation, secret scanning, logging, or repo-specific behavior, while scoped tokens for Business and Enterprise teams support CI/CD, release workflows, and internal automation with expiration or revocation.

Why it matters: HKR-H/K/R all pass: Codex gains concrete automation hooks and programmatic tokens for CI/CD. Score stays in the 72–77 band because the post discloses workflow fit, not pricing, permission detail, or impact data.

The Verge · AI

OpenAI’s Codex is now in the ChatGPT mobile app

OpenAI will let users access Codex from the ChatGPT mobile app; the RSS snippet says Codex can write code and use apps on a computer, but the post does not disclose launch timing, pricing, or the full mobile feature scope.

Why it matters: OpenAI added Codex access to ChatGPT mobile, a mid-weight product update. HKR-H/K/R pass through the mobile coding-agent hook, concrete app-control claim, and developer workflow nerve; missing timing, pricing, and support scope keep it at the featured floor.

TechCrunch · AI

OpenAI is reportedly preparing legal action against Apple; it would not be the first partner to feel burned

OpenAI is reportedly exploring legal action against Apple over a ChatGPT integration that failed to deliver expected subscribers and prominence. The RSS snippet does not disclose the legal claims, filing timeline, contract terms, subscriber targets, or Apple’s response.

Why it matters: HKR-H/K/R pass: a reported OpenAI-Apple legal fight has hook, motive, and platform-risk resonance. Claims and contract details are not disclosed, so it stays below must-write.

The Verge · AI

Microsoft starts canceling Claude Code licenses

Microsoft plans to remove most Claude Code licenses and push many developers toward Copilot CLI; the snippet says Microsoft opened access in December to thousands of internal developers, but the post does not disclose the exact license count, pricing, or migration schedule.

Why it matters: HKR-H comes from Microsoft dropping a rival coding tool; HKR-K adds the Dec rollout to thousands of internal devs; HKR-R hits Claude Code vs. Copilot competition. Strong featured, not a major release.

AI HOT (Curated Pool)

The Founder's Playbook: Building an AI-Native Startup

Anthropic published an AI-native startup playbook covering four stages—ideation, MVP, launch, and scaling—with goals, exit criteria, failure modes, and Claude-based exercises for validation, customer discovery, technical debt control, product-market fit checks, and workflow automation.

Why it matters: HKR-H/K/R all pass, but this is an Anthropic playbook rather than a model or product capability release. The concrete value is the 4-stage framework, exit criteria, and Claude-driven exercises, so it lands at the featured floor.

AI HOT (Curated Pool)

Using Claude Code Effectively in Large Codebases: Best Practices and Where to Start

Claude Code is used in million-line monorepos, legacy systems, and distributed architectures, and the post says its large-codebase workflow relies on five extension points: CLAUDE.md, hooks, skills, plugins, and MCP servers for agentic search on local codebases.

Why it matters: HKR-H/K/R pass: official Claude Code guidance, five concrete extension points, and a direct coding-agent workflow nerve. It is a high-quality tutorial, not a new model or major capability release, so it stays in the 72–77 band.

AI HOT (Curated Pool)

OpenEvidence reaches 65% of U.S. doctors, drawing attention to shadow AI use

OpenEvidence reaches 65% of U.S. doctors and recorded 27 million clinical uses in April; doctors registered on mobile with license numbers, while hospitals were initially unaware of the shadow AI adoption pattern.

Why it matters: HKR-H/K/R all pass: 65% doctor reach and 27M April uses are unusually strong adoption data, with the hospital-unaware angle adding shadow-AI tension. Kept below 85 because methodology, revenue, and liability controls are not disclosed.

AI HOT (Curated Pool)

Genkit launches middleware system to improve control in agentic AI apps

Google’s open-source Genkit framework added a middleware system that intercepts generation calls, models, and tools, with support for TypeScript, Go, Dart, and Python.

Why it matters: HKR-H/K/R all pass: the Google Genkit update adds concrete middleware hooks for agentic apps across generation, model, and tool layers. Scope stays within Genkit, so this sits at the featured threshold rather than a must-write release.

r/LocalLLaMA

inclusionAI/Ring-2.6-1T on Hugging Face

inclusionAI released Ring-2.6-1T, a 1T-parameter reasoning model on Hugging Face; it supports high and xhigh reasoning effort levels, targets agent workflows and long-horizon tasks, and uses Async RL with the IcePop algorithm for reinforcement-learning training stability.

Why it matters: HKR-H/K/R pass: a 1T HF model with two reasoning modes and named training methods is real signal. Benchmarks, license, and inference cost are not disclosed, so this stays at the lower edge of featured.

May 14Thursday

AI HOT (Curated Pool)

Anthropic and Gates Foundation form $200M partnership for global health and education

Anthropic and the Gates Foundation formed a four-year, $200 million partnership that provides funding, Claude credits, and technical support for global health, life sciences, education, and economic mobility projects.

Why it matters: HKR-H is the $200M Gates partnership; HKR-K is the four-year structure plus funding, Claude credits and technical support. HKR-R is weak: no new model capability, pricing, or developer impact. No hard exclusion; high-authority partnership sits at featured threshold.

AI HOT (Curated Pool)

Kimi launches Web Bridge browser extension for multi-platform interaction

Kimi launched the Web Bridge browser extension, which lets agents search, scroll, click, type, and complete website tasks, with support for Kimi Code CLI, Claude Code, Cursor, Codex, and Hermes.

Why it matters: HKR-H/K/R all pass: the product hook, action list, and workflow relevance are clear. Kept in the 72–77 band because this is a mid-weight tool update, not a model release, and safety or performance details are not disclosed.

AI HOT (Curated Pool)

Use Codex from anywhere

OpenAI added Codex to the ChatGPT mobile app, letting users monitor, guide, and approve remote coding tasks across devices.

Why it matters: OpenAI added Codex controls to ChatGPT mobile for monitoring, guiding, and approving remote coding tasks. This clears HKR-H/K/R as a mid-weight product update, but pricing, permission details, and task limits are not disclosed, so it stays below a major release.

r/LocalLLaMA

Automated AI researcher running locally with llama.cpp

Hugging Face’s ml-intern added local-model support through llama.cpp and ollama; the post says Qwen3.6-35B-A3B can orchestrate CPU/GPU sandboxes and Hub jobs to run an end-to-end SFT workflow.

Why it matters: HKR-H/K/R all pass, but this is a Reddit-sourced open-source tool update, not a major model release. Local sandbox and Hub-job orchestration for SFT put it just above the featured threshold.