Skip to content

Anthropic / Claude

Everything Anthropic: the Claude models, Claude Code, its safety research agenda and company news.

Latest picks

861–880 of 1,304

Jun 4Thursday

Hacker News front page

Show HN: Cost.dev (YC W21) Makes Agents Cost-Aware and Cheaper to Call

Infracost launched Cost.dev, a local CLI for cloud-cost estimates in coding-agent workflows, and says it cut Claude output-token use by up to 79% and API cost by up to 67% versus a bare-Claude baseline.

Why it matters: HKR-H/K/R all pass: the local CLI cost-estimation mechanism and 79%/67% reduction claims are concrete. It is still a small vendor launch, so it sits at the featured floor, not same-day news.

Xinzhiyuan · WeChat

Claude Mythos Hits 3 Hours 6 Minutes Before Experts’ Year-End Forecast

Anthropic Claude Mythos completed 186 minutes of autonomous tasks at an 80% success rate on the METR benchmark, and the post says this matches the 3–4 hour median forecast that experts had placed at the end of 2026.

Why it matters: HKR-H/K/R all pass: the 3h06m autonomy result is a strong hook, METR 80%/186 minutes gives concrete signal, and agent safety lands with practitioners. Single-source coverage without release details or reproducible setup keeps it below p1.

AI HOT (Curated Pool)

Microsoft AI chief says Anthropic models are too expensive and is building cheaper alternatives

Microsoft’s AI chief said Anthropic models cost too much and the company is developing cheaper internal alternatives; the post does not disclose model names, cost figures, or a launch timeline.

Why it matters: HKR-H/K/R pass on a Bloomberg-reported Microsoft cost-and-replacement claim. Missing model name, cost figures, and launch timing keep it in the lower featured band.

AI HOT (Curated Pool)

How Anthropic Enables Self-Service Data Analytics with Claude

Anthropic uses Claude to automate 95% of business analytics queries with about 95% accuracy; its agentic analytics stack uses a data foundation layer, validation workflows, and skills to handle ambiguity, stale data, and retrieval failures.

Why it matters: HKR-H/K/R all pass: the official post has marketing tone, but gives 95% automation, ~95% accuracy, and an agentic analytics stack. No new model or product release keeps it in the 72–77 band.

Jun 3Wednesday

Xinzhiyuan · WeChat

OpenAI’s Greg Brockman and a 9-Year Rift With Anthropic Co-Founder Dario Amodei

A WSJ-based profile says Dario Amodei once barred Greg Brockman from an internal OpenAI project that later led to ChatGPT, and the article says Brockman now oversees OpenAI product strategy with nearly 1,500 people under that function.

Why it matters: HKR-H/K/R all pass: the WSJ-sourced ban detail and the nearly 1,500-person scope give this more signal than gossip. It is not a model release or current executive departure, so it stays in the good-quality featured band.

AI HOT (Curated Pool)

Sensor Tower: ChatGPT surpasses 1B monthly active users, fastest ever

Sensor Tower estimates ChatGPT surpassed 1 billion global monthly active users in May 2025, while Anthropic’s Claude reached 56 million monthly active users in the same period with about 640% year-over-year growth.

Why it matters: HKR-H/K/R all pass: the article adds concrete adoption estimates for ChatGPT and Claude. It stays below P1 because these are third-party usage metrics, not a model or product capability release.

AI HOT (Curated Pool)

Intelligence Cost-Performance

Microsoft added average token usage to its model release card; the model scored 71.6 on SWE-Bench Verified while using about one-third of Claude Haiku 4.5’s tokens.

Why it matters: HKR-H/K/R all pass: the score-per-token angle is clickable, with concrete 71.6 and one-third-token claims. The article is thin on full test setup and pricing, so it lands at 78.

AI HOT (Curated Pool)

Claude Code Adds Dynamic Workflows

Claude Code added dynamic workflows that execute JavaScript files at runtime to create and coordinate multiple subagents; each subagent has its own context window, and the feature is described for research, security analysis, and code review tasks.

Why it matters: HKR-H/K/R all pass: Claude Code gets runtime JS workflows coordinating isolated-context subagents. Anthropic update earns a bump, but this is a feature release rather than a model or platform launch, so it sits in the 78–84 band.

AI HOT (Curated Pool)

Claude Code launches dynamic workflows for task-specific frameworks

Claude Code added dynamic workflows that execute JavaScript files to coordinate subagents, with configurable model choice and workspace isolation level, but the post does not disclose token overhead figures or release availability details.

Why it matters: HKR-H/K/R all pass, but the post gives mechanism-level detail only; token overhead, rollout scope, and pricing are not disclosed. Claude Code relevance lifts this to the high end of a mid-weight product update.

AI HOT (Curated Pool)

Alphabet Plans $80 Billion Raise; Anthropic Files for IPO

Alphabet plans to raise $80 billion through equity financing for AI infrastructure expansion, while Anthropic has confidentially filed for an IPO; the post does not disclose valuation, listing timeline, or underwriters.

Why it matters: HKR-H/K/R all pass; Anthropic filing for an IPO is near the foundation-model IPO band, with Alphabet’s $80B AI infra financing adding hard capex signal. Valuation, timing, and banks are not disclosed, so the score stays below the very top.

AI HOT (Curated Pool)

Claude Platform Adds CLI Tool

Claude Platform added a CLI that runs every API endpoint from the terminal, calls the Messages API, launches Claude-hosted agents, and pipes results directly into the shell.

Why it matters: Claude Platform CLI clears HKR-H/K/R as a practical developer-tooling update, but the post only gives capability scope; install flow, permissions, safety limits, and pricing are not disclosed.

Financial Times · Technology

Anthropic to Expand Mythos Access to More Than 15 Countries

Anthropic will expand Mythos access to more than 15 countries, and about 150 organizations will receive the advanced cybersecurity model after requests from around the world.

Why it matters: HKR-H/K/R pass: Anthropic’s Mythos expansion has concrete scale and security resonance. It stays at the lower featured band because the post gives access numbers, not new capability details, country list, or usage terms.

AI HOT (Curated Pool)

Claude Code Team Practice: How Agentic Coding Changes Engineering Organizations and Processes

The Claude Code engineering team described process changes after making agentic coding the default at Code w/ Claude SF 2026: JIT planning, asking Claude first for context collection, Claude handling style and tests in code review, and humans focusing on legal and safety judgments.

Why it matters: First-party Claude Code workflow post with concrete engineering mechanisms and strong HKR-H/K/R fit. It is not a model or major product release, so it stays in the 78–84 band.

r/LocalLLaMA

Benchmarks of 20 Small LLMs on a 6GB RTX 4050

The author benchmarked 20 small LLMs on a 6GB RTX 4050 using LM Studio’s OpenAI-compatible API, with N=5 speed runs at 1k, 8k, and 32k context; unsloth/lfm2.5-vl-1.6b led throughput at 207 tok/s on 1k context while using 3.0GB VRAM.

Why it matters: HKR-H/K/R all pass: the low-VRAM GPU hook is concrete, the post gives speed/context/VRAM numbers, and it speaks to local-inference cost pressure. Source authority is a Reddit post, so it stays in the lower featured band.

Jun 2Tuesday

TechCrunch · AI

Anthropic scales Claude Mythos to critical infrastructure in 15+ countries

Anthropic is expanding Project Glasswing and Mythos access to 150 organizations across 15 countries, targeting power, water, healthcare, and communications infrastructure where a cyberattack could affect 100 million people.

Why it matters: HKR-H/K/R all pass: the story has scale, named sectors, and a security nerve. It stays below 85 because the post discloses rollout scope, not Mythos mechanisms, controls, or evaluation results.

Ben's Bites

Opus 4.8

Ben’s Bites says Claude Opus 4.8 is out, and Claude Code can write an orchestration script before launching subagents in parallel to work through complex tasks.

Why it matters: HKR-H/K/R all pass for a substantive Anthropic/Claude release and Claude Code agent update. The post is thin on benchmarks, pricing, and context window, so it stays low in the 85–94 band.

AI HOT (Curated Pool)

Anthropic Expands Project Glasswing Program

Anthropic expanded Project Glasswing to about 150 new organizations across more than 15 countries, covering electricity, water, healthcare, communications, and hardware infrastructure, after an initial group of about 50 partners.

Why it matters: Anthropic expanded Project Glasswing to about 150 new organizations across 15+ countries, giving HKR-H/K/R enough substance. No concrete safety mechanism or Claude capability change is disclosed, so it stays in the lower featured band.

r/LocalLLaMA

Replaced Claude with local Qwen3.6-27B in my multi-agent orchestrator for 2 weeks

The author ran Qwen3.6-27B on one RTX 3090 across 47 multi-step coding workflows. Plan generation reached about 95% schema validity, but tool-call formatting errors were about 12%, and practical long-context use degraded past about 12k tokens.

Why it matters: HKR-H/K/R all pass: a named first-person local-vs-Claude experiment with concrete numbers. The single Reddit source and 47-workflow scope keep it below the 78–84 band.

Financial Times · Technology

Top AI Labs Expand Research Into Machine “Consciousness”

Google DeepMind, Anthropic, and Meta are studying whether AI can become conscious and the human implications, but the post does not disclose methods, timelines, or evaluation criteria.

Why it matters: HKR-H and HKR-R pass because top labs studying machine consciousness is a live safety debate. HKR-K fails: the body names labs but gives no method, timeline, or criterion, so this stays at the 72 featured floor.

Xinzhiyuan · WeChat

Pope and Anthropic warn of AGI by 2030 and a three-year governance window

Xinzhiyuan says Pope Leo XIV and Anthropic co-founder Christopher Olah backed AI governance, citing AGI by 2030, a 1,500-day window, and a proposed FATF-style international audit framework for AI oversight.

Why it matters: HKR-H/K/R all pass, but this is governance commentary and timeline warning, not a model launch or binding policy. The concrete hooks are 2030, 1,500 days, and a FATF-style audit frame, so it lands in low featured.