Skip to content

Anthropic / Claude

Everything Anthropic: the Claude models, Claude Code, its safety research agenda and company news.

Latest picks

681–700 of 1,304

Jun 25Thursday

Hacker News front page

Anthropic says Alibaba illicitly extracted Claude AI model capabilities

Reuters reports that Anthropic accuses Alibaba of illicitly extracting Claude model capabilities. The article does not detail the extraction method or which model versions are involved. No response from Alibaba is included, and no technical specifics are provided. I'd hold off until both sides show evidence.

Why it matters: Anthropic publicly accusing Alibaba of illicitly extracting Claude capabilities is a high-conflict event with cross-source cluster formation. Currently only the Reuters headline is available — no technical details or Alibaba response yet, so the knowledge axis can't carry weig...

Bloomberg Technology

Anthropic Accuses Alibaba of Illicitly Accessing Its AI Models

Anthropic publicly accuses Alibaba of illicitly accessing its AI models. The post does not specify whether this refers to model weights or API outputs. Only the headline is available so far—I'll wait for more details before drawing conclusions.

Why it matters: Anthropic's public accusation against Alibaba is a high-conflict story with strong discussion value. But the body currently only has a headline, with key facts missing — what exactly was accessed and whether Alibaba has responded — so the score stays conservative.

Bloomberg Technology

Google set to lose two more senior AI staffers to Anthropic

Bloomberg reports Google is about to lose two more senior AI researchers to Anthropic. This is the latest in a string of Google-to-Anthropic moves, though the article does not name the two staffers or specify their roles.

Why it matters: Bloomberg exclusive on two more senior Google AI researchers heading to Anthropic — the talent-flow narrative has pull. But without names or roles, the knowledge axis is thin. H and R hit, K misses, landing right at the featured threshold.

The Verge · AI

The $27 million AI proxy war over Alex Bores ends in a draw

In New York's 12th District Democratic primary, the candidate backed by Anthropic and the one backed by OpenAI fought to a draw. Anthropic's pick Alex Bores didn't win, but OpenAI didn't crush him either. Combined spending hit $27 million; the post doesn't break down how much each side put in. The race was seen as a stress test of both AI companies' political influence, and neither walked away with a decisive win.

Why it matters: Two top AI labs spent $27M on a congressional primary proxy war and ended in a draw—that's inherently a story. Hits all three HKR axes, but the article doesn't break down how much each company spent, so the score stays at the featured threshold.

AI HOT (Curated Pool)

Figma Config 2026 bets on human judgment while AI costs eat margins and models come from competitors

At Config 2026, Figma turned its canvas into a workspace for code, motion, 3D, and shaders. Code Layers puts design and production code side by side; Motion brings animation timelines into collaborative editing; Shader uses WebGPU for material effects. But the company admits high inference costs from third-party AI models are squeezing margins, and those models come from providers like Anthropic that are building competing products. Figma's bet is on AI that produces tweakable tools rather than one-shot outputs, plus team-shared prompts and plugins to cut token use. The post doesn't spell out progress on in-house models.

Why it matters: Figma Config 2026 product updates are substantive (Code Layers / Motion / Shader), but the real news is the company openly admitting third-party AI inference costs are eroding margins, with models coming from Anthropic and others who are building competing products. HKR all hi...

AI HOT (Curated Pool)

Gemini 3.5 Flash now has built-in computer use

Google added a native computer-use tool to Gemini 3.5 Flash, letting the model take screenshots, click buttons, and fill forms to operate web and desktop UIs. It joins Anthropic and OpenAI in baking screen control directly into a model. The post claims 3.5 Flash beats Claude Sonnet 4 and GPT-5 on the WebVoyager benchmark, but Google didn't release full eval details or reproduction steps—hold for third-party tests. Available now via Gemini API and Google AI Studio; pricing and rate limits aren't disclosed in the post.

Why it matters: Google natively integrates computer use into Gemini 3.5 Flash, directly competing with Claude Sonnet 4 and OpenAI's equivalent, with benchmark numbers provided. The gap: it's a blog announcement with no API pricing or real latency data yet — one step short of production-readin...

Jun 24Wednesday

Hacker News front page

NSA lost access to Anthropic's Mythos model after export controls

NSA cybersecurity analysts had been testing Anthropic's Mythos model and found it could break into nearly all of their classified systems within hours. After the Trump administration imposed export controls on Anthropic this month, Mythos 5 and Fable 5 were pulled back, cutting off NSA's access. Senator Mark Warner cited NSA chief Gen. Joshua Rudd at a hearing, saying Mythos breached classified networks 'not in weeks, but in hours.' The article doesn't say whether NSA has a backup plan or what specifically triggered the export controls.

Why it matters: NYT exclusive with NSA test results and Senate hearing details. The 'hours not weeks' breach claim is hard new info, and the export-control cutoff adds policy tension. Downside: the article doesn't specify which systems were breached or how, so the technical picture is still c...

Latent Space

Anthropic launches Claude Tag: async, proactive Claude agents inside Slack

Anthropic put Claude into Slack as a persistent team member that can work across channels, follow up proactively, and wait days for dependencies. Internally it already writes 65% of product PRs, including most of Claude Tag itself. Beta is limited to Enterprise and Team plans; the post doesn't disclose pricing.

Why it matters: Anthropic launched Claude Tag, letting Claude join Slack as a team member with cross-channel @ mentions, proactive follow-ups, and the ability to wait days for blockers to resolve. Internal data shows it wrote 65% of the product team's PRs, including most of Tag's own code. Th...

Computing Life · Share · Yage

Claude Tag deconstructed: the tech isn't new, but enterprises are now treating agents as governed entities

Anthropic's Claude Tag puts a persistent AI teammate in Slack that can proactively chime in and retain channel context. Under the hood, ambient mode is still scheduled triggers hitting an HTTP endpoint, and memory is raw chat logs, not organizational knowledge. The real shift is governance: enterprises are assigning agents independent identities, permissions, budgets, and audit trails. Microsoft's Entra Agent ID treats agents as first-class directory citizens. Pricing is moving from per-seat to consumption-based. One Reddit user burned $41,952 in a month running persistent agents, 99.6% spent replaying context. Continuous learning isn't here yet—Gartner predicts over 40% of agent projects will be abandoned by 2027—but the direction is set: enterprise software's runtime object is shifting from apps to governed agents.

Why it matters: A sober technical deconstruction of Claude Tag that identifies the real shift: enterprise authorization moving from 'who installed the app' to 'what identity does this agent have, what can it access, who's responsible.' Concrete dissection (HTTP endpoints, external orchestrati...

Bloomberg Technology

Anthropic Customer Sues US Government Over Fable AI Access Ban

An Anthropic customer is suing the US government after losing access to the Fable model. The article doesn't specify which Anthropic model Fable refers to. The plaintiff argues the government cut off a paid service without due process. Only the lawsuit filing is confirmed so far; case details and the government's response aren't public yet.

Why it matters: An Anthropic customer suing the US government over revoked model access is a rare policy-vs-business conflict, hitting H and R. But the body is thin — no model identity or government rationale disclosed — so K is absent, keeping the score at the featured threshold.

Hacker News front page

Anthropic launches Claude Tag: @Claude in Slack as a proactive team member

Claude Tag lets teams @Claude in Slack channels to delegate tasks. The model breaks down requests, uses connected tools, and works asynchronously. It retains channel context so you don't repeat yourself. Anthropic says 65% of its product team's code is now created by an internal version of Claude Tag, with use cases extending to metrics, support tickets, and bug hunting. It runs on Opus 4.8 and is in beta for Enterprise and Team customers. The post doesn't specify a timeline for expanding beyond Slack.

Why it matters: Anthropic product launch that moves Claude from chat UI into team collaboration tools, backed by the hard stat of 65% internal code generation. Hits all three HKR axes, same-day must-write. Not scoring higher because real-world external team data isn't in yet.

TechCrunch · AI

Anthropic’s Claude Tag learns your company by reading Slack messages

Anthropic launched Claude Tag, an always-on AI teammate inside Slack. It reads channel messages over time to build institutional knowledge, not just answer one-off questions. The play is clear: turn organizational context into Claude's long-term memory and embed the model into daily workflows. The post doesn't disclose pricing or launch date. Privacy and permission controls are mentioned only briefly—admins can configure them, but details are thin.

Why it matters: Anthropic launched Claude Tag, a Slack integration that continuously reads channel history to build business context — not a Q&A bot. The product shape is novel, the mechanism differs from standard RAG, and Slack-heavy teams will feel both curiosity and privacy concern. TechCr...

Jun 23Tuesday

Hacker News front page

AI's Affordability Crisis

David Rosenthal breaks down the subsidy math behind AI platforms. SemiAnalysis found a $200/month subscription can burn $14,000 in OpenAI tokens or $8,000 in Anthropic tokens—a 40–70x subsidy. Ed Zitron obtained OpenAI's 2025 financials: $13.07B revenue against $34B in costs, a $38.5B net loss, with sales and marketing alone eating 44% of revenue. The post argues the 'first one's free' drug-dealer model is hitting a wall as business adoption fails to keep pace with the cash burn.

Why it matters: David Rosenthal lays out a clear cost analysis using SemiAnalysis test data and Ed Zitron's leaked OpenAI financials: users pay $200/month while platforms subsidize 40-70x the actual cost. Not a new argument (Sequoia's Cahn raised it in 2023), but the data is updated and the c...

AI Chat-Group Daily (群聊日报)

Anthropic's 400K Claude Code sessions report: managers beat engineers on verified success

Anthropic analyzed ~400K real Claude Code sessions and found managers scored highest on verified success—tasks requiring git commits, merged PRs, and passing tests—beating software engineers. The group flagged selection bias: managers chase aha moments, not corner cases. Sakana AI launched Fugu, a multi-model orchestration product hitting SWE Bench Pro 73.7, explicitly marketed as free from US export controls. OpenCode's public dashboard shows 136K DAU with DeepSeek at 53.6% share—domestic Chinese models dominate. On the practical side, the root cause of Gemini 3.5 Flash output truncation was traced to thinking monologues consuming tokens before the response could start; increasing the output limit helps but remains unstable.

Why it matters: Anthropic's research on 400K real Claude Code sessions shows managers scoring higher than engineers on verified success, with selection bias already flagged in the chat. HKR all hit, but this is a second-hand digest from a daily roundup rather than a primary source, so 72 at t...

Latent Space

SpaceX hits $28B/yr GPU rental run rate, rivaling top neoclouds

SpaceX signed its third GPU deal, leasing GB300s to Reflection AI at $150M/month. Combined with earlier Anthropic ($1.25B/month) and Google ($920M/month) contracts, Jamin Ball estimates SpaceX's GPU revenue at $2.32B/month, annualizing to ~$28B — roughly double CoreWeave's current revenue. Blackwell pricing exceeds $10/hour, which is on the high side. The post doesn't disclose total GPU capacity, utilization, or whether these deals are exclusive.

Why it matters: SpaceX's third major GPU deal lands, with annualized revenue hitting $28B — the numbers are concrete and the tenant breakdown is useful for anyone tracking the compute supply chain. Not scoring higher because this is a second-order financial analysis, not a first-party product...

Hacker News front page

Benchmarking Mythos: can other models find the same security bugs?

The author built a blind benchmark from 9 real bugs found by Anthropic's Mythos. Models get the source tree and a target file, no hints. All models underperformed expectations. Gemma 4 MoE led with 4/9 detections but had multiple retries due to crashes; GPT 5.5 Pro burned $100 after only 4 cases. The post doesn't report Mythos's own score on the same setup, nor whether these bugs were found in one shot. Sample is tiny and single-run, so treat rankings as directional.

Why it matters: The author built a blind benchmark using 9 real Mythos-discovered bugs from Anthropic's own disclosures. All models underperformed expectations—Gemma 4 MoE led at 4/9 but crashed repeatedly, GPT 5.5 Pro burned $100 on just 4 cases. First public blind test using actual Mythos v...

Computing Life · Share · Yage

Sakana AI's Fugu is a trained orchestrator that learns to manage other models, moving coordination from code into weights

Sakana AI's Fugu is a trained multi-model orchestrator. It decides which models to call, how to assign roles, and how to verify results—no hand-coded rules. Two ICLR 2026 papers back it: TRINITY trains a sub-20K-parameter coordinator via evolutionary strategy, Conductor trains a 7B orchestrator via RL. Fugu Ultra scores 93.2 on LiveCodeBench, beating Fable 5's 89.8, but trails on SWE-Bench Pro and HLE. Sakana deliberately hides which models are called per request and their raw outputs, calling routing info proprietary. All data is self-reported with no independent third-party replication and no head-to-head against OpenRouter Fusion. The direction holds, but transparency is still missing.

Why it matters: Sakana AI turned multi-model orchestration from hand-coded rules into a trained capability, backed by two ICLR 2026 papers with concrete mechanisms. Score held back because the post doesn't disclose LiveCodeBench numbers or pricing, and the product just launched without commun...

AI HOT (Curated Pool)

Full Claude Desktop experience now available on AWS, Google Cloud, and Microsoft Foundry

Anthropic is bringing the full Claude Desktop experience to AWS, Google Cloud, and Microsoft Foundry. Enterprises can now deploy Claude Desktop directly inside their own cloud environment without routing through the public internet. The post doesn't spell out pricing or exact launch dates, but confirms this targets Claude Enterprise customers. For teams with strict compliance needs, keeping data inside their own cloud account is a concrete win.

Why it matters: Anthropic is bringing the full Claude Desktop experience inside AWS, Google Cloud, and Microsoft Foundry, letting enterprises run it within their own VPC. Real compliance win for finance and healthcare, but no pricing or GA date disclosed — Enterprise-only for now, so rollout ...

Hacker News front page

AI Has Already Killed Academia as We Know It

A tenured professor and editor-in-chief argues academia's reward system is broken under AI. Students use two paid models to polish assignments into undetectable, high-scoring work, penalizing honest writers. In research, someone with Consensus and Claude can produce a publishable paper a day, overwhelming peer review. Grant applications face the same volume attack: a five-person team can submit ten proposals per cycle by rotating lead applicants, with AI fixing budget errors and missed citations. The author says institutions are moving too slowly—teaching assessments are being redesigned, but research metrics still treat AI-inflated output as real.

Why it matters: First-person account from a tenured professor showing how academic reward systems have collapsed under AI, with specific cheating chains and research production details — not vague hand-wringing. Docked slightly for being a personal blog rather than institutional release, and ...

Jun 22Monday

Hacker News front page

Claude Code's 'extended thinking' is a summary, not the model's real reasoning

Patrick McCanna inspected Claude Code's local session logs and found that 'thinking blocks' contain only a 600-character signature, not readable reasoning. Anthropic encrypts the actual reasoning into that signature, holds the decryption key server-side, and the API returns a summary. Full thinking output requires an enterprise agreement. The 'extended thinking' you see in the terminal is a post-hoc summary by Fable/Opus, not the raw reasoning that drove the agent's actions. Don't count on this as an audit trail, and the docs are indirect enough that you might miss the caveat without coffee.

Why it matters: The author dug into Claude Code's local session logs and found that thinking blocks contain only encrypted signatures — the API returns a summary generated by Fable/Opus, not the raw reasoning. This is a real constraint for teams relying on thinking output for audits or debugg...