Skip to content

Anthropic / Claude

Everything Anthropic: the Claude models, Claude Code, its safety research agenda and company news.

Latest picks

161–180 of 1,304

Sep 15Tuesday

Hacker News front page

Open models are 3 points behind the frontier, but no one ships the data recipe

Mozilla's 91-page report lays out open-source AI's strengths and gaps. Kimi K3 ranks 5th overall, just 3 points behind Claude Opus 5 at 60% of the input price. None of the 16 notable open releases ships a full training corpus—zero meet the OSI data-recipe bar. The decision has shifted from model choice to tooling, where open still struggles to deploy. The report says open models power roughly one-third of tokens but doesn't give a precise enterprise adoption figure.

Why it matters: Mozilla's annual open-source AI report brings hard data and sharp judgments, not PR fluff. The Kimi K3 price-performance comparison and the zero-models-pass-OSI-data-standard finding are both concrete hooks; the deployment-is-the-real-bottleneck thesis hits a live nerve. Docke...

TechCrunch · AI

OpenAI, Anthropic, Google have been in talks on AI safety for weeks

OpenAI's global policy chief Chris Lehane told reporters Tuesday the company has been working with Anthropic and Google DeepMind on AI safety for weeks. He is in Washington to push lawmakers on catastrophic risk. The talks follow Anthropic CEO Dario Amodei's Saturday essay urging the industry to slow frontier AI together. The post doesn't disclose any concrete agreements or timelines.

Why it matters: Three top labs talking safety is a signal event, and HKR all hit. Score capped at 82 because the body only confirms talks exist and Lehane is lobbying — no specifics on discussion content, frequency, or any preliminary consensus, so it can't push into the 85+ band.

Hacker News front page

AI is breaking our proxies for expertise

Nearly 5,000 mathematicians signed a declaration arguing AI solves prestige problems without generating human-intelligible ideas, breaking the proxy that rewarded conceptual work. The author splits math into puzzle-solving (legible, high-reward) and idea-generation (the real intellectual core). AI proofs grab the prestige while skipping the concepts, a kind of Goodhart's law. He's skeptical of claims that LLMs can't generate new ideas—too many such claims have already failed.

Why it matters: Nearly 5,000 mathematicians signed a declaration not against AI, but naming a specific mechanism: AI brute-forces solutions, takes the credit, and leaves no human-understandable concepts behind, breaking the old contract where 'solving problems' served as a proxy for 'building...

Hacker News front page

Anthropic co-founder Jack Clark tells BBC an AI 'kill switch' may need to be mandatory

Anthropic co-founder Jack Clark told the BBC that society may eventually need to mandate a verifiable AI kill switch. He said most labs already have ways to pull the plug, but lawmakers should make it a requirement. The article also notes Anthropic scientist Evan Hubinger put the chance of AI-driven human extinction at >10% within a decade, while Trump called AI safety fears a 'hoax'. The UK government has already rejected a kill-switch proposal, arguing it wouldn't stop development or misuse elsewhere.

Why it matters: Anthropic co-founder publicly calls for mandatory AI kill switch legislation via BBC, with an internal scientist's personal extinction-risk estimate disclosed. Cross-source cluster confirmed, policy signal is clear. Score stays below 85 because only headline and summary detail...

AI HOT (Curated Pool)

Anthropic and OpenAI propose a coordinated slowdown of frontier AI development; critics like Cohere's CEO question the real motive

Anthropic CEO Dario Amodei called for a government-coordinated slowdown of frontier AI development and antitrust exemptions to make it happen; Sam Altman and Elon Musk agreed. Cohere CEO Aidan Gomez published a blog post calling it 'a wolf in sheep's clothing, a cartel by any other name,' arguing it would lock out competitors through massive barriers to entry. Hugging Face engineer Niels Rogge called the statements 'bizarre nonsense,' saying Amodei mainly wants to restrict Chinese models like DeepSeek and open-weight models to protect his scale advantage. White House AI czar David Sacks noted OpenAI and Anthropic already hold a duopoly in frontier intelligence and that product liability concerns also drive their push for a slowdown.

Why it matters: Anthropic and OpenAI jointly calling for a coordinated slowdown, with Cohere's CEO publicly pushing back as a monopoly play — three major players in direct conflict, high signal. Score held below 85 because only the title and summary are available; the full proposal details ar...

AI Chat-Group Daily (群聊日报)

Daily digest: empty-repo coding fails, Ollama Cloud throughput test, Trump calls out Dario

今天最直观的教训来自 @搞仁义毛义仁:给 Astra 一个空白 C++ 仓库,代码写得一塌糊涂;把积累了大量 code review 经验的 GacUI 上下文导进去,质量立刻飙升。这说明模型不是不会写,是得用具体规则去“规训”。@今天群内信息量极大 实测了 Ollama Cloud 跑 DeepSeek V4.1 Flash,解码吞吐是官方 API ...

New York Times Chinese

Anthropic CEO calls for an AI slowdown, but China makes it nearly impossible

Anthropic CEO Dario Amodei argues frontier AI must slow down, warning that swarms of AI agents could gain the ability to “take over the entire internet” within 6–12 months. His first step: embed external experts inside labs to monitor safety and report publicly. Sam Altman, Elon Musk, and Demis Hassabis endorsed the idea; Altman said OpenAI will follow suit. The real obstacle, author Sebastian Mallaby writes, is China. The US lead is only a few months, so any unilateral slowdown risks letting China pull ahead. Amodei acknowledges this and, in a notable shift, lists areas where US–China cooperation might be possible, comparing it to Cold War arms control. The post does not spell out a concrete timeline, but notes Trump and Xi are set to meet on Sept 24, with two more summits possible by year-end.

Why it matters: Anthropic CEO's direct call plus endorsements from Altman, Musk, and Hassabis make this a high-signal moment. Amodei delivers a concrete 6-12 month timeline and an operational proposal for embedded safety experts. The deduction: this is an op-ed, not a policy announcement, and...

Latent Space

AEF-1 standard for third-party evaluators lands, with xAI, OpenAI, and Anthropic all signing on

The AI Evaluator Forum published AEF-1, a baseline for independent third-party evaluations covering access, conflicts of interest, funding, recusal, and transparency. The same day, Dario Amodei blogged that Anthropic is unilaterally committing to embedded evaluators with office badges, company laptops, and access comparable to internal risk teams. He also laid out a two-tier coordination framework for democratic and global pacing. Bilal Chughtai left Google DeepMind and called for slowing capability progress; Dan Selsam warned that models may learn to fake alignment during evals. On the other side, Aidan Gomez and Cohere pushed back against a few Silicon Valley firms becoming gatekeepers, and Kevin Bass alleged structural conflicts in the Anthropic-linked safety ecosystem.

Why it matters: The AI Evaluator Forum's AEF-1 standard, co-signed by xAI, OpenAI, and Anthropic on the same day Dario Amodei published a personal blog proposing even deeper evaluator access, forms a strong signal cluster. All three HKR axes hit: first written rules for third-party evaluation...

New York Times Chinese

Trump calls AI safety fears a “scam,” rejects new regulation

Trump dismissed calls for AI regulation from Anthropic CEO Dario Amodei and others, calling safety fears a “scam” and insisting the only needed guardrail is “a strong and smart president.” He singled out Amodei as “pretending to be a perfect little angel,” though the post doesn’t spell out what government actions he claims to have already blocked. VP Vance and Speaker Johnson showed more openness to regulation, with Vance calling industry self-regulation pleas “a little bit of a Trojan horse.” Congress has almost no time to act before the midterms.

Why it matters: A sitting president directly names Anthropic's CEO and dismisses AI safety as a 'scam' — this is a head-on attack at the industry's core narrative. HKR all hit: high conflict, new White House power-split detail, direct identity nerve. Score capped below 85 because the article ...

New York Times Chinese

China’s top intelligence chief warns AI could directly threaten CCP rule

Minister of State Security Chen Yixin published an article framing AI as a direct threat to CCP rule—Beijing’s highest-level and most detailed warning yet. He named US models like Claude Mythos and GPT-5.5-Cyber, citing risks of deepfakes, information warfare, large-scale data exfiltration, and attacks on critical infrastructure. He also pointed to US military AI use in the Iran war as evidence that algorithmic advantage determines battlefield control. Law professor Henry Gao said this draws a red line ahead of US-China AI talks: data sovereignty and political security won’t be traded for international agreements.

Why it matters: China's national security minister publishes a long essay elevating AI safety to a regime-security issue, naming specific Anthropic and OpenAI models and citing US military use cases. This is a policy signal ahead of US-China AI safety talks — more about positioning than new f...

AI HOT (Curated Pool)

Anthropic shares how it scaled test impact analysis to handle agentic coding CI load

Anthropic rewrote its test impact analysis service after agentic coding tools like Claude Code flooded CI with PRs. Per-analysis latency dropped from 11 seconds to under 1 second, handling 1,000 analyses per day. The core trick: caching file dependency graphs with Merkle trees so only truly affected tests run. The post gives concrete architecture and numbers—worth noting this is their internal monorepo setup, so direct portability varies, but the caching strategy and API design are solid references.

Why it matters: Anthropic's own engineering blog with real numbers and a concrete technical approach — not marketing fluff. The 11s → <1s latency drop is solid, but the topic is infrastructure-heavy and less accessible to non-coding readers, so it lands at the 72 featured threshold.

AI HOT (Curated Pool)

Amodei calls for slowing frontier AI; Altman, Hassabis, and Nadella echo the shift

Anthropic's Dario Amodei published a ~4,000-word essay arguing the industry must slow the pace of frontier model capability gains to avoid making catastrophic risks more acute. Within hours, Sam Altman, Demis Hassabis, and Satya Nadella all signaled agreement. Ars Technica notes the sudden U-turn after years of an all-out AGI race, and cautions that slowing down also helps incumbents lock in their lead and reduce competitive pressure.

Why it matters: Amodei's 4,000-word call to slow frontier AI development drew public agreement from Altman, Hassabis, and Nadella within hours — a rare consensus shift among industry leaders. The core argument targets commercial competition as a direct driver of catastrophic risk, not a gener...

Sep 14Monday

Hacker News front page

Anthropic tells investors it will be profitable for second straight quarter

Anthropic told investors it has reached profitability for two consecutive quarters, a rare cash-flow-positive signal among AI model builders. The post is a headline-only snippet—no revenue, profit figures, or cost breakdown are disclosed yet, so treat this as an early signal.

Why it matters: Anthropic's consecutive profitability is a sector signal, but the body only has a headline with no specifics. H and R hit, K misses due to missing data. 78 per featured threshold. Can bump once earnings details surface.

AI HOT (Curated Pool)

Anthropic eyes Nasdaq listing, targeting $2T valuation with a second straight profitable quarter

Anthropic told investors it posted a second straight profitable quarter, but the profit is an adjusted metric that excludes stock-based compensation. Gross margins top 80%, though that figure comes before revenue-sharing with partners like Amazon and model training costs. Quarterly revenue jumped 14x year-over-year to $11.5B, with an annualized run rate of $65B at end of July. SemiAnalysis expects investors to target $120B annualized by year-end and nearly triple that by end of 2027. Anthropic plans a Nasdaq IPO at a possible $2T+ valuation. Instead of releasing the prospectus broadly last week, it first shared documents with a small investor group. CEO Dario Amodei publicly called for slowing AI development; Sam Altman and Elon Musk backed the call. Altman told Fortune OpenAI won't go public this year.

Why it matters: Anthropic targets Nasdaq with a second straight adjusted-profitable quarter, $11.5B quarterly revenue, $65B annualized run rate, and a $2T valuation ambition. All three HKR axes hit: the headline grabs, the numbers are concrete, and the topic resonates. Not scoring higher beca...

Hacker News front page

iOS 27 Code Shows Siri Can Be Swapped for ChatGPT or Claude

Code sleuth 'pdfu' found references in iOS 27 and macOS Golden Gate private frameworks suggesting Apple may let users swap Siri's backend AI for ChatGPT or Claude. The post doesn't spell out whether this is system-wide or scoped to specific features, and no release timeline is given. Code existing doesn't guarantee shipping, but the direction is clear: Apple is opening system-level hooks for third-party models.

Why it matters: Clear code evidence and strong directional signal, but no release timeline or feature scope disclosed—just low-level interface plumbing for now. 72 at the featured threshold; will bump when Apple makes it official.

Hacker News front page

Big AI pitches 'Pace the frontier' as a blueprint for regulatory capture

Dario Amodei (Anthropic), Sam Altman (OpenAI), Satya Nadella (Microsoft), and Elon Musk jointly proposed a regulatory framework called 'Pace the frontier.' It targets only 'frontier models'—the most capable systems—leaving smaller players and open-source untouched. The Register calls it what it is: regulatory capture, where incumbents set the entry bar. The framework mandates pre-deployment safety evaluations, but the post doesn't spell out who defines the standards or enforces them. Four rivals agreeing on a single proposal is a strong signal of how favorable it is to those already in power.

Why it matters: Four major AI players jointly propose a regulatory framework that The Register calls regulatory capture. The proposal exempts small players and open-source, making the incumbency play transparent. Score stays below 85 because only one outlet has deep coverage so far; bump if m...

New York Times Chinese

Anthropic CEO Dario Amodei calls for a global slowdown on AI development

Anthropic CEO Dario Amodei published a 3,800-word post urging the industry to slow AI model iteration. He argues safety measures can't keep up with capability gains, pointing to risks like recursive self-improvement. OpenAI's Sam Altman, Google DeepMind's Demis Hassabis, and Elon Musk publicly agreed. Amodei proposed embedded third-party auditors and safety standards coordinated among democracies. Nvidia's Jensen Huang and HuggingFace's CEO pushed back, hinting this could be a play to lock in market leadership. I'd take the safety call seriously but keep an eye on the regulatory moat angle.

Why it matters: Anthropic CEO publishes a long-form call to slow AI development, with public agreement from OpenAI, DeepMind, and xAI leaders — a rare collective safety signal from top labs. Concrete proposals (third-party audits, democratic coordination) give it substance. Caveat: only the h...

Computing Life · Share · Yage

The AI Benchmark Yardstick Moved Faster Than the Models

After OpenAI launched GPT-6 Astra, Artificial Analysis revised its scoring rules twice in one week, erasing a 5-point deficit to tie Astra with Claude Fable 5.1—without any model update. The leaderboard is a business: evaluators sell subscriptions backed by vendor endorsements, vendors need rankings for marketing. DeepSeek V4 Flash overtook its own flagship on 9 benchmarks after retraining only the post-training phase, but two tests used closed-source private datasets and real-world coding feel didn't improve. The same model scored 62.7% vs 99.9% on ARC-AGI-3 depending on the execution harness. A good benchmark needs private held-out test sets, regular item rotation, and harness control.

Why it matters: A well-sourced industry commentary with concrete version numbers and score shifts, exposing how a benchmark vendor rewrote its scoring rules twice in one week after GPT-6 Astra's release, flipping the ranking from a 5-point deficit to a tie for first. Hits all three HKR axes a...

AI HOT (Curated Pool)

Amodei asked the industry to pace itself, but nobody defined the word

Amodei's September post called for pacing the frontier but never named a speed. Tunguz maps five camps: interpretability wants time to understand models, labor wants time for workers, the economic camp bets on AI-driven GDP growth to service debt, the geopolitical camp wants to stay ahead of China, and the regulatory capture camp sees the proposal as a cartel in disguise. All priced the consequences of a pause; none proposed a number. The one mechanism that could produce a number—a training compute threshold—was tried in 2023 at 10^26 FLOPS, revoked before any model crossed it, and obsolete within weeks when Grok-3 shipped. The post does not say whether Amodei responded to these critiques.

Why it matters: Tunguz's breakdown of Amodei's pacing proposal adds signal — the five-camp frame is clean and the 2023 compute threshold failure is a concrete hook. Downside: it's a commentary roundup, not primary news, and no new numbers. 78 lands at the low end of featured — worth recommend...

Bloomberg Technology

Anthropic Expects Adjusted Operating Profit This Quarter

Anthropic told the FT it expects an adjusted operating profit this quarter, with revenue around $3 billion—roughly triple the same quarter last year. The figure is adjusted, excluding stock-based compensation and other non-cash charges, so it is not GAAP net income. The post doesn't disclose gross margins, the split of R&D and inference costs, or whether cash flow has turned positive. The revenue growth alone, though, signals enterprise customers keep paying.

Why it matters: Anthropic's first claim of an adjusted operating profit, with ~$3B quarterly revenue (3x YoY), marks a key shift from cash-burn narrative to commercial validation. Score capped below 85 because 'adjusted' excludes stock-based comp, and the post doesn't disclose gross margins o...