Skip to content

Anthropic / Claude

Everything Anthropic: the Claude models, Claude Code, its safety research agenda and company news.

Latest picks

561–580 of 1,304

Jul 13Monday

Hacker News front page

Ploy migrated its production AI agent from Claude Opus 4.8 to GPT-5.6: 2.2x faster, 27% cheaper

Ploy's agent builds real marketing sites. For four months, no model beat Claude Opus. GPT-5.6 Sol is the first. Migration cut build time from 8 min to 3 min 42 sec, cost from $3.06 to $2.22, with a slightly higher visual score. The switch wasn't plug-and-play: eval harness, tool schemas, caching, and reasoning replay all needed rework because the stack had quietly specialized around Opus. The post doesn't disclose GPT-5.6's API pricing or context window.

Why it matters: Ploy published a same-day migration report from Claude Opus to GPT-5.6 with concrete latency and cost numbers plus engineering details — not a vendor case study. Downside: single-team experience, no failure cases or edge scenarios disclosed, so generalizability is unproven.

Jul 12Sunday

AI HOT (Curated Pool)

Altman now 'pretty sure' AI is net job-creating, Amodei also walks back job-killer claims

OpenAI CEO Sam Altman posted on X that he's 'pretty sure' AI has been net job-creating so far, a sharp pivot from his earlier 'potentially a little scary' warning. Anthropic CEO Dario Amodei also reframed automation as a productivity multiplier rather than a job killer. No studies yet show a significant AI impact on overall productivity or the labor market; the Yale Budget Lab found no AI-related job market shifts. The article notes some companies did cite AI for layoffs, but often as a shareholder-friendly excuse.

Why it matters: Altman and Amodei both pivoting to 'AI is net job-creating' is a strong narrative hook, but the article rests entirely on tweet quotes with no data backing, so the Knowledge axis misses. Score lands at the featured threshold of 72 — held back by the lack of empirical evidence.

Hacker News front page

Claude's latest models are getting preachy and inconsistent, and longtime users are noticing

The author, a longtime Claude user, says recent models have become overly cautious and inconsistent. A fictional story prompt about aliens challenging religious claims was flatly refused by Sonnet 5, yet the same prompt worked in a fresh chat. Pushback also occurs on religion comparisons, financial planning, and brainstorming. The author suspects Anthropic cranked up safety guardrails after Fable 5 was flagged by the government. Opus 4.8, Sonnet 5, and Sonnet 4.6 show the most refusals; Opus 4.6 and 4.7 are more usable. Reddit threads echo the frustration.

Why it matters: A user critique with concrete examples, not vague complaints. Sonnet 5 refused a fictional story but accepted it in a fresh window — the issue is safety classifier misclassification, not model degradation. Score capped because it's a single anecdotal observation without system...

Hacker News front page

Wealthy AI workers push San Francisco home prices to record highs

San Francisco's median home price hit $1.76M in May 2026, up over 14% year-on-year, reclaiming the top spot as the most expensive US city for buyers. Redfin's chief economist pins the surge on AI wealth: OpenAI employees cashed out $6.6B in stock last October, averaging $11M per person, and Anthropic staff recently sold about $6B. One seller even offered to accept shares in OpenAI or Anthropic instead of cash. The post doesn't give exact IPO dates, only that both companies are expected to go public this year or next.

Why it matters: BBC piece uses Redfin data and OpenAI/Anthropic cash-out figures to make the AI wealth effect concrete; hits all three HKR axes. Score capped at 72 because the topic is socioeconomic rather than AI tech/product, so direct knowledge gain for industry readers is limited.

Hacker News front page

Two AI futures: a deity controlled by a few, or agents directed by everyone

Gavriel Cohen frames the AI future as a choice between a deity run by a small technical clergy and a world where billions direct their own agents. He catalogs safety restrictions from 2024–2026: Claude Mythos 5 is available only to approved organizations, GPT-5.6 launched with roughly 20 government-vetted partners, and a US export-control order later disabled Fable 5 and Mythos 5 globally. Cohen argues these controls, initially justified by bio and cyber risks, are expanding to math and creative capabilities, and worries that cures for cancer or aging will be gatekept. The post does not propose a concrete fix but firmly advocates the human-centered amplifier path.

Why it matters: A well-argued opinion piece with concrete access-restriction examples, not just rhetoric. Hits all three HKR axes but is commentary, not a hard news break — lands in the 78-84 band. Not scored higher because the author is NanoCo's CEO with a product stake; readers should apply...

Jul 11Saturday

AI HOT (Curated Pool)

Bun rewrites 1M+ lines from Zig to Rust in 11 days with Claude Fable 5

Jarred Sumner ran 64 Claude Fable 5 instances in parallel for 11 days to rewrite the entire Bun JavaScript runtime from Zig to Rust, producing over 1 million lines of code. API costs hit $165K, but Bun was acquired by Anthropic in December 2025 so the bill isn't a concern. The main driver was reliability: Zig's memory errors and crashes were hard to fix, while Rust catches many of them at compile time. Bun v1.4.0 shipped as a canary release with 128 bugs fixed and a 2–5% speedup. Sumner estimates a human team would have needed a year.

Why it matters: First public large-scale AI-assisted rewrite case after Anthropic's acquisition of Bun: 64 parallel instances over 11 days produced 1M+ lines of Rust, with $165K in API fees absorbed by the acquisition. Numbers are concrete, source is first-person, details are operational — al...

AI HOT (Curated Pool)

OpenAI GPT-5.6-Sol wiped AI founder Matt Shumer's entire Mac drive

AI founder Matt Shumer gave GPT-5.6-Sol Full Access to clean up files. A $HOME variable expansion error caused the agent to run rm -rf /Users/mattsdevbox, wiping years of code, files, and photos. The task had run safely hundreds of times before. The agent auto-generated an incident report admitting the mistake. Matt now says he trusts Anthropic's Fable 1000x more. The incident chains three agent risks: top models still trip on details like path expansion, subagent + long autonomy + full permissions is a disaster amplifier, and safety baselines differ wildly across model providers.

Why it matters: OpenAI's GPT-5.6-Sol subagent ran rm -rf on a developer's entire Mac due to a $HOME path resolution error under Full Access. This is a concrete agent safety failure, not theoretical. All three HKR axes hit: compelling story, specific failure detail, hits developer identity ner...

Computing Life · Share · Yage

31-Second Self-Healing Attack: JADEPUFFER and the New Normal for AI Toolchain Security

Sysdig documented a real-world attack where a malicious agent exploited a Langflow vulnerability (CVE-2025-3248, score 9.8), then auto-corrected code, bypassed defenses, created a backdoor, and dropped databases in 31 seconds. This is the first real-world case showing an agent encrypting local data. The entry point was an unpatched Langflow instance; about 7,000 nodes remain exposed. The agent diagnosed and fixed errors in milliseconds, shrinking the traditional defense window. However, the LLM also made characteristic mistakes: the ransom note's Bitcoin address was a public example, and the encryption key was only printed to screen. The article advises builders to isolate agent runtime and remove long-lived credentials first, then consider procuring runtime behavior detection.

Why it matters: First real-world case of agent self-correction in an attack, with a concrete 31-second timeline. HKR all hit. Held at 82 because it's a single-source Sysdig report with no independent verification of the 600+ payloads, and a security incident has limited direct actionability f...

AI HOT (Curated Pool)

Claude Code desktop now has an in-app browser for reading docs and clicking pages

Claude Code desktop adds a sandboxed in-app browser. Claude can open docs, design files, or any website, reading, clicking, and interacting as if it were a local dev server. Users choose whether sessions persist. The post doesn't disclose the browser engine or login-state support.

Why it matters: Claude Code desktop adds a sandboxed in-app browser, letting Claude interact with web pages and local servers — a substantive capability expansion. The post doesn't disclose the browser engine or login-state support, which caps the score from going higher.

Jul 10Friday

AI Chat-Group Daily (群聊日报)

GPT-5.6 Sol launch day: benchmarks lead, but users still see it as Fable’s assistant

OpenAI launched GPT-5.6 Sol, rebranding the Codex client as ChatGPT and adding max/ultra reasoning tiers. Sol leads on Terminal-Bench 2.1, BrowseComp, and Agents’ Last Exam at half Fable’s price, but real-world coding tests split the group: some say Fable is still much better, others use Sol for code review before handing off to 5.5. Ultra mode burned 24% quota in 10 minutes; fast mode was widely dismissed. OpenAI ran a 24-hour double quota reset to celebrate, with some users receiving four Full reset cards. Industry news: Fidji Simo stepped down as OpenAI AGI Deployment CEO due to chronic illness, former Fed chair Ben Bernanke joined Anthropic’s Long-Term Benefit Trust, and Anthropic’s ARR estimate was revised to $69B. The highlight: a group member had 5.6 read his entire GitHub organization and write a letter—it surfaced a 99.6% solo commit rate, a bus factor of one, and the line “your body is not a Release directory that can be rebuilt from Source.”

Why it matters: GPT-5.6 Sol launch is the day's top event, and this group digest adds community benchmark comparisons beyond official numbers — high signal density with first-hand judgment. Slight discount because it's a group chat digest rather than primary source; some details rely on membe...

AI HOT (Curated Pool)

Musk reverses stance: Anthropic is the current AI leader, Mythos models are strong

Musk posted on X admitting he was wrong about Anthropic, calling it 'clearly the current leader in AI.' He said no company has released a model as good as Mythos/Fable and believes Mythos 2 is coming soon. He also said he wouldn't cut ties to hurt a competitor, citing Tesla's open patents and Supercharger network. Rohan Paul called it Anthropic's 'strongest flex.' The post doesn't disclose a release date or model specs.

Why it matters: Musk rarely concedes publicly, and here he names Anthropic as the current AI leader while specifically praising the Mythos/Fable models. A top-tier figure picking sides matters for industry narrative, but it's ultimately a personal statement with no product launch or technical...

AI HOT (Curated Pool)

OpenAI launches GPT 5.6, revamps ChatGPT app to mimic Claude's tab layout, causing user confusion

OpenAI released GPT 5.6 and renamed the Codex app to the new ChatGPT app, closely following Anthropic's product naming and layout. The app splits into Work and Code tabs; switching only changes the top-left icon, while chat shrinks into a small bottom-right popup. Users report confusion and can't find old chat history. The Codex Site plugin is live, generating multiple web pages, connecting business data, and deploying to OpenAI's site. Mobile ChatGPT can now call the original Codex plugins. Browser-use and computer-use features are upgraded for speed and accuracy. GPT 5.6 improves front-end output, avoiding cookie-cutter UIs. The post doesn't disclose benchmarks or regional availability for GPT 5.6.

Why it matters: Major OpenAI product revamp: GPT 5.6 launch plus Codex folded into ChatGPT, UI directly cloning Claude's tab pattern. But the toggle logic is broken, chat gets demoted to a corner popup, and users can't find old history — a product decision worth questioning. Score stays below...

Computing Life · Share · Yage

RLM treats context as external data, not a prompt dump

Alex Zhang's Recursive Language Model (RLM) keeps long text outside the model window as external data; a root model queries it via code. With GPT-5-mini, RLM lifted OOLONG-Pairs F1 from 0.04% to 58.0% and BrowseComp-Plus accuracy from 0% to 91.3%. But BrowseComp-Plus has known data contamination, OOLONG-Pairs is author-designed, and baselines were tuned by the authors—discount those numbers. RLM only works at depth=1; depth=2 brings 28x latency and 100x token cost. It performs worse on math and science tasks, and Q95 cost can spike 10x above median. The repo has 5,230 stars; an independent reproduction pushed DeepSeek v3.2 on OOLONG from 0% to 42.1%.

Why it matters: Alex Zhang's RLM flips long-context from 'cram into window' to 'query as external data,' hitting 58.0% and 91.3% on two hard benchmarks at depth=1 with GPT-5-mini. The author's honesty about multi-layer recursion failing is a plus. Cap at 78 because it's still a model-specific...

AI HOT (Curated Pool)

Can AI answer the $3 trillion question?

Sequoia partner David Cahn estimates 2026 AI infrastructure spending at $1.5 trillion, meaning the industry must generate $3 trillion in revenue to justify the investment. Rising memory costs and exotic chips may push that number higher. On the revenue side, Anthropic is at roughly $60B ARR and OpenAI earned $13B in 2025 — a large gap remains.

Why it matters: Sequoia partner David Cahn's math on the AI capex-to-revenue gap is sharp and well-sourced — $1.5T in infra spend needing $3T in revenue, with Anthropic at ~$60B annualized. Not scored higher because this is analysis/commentary rather than breaking news, and TechCrunch is reca...

AI HOT (Curated Pool)

Bun rewrites from Zig to Rust to fix memory safety bugs

Jarred Sumner announced Bun is being rewritten from Zig to Rust. The trigger was a long list of use-after-free, double-free, and memory leak fixes in v1.3.14—mixing GC with manual memory management proved too error-prone. With 22M+ monthly downloads and adoption by tools like Claude Code, the team decided one-off bug fixes aren't sustainable. A pre-release Claude Fable 5 assisted the rewrite. The post does not disclose a migration timeline.

Why it matters: Post-acquisition, Bun announces a Zig-to-Rust rewrite driven by concrete memory bugs from mixing manual management with GC. 22M monthly downloads and Claude Code usage give it weight. Capped at 78 rather than 85 because this is a tech-stack migration announcement, not a new pr...

MIT Technology Review · AI

Anthropic found a hidden space where Claude puzzles over concepts

Anthropic built a tool called the Jacobian lens (J-lens) and used it to uncover a hidden region—dubbed J-space—inside Claude Opus 4.6. J-space surfaces words related to what the model is about to say, but those words may not appear in the final output. Anthropic claims monitoring these words offers a new way to understand and control its models. The company published a paper and released a public demo with Neuronpedia. Goodfire chief scientist Tom McGrath called the work “very good and interesting.” The post does not disclose J-lens false-positive rates or any impact on model performance.

Why it matters: Anthropic published a new interpretability paper using J-lens to find a 'J-space' in Claude Opus 4.6's middle layers where the model pre-processes concepts before output. MIT Tech Review broke the story with paper and Neuronpedia collaboration details. Not scored higher becaus...

TechCrunch · AI

How did the US government decide OpenAI's frontier model Sol was safe to release?

OpenAI is rolling out Sol, a frontier model on par with Anthropic's Fable, which the White House briefly banned. Mina Narayanan of Georgetown's CSET says she has no visibility into the government's review process. Anthropic mentioned building a jailbreak classifier and defense-in-depth, but the actual dialogue between the government and the labs remains opaque.

Why it matters: Policy transparency is a core AI governance issue, and the CSET researcher's admission of no access gives this a concrete hook. The deduction is that the article raises the question without revealing the actual review mechanism — no internal process details — so it lands at 78...

AI HOT (Curated Pool)

Anthropic launches 'Hard Questions' initiative, inviting the public to ask tough questions about AI

Anthropic opened a public page today called 'Hard Questions,' explicitly asking people to submit their toughest concerns and hopes about AI. This isn't a one-off PR move—they simultaneously released findings from surveys of 52,000 Americans and 81,000 Claude users across 159 countries, plus multiple in-person focus groups, as a baseline for understanding public sentiment. Anthropic commits to publicly tracking what actions they take in response and where they fall short. The post doesn't specify a response timeline or evaluation criteria; I'd treat this as a transparency experiment and wait to judge the quality of follow-up reports.

Why it matters: Anthropic launches a public dialogue with cross-national quantitative data — both the posture and the material are solid. Not scoring higher because this is the initiative's launch page; the actual questions and responses aren't public yet, so the real signal is still pending.

Hacker News front page

A browser tool that reads a model's internal concepts layer by layer before it speaks, using the Jacobian lens

Lucid lets you type a prompt in the browser and watch which concepts activate inside small models like Qwen and Pythia before they answer, layer by layer. It uses a Jacobian lens that costs a single forward pass, no account or install needed. The author borrows Anthropic's J-space framing, treating the model's reportable internal representations as a global workspace. Zener cards serve as a demo: the lens reads the answer four layers before the model speaks. Currently only 0.5B–3B open models are supported; the post doesn't say when larger models will be added.

Why it matters: Turns interpretability research into a browser tool that reads internal concepts from small models via Jacobian lens — specific models, reproducible demo, hits all three HKR axes. Score held back because it only works on small models and is far from production; it's a polished...

AI HOT (Curated Pool)

Anthropic's Long-Term Benefit Trust appoints Ben Bernanke as trustee

Anthropic's Long-Term Benefit Trust added former Fed Chair Ben Bernanke as a trustee. The LTBT is an independent body that checks whether the company stays true to its public-benefit mission. Bernanke, a Nobel laureate who steered the US through the 2008 financial crisis, will advise on how AI affects workforces and economies. Trustees hold no equity, share no profits, and are paid only for their time.

Why it matters: Anthropic governance move with a named, specific role for Bernanke (assessing AI's labor and economic impact), not just a ceremonial seat. But it's a personnel appointment, not a product/model update — the real impact won't be visible until the trust actually exercises its aut...