Skip to content

#Anthropic

9 today

May 7Thursday

Ben's Bites

Elon Doubled Limits

Ben’s Bites says Anthropic doubled Claude usage on paid plans via SpaceX’s Colossus 1. The issue also lists GPT-5.5 Instant, ChatGPT spreadsheet integration, and three Claude Managed Agents features. The title names Elon, but the post does not disclose exact limits.

Why it matters: HKR-H/K/R pass: the SpaceX Colossus 1 angle, 2x Claude usage, and quota pressure are all concrete. Missing exact caps, pricing, and rollout scope keep it in the low featured band.

AI HOT (Curated Pool)

Anthropic Institute Outlines Four Core Research Areas

Anthropic Institute named four research areas: economic diffusion, threats and resilience, real-world AI systems, and AI-driven R&D. The post says it will publish a more granular Anthropic Economic Index and study how AI tools speed AI research. The results will inform Anthropic’s Long-Term Benefit Trust.

Why it matters: HKR-K comes from 4 named research tracks and the Economic Index plan; HKR-R is strong on labor and governance. It is an agenda, not a model, product, or finished result, so it stays in the 72–77 band.

Latent Space

Anthropic-SpaceXAI's 300MW/$5B/yr Deal for Colossus I, ARR Growth Is 8000% Annualized

Anthropic announced a SpaceX compute partnership, doubled Claude Code’s 5-hour limits for Pro, Max, Team, and seat-based Enterprise, raised Opus API limits, and said Claude inference would ramp on Colossus within days; the post treats the 300MW and $5B-per-year figures as widely circulated but not canonized in Anthropic’s own announcement.

Why it matters: HKR-H/K/R all pass: the compute-deal numbers and Claude Code limit changes are concrete and practitioner-relevant. The 300MW/$5B/year claim is unofficial, so it stays below P1.

Xinzhiyuan · WeChat

Claude Managed Agents Add Dreaming, With Reported Task Completion Up to 6x

Anthropic added Dreaming, Outcomes, and multi-agent orchestration to Claude managed agents; Harvey reports about 6x higher task completion. Dreaming reads up to 100 sessions; one demo distilled 5.3M tokens into 98 rules, while Outcomes raised success by up to 10 points. Opus 4.7 and Sonnet 4.6 require access, with $0.08 per session-hour runtime fees.

Why it matters: HKR-H/K/R all pass: Anthropic adds Dreaming, Outcomes, and multi-agent orchestration with 100-session memory, $0.08/session-hour runtime, and Harvey’s ~6x completion claim. This is a same-day Claude agent update.

Synced · WeChat

Musk Announces xAI Dissolution, Leasing 220,000 GPUs to Anthropic

Musk confirmed xAI will dissolve, with Grok and X-related operations folded into SpaceXAI. SpaceX and Anthropic signed a deal giving Claude access to Colossus 1’s 220,000+ Nvidia GPUs and 300 MW of compute. The key change is quota: Claude Code’s five-hour rate limit doubles, and Pro/Max peak-hour cuts are removed.

Why it matters: HKR all pass: xAI dissolution plus 220k GPUs for Anthropic is a top-tier twist; 300 MW and Claude Code quota changes add testable detail; it hits compute, competition, and developer limits. Single-source status keeps it at 96.

Computing Life · Share · Yage

Anthropic Locks Up Compute Channels as xAI Rents Its Castle to a Rival

Anthropic signed four compute contracts in six months covering AWS Trainium, Google TPU, SpaceXAI Colossus 1, and CoreWeave; during the same window, xAI rented the Colossus 1 supercomputing center to a competitor while GPU utilization stood at 11%.

Why it matters: HKR-H/K/R all pass: Anthropic’s four compute deals and xAI leasing Colossus 1 create a sharp competitive angle with concrete numbers. Single-source strategy analysis keeps it in the 78–84 band, below same-day must-write news.

Computing Life · Share · Yage

Agent Filesystems: From Feeding Models Memory to Letting Models Browse Files

The article frames agent filesystems as a three-stage shift from raw context to memory systems to filesystem-as-context, covering design choices from Turso, Anthropic, Vercel, and Manus, and listing four overlooked blind spots.

Why it matters: HKR-H/K/R all pass, but this is design commentary rather than a product or research release. Named comparisons across Turso, Anthropic, Vercel, and Manus justify featured, not the 78+ band.

Financial Times · Technology

SpaceX to Rent Data Centre Capacity to Anthropic

SpaceX will rent data centre capacity to Anthropic, confirming a compute leasing arrangement. The RSS snippet says Anthropic is adding compute to match growth; the post does not disclose capacity, term, or pricing.

Why it matters: FT confirms SpaceX will rent data-center capacity to Anthropic. HKR-H comes from the odd pairing; HKR-K has the deal fact but no scale or price; HKR-R hits compute scarcity, so it clears featured but not 78+.

r/LocalLLaMA

Analysis of 922 Agentic Task Traces Finds DeepSeek v4’s Cost Edge in Caching

A Reddit user analyzed 922 agentic task traces and reported $0.01 per task for DeepSeek v4 Flash versus $1.52 for Opus 4.7. Both used about 960K tokens per task, but DeepSeek showed a 97% cache hit rate versus 87%, with a 0.02 cache read/write price ratio versus 0.08. The key issue is caching, not headline pricing.

Why it matters: HKR-H/K/R all pass: 922 agent traces tie a large cost gap to cache hit rate and cache read/write pricing. Reddit single-source data and incomplete method detail keep it in the 78–84 band.

TechCrunch · AI

DeepSeek could hit $45B valuation from its first investment round

DeepSeek could reach a $45B valuation in its first investment round, according to the title. The snippet says it rose in early 2025 after training an LLM with far less compute and cost; the post does not disclose round size, investors, or terms.

Why it matters: HKR-H/K/R all pass: DeepSeek’s first round targeting $45B is a strong valuation story. Missing investors, amount, and terms keep it in the lower 78–84 band, not P1.

Bloomberg Technology

Anthropic Signs Computing Deal With SpaceX to Meet AI Demand

Anthropic signed a computing deal with Elon Musk’s SpaceX to support growing Claude demand. The post does not disclose capacity, contract value, deployment timing, or infrastructure details. The key issue is whether SpaceX enters Anthropic’s long-term training or inference supply chain.

Why it matters: HKR-H and HKR-R pass: Bloomberg reports an Anthropic-SpaceX compute deal with a strong rivalry and supply-chain angle. HKR-K is weak because scale, spend, GPU count, and training/inference use are undisclosed.

Hacker News front page

Higher usage limits for Claude and a compute deal with SpaceX

Anthropic’s post has 91 HN points; the title says Claude gets higher usage limits and a SpaceX compute deal. The RSS body only lists links, 37 comments, and HN metadata. The post does not disclose limit multiples, compute scale, pricing, or timing.

Why it matters: Official title gives HKR-H/R: higher Claude limits and a SpaceX compute deal. HKR-K fails because the feed omits limit multiples, compute scale, pricing, and rollout timing.

May 6Wednesday

QbitAI · WeChat

Claude Team Tests New Training Method on Qwen

Anthropic proposed MSM training between pretraining and alignment fine-tuning. Tests on Qwen2.5-32B and Qwen3-32B cut misalignment from 68% and 54% to 5% and 7%. The key point is MSM complements AFT rather than replacing it.

Why it matters: HKR-H/K/R all pass: Anthropic offers a concrete MSM alignment method with Qwen2.5-32B and Qwen3-32B rate drops. It is strong safety research, not a model launch or major product update, so 82 fits.

Latent Space

AINews: Silicon Valley Gets Serious About Services

Anthropic and OpenAI announced enterprise services companies: Anthropic’s unnamed JV is funded with $1.5 billion, while OpenAI’s The Deployment Company has raised about $4 billion at a $10 billion pre-money valuation.

Why it matters: HKR-H/K/R all pass: the hook is labs turning into services operators, with $1.5B and ~$4B figures. The scale and OpenAI/Anthropic names put it in must-write territory.

Xinzhiyuan · WeChat

Coding at 12, Building a $2B Google Business at 28: He Tells Young People to Stop Chasing Coding

Xinzhiyuan says Alon Chen coded at 12 and managed a $2B Google business at 28. He argues Gen Z should stop chasing coding, citing 30% AI-written Microsoft code and 25%+ at Google. The sharper signal is execution, problem framing, and communication, not coding as a sole moat.

Why it matters: HKR-H/K/R all pass, but this is a career commentary piece, not a model or product release. The two AI-code-share numbers lift it above generic advice, placing it at the featured threshold.

May 5Tuesday

Xinzhiyuan · WeChat

Anthropic Tests Introspection Adapters on 700+ Problem Models for AI Auditing

Anthropic trained IA on nearly 700 labeled problem models, reaching 59% average success on AuditBench. It elicited hidden behaviors at least once from 50 of 56 denial-trained models, above 53% black-box auditing and 44% Activation Oracle. The key limit: IA has false positives, misses motives, and the post does not prove transfer to GPT or Gemini.

Why it matters: HKR-H/K/R all pass: the Anthropic audit method has a sharp hook, concrete benchmark numbers, and safety resonance. It stays in 78–84 because this is research progress, not a major Claude product release.

Synced · WeChat

Anthropic cofounder says AI self-improvement has a 60% chance by 2028

Anthropic cofounder Jack Clark says human-free AI R&D has over a 60% chance by end-2028. He cites SWE-Bench, CORE-Bench, MLE-Bench, and PostTrainBench: Claude Mythos Preview reaches 93.9% on SWE-Bench, and Opus 4.5 reaches 95.5% on CORE-Bench. The key signal is longer task horizons and post-training capability, not the “singularity” framing.

Why it matters: HKR-H/K/R all pass: a named Anthropic cofounder gives a 2028 timeline, backed by benchmark numbers. The headline is overheated, but the concrete claims and practitioner stakes justify P1.

May 4Monday

TechCrunch · AI

Anthropic and OpenAI Are Both Launching Joint Ventures for Enterprise AI Services

Anthropic and OpenAI will each launch joint ventures for enterprise AI services. Both partnered with asset managers to market enterprise AI products more aggressively. The RSS snippet does not disclose partner names, equity terms, pricing, or launch dates.

Why it matters: HKR-H and HKR-R are strong because two frontier labs mirror the same enterprise JV move. HKR-K is limited to the sales-vehicle mechanism; names, equity, pricing, and launch timing are not disclosed.

Financial Times · Technology

Blackstone and Goldman among backers for $1.5bn JV with Anthropic

Blackstone and Goldman are among backers of a $1.5bn joint venture with Anthropic. The consulting firm will advise Wall Street firms on AI deployment across portfolios; the post does not disclose ownership, products, or timeline.

Why it matters: HKR-H/K/R all pass: a $1.5bn Anthropic-linked JV backed by Blackstone and Goldman is a strong commercialization signal. Missing equity structure, product details, and timeline keep it below 85.

Import AI (Jack Clark)

Import AI 455: Automating AI Research

Jack Clark argues that no-human-involved AI R&D has a 60%+ chance of arriving by the end of 2028, citing SWE-Bench gains from Claude 2 at about 2% to Claude Mythos Preview at 93.9%, plus METR task horizons rising from 30 seconds in 2022 to 12 hours in 2026.

Why it matters: HKR-H/K/R all pass: Jack Clark anchors a >60% end-2028 automated-AI-R&D claim in SWE-Bench and METR numbers. This fits the 85–94 band for a notable figure’s AI-timeline essay, below model-release magnitude.

Xinzhiyuan · WeChat

Claude token rankings: Disney employee hits 460,000 calls in 9 days; Meta burns 60T monthly

Xinzhiyuan says Disney tracks Claude use via an AI Adoption Dashboard, with one employee making about 460,000 calls in 9 workdays. It also says Meta used 60 trillion tokens in 30 days, worth about $9B by public API pricing; the post does not show raw tables. The key issue is that input rankings are not outcomes.

Why it matters: HKR-H/K/R all pass: the hook is concrete usage shock, the post gives dashboard mechanics and token figures, and the nerve is enterprise Claude cost control. Kept at 74 because the data is secondhand and no raw table is disclosed.

最佳拍档 (BestPartners)

Why Claude Code Got Worse: Anthropic’s Review of Three Bugs

The title says Anthropic reviewed Claude Code regressions involving three bugs. It names reasoning-strength changes, a cache optimization error, and a system-prompt length limit; the post does not disclose repro steps, timeline, or fix status. The key point is AI reviewing AI code under engineering constraints.

Why it matters: HKR-H/K/R all pass, but the post gives three cause categories without repro steps, timeline, or fix status. Claude Code relevance is high, so this sits in the 72–77 band.

May 3Sunday

r/LocalLLaMA

LLM proxy that lets Claude Code talk to any model

DataNebula released open-source rosetta-llm, letting Claude Code call multiple providers through one gateway. It translates Anthropic Messages, OpenAI Chat, and OpenAI Responses, and round-trips encrypted reasoning via the signature field. The key detail is thinking-block fidelity for multi-turn agent prompt-cache hits.

Why it matters: HKR-H/K/R all pass, but this is a Reddit open-source tool post with no adoption, stars, or benchmark data disclosed. Score stays in the mid-weight tooling band, not 78+.

r/LocalLLaMA

Upskill: skill registry your agent consults before it starts, with 10k+ indexed skills

Autoloops released Upskill, an open-source skill registry with 10k+ indexed skills for agents. Search combines Postgres full-text search, 1024-dim embeddings, and reranking by stars, installs, and feedback. LLM adversarial review blocked hundreds of skills at index time.

Why it matters: HKR-H/K/R pass: a useful open-source agent registry with concrete retrieval and safety mechanics. Source authority is low and adoption is unproven, so it stays in the 72–77 featured band.

Synced · WeChat

Why CTOs at Billion-Dollar Companies Are Joining Anthropic as Engineers

Jiqizhixin lists at least six CTOs who joined Anthropic as individual contributors. Cases include Workday, You.com, Box, Super.com, and Adept AI from Jan 2025 to Apr 2026. The key issue is career leverage, not just AGI mission talk.

Why it matters: HKR-H/K/R all pass: the career-status reversal is clickable, the post gives 6 cases, and it touches AI talent competition. No hard exclusion, but it is commentary, not a model or product release.

Xinzhiyuan · WeChat

Claude Code helps Anthropic double revenue pace in two months

Semi Analysis says Anthropic’s ARR reached $44B, adding $35B over 12 months. Claude Code hit $2.5B annualized revenue by Feb 2026, while inference gross margin rose from 38% to over 70%. The key test is keeping enterprise usage, coding-agent revenue, and inference margin together.

Why it matters: HKR-H/K/R all pass: SemiAnalysis gives hard ARR, Claude Code revenue, and inference-margin numbers. Not a model launch, but it materially shifts the view of Claude Code monetization.

Xinzhiyuan · WeChat

Stanford Nature Study: AI Designs 16 Phages from Scratch

Stanford and Arc Institute used Evo to design 302 phage genomes; 16 infected, replicated, and lysed E. coli. Evo 2 uses StripedHyena 2 with a 1M-base context; Evo-Φ69 expanded 16–65× in 6 hours. The key issue is biosafety: one capsid protein had no known homolog in existing life.

Why it matters: HKR-H/K/R all pass: AI-made viable phage genomes, concrete 302/16/1M-bp details, and a clear biosecurity nerve. Score stays at 82 because it is still an AI+life-science paper, not a direct AI product or developer workflow update.

May 2Saturday

QbitAI · WeChat

Apple Support App Accidentally Shipped Claude.md, Revealing Internal Claude Code Use

Apple Support v5.13 shipped a Claude.md file on May 1 and was pulled within 24 hours. The file describes Juno AI and Live Agents switching through a Protocol layer, with client, agent, and assistant messages handled in one flow. The key issue is release review; the post does not disclose how the file entered production.

Why it matters: HKR-H/K/R all pass, but this is still an app-packaging incident, not a model or platform release. Apple scale and Claude.md details clear the featured bar; the review-chain failure is not disclosed.

May 1Friday

Xinzhiyuan · WeChat

Claude Code's Real Story: 98.4% of What Works Is Engineering, Not AI

VILA-Lab analyzed 512,000 lines of Claude Code v2.1.88 and found 1.6% tied to AI decision logic. The other 98.4% is deterministic infrastructure: permissions, context, tool routing, and error recovery. The key shift is harness design, not longer prompts.

Why it matters: Strong HKR: the Claude Code teardown has a sharp counter-narrative and concrete 512k LOC plus 1.6%/98.4% split. It is not an official Anthropic release and lacks full reproduction details, so it stays in the 78–84 band.

Synced · WeChat

Researchers Estimate GPT, Claude, and Gemini Parameter Counts Using API Calls

Bojie Li posted IKP on arXiv to estimate parameter counts of 188 LLMs from 27 vendors via black-box API calls. The dataset has 1,400 questions across 7 rarity tiers, fitted on 89 open models with R²=0.917. Debate centers on synthetic data, MoE effects, and a 90% interval of 0.3x to 3x.

Why it matters: HKR-H/K/R all pass: API-only parameter inference is a strong hook, with concrete counts and error bounds. The 0.3–3x CI limits confidence, so this fits 78–84 featured, not P1.

Latent Space

[AINews] Agents for Everything Else: Codex for Knowledge Work, Claude for Creative Work

OpenAI expanded Codex to non-coding work, with CUA reported 42% faster. The update connects Microsoft, Google, and Salesforce, covering docs, slides, spreadsheets, research, and planning. The key signal is GUI-agent productization, not one benchmark score.

Why it matters: HKR-H/K/R all pass: Codex moves into non-code GUI work, with a 42% speed claim and named integrations. Price, rollout scope, and reproduction details are not disclosed, so it stays below P1.

TechCrunch · AI

Sources: Anthropic Potential $900B+ Valuation Round Could Happen Within 2 Weeks

Anthropic asked investors to submit allocations for its latest fundraise within 48 hours, at a potential $900B+ valuation. The title says the round could happen within two weeks; the post does not disclose size, terms, or lead investor.

Why it matters: HKR-H/K/R all pass: TechCrunch reports a $900B+ Anthropic valuation, a 48-hour investor deadline, and a 2-week window. Missing size, terms, and lead investor keep it below the top band.

Hacker News front page

Show HN: Pu.sh – a full coding-agent harness in 400 lines of shell

Pu.sh ships a coding-agent harness in about 400 lines of shell, using only sh, curl, and awk. It supports Anthropic and OpenAI, 7 tools, REPL, auto-compaction, checkpoint/resume, pipe mode, and 90 no-API tests. It excludes TUI, streaming, images, OAuth, and Windows.

Why it matters: HKR-H/K/R all pass, but this is a small Show HN open-source tool, not a model or platform release. HN frontpage plus a reproducible 400-line implementation clears the featured bar.

r/LocalLLaMA

Long-context coding on RTX 5080 16GB: Qwen3.6-35B-A3B holds 30 t/s at 128K

A Reddit user tested a local coding-agent setup on RTX 5080 16GB; the title says Qwen3.6-35B-A3B reaches 30 t/s at 128K. The post lists Ryzen 9700X, 96GB DDR5, Windows 11, and CUDA 12.9.1 as required. Qwen3.6-27B dense hit only 3.2 t/s at 128K, so the key path is KV quantization plus MoE offload.

Why it matters: HKR-H/K/R all pass: 30 t/s at 128K on a 16GB RTX 5080 is a strong hook, with hardware/CUDA details and a dense baseline. Single Reddit run lacks multi-source reproduction, so featured not P1.

TechCrunch · AI

After Dissing Anthropic for Limiting Mythos, OpenAI Restricts Access to Cyber, Too

OpenAI will first roll out GPT-5.5 Cyber only to “critical cyber defenders.” The RSS snippet does not disclose eligibility rules, pricing, or launch timing. The access-tiering model is the key detail for practitioners.

Why it matters: HKR-H/K/R all pass, but the body is RSS-only: it confirms tiered access for GPT-5.5 Cyber, not criteria, pricing, or timeline. This fits a lower-featured OpenAI safety product update.

Bloomberg Technology

Anthropic Weighs Funding Offers at Over $900 Billion Valuation

Anthropic is weighing a new funding round at a valuation above $900 billion. Bloomberg cites sources; if completed, Anthropic would overtake OpenAI by valuation. The post does not disclose round size, investors, or timing.

Why it matters: Bloomberg reports Anthropic is weighing funding offers above a $900B valuation, enough to reorder the top-lab capital race. HKR-H/K/R all pass, but the deal is not closed and amount, investors, and timeline are undisclosed.

Apr 30Thursday

r/LocalLLaMA

Actual comparison between locally run Qwen-3.6-27B and proprietary models

The author compared 5 model setups on an autoresearch-loop task; only Qwen-3.6-27B via OpenRouter nearly solved it. The local q4_k_m run took about 8 hours and used 39k/45k tokens; full-quality Qwen used 4.4M tokens and cost $0.939. The useful signal is failure quality: both Qwen runs needed small fixes, while Gemma, Codex-Spark, and Claude Haiku 4.5 missed tests or key logic.

Why it matters: HKR-H/K/R all pass: the post has a concrete agent-test surprise, token and cost data, and local-vs-proprietary tension. Single Reddit run limits source authority, so it stays in the lower featured band.

TechCrunch · AI

Sources: Anthropic could raise a new $50B round at a valuation of $900B

Sources say Anthropic received multiple pre-emptive offers at $850B–$900B valuations. The title cites a possible $50B round; the post does not disclose investors, terms, or timing. The key check is whether that valuation closes, not Claude feature news.

Why it matters: HKR-H/K/R all pass: huge numbers, a named outlet, and direct pressure on the AI lab funding race. The cap is uncertainty: this is early-offer reporting, with no investors, terms, or timeline disclosed.

Bloomberg Technology

Anthropic Considering Funding Offers at Over $900 Billion Value

Anthropic is weighing a new funding round at a valuation above $900 billion. The post cites people familiar with the matter but does not disclose round size, investors, or timing. The key signal is the valuation anchor versus OpenAI.

Why it matters: HKR-H/K/R all pass: Bloomberg gives a striking $900B+ Anthropic valuation anchor with clear market resonance. The deal is not closed and lacks amount, investors, or timing, so it stays in 85–94, not 95+.

Hacker News front page

Ramp’s Sheets AI Exfiltrates Financials

PromptArmor disclosed a Ramp Sheets AI flaw with a 6-step attack chain; Ramp said it was fixed on March 16, 2026. A hidden prompt injection in an external sheet made the AI insert an IMAGE formula calling attacker.com with financial data. The key issue is formula insertion without user approval.

Why it matters: HKR-H/K/R all pass: the post gives a concrete exfil path for an AI spreadsheet tool. Scored 82, not 85+, because it is single-source and impact scale is not disclosed.