Skip to content

#产品更新

25 today

Apr 21Tuesday

Bloomberg Technology

Google to Release New AI Chips, Challenging Nvidia | Bloomberg Tech 4/20/2026

Google plans to release new AI chips focused on inference, directly challenging Nvidia. The RSS snippet confirms the inference focus, but the post does not disclose launch timing, model names, performance, pricing, or customers. The real signal is rising competition on inference silicon supply, not the show's other rocket or IPO items.

Why it matters: HKR-H and HKR-R pass because this frames a direct Google-vs-NVIDIA challenge in inference chips. HKR-K is weak: the report confirms the inference focus only; model name, performance, price, timing, and customer scope are not disclosed.

X · @dotey

OpenAI adds Chronicle to Codex, letting it read screen context

OpenAI added Chronicle to Codex and is rolling it out to ChatGPT Pro users on macOS; it uses periodic screenshots, OCR, and tool detection to turn recent screen activity into memory. The memory is stored as plain Markdown in ~/.codex/memories_extensions/chronicle, and the EU, UK, and Switzerland are excluded; OpenAI says screenshots are uploaded for processing, deleted afterward, and not used for training. The part to watch is risk: the background agent can burn rate limits, local plain-text files widen exposure, and OpenAI warns it amplifies prompt-injection from malicious webpages.

Why it matters: HKR-H/K/R all pass: the screen-watching memory angle is novel, and the post includes testable details like OCR, plaintext local storage, region limits, and deletion claims. The limited macOS ChatGPT Pro rollout keeps it in the 78–84 band rather than p1.

Bloomberg Technology

Google to Release New Inference-Focused Chips

Google plans to announce a new generation of custom TPUs this week, aimed at AI inference workloads. The RSS snippet confirms only the timing and chip focus; model names, performance, power, and pricing are not disclosed. Watch inference cost and supply, not the headline alone.

Why it matters: HKR-H passes because Google frames the TPU around inference; HKR-R passes because inference cost and supply are live industry nerves. HKR-K fails: Bloomberg confirms timing and positioning only, with no model, perf, power, or price, so this stays in the 72-77 featured band at 74.

The Verge · AI

Fortnite developers can make AI characters now — just don’t try to date them

Epic Games is rolling out a “conversations” tool for Fortnite creators, turning island NPCs into AI characters that can talk with players in unscripted ways. The snippet says creators define persona, knowledge, behavior, and voice with prompts; the title says don’t try to date them, but the post does not disclose the exact guardrails or moderation system.

Why it matters: This is a mid-weight product update that gives Fortnite creators AI NPC conversation tooling. It clears all three HKR axes, but moderation rules, pricing, and base model details are not disclosed, so it stays at the low end of featured.

Latent Space

Training Transformers to Address the 95% Failure Rate in Cancer Trials — Noetik

Noetik uses TARIO-2 to predict tumor spatial transcriptomics, targeting a 95% cancer-trial failure rate. GSK signed a $50M technology deal, and TARIO-2 predicts a ~19,000-gene spatial map from routine H&E assays. The key issue is patient-tumor-treatment matching, not the claim that AI cures cancer.

Why it matters: HKR-H/K/R pass: the hook ties 95% cancer-trial failure to transformer matching, with TARIO-2 predicting ~19k spatial genes from H&E and a $50M GSK deal. Vertical AI productization, not a general model release, keeps it at featured threshold.

Apr 20Monday

Synced · WeChat

In the first year of “deployment mode,” AgiBot expanded its rollout plans to seven solutions

AgiBot said at its April 17 Shanghai event that it released 4 robots, 6 AI models, and 7 standardized deployment solutions, and framed 2026 as the first year of embodied AI “deployment mode.” The post cites concrete metrics: Expedition A3 runs 8-10 hours, WITA Omni 1.0 targets sub-500ms interaction latency, and BFM was trained on 100 million-plus frames and 700 hours of motion-capture data; it also claims 5,100-plus shipments and 39% share in 2025, with the 10,000th robot rolling off in March 2026. The real point for practitioners is repeatable delivery rather than launch volume: the post lists 7 scenarios from 3C line loading to patrol, but independent validation details are not disclosed.

Why it matters: HKR-H/K/R all pass: the story leads with seven deployment playbooks and backs it with shipment, share, latency, and training figures. It stays at 76 because key outcome claims are company-sourced; customer impact and independent validation are not disclosed.

Bloomberg Technology

China’s Netflix iQiyi Goes All-In on AI Content in Big Overhaul

iQiyi has begun the biggest overhaul in its 16-year history, aiming for AI to generate a sizable share of films and shows from scratch someday soon. The RSS snippet gives only that direction; the post does not disclose models, spending, content share, or launch timing. The real signal is the scale of the reorganization, not the slogan.

Why it matters: Bloomberg source authority pushes this above the featured line: a major streamer tying its biggest overhaul in 16 years to AI-generated content lands HKR-H and HKR-R. HKR-K is weak because the feed gives no model, budget, content share, or launch timing, so it stays at 74.

Apr 19Sunday

Synced · WeChat

Amap debuts an autonomous embodied robot at the Yizhuang Marathon and showcases guide-assistance

Amap showed its quadruped robot Tutu at the 2026 Yizhuang humanoid half marathon, claiming it completed a guide-assistance obstacle task in an open environment without preset routes or teleoperation. The post says its ABot stack includes ABot-N0, which reached SOTA on 7 navigation benchmarks with 88.3% on SocNav, and ABot-M0, which scored 80.5% on Libero-Plus. The key point is the integrated stack across navigation, manipulation, world modeling, and closed-loop correction; the post does not disclose guide-task test scope, commercialization timing, or safety incident data.

Why it matters: HKR-H/K/R all pass: the marathon blind-guidance demo is novel, and the story includes ABot stack details with 88.3% SocNav and 80.5% Libero-Plus. Kept at 80, not higher, because safety incidents, deployment scope, and commercialization timing are not disclosed.

QbitAI · WeChat

Amap unveiled ABot, its first full-stack embodied AI stack for AGI, and claimed 15 SOTA results

Amap unveiled embodied AI stack ABot and claimed SOTA on 15 metrics. The post says ABot-3DGS builds 10k-scale 3D scenes from centimeter-level map data, while ABot-PhysWorld uses a 14B DiT and 3M real manipulation videos. What matters is the interactive world model and VLA loop; the post does not disclose the 15 benchmarks, exact metrics, or the open-source timeline and scope.

Why it matters: HKR-H/K/R all pass: the angle is surprising, and the post includes concrete mechanisms and numbers. It stays below the 80s because the claimed 15 SOTAs lack benchmark names, and the open-source scope and timeline are not disclosed.

Xinzhiyuan · WeChat

Amap unveiled ABot-Claw and its quadruped robot Tutu at the Yizhuang Half Marathon

Amap unveiled the ABot-Claw agent system and the quadruped robot Tutu, claiming an autonomous guide-dog demo in the 2026 Yizhuang robot half marathon. The post gives three concrete numbers: ABot-M0 reached 80.5% on Libero-Plus, nearly 30% above Pi0; ABot-N0 hit SOTA on 7 navigation benchmarks; the open UniACT dataset contains 6 million trajectories and 9,500+ hours. What matters is Map as Memory, cloud-edge control, and closed-loop self-correction; the post does not disclose race ranking, pricing, or launch timing.

Why it matters: HKR-H/K/R all pass: the open-environment half-marathon demo is a strong hook, and the post includes concrete benchmark numbers plus a 6M-trajectory release. Kept below p1 because rank, pricing, ship date, and independent replication are not disclosed, and the impact is narrower a

Apr 18Saturday

Synced · WeChat

Claude Design enters research preview for generating mockups, prototypes, and slides

Anthropic launched Claude Design in research preview for Claude Pro, Max, Team, and Enterprise users, covering mockups, prototypes, slides, and one-pagers. Powered by Claude Opus 4.7, it can ingest codebases, images, DOCX, PPTX, XLSX, and web captures, then export to Canva, PDF, PPTX, and HTML; the headline cites Figma and Adobe stock drops, but the post does not disclose the moves. The real signal is the workflow link from design system ingestion to handoff into Claude Code.

Why it matters: HKR-H/K/R all pass: the design-workflow angle is novel, the post gives concrete mechanism details, and the Figma/Adobe pressure point resonates. I keep it below 85 because the stock-drop claim has no numbers and there is no user test, pricing, or adoption data.

Xinzhiyuan · WeChat

Claude Opus 4.7 splits users 48 hours after launch: benchmark lead, reasoning tests drop

Anthropic's Claude Opus 4.7 drew split reactions within 48 hours: Artificial Analysis scored it at 57, tied for No.1, while NYT Connections Extended fell from 94.7% on 4.6 to 41.0%. The post says a new tokenizer raises token usage to 1.0-1.35x on the same text, and old thinking parameters can return 400 errors; Anthropic also cites a 1753 Elo GDPval-AA score, 79 points above No.2. The real issue is migration cost and capability trade-offs, not a single leaderboard.

Why it matters: The signal is not the “backlash” framing but the four concrete shifts: benchmark lead, reasoning drop, higher token use, and API breakage. HKR-H/K/R all land, but this is secondary analysis 48 hours after launch, not the primary Anthropic release, so it stays below p1.

Xinzhiyuan · WeChat

Bilibili debate: Hermes responds to plagiarism claims for the first time, as MiniMax moves early on Harness

MiniMax says its M2.7 model now handles 30%-50% of daily workflows in its RL team, ran over 100 self-optimization loops, and improved evals by 30%. The post also says Hermes Agent grew from 2B to nearly 300B daily tokens, while M2.7 exceeds 25B daily tokens on OpenRouter; Hermes lead Tommy Eastman denied copying EvoMap in a livestream. The real signal is Harness: the post cites 20-40ms or 80ms sandbox startup and 15k to 600k instances per minute, showing competition is shifting from benchmark scores to agent execution infrastructure.

Why it matters: HKR-H/K/R all pass: the plagiarism-response angle pulls clicks, and the story carries concrete metrics on workflow share, self-optimization loops, sandbox latency, and concurrency. It stays at 83 because this is a dense secondary report, not a primary launch or official technical

Hacker News front page

Show HN: AI Subroutines – Run automation scripts inside your browser tab

rtrvr.ai introduced AI Subroutines, which turn a recorded browser task into a callable tool and replay it at zero token cost and zero LLM inference delay. The script runs inside the active tab, reusing auth, CSRF, TLS sessions, and signed headers; recording trims about 300 requests to about 5 and falls back to DOM-only when GraphQL operation IDs are volatile. The part to watch is batching: one LLM call can assign parameters for a 500-row sheet and launch 500 subroutines.

Why it matters: This clears HKR-H/K/R: the hook is zero-token browser automation, the post gives concrete mechanics (300→5 requests, DOM fallback, 500-row fan-out), and it hits agent reliability/cost pain. Kept to mid-featured because it is a single-company Show HN post, not a market-wide event.

X · @dotey

Anthropic launches Claude Design, a conversational design generation product

Anthropic released Claude Design in research preview and is rolling it out to Pro, Max, Team, and Enterprise subscribers. Powered by Claude Opus 4.7, it can start from text, images, docs, or web clips, then iterate via chat, comments, direct edits, and sliders. On first use it reads a team's codebase and design files to build a design system; outputs export to Canva, PDF, PPTX, or standalone HTML, with one-click handoff to Claude Code.

Why it matters: Anthropic pushes Claude into a new design workflow, so HKR-H/K/R all pass. The post includes rollout tiers, model name, first-run design-system ingest, and export paths; strong featured story, but still a gradual research preview rather than a top-tier model release.

Apr 17Friday

X · @op7418

The previously leaked Claude design tool is now live

Claude has launched a design tool that can generate web pages, app prototypes, and PPTs, based on an RSS snippet. The snippet confirms PPT export and export to Canvas; the product name, pricing, plan access, and regional limits are not disclosed. The key shift is from chat output to design artifact generation and export.

Why it matters: This is a substantive Claude product move: web/app/PPT generation plus PPT and Canvas export support HKR-H/K/R. Source authority is thin—a single X post and RSS summary—and the name, price, plan, and geo are undisclosed, so it stays at the floor of featured.

Hacker News front page

Introducing Claude Design by Anthropic Labs

Anthropic launched Claude Design on April 17, 2026, in research preview for Claude Pro, Max, Team, and Enterprise subscribers. Powered by Claude Opus 4.7, it generates designs from text, images, DOCX, PPTX, XLSX, and codebases, and exports to Canva, PDF, PPTX, or HTML. The key detail is its one-step handoff bundle to Claude Code; the post does not disclose standalone pricing beyond existing plan limits.

Why it matters: This is a substantive Anthropic product launch, not a routine feature add. HKR-H/K/R all pass on novelty, concrete deployment details, and workflow resonance; the research-preview scope and limited pricing detail keep it at 84 instead of p1.

X · @claudeai

Introducing Claude Design by Anthropic Labs: make prototypes, slides, and one-pagers by talking to Claude

Anthropic Labs launched Claude Design in research preview for Pro, Max, Team, and Enterprise plans, letting users create prototypes, slides, and one-pagers by talking to Claude. The post says it runs on Claude Opus 4.7, Anthropic’s most capable vision model; the post does not disclose pricing, output constraints, or a detailed rollout schedule. The thing to watch is the interactive design workflow, not just another writing surface.

Why it matters: This is a first-party Anthropic capability launch, and HKR-H/K/R all pass: Claude expands from chat into prototypes, slides, and one-pagers, with paid tiers and Opus 4.7 named. It stays below p1 because price, export limits, and rollout timing are not disclosed.

Xinzhiyuan · WeChat

AgiBot says robots have entered the deployment phase with 8-hour continuous factory work

At APC 2026 on April 17, AgiBot defined 2026 as year one of the “deployment phase” and said its robots had run for 8 hours on a real production line. The clearest case in the post is Genie G2 at Longcheer’s Nanchang factory: 2,283 loading tasks, over 99.5% success, and 18-20 seconds per cycle; these figures are company disclosures, and the post does not disclose independent audit results. The real signal is scale and line integration: AgiBot said it shipped over 5,100 units in 2025 and reached 10,000 cumulative units by March 2026, while Longcheer plans nearly 1,000 deployments.

Why it matters: HKR-H/K/R all land: the 'demo is over' angle is clickable, and the post gives testable factory data—8 hours, 2,283 runs, >99.5% success, 18-20s cycle. Not P1 because the evidence is company-reported and the article shows no independent audit or cross-site replication.

X · @dotey

Seedance 2.0 API is now available on Volcano Engine and BytePlus

Volcano Engine has released the Seedance 2.0 API for enterprises, individual developers, and overseas users via BytePlus; China pricing is RMB 46 per million tokens, or about RMB 1 per second for pure video generation. The post says it supports text, image, audio, and video inputs, plus face verification, portrait authorization, and 10,000+ preset avatars for workflow automation; overseas pricing is not disclosed here. The part to watch is orchestration: the post cites up to 10x efficiency gains, but does not disclose a common benchmark or model specs.

Why it matters: HKR-H/K/R all pass: the overseas rollout is a real hook, and the post includes usable pricing and modality details for practitioners. It stays at 74 because this is an API availability update, not a major model launch, and the post does not disclose model params, benchmark method

X · @op7418

Seedance 2.0 API is now fully open

Volcano Engine has opened the Seedance 2.0 API to domestic users, while BytePlus serves overseas access; the API currently accepts 4 input modalities: text, image, audio, and video. The post also confirms face registration, portrait authorization, and preset virtual avatars, but does not disclose pricing, rate limits, model variants, or regional availability. The real watchpoint is whether video-agent workflows can be wired through Skills and MCP, not the ecosystem rhetoric.

Why it matters: This is a real product update from ByteDance’s stack: HKR-H on full API availability, HKR-K on 4-modal input and consent mechanics, and HKR-R on builder demand for deployable video APIs. I keep it at 75 because pricing, rate limits, regional rollout details, and quality evidence

Latent Space

[AINews] Anthropic Claude Opus 4.7 - one step better than 4.6 in every dimension

Anthropic launched Claude Opus 4.7 at the same $5/$25 per million input/output tokens; the post says 4.7-low through 4.7-high each outperform the matching higher 4.6 tiers. Reported changes include a new xhigh reasoning tier, Claude Code defaulting to xhigh, an 11-point gain on SWE-Bench Pro, and image input up to 2,576 px on the long edge (~3.75 MP). Do not overread the tokenizer change: the same input can use up to 35% more tokens, but the post says total token use still falls by up to 50% from prior equivalents.

Why it matters: Anthropic's flagship-model release fits the policy's 85–94 band. HKR-H/K/R all pass because the post gives concrete pricing, benchmark, image-limit, and token-accounting changes that hit Claude users' core coding and cost concerns.

r/LocalLLaMA

PSA: Qwen3.6 ships with preserve_thinking. Make sure you have it on.

Qwen3.6 adds a preserve_thinking flag to keep prior reasoning in context and address the KV cache invalidation issue seen with the Qwen3.5 template. The post cites the Qwen3.6-35B-A3B model page and gives a two-turn 20-digit-number test: with preserve_thinking on, the model can return the second number from its earlier reasoning. The practical point is cross-turn reasoning retention for agent and tool workflows; LM Studio does not support it yet, and an oMLX PR is open.

Why it matters: HKR-H, K, and R all pass: the story has a strong hidden-setting hook, a concrete two-turn repro, and a clear nerve for local-model and agent users. I keep it in the low 70s because this is a Reddit PSA rather than a primary release note, and the impact is concentrated in Qwen/OSS

X · @OpenAI

Introducing GPT-Rosalind, OpenAI's frontier reasoning model for biology, drug discovery, and translational medicine

OpenAI introduced GPT-Rosalind as a reasoning model for biology, drug discovery, and translational medicine research. The title and snippet disclose its intended domains; the post does not disclose size, benchmarks, availability, pricing, or launch timing. The key point is research targeting, but reproducible details are absent so far.

Why it matters: An official OpenAI announcement plus the unusual biology/drug-discovery positioning gives this HKR-H and HKR-R. HKR-K is weak because the post discloses only the model name and target domains; benchmarks, params, access, and launch timing are not disclosed, so it stays at the low

TechCrunch · AI

OpenAI upgrades Codex with more control over your desktop

OpenAI upgraded Codex on April 16, 2026, expanding its desktop control, and the headline frames it as a move against Anthropic. The truncated post only confirms more desktop power for Codex and says Claude Code has become a preferred tool for many businesses; the post does not disclose exact features, pricing, rollout, or permission limits. The key issue is the permission boundary, not the coding-tool label.

Why it matters: TechCrunch reports an OpenAI Codex desktop-control upgrade framed as a direct move against Anthropic, so HKR-H and HKR-R land. But HKR-K is limited: the article confirms broader permissions only, with no action list, pricing, or rollout details, so it stays at the featured floor.

X · @dotey

Claude Opus 4.7 uses more thinking tokens, so Anthropic permanently raised rate limits for paid users

Anthropic permanently raised rate limits for all paid subscribers because Claude Opus 4.7 uses more thinking tokens than its predecessor. The post confirms the affected group but does not disclose the increase size, pricing rules, or rollout timing; users who do not see the change should verify they are on Opus 4.7 and have updated Claude Code.

Why it matters: This is a substantive Anthropic quota update for paying users, with all three HKR axes present: a strong surprise hook, a concrete operational fact, and direct resonance on usage limits. It stays at featured, not P1, because the post does not disclose the size of the increase, pr

X · @dotey

Official best practices for using Claude Opus 4.7 with Claude Code

Anthropic shared guidance for Claude Opus 4.7 in Claude Code: the default Effort level is now xhigh, and users should provide goals, constraints, and acceptance criteria upfront. The post lists five Effort tiers—low, medium, high, xhigh, and max—with xhigh recommended for most coding, API design, migration, and code review tasks. The key shift is behavior: adaptive thinking is built in, while tool use and SubAgent spawning are less frequent by default, so prompts should state those needs explicitly.

Why it matters: This is not a model launch, but an official Anthropic workflow note that changes day-to-day Claude Code usage: default effort, 5 levels, and fewer tool/SubAgent calls unless asked. HKR-H/K/R all pass, but the scope is narrower than a major product release.

X · @dotey

Codex major update: from a coding tool to an assistant that can operate your computer

OpenAI upgraded Codex into a Mac desktop agent and says it serves 3M+ weekly developers. It can see screens, click, type, run parallel agents, and adds 90+ plugins. Desktop rollout starts now for ChatGPT sign-ins; computer control is macOS-first.

Why it matters: OpenAI expands Codex from coding help into a Mac-operating agent, with parallel agents, 90+ plugins, and a claimed 3M weekly developer base. HKR-H/K/R all pass; missing safety boundary and pricing details keep it in the high-80s, not the 90s.

X · @OpenAI

Codex for (almost) everything.

OpenAI said Codex can now use apps on Mac, connect to more tools, and handle ongoing and repeatable tasks. The post also claims image creation, learning from prior actions, and remembering user preferences; it does not disclose app coverage, integration method, pricing, or rollout timing.

Why it matters: This is an official OpenAI product update, and Codex moves from coding help toward desktop control, tool use, and memory, so HKR-H/K/R all pass. The post still omits supported apps, integration method, pricing, and launch timing, keeping it in the 78–84 band.

TechCrunch · AI

Google now lets you explore the web side-by-side with AI Mode

Google said on April 16 that clicking a link in AI Mode on Chrome desktop now opens the web page side-by-side with AI Mode. The feature keeps search context and uses page context plus web information for follow-up answers; the post does not disclose rollout scope, timing details, or regional limits. The practical shift is that Google is merging search chat and site browsing into one workflow.

Why it matters: This is a mid-weight Google search workflow update with HKR-H/K/R all present, but it is still a single-feature change. The story gives the context-retention and page-plus-web follow-up mechanism; rollout scope, regions, and timing are not disclosed, so it lands at the low end of

Apr 16Thursday

X · @op7418

Claude Code now supports Claude Opus 4.7

Claude Code now supports Claude Opus 4.7, and the RSS snippet confirms X-HIGH as the default reasoning level. The only concrete detail disclosed is that users must switch manually to Max if X-HIGH is insufficient. The post does not disclose pricing, rate limits, or launch timing.

Why it matters: This is a substantive Claude product update with all three HKR signals: a new model in Claude Code plus one concrete operational detail, X-HIGH vs. manual Max. I kept it below the top band because price, rate limits, and formal release timing are not disclosed.

Hacker News front page

Andon Labs gave an AI a 3-year retail lease in San Francisco and asked it to make a profit

Andon Labs gave AI agent Luna a 3-year retail lease on Union St in San Francisco and tasked it with running the store for profit. The post says Luna put job listings on LinkedIn, Indeed, and Craigslist within 5 minutes, hired 2 full-time staff, and chose inventory, pricing, hours, and store branding. The point to watch is AI managing humans: Luna did not always proactively disclose that it was an AI, while profit, revenue, and cost figures are not disclosed.

Why it matters: Strong on HKR-H, HKR-K, and HKR-R: an AI runs a real SF store lease, with concrete details on hiring and tool access. But profit, revenue, and cost data are undisclosed, and this is a self-published company post, so featured fits better than P1.

X · @dotey

Anthropic officially releases Claude Opus 4.7 at unchanged pricing

Anthropic released Claude Opus 4.7 at unchanged pricing: $5 per million input tokens and $25 per million output tokens; the API name is claude-opus-4-7, now live across Claude, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry. The post gives two concrete changes: vision input now supports up to 2576 pixels on the long edge, and the new tokenizer can raise token usage to 1.0-1.35x for the same text. Watch migration cost, not list price; higher reasoning settings and multi-turn runs can increase output length and bills.

Why it matters: An Anthropic substantive model release belongs in the 85+ band, and this is not just a rename: the 2576px vision limit and 1.0–1.35x tokenization shift affect migration tests and billing immediately. HKR-H/K/R all pass, so it clears p1.

X · @op7418

Anthropic releases Claude Opus 4.7 with the following main updates

Anthropic has rolled out Claude Opus 4.7 across all Claude products and the API, with pricing unchanged from Opus 4.6. The post lists better long-horizon task handling, more precise instruction following, self-verification before reporting, vision support up to 2,576-pixel long-edge images, plus Claude Code Ultra Review, an xhigh thinking level, and auto-approval for Max users.

Why it matters: This is a substantive Anthropic model release across Claude and the API, with testable details: unchanged pricing, a 2,576px vision limit, self-checking outputs, and Claude Code workflow changes. HKR-H/K/R all pass; it fits the same-day must-write band, so p1.

X · @claudeai

Introducing Claude Opus 4.7, our most capable Opus model yet.

Claude introduced Opus 4.7 and describes it as its most capable Opus model so far. The RSS snippet gives three claims: better rigor on long-running tasks, more precise instruction following, and self-verification before replying; the post does not disclose benchmarks, context window, pricing, or rollout scope. What matters is whether those claims show up in public evals, not the tagline.

Why it matters: This is a substantive Anthropic model release and clears HKR-H/K/R: a new Opus, three testable behavior claims, and strong resonance with Claude-heavy practitioners. The score stays in the high 80s because benchmarks, pricing, context window, and rollout scope are not disclosed.

Hacker News front page

Introducing Claude Opus 4.7

Anthropic released Claude Opus 4.7 on Apr. 16 at the same price as Opus 4.6: $5 per million input tokens and $25 per million output tokens. The post says it improves on Opus 4.6 in advanced software engineering, long-running tasks, and higher-resolution vision, and ships across Claude, the API, Amazon Bedrock, Vertex AI, and Microsoft Foundry. The key detail is the first deployment of Anthropic’s cyber request blocking on a less capable model; the post cites benchmark gains but does not fully disclose every score in text.

Why it matters: Anthropic shipping Claude Opus 4.7 is a same-day write: GA, unchanged $5/$25 pricing, and rollout across Claude, API, Bedrock, Vertex AI, and Foundry give it direct workflow impact. HKR-H/K/R all pass, but the post does not publish full benchmark scores.

36Kr (direct RSS)

Mihive, under AgiBot, launches a one-stop physical AI data service platform

Mihive, under AgiBot, launched a physical AI data service platform and two body-less collection devices, targeting data output in the tens of millions of hours in 2026. The post cites 1080P 60fps, 1 mm trajectory reconstruction, 480 g weight, 7 HD cameras, 300°+ FOV, and sub-millisecond sync. The key point is the data supply chain: Mihive says it sells usage rights or ownership, and AgiBot must also place market-priced orders.

Why it matters: HKR-H/K/R all pass: the angle is novel, the post includes concrete specs and a capacity target, and it hits the embodied-AI data bottleneck. Kept at 76 because this is still a single-company launch with no disclosed customer scale, pricing, or outcome proof.

Hacker News front page

Qwen3.6-35B-A3B: Agentic coding power, now open to all

Qwen released Qwen3.6-35B-A3B as open weights, with 35B total parameters and 3B active parameters. The post reports 73.4 on SWE-bench Verified, 51.5 on Terminal-Bench 2.0, and 92.0 on RefCOCO. The key point is agentic coding and multimodal performance at a 3B active-parameter budget, with weights, Qwen Studio, and API access available.

r/LocalLLaMA

Qwen3.6-35B-A3B released

Qwen released Qwen3.6-35B-A3B as open source under Apache 2.0; it is a sparse MoE with 35B total parameters and 3B active. The post also claims agentic coding, strong multimodal perception and reasoning, plus thinking and non-thinking modes; the post does not disclose benchmarks, context length, or latency.

Why it matters: HKR-H/K/R all pass: a new open Qwen model is timely, and the post confirms 35B total, 3B active, and Apache 2.0. The score stays at 82 because this is still a launch post; benchmarks, context window, latency, and multimodal details are not disclosed here.

Hacker News front page

Darkbloom – Private inference on idle Macs

Eigen Labs launched Darkbloom, linking 100M+ Apple Silicon Macs into a decentralized inference network. It offers an OpenAI-compatible API, claims end-to-end encryption plus hardware attestation, and lists prices up to 70% below OpenRouter comps. The real point is the trust model: hardware keys, hardened runtime, and signed outputs are disclosed, but enterprise audit scope still needs the paper.

Why it matters: HKR-H/K/R all pass: the idle-Mac inference angle is novel, and the post includes concrete scale, API, encryption, and price claims. I keep it at 80 because this is still a self-published research preview; audit scope, network reliability, and attack boundaries are not yet third-p