Skip to content

All news

25 today

May 22Friday

AI HOT (Curated Pool)

Project Genie and Google Maps Street View launch interactive worlds

Project Genie partnered with Google Maps Street View to turn real U.S. locations into interactive worlds; the post does not disclose supported cities, generation mechanics, pricing, or access scope.

Why it matters: Google DeepMind’s official post says Genie × Street View turns real US locations into interactive worlds, so HKR-H and HKR-R pass. HKR-K fails because cities, generation method, and access are not disclosed.

Hacker News front page

Launch HN: Superset (YC P26) – IDE for the agents era

Superset launched an open-source agentic IDE that runs coding agents such as Claude Code, Codex, and OpenCode in parallel through git worktrees, and the team added Remote Workspaces in beta for running agents on remote machines while managing work from the desktop app.

Why it matters: HKR-H/K/R all pass, but Superset is still a new YC launch and the post lacks usage, pricing, or performance data. The git-worktree agent workflow clears the featured bar, not the must-write band.

Mistral AI

Mistral launches Connectors in Studio with built-in and custom MCP

Mistral launched Connectors in Studio. All built-in connectors and custom MCP are now callable through the API/SDK by every model and agent. New features include direct tool calling, human-in-the-loop approval flows, and programmatic access to create, modify, list and delete connectors.

Why it matters: The original gives the API usage and code examples for Connectors, enough to judge how enterprise MCP integration gets built.

AI HOT (Curated Pool)

Karpathy’s CLAUDE.md Four Rules Raise AI Coding Accuracy to 94%

Karpathy published a 65-line CLAUDE.md with four rules that raised AI coding accuracy from 65% to 94%, and the file received over 220,000 GitHub stars.

Why it matters: HKR-H/K/R all pass: a notable name, a claimed accuracy jump, and a rules-based Claude Code workflow. It stays below 85 because the body only gives summary-level numbers; task set, evaluation method, and the four rules are not disclosed.

AI HOT (Curated Pool)

Alibaba Qianwen App, PC, and Web Add Qwen3.7-Max

Alibaba added Qwen3.7-Max to the Qianwen app, PC client, and web client, with free access after updating the app to version 6.9.7 or later, and the official test reports a 35-hour autonomous kernel optimization run with more than 1,000 tool calls.

Why it matters: HKR-H/K/R all pass: Alibaba ships Qwen3.7-Max across three Qianwen clients, with v6.9.7+ free access and a 35-hour, 1,000+ tool-call claim. Benchmarks, context window, and API pricing are not disclosed, so it stays below 90.

MIT Technology Review · AI

Google I/O showed how the path for AI-driven science is shifting

MIT Technology Review says Google used I/O to shift its scientific AI framing toward Gemini for Science, a package that groups AI Co-Scientist and AlphaEvolve, while researchers can now apply for access and older specialized systems like AlphaFold and WeatherNext remain active.

Why it matters: HKR-H and HKR-K pass: MIT Technology Review frames a real Google science-AI product shift with named components and access conditions. HKR-R is weak because the impact is mostly research-facing, not practitioner-wide.

Xinzhiyuan · WeChat

Microsoft, after investing $13B in OpenAI, saw its engineers run up Claude Code costs

Microsoft plans to end Claude Code subscriptions by the end of June for its Experiences and Devices teams and move nearly 100,000 engineers to GitHub Copilot CLI, with the article attributing the change to external token-based billing costs.

Why it matters: HKR-H/K/R all pass: the OpenAI-Claude contrast hooks, the story gives end-June migration, nearly 100k engineers and token-billing, and it hits enterprise coding-agent cost control. Not a model release or official major launch, so 78–84 fits.

Xinzhiyuan · WeChat

Enterprise Agent Operations Begin? Anthropic Updates Architecture, Chinese Tech Firms Have It Running

Alibaba Cloud JVS Crew splits Agent, Environment, and Session into three layers, with sandboxes, snapshot recovery, RBAC, and usage-based billing. Anthropic added self-hosted sandboxes to Claude Managed Agents on May 19, while the article cites 2-week deployments and 5x or 10x efficiency gains in several Chinese customer cases.

Why it matters: HKR-H/K/R all pass, but the facts are an enterprise agent-infra comparison: Anthropic self-hosted sandboxes and Alibaba Cloud JVS Crew architecture. This is featured-level, not a must-write model release.

AI HOT (Curated Pool)

OpenAI Codex /goal Feature Officially Launches with Usage Guide

OpenAI moved Codex /goal mode from experiment to stable release, letting users set milestones in the Codex app, IDE extension, or CLI and keep tasks running for hours or days with progress checks, direction changes, and pause controls.

Why it matters: HKR-H/K/R all pass: OpenAI Codex /goal is now stable, with milestones across app, IDE extension, and CLI. The article is thin on permissions, safety limits, and tier access, so it stays in the lower featured band.

Hacker News front page

Show HN: Spec-Driven Development Workflow for Claude Code

The sddw author released a Claude Code plugin that splits work into requirements, code analysis, and design specs, then clears context after each step to keep cost and context focused.

Why it matters: HKR-H/K/R all pass for a Claude Code workflow with a concrete spec-and-context mechanism. It stays in the 72–77 featured band because the post lacks benchmarks, adoption data, or an official Anthropic release.

Computing Life · Share · Yage

How to Run DeepSeek V4 Flash Locally on Mac: DS4 Engine Explained

DS4 provides a macOS local runtime path for DeepSeek V4 Flash; the post only discloses three mechanisms—multi-agent integration, KV cache disk persistence, and activation steering—and does not disclose performance numbers, hardware requirements, or pricing.

Why it matters: HKR-H/K/R all pass, but the body only names DS4 mechanisms and omits performance, model size, Mac support, and reproducible tests; this fits the featured threshold for a local-inference tutorial.

Computing Life · Share · Yage

The Technology Behind GLM-5.1 Reaching 400 Tokens/s

Zhipu GLM-5.1 high-speed API claims 400 tokens/s, and the post says TileRT reconstructs GPU inference at the execution-model level; the RSS snippet does not disclose benchmark conditions, hardware, pricing, or latency distribution.

Why it matters: HKR-H/K/R all pass: 400 tokens/s is a concrete hook, TileRT adds mechanism, and latency/cost resonates with builders. It stays at 78 because the speed is claimed, with no independent test or pricing condition disclosed.

AI HOT (Curated Pool)

v2.1.147 Release Update

Claude Code v2.1.147 adds a Workflow tool, disabled by default, for deterministic multi-agent orchestration, and renames /simplify to /code-review with code-correctness reporting and GitHub PR inline-comment generation.

Why it matters: HKR-H/K/R all pass: the official Claude Code release adds a default-off Workflow tool for deterministic multi-agent orchestration. No performance data, pricing, or scope limits are disclosed, so this stays in the mid product-update band.

Latent Space

Giving Agents Computers — Ivan Burazin, Daytona

Daytona provides composable computers for AI agents, with one sandbox starting in about 60 ms, 50,000 sandboxes in about 75 seconds, and its largest customer running roughly 850,000 sandboxes per day.

Why it matters: HKR-H/K/R all pass: the agent-computer framing is clickable, and the sandbox scale numbers are concrete. Still, this is a startup infrastructure story, not a major model or platform release.

AI HOT (Curated Pool)

ChatGPT now supports creating and editing presentations directly in PowerPoint

ChatGPT is testing PowerPoint support for creating and editing presentations directly, including building, updating, understanding, and refining editable slides; the post does not disclose pricing, rollout scope, or availability conditions.

Why it matters: HKR-H/K/R all pass: OpenAI shows ChatGPT creating and editing editable PowerPoint slides. Pricing, rollout scope, and enterprise controls are not disclosed, so this stays featured rather than P1.

AI HOT (Curated Pool)

Datasette Agent

Datasette released Datasette Agent as its first extensible AI assistant, offering conversational data queries, plugin-based chart generation, official plugins for charts, AI image creation, and sandboxed code execution, with support for Gemini 3.1 Flash-Lite cloud models and local open-source models through LM Studio.

Why it matters: HKR-H/K/R all pass: a concrete Datasette agent with chart plugins and LM Studio local execution. The audience is narrower than major lab releases, so it sits in the 72–77 featured band.

Bloomberg Technology

SpaceX Aims to Build 10-Gigawatt Solar Factory Near Austin

SpaceX plans to build a 10-gigawatt solar manufacturing facility near Austin to supply power for Elon Musk’s proposed artificial intelligence data centers in space.

Why it matters: HKR-H/K/R pass: the space-AI-data-center angle is novel, the 10GW Austin-area factory is concrete, and power is a live AI-infra concern. Kept in the 72–77 band because cost, timeline, and buildout details are not disclosed.

AI HOT (Curated Pool)

Codex Enables Secure Cross-Device Mac Control Around the Clock

OpenAI Devs says Codex can use apps on a Mac from a phone while the Mac remains locked and the screen is off; the post does not disclose permission boundaries, pricing, or a release timeline.

Why it matters: HKR-H/K/R all pass: OpenAI Devs disclosed a concrete Codex Mac-control condition. Missing permission boundaries, pricing, and launch timing keep it below the 85+ band.

AI HOT (Curated Pool)

Aleph 2.0 and Edit Studio

Runway released Aleph 2.0 and Edit Studio, combining generation, editing, and post-production into one platform; the post does not disclose pricing, technical parameters, or rollout scope.

Why it matters: Runway is a major AI video vendor, and Aleph 2.0 plus Edit Studio is a mid-weight product update. HKR-H/K/R pass, but missing price, specs, and rollout keep it at the featured threshold.

AI HOT (Curated Pool)

Claude now supports more security and compliance tools

Anthropic added 28 security and compliance integrations for Claude Enterprise and its platform, using the Claude Compliance API to provide conversation content and activity events to DLP, SIEM, and existing enterprise monitoring workflows.

Why it matters: Official Anthropic product update with 28 compliance integrations and Compliance API event routing, so HKR-K/R pass. It is enterprise governance rather than a model capability jump, keeping it near the featured threshold.

AI HOT (Curated Pool)

Kotlin ADK and Android ADK 0.1.0 Released for Building AI Agents

Google released Kotlin ADK and Android ADK 0.1.0 for developers, with Kotlin ADK targeting backend agent workflows and Android ADK providing mobile-specific functions for building AI agents.

Why it matters: Google’s Kotlin ADK and Android ADK 0.1.0 release is a mid-weight agent tooling update. HKR-H/K/R pass, but the disclosed facts stop at platforms and version, with no performance data, examples, or ecosystem scale.

Hacker News front page

Launch HN: Runtime (YC P26) – Sandboxed coding agents for everyone on a team

Runtime launched open-source sandbox infrastructure for coding agents, supporting Claude Code, Codex, Cursor, Copilot, Gemini, and Devin, with hosted access, a free tier, and pricing based on a flat platform fee plus compute without token markup.

Why it matters: HKR-H/K/R pass: this is not a major-lab launch, but open-source sandboxes, six coding-agent types, and no token markup give teams concrete adoption signals. No usage data or marquee customers keeps it near the featured floor.

NVIDIA Blog

NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI

NVIDIA won four COMPUTEX 2026 Best Choice Awards for Vera Rubin NVL72, Jetson Thor, and Alpamayo; Vera Rubin NVL72 connects 36 Vera CPUs and 72 Rubin GPUs, and NVIDIA says it delivers up to 10x higher inference performance per watt and 10x lower cost per token.

Why it matters: HKR-H/K/R all pass: NVIDIA gives concrete Vera Rubin NVL72 specs and a 10x inference-efficiency claim, directly tied to AI compute costs. The source is NVIDIA’s event blog, so this stays below the 85 same-day must-write band.

May 21Thursday

The Verge · AI

Spotify is launching AI-generated remixes

Spotify and UMG announced a licensing deal that lets Premium subscribers pay for AI-generated remixes and covers of streaming songs; artists can opt out, while participating artists collect royalties from these AI remixes.

Why it matters: HKR-H/K/R all pass: Spotify and UMG add licensed AI covers/remixes with paid use, opt-out, and royalties. It is consumer audio, not a model or dev-tool release, so it sits just above the featured threshold.

Financial Times · Technology

Spotify targets high-spending superfans with AI-generated music

Spotify and Universal Music Group struck a licensing deal for a paid AI-generated music add-on inside Spotify’s app, targeting high-spending superfans; the RSS snippet does not disclose pricing, launch timing, supported markets, or model details.

Why it matters: HKR-H/K/R all pass: Spotify-UMG licensing turns AI music into a paid in-app product, not just a demo. Pricing, launch date, and revenue split are not disclosed, so this stays below must-write range.

MIT Technology Review · AI

Anthropic’s Code with Claude Showed Off Coding’s Future—Whether You Like It or Not

Anthropic used its two-day Code with Claude event in London to show Claude Code automation, with nearly half the room saying they shipped a pull request fully written by Claude in the past week, and many keeping their hands raised when asked whether they had shipped it without reading the code.

Why it matters: HKR-H/K/R all pass: the MIT Tech Review piece has a strong Claude Code hook, a concrete developer-behavior number, and clear resonance for programmers. It is not a model release or major product launch, so it stays in the 78–84 band.

r/LocalLLaMA

LLM planner: pick a rig by use case, model, or budget, or pick models for your rig

totosse17 published the LLMRequirements hardware planner with 60+ build configs, 50+ models, 130 cited tokens-per-second sources, 150+ reviewer videos, multi-region prices, idle and active watts, and a public GitHub data repo.

Why it matters: HKR-H/K/R all pass, but this is a Reddit community tool for local LLM rigs, not a broad platform release. The concrete dataset earns a featured-threshold score, not the 78+ band.

The Verge · AI

I Can’t Believe How Fast Google Vibe Coded My First Android App

The Verge’s Sean Hollister used Google AI Studio to generate three Android apps in one afternoon; one app came from a 148-word browser prompt and installed about 10 minutes later on an Android phone prepared with USB debugging and a PC connection.

Why it matters: HKR-H/K/R all pass: the story has a personal-test hook plus concrete timing and prompt details. This is not a major Google launch, so it fits the high-quality first-person experiment band, not same-day must-write.

Hacker News front page

Google officially announces ads in AI Mode search results

Google announced that AI Mode search results will include ads; the RSS snippet only lists 78 points and 66 comments, and the post does not disclose ad formats, targeting mechanics, or rollout timing.

Why it matters: HKR-H lands on the clean-AI-search twist; HKR-K has one concrete Google confirmation. HKR-R is strong for SEO and ad budgets, but missing format, auction logic, and launch timing keeps it below P1.

Xinzhiyuan · WeChat

Anthropic Acquires SDK Toolmaker Stainless, Leaving OpenAI and Others to Maintain SDKs

Anthropic has completed its acquisition of Stainless, an SDK generation company used by OpenAI, Anthropic, Meta, Cloudflare, and other infrastructure vendors; Stainless says prior SDK ownership remains with customers, but it will shut down hosted products including SDK generator and stop providing ongoing support.

Why it matters: HKR-H/K/R all pass: the deal targets API SDK generation, names OpenAI/Meta/Cloudflare as customers, and says hosted products will shut down. Anthropic bump applies, but this is not a model or core capability release, so it fits 78–84.

Synced · WeChat

Zhipu deploys ZCube, raising inference throughput 15% on the same GPUs

Zhipu deployed ZCube in a thousand-GPU GLM-5.1 production inference cluster, replacing ROFT while keeping GPUs, software stack, and business code unchanged; throughput rose by over 15%, TTFT P99 fell 40.6%, and switch plus optical module costs dropped by one third.

Why it matters: HKR-H/K/R all pass: Zhipu reports ZCube in a GLM-5.1 1k-GPU production inference cluster with +15% throughput and 40.6% lower TTFT P99. Single-source infra optimization keeps it below major model-release weight.

AI HOT (Curated Pool)

FSD officially launches in mainland China

The title says FSD has launched in mainland China, while the post only states an official entry into the mainland and does not disclose eligible cities, vehicle models, pricing, or regulatory conditions.

Why it matters: HKR-H and HKR-R pass: FSD’s China entry is a high-attention rollout with autonomy regulation and competition stakes. HKR-K fails because the post gives no rollout scope, pricing, or approval details.

AI HOT (Curated Pool)

Equipping AI with a Scientific Toolkit to Accelerate Research Workflows

Google DeepMind released Science Skills for Google Antigravity, integrating insights from more than 30 life-science sources, including the UniProt and AlphaFold databases.

Why it matters: Google DeepMind has a concrete product update: Science Skills adds 30+ life-science sources to Antigravity, hitting HKR-H/K. It is a vertical toolkit rather than a model release, so it sits at the low featured band.

AI HOT (Curated Pool)

Tencent Launches OS-Level AI Assistant Mavis on Windows, Mac, and Android

Tencent launched the OS-level AI assistant Mavis on May 21 across Windows, Mac, and Android, with document parsing, image recognition, system maintenance, partial offline use, model dispatching, and desktop control of mobile apps listed as supported functions.

Why it matters: HKR-H/K/R all pass: Tencent’s OS-level assistant spans Windows, Mac, and Android with concrete tool abilities. Model, pricing, and permission design are not disclosed, so it stays at the lower featured band.

Latent Space

Railway: The Agent-Native Cloud — Jake Cooper

Railway serves 3 million users with a 35-person team, adds about 100,000 signups per week, has raised $124 million, and has moved most workloads to its own bare-metal data centers with a reported three-month payback versus rented cloud capacity.

Why it matters: HKR-H/K/R all pass: the Railway interview has concrete growth, funding, and bare-metal details tied to agent infrastructure. It is a strong practitioner story, not a core model or major AI product release, so 74 fits the featured threshold.

AI HOT (Curated Pool)

Google Stitch update: AI design assistant supports end-to-end building

Google updated its AI design partner Stitch with real-time streaming design builds, direct edits and feedback, codebase or Design.md imports, dynamic UI generation, shareable URL exports, and global availability.

Why it matters: HKR-H/K/R pass: Google Stitch adds streaming builds, codebase/Design.md import, and global access. It stays at the featured threshold because model details, pricing, and measured output quality are not disclosed.

Bloomberg Technology

Nvidia Beats on Earnings, Revenue Projected at $91 Billion

Nvidia reported fiscal first-quarter earnings of $1.87 per share, above the $1.77 estimate; the company projected revenue of $91 billion for the quarter ending in July, above Wall Street expectations of about $87.4 billion.

Why it matters: NVIDIA earnings are an AI infrastructure temperature check: the $91B guide gives HKR-H/K/R real signal. It is not a model or capability release, so it stays in the good-quality featured band.

AI HOT (Curated Pool)

Nvidia fiscal Q1 2027 net income reached $58.321 billion, up 211% YoY

Nvidia reported fiscal Q1 2027 revenue of $81.615 billion and net income of $58.321 billion, while data center revenue reached $75.2 billion and the company guided fiscal Q2 revenue to $91 billion.

Why it matters: HKR-H/K/R all pass: NVIDIA’s earnings carry hard numbers tied to AI infrastructure economics. It stays below 85 because this is a financial result, not a model or product capability release.

AI HOT (Curated Pool)

GPT-5 Is Coming Soon

ChatGPTapp says GPT-5 is coming soon; the post repeats only that line and does not disclose the release date, parameters, pricing, context window, or API conditions.

Why it matters: HKR-H and HKR-R pass: an official-sounding “GPT-5 is coming” post is highly clickable and hits practitioner expectations. HKR-K fails because the post gives no date, capability, pricing, or API detail, so it sits at the low featured threshold.

AI HOT (Curated Pool)

Context compression improves search efficiency and accuracy

Perplexity has deployed query-aware compression in production, reducing context tokens by up to 70% while improving search answer quality.

Why it matters: HKR-H/K/R all pass: a counterintuitive production search update with a 70% token-cut claim and a direct cost-latency-quality hook. Single-source X post lacks benchmarks and reproducible setup, so it stays in the lower good-quality band.