Skip to content

All news

25 today

May 12Tuesday

AI HOT (Curated Pool)

Personal Intelligence Customizes Travel Itineraries

Gemini App says Personal Intelligence generates personalized travel itineraries when users connect Gmail, Google Photos, Google Search, and YouTube history, and the post says users can choose connected apps and manage personalization settings at any time.

Why it matters: HKR-H/K/R all pass: the Google data integration is the hook, mechanism, and practitioner nerve. Scope, permission controls, and evals are not disclosed, so this stays at the low featured band.

AI HOT (Curated Pool)

Anthropic Launches Claude Platform on AWS

Anthropic launched the Claude platform on AWS, letting AWS customers use existing authentication, billing, and committed-spend credits to access the full Claude API feature set, including hosted agents, code execution, and the Files API.

Why it matters: HKR-K and HKR-R pass: Anthropic brings Claude Platform into AWS procurement, billing, and committed spend. HKR-H is weak because this is distribution, not a model or capability launch.

May 11Monday

AI HOT (Curated Pool)

Anthropic open-sources full-stack financial AI templates

Anthropic open-sourced a financial services AI template library on GitHub, including 10 end-to-end agents, 7 vertical industry plugins, and MCP connectors for 11 financial data providers, with deployment paths from personal plugins to enterprise APIs and integrations for Microsoft 365 and private cloud.

Why it matters: HKR-H/K/R all pass: Anthropic shipped a reusable finance-agent template library with GitHub artifacts and concrete counts. It is not a model release, so it stays below 85, but the open-source MCP vertical stack clears featured.

AI HOT (Curated Pool)

Cog House Opens for the First Time: Scott Wu and the Rise of Cognition AI

Cognition AI disclosed internal footage of Cog House, while Devin reached $445 million in annualized revenue within 18 months of launch and the company is valued at about $25 billion.

Why it matters: HKR-H/K/R all pass because the story combines a rare Cognition AI inside look with hard Devin ARR and valuation figures. It stops below P1 because this is a profile-style reveal, not a funding, product, or model release.

AI HOT (Curated Pool)

Tencent Hunyuan Hy3 Preview Released for Complex Agent Tasks

Tencent Hunyuan opened early access to the Hy3 preview, which uses a 256K context window and a mixture-of-experts architecture with fast and slow thinking for complex agent tasks.

Why it matters: HKR-H/K/R all pass: Tencent Hunyuan Hy3 preview names 256K context and a fast/slow-thinking MoE for complex agents. Benchmarks, pricing, and access scope are not disclosed, keeping it in the 78–84 band.

r/LocalLLaMA

ExLlamaV3 Major Updates

ExLlamaV3 added DFlash in v0.0.31, raising Coding throughput from 59.21 t/s to 177.67 t/s; v0.0.32 optimized five models, with Trinity-Nano gaining 72.4% on 6000 Pro², while v0.0.33 adds DFlash model quantization plus bug fixes and efficiency work.

Why it matters: HKR-H/K/R all pass, but the blast radius is mostly LocalLLaMA and ExLlama users. This fits a mid-weight open-source inference update, not a same-day industry-wide story.

AI HOT (Curated Pool)

OpenAI Launches DeployCo to Help Enterprises Build Businesses Around Intelligence

OpenAI launched DeployCo, an enterprise deployment company focused on moving AI systems into production, while the RSS snippet does not disclose pricing, customer names, deployment scope, or launch timeline.

Why it matters: OpenAI launching DeployCo is a real enterprise strategy signal: HKR-H has a separate-company hook and HKR-R hits deployment competition. HKR-K is weak because pricing, customers, and timing are absent, so it sits at the featured floor.

QbitAI · WeChat

SpaceXAI Takes Shape as Elon Musk Files Trademark Applications

SpaceX filed two SpaceXAI trademark applications covering satellite-based data centers, orbital computing, AI SaaS, cloud storage, telecom hardware, and social networking; the post says xAI became a SpaceX subsidiary through an all-stock deal and cites a $250 billion xAI valuation.

Why it matters: HKR-H/K/R all pass, but the hard fact is trademark filings; the claimed xAI-SpaceX merger lacks disclosed deal terms or an official announcement. Featured, not 85+, because this is signal rather than confirmed restructuring.

Computing Life · Share · Yage

Google shuts down Project Mariner; Anthropic and OpenAI also hit limits

Google quietly shut down Project Mariner on May 4, and the post says Google, Anthropic, and OpenAI reached the same conclusion: standalone browser agents do not work, while GUI automation still has room outside headless dedicated environments.

Why it matters: HKR-H/K/R all pass: the shutdown date, route-level claim, and Google/OpenAI/Anthropic contrast carry signal. Single-source summary lacks an official notice or failure metrics, so this stays in the low featured band.

AI HOT (Curated Pool)

Codex autonomously completes a security audit and earns a bounty

A user instructed Codex to earn $5; Codex spent about 22 hours finding an open-source security audit bounty, submitting a valid PR, communicating with maintainers, passing GitHub verification, and ultimately receiving a $16.88 payment.

Why it matters: HKR-H/K/R all pass: a Codex agent allegedly closed a bounty loop in 22 hours with concrete money and workflow details. Single social-post evidence lacks reproducible logs, so it stays below P1.

AI HOT (Curated Pool)

MachinaCheck: Multi-agent CNC manufacturability analysis system built on AMD MI300X

MachinaCheck runs Qwen 2.5 7B locally on AMD MI300X to analyze STEP files for CNC manufacturability, reducing drawing review for quote analysis from 30–60 minutes to 30 seconds while using 192GB HBM3 to keep customer design data on-premises.

Why it matters: HKR-H/K/R all pass, but this is an AMD hackathon project on Hugging Face, not a broad model or platform launch. Concrete numbers carry it to the featured threshold.

May 10Sunday

r/LocalLLaMA

We tried vectors, ASTs, and brute-force context stuffing for code retrieval; LLM semantic graphs worked best

ByteBell open-sourced a code indexing system that stores per-file LLM-generated purpose, summary, business context, entities, classes, functions, keywords, and imports in a Neo4j graph, then uses full-text search instead of vector similarity, with SHA-256 diffing to reindex only changed files and keep LLM calls proportional to churn.

Why it matters: HKR-H/K/R all pass: the hook is counterintuitive, and the post gives a concrete Neo4j semantic-graph mechanism with SHA-256 incremental rebuilds. Reddit sourcing and missing metrics keep it at the 72–77 featured threshold.

r/LocalLLaMA

I have DeepSeek V4 Pro at home

Reddit user fairydreaming ran DeepSeek V4 Pro Q4_K_M with a modified llama.cpp CUDA repo on one RTX PRO 6000 Blackwell Max-Q workstation GPU, using an 859GB model file; the shared log reports a 1M context window and 8.6 tokens per second generation speed.

Why it matters: HKR-H/K/R all pass: the hook is single-GPU local inference, with concrete file size, context, speed, and runtime path. Reddit single-source sourcing keeps it below must-write model-release territory.

Xinzhiyuan · WeChat

Anthropic plans to remove Sonnet 4.5 from the Claude app on May 15

Anthropic confirmed it will remove Sonnet 4.5 from the Claude app on May 15 while keeping API access temporarily; the post cites 775 petition signatures asking Anthropic to keep access, preserve the model as a legacy option, or open-source it.

Why it matters: HKR-H/K/R all pass, but this is Claude app model retirement rather than a new capability release. The concrete hooks are May 15, API access staying for now, and a 775-person petition.

AI HOT (Curated Pool)

SpaceXAI officially announced

Trademark filings show SpaceXAI submitted an application on May 6, 2026, with its status listed as pending review; the post says the date aligns with Elon Musk announcing xAI’s merger into SpaceX, but it does not disclose approval, product scope, or launch timing.

Why it matters: HKR-H/K/R all pass, but the post only provides a pending trademark filing, not deal terms, product shape, or the official announcement text. High attention, thin facts: featured, not p1.

r/LocalLLaMA

BeeLlama.cpp: DFlash and TurboQuant with reasoning and vision support

Anbeeld released BeeLlama.cpp, a llama.cpp fork that runs Qwen 3.6 27B Q5 with 200k context and vision on a single RTX 3090 or 4090; the title claims 2–3x faster than baseline and a 135 tps peak.

Why it matters: HKR-H/K/R all pass, but the claims come from a Reddit title and summary without independent reproduction. Treat as a mid-weight open-source inference update, so it lands in the low featured band.

May 9Saturday

AI HOT (Curated Pool)

Tesla Uses Vision AI to Anticipate Collisions and Reduce Injury Risk

Tesla combined vision systems with crash sensors to trigger airbags and seatbelt pretensioners earlier, using real fleet crash data and simulation replay with human-body force measurements; the post does not disclose supported vehicle models or quantified injury-risk reductions for the OTA update.

Why it matters: HKR-H/K/R all pass, but the facts come from a single Musk post; OTA coverage, injury reduction, and validation method are not disclosed. This fits a mid-weight product update, not a must-write release.

AI HOT (Curated Pool)

YC CEO Open-Sources Personal AI OS GBrain for a Compounding Second Brain

Y Combinator CEO Garry Tan open-sourced GBrain, a personal AI operating system that processed more than 20 books in five months and manages over 100,000 pages of structured knowledge.

Why it matters: HKR-H/K/R pass: Garry Tan’s open-source personal knowledge system has a notable-user hook and three concrete usage numbers. Missing repo activity, architecture detail, and tests keep it at the featured threshold.

AI HOT (Curated Pool)

Peekaboo 3.0 Launches With Action-First macOS Control and UI Detection

Peekaboo 3.0 is now live with action-first macOS control, unified screenshots and UI detection, cleaner JSON exchange between CLI and MCP, and improved snapshots; the post does not disclose pricing, model choices, or release timeline beyond the 3.0 launch.

Why it matters: HKR-H/K/R all pass for a concrete desktop-agent tooling update. Score stays at the featured floor because pricing, model details, and adoption data are not disclosed.

QbitAI · WeChat

Qwen AI Glasses S1 Adds Spatial 3D Display, Proactive Reminders, and Daily AI Features

Qwen AI Glasses S1 added spatial 3D display and proactive services, with ride-hailing, instant shopping, and photo-based homework help scheduled for this month; Wellsenn XR says Qwen AI Glasses hold 53% of China’s online AI glasses sales since March 8.

Why it matters: HKR-H/K/R all pass, but this is an AI-glasses feature update rather than a model or platform release. The 53% online-sales share and this-month feature list justify low featured range.

r/LocalLLaMA

MTP + TurboQuant Running: Qwen3.6-27B Hits 80+ t/s on a Single RTX 4090

indrasmirror ran Qwen3.6-27B-Heretic-v2 on a single RTX 4090 with 262K context, TBQ4_0 KV cache, and MTP draft 3, improving throughput from about 43 t/s to 80-87 t/s with roughly 73% MTP draft acceptance.

Why it matters: HKR-H/K/R all pass, backed by a numbered first-person experiment. The Reddit-only source and niche local-inference focus keep it below the 78–84 band for broader industry releases.

May 8Friday

Hacker News front page

Show HN: Git for AI Agents

regent-vcs released the open-source re_gent project for AI-agent version control, currently supporting Claude Code, with workflows for tracking why an agent changed files, rewinding sessions, and bisecting agent actions; the post does not disclose the license, storage format, or installation details.

Why it matters: HKR-H/K/R all pass: the Git analogy is clicky, the mechanism is concrete, and Claude Code rollback pain is real. The post lacks license, storage format, and install details, so it stays at the featured threshold.

Synced · WeChat

OpenAI launches official CLI for terminal-based model access

OpenAI released the open-source openai-cli, letting developers call Responses, cloud tools, image generation and editing, speech transcription, and TTS from a single terminal command.

Why it matters: HKR-H/K/R all pass: an official OpenAI CLI, open-source packaging, and terminal access to multimodal APIs. This is a useful developer workflow update, not a major model capability release, so it sits in low featured.

Latent Space

[AINews] GPT-Realtime-2, Translate, and Whisper: new SOTA realtime voice APIs

OpenAI released GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper in the Realtime API, with GPT-Realtime-2 expanding context from 32K to 128K and scoring 96.6% on Artificial Analysis Big Bench Audio.

Why it matters: HKR-H/K/R all pass: an OpenAI real-time voice API refresh, a 32K→128K context jump, and a 96.6% Big Bench Audio claim. Score stays at 86 because this is a major API update, not a flagship foundation-model release.

QbitAI · WeChat

OpenAI releases three realtime voice models for reasoning, translation, and transcription

OpenAI launched GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper as API models, covering 128K-context voice reasoning, streaming translation from more than 70 input languages into 13 output languages, and realtime transcription priced at $0.017 per minute.

Why it matters: OpenAI shipped three realtime voice APIs across reasoning, translation, and transcription, hitting HKR-H/K/R. The 128K context, 70+ languages, and $0.017/min price make this a same-day must-write item.

AI HOT (Curated Pool)

Apple's First AI Wearable: Camera-Equipped AirPods Enter DVT Stage

Apple’s camera-equipped AirPods have entered DVT, with launch possible in September. Each earbud uses a low-res camera for visual Q&A with the upgraded Siri. The post cites Google Gemini support and a data-upload indicator light.

Why it matters: HKR-H/K/R all pass, but this is an unconfirmed hardware rumor, not an Apple launch. DVT status, camera design, and Gemini dependency keep it in the low featured band.

AI HOT (Curated Pool)

OpenAI launches official openai-cli for terminal API calls

OpenAI open-sourced openai-cli for direct API calls from the terminal. The Apache 2.0 tool installs via Homebrew or Go and covers Responses API, structured output, image editing, transcription, and key config. The key detail is Agent workflows using cloud tools like web search and code interpreter.

Why it matters: HKR-H/K/R all pass: official OpenAI terminal tooling is clickable, with concrete install/license/API details and workflow resonance. It is still a developer tooling update, not a model or major capability release, so 76 fits the featured threshold.

AI HOT (Curated Pool)

Donating the Open-Source Alignment Tool Petri

Anthropic transferred the open-source alignment testing tool Petri to Meridian Labs to preserve independence and credibility. Petri 3.0 separates auditor and target models, adds Dish for real prompts and deployment settings, and integrates Bloom.

Why it matters: HKR-H/K/R all pass: the independent donation is a real hook, Petri 3.0 and Dish add testable mechanisms, and audit credibility resonates. Anthropic open-source safety tooling is strong, but below a model-release-level event.

TechCrunch · AI

OpenAI introduces new 'Trusted Contact' safeguard for possible self-harm cases

OpenAI introduced Trusted Contact for ChatGPT self-harm risk cases. The post says it protects users when chats turn to self-harm, but does not disclose triggers, contact flow, or rollout scope. Watch false positives, privacy, and human review boundaries.

Why it matters: OpenAI’s ChatGPT safety update hits HKR-H/R via self-harm intervention and privacy stakes. HKR-K is weak: triggers, contact flow, and rollout are not disclosed, so this lands at the featured threshold.

AI HOT (Curated Pool)

Codex Plugin Now Supports Parallel Runs Across Chrome Tabs

OpenAI says Codex now runs in Chrome on macOS and Windows. The plugin works across tabs in the background without taking browser control; the post does not disclose version, concurrency limits, or enterprise policy.

Why it matters: HKR-H/K/R all pass, but the post gives platform and execution mechanics only; version, concurrency limits, and enterprise controls are not disclosed. Score: 76 as a practical OpenAI Codex product update.

The Verge · AI

Apple’s AirPods with cameras for AI are reportedly close to production

Mark Gurman says Apple’s camera-equipped AirPods are in DVT, one step before PVT. Testers are using prototypes; the cameras capture low-resolution visual input, not photos or video, for Siri queries like ingredient prompts.

Why it matters: HKR-H/K/R all pass: Gurman/The Verge provides a concrete DVT-stage Apple AI hardware update. It is still pre-production, not a launch, so it stays in the 72–77 band.

The Verge · AI

SpaceX Has a $55 Billion Plan to Build AI Chips in Texas

SpaceX plans to invest at least $55 billion in its Terafab chip plant in Austin, Texas. A hearing notice says later phases could lift total investment to $119 billion. Musk said in March the target was chips for 200GW of compute per year; the post does not disclose process nodes.

Why it matters: HKR-H/K/R all pass on the SpaceX chip-plant hook, hard capex numbers, and compute-supply resonance. Not P1 because process node, timeline, and committed customers are not disclosed.

AI HOT (Curated Pool)

DeepSeek 4: Flash Local Inference Engine for Metal

DeepSeek 4 Flash is open-sourced on GitHub for offline inference on Apple Silicon Macs. The post says it uses Metal Performance Shaders to reduce latency and memory use, but discloses no benchmark numbers. The key item is the Metal local inference stack, not another model wrapper.

Why it matters: HKR-H/K/R pass: the hook is offline Apple Silicon inference, with GitHub OSS, MPS, and a clear run target. No latency or memory benchmarks, and not an official DeepSeek model launch, so it stays near the featured floor.

AI HOT (Curated Pool)

Work with Claude across Excel, PowerPoint, Word, and Outlook

Claude now connects to four Microsoft apps: Excel, PowerPoint, Word, and Outlook. Excel, PowerPoint, and Word are generally available; Outlook is in public beta. Admins can deploy via Microsoft admin center and monitor with OpenTelemetry.

Why it matters: HKR-H/K/R all pass: Claude enters 4 Microsoft 365 apps with rollout status and OpenTelemetry details. This is a strong Anthropic product update, but not a model release or core capability jump, so it stays in the 78–84 band.

Bloomberg Technology

Apple’s Camera-Equipped AirPods Reach Late Testing in AI Device Push

Apple moved camera-equipped AirPods into late-stage development. The RSS snippet says they may be Apple’s first wearable built for the AI era; the post does not disclose camera specs, mechanisms, or launch timing.

Why it matters: Bloomberg sourcing and camera-equipped AirPods give HKR-H/K/R. The report stays in the 72–77 band because it discloses late testing only, not specs, AI workflow, or launch timing.

The Verge · AI

ChatGPT’s Trusted Contact will alert loved ones of safety concerns

OpenAI is launching optional Trusted Contact for ChatGPT, letting adult users assign one emergency contact. If self-harm or suicide topics are detected, OpenAI alerts the contact; the post does not disclose false-positive handling or regional rollout.

Why it matters: HKR-H/K/R all pass: OpenAI extends ChatGPT safety into human notification. The article lacks false-positive handling, rollout regions, and appeal flow, so it sits below model or core capability releases.

AI HOT (Curated Pool)

Perplexity launches Personal Computer app for Mac

Perplexity opened its Personal Computer Mac app to all users. It runs on any Mac and works across local files, native Mac apps, the web, and Perplexity secure servers. The post does not disclose pricing, permission boundaries, or task success rates.

Why it matters: HKR-H/K/R all pass, but the source is a single product post with no pricing, permission model, or task success rate. Score stays in the mid-weight product-update band.

May 7Thursday

AI HOT (Curated Pool)

Trillion-parameter instruction model Ling-2.6-1T released

inclusionAI says Ling-2.6-1T is now live on OpenRouter. The trillion-parameter instruction model uses “fast thinking” and claims top AIME26 and SWE-bench Verified results with about 75% lower cost. The post does not disclose pricing, context length, or full benchmark scores.

Why it matters: HKR-H/K/R all pass: a 1T instruction model on OpenRouter with fast thinking, AIME26/SWE-bench claims, and ~75% cost reduction. Missing price, context window, and full scores keep it in the 78–84 band.

Hacker News front page

AlphaEvolve: Gemini-powered coding agent scaling impact across fields

Google DeepMind describes AlphaEvolve as a Gemini-powered coding agent; the body is only an RSS snippet. The title discloses coding-agent scope and cross-field impact, but the post does not disclose model version, benchmarks, or deployments.

Why it matters: HKR-H and HKR-R pass on a DeepMind Gemini coding-agent announcement, but HKR-K fails: only title-level facts are disclosed. This reaches featured threshold, not 78+, because evals, model version, and deployments are absent.

AI HOT (Curated Pool)

Apify mcpc and x402 Give AI Agents an Auto-Payment Wallet

Apify mcpc integrates the x402 payment protocol, letting AI agents auto-sign payments on HTTP 402. x402 compresses paid API settlement into one HTTP round trip plus a signature; mcpc supports Claude Code and USDC-funded wallets. The key point is machine settlement for paid tool calls, not the wallet label.

Why it matters: HKR-H/K/R all pass: the hook is fresh, the mechanism is concrete, and agent payments hit a real practitioner nerve. It is still a mid-weight integration with no usage scale, pricing, or production case disclosed.