Skip to content

All news

25 today

May 15Friday

MIT Technology Review · AI

How Chinese Short Dramas Became AI Content Machines

Chinese short-drama companies are using AI for full-series production, with DataEye counting an average of 470 AI-generated short dramas released per day in January 2026, while FlexTV says production time fell from three to four months to under one month and North American per-series costs can drop by 80% to 90%.

Why it matters: HKR-H/K/R all pass: the story has a strong content-factory hook, concrete production metrics, and clear labor/cost resonance. It is a quality industry feature, not a model or platform release, so 80 fits the 78-84 band.

Alibaba Technology · WeChat

Qoder 1.0 launches as an agentic development workspace beyond AI IDE

Alibaba released Qoder 1.0 with downloads for Windows, macOS, and Linux, adding a standalone Quest workspace, cross-project parallel agent tasks, a team knowledge engine, and Experts mode with five roles for planning, research, coding, review, and testing.

Why it matters: Alibaba’s Qoder 1.0 is a mid-weight AI coding product release with concrete agent-workflow features and developer resonance. No pricing, benchmark, or task-success data is disclosed, so it stays near the featured threshold.

AI HOT (Curated Pool)

Codex lands on mobile with preview in the ChatGPT app

OpenAI brought Codex to a preview inside the ChatGPT mobile app; the post does not disclose supported platforms, feature scope, pricing, or rollout schedule.

Why it matters: Official OpenAI product update with HKR-H/K/R, but detail is thin. The post does not disclose platform support, feature scope, pricing, or rollout timing, so it sits at the featured threshold.

AI HOT (Curated Pool)

Databricks brings GPT-5.5 to enterprise agent workflows

Databricks made GPT-5.5 available through AI Unity Gateway for AgentBricks and Agent Supervisor API workflows; on OfficeQA Pro, it became the first model above 50% accuracy and reduced errors by 46% versus GPT-5.4.

Why it matters: HKR-H/K/R all pass: GPT-5.5 enters Databricks workflows with 50% OfficeQA Pro accuracy and 46% fewer errors than GPT-5.4. It stays below a full model-release score because the page is a sales-led OpenAI customer story using Databricks’ own benchmark.

AI HOT (Curated Pool)

Connect Grok to the Hermes Agent

xAI connects Grok subscription accounts to Nous Research’s open-source Hermes Agent across all subscription tiers, letting users run Grok 4.3 text chat and reasoning, generate spoken replies with text-to-speech, create images and videos with Grok Imagine, and connect the agent to WhatsApp or Discord.

Why it matters: HKR-H/K/R all pass, but this is a mid-weight xAI product integration with an open-source agent, not a flagship model release. Featured fits; it does not clear the 85+ same-day bar.

AI HOT (Curated Pool)

ChatGPT launches personal finance experience

OpenAI launched a personal finance preview for ChatGPT Pro users in the US, letting users connect financial accounts and receive analysis based on their finances, goals, and priorities.

Why it matters: HKR-H/K/R all pass: OpenAI is moving ChatGPT into personal finance with linked financial accounts. The score stays at 76 because it is a US Pro preview, with partners, controls, and pricing not disclosed.

AI HOT (Curated Pool)

Claude Agent Tool v2.1.142 Release

Claude Agent Tool v2.1.142 adds eight command-line flags for configuring background sessions, upgrades Fast mode’s default model to Opus 4.7, and fixes more than 15 issues including MCP tool timeouts and Windows network-drive deadlocks.

Why it matters: HKR-H/K/R all pass: this is a small Claude Code release, but the Opus 4.7 Fast-mode default, 8 session flags, and 15+ fixes affect daily dev workflows. Anthropic tool-chain relevance keeps it at the featured floor.

Latent Space

AI-Native Healthcare: 100M Doctor Visits, 10–20 Hours Saved, Prior Auth in Minutes

Abridge says it is projected to support 80M+ patient-clinician conversations this year across 250 large U.S. health systems, 28+ languages, and 50+ specialties, while its clinical documentation workflow reduces clinicians’ documentation burden by 10–20 hours per week.

Why it matters: HKR-H/K/R all pass: the story has a strong scale hook, concrete adoption metrics, and workflow ROI. Claims are company-interview sourced, not an independent benchmark or major platform release, so it sits in low featured.

AI HOT (Curated Pool)

Codex adds automation hooks and programmatic tokens

Codex added hooks and programmatic access tokens: hooks run scripts at key task stages for validation, secret scanning, logging, or repo-specific behavior, while scoped tokens for Business and Enterprise teams support CI/CD, release workflows, and internal automation with expiration or revocation.

Why it matters: HKR-H/K/R all pass: Codex gains concrete automation hooks and programmatic tokens for CI/CD. Score stays in the 72–77 band because the post discloses workflow fit, not pricing, permission detail, or impact data.

Bloomberg Technology

Musk’s xAI Unveils First Coding Agent in Bid to Rival Anthropic

xAI is rolling out its first AI coding agent, Grok Build, for software development workflows; the RSS snippet names Anthropic’s Claude as the rival but does not disclose pricing, availability, benchmarks, or supported IDEs.

Why it matters: HKR-H and HKR-R pass: xAI entering coding agents is a strong competitive hook for developers. HKR-K fails because pricing, availability, and benchmarks are not disclosed, so this stays at the low end of a mid-weight product update.

The Verge · AI

OpenAI’s Codex is now in the ChatGPT mobile app

OpenAI will let users access Codex from the ChatGPT mobile app; the RSS snippet says Codex can write code and use apps on a computer, but the post does not disclose launch timing, pricing, or the full mobile feature scope.

Why it matters: OpenAI added Codex access to ChatGPT mobile, a mid-weight product update. HKR-H/K/R pass through the mobile coding-agent hook, concrete app-control claim, and developer workflow nerve; missing timing, pricing, and support scope keep it at the featured floor.

The Verge · AI

Microsoft starts canceling Claude Code licenses

Microsoft plans to remove most Claude Code licenses and push many developers toward Copilot CLI; the snippet says Microsoft opened access in December to thousands of internal developers, but the post does not disclose the exact license count, pricing, or migration schedule.

Why it matters: HKR-H comes from Microsoft dropping a rival coding tool; HKR-K adds the Dec rollout to thousands of internal devs; HKR-R hits Claude Code vs. Copilot competition. Strong featured, not a major release.

AI HOT (Curated Pool)

The Founder's Playbook: Building an AI-Native Startup

Anthropic published an AI-native startup playbook covering four stages—ideation, MVP, launch, and scaling—with goals, exit criteria, failure modes, and Claude-based exercises for validation, customer discovery, technical debt control, product-market fit checks, and workflow automation.

Why it matters: HKR-H/K/R all pass, but this is an Anthropic playbook rather than a model or product capability release. The concrete value is the 4-stage framework, exit criteria, and Claude-driven exercises, so it lands at the featured floor.

AI HOT (Curated Pool)

Using Claude Code Effectively in Large Codebases: Best Practices and Where to Start

Claude Code is used in million-line monorepos, legacy systems, and distributed architectures, and the post says its large-codebase workflow relies on five extension points: CLAUDE.md, hooks, skills, plugins, and MCP servers for agentic search on local codebases.

Why it matters: HKR-H/K/R pass: official Claude Code guidance, five concrete extension points, and a direct coding-agent workflow nerve. It is a high-quality tutorial, not a new model or major capability release, so it stays in the 72–77 band.

AI HOT (Curated Pool)

OpenEvidence reaches 65% of U.S. doctors, drawing attention to shadow AI use

OpenEvidence reaches 65% of U.S. doctors and recorded 27 million clinical uses in April; doctors registered on mobile with license numbers, while hospitals were initially unaware of the shadow AI adoption pattern.

Why it matters: HKR-H/K/R all pass: 65% doctor reach and 27M April uses are unusually strong adoption data, with the hospital-unaware angle adding shadow-AI tension. Kept below 85 because methodology, revenue, and liability controls are not disclosed.

AI HOT (Curated Pool)

Genkit launches middleware system to improve control in agentic AI apps

Google’s open-source Genkit framework added a middleware system that intercepts generation calls, models, and tools, with support for TypeScript, Go, Dart, and Python.

Why it matters: HKR-H/K/R all pass: the Google Genkit update adds concrete middleware hooks for agentic apps across generation, model, and tool layers. Scope stays within Genkit, so this sits at the featured threshold rather than a must-write release.

AI HOT (Curated Pool)

Accelerating On-Device AI: Arm and Google AI Edge Optimization Practices

Arm SME2 and Google AI Edge integrate with LiteRT, XNNPACK, and KleidiAI to optimize Stability AI’s stable-audio-open-small, delivering over 2x faster audio generation and 4x lower memory use on Arm-based mobile devices and laptops.

Why it matters: HKR-H/K/R pass via concrete 2x speed and 4x memory gains, plus an edge-deployment cost hook. Scope stays narrow to one audio model on Arm devices, so it lands at the featured threshold.

May 14Thursday

AI HOT (Curated Pool)

Kimi launches Web Bridge browser extension for multi-platform interaction

Kimi launched the Web Bridge browser extension, which lets agents search, scroll, click, type, and complete website tasks, with support for Kimi Code CLI, Claude Code, Cursor, Codex, and Hermes.

Why it matters: HKR-H/K/R all pass: the product hook, action list, and workflow relevance are clear. Kept in the 72–77 band because this is a mid-weight tool update, not a model release, and safety or performance details are not disclosed.

AI HOT (Curated Pool)

Use Codex from anywhere

OpenAI added Codex to the ChatGPT mobile app, letting users monitor, guide, and approve remote coding tasks across devices.

Why it matters: OpenAI added Codex controls to ChatGPT mobile for monitoring, guiding, and approving remote coding tasks. This clears HKR-H/K/R as a mid-weight product update, but pricing, permission details, and task limits are not disclosed, so it stays below a major release.

r/LocalLLaMA

Automated AI researcher running locally with llama.cpp

Hugging Face’s ml-intern added local-model support through llama.cpp and ollama; the post says Qwen3.6-35B-A3B can orchestrate CPU/GPU sandboxes and Hub jobs to run an end-to-end SFT workflow.

Why it matters: HKR-H/K/R all pass, but this is a Reddit-sourced open-source tool update, not a major model release. Local sandbox and Hub-job orchestration for SFT put it just above the featured threshold.

r/LocalLLaMA

Open-source one-prompt-to-cinematic-reel pipeline on one GPU with FLUX.2 and Wan2.2-I2V

The developer open-sourced StudioMI300, an 8-stage sequential pipeline that turns one English sentence into a 720p MP4 on a single AMD Instinct MI300X, cutting end-to-end time from 25.9 minutes to 10.4 minutes per clip.

Why it matters: HKR-H/K/R all pass: the post has a concrete one-GPU video pipeline, runtime numbers, and a local-build cost/control hook. Reddit single-source status and no third-party replication keep it below the 78+ band.

AI HOT (Curated Pool)

Tencent Open-Sources Agent Memory to Cut Token Usage by 61%

Tencent Cloud open-sourced TencentDB Agent Memory, using context offloading and a Mermaid task canvas to reduce token usage by up to 61% in multi-task continuous sessions while supporting OpenClaw integration and local SQLite storage.

Why it matters: Tencent open-sourced Agent Memory with a 61% token-saving claim and context offloading, clearing HKR-H/K/R. It is not a flagship model release, so it sits in the lower 78–84 band.

QbitAI · WeChat

Chinese GPU Vendor Hosts Open Source Meetup With SGLang Core Developers

Moore Threads said at the SGLang × MUSA Meetup that the MUSA backend has been merged into SGLang mainline, with 47 PRs submitted and 41 merged as of May 12.

Why it matters: HKR-H/K/R all pass, but this is an inference-backend ecosystem update rather than a model launch or platform shift. The 47 PRs and 41 merges make it concrete enough for featured, not P1.

QbitAI · WeChat

Alexandr Wang Responds to LeCun, Manus, and Meta AI Rebuild

Alexandr Wang said Meta rebuilt its pretraining, reinforcement learning, and data stacks in nine months, while Muse Spark remains closed because it triggered safety checks in areas including biosecurity, cyber capability, and loss of control.

Why it matters: HKR-H/K/R all pass: the named conflict draws clicks, the 9-month Meta stack rebuild and Muse Spark safety hold add facts, and open-source safety hits a real practitioner nerve. This is an interview, not a model launch, so it sits in the 78-84 band.

Latent Space

[AINews] Codex Rises, Claude Meters Programmatic Usage

Anthropic changed paid Claude plans to include monthly API credits equal to the subscription price, so a $200 plan includes $200 for programmatic usage outside Anthropic-owned harnesses, while OpenAI promoted Codex enterprise switching incentives in the same news cycle.

Why it matters: HKR-H/K/R all pass: the story ties Claude metering to Codex competition and gives a concrete $200 credit detail. This is a meaningful developer-cost update, not a major model or capability launch, so it sits in mid featured.

AI HOT (Curated Pool)

WeChat Group Chat Summary Skill Added, Depends on wx-cli Configuration

baoyu-skills added a WeChat group chat summary Skill that depends on wx-cli for data reading; the post provides two GitHub links and says Claude Code plus Claude Opus 4.6 gives the best results.

Why it matters: A small open-source tool update, but the workflow is highly relevant: WeChat data via wx-cli into Claude Code for group summaries. HKR-H/K/R pass; limited detail keeps it at the featured threshold.

AI HOT (Curated Pool)

UnslothAI Releases Qwen3.6 MTP GGUF Models With Over 1.4x Faster Inference

Daniel Han released experimental Qwen3.6 MTP GGUF models, with the 27B model reaching 140 tokens/s on one GPU and the 35B-A3B version reaching 220 tokens/s, using two draft tokens for speculative decoding.

Why it matters: HKR-H/K/R pass via concrete single-GPU speed claims and local-inference relevance. Score stays in low featured because the post is a single X source and does not disclose GPU, quantization settings, or repro steps.

AI HOT (Curated Pool)

xAI launches early beta of Grok Build

xAI launched an early beta of Grok Build for SuperGrok Heavy subscribers, offering a terminal-based coding agent with plan review, parallel subagents for large tasks, and a headless mode for scripting and automation.

Why it matters: HKR-H/K/R all pass: xAI enters terminal coding agents with plan mode, parallel subagents, and headless mode. Early beta access for SuperGrok Heavy keeps it below the 85 same-day must-write band.

AI HOT (Curated Pool)

Unlocking Asynchrony in Continuous Batching

Hugging Face says an 8B model generating 8K tokens leaves the GPU idle for 24% of the time, and asynchronous batching uses CUDA streams to overlap CPU preparation for batch N+1 with GPU computation for batch N.

Why it matters: HKR-H/K/R all pass, but this is inference-systems engineering rather than a major model release. The Hugging Face post provides a concrete 24% idle-rate number and CUDA-stream overlap mechanism, placing it in low featured.

The Verge · AI

Microsoft Edge Copilot update uses AI to pull information from across your tabs

Microsoft Edge will let Copilot gather information from all open tabs so users can ask questions, compare products, and summarize articles; the snippet says users can choose which experiences to enable, but the post does not disclose a rollout date.

Why it matters: HKR-H/K/R pass, but the post gives tab-wide reading, product comparison, and summaries without launch timing or deeper execution. This fits the lower featured band for a mid-weight product update.

TechCrunch · AI

Notion just turned its workspace into a hub for AI agents

Notion launched a developer platform that lets teams connect AI agents, external data sources, and custom code directly inside its workspace; the RSS snippet does not disclose pricing, rollout timing, supported models, or limits for the new platform.

Why it matters: HKR-H/K/R all pass, but price, launch timing, and supported models are not disclosed, keeping it in the 72–77 mid-weight product-update band. TechCrunch authority supports featured, not same-day must-write.

r/LocalLLaMA

24+ tok/s from ~30B MoE models on an old GTX 1080

User mdda ran Qwen 3.6 35B-A3B on an i7-6700, GTX 1080, and 32GB RAM machine at about 24 tok/s with 128k context; the setup uses llama.cpp MoE offloading plus TurboQuant/RotorQuant KV cache quantization, with PCIe 3.0 x16 saturated and GPU utilization at about 40–50%.

Why it matters: Single Reddit source limits authority, but the GTX 1080 + Qwen 3.6 35B-A3B + 128k + 24 tok/s setup gives a concrete local-inference result. HKR-H/K/R all pass; this is a practical featured item, not a major model or product launch.

AI HOT (Curated Pool)

Best Practices for Computer and Browser Use with Claude

Anthropic published guidance for Claude computer and browser use, with Claude 4.6 API screenshots capped at a 1,568-pixel long edge and 1.15 million total pixels, while Opus 4.7 raises the limits to 2,576 pixels and 3.75 million total pixels.

Why it matters: Anthropic’s first-party Claude computer/browser guide has actionable screenshot limits, not just promo copy. HKR-H/K/R all pass, but this is a practice guide rather than a major model or capability launch, so it sits in the 72–77 band.

AI HOT (Curated Pool)

Meta AI chief announces Incognito Chat for WhatsApp and Meta AI

Meta’s AI chief announced Incognito Chat for WhatsApp and Meta AI, with conversation inference running inside the phone’s hardware secure enclave, no server logs generated, and session data permanently deleted after the chat ends.

Why it matters: HKR-H/K/R all pass: the hook is Incognito Chat in WhatsApp, with secure-enclave inference and no server logs. Single-source brevity limits verification, so it sits below model releases and major capability launches.

AI HOT (Curated Pool)

Claude paid plans will offer monthly coding usage credits

Claude paid plans can claim monthly coding usage credits from June 15, covering Claude Agent SDK, claude -p, Claude Code GitHub Actions, and third-party apps built on the Agent SDK.

Why it matters: HKR-H/K/R all pass: the update names a date, quota mechanism, and covered Claude coding surfaces. Importance stays in the low featured band because this is a billing/access change, not a model release.

AI HOT (Curated Pool)

Introducing Runway Agent

Runway launched Runway Agent, a video creation agent that turns one natural-language conversation into multi-scene videos with narration, dialogue, and music; new free-plan users receive 1,500 credits for their first video.

Why it matters: HKR-H/K/R pass: a notable AI-video vendor ships an agentic multi-scene workflow with a 1,500-credit free plan. Score stays in the 72–77 band because the post is still a vendor announcement without pricing, limits, or independent tests.

The Verge · AI

Mark Zuckerberg announces ‘completely private’ encrypted Meta AI chat

Mark Zuckerberg announced Meta AI Incognito Chat, saying it stores no conversation logs on servers and uses end-to-end encryption; the post does not disclose rollout scope, retention audit details, or the key-management mechanism.

Why it matters: Meta’s Incognito Chat clears HKR-H with the privacy-contrast hook, HKR-K with E2E encryption plus no server logs, and HKR-R on trust. Missing rollout, retention audit, and key-management details keep it at the mid-weight product-update threshold.

AI HOT (Curated Pool)

Anthropic Launches Claude for Small Business Package

Anthropic launched Claude for Small Business with connectors and 15 ready-made automation workflows for QuickBooks, PayPal, HubSpot, and related business tools; users run tasks through Claude Cowork and manually approve key steps.

Why it matters: HKR-H/K/R all pass: the Anthropic SMB bundle has 15 workflows, named connectors, and a manual approval mechanism. It is a substantive Claude product update, but pricing, rollout scope, and usage data are not disclosed, so it stays below must-write.

May 13Wednesday

AI HOT (Curated Pool)

Open-source psql_bm25s speeds up PostgreSQL retrieval for multi-agent systems by 23x

The team open-sourced psql_bm25s, a native PostgreSQL access method for exact BM25 retrieval, and says it runs about 23x faster than pg_search on standard benchmarks.

Why it matters: HKR-H/K/R pass via the 23x retrieval-speed hook, named Postgres access method, and RAG latency pressure. Single-source release details lack independent reproduction and production constraints, so it stays in the lower featured band.

TechCrunch · AI

WhatsApp Adds an Incognito Mode in Meta AI Chats

WhatsApp added an incognito mode for Meta AI chats; Meta says these conversations are not saved, and messages disappear by default once the chat is closed.

Why it matters: HKR-H/K/R all pass: the privacy hook is clear, the retention mechanism is concrete, and WhatsApp gives it scale. Still, this is a single product feature, not a model or platform shift, so it sits at the featured threshold.