Skip to content

#MCP/工具调用

5 today

Mar 17Tuesday

OpenAI News

Introducing GPT-5.4 mini and nano

OpenAI released GPT-5.4 mini and nano on March 17, 2026 for coding and subagents; mini runs over 2x faster than GPT-5 mini. In the API, mini has a 400k context window and costs $0.75/$4.50 per 1M input/output tokens, while nano is API-only at $0.20/$1.25. The key signal is performance per latency: mini scores 54.4% on SWE-Bench Pro versus GPT-5.4 at 57.7%.

Why it matters: This is an official OpenAI model launch, not a routine patch. It includes concrete numbers—>2x speed, 400k context, API pricing, and 54.4% vs 57.7% on SWE-Bench Pro—so HKR-H/K/R all pass; scored at the low end of the 85–94 band.

MIT Technology Review · AI

Where OpenAI’s technology could show up in Iran

Just over two weeks after OpenAI’s classified-use deal with the Pentagon, MIT Technology Review outlined three places its tech could surface in Iran-related conflict. The post names target prioritization, Anduril counter-drone analysis, and GenAI.mil back-office use; it does not disclose when classified integration will finish or confirm deployment in Iran.

Why it matters: MIT Technology Review maps OpenAI’s classified-defense deal to 3 Iran-linked scenarios, giving it strong HKR-H and HKR-R. HKR-K is weaker because the piece does not confirm deployment, integration timing, or system limits, so it lands at the featured floor.

Mar 16Monday

MIT Technology Review · AI

Nurturing agentic AI beyond the toddler stage

The article says no-code tools and the open-source agent OpenClaw pushed agentic AI into a more autonomous stage between Dec. 2025 and Jan. 2026. It cites California AB 316 taking effect on Jan. 1, 2026, so firms cannot dodge liability by blaming AI, and an IDC survey sponsored by Data Robot reporting 96% of generative AI deployments and 92% of agentic AI deployments cost more than expected. The real issue is workflow-level governance: permission drift, orphaned agents, long-lived tokens, and sessions that can reach $100,000.

Mar 11Wednesday

MIT Technology Review · AI

Hustlers are cashing in on China’s OpenClaw AI craze

Beijing engineer Feng Qingyang turned OpenClaw installation support into a 100+ person business after starting in January, handling 7,000 orders at about RMB 248 each. Taobao and JD now show hundreds of related listings priced at RMB 100-700; the real story is setup friction and data-isolation risk turning an open-source agent into a service market.

Why it matters: Featured. HKR-H/K/R all pass: the side-gig-to-100-person-team angle is clickworthy, the piece adds hard market numbers, and the data-isolation risk gives it real industry resonance. This is not a product launch, but it is strong field reporting.

OpenAI News

From model to agent: Equipping the Responses API with a computer environment

OpenAI said on March 11, 2026 that Responses API now works with a shell tool and hosted container workspace, so models can execute commands in an isolated loop. The post says GPT-5.2 and later are trained to propose shell commands, while the API streams outputs and can run multiple commands concurrently across sessions; the container includes a filesystem, optional SQLite, and restricted network access. The key change is orchestration, not the “agent” label; pricing, quotas, and full security details are not disclosed in the visible post.

Why it matters: Substantive OpenAI developer update: the Responses API moves from tool calls to a managed computer environment with shell execution, streaming, parallel runs, and context compaction, so HKR-H/K/R all pass. The post is truncated and omits pricing, quotas, and full safety details,【

Mar 10Tuesday

NVIDIA Blog

NVIDIA and Thinking Machines Lab Announce Long-Term Gigawatt-Scale Strategic Partnership

NVIDIA and Thinking Machines Lab formed a multiyear deal to deploy at least 1 gigawatt of NVIDIA Vera Rubin systems, targeted for early next year, for frontier model training and customizable AI platforms. The partnership also covers training and serving system design for NVIDIA architectures and broader access to frontier and open models for enterprises and researchers; the post does not disclose the investment size. The key signal is the explicit 1-gigawatt compute commitment, not a routine cloud purchase.

Why it matters: The 1GW Vera Rubin commitment lifts this above routine partnership PR: HKR-H on scale, HKR-K on a named system with a dated deployment target, and HKR-R on frontier compute competition. It stays below P1 because the source is a vendor blog and key details—spend, ownership, and ph

OpenAI News

New ways to learn math and science in ChatGPT

OpenAI launched interactive math and science visualizations in ChatGPT on March 10, 2026, covering 70+ core concepts and rolling out globally across all plans. Users can adjust variables, manipulate formulas, and see graphs update in real time; OpenAI says 140 million people use ChatGPT weekly for math and science learning. The key point is productized interactivity, while the post does not disclose the underlying model, evaluation method, or outcome data.

Why it matters: HKR-H lands on the interactive-visual hook, HKR-K on 140M weekly learners plus 70+ concepts and live manipulation, and HKR-R on the product and edtech nerve. It is still a mid-weight product update; model details and learning-outcome evaluation are not disclosed, so it stays in a

Mar 9Monday

MIT Technology Review · AI

How AI Is Turning the Iran Conflict Into Theater

The author reviewed more than a dozen Iran-war dashboards in one week and argues they turn satellite data, ship tracking, AI summaries, and betting links into a real-time war spectator interface. The post cites a dashboard built by two Andreessen Horowitz staffers that pulls in Kalshi bets, while Craig Silverman has logged 20 similar dashboards. The point to watch is information quality: the piece cites Financial Times reporting on AI-generated satellite images spreading online, while these dashboards lack the human vetting and historical context used by intelligence agencies.

Why it matters: HKR-H lands on the war-dashboard-plus-betting hook; HKR-K lands on the named examples, counts, and Kalshi mechanism; HKR-R lands on reliability and ethics nerves for AI builders. Strong reported commentary, but not a product, model, or research milestone, so it ranks as featured,

OpenAI News

OpenAI to acquire Promptfoo

OpenAI said it will acquire Promptfoo and integrate its technology into OpenAI Frontier after closing. The post discloses that Promptfoo is used by over 25% of Fortune 500 companies, and the deal is still subject to customary closing conditions. The key signal is native agent security testing, red-teaming, and traceability in Frontier; the post does not disclose price or timeline.

Why it matters: This is not a routine partnership; OpenAI is absorbing a known eval and red-team vendor into Frontier. HKR-H/K/R all pass on novelty, concrete adoption data, and strong resonance with agent teams, but price, timing, and integration scope are still undisclosed, so it stays below p

Mar 7Saturday

Bloomberg Technology

Oracle and OpenAI End Plans to Expand Flagship Data Center

Oracle and OpenAI ended talks to expand a flagship AI data center in Abilene, Texas, after financing delays and OpenAI's changing needs. Meta is considering leasing the site from Crusoe, and Nvidia helped facilitate talks; the post only says such projects cost tens of billions of dollars.

Why it matters: Bloomberg reports that OpenAI and Oracle ended talks to expand the Abilene flagship site, with Meta potentially taking the parcel. HKR-H/K/R all pass: the reversal is strong, the story adds financing and demand detail, and the compute-capex angle will travel, but it is still an i

Bloomberg Technology

OpenAI, Oracle Won't Expand Flagship AI Data Center in Texas

OpenAI and Oracle have scrapped plans to expand a flagship AI data center in Texas after financing talks dragged and OpenAI's needs changed. The RSS snippet confirms only the Texas site; the post does not disclose the facility name, target capacity, capex, or revised timeline. The signal to watch is shifting compute demand, not just a stalled real estate project.

Why it matters: Bloomberg reports OpenAI and Oracle dropped a flagship Texas data-center expansion, citing financing delays and shifting OpenAI demand. HKR-H/K/R all pass and source authority helps, but missing capacity, capex, and timeline details keep it in the low 80s.

Bloomberg Technology

Oracle and OpenAI End Plans to Expand Flagship Data Center

Oracle and OpenAI ended plans to expand a flagship AI data center in Texas. The RSS snippet says talks dragged over financing and OpenAI’s changing needs; the post does not disclose the site’s size, budget, or timeline. The real signal is financing friction plus a demand reassessment.

Why it matters: Bloomberg reports a meaningful infrastructure reversal, so HKR-H and HKR-R land: it is unexpected and it hits compute-supply and capex concerns around OpenAI. HKR-K is limited because the writeup omits size, spend, and timing, keeping this near the featured threshold.

MIT Technology Review · AI

Is the Pentagon allowed to surveil Americans with AI?

MIT Technology Review reports that the Pentagon sought to use Anthropic Claude to analyze bulk commercial data on Americans, triggering a public clash; OpenAI then revised its contract to bar intentional domestic surveillance of U.S. persons. The key mechanism disclosed is that the U.S. government can buy commercial location and browsing data, and if collection is deemed lawful, current law often does not restrict feeding it into AI for aggregation and profiling. The real issue is that contract red lines may not bind the DoD; OpenAI has not released the full contract, and the post does not disclose how its safety stack would be enforced.

Why it matters: Full HKR-H/K/R: strong Pentagon-surveillance hook, a concrete legal mechanism on commercial data reuse, and clear resonance for defense-contract and safety-boundary debates. It stops short of 85 because the new OpenAI contract text and enforcement details are not disclosed.

Bloomberg Technology

OpenAI Releases AI Agent Security Tool for Research Preview

OpenAI released a research-preview AI agent for security teams to find and patch vulnerabilities in large databases. The RSS snippet discloses the use case and preview status, but the post does not disclose the model name, supported databases, pricing, or rollout timeline. Watch the deployment boundary, not the headline alone.

Why it matters: HKR-H lands because OpenAI is shipping an agent for vuln discovery and patching; HKR-R lands because security automation is a live enterprise nerve. HKR-K is weak: the preview lacks model, coverage, pricing, and rollout details, so this stays at the featured floor.

Bloomberg Technology

Anthropic Unveils Amazon-Inspired Marketplace for AI Software

Anthropic is launching a platform for enterprise customers to buy third-party software, expanding its AI offerings. The RSS snippet confirms the audience and purpose, but the post does not disclose launch timing, revenue terms, or software scope. The key signal is a move from selling models toward channel distribution, as the company faces business uncertainty tied to a Pentagon standoff.

Why it matters: This is a meaningful channel move: Anthropic is extending from model sales into enterprise software distribution. HKR-H and HKR-R pass, but HKR-K is limited because launch timing, rev share, and catalog scope are not disclosed, so it lands at the low end of featured.

Mar 6Friday

Ruan YiFeng's Weblog

Technology Enthusiast Weekly Issue 387: You Are Ahead

Ruanyifeng says that, out of 8.1 billion people, only 1.38 billion have used AI, or 16%; just 15 to 25 million pay for AI services, or 0.3%. The post adds that only 2 to 5 million people have used AI to create their own coding projects, or 0.04%. The real signal is the adoption gap, not the idea that everyone already uses AI.

Why it matters: This is data-backed commentary, not a product launch or primary reporting. HKR-H/K/R all pass: the angle punctures the 'everyone uses AI' narrative and supplies 16% / 0.3% / 0.04% adoption estimates, but the source basis is unclear here, so it sits at the low end of featured.

Mar 5Thursday

OpenAI News

Introducing ChatGPT for Excel and new financial data integrations

OpenAI launched ChatGPT for Excel beta on March 5, 2026, bringing GPT-5.4 into Excel workbooks and finance workflows. The post says it can build and update models, trace changes to cells, and is off by default for Enterprise and Edu admins; OpenAI's internal banking benchmark rose from 43.7% with GPT-5 to 87.3% with GPT-5.4 Thinking. The key move is data access: Moody’s, Dow Jones Factiva, MSCI, Third Bridge, and MT Newswires are live, while FactSet is listed as coming soon.

Why it matters: This is more than a routine add-on: OpenAI puts ChatGPT into Excel, names major finance data feeds, and cites a 43.7%→87.3% internal banking benchmark gain. HKR-H/K/R all pass; importance lands at 82 because this is a strong vertical workflow move, not a market-wide model release

Mar 3Tuesday

OpenAI News

GPT-5.3 Instant: Smoother, more useful everyday conversations

OpenAI released GPT-5.3 Instant on March 3, 2026 as an update to ChatGPT’s most-used model, aiming for fewer unnecessary refusals, fewer disclaimers, and more accurate everyday answers. The post shows one concrete contrast: GPT-5.2 Instant refused long-range archery trajectory help, while GPT-5.3 Instant requested parameters and gave a no-drag example at 300 fps (about 91 m/s), 45°, and 845 m; the key issue is the safety-boundary shift, while the post does not disclose benchmark scores, system card details, or API pricing.

Why it matters: OpenAI updated a core ChatGPT everyday model, and the story clears HKR-H/K/R because the refusal-boundary shift is concrete and widely relevant. The post includes a specific 5.2 vs 5.3 behavior example, but no system card, benchmark table, or API pricing, so it lands below the 85

Feb 27Friday

MIT Technology Review · AI

AI is rewiring how the world’s best Go players think

AI has become standard in pro Go training in South Korea, and the piece says competing professionally without it is now essentially impossible. It cites two figures: Shin Jin-seo matches AI moves 37.5% of the time versus a 28.5% player average, and AlphaGo Zero beat AlphaGo Lee 100-0 after three days of training. The shift to watch is training, not hype: KataGo is now a common tool, opening moves often mirror AI for the first 50 turns, and even top players still cannot fully explain its choices.

Why it matters: Strong HKR-H/K/R: the novelty is elite cognition shifting under AI, and the story brings concrete numbers plus a named tool. It is a reported commentary rather than a new model or product move, so it sits at the low end of featured.

OpenAI News

OpenAI and Amazon announce strategic partnership

OpenAI and Amazon announced a multi-year partnership, with Amazon investing $50 billion in OpenAI: $15 billion upfront and $35 billion tied to conditions. They will launch a Stateful Runtime Environment on Amazon Bedrock, and OpenAI will consume about 2 gigawatts of Trainium capacity on AWS. The part to watch is distribution plus compute lock-in: AWS becomes the exclusive third-party cloud distributor for OpenAI Frontier.

Why it matters: This is not a routine partnership post. The disclosed $50B staged investment, Bedrock runtime, and ~2GW Trainium commitment change OpenAI's distribution and compute posture; HKR-H/K/R all pass, so this lands in P1.

36Kr (direct RSS)

From short video to long-form: Douyin is also handing news to AI

Douyin launched long-form posts in late 2025, raising the cap from 4,000 to 8,000 Chinese characters, and added “AI-selected news” summaries in its Hot topics tab. Long-form publishing is web-only for now, and the post says AI news will enter the main feed, but it does not disclose ranking weight, licensing scope, or fact-checking rules. The real issue is distribution and accountability: AI summaries and original articles will compete in the same traffic pool.

Why it matters: This clears HKR-H/K/R: Douyin putting AI summaries into its hot-news surface is a strong hook, and the piece includes concrete mechanics like the 8,000-word cap, web-only publishing, and follow-up queries. The real industry angle is distribution, copyright, and fact-checking, but

Ruan YiFeng's Weblog

Weekly for Technology Enthusiasts #386: When Delivery Workers Plug Into AI

Waymo placed a $6.25 task on a delivery platform to send a rider 1 km away to close a robotaxi door, with another $5 after completion. The post frames this as software dispatching human labor, not a one-off gig, and argues platform workers are becoming a human API inside automated workflows. The point to watch is the AI-plus-labor loop; the post does not disclose Waymo's scale, frequency, or formal product design.

Why it matters: Not a primary-source scoop, but the $6.25+$5 Waymo case makes the “humans as API” mechanism concrete. HKR-H/K/R all pass; score stays at the low end of featured because this is commentary and scale, frequency, and a formal product path are not disclosed.

Feb 26Thursday

New York Times Chinese

Where Is the U.S. Losing to China in AI?

The piece argues China has embedded AI into manufacturing, with 30,000+ smart factories, and over half of all industrial robots installed globally in 2024 going to Chinese plants. It cites shop-floor data: Zeekr's Ningbo plant uses 800+ robots, Xiaomi says its Beijing factory produces one car every 76 seconds, while only 18% of U.S. manufacturers report a formal AI strategy and two-thirds struggle to scale pilots. The real point is not frontier models but AI deployment in factory automation, scheduling, and inspection.

Why it matters: Data-backed commentary with all three HKR axes: a strong US-vs-China hook, concrete factory metrics, and direct resonance on AI deployment and competitiveness. Not a new product, model, or research release, so it stays in the low featured band.

OpenAI News

OpenAI Codex and Figma launch code-to-design roundtrip workflow

OpenAI and Figma launched a Codex integration on Feb. 26, 2026 that turns code into editable Figma designs and brings Figma Design, Figma Make, and FigJam content back into code. The workflow uses MCP via the Figma MCP Server in the Codex desktop app; OpenAI says Codex has 1M+ weekly users and usage is up 400%+ since the start of the year. The key issue is whether roundtrip context stays intact; the post does not disclose supported models, permission boundaries, or pricing.

Why it matters: This is a solid OpenAI/Figma workflow update with clear HKR-H/K/R: a bidirectional code↔design loop via MCP and Figma MCP Server. It stays below 85 because the post does not disclose model support, permission boundaries, pricing, or roundtrip reliability.

Feb 20Friday

Hugging Face Blog

GGML and llama.cpp join Hugging Face to support the long-term progress of Local AI

Hugging Face said the GGML and llama.cpp team is joining the company, while Georgi Gerganov’s team will still spend 100% of its time maintaining llama.cpp. The post says the project remains 100% open source and community driven, with full technical and community autonomy. The key angle is tighter delivery from transformers model definitions into llama.cpp, aiming for near “single-click” shipping; the post does not disclose timeline, team size, or deal terms.

Why it matters: This is a meaningful local-AI infrastructure move: HF brings in the GGML/llama.cpp team, so HKR-H/K/R all pass. I kept it at 78 because the post confirms staffing and integration direction, but not a ship date, team size, or deal terms.

MIT Technology Review · AI

Microsoft has a new plan to prove what’s real and what’s AI online

Microsoft evaluated 60 combinations of provenance, watermarking, and fingerprinting methods, and shared a blueprint with MIT Technology Review for labeling AI-manipulated content online. The plan only indicates origin and manipulation, not truthfulness; an audit found just 30% of test posts were labeled correctly, so the real issue is adoption and execution by platforms.

Why it matters: HKR-H/K/R all pass: strong hook, two concrete facts (60 combinations tested, 30% correct labels), and a live trust-infrastructure debate. It stays at featured, not higher, because this is a blueprint and standards problem, not a deployed product or binding rule.

Feb 15Sunday

Computing Life · Yage

OpenClaw deep dive: why it suddenly took off, and what it means for us

OpenClaw surged in late January 2026 because it plugged local coding agents into Slack, WhatsApp, and Feishu, giving non-technical users file access, command execution, and persistent memory in a chat UI. The article also names the costs: 12% of third-party skills contained malicious code, and the $CLAWD token scam took $16 million; the chat interface remains linear, low-density, and hard to observe. The real takeaway is not to copy OpenClaw blindly, but to reuse its unified context, file-based memory, and composable skills in a controllable stack like OpenCode.

Why it matters: This is more than a recap: it breaks down OpenClaw's adoption mechanism, downside, and reusable design pattern. HKR-H/K/R all pass with two hard facts—12% malicious skills and a $16M scam—but as a personal analysis rather than an official release or industry event, it lands in `+

Computing Life · Yage

OpenClaw Deep Dive: Why It Went Viral and What It Means for You

The post says OpenClaw went viral in late January 2026, changed names 3 times in one week, and a $CLAWD scam token took $16 million. It cites two concrete risks: 12% of third-party skills had malicious code, and some users exposed consoles to the public internet without passwords. The excerpt is truncated, but the core claim is distribution: OpenClaw put agentic AI into WhatsApp, Slack, and Lark for non-technical users.

Why it matters: HKR-H/K/R all pass: the viral arc is dramatic, the post includes a 12% malicious-skills figure and a specific exposed-console risk, and the distribution angle matters to agent builders. It is still a secondary deep-dive, not a primary launch or official research, so 78 and tiered

Feb 14Saturday

Ruan YiFeng's Weblog

Using ByteDance's Seed 2.0 and TRAE with Skills for app building and deployment

Ruanyifeng used ByteDance's Seed 2.0 Code and TRAE to generate one ASCII-to-Excalidraw web app and preview it at localhost:8080. The post says Seed 2.0 includes Pro, Lite, Mini, and Code models, and shows Skills as YAML-headed Markdown files, including Anthropic's frontend-design and Vercel deploy examples.

Why it matters: HKR-H and HKR-K land because the post turns Seed 2.0 Code + TRAE into a runnable mini app and explains the Skill mechanism with concrete setup details. HKR-R also lands for coding-agent workflow reuse, but this is a strong tutorial, not a major ByteDance launch, so it sits at the

MIT Technology Review · AI

ALS stole this musician’s voice. AI let him sing again.

Patrick Darling, 32, returned to the stage on February 11 in London after two years without singing, using an AI voice clone rebuilt from old recordings. The post says speech cloning typically needs about 10 minutes of clean audio; his singing clone was built from noisy phone clips and kitchen recordings, then refined with Eleven Music over about six weeks. The practical signal is access, not sentiment: ElevenLabs offers the tools free to people who lost their voices to ALS and similar conditions, but the post does not disclose model details.

Why it matters: HKR-H/K/R all land: the hook is strong, the story gives concrete reproducible details, and the use case hits accessibility plus voice-rights nerves. Still, this is a strong application story, not a major model, product, or research release, so it stays in low featured.

Feb 12Thursday

Lex Fridman (YouTube RSS)

OpenClaw: The Viral AI Agent Behind the Hype - Peter Steinberger | Lex Fridman Podcast #491

Lex Fridman’s episode #491 interviews Peter Steinberger about the open-source AI agent OpenClaw; the transcript says it reached 175k-180k GitHub stars. The post says it can connect to Telegram, WhatsApp, Signal, and iMessage, and use models such as Claude Opus 4.6 and GPT 5.3 Codex; it does not fully disclose the architecture, evals, or security boundaries. The real point is system-level access and self-modifying behavior: this is not chat, but an agent that can take actions.

Why it matters: This is more than a routine podcast. OpenClaw scores on HKR-H/K/R with 175k-180k GitHub stars, messaging integrations, and self-modifying behavior. It stays at featured, not p1, because the post does not disclose architecture, evaluations, or safety boundaries.

MIT Technology Review · AI

Is a secure AI assistant possible?

OpenClaw was uploaded to GitHub in November 2025 and went viral in late January, extending LLMs into email, browsing, and local files with larger security risks. The post names prompt injection as the central threat, says there are likely “hundreds of thousands” of OpenClaw agents online, and notes a public warning from the Chinese government. The key point: the article says there is no silver-bullet defense yet, and the truncated body does not disclose the full mitigation details.

Why it matters: This is not a launch, but it clears HKR-H/K/R: the question is a strong hook, the piece adds concrete scale plus 'no silver-bullet' defense, and it hits the agent-builder safety nerve. Featured, not p1, because the article does not disclose reproducible mitigations.

Feb 10Tuesday

36Kr (direct RSS)

OpenAI to integrate ChatGPT into the U.S. Department of Defense's generative AI platform

The U.S. Department of Defense will work with OpenAI to integrate ChatGPT into GenAI.mil for about 3 million personnel. The RSS snippet discloses the integration, platform name, and user count, but not the model version, access controls, scope, or launch date.

Why it matters: A DoD distribution deal for ChatGPT at roughly 3M-seat scale clears HKR-H, K, and R. The ceiling stays below p1 because the post confirms the platform and reach only; model version, access controls, and launch timing are not disclosed.

36Kr (direct RSS)

Embodied AI company Noematrix raises several hundred million yuan in Series A, with overseas funds joining

Noematrix closed a Series A worth several hundred million yuan, led by C Capital, with Sea Limited and Puhua Capital participating, and Prosperity7 Ventures increasing its stake. Founded in Nov. 2023, the company says its Noematrix Brain has been deployed on wheeled single-arm, wheeled dual-arm, and humanoid dual-arm robots in retail pharmacies and hotel laundries; the post does not disclose valuation or revenue. The sharper signal is its claimed hundreds of thousands of hours of real-robot data and its data-model-scenario loop.

Why it matters: HKR-H/K/R all pass: the funding hook is strong, and the body adds real-world data plus deployed robot forms and scenarios. It stays at the low end of featured because this is still a single-company financing scoop, and valuation, revenue, and customer counts are not disclosed.

Feb 9Monday

36Kr (direct RSS)

Voice Ask is live: why is Xiaohongshu pushing search-by-question?

Xiaohongshu fully launched Voice Ask on Jan. 27, letting users long-press to speak on the search page and get structured answers distilled from in-app user experience posts. The post says it can handle 3-minute spoken queries, foreign languages, and dialects, but does not disclose the model, ASR stack, latency, or accuracy. The real shift is from 3-4 character keyword search to longer spoken questions, widening search intent capture and scenario coverage.

36Kr (direct RSS)

Qwen’s 10 Million Milk Teas: How Alibaba’s Massive AI Freebie Campaign Unfolded

Alibaba’s Qwen drove over 10 million orders via a Feb. 6 free-order campaign, but the app slowed and crashed from 10 a.m. to noon as load exceeded capacity; orders had already passed 2 million before noon. 36Kr says initial server capacity was only about one-third of the expected peak, and the subsidy pool was framed as 3 billion yuan; the real signal is not a model leap but a paid test of AI commerce entry and consumer acquisition.

Why it matters: HKR-H lands on the free-milk-tea plus outage hook, while HKR-K lands on concrete scale and capacity numbers. HKR-R also lands because the story speaks to AI distribution, subsidy economics, and infra reliability, but it remains a single-company promo test rather than a market-shi

Feb 7Saturday

MIT Technology Review · AI

Moltbook was peak AI theater

Moltbook went viral within hours, and the platform says it now has 1.7 million agent accounts, 250,000 posts, and 8.5 million comments, but the article argues the activity is mostly human-scripted mimicry. It says OpenClaw can connect Claude, GPT-5, or Gemini to tools like email and browsers; cited operators say the agents lack shared goals, shared memory, and self-directed autonomy, and some viral posts were written by humans posing as bots. The key takeaway is risk: agents tied to private data such as passwords or bank details were active on a site filled with spam and potentially malicious instructions.

Why it matters: This is strong anti-hype commentary, not a market-moving event. HKR-H/K/R all pass: the hook is sharp, the piece adds 1.7M/250k/8.5M plus concrete critique on memory and goals, and the security angle lands with practitioners, so it clears featured but stays mid-70s.

Feb 6Friday

TechCrunch · AI

OpenAI launches new agentic coding model minutes after Anthropic releases its own

OpenAI launched an agentic coding model minutes after Anthropic released a similar one, and the model is meant to accelerate Codex, which OpenAI launched earlier this week. The RSS snippet gives only the timing and purpose; the post does not disclose the model name, benchmarks, pricing, context length, or availability. The signal is direct competition in agentic coding, not a substantiated performance claim.

Why it matters: Major-lab product news plus a minutes-apart Anthropic clash gives this HKR-H and HKR-R. The score stays in the low featured band because HKR-K is weak: the post lacks the model name, benchmarks, price, context window, and availability.

Feb 4Wednesday

TheValley101 (硅谷101)

E224 | Why Clawdbot became the first breakout product of 2026 amid the Mac mini rush | Moltbot | MoltBook | OpenClaw

The podcast says Clawdbot passed 100k GitHub stars within days and reached 146k on Feb. 2, while being renamed to Moltbot and then OpenClaw within a week. It attributes the traction to a stack of Claude, long-term memory, IM-based messaging, and proactive heartbeat workflows; the title mentions a Mac mini rush, but the post does not disclose sales figures. The real signal is the interaction layer rather than a new model release: this is industry commentary and user anecdotes, not an official spec sheet.

Why it matters: This is a commentary-led breakdown of a hot agent phenomenon, not a primary launch. HKR-H/K/R all pass: the 146k-star surge and rename chain are novel, the post explains memory + IM + heartbeat mechanics, and it hits nerves on agent UX, dedicated hardware, and security bills; the