Skip to content

#OpenAI

45 today

May 12Tuesday

Xinzhiyuan · WeChat

OpenAI releases GPT-Realtime-2, described as a GPT-5-level reasoning audio model

OpenAI released GPT-Realtime-2 alongside Realtime-Translate and Realtime-Whisper, with a 128K context window, five reasoning-effort levels, and API pricing of $32 per million input tokens and $64 per million output tokens.

Why it matters: HKR-H/K/R all pass: realtime audio reasoning is a strong hook; 128K context, five reasoning levels, and $32/$64 per 1M tokens add substance; voice-agent cost and stack choices hit practitioners. This is a same-day OpenAI product update.

Latent Space

Thinking Machines' Native Interaction Models: TML-Interaction-Small 276B-A12B Advances Realtime Voice

Thinking Machines released TML-Interaction-Small, a 276B-parameter MoE model with 12B active parameters, and the post says it advances realtime voice through 200ms time-aligned microturns, encoder-free early fusion for audio and images under 200ms, and benchmark wins over GPT-Realtime-2 and Gemini 3.1-Flash.

Why it matters: HKR-H/K/R all pass: TML-Interaction-Small gives architecture, active parameters, 200ms interaction, and named rivals. Benchmarks still need replication, but a real-time voice SOTA claim is same-day material.

AI HOT (Curated Pool)

Thinking Machines Releases Native Multimodal Interaction Model for Real-Time Human-AI Collaboration

Thinking Machines released an interaction model that natively receives audio, video, and text input, processes foreground interaction at 200-millisecond intervals, and uses a background reasoning model for long-horizon planning and tool calls.

Why it matters: HKR-H/K/R all pass: this is more than a model notice, with a two-layer foreground/background interaction design. Pricing, access scope, and benchmarks are missing, so it sits at the lower end of 85-94.

AI HOT (Curated Pool)

What Parameter Golf Taught Us About AI-Assisted Research

OpenAI’s Parameter Golf brought together over 1,000 participants and more than 2,000 submissions to test AI-assisted machine learning research, coding agents, model quantization, and model design under strict parameter constraints.

Why it matters: OpenAI’s Parameter Golf recap clears HKR-H/K/R with a concrete contest, 1,000+ participants, and 2,000+ submissions. It is research/benchmark signal, not a model or product launch, so 78 fits the lower featured band.

The Verge · AI

OpenAI just released its answer to Claude Mythos

OpenAI launched Daybreak, a security initiative that uses the Codex Security AI agent released in March to model an organization’s code, validate likely vulnerabilities, and automate detection of higher-risk issues before attackers find them.

Why it matters: HKR-H/K/R all pass: Daybreak has a rivalry hook, concrete agent workflow, and code-security resonance. It is narrower than a model or ChatGPT capability release, so it stays in the 78–84 band.

The Verge · AI

Here’s What Mira Murati’s AI Company Is Up To

Thinking Machines announced work on “interaction models” that continuously take in audio, video, and text and respond or act in real time; the post does not disclose model size, release timing, pricing, or the final product format.

Why it matters: HKR-H/K/R all pass, but the body lacks parameters, launch timing, and product form. This is a high-interest startup direction reveal, not a usable model release, so it stays at the top of the 72–77 band.

AI HOT (Curated Pool)

Introducing Daybreak: Frontier AI for Cyber Defenders

OpenAI introduced Daybreak for cyber defenders, combining OpenAI models, Codex, and security partners; the post does not disclose pricing, launch timing, or concrete defense metrics.

Why it matters: OpenAI’s Daybreak announcement clears HKR-H/R as a security-focused product hook, but HKR-K fails: no defense metrics, access terms, or pricing. That keeps it in the 72–77 product-update band.

Bloomberg Technology

Microsoft Targeted $92 Billion Return on Early OpenAI Investment

Microsoft targeted a $92 billion return from its early OpenAI investments, according to the RSS snippet; the post does not disclose the investment terms, payout schedule, or actual realized return.

Why it matters: HKR-H/K/R all pass because Bloomberg adds the $92B target and the Microsoft-OpenAI economics are high-salience. Terms, timing, and realized returns are not disclosed, so it stays below must-write.

Bloomberg Technology

Sutskever Says His OpenAI Stake Is Worth About $7 Billion

Ilya Sutskever said his OpenAI stake is worth roughly $7 billion, making him one of its largest individual shareholders; the RSS snippet does not disclose his ownership percentage, valuation method, or transaction terms.

Why it matters: HKR-H/K/R all pass: Bloomberg gives a hard $7B number for Sutskever’s OpenAI stake. Missing stake percentage, valuation basis, and transaction terms keep it in the featured-threshold band, not must-write.

May 11Monday

AI HOT (Curated Pool)

Fields Medalist Tests ChatGPT 5.5 Pro: Paper-Level Result in 17 Minutes

Timothy Gowers tested ChatGPT 5.5 Pro and said it independently solved an open additive number theory problem in 17 minutes with only a simple prompt, producing PhD thesis-level work; he warned that this pace threatens mathematics research training, while Terence Tao said human value lies in digesting and deeply understanding proofs.

Why it matters: HKR-H/K/R all pass: a named mathematician, a 17-minute result, and a PhD-training warning. The exact problem, prompt, and verification path are not disclosed, keeping it below P1.

AI HOT (Curated Pool)

Cog House Opens for the First Time: Scott Wu and the Rise of Cognition AI

Cognition AI disclosed internal footage of Cog House, while Devin reached $445 million in annualized revenue within 18 months of launch and the company is valued at about $25 billion.

Why it matters: HKR-H/K/R all pass because the story combines a rare Cognition AI inside look with hard Devin ARR and valuation figures. It stops below P1 because this is a profile-style reveal, not a funding, product, or model release.

AI HOT (Curated Pool)

OpenAI Launches DeployCo to Help Enterprises Build Businesses Around Intelligence

OpenAI launched DeployCo, an enterprise deployment company focused on moving AI systems into production, while the RSS snippet does not disclose pricing, customer names, deployment scope, or launch timeline.

Why it matters: OpenAI launching DeployCo is a real enterprise strategy signal: HKR-H has a separate-company hook and HKR-R hits deployment competition. HKR-K is weak because pricing, customers, and timing are absent, so it sits at the featured floor.

QbitAI · WeChat

OpenAI backs Cerebras as the Nvidia challenger targets a $35B IPO valuation

Cerebras raised its IPO price range to $150-$160 per share, targeting about a $35 billion valuation at the top end, after OpenAI signed a 750-megawatt AI compute purchase agreement with deliveries through 2028.

Why it matters: HKR-H/K/R all pass: this is not a routine IPO note, since OpenAI’s 750MW purchase agreement anchors Cerebras at a reported $35B valuation and feeds the NVIDIA-alternative compute story.

AI HOT (Curated Pool)

Cerebras IPO reportedly over 20 times oversubscribed, with pricing set to rise nearly 30%

Cerebras received more than 20 times oversubscription for its IPO and plans to raise the share count from 28 million to 30 million while increasing the price range to $150-$160.

Why it matters: HKR-H/K/R all pass: the Cerebras IPO repricing has rare demand numbers and clear AI-infrastructure resonance. It stays in the lower 85-94 band because this is pricing news, not the actual listing or a new chip launch.

Computing Life · Share · Yage

DeployCo Arrives: OpenAI and Anthropic Form AI Deployment JVs with PE on the Same Day

OpenAI and Anthropic announced AI deployment joint ventures with private equity on May 4, and the snippet cites divergent terms, including a 17.5% guaranteed return versus no guaranteed return.

Why it matters: HKR-H/K/R all pass: the angle has tension, the facts include PE JVs and a 17.5% floor, and the nerve is model-lab commercialization. Single-source commentary keeps it in the 78–84 band, not must-write.

Computing Life · Share · Yage

Google shuts down Project Mariner; Anthropic and OpenAI also hit limits

Google quietly shut down Project Mariner on May 4, and the post says Google, Anthropic, and OpenAI reached the same conclusion: standalone browser agents do not work, while GUI automation still has room outside headless dedicated environments.

Why it matters: HKR-H/K/R all pass: the shutdown date, route-level claim, and Google/OpenAI/Anthropic contrast carry signal. Single-source summary lacks an official notice or failure metrics, so this stays in the low featured band.

AI HOT (Curated Pool)

Older AI Model Outperforms Human Doctors in Emergency Diagnosis

A Science study reports that OpenAI o1 reached a 67% correct or near-correct diagnosis rate on real emergency department data, exceeding doctors at 50-55%, but the study did not cover long-term inpatient data or imaging diagnosis.

Why it matters: HKR-H/K/R all pass: a Science-linked ER benchmark reports o1 at 67% versus doctors at 50-55%. It stays below P1 because it is one diagnostic study and excludes inpatient and imaging settings.

May 10Sunday

Financial Times · Technology

OpenAI trial lays bare rivalries behind start-up’s $852bn rise

The title says OpenAI’s rise reached an $852bn valuation; the RSS snippet only discloses that Elon Musk’s lawsuit is entering its final week in court and that Sam Altman is due to testify.

Why it matters: HKR-H/K/R all pass: FT ties the OpenAI trial to Musk/Altman rivalry and an $852bn valuation. No model, product, IPO, or executive-change trigger, so it lands in the good-quality featured band, not P1.

Xinzhiyuan · WeChat

Harsh Claim: Top Silicon Valley AI Is One Year Ahead of the World

Elad Gil claims top AI lab employees are 3-4 months ahead of Silicon Valley, while Silicon Valley is 3-6 months ahead of New York; the post cites Mythos’ 73% success rate in expert cyberattack simulations as evidence in a disputed “geographic time gap” argument.

Why it matters: HKR-H/K/R all pass: the lab-to-user lag hook is clickable, and the post cites 3–4 months, 3–6 months, and a 73% Mythos figure. It is secondhand commentary, not a model or product release, so it stays in the 72–77 threshold band.

May 9Saturday

TechCrunch · AI

Nvidia has already committed $40B to equity AI deals this year

Nvidia has committed more than $40 billion to equity investments in AI companies in 2026, including a $30 billion investment in OpenAI, seven multi-billion-dollar public-company deals, and around two dozen private startup rounds, according to CNBC and FactSet data cited by TechCrunch.

Why it matters: HKR-H/K/R all pass: the $40B hook is strong, with $30B to OpenAI and about 24 deals disclosed. It is a capital-structure signal, not a model or product launch, so it sits in the 78–84 band.

Synced · WeChat

OpenAI's Jiayi Weng: Is the Next AI Training Paradigm Beyond Gradients?

OpenAI researcher Jiayi Weng proposes Heuristic Learning: codex gpt-5.4 reached a perfect 864 score on Breakout and generated 342 search trajectories across Atari 57, with updates applied to code, tests, replays, and memory rather than neural-network weights.

Why it matters: HKR-H/K/R all pass: an OpenAI researcher proposes Heuristic Learning with concrete hooks like Breakout 864 and 342 Atari 57 trajectories. This is strong research/commentary signal, not an official model or product release, so it stays in the 78–84 band.

Latent Space

Anthropic growing 10x/year while others lay off over 10% of staff

Anthropic is described as growing 10x annually and being valued at $1T-$1.2T, while the post cites layoffs of 40% at Block, 14% at Coinbase, and 20% at Cloudflare under AI-readiness framing.

Why it matters: HKR-H/K/R all pass: the title has contrast, the post gives growth, valuation, and layoff figures, and it hits jobs plus AI-capital concentration. It is high-signal industry commentary, not an official funding or product event, so 78-84 fits.

AI HOT (Curated Pool)

DeepSeek Raises $7 Billion at Record Scale, Founder Personally Invests $3 Billion

DeepSeek is raising up to $7 billion at a $50 billion valuation, with founder Liang Wenfeng personally contributing $3 billion, or 40% of the round, while the company says the funding will target large-scale compute, V4.1 model releases, enterprise products, and a path toward positive revenue.

Why it matters: HKR-H/K/R all pass: the $3B founder check is a strong hook, with concrete funding numbers and clear competitive resonance. Single X-source sourcing leaves lead investor, terms, and confirmation undisclosed, so it stays below P1.

MIT Technology Review · AI

Musk v. Altman Week 2: OpenAI Fires Back, and Shivon Zilis Says Musk Tried to Poach Sam Altman

Greg Brockman testified that Elon Musk pushed OpenAI in 2017 to create a for-profit arm and sought majority equity, board control, and the CEO role; Musk now asks the court to remove Sam Altman and Brockman, unwind OpenAI’s restructuring, and award up to $134 billion from OpenAI and Microsoft.

Why it matters: HKR-H/K/R all pass: the story adds testimony on Musk’s 2017 control push and a $134B claim against OpenAI and Microsoft. It is a strong legal-governance update, not a model or product release, so it stays in the 78–84 band.

The Verge · AI

All the Latest Updates on AI Data Centers

The Verge tracks AI data center disputes with specific updates: 43% of Americans blame data centers for rising power bills, a 40,000-acre Utah project won approval despite local opposition, and Anthropic says it will invest $50 billion in US AI data centers.

Why it matters: HKR-H/K/R all pass, but this is a Verge running roundup rather than a single breakout event. The concrete power-grid and capex numbers place it at the upper end of industry reporting.

May 8Friday

AI HOT (Curated Pool)

Robotics Endgame: A Physical AGI Roadmap and LLM Analogy

The speaker presented a physical AGI roadmap with six named components: video world models, WAM, EgoScale, dexterity scaling laws, physical reinforcement learning, and DreamDojo; the snippet also mentions a 2016 OpenAI DGX-1 signing story with Jensen and Elon.

Why it matters: HKR-H/K/R all pass: the physical-AGI endgame hook is strong, the post gives a 6-part roadmap, and robotics practitioners will debate the path. It is still a personal roadmap, not a release or benchmark, so it sits in 78–84.

AI HOT (Curated Pool)

Running Codex Safely at OpenAI

OpenAI runs Codex with four safeguards: sandbox isolation, human approval, strict network policies, and native agent telemetry; the post does not disclose evaluation metrics, incident rates, or enterprise deployment requirements.

Why it matters: HKR-H/K/R all pass: the OpenAI Codex post gives concrete safety mechanisms for code agents. I keep it at 74 because it lacks eval data, incident rates, or enterprise rollout details.

Alibaba Technology · WeChat

The AI-Native Era: Where R&D Organizations Go Next

Xu Xiaobin cites internal interviews showing that engineers who use AI heavily cut coding time from 30% to 5%, raised Agent conversation time from 5% to 60%, and increased end-to-end delivery efficiency by 2 to 3 times, while pure coding efficiency rose 10 times.

Why it matters: Alibaba Tech’s internal-interview numbers make HKR-H/K/R pass, but this is org-methodology commentary rather than a product or model release, so it sits just above the featured threshold.

Synced · WeChat

OpenAI launches official CLI for terminal-based model access

OpenAI released the open-source openai-cli, letting developers call Responses, cloud tools, image generation and editing, speech transcription, and TTS from a single terminal command.

Why it matters: HKR-H/K/R all pass: an official OpenAI CLI, open-source packaging, and terminal access to multimodal APIs. This is a useful developer workflow update, not a major model capability release, so it sits in low featured.

Latent Space

[AINews] GPT-Realtime-2, Translate, and Whisper: new SOTA realtime voice APIs

OpenAI released GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper in the Realtime API, with GPT-Realtime-2 expanding context from 32K to 128K and scoring 96.6% on Artificial Analysis Big Bench Audio.

Why it matters: HKR-H/K/R all pass: an OpenAI real-time voice API refresh, a 32K→128K context jump, and a 96.6% Big Bench Audio claim. Score stays at 86 because this is a major API update, not a flagship foundation-model release.

AI HOT (Curated Pool)

Anthropic reportedly seeks tens of billions this summer, targeting a $1T valuation above OpenAI

Anthropic plans to raise up to $50B this summer, with a pre-money valuation near $900B. The round would put it near $1T, above OpenAI’s reported $852B valuation. Watch compute expansion and pre-IPO positioning.

Why it matters: HKR-H/K/R all pass: the hook is Anthropic nearing $1T and overtaking OpenAI, backed by $50B, $900B pre-money, and $852B figures. No closed-round confirmation, so this sits at the low end of 85–94.

QbitAI · WeChat

OpenAI releases three realtime voice models for reasoning, translation, and transcription

OpenAI launched GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper as API models, covering 128K-context voice reasoning, streaming translation from more than 70 input languages into 13 output languages, and realtime transcription priced at $0.017 per minute.

Why it matters: OpenAI shipped three realtime voice APIs across reasoning, translation, and transcription, hitting HKR-H/K/R. The 128K context, 70+ languages, and $0.017/min price make this a same-day must-write item.

Financial Times · Technology

Anthropic weighs deal for near $1tn valuation as revenue surges

Anthropic is fielding inbound investment offers that could value it near $1 trillion and surpass OpenAI, while the RSS snippet does not disclose revenue growth, deal size, investor names, or terms.

Why it matters: HKR-H/K/R all pass: the FT reports Anthropic weighing a deal near a $1tn valuation, potentially above OpenAI. The score stays low in the 85-94 band because revenue growth, funding size, and terms are not disclosed.

AI HOT (Curated Pool)

OpenAI launches official openai-cli for terminal API calls

OpenAI open-sourced openai-cli for direct API calls from the terminal. The Apache 2.0 tool installs via Homebrew or Go and covers Responses API, structured output, image editing, transcription, and key config. The key detail is Agent workflows using cloud tools like web search and code interpreter.

Why it matters: HKR-H/K/R all pass: official OpenAI terminal tooling is clickable, with concrete install/license/API details and workflow resonance. It is still a developer tooling update, not a model or major capability release, so 76 fits the featured threshold.

AI HOT (Curated Pool)

WIRED examines why ChatGPT keeps saying “I’ve got you” in Chinese replies

ChatGPT repeatedly uses phrases like “I’ll steadily catch you” in Chinese chats. WIRED links it to mode collapse, translation mismatch, and RLHF rewards for pleasing replies. Similar phrases appear in Claude and DeepSeek; the post does not disclose sample size.

Why it matters: HKR-H comes from the odd “I’ll catch you steadily” meme; HKR-K names three mechanisms; HKR-R touches alignment and Chinese UX concerns. No sample size is disclosed, so this stays in the lower featured band.

TechCrunch · AI

OpenAI introduces new 'Trusted Contact' safeguard for possible self-harm cases

OpenAI introduced Trusted Contact for ChatGPT self-harm risk cases. The post says it protects users when chats turn to self-harm, but does not disclose triggers, contact flow, or rollout scope. Watch false positives, privacy, and human review boundaries.

Why it matters: OpenAI’s ChatGPT safety update hits HKR-H/R via self-harm intervention and privacy stakes. HKR-K is weak: triggers, contact flow, and rollout are not disclosed, so this lands at the featured threshold.

AI HOT (Curated Pool)

Codex Plugin Now Supports Parallel Runs Across Chrome Tabs

OpenAI says Codex now runs in Chrome on macOS and Windows. The plugin works across tabs in the background without taking browser control; the post does not disclose version, concurrency limits, or enterprise policy.

Why it matters: HKR-H/K/R all pass, but the post gives platform and execution mechanics only; version, concurrency limits, and enterprise controls are not disclosed. Score: 76 as a practical OpenAI Codex product update.

The Verge · AI

Mira Murati’s Deposition Pulled Back the Curtain on Sam Altman’s Ouster

The Verge reports Mira Murati’s deposition on Sam Altman’s 2023 ouster from OpenAI before Thanksgiving. The material comes from Musk v. Altman exhibits and centers on the board’s claim that Altman was not consistently candid. The RSS snippet does not disclose full testimony details.

Why it matters: HKR-H/K/R all pass, but the core event is a 2023 ouster with a new deposition angle. The RSS summary lacks full testimony details, so this sits above the featured threshold, not in P1.

The Verge · AI

ChatGPT’s Trusted Contact will alert loved ones of safety concerns

OpenAI is launching optional Trusted Contact for ChatGPT, letting adult users assign one emergency contact. If self-harm or suicide topics are detected, OpenAI alerts the contact; the post does not disclose false-positive handling or regional rollout.

Why it matters: HKR-H/K/R all pass: OpenAI extends ChatGPT safety into human notification. The article lacks false-positive handling, rollout regions, and appeal flow, so it sits below model or core capability releases.

r/LocalLLaMA

WARNING: Open-OSS/privacy-filter Malware

A Reddit user says Hugging Face repo Open-OSS/privacy-filter is an infostealer. It mimics OpenAI's privacy filter, uses loader.py to fetch PowerShell, then downloads an EXE and runs it via Task Scheduler. The author says they reported it to Microsoft and Hugging Face; the post says Linux is unaffected.

Why it matters: HKR-H/K/R all pass: malware disguised as an OpenAI privacy filter has a concrete Windows execution chain. Single Reddit sourcing keeps it at the 72-77 featured threshold.