Skip to content

OpenAI / ChatGPT

Everything OpenAI: the GPT models, ChatGPT and Sora, company strategy and people moves.

Latest picks

921–940 of 1,549

May 31Sunday

r/LocalLLaMA

PolyRange: Contamination-resistant offensive-AI benchmark for web targets

PolyRange v1.0 ships 84 WSTG-derived classes across 12 OWASP testing-guide categories. It generates fresh targets per deploy with a chosen LLM, adds two defense tiers, uses an agent-submits-flag oracle, and runs via a single-command CLI on Fly.io or Docker.

Why it matters: HKR-H/K/R all pass: PolyRange turns web-security targets into a dynamic agent benchmark with 84 WSTG classes and two defense levels. Single-source Reddit origin and security niche keep it at 78.

r/LocalLLaMA

Use any model and provider with the official OpenAI Codex Desktop App without modifying its code

Reddit user thibautrey describes a 3-step setup: edit Codex Desktop config.toml, store an API key, and use a multicodex proxy alias to map gpt-5.3-codex to MiniMax-Latest. The post lists a local base_url of 127.0.0.1:1455 and says the proxy disguises returned model names as gpt-5.3-codex.

Why it matters: This is a reproducible developer workflow trick, not an official release. HKR-H comes from the lock-in workaround, HKR-K has concrete config details, and HKR-R hits cost and model-choice pressure, placing it at the tutorial featured threshold.

Synced · WeChat

Microsoft open-sources SkillOpt for training Agent skill documents, reaching 3.3k stars in a week

Microsoft open-sourced SkillOpt, a text-space optimization framework that trains Agent skill documents without changing model weights; the paper reports best or tied-best results across 52 combinations covering 7 target models, 6 benchmarks, and 3 execution environments.

Why it matters: Microsoft’s open-source SkillOpt is a strong Agent tooling and research release. HKR-H has the 3.3k-star/trainable-skill hook, HKR-K has the text-parameter mechanism and 52 eval setups, and HKR-R hits agent engineering pain, so it lands in featured at 82.

May 30Saturday

AI HOT (Curated Pool)

OpenAI launches real-time translation model with 70+ input languages

OpenAI launched gpt-realtime-translate, a speech translation model that accepts 70+ input languages and outputs speech in 13 target languages; the post says the feature is running on smart glasses.

Why it matters: HKR-H/K/R all pass: OpenAI has a concrete realtime-translation model with numbers and a wearable demo. Missing latency, pricing, and API availability keep it below P1.

Bloomberg Technology

OpenAI Has Discussed Adding Citigroup, JPMorgan to Bank Lineup for IPO

The title says OpenAI discussed adding Citigroup and JPMorgan to its IPO bank lineup; the body only shows a Bloomberg 403 anti-bot page and does not disclose timing, valuation, mandate status, or the roles of the two banks.

Why it matters: HKR-H/K/R all pass, but the body is a Bloomberg 403 page; only the title gives the bank names, with no timeline, valuation, or role details. OpenAI IPO relevance is high, yet this is bank-lineup discussion, not a filing or priced deal.

AI HOT (Curated Pool)

Codex now supports computer use on Windows

OpenAI added Windows computer-use support for Codex, letting users start, review, and guide tasks on a Windows PC through the ChatGPT mobile app; the post states this is an early experience and does not disclose pricing or rollout scope.

Why it matters: HKR-H/K/R all pass: OpenAI adds Windows computer use to Codex, controlled through ChatGPT mobile. The post gives the workflow and early-stage condition, but not permissions, pricing, or rollout scope, so this stays at the featured threshold.

Bloomberg Technology

Anthropic Valuation of $965 Billion Passes OpenAI

Bloomberg’s title says Anthropic reached a $965 billion valuation and passed OpenAI; the post does not disclose the valuation source, funding round, deal terms, or the comparison basis for OpenAI.

Why it matters: HKR-H/K/R all pass: the $965B Anthropic-over-OpenAI claim is a major capital signal. The score stays below P1 because the article body is a video shell and does not disclose source, round, terms, or comparison basis.

Bloomberg Technology

Anthropic Raises at $965 Billion Valuation, Eclipsing OpenAI

Anthropic raised $65 billion at a $965 billion post-money valuation, surpassing OpenAI’s value for the first time; Altimeter Capital, Dragoneer, Greenoaks, and Sequoia Capital led the round, and each lead investor put in more than $2 billion.

Why it matters: Bloomberg reports Anthropic raising $65B at a $965B valuation with named lead investors; if closed, this is a top-tier foundation-model capital event. HKR-H, HKR-K, and HKR-R all pass, placing it in the 95+ band.

May 29Friday

New York Times Chinese

Anthropic Tops OpenAI Valuation to Become the Most Valuable AI Startup

Anthropic raised $65 billion at a $900 billion pre-money valuation, above OpenAI’s last $730 billion valuation. Claude Opus 4.8 also scored 10% higher than Anthropic’s previous model on Vals AI’s vibe-coding benchmark.

Why it matters: Anthropic topping OpenAI with $65B financing and a $900B pre-money valuation is a foundation-model market-structure event. HKR-H/K/R all pass, with NYT source authority supporting p1.

Xinzhiyuan · WeChat

Claude Opus 4.8 tests split users: strong at high effort, costly under rate limits

The article says Claude Opus 4.8 scores 63 on an Extra-High senior engineering benchmark, 30 points above Opus 4.7, but drops to 42 at High effort, while $200/month Max users report hitting rate limits within hours on complex agent tasks.

Why it matters: Anthropic/Claude relevance plus concrete test numbers clears HKR-H/K/R: the hook is strength versus cost, K has benchmark and quota details, and R hits agent-budget anxiety. Source is a media test rather than an official release, so this lands at low P1.

AI HOT (Curated Pool)

Strengthening Societal Resilience with Rosalind Biodefense

OpenAI launched Rosalind Biodefense and provides trusted GPT-Rosalind access to vetted developers and U.S. government partners; the post does not disclose model parameters or pricing.

Why it matters: HKR-H/K/R all pass: OpenAI launched GPT-Rosalind access for vetted developers and US government partners. Missing parameters, pricing, and eval results keep it below a major capability release.

Latent Space

Anthropic raises $65B Series H, releases Opus 4.8 and Dynamic Workflows

Anthropic announced a $65B Series H at a $965B post-money valuation, disclosed a $47B revenue run rate, and released Claude Opus 4.8 plus Claude Code Dynamic Workflows as a research preview for parallel subagent orchestration.

Why it matters: HKR-H/K/R all pass: this combines a frontier-lab financing event with an Anthropic model and Claude Code workflow release. I score using the summary’s $65B raise and $965B post-money valuation because the title’s dollar figure conflicts with it.

Ruan YiFeng's Weblog

Technology Enthusiasts Weekly Issue 398: Token Costs Are Hard to Afford

Peter Steinberger posted one month of usage showing 7.6 million requests and 603 billion tokens, with CodexBar estimating a $1.3 million value under preset rates rather than his actual spend as an OpenAI employee.

Why it matters: HKR-H/K/R all pass: the CodexBar case turns token economics into concrete usage and cost. This is strong practitioner commentary, not a model or platform release, so it fits the 72–77 featured band.

AI HOT (Curated Pool)

Skill distillation

Skill distillation has Opus 4.7, GPT-5.1, and Gemini 3 Pro write standardized SKILL.md procedure files, while local Qwen 35B and Gemma 26B models execute those files step by step.

Why it matters: HKR-H/K/R pass: the agent-skill distillation pattern is concrete and practitioner-relevant. The summary lacks success rates, cost data, or task outcomes, so it sits at the featured threshold, not must-write.

Financial Times · Technology

Anthropic finalises $65bn funding deal to surpass OpenAI’s valuation

Anthropic finalised a $65bn funding deal that values the Claude AI maker at $965bn including the new money, taking its valuation above OpenAI’s, according to the RSS snippet.

Why it matters: HKR-H/K/R all pass: FT reports Anthropic finalized a $65bn funding deal at a $965bn valuation, overtaking OpenAI. That is a top-model-lab capital reset, above the must-write band.

Bloomberg Technology

Anthropic Raises $65 Billion in Funding Round, Eclipses OpenAI

Anthropic raised $65 billion at a $965 billion post-money valuation, and Bloomberg says the round put its value above OpenAI for the first time.

Why it matters: HKR-H/K/R all pass: $65B raised and a $965B valuation would reorder frontier-lab capital rankings. The item stays below 95 because investor names and terms are not disclosed.

Bloomberg Technology

Anthropic Eclipses OpenAI With Valuation of $965 Billion

Anthropic raised $65 billion at a $965 billion post-money valuation, surpassing OpenAI’s valuation for the first time; the RSS snippet does not disclose investors, deal terms, or a timeline.

Why it matters: HKR-H/K/R all pass: Bloomberg reports a $65B raise at a $965B valuation, putting Anthropic above OpenAI. Investors and terms are not disclosed, but the scale makes it same-day top AI funding news.

May 28Thursday

AI HOT (Curated Pool)

OpenAI Frontier Governance Framework

OpenAI published its Frontier Governance Framework to align its AI safety, security, and risk management practices with new EU and California regulations; the post does not disclose specific evaluation metrics, implementation timelines, or the list of covered frontier models.

Why it matters: HKR-K/R pass: an official OpenAI frontier-governance framework carries safety and regulatory signal, but metrics, timeline, and covered models are not disclosed, so it stays at the lower featured band.

AI HOT (Curated Pool)

Cognition AI raises over $1B, targets 10x software engineering productivity

Cognition AI raised over $1 billion at a $26 billion pre-money valuation, while annualized revenue grew from $37 million to about $492 million in one year, and Devin is positioned as an autonomous junior engineer that can plan, test, and deploy through multi-step workflows.

Why it matters: HKR-H/K/R all pass: the story has hard numbers on funding, valuation, and ARR, plus a direct junior-engineer automation angle. Single-post sourcing keeps it below the 95+ industry-shaking band.

AI HOT (Curated Pool)

OpenAI Products Support Secure Connections to Private MCP Servers

OpenAI supports ChatGPT, Codex, and the Responses API connecting to internal MCP servers through outbound-only HTTPS, while teams keep those servers inside private networks.

Why it matters: HKR-H/K/R pass: OpenAI adds private MCP server support with outbound-only HTTPS, a concrete enterprise agent integration mechanism. Missing permission model, pricing, and rollout details keep it in the lower featured band.