Skip to content

#OpenAI

37 today

Oct 1, 2024Tuesday

OpenAI News

Prompt Caching in the API

OpenAI added automatic prompt caching to GPT-4o, GPT-4o mini, o1-preview, and o1-mini API models, giving a 50% discount on recently reused input prefixes. Caching starts at 1,024 tokens and grows in 128-token increments; caches are often cleared after 5-10 minutes of inactivity and always within 1 hour of last use. The field to watch is cached_tokens in the API usage response.

Why it matters: A substantive OpenAI API update: not a new model, but it ships a 50% input discount, a 1,024-token threshold, 128-token cache steps, and cached_tokens telemetry, so HKR-H/K/R all pass. It is highly relevant to builder cost and latency, strong enough for featured, but not a same‑y

OpenAI News

Model Distillation in the API

OpenAI launched an API distillation workflow on October 1, 2024, letting developers use outputs from GPT-4o and o1-preview to fine-tune cheaper models such as GPT-4o mini. The suite includes Stored Completions, Evals in beta, and fine-tuning; setting store:true auto-saves input-output pairs with no added latency, per the post. Pricing includes 2M free GPT-4o mini training tokens per day and 1M for GPT-4o through October 31; Evals are free up to 7 runs per week through year-end if shared with OpenAI.

Sep 26, 2024Thursday

OpenAI News

Upgrading the Moderation API with OpenAI's new multimodal moderation model

OpenAI released omni-moderation-latest on September 26, 2024, a GPT-4o-based Moderation API model for text and image inputs that is free for all developers. It adds illicit and illicit/violent text categories, supports image moderation in 6 subcategories, and improves 42% on an internal 40-language eval, with gains in 98% of languages tested.

Why it matters: Official OpenAI developer product update with strong HKR-K: new moderation classes, image coverage, and a concrete +42% result across 40 languages. HKR-R also lands because moderation and compliance affect shipping teams directly; HKR-H is weak, so this sits at the low end of the

Sep 16, 2024Monday

OpenAI News

An update on OpenAI's safety and security practices

OpenAI said on September 16, 2024 that its Safety and Security Committee will become an independent board oversight committee, chaired by Zico Kolter, for critical safeguards in model development and deployment. The committee can review major model safety evaluations and delay launches until concerns are addressed; the post also cites a 90-day review, evaluation of an AI-sector ISAC, and work with Los Alamos National Laboratory.

Why it matters: HKR-H/K/R all pass. OpenAI says an independent board committee can review major safety evaluations and delay release, which is more concrete than a generic safety post. It stays below P1 because there is no new model, external audit result, or reproducible benchmark data.

Sep 12, 2024Thursday

OpenAI News

Introducing OpenAI o1

OpenAI released o1-preview and o1-mini on Sept. 12, 2024, with access for ChatGPT Plus, Team, and tier-5 API developers. The post cites 83% vs 13% on an IMO qualifier, 84 vs 22 on a jailbreak test, and says o1-mini is 80% cheaper than o1-preview. The tradeoff is clear: the API lacks function calling, streaming, and system messages, and the models do not yet support browsing or file and image uploads.

Why it matters: A major OpenAI reasoning-model launch with all three HKR signals: HKR-H from the new “think before answering” hook, HKR-K from concrete benchmark, safety, and pricing numbers, and HKR-R from the tradeoff practitioners must manage between stronger reasoning and missing API basics.

OpenAI News

Learning to reason with LLMs

OpenAI released o1-preview and reported 74% single-sample accuracy on AIME 2024, versus 12% for GPT-4o. The post says o1 reached the 89th percentile on Codeforces and exceeded human PhD experts on GPQA Diamond; it attributes this to large-scale RL and gains from both train-time and test-time compute. The key signal is scaling reasoning with compute, not just pretraining a larger base model.

Why it matters: This is a substantive OpenAI research release with product implications. HKR-H lands on the new reasoning line, HKR-K on the disclosed benchmark jumps and compute-scaling mechanism, and HKR-R on the direct impact to model strategy and inference economics; strong 90s, not 95+.

OpenAI News

OpenAI o1-mini

OpenAI released o1-mini on Sept. 12, 2024 for Tier 5 API users at 80% lower cost than o1-preview. The post reports 70.0% on AIME and 1650 Codeforces Elo, close to o1 at 74.4% and 1673, with about 3-5x faster answers than o1-preview in one word-reasoning example. The key tradeoff is explicit: it targets STEM reasoning, while non-STEM factual knowledge is only comparable to small models like GPT-4o mini.

Why it matters: OpenAI shipped a substantive model release, so this lands in the must-write band. HKR-H comes from the 80%-cheaper/nearly-o1 tradeoff; HKR-K from AIME 70.0 and Codeforces 1650; HKR-R from immediate developer cost/performance implications.

Aug 20, 2024Tuesday

OpenAI News

OpenAI partners with Condé Nast

OpenAI said on August 20, 2024 that it partnered with Condé Nast to surface content from brands such as Vogue, The New Yorker, and Wired in ChatGPT and the SearchGPT prototype. The post names at least nine Condé Nast brands and says SearchGPT links directly to source stories; it does not disclose deal value, licensing scope, revenue terms, or launch regions.

Why it matters: This is a meaningful OpenAI licensing/distribution move: ChatGPT and SearchGPT will show Condé Nast content with direct links. HKR-K and HKR-R pass, but HKR-H is limited because deal terms, rollout scope, and revenue share are undisclosed, so it lands at low-featured.

OpenAI News

Fine-tuning now available for GPT-4o

OpenAI has opened GPT-4o fine-tuning to developers on all paid tiers, with 1M free training tokens per org per day through September 23. Training costs $25 per 1M tokens, and inference costs $3.75 per 1M input tokens and $15 per 1M output tokens on gpt-4o-2024-08-06. The signal for practitioners: partners reported 43.8% on SWE-bench Verified and 71.83% on BIRD-SQL with fine-tuned GPT-4o.

Why it matters: This is a substantive OpenAI developer release with concrete details: temporary free training quota, train/inference prices, base model version, and two benchmark datapoints. HKR-H/K/R all pass, but this is an API capability expansion, not a new frontier-model launch or platform-

Aug 16, 2024Friday

OpenAI News

Disrupting a covert Iranian influence operation

OpenAI said it banned ChatGPT accounts tied to the Iranian influence operation Storm-2035 in August 2024 after they generated election and geopolitics content for X, Instagram, and five websites. The company identified 12 X accounts and one Instagram account; on Brookings' Breakout Scale, the operation ranked at the low end of Category 2, with most posts getting few or no likes, shares, or comments. What matters is the workflow: the models were used for long articles, comment rewrites, and English-Spanish posting, not for meaningful audience reach.

Why it matters: HKR-H lands on the covert election-influence angle; HKR-K lands on the account counts, sites, languages, and Breakout Scale 2. HKR-R lands via model-abuse governance, but the score stays at 76 because OpenAI reports no meaningful audience reach.

Aug 13, 2024Tuesday

OpenAI News

Introducing SWE-bench Verified

OpenAI released SWE-bench Verified, a human-validated subset built with the benchmark’s authors to assess real software issue resolution more reliably. The post names 3 failure modes in SWE-bench: overly narrow tests, underspecified issue statements, and unreliable environment setup; as of Aug. 5, 2024, top agents scored about 20% on SWE-bench and 43% on SWE-bench Lite. The key point is that the original benchmark can systematically underestimate coding-agent ability.

Why it matters: This is a strong benchmark release, not a routine post: OpenAI re-audited SWE-bench with the original authors, named 3 defect classes, and reported new score ceilings of 20% and 43%. HKR-H/K/R all pass because it changes how builders read code-agent leaderboards.

Aug 8, 2024Thursday

OpenAI News

Zico Kolter Joins OpenAI's Board of Directors

OpenAI appointed Carnegie Mellon professor Zico Kolter to its board on August 8, 2024, and added him to the Safety and Security Committee. The post says he will advise on critical safety and security decisions across all OpenAI projects alongside Bret Taylor, Sam Altman, and other members. The signal here is governance adding AI safety and robustness expertise, not a product launch.

Why it matters: The real signal is governance: OpenAI added a director with AI safety and robustness credentials and placed him on the Safety & Security Committee. HKR-K and HKR-R pass, but HKR-H is limited because this is a straightforward appointment notice, so it lands in low featured.

OpenAI News

GPT-4o System Card

OpenAI published the GPT-4o System Card on August 8, 2024, reporting 3 of 4 Preparedness categories as low risk and persuasion as borderline medium. The post says GPT-4o accepts text, audio, image, and video inputs, responds to audio in as little as 232 ms with a 320 ms average, and is 50% cheaper than GPT-4 Turbo in the API. The key issue for practitioners is voice safety: the card names unauthorized voice generation, speaker identification, and sensitive trait attribution, and says only models with post-mitigation scores at medium or below can be deployed.

Why it matters: This is not a routine post: it adds concrete preparedness ratings, 232ms voice latency, and a clear deployment threshold. HKR-H/K/R all pass, but it is a safety disclosure rather than a new model or major launch, so it lands as featured, not p1.

Aug 6, 2024Tuesday

OpenAI News

Introducing Structured Outputs in the API

OpenAI released Structured Outputs on Aug 6, 2024, making model outputs conform to developer-supplied JSON Schemas; `gpt-4o-2024-08-06` scored 100% on complex schema-following evals versus under 40% for `gpt-4-0613`. The feature is enabled with `strict: true` in function calling and works on tool-supporting models including `gpt-4-0613`, `gpt-3.5-turbo-0613`, and later. The key shift is constrained decoding plus schema training, not just valid JSON from JSON mode.

Why it matters: HKR-H/K/R all pass: OpenAI moves from 'valid JSON' to strict schema adherence and publishes a 100% vs <40% reliability gap. I keep it at 84 because this is a high-value API capability update, not a new frontier-model launch or company-level industry event.

Jul 25, 2024Thursday

OpenAI News

SearchGPT is a prototype of new AI search features

OpenAI began testing the SearchGPT prototype on July 25, 2024 with a small group of users and publishers. It answers with real-time web information, named inline citations, source links in a sidebar, and follow-up queries in shared context. The key detail is scope: this is a temporary prototype planned for future ChatGPT integration; the post does not disclose the model, rollout size, or commercial timeline.

Why it matters: Scored in the 85–94 band: OpenAI is testing a standalone AI-search prototype with live web answers and publisher participation, which is a same-day write for the industry. HKR-H/K/R all pass, but key rollout details, model identity, and commercialization timing are not disclosed.

Jul 24, 2024Wednesday

OpenAI News

Improving Model Safety Behavior with Rule-Based Rewards

OpenAI said on July 24, 2024 it uses Rule-Based Rewards in the RLHF pipeline to reduce repeated human feedback for safety alignment. The post defines three response types—hard refusal, soft refusal, and comply—and says the method has been part of OpenAI’s safety stack since GPT-4, including GPT-4o mini. The key point is maintainability when policies change; the post excerpt does not disclose quantitative gains.

Why it matters: HKR-H/K/R all pass: explicit rules inside RLHF is a strong hook, and the post adds three response modes plus paper/code. I keep it in the 78–84 band because the excerpt does not disclose effect sizes, baselines, or failure-case detail.

Jul 18, 2024Thursday

OpenAI News

GPT-4o mini: advancing cost-efficient intelligence

OpenAI released GPT-4o mini on July 18, 2024 at $0.15 per 1M input tokens and $0.60 per 1M output tokens, replacing GPT-3.5 in ChatGPT. It supports text and vision, offers a 128K context window and 16K max output, scores 82.0% on MMLU and 87.2% on HumanEval. The key detail for builders is that its API version is the first to use instruction hierarchy against jailbreaks and prompt injection.

Why it matters: This is a substantive OpenAI model launch, not a minor refresh: GPT-4o mini adds $0.15/$0.60 pricing, 128K context, 16K max output, benchmark details, and instruction hierarchy, then replaces GPT-3.5 in ChatGPT. HKR-H/K/R all pass, so it lands in P1.

OpenAI News

New compliance and administrative tools for ChatGPT Enterprise

OpenAI on July 18, 2024 launched an Enterprise Compliance API, eight third-party compliance integrations, and SCIM user management for ChatGPT Enterprise. The post confirms timestamped exports for conversations, files, GPT configs, memories, and users, plus support for Okta, Microsoft Entra ID, Google Workspace, and Ping; details on “Expanded GPT controls” are not disclosed in the provided body.

Jul 17, 2024Wednesday

OpenAI News

Prover-Verifier Games improve legibility of language model outputs

OpenAI trained GPT-4-family prover-verifier games so stronger models write solutions weaker models can verify; under time-limited human review, correctness-only optimization led to nearly 2x more evaluation errors. The post says the large and small models differ by about 3 orders of magnitude in pretraining compute, and checkability training recovers about half the performance gain of correctness-only optimization; the full experimental numbers are not fully disclosed in the provided text.

Why it matters: This is a substantive OpenAI research release with HKR-H/K/R all present: novel setup, clear mechanism, and strong relevance to scalable oversight. The excerpt confirms the method and the human-evaluation effect, but not the full experimental tables, so it fits the 78–84 band, نه

Jun 13, 2024Thursday

OpenAI News

OpenAI appoints Retired U.S. Army General Paul M. Nakasone to Board of Directors

OpenAI appointed retired U.S. Army General Paul M. Nakasone to its board, adding 1 director. The body is empty, so the post does not disclose the effective date, scope, or term. The signal here is governance, not a product update.

Why it matters: OpenAI's official post gives this board move real weight: retired general Paul M. Nakasone joins the board. HKR-H and HKR-R pass on the unusual security angle, but HKR-K fails because the post discloses little beyond the appointment, so it lands at the featured floor.

Jun 10, 2024Monday

OpenAI News

OpenAI and Apple announce partnership

OpenAI and Apple announced a partnership, and the title confirms only the two companies and the partnership action. The RSS item has no body, so scope, products, timeline, and commercial terms are not disclosed. This is not a product rollout yet; it is a partnership claim with missing details.

Why it matters: An official post confirms an Apple–OpenAI partnership, so HKR-H and HKR-R pass on entity weight and distribution stakes. I keep it at the low featured edge because HKR-K fails: the post gives no scope, mechanism, timeline, or commercial terms.

OpenAI News

OpenAI welcomes Sarah Friar (CFO) and Kevin Weil (CPO)

OpenAI says Sarah Friar will serve as CFO and Kevin Weil as CPO, adding 2 executives. Only the title is available; the post does not disclose start dates, scope, or reporting lines. The key signal is simultaneous hires across finance and product.

Why it matters: This is a strong official personnel signal from OpenAI: naming a CFO and CPO together gives it HKR-H and HKR-R, with top source authority. HKR-K is weak because the provided text confirms only names and titles; start dates, remit, and reporting lines are not disclosed, so it sits

May 28, 2024Tuesday

OpenAI News

OpenAI Board Forms Safety and Security Committee

OpenAI's board formed a Safety and Security Committee; that action is the only confirmed fact so far. The source provides only a title, and the post does not disclose members, authority, reporting lines, or timing. Watch governance power, not the committee name.

Why it matters: This is an official board-level OpenAI governance move with HKR-H and HKR-R. It stays in the low featured band because HKR-K is weak: the post confirms the committee exists, but gives no members, remit, reporting line, or effective date.

May 15, 2024Wednesday

OpenAI News

Ilya Sutskever to leave OpenAI, Jakub Pachocki announced as Chief Scientist

OpenAI announced that Ilya Sutskever will leave and Jakub Pachocki will become Chief Scientist. Only the title confirms these 2 personnel changes; the post body is empty and does not disclose timing, transition terms, or scope of responsibilities.

Why it matters: This is a 95+ personnel story: OpenAI's cofounder and chief scientist is leaving, which fits the policy's top band. HKR-H/K/R all pass, but the body does not disclose timing, scope, or transition details, so it stops short of a higher score.

May 13, 2024Monday

OpenAI News

Introducing GPT-4o and more tools to ChatGPT free users

OpenAI says it is bringing GPT-4o and more tools to ChatGPT free users, with free-tier access as the stated condition. Only the title is available; the post does not disclose tool list, usage limits, rollout timing, or regions.

Why it matters: This is a high-weight OpenAI product update, with HKR-H/K/R all passing: strong access hook, a concrete new availability fact, and clear competitive resonance. I stopped below the top of the band because the body is empty: tools, quotas, regions, and rollout conditions are notdis

Apr 24, 2024Wednesday

OpenAI News

GPT-4 API general availability and deprecation of older models in the Completions API

OpenAI says the GPT-4 API is generally available and older models in the Completions API will be deprecated. Only the title confirms these two facts; the post body is empty and does not disclose scope, timeline, or affected model names. The real issue to watch is migration cost: this is both an API and model transition.

Why it matters: This is a meaningful OpenAI platform update with strong HKR-K and HKR-R: GPT-4 API GA plus Completions deprecations affects developers immediately. It stays below p1 because the body is absent, so rollout scope, deadlines, and the affected model list are not disclosed.

Mar 8, 2024Friday

OpenAI News

Review completed; Altman and Brockman to continue to lead OpenAI

OpenAI says its review is complete, and Sam Altman and Greg Brockman will continue leading the company. Only the title is disclosed; the post does not disclose the review scope, evidence, or effective timeline. The key signal is leadership continuity, not strategy detail.

Why it matters: Official OpenAI governance news with strong HKR-H and HKR-R: it resolves the core suspense from the board crisis and matters to roadmap and partner trust. HKR-K is limited because the post discloses the outcome only; scope, evidence, and governance changes are not provided.

Feb 13, 2024Tuesday

OpenAI News

Memory and new controls for ChatGPT

OpenAI says ChatGPT is getting memory and new controls, with 2 changes disclosed in the title. The body is empty, so default state, opt-out scope, and user-tier availability are not disclosed. The key issue is control granularity; the title alone is not enough to judge product impact.

Why it matters: An official OpenAI post confirms ChatGPT memory plus new controls, so HKR-H and HKR-R pass on a core product readers already use. HKR-K fails because the body does not disclose defaults, rollout scope, user tiers, or control granularity, keeping this at the low featured edge.

Nov 29, 2023Wednesday

OpenAI News

Sam Altman returns as CEO, OpenAI has a new initial board

Sam Altman returns as OpenAI CEO, and OpenAI has a new initial board. The confirmed facts come only from the title; the post body is empty and does not disclose board size, member names, or timing. The real issue is governance change, but this item gives no verifiable detail.

Why it matters: This is a 95–100 band governance event: Sam Altman returns as CEO and OpenAI resets its initial board. HKR-H/K/R all pass on the reversal and industry impact, but the post does not disclose board members or timing, so it stays below a perfect score.

Sep 25, 2023Monday

OpenAI News

ChatGPT can now see, hear, and speak

OpenAI says ChatGPT now supports seeing, hearing, and speaking. The post body is empty, so it does not disclose model versions, rollout timing, regional limits, pricing, or API scope. The real watchpoints are voice latency, vision limits, and access paths.

Why it matters: This is a substantive OpenAI product update: the title confirms vision input, voice input, and speech output for ChatGPT, so HKR-H/K/R all pass. The copy provided here omits tiers, rollout scope, latency, and pricing, which keeps it at 88 rather than the top of the band.