Skip to content

Product updates

New features, redesigns and pricing in AI products — whose product got better, pricier or finally usable.

841 picksRelated topicsModel releasesIndustryAI coding

Latest picks

821–840 of 841

Jan 23, 2025Thursday

OpenAI News

Introducing Operator

OpenAI released Operator on Jan 23, 2025 as a research preview for U.S. Pro users; it uses its own browser to click, type, and scroll through web tasks. It runs on Computer-Using Agent, combining GPT-4o vision with RL-based reasoning; the post says it sets SOTA on WebArena and WebVoyager but does not disclose scores. The key boundary is control: login, payment, and CAPTCHA flows hand control back to users, and a July 17 update says it was folded into ChatGPT agent.

Why it matters: OpenAI's Operator is a same-day, must-write product release: a browser-using agent moves ChatGPT from answering to acting. HKR-H/K/R all pass; the post gives the own-browser setup, GPT-4o+RL, and user handoff for login/payments, but US Pro limits and missing benchmark scores keep

Dec 17, 2024Tuesday

OpenAI News

OpenAI o1 and new tools for developers

OpenAI released o1 in the API, updated the Realtime API, added Preference Fine-Tuning, and shipped beta Go/Java SDKs; o1 is rolling out first to usage tier 5 developers. Disclosed details include 60% fewer reasoning tokens than o1-preview on average, and a 60% GPT-4o audio price cut in Realtime API to $40/1M input and $80/1M output tokens. The key shift is production support for function calling, Structured Outputs, developer messages, vision, and a reasoning_effort parameter; the post is truncated, so some GPT-4o mini realtime pricing details are not disclosed here.

Why it matters: This is a substantive OpenAI developer release: o1 reaches the API with function calling, Structured Outputs, vision, and developer messages, which materially improves production readiness. HKR-H/K/R all pass; the excerpt includes concrete token and pricing data, but later GPT-4o

Dec 9, 2024Monday

OpenAI News

Sora is here

OpenAI moved Sora out of research preview on December 9, 2024 and rolled it out to ChatGPT Plus and Pro users. Sora Turbo supports up to 1080p and 20-second videos; Plus includes up to 50 monthly 480p videos or fewer 720p generations. The key detail for practitioners is deployment scope: the UK, Switzerland, and the EEA are excluded, person uploads are limited, and OpenAI says physics and long complex actions remain weak.

Why it matters: OpenAI moved Sora from preview to paid availability, so HKR-H/K/R all pass: high-curiosity launch, concrete specs and limits, and clear impact on creator workflows. I stop below 95 because the post itself notes region blocks, restrictions on uploads with people, and instabilityon

Dec 5, 2024Thursday

OpenAI News

Introducing ChatGPT Pro

OpenAI launched ChatGPT Pro at $200 per month, with unlimited access to OpenAI o1, o1-mini, GPT-4o, Advanced Voice, and a higher-compute o1 pro mode. The post specifies a stricter 4/4 reliability metric, where a question counts only if the model answers correctly in all four attempts, but it does not disclose concrete quotas or latency figures. The key signal is compute tiering: longer reasoning time is now a paid product feature.

Oct 31, 2024Thursday

OpenAI News

Introducing ChatGPT search

OpenAI launched ChatGPT search on Oct. 31, 2024 for Plus, Team, and SearchGPT waitlist users, adding web answers with source links inside ChatGPT. It can trigger web search automatically or manually, shows a Sources sidebar, and uses a fine-tuned GPT-4o post-trained with distilled outputs from o1-preview. The shift to watch is distribution: search is folded into chat, not a separate search engine hop.

Why it matters: This is a same-day OpenAI product launch, not a minor feature tweak; search is merged into the chat UI, so HKR-H/K/R all pass. The post confirms source-linked web answers and launch conditions, and the move hits search distribution directly, which pushes it to P1.

Oct 3, 2024Thursday

OpenAI News

Introducing canvas, a new way to write and code with ChatGPT

OpenAI launched the canvas beta on October 3, 2024 for ChatGPT Plus and Team users, adding a GPT-4o-based workspace for writing and coding beyond chat. The post says canvas can auto-trigger or open via “use canvas,” supports targeted edits, version restore, and shortcuts like code review and bug fixing. The key signal is model training: across 20+ internal evals, trigger accuracy reached 83% for writing and 94% for coding, targeted edits beat baseline by 18%, and comment accuracy and quality improved by 30% and 16%.

Oct 1, 2024Tuesday

OpenAI News

Introducing the Realtime API

OpenAI launched a public beta of the Realtime API on Oct. 1, 2024 for all paid developers, using a persistent WebSocket to stream low-latency speech-to-speech interactions with GPT-4o. It supports function calling and interruption handling, priced at $5/1M text input tokens and $100/1M audio input tokens; the post also says audio I/O for Chat Completions would arrive in the following weeks.

Why it matters: OpenAI moved voice apps from stitched ASR+TTS calls to a persistent GPT-4o session, with function calling, interruption handling, and published audio/token pricing. HKR-H/K/R all pass, so this is a same-day must-write developer platform update and clears p1.

OpenAI News

Introducing vision to the fine-tuning API

OpenAI launched GPT-4o vision fine-tuning on Oct 1, 2024, letting paid-tier developers train with images plus text, starting from as few as 100 images. The post cites Grab improving lane-count accuracy by 20% and speed-limit sign localization by 13%, while Automat raised RPA success from 16.60% to 61.67%. The notable shift is multimodal customization in the main API; the pricing section is truncated, so full price details are not disclosed.

Why it matters: OpenAI shipped a substantive API update: GPT-4o vision fine-tuning with a 100-image floor and named gains from Grab and Automat, so HKR-H/K/R all pass. Scope is strong for builders, but the blast radius is narrower than a flagship model launch, and pricing is incomplete in the ex

OpenAI News

Prompt Caching in the API

OpenAI added automatic prompt caching to GPT-4o, GPT-4o mini, o1-preview, and o1-mini API models, giving a 50% discount on recently reused input prefixes. Caching starts at 1,024 tokens and grows in 128-token increments; caches are often cleared after 5-10 minutes of inactivity and always within 1 hour of last use. The field to watch is cached_tokens in the API usage response.

Why it matters: A substantive OpenAI API update: not a new model, but it ships a 50% input discount, a 1,024-token threshold, 128-token cache steps, and cached_tokens telemetry, so HKR-H/K/R all pass. It is highly relevant to builder cost and latency, strong enough for featured, but not a same‑y

Sep 26, 2024Thursday

OpenAI News

Upgrading the Moderation API with OpenAI's new multimodal moderation model

OpenAI released omni-moderation-latest on September 26, 2024, a GPT-4o-based Moderation API model for text and image inputs that is free for all developers. It adds illicit and illicit/violent text categories, supports image moderation in 6 subcategories, and improves 42% on an internal 40-language eval, with gains in 98% of languages tested.

Why it matters: Official OpenAI developer product update with strong HKR-K: new moderation classes, image coverage, and a concrete +42% result across 40 languages. HKR-R also lands because moderation and compliance affect shipping teams directly; HKR-H is weak, so this sits at the low end of the

Sep 12, 2024Thursday

OpenAI News

OpenAI o1-mini

OpenAI released o1-mini on Sept. 12, 2024 for Tier 5 API users at 80% lower cost than o1-preview. The post reports 70.0% on AIME and 1650 Codeforces Elo, close to o1 at 74.4% and 1673, with about 3-5x faster answers than o1-preview in one word-reasoning example. The key tradeoff is explicit: it targets STEM reasoning, while non-STEM factual knowledge is only comparable to small models like GPT-4o mini.

Why it matters: OpenAI shipped a substantive model release, so this lands in the must-write band. HKR-H comes from the 80%-cheaper/nearly-o1 tradeoff; HKR-K from AIME 70.0 and Codeforces 1650; HKR-R from immediate developer cost/performance implications.

Aug 20, 2024Tuesday

OpenAI News

Fine-tuning now available for GPT-4o

OpenAI has opened GPT-4o fine-tuning to developers on all paid tiers, with 1M free training tokens per org per day through September 23. Training costs $25 per 1M tokens, and inference costs $3.75 per 1M input tokens and $15 per 1M output tokens on gpt-4o-2024-08-06. The signal for practitioners: partners reported 43.8% on SWE-bench Verified and 71.83% on BIRD-SQL with fine-tuned GPT-4o.

Why it matters: This is a substantive OpenAI developer release with concrete details: temporary free training quota, train/inference prices, base model version, and two benchmark datapoints. HKR-H/K/R all pass, but this is an API capability expansion, not a new frontier-model launch or platform-

Aug 6, 2024Tuesday

OpenAI News

Introducing Structured Outputs in the API

OpenAI released Structured Outputs on Aug 6, 2024, making model outputs conform to developer-supplied JSON Schemas; `gpt-4o-2024-08-06` scored 100% on complex schema-following evals versus under 40% for `gpt-4-0613`. The feature is enabled with `strict: true` in function calling and works on tool-supporting models including `gpt-4-0613`, `gpt-3.5-turbo-0613`, and later. The key shift is constrained decoding plus schema training, not just valid JSON from JSON mode.

Why it matters: HKR-H/K/R all pass: OpenAI moves from 'valid JSON' to strict schema adherence and publishes a 100% vs <40% reliability gap. I keep it at 84 because this is a high-value API capability update, not a new frontier-model launch or company-level industry event.

Jul 25, 2024Thursday

OpenAI News

SearchGPT is a prototype of new AI search features

OpenAI began testing the SearchGPT prototype on July 25, 2024 with a small group of users and publishers. It answers with real-time web information, named inline citations, source links in a sidebar, and follow-up queries in shared context. The key detail is scope: this is a temporary prototype planned for future ChatGPT integration; the post does not disclose the model, rollout size, or commercial timeline.

Why it matters: Scored in the 85–94 band: OpenAI is testing a standalone AI-search prototype with live web answers and publisher participation, which is a same-day write for the industry. HKR-H/K/R all pass, but key rollout details, model identity, and commercialization timing are not disclosed.

Jul 23, 2024Tuesday

Hugging Face Blog

Llama 3.1: 405B, 70B & 8B with multilinguality and long context

Meta released Llama 3.1 with 405B, 70B, and 8B sizes, and the title says it adds multilingual support and long context. Only the title is available; the post does not disclose context length, languages, license terms, or benchmark results. Watch the 405B release terms and real inference cost.

Why it matters: Meta's Llama 3.1 is a major flagship open-model release, and the title already gives concrete sizes plus multilingual and long-context positioning. HKR-H/K/R all pass; missing license, exact context window, and benchmark detail keep it at the low end of the 85-94 band.

Jul 18, 2024Thursday

OpenAI News

GPT-4o mini: advancing cost-efficient intelligence

OpenAI released GPT-4o mini on July 18, 2024 at $0.15 per 1M input tokens and $0.60 per 1M output tokens, replacing GPT-3.5 in ChatGPT. It supports text and vision, offers a 128K context window and 16K max output, scores 82.0% on MMLU and 87.2% on HumanEval. The key detail for builders is that its API version is the first to use instruction hierarchy against jailbreaks and prompt injection.

Why it matters: This is a substantive OpenAI model launch, not a minor refresh: GPT-4o mini adds $0.15/$0.60 pricing, 128K context, 16K max output, benchmark details, and instruction hierarchy, then replaces GPT-3.5 in ChatGPT. HKR-H/K/R all pass, so it lands in P1.

OpenAI News

New compliance and administrative tools for ChatGPT Enterprise

OpenAI on July 18, 2024 launched an Enterprise Compliance API, eight third-party compliance integrations, and SCIM user management for ChatGPT Enterprise. The post confirms timestamped exports for conversations, files, GPT configs, memories, and users, plus support for Okta, Microsoft Entra ID, Google Workspace, and Ping; details on “Expanded GPT controls” are not disclosed in the provided body.

May 13, 2024Monday

OpenAI News

Introducing GPT-4o and more tools to ChatGPT free users

OpenAI says it is bringing GPT-4o and more tools to ChatGPT free users, with free-tier access as the stated condition. Only the title is available; the post does not disclose tool list, usage limits, rollout timing, or regions.

Why it matters: This is a high-weight OpenAI product update, with HKR-H/K/R all passing: strong access hook, a concrete new availability fact, and clear competitive resonance. I stopped below the top of the band because the body is empty: tools, quotas, regions, and rollout conditions are notdis

Apr 24, 2024Wednesday

OpenAI News

GPT-4 API general availability and deprecation of older models in the Completions API

OpenAI says the GPT-4 API is generally available and older models in the Completions API will be deprecated. Only the title confirms these two facts; the post body is empty and does not disclose scope, timeline, or affected model names. The real issue to watch is migration cost: this is both an API and model transition.

Why it matters: This is a meaningful OpenAI platform update with strong HKR-K and HKR-R: GPT-4 API GA plus Completions deprecations affects developers immediately. It stays below p1 because the body is absent, so rollout scope, deadlines, and the affected model list are not disclosed.

Feb 13, 2024Tuesday

OpenAI News

Memory and new controls for ChatGPT

OpenAI says ChatGPT is getting memory and new controls, with 2 changes disclosed in the title. The body is empty, so default state, opt-out scope, and user-tier availability are not disclosed. The key issue is control granularity; the title alone is not enough to judge product impact.

Why it matters: An official OpenAI post confirms ChatGPT memory plus new controls, so HKR-H and HKR-R pass on a core product readers already use. HKR-K fails because the body does not disclose defaults, rollout scope, user tiers, or control granularity, keeping this at the low featured edge.