Skip to content

#产品更新

0 today

Oct 1, 2024Tuesday

OpenAI News

Prompt Caching in the API

OpenAI added automatic prompt caching to GPT-4o, GPT-4o mini, o1-preview, and o1-mini API models, giving a 50% discount on recently reused input prefixes. Caching starts at 1,024 tokens and grows in 128-token increments; caches are often cleared after 5-10 minutes of inactivity and always within 1 hour of last use. The field to watch is cached_tokens in the API usage response.

Why it matters: A substantive OpenAI API update: not a new model, but it ships a 50% input discount, a 1,024-token threshold, 128-token cache steps, and cached_tokens telemetry, so HKR-H/K/R all pass. It is highly relevant to builder cost and latency, strong enough for featured, but not a same‑y

Sep 26, 2024Thursday

OpenAI News

Upgrading the Moderation API with OpenAI's new multimodal moderation model

OpenAI released omni-moderation-latest on September 26, 2024, a GPT-4o-based Moderation API model for text and image inputs that is free for all developers. It adds illicit and illicit/violent text categories, supports image moderation in 6 subcategories, and improves 42% on an internal 40-language eval, with gains in 98% of languages tested.

Why it matters: Official OpenAI developer product update with strong HKR-K: new moderation classes, image coverage, and a concrete +42% result across 40 languages. HKR-R also lands because moderation and compliance affect shipping teams directly; HKR-H is weak, so this sits at the low end of the

Sep 12, 2024Thursday

OpenAI News

OpenAI o1-mini

OpenAI released o1-mini on Sept. 12, 2024 for Tier 5 API users at 80% lower cost than o1-preview. The post reports 70.0% on AIME and 1650 Codeforces Elo, close to o1 at 74.4% and 1673, with about 3-5x faster answers than o1-preview in one word-reasoning example. The key tradeoff is explicit: it targets STEM reasoning, while non-STEM factual knowledge is only comparable to small models like GPT-4o mini.

Why it matters: OpenAI shipped a substantive model release, so this lands in the must-write band. HKR-H comes from the 80%-cheaper/nearly-o1 tradeoff; HKR-K from AIME 70.0 and Codeforces 1650; HKR-R from immediate developer cost/performance implications.

Aug 20, 2024Tuesday

OpenAI News

Fine-tuning now available for GPT-4o

OpenAI has opened GPT-4o fine-tuning to developers on all paid tiers, with 1M free training tokens per org per day through September 23. Training costs $25 per 1M tokens, and inference costs $3.75 per 1M input tokens and $15 per 1M output tokens on gpt-4o-2024-08-06. The signal for practitioners: partners reported 43.8% on SWE-bench Verified and 71.83% on BIRD-SQL with fine-tuned GPT-4o.

Why it matters: This is a substantive OpenAI developer release with concrete details: temporary free training quota, train/inference prices, base model version, and two benchmark datapoints. HKR-H/K/R all pass, but this is an API capability expansion, not a new frontier-model launch or platform-

Aug 6, 2024Tuesday

OpenAI News

Introducing Structured Outputs in the API

OpenAI released Structured Outputs on Aug 6, 2024, making model outputs conform to developer-supplied JSON Schemas; `gpt-4o-2024-08-06` scored 100% on complex schema-following evals versus under 40% for `gpt-4-0613`. The feature is enabled with `strict: true` in function calling and works on tool-supporting models including `gpt-4-0613`, `gpt-3.5-turbo-0613`, and later. The key shift is constrained decoding plus schema training, not just valid JSON from JSON mode.

Why it matters: HKR-H/K/R all pass: OpenAI moves from 'valid JSON' to strict schema adherence and publishes a 100% vs <40% reliability gap. I keep it at 84 because this is a high-value API capability update, not a new frontier-model launch or company-level industry event.

Jul 25, 2024Thursday

OpenAI News

SearchGPT is a prototype of new AI search features

OpenAI began testing the SearchGPT prototype on July 25, 2024 with a small group of users and publishers. It answers with real-time web information, named inline citations, source links in a sidebar, and follow-up queries in shared context. The key detail is scope: this is a temporary prototype planned for future ChatGPT integration; the post does not disclose the model, rollout size, or commercial timeline.

Why it matters: Scored in the 85–94 band: OpenAI is testing a standalone AI-search prototype with live web answers and publisher participation, which is a same-day write for the industry. HKR-H/K/R all pass, but key rollout details, model identity, and commercialization timing are not disclosed.

Jul 23, 2024Tuesday

Hugging Face Blog

Llama 3.1: 405B, 70B & 8B with multilinguality and long context

Meta released Llama 3.1 with 405B, 70B, and 8B sizes, and the title says it adds multilingual support and long context. Only the title is available; the post does not disclose context length, languages, license terms, or benchmark results. Watch the 405B release terms and real inference cost.

Why it matters: Meta's Llama 3.1 is a major flagship open-model release, and the title already gives concrete sizes plus multilingual and long-context positioning. HKR-H/K/R all pass; missing license, exact context window, and benchmark detail keep it at the low end of the 85-94 band.

Jul 18, 2024Thursday

OpenAI News

GPT-4o mini: advancing cost-efficient intelligence

OpenAI released GPT-4o mini on July 18, 2024 at $0.15 per 1M input tokens and $0.60 per 1M output tokens, replacing GPT-3.5 in ChatGPT. It supports text and vision, offers a 128K context window and 16K max output, scores 82.0% on MMLU and 87.2% on HumanEval. The key detail for builders is that its API version is the first to use instruction hierarchy against jailbreaks and prompt injection.

Why it matters: This is a substantive OpenAI model launch, not a minor refresh: GPT-4o mini adds $0.15/$0.60 pricing, 128K context, 16K max output, benchmark details, and instruction hierarchy, then replaces GPT-3.5 in ChatGPT. HKR-H/K/R all pass, so it lands in P1.

OpenAI News

New compliance and administrative tools for ChatGPT Enterprise

OpenAI on July 18, 2024 launched an Enterprise Compliance API, eight third-party compliance integrations, and SCIM user management for ChatGPT Enterprise. The post confirms timestamped exports for conversations, files, GPT configs, memories, and users, plus support for Okta, Microsoft Entra ID, Google Workspace, and Ping; details on “Expanded GPT controls” are not disclosed in the provided body.

May 13, 2024Monday

OpenAI News

Introducing GPT-4o and more tools to ChatGPT free users

OpenAI says it is bringing GPT-4o and more tools to ChatGPT free users, with free-tier access as the stated condition. Only the title is available; the post does not disclose tool list, usage limits, rollout timing, or regions.

Why it matters: This is a high-weight OpenAI product update, with HKR-H/K/R all passing: strong access hook, a concrete new availability fact, and clear competitive resonance. I stopped below the top of the band because the body is empty: tools, quotas, regions, and rollout conditions are notdis

Apr 24, 2024Wednesday

OpenAI News

GPT-4 API general availability and deprecation of older models in the Completions API

OpenAI says the GPT-4 API is generally available and older models in the Completions API will be deprecated. Only the title confirms these two facts; the post body is empty and does not disclose scope, timeline, or affected model names. The real issue to watch is migration cost: this is both an API and model transition.

Why it matters: This is a meaningful OpenAI platform update with strong HKR-K and HKR-R: GPT-4 API GA plus Completions deprecations affects developers immediately. It stays below p1 because the body is absent, so rollout scope, deadlines, and the affected model list are not disclosed.

Feb 13, 2024Tuesday

OpenAI News

Memory and new controls for ChatGPT

OpenAI says ChatGPT is getting memory and new controls, with 2 changes disclosed in the title. The body is empty, so default state, opt-out scope, and user-tier availability are not disclosed. The key issue is control granularity; the title alone is not enough to judge product impact.

Why it matters: An official OpenAI post confirms ChatGPT memory plus new controls, so HKR-H and HKR-R pass on a core product readers already use. HKR-K fails because the body does not disclose defaults, rollout scope, user tiers, or control granularity, keeping this at the low featured edge.

Sep 25, 2023Monday

OpenAI News

ChatGPT can now see, hear, and speak

OpenAI says ChatGPT now supports seeing, hearing, and speaking. The post body is empty, so it does not disclose model versions, rollout timing, regional limits, pricing, or API scope. The real watchpoints are voice latency, vision limits, and access paths.

Why it matters: This is a substantive OpenAI product update: the title confirms vision input, voice input, and speech output for ChatGPT, so HKR-H/K/R all pass. The copy provided here omits tiers, rollout scope, latency, and pricing, which keeps it at 88 rather than the top of the band.