Skip to content

#MCP/工具调用

0 today

Apr 3Friday

X · @claudeai

Microsoft 365 connectors are now available on every Claude plan

Anthropic made Microsoft 365 connectors available on every Claude plan, covering Outlook, OneDrive, and SharePoint. The post confirms plan coverage and supported apps; it does not disclose pricing, permission boundaries, regional limits, or admin requirements. The real signal is broad rollout across all plans, not a new standalone connector.

Why it matters: This is a mid-weight Claude product update: Anthropic expanded Microsoft 365 connectors to every Claude plan, which changes real Outlook, OneDrive, and SharePoint access. HKR-H/K/R all pass, but missing price, permission, region, and admin details keeps it at low-end featured.

X · @claudeai

Computer use in Claude Cowork and Claude Code Desktop is now available on Windows

Claude has brought computer use in Claude Cowork and Claude Code Desktop to Windows. The post confirms the Windows rollout, but does not disclose supported versions, permission model, latency, pricing, or release timing. What matters is the reliability boundary for desktop agents on Windows, and the post gives no reproducible conditions yet.

Why it matters: HKR-H lands on the Windows rollout hook, and HKR-R lands because desktop agents on Windows map to real workflows. Score stays at 74: this is an official Claude update, but the post confirms availability only; versions, permissions, latency, and price are not disclosed.

X · @OpenAI

ChatGPT is now available in CarPlay

OpenAI is rolling out ChatGPT in CarPlay to iPhone users on iOS 26.4+ where CarPlay is supported. The post confirms voice mode is available in-car, but does not disclose regions, vehicle coverage, or feature limits. The key shift is distribution into the driving interface, not a new model launch.

Why it matters: This matters more as a distribution-surface shift than a model update. HKR-H and HKR-R pass on the CarPlay hook and assistant-entry competition; HKR-K stays limited because the post gives iOS 26.4+ rollout only, not regions, car support, or full feature bounds.

Mar 31Tuesday

Mistral AI

Spaces: A CLI Built for Humans and Agents

Mistral AI 发布 Spaces CLI,同时面向人类开发者与编码智能体。它通过 `spaces init`、`spaces dev` 等命令快速搭建多服务项目,并为每个交互式提示提供对应的 flag 与 `-y` 选项,使智能体可自主完成配置与部署。每次 init 还会生成 context.json 和 AGENTS.md,为智能体提供项目上下文与操作规则。

Mar 17Tuesday

OpenAI News

Introducing GPT-5.4 mini and nano

OpenAI released GPT-5.4 mini and nano on March 17, 2026 for coding and subagents; mini runs over 2x faster than GPT-5 mini. In the API, mini has a 400k context window and costs $0.75/$4.50 per 1M input/output tokens, while nano is API-only at $0.20/$1.25. The key signal is performance per latency: mini scores 54.4% on SWE-Bench Pro versus GPT-5.4 at 57.7%.

Why it matters: This is an official OpenAI model launch, not a routine patch. It includes concrete numbers—>2x speed, 400k context, API pricing, and 54.4% vs 57.7% on SWE-Bench Pro—so HKR-H/K/R all pass; scored at the low end of the 85–94 band.

Mar 11Wednesday

OpenAI News

From model to agent: Equipping the Responses API with a computer environment

OpenAI said on March 11, 2026 that Responses API now works with a shell tool and hosted container workspace, so models can execute commands in an isolated loop. The post says GPT-5.2 and later are trained to propose shell commands, while the API streams outputs and can run multiple commands concurrently across sessions; the container includes a filesystem, optional SQLite, and restricted network access. The key change is orchestration, not the “agent” label; pricing, quotas, and full security details are not disclosed in the visible post.

Why it matters: Substantive OpenAI developer update: the Responses API moves from tool calls to a managed computer environment with shell execution, streaming, parallel runs, and context compaction, so HKR-H/K/R all pass. The post is truncated and omits pricing, quotas, and full safety details,【

Mar 10Tuesday

NVIDIA Blog

NVIDIA and Thinking Machines Lab Announce Long-Term Gigawatt-Scale Strategic Partnership

NVIDIA and Thinking Machines Lab formed a multiyear deal to deploy at least 1 gigawatt of NVIDIA Vera Rubin systems, targeted for early next year, for frontier model training and customizable AI platforms. The partnership also covers training and serving system design for NVIDIA architectures and broader access to frontier and open models for enterprises and researchers; the post does not disclose the investment size. The key signal is the explicit 1-gigawatt compute commitment, not a routine cloud purchase.

Why it matters: The 1GW Vera Rubin commitment lifts this above routine partnership PR: HKR-H on scale, HKR-K on a named system with a dated deployment target, and HKR-R on frontier compute competition. It stays below P1 because the source is a vendor blog and key details—spend, ownership, and ph

OpenAI News

New ways to learn math and science in ChatGPT

OpenAI launched interactive math and science visualizations in ChatGPT on March 10, 2026, covering 70+ core concepts and rolling out globally across all plans. Users can adjust variables, manipulate formulas, and see graphs update in real time; OpenAI says 140 million people use ChatGPT weekly for math and science learning. The key point is productized interactivity, while the post does not disclose the underlying model, evaluation method, or outcome data.

Why it matters: HKR-H lands on the interactive-visual hook, HKR-K on 140M weekly learners plus 70+ concepts and live manipulation, and HKR-R on the product and edtech nerve. It is still a mid-weight product update; model details and learning-outcome evaluation are not disclosed, so it stays in a

Mar 9Monday

OpenAI News

OpenAI to acquire Promptfoo

OpenAI said it will acquire Promptfoo and integrate its technology into OpenAI Frontier after closing. The post discloses that Promptfoo is used by over 25% of Fortune 500 companies, and the deal is still subject to customary closing conditions. The key signal is native agent security testing, red-teaming, and traceability in Frontier; the post does not disclose price or timeline.

Why it matters: This is not a routine partnership; OpenAI is absorbing a known eval and red-team vendor into Frontier. HKR-H/K/R all pass on novelty, concrete adoption data, and strong resonance with agent teams, but price, timing, and integration scope are still undisclosed, so it stays below p

Mar 5Thursday

OpenAI News

Introducing ChatGPT for Excel and new financial data integrations

OpenAI launched ChatGPT for Excel beta on March 5, 2026, bringing GPT-5.4 into Excel workbooks and finance workflows. The post says it can build and update models, trace changes to cells, and is off by default for Enterprise and Edu admins; OpenAI's internal banking benchmark rose from 43.7% with GPT-5 to 87.3% with GPT-5.4 Thinking. The key move is data access: Moody’s, Dow Jones Factiva, MSCI, Third Bridge, and MT Newswires are live, while FactSet is listed as coming soon.

Why it matters: This is more than a routine add-on: OpenAI puts ChatGPT into Excel, names major finance data feeds, and cites a 43.7%→87.3% internal banking benchmark gain. HKR-H/K/R all pass; importance lands at 82 because this is a strong vertical workflow move, not a market-wide model release

Mar 3Tuesday

OpenAI News

GPT-5.3 Instant: Smoother, more useful everyday conversations

OpenAI released GPT-5.3 Instant on March 3, 2026 as an update to ChatGPT’s most-used model, aiming for fewer unnecessary refusals, fewer disclaimers, and more accurate everyday answers. The post shows one concrete contrast: GPT-5.2 Instant refused long-range archery trajectory help, while GPT-5.3 Instant requested parameters and gave a no-drag example at 300 fps (about 91 m/s), 45°, and 845 m; the key issue is the safety-boundary shift, while the post does not disclose benchmark scores, system card details, or API pricing.

Why it matters: OpenAI updated a core ChatGPT everyday model, and the story clears HKR-H/K/R because the refusal-boundary shift is concrete and widely relevant. The post includes a specific 5.2 vs 5.3 behavior example, but no system card, benchmark table, or API pricing, so it lands below the 85

Feb 27Friday

OpenAI News

OpenAI and Amazon announce strategic partnership

OpenAI and Amazon announced a multi-year partnership, with Amazon investing $50 billion in OpenAI: $15 billion upfront and $35 billion tied to conditions. They will launch a Stateful Runtime Environment on Amazon Bedrock, and OpenAI will consume about 2 gigawatts of Trainium capacity on AWS. The part to watch is distribution plus compute lock-in: AWS becomes the exclusive third-party cloud distributor for OpenAI Frontier.

Why it matters: This is not a routine partnership post. The disclosed $50B staged investment, Bedrock runtime, and ~2GW Trainium commitment change OpenAI's distribution and compute posture; HKR-H/K/R all pass, so this lands in P1.

Feb 26Thursday

OpenAI News

OpenAI Codex and Figma launch code-to-design roundtrip workflow

OpenAI and Figma launched a Codex integration on Feb. 26, 2026 that turns code into editable Figma designs and brings Figma Design, Figma Make, and FigJam content back into code. The workflow uses MCP via the Figma MCP Server in the Codex desktop app; OpenAI says Codex has 1M+ weekly users and usage is up 400%+ since the start of the year. The key issue is whether roundtrip context stays intact; the post does not disclose supported models, permission boundaries, or pricing.

Why it matters: This is a solid OpenAI/Figma workflow update with clear HKR-H/K/R: a bidirectional code↔design loop via MCP and Figma MCP Server. It stays below 85 because the post does not disclose model support, permission boundaries, pricing, or roundtrip reliability.

Feb 20Friday

Hugging Face Blog

GGML and llama.cpp join Hugging Face to support the long-term progress of Local AI

Hugging Face said the GGML and llama.cpp team is joining the company, while Georgi Gerganov’s team will still spend 100% of its time maintaining llama.cpp. The post says the project remains 100% open source and community driven, with full technical and community autonomy. The key angle is tighter delivery from transformers model definitions into llama.cpp, aiming for near “single-click” shipping; the post does not disclose timeline, team size, or deal terms.

Why it matters: This is a meaningful local-AI infrastructure move: HF brings in the GGML/llama.cpp team, so HKR-H/K/R all pass. I kept it at 78 because the post confirms staffing and integration direction, but not a ship date, team size, or deal terms.

Jan 28Wednesday

Mistral AI

Mistral releases terminal coding agent Mistral Vibe 2.0

Mistral released Mistral Vibe 2.0, a terminal coding agent powered by the Devstral 2 model family. It adds custom subagents, multi-option clarification, slash-command skills, a unified agent mode and automatic updates.

Why it matters: The post lists Vibe 2.0's custom subagents, slash-command skills and subscription entry point, enough to judge how terminal coding agent workflows change.

Jan 21Wednesday

NVIDIA Blog

Jensen Huang on AI’s “Five-Layer Cake” at Davos: the largest infrastructure buildout in human history

Jensen Huang said at Davos that global VC investment topped $100 billion in 2025, with most capital going to AI-native startups building the AI stack’s application and infrastructure layers. He described AI as a five-layer stack: energy, chips and computing infrastructure, cloud data centers, models, and applications, and cited a US nursing shortage of about 5 million where AI can handle charting and transcription. The key point for practitioners is that the bottleneck is not just models, but the full infrastructure and labor chain.

Why it matters: This clears HKR-H/R because Jensen's Davos framing is a strong, discussable hook for practitioners. HKR-K also passes on specific facts (> $100B VC, five-layer stack, 5M nurse gap), but it is still executive commentary, not a model or product launch, so it stays in the 78-84 band

Oct 22, 2025Wednesday

Hugging Face Blog

Hugging Face and VirusTotal collaborate to strengthen AI security

Hugging Face said on Oct. 22, 2025 it is continuously scanning more than 2.2 million public model and dataset repositories on the Hub through a VirusTotal collaboration. The Hub checks file hashes against VirusTotal and returns status, detection counts, and threat intel without sending raw file contents. The key point is earlier supply-chain visibility before download; the post does not disclose false-positive rates, scan latency, or remediation flow.

Why it matters: HKR-H/K/R all pass: the story moves threat visibility to before download across 2.2M+ public repos and explains the hash-based integration. It stays below must-write because false-positive rate, scan latency, and remediation flow are not disclosed.

Oct 21, 2025Tuesday

Hugging Face Blog

Unlock the power of images with AI Sheets

Hugging Face added vision support to its open-source AI Sheets, letting users analyze images, extract data, generate visuals, and edit images inside a spreadsheet. The post says AI Sheets uses Inference Providers to access thousands of open models, and manual edits plus thumbs-up feedback become few-shot examples; outputs can be exported as CSV or Parquet. What matters is the unified data workflow, not a standalone demo.

Why it matters: Direct-source Hugging Face product update with concrete mechanics: AI Sheets now handles OCR, image understanding, generation, and editing in one spreadsheet flow, and corrections become few-shot examples. HKR-H and HKR-K pass; HKR-R is weaker because the impact is workflow-level

OpenAI News

Introducing ChatGPT Atlas, the browser with ChatGPT built in

OpenAI launched ChatGPT Atlas on October 21, 2025, with a worldwide macOS release for Free, Plus, Pro, and Go users. Atlas embeds ChatGPT, browser memories, and page-visibility controls into the browser; agent mode preview is available for Plus, Pro, and Business. The key shift is persistent browsing context: web content is excluded from training by default unless users opt in.

Why it matters: OpenAI moving ChatGPT into its own browser is a distribution-layer product move, not a routine feature drop, so this lands at 88 and p1. HKR-H/K/R all pass: novel hook, concrete rollout/privacy details, and clear resonance around browser control, retention, and data boundaries.

Oct 13, 2025Monday

OpenAI News

OpenAI and Broadcom announce collaboration to deploy 10 gigawatts of OpenAI-designed AI accelerators

OpenAI and Broadcom announced a multi-year deal to deploy 10 gigawatts of OpenAI-designed AI accelerators, with rack deployments starting in H2 2026 and completing by the end of 2029. OpenAI will design the accelerators and systems, while Broadcom provides accelerator deployment plus Ethernet, PCIe, and optical networking for OpenAI sites and partner data centers. The key signal is OpenAI's custom-chip plus Ethernet cluster path, but the post does not disclose process node, chip specs, or capex.

Why it matters: Not a routine partnership story: OpenAI put a 10GW custom-chip plan and a 2026-2029 deployment schedule on record. HKR-H/K/R all pass, but process node, per-chip specs, and capex are still undisclosed, so this lands in p1 rather than 95+.

Oct 6, 2025Monday

OpenAI News

Codex is now generally available

OpenAI said on October 6, 2025 that Codex is now generally available, with a Slack integration, a Codex SDK, and new admin controls. The post says daily Codex usage is up more than 10x since early August, and GPT-5-Codex served over 40 trillion tokens in three weeks; starting October 20, cloud tasks count toward usage, but the post does not disclose pricing details. The signal for practitioners is enterprise uptake: OpenAI says nearly all of its engineers use Codex, and they merge 70% more pull requests per week.

OpenAI News

Introducing apps in ChatGPT and the new Apps SDK

OpenAI launched apps inside ChatGPT on October 6, 2025 and previewed the Apps SDK for developers, for logged-in users outside the EEA, Switzerland, and the UK on Free, Go, Plus, and Pro plans. Seven partners are live and 11 more are due later this year; the SDK is open source and built on MCP, while the post does not disclose app review, listing, or revenue-share details.

Why it matters: This is a major OpenAI platform move: ChatGPT gains an app layer and developers get an SDK, so HKR-H/K/R all pass. Concrete facts include plan coverage, region limits, 7+11 partners, and an open-source MCP base; listing, review, and revenue-share terms are still undisclosed.

OpenAI News

Introducing AgentKit, new Evals, and RFT for agents

OpenAI launched AgentKit on October 6, 2025 with three agent-building components: Agent Builder, Connector Registry, and ChatKit. The post says Evals adds datasets, trace grading, automated prompt optimization, and third-party model support; Connector Registry covers Dropbox, Google Drive, SharePoint, Microsoft Teams, and third-party MCPs. The real signal is workflow versioning and safety governance; the title mentions RFT, but the provided post does not disclose its training details, pricing, or rollout scope.

Why it matters: This is a substantial OpenAI release for agent builders, with HKR-H/K/R all passing. It provides concrete mechanisms across Agent Builder, connectors, ChatKit, and Evals, but the excerpt does not disclose RFT mechanics, pricing, or rollout scope, so it stays at 84 rather than p1.

Oct 1, 2025Wednesday

OpenAI News

Samsung and SK join OpenAI’s Stargate initiative to expand global AI infrastructure

OpenAI said on Oct. 1, 2025 that Samsung and SK joined Stargate, with the partnership centered on Korea’s AI chip supply and data center expansion. The post gives one hard target: Samsung Electronics and SK hynix plan to scale advanced memory output to 900,000 DRAM wafer starts per month, while OpenAI also signed Korean data center exploration agreements with MSIT, SK Telecom, and Samsung affiliates. The key gap is execution detail: the post does not disclose investment size, timeline, or facility scale.

Why it matters: OpenAI adding Samsung and SK to Stargate is more than a routine partnership: the post gives a 900k DRAM wafer-start target and concrete data-center assessment ties. HKR-H/K/R all pass, but missing capex, timeline, and site scale keeps it featured, not p1.

Sep 29, 2025Monday

OpenAI News

Buy it in ChatGPT: Instant Checkout and the Agentic Commerce Protocol

OpenAI launched Instant Checkout in ChatGPT on September 29, 2025, letting U.S. Plus, Pro, and Free users buy from U.S. Etsy sellers in chat; it currently supports single-item purchases. OpenAI says ChatGPT has over 700 million weekly users and open-sourced the Agentic Commerce Protocol with Stripe; Stripe merchants can enable it with as little as one line of code, while the post does not disclose the fee rate merchants pay.

Why it matters: This is a high-weight ChatGPT product expansion from discovery to completed purchases, so HKR-H/K/R all pass. The post confirms U.S. Free/Plus/Pro checkout with U.S. Etsy sellers and a Stripe-backed protocol layer; merchant fee details are not disclosed, so it stays high but sub-

Sep 25, 2025Thursday

OpenAI News

More ways to work with your team and tools in ChatGPT

OpenAI rolled out shared projects for ChatGPT Business on September 25, 2025, and made them available for Enterprise and Edu plans. Shared projects support email or link invites, two access levels, and private project memory; Enterprise and Edu have them off by default under admin control. OpenAI also added Gmail, Google Calendar, Outlook, Teams, SharePoint, GitHub, Dropbox, and Box connectors, and said ChatGPT can now choose connectors automatically per prompt.

Why it matters: HKR-H/K/R all pass: shared projects, 8 connectors, and prompt-routed connector selection are concrete workflow changes with clear admin controls. I keep it below 85 because this is a collaboration-layer product update, not a model release or a broad capability jump.

OpenAI News

OpenAI introduces GDPval to measure model performance on real-world tasks

OpenAI introduced GDPval, an eval covering 44 occupations and 1,320 real-world work tasks, with 220 gold tasks open-sourced. It spans the top 9 U.S. GDP industries, uses tasks built and vetted by professionals averaging 14+ years of experience, and is limited to one-shot evaluation rather than iterative workflows. The key shift is from exam-style prompts to real deliverables like docs, slides, spreadsheets, diagrams, and multimedia.

Why it matters: OpenAI's GDPval is a strong HKR-H/K/R story: the hook is evaluation on real work outputs, the post adds concrete dataset numbers and limits, and it hits the automation-of-knowledge-work nerve. It is not a model launch or executive event, so it stays featured rather than p1.

OpenAI News

Introducing ChatGPT Pulse

OpenAI previewed ChatGPT Pulse for Pro users on mobile on September 25, 2025, with one daily proactive research update. It uses memory, chat history, feedback, and optional Gmail and Google Calendar connections to generate visual cards; integrations are off by default and outputs pass safety checks. The shift to async delivery matters more than the headline, but the post does not disclose the model, pricing changes, or a Plus launch date.

Why it matters: HKR-H/K/R all pass: the novel angle is proactive outreach, and the post gives concrete scope and input sources. This is a meaningful ChatGPT product update, but model details, rollout beyond Pro mobile, update cadence, and pricing changes are not disclosed, so it stays featured,

Sep 24, 2025Wednesday

OpenAI News

SAP and OpenAI partner to launch sovereign 'OpenAI for Germany'

SAP and OpenAI announced OpenAI for Germany for the German public sector, planned for 2026 and hosted by Delos Cloud on Microsoft Azure. SAP plans to expand Delos Cloud in Germany to 4,000 GPUs for AI workloads; the post does not disclose model names, pricing, or contract size. The key point is delivery: this is a sovereign public-sector deployment focused on compliance, data residency, and AI agents inside existing workflows.

Why it matters: HKR-H/K/R all pass: the story pairs a novel sovereign-deployment angle with concrete facts like a 2026 launch, Delos Cloud on Azure, and 4,000 GPUs. It matters because sovereignty and public-sector procurement are live issues, but missing model, pricing, and deal-scope details it

Sep 23, 2025Tuesday

OpenAI News

OpenAI, Oracle, and SoftBank expand Stargate with five new AI data center sites

OpenAI, Oracle, and SoftBank announced five new U.S. Stargate AI data center sites, bringing planned capacity to nearly 7 GW and investment to over $400 billion in three years. The post says this keeps Stargate on track to reach its full $500 billion, 10 GW commitment by the end of 2025; Oracle-linked sites account for over 5.5 GW, while two SoftBank-OpenAI sites can scale to 1.5 GW in 18 months. The key signal is supply progress: Abilene is already running early training and inference workloads with first NVIDIA GB200 racks delivered in June.

Sep 22, 2025Monday

OpenAI News

OpenAI and NVIDIA announce strategic partnership to deploy 10 gigawatts of NVIDIA systems

OpenAI and NVIDIA signed a letter of intent to deploy at least 10 gigawatts of NVIDIA systems for OpenAI’s next-generation AI infrastructure. NVIDIA plans to invest up to $100 billion into OpenAI as each gigawatt is deployed, and the first 1 GW phase is targeted for H2 2026 on the Vera Rubin platform. The key detail is execution: this is still an LOI, and final terms are not yet closed.

Why it matters: Strong HKR-H/K/R: the official post discloses 10 GW, millions of GPUs, up to $100B intended investment, and a first 1 GW phase in H2 2026 on Vera Rubin. It is still a letter of intent, not a signed final deal, so it stays below the 95+ band; the scale still makes it p1.

Sep 15, 2025Monday

OpenAI News

Introducing upgrades to Codex

OpenAI released GPT-5-Codex and made it the default model for Codex cloud tasks and code review; in testing, it worked independently for more than 7 hours on complex tasks. OpenAI says it used 93.7% fewer tokens than GPT-5 on the lowest 10% of employee turns, while spending 2x longer reasoning, editing, and testing on the highest 10%. The key point is one model now spans interactive coding and long-running agentic execution; pricing and full availability details are not fully disclosed in the provided body.

Why it matters: This is a substantive OpenAI developer-tool update: GPT-5-Codex becomes the default for Codex cloud tasks and code review, with concrete numbers on 7-hour autonomy and token use. HKR-H/K/R all pass; pricing and full availability are not fully disclosed in the excerpt, so it stays

OpenAI News

How people are using ChatGPT

OpenAI and Harvard economist David Deming released a study of 1.5 million ChatGPT conversations, framed as the largest consumer-usage analysis to date against ChatGPT’s 700 million weekly active users. The paper says feminine-name users rose from 37% in Jan 2024 to 52% in Jul 2025; 49% of messages were Asking, 40% Doing, 11% Expressing, and about 30% of usage was work-related. The shift to watch is distribution: by May 2025, adoption growth in the lowest-income countries was over 4x that of the highest-income countries, while the study covers consumer plans only.

Why it matters: HKR-H/K/R all pass: the story has a strong hook, concrete usage splits, and clear relevance to workplace adoption and global diffusion. I stop at 82 because this is a consumer-usage study, not a model or product change, so it is high-signal context rather than same-day must-cover

Sep 2, 2025Tuesday

Mistral AI

Make Memory work for you.

Mistral AI 为 Le Chat 上线 Memories(beta)记忆功能,可自动保存有用信息,但回忆过程可见、可溯源,并附来源链接。用户可随时关闭记忆、开启不使用记忆的隐身对话、编辑或删除单条记忆,还能导出和从外部导入记忆。同步推出 Memory Insights,基于用户自身数据提示记忆趋势与摘要。

OpenAI News

Vijaye Raji to become CTO of Applications with acquisition of Statsig

OpenAI said it will acquire Statsig, and Vijaye Raji will become CTO of Applications once the deal closes. Raji will report to Fidji Simo and lead product engineering for ChatGPT and Codex, including infrastructure and Integrity. Statsig staff will join OpenAI after closing, but the platform will keep operating independently from Seattle; regulatory approval is still pending.

Why it matters: OpenAI is acquiring Statsig and naming Vijaye Raji as CTO of Applications, a high-signal personnel plus M&A story tied to ChatGPT and Codex engineering. HKR clears all three; the post gives scope and close structure but omits price and integration timeline, so this is must-write,

Mistral AI

Mistral adds custom MCP connectors and Memories to Le Chat

Mistral launched a catalog of 20+ secure MCP-based connectors for Le Chat (beta), covering data, productivity, development, automation and business categories. It supports tools including Databricks, Snowflake, GitHub, Atlassian, Asana, Outlook, Box, Stripe and Zapier, and lets users add custom MCP connectors or connect to any remote MCP server.

Why it matters: The post lists the specific tool categories and deployment methods behind the 20+ connectors, showing how far Le Chat reaches into enterprise workflows.

Aug 28, 2025Thursday

OpenAI News

Introducing gpt-realtime and Realtime API updates for production voice agents

OpenAI released the speech-to-speech model gpt-realtime and made the Realtime API generally available, adding remote MCP server support, image input, and SIP phone calling. The post reports 82.8% on Big Bench Audio versus 65.6% for the December 2024 model, and 30.5% on the audio MultiChallenge benchmark versus 20.6%. The key change is that tool access and phone connectivity now ship in the same production API.

Why it matters: This is a substantive OpenAI model + API release, not a minor refresh. HKR-H/K/R all pass: the release has a clear hook, hard benchmark deltas, and direct deployment impact for production voice agents, so it reaches p1.

Aug 7, 2025Thursday

OpenAI News

Introducing GPT-5 for developers

OpenAI released GPT-5 in its API on August 7, 2025, in three sizes: gpt-5, gpt-5-mini, and gpt-5-nano. The post reports 74.9% on SWE-bench Verified, 88% on Aider polyglot, 96.7% on τ2-bench telecom, plus new verbosity, minimal reasoning_effort, and custom tools; pricing and full availability details are not disclosed in the provided text. The real developer signal is the API surface change, not just a model rename.

Why it matters: This is an OpenAI flagship-model API launch, so it belongs in the 95–100 band. HKR-H lands on the GPT-5 debut; HKR-K lands on concrete benchmark scores and new controls; HKR-R lands on immediate developer concerns around migration, tooling, and model comparison; the excerpt omits

OpenAI News

Introducing GPT-5

OpenAI launched GPT-5 on August 7, 2025 and made it available to all ChatGPT users. The system combines a base model, GPT-5 thinking, and a real-time router; Plus gets higher limits, while Pro gets GPT-5 pro. The key change is unified routing with built-in reasoning; the post does not disclose pricing, context window, or API specifics.

Why it matters: An OpenAI frontier-model launch is a top-band event on its own. The excerpt confirms a unified system (base model + GPT-5 thinking + router) and rollout to all ChatGPT users; HKR-H/K/R all pass, and missing price/context/API details do not block p1.