When chat is the wrong UI
GitHub Copilot 应用推出 canvas,一种运行在应用内、无浏览器外壳的全栈小应用,可与 Copilot 智能体双向通信,并能在本地执行代码、调用第三方 API。作者认为聊天只是 AI 的通用兜底界面,用户明确任务时更该让智能体生成可复用工具,而非把智能体本身当工具、白白消耗 token。示例包括 Connect 4 游戏、Winget 包管理、SQLite 操作和开发工作流自动化。
GitHub Copilot 应用推出 canvas,一种运行在应用内、无浏览器外壳的全栈小应用,可与 Copilot 智能体双向通信,并能在本地执行代码、调用第三方 API。作者认为聊天只是 AI 的通用兜底界面,用户明确任务时更该让智能体生成可复用工具,而非把智能体本身当工具、白白消耗 token。示例包括 Connect 4 游戏、Winget 包管理、SQLite 操作和开发工作流自动化。
GitHub Copilot 应用重建了 pull request 视图,以流畅渲染含 2,200 个文件、超百万行改动和 400 多条行内评论的超大 PR。其做法是把文档高度拆成确定性的代码几何与动态评论块两套几何:代码行高提前精确算好,评论高度按块懒测量并锚定到文件、行与侧,避免滚动跳动。
GitHub Copilot 应用内置 diff、终端和浏览器三个面板,让用户无需离开应用即可审查、运行和预览 AI 智能体生成的代码变更。diff 面板以绿色和红色高亮显示代码的增删改,终端面板支持直接运行项目命令并可通过 Run 按钮配置脚本,浏览器面板则提供 Pick & Polish 工具来选取页面元素并让智能体调整。
GitHub Copilot 应用支持同时运行多个智能体会话,每个会话运行在独立的 Git worktree 上,互不干扰且各自保留上下文,可随时切换并从中断处继续。用户可在会话视图中查看各任务标题与进度,例如在同一项目上并行执行 funded sort 开发、无障碍审查和测试运行。
OpenAI published the first measured results for Jalapeño, its custom inference chip. On the InferenceX benchmark running GPT‑OSS 120B, it delivered higher peak throughput per kilowatt and lower token latency than the commercial systems compared, with strong results on DeepSeek R1 and Kimi K2 as well. The post frames this as a working first-party silicon path that gives OpenAI direct control over serving economics. It also details a multi-supplier compute portfolio—Microsoft, NVIDIA, AWS, AMD, Broadcom, Cerebras, CoreWeave, Oracle, SB Energy, SoftBank—and a self-built data center in Georgia called Project Camellia. The core argument: co-designed hardware and software lower the cost of useful intelligence, which expands usage, funds further R&D, and creates a compounding advantage.
Why it matters: OpenAI's first public benchmarks for its custom Jalapeño inference chip show better per-kW throughput and per-token latency than commercial alternatives on GPT-OSS 120B, with solid results on DeepSeek R1 and Kimi K2. This marks a key step from pure model company to full-stack ...
Hugging Face released Discover Tool, a reference implementation of the Agentic Resource Discovery (ARD) spec. ARD is an open draft co-developed by Microsoft, Google, GoDaddy, Hugging Face, and others. It lets agents find MCP tools, A2A agents, or skills at runtime via natural-language search instead of hardcoding each one. Hugging Face's implementation wraps the Hub's existing semantic search and Agent Skills into an ARD catalog, exposed as a REST API and an MCP Tool. The post does not disclose pricing, search latency, or accuracy figures.
Why it matters: ARD tackles a real pain point—agent tool discovery—with cross-vendor backing from Microsoft, Google, and Hugging Face, plus a working reference implementation. Not scoring higher because it's still an open draft, not a ratified standard, and the post doesn't spell out adoption...
NVIDIA and Microsoft announced a unified agentic AI deployment stack at Build across Windows, Azure, and local environments; RTX Spark provides 1 petaflop of AI performance, while DGX Station for Windows offers 20 petaflops of FP4 performance and up to 748GB of coherent memory.
Why it matters: HKR-H/K/R pass: the NVIDIA-Microsoft stack spans Windows, Azure, and local devices, with 1 PFLOP and 20 PFLOPs FP4 specs. Vendor-source limits the score: pricing, benchmarks, and migration details are not disclosed.
NVIDIA added MRC support to Spectrum-X Ethernet, letting one RDMA connection spread traffic across multiple paths. MRC ran in Blackwell deployments, with microsecond failure bypass and hardware rerouting. The key detail is the OCP open specification and multiplane support for clusters up to hundreds of thousands of GPUs.
Why it matters: HKR-K/R are solid: MRC stripes one RDMA flow across paths, detects failures in microseconds, and is tied to Blackwell deployments. HKR-H is narrow and the source is vendor-owned, so this stays below major release level.
Anthropic made Microsoft 365 connectors available on every Claude plan, covering Outlook, OneDrive, and SharePoint. The post confirms plan coverage and supported apps; it does not disclose pricing, permission boundaries, regional limits, or admin requirements. The real signal is broad rollout across all plans, not a new standalone connector.
Why it matters: This is a mid-weight Claude product update: Anthropic expanded Microsoft 365 connectors to every Claude plan, which changes real Outlook, OneDrive, and SharePoint access. HKR-H/K/R all pass, but missing price, permission, region, and admin details keeps it at low-end featured.
OpenAI and Microsoft issued a joint statement. The provided content includes only the headline and no body text, so the only confirmed fact is that the statement came from the two companies; its subject, actions, and timing are not stated.
Why it matters: An official statement gives this enough weight: it says OpenAI's new funding and partners do not change Microsoft's existing terms. HKR-K and HKR-R pass because the alliance shapes cloud distribution and market power; HKR-H is weak and detail density is limited.
SAP and OpenAI announced OpenAI for Germany for the German public sector, planned for 2026 and hosted by Delos Cloud on Microsoft Azure. SAP plans to expand Delos Cloud in Germany to 4,000 GPUs for AI workloads; the post does not disclose model names, pricing, or contract size. The key point is delivery: this is a sovereign public-sector deployment focused on compliance, data residency, and AI agents inside existing workflows.
Why it matters: HKR-H/K/R all pass: the story pairs a novel sovereign-deployment angle with concrete facts like a 2026 launch, Delos Cloud on Azure, and 4,000 GPUs. It matters because sovereignty and public-sector procurement are live issues, but missing model, pricing, and deal-scope details it
OpenAI and NVIDIA signed a letter of intent to deploy at least 10 gigawatts of NVIDIA systems for OpenAI’s next-generation AI infrastructure. NVIDIA plans to invest up to $100 billion into OpenAI as each gigawatt is deployed, and the first 1 GW phase is targeted for H2 2026 on the Vera Rubin platform. The key detail is execution: this is still an LOI, and final terms are not yet closed.
Why it matters: Strong HKR-H/K/R: the official post discloses 10 GW, millions of GPUs, up to $100B intended investment, and a first 1 GW phase in H2 2026 on Vera Rubin. It is still a letter of intent, not a signed final deal, so it stays below the 95+ band; the scale still makes it p1.
OpenAI said its nonprofit will keep control of its PBC and receive an equity stake exceeding $100 billion. The post also confirms a first $50 million grant program across AI literacy, community innovation, and economic opportunity; it does not disclose the valuation method, stake size, or closing timeline. The real issue is governance: the statement says safety decisions must follow OpenAI's mission, and OpenAI is working with the California and Delaware Attorneys General.
OpenAI and Microsoft said on September 11, 2025 they signed a non-binding MOU for the next phase of their partnership and are working on a definitive agreement. The post discloses only the MOU and the plan to finalize terms, not funding, term length, compute, or equity changes. This is not a closed contract yet; it is a signal that talks continue.
Why it matters: Primary-source signal on the industry's most important AI partnership. HKR-H/R pass because the OpenAI-Microsoft reset is inherently clickable and strategic; HKR-K fails because the post gives only a non-binding MOU and a pending definitive agreement, with no economics, term, or
OpenAI said on May 5 that its nonprofit will keep control of OpenAI, while its for-profit LLC will convert into a Public Benefit Corporation. The post says the nonprofit will remain the controller and become a large shareholder of the PBC, after talks with the California and Delaware attorneys general. The key point is governance did not shift, but the post does not disclose the ownership split, PBC timeline, or Microsoft-specific terms.
Why it matters: This is a high-signal OpenAI governance update: nonprofit control remains, the for-profit LLC converts to a PBC, and the plan was discussed with California and Delaware AG offices. HKR-H/K/R all land; undisclosed equity split, timing, and Microsoft terms keep it below the top bin
OpenAI disclosed in a February 2025 threat report that it banned dozens of accounts tied to a deceptive employment scheme. The accounts used its models to generate fake résumés, fake references, and real-time interview answers to land remote jobs at Western companies. The tactics match what Microsoft and Google previously attributed to North Korean IT-worker fraud, though OpenAI says it cannot confirm the actors' locations or nationalities. Once hired, they kept using the models for coding tasks and to invent cover stories for skipping video calls.
Why it matters: OpenAI's own threat intel report details account bans tied to a deceptive hiring scheme—fake resumes, real-time interview cheating, and post-hire cover stories—with links to DPRK IT worker activity. It's a first-party enforcement action with concrete TTPs, not a generic safety...
OpenAI said on January 30, 2025 it signed an agreement with the U.S. National Laboratories to deploy o1 or another o-series model on Venado, an NVIDIA supercomputer at Los Alamos, for a system that includes about 15,000 scientists. The resource will be shared across Los Alamos, Lawrence Livermore, and Sandia for science, cybersecurity, energy, and nuclear-security work; the key detail is that nuclear and broader CBRN use cases will receive selective review and safety consultation from OpenAI researchers with security clearances.
Why it matters: Strong HKR-H/K/R: the national-lab + nuclear-review angle is clickable, and the post adds concrete facts—15,000 scientists, Venado, three labs, and selective CBRN review. Not P1 because this is a partnership deployment, not a new model release or major capability jump.
OpenAI says its board is evaluating changes to its nonprofit/for-profit structure, after estimating in 2019 that AGI would require about $10B. The post cites ChatGPT’s 300M+ weekly users and $137M in 2015 donations, but the specific final structure under consideration is not fully disclosed in the provided text. The key signal is financing pressure: OpenAI says investors at this scale want more conventional equity.
OpenAI said it banned ChatGPT accounts tied to the Iranian influence operation Storm-2035 in August 2024 after they generated election and geopolitics content for X, Instagram, and five websites. The company identified 12 X accounts and one Instagram account; on Brookings' Breakout Scale, the operation ranked at the low end of Category 2, with most posts getting few or no likes, shares, or comments. What matters is the workflow: the models were used for long articles, comment rewrites, and English-Spanish posting, not for meaningful audience reach.
Why it matters: HKR-H lands on the covert election-influence angle; HKR-K lands on the account counts, sites, languages, and Breakout Scale 2. HKR-R lands via model-abuse governance, but the score stays at 76 because OpenAI reports no meaningful audience reach.
OpenAI on July 18, 2024 launched an Enterprise Compliance API, eight third-party compliance integrations, and SCIM user management for ChatGPT Enterprise. The post confirms timestamped exports for conversations, files, GPT configs, memories, and users, plus support for Okta, Microsoft Entra ID, Google Workspace, and Ping; details on “Expanded GPT controls” are not disclosed in the provided body.