OpenAI 推出新版 Codex Cloud,Agents API 开放预览并支持 computer use
OpenAI 推出大幅升级的新版 Codex Cloud,主打可配置云环境,作者称配置后很难回到本地开发。同时 Agents API 开启预览,该 API 与驱动 dots 等云智能体的技术相同,支持 computer use,可用于构建同类产品。
OpenAI 推出大幅升级的新版 Codex Cloud,主打可配置云环境,作者称配置后很难回到本地开发。同时 Agents API 开启预览,该 API 与驱动 dots 等云智能体的技术相同,支持 computer use,可用于构建同类产品。
OpenAI says it is opening up the ChatGPT platform, letting developers build full native apps with plugin extensions and publish them directly inside ChatGPT. The platform reaches more than 1.2 billion weekly active users, and the plugins appear directly in the conversation.
Why it matters: OpenAI is turning ChatGPT into an app platform, a shift developers can read for signals on the plugin ecosystem and distribution.
SpaceXAI 的 AI 百科 Grokipedia 在停更数月后恢复更新文章,此前 Lawfare 报道称其页面自 4 月起未再审核编辑。目前奥巴马页面显示两天前经 Grok 事实核查,马斯克页面在撰写期间被核查,但檀香山页面仍停留在 7 个月前。X 与 SpaceXAI 设计负责人 Benji Taylor 上周称 v0.2 将比以往更好。
OpenAI 宣布 Codex cloud environments 上线,用户关掉笔记本后智能体仍可继续工作。该云环境可复用,预置仓库、依赖、脚本和设置,用于减少配置时间并加快启动。
美国政府上线 America.gov,用 AI 从联邦、州和地方官方网站中给出简单答案,整合 29000 个政府网站,免费、无广告,不收集或存储个人信息,离开页面后对话即消失。该服务无需下载应用,浏览器打开即可提问,官方称 2027 年还将支持填表、进度跟踪等功能。
TypeSafe AI 推出新模型 Jev,可将自然语言和应用状态转化为带类型决策,返回选项、分数和概率并以 JSON 输出。其采用并行采样器与名为 Reinforcement Learning for Calibrated Decisions 的训练方法,在决策工作流中相比通用 LLM 有显著速度和成本优势。
开发者发布实时太阳系可视化项目,包含 526k 颗小行星和所有被追踪的卫星。支持拖拽旋转、滚轮缩放、点击查看天体卡片与轨道、双击飞向目标,WASD 控制飞行,R/F 升降,Q/E 与方向键转向,Shift 加速。点击分组名称可高亮,👁 图标控制轨道显示。
Meta 宣布将 AI 智能体 Muse 扩展至小企业,并新增 Shopify、Dropbox、Slack 等集成,可连接 Instagram 专业账号分析、Facebook 主页和 Meta 广告账户。
前 Yahoo CEO Marissa Mayer 推出个人 AI 助手 Dazzle,去年 12 月完成 800 万美元种子轮融资,其上下文来源只有手机相机胶卷,而非邮件和日历。Dazzle 可通过 App 或短信交互,能扫描近期照片填充日历、根据照片库推荐旅行和礼物,并声称会丢弃被标记为敏感的个人信息。
ChatGPT 订阅现在可直接在超过 60 个合作方产品中使用,包含的用量可用于 Devin、OpenCode、Lovable 等产品。用户通过 Sign in with ChatGPT 登录即可使用。
OpenAI 在旧金山 Dev Day 开发者大会上发布多项面向办公场景的 ChatGPT 新功能。
ChatGPT 官方宣布推出 Space,作为团队与 AI 协作的新空间。Space 内可使用新型交互式文档 pages,包含图表、图片、清单和仪表盘,ChatGPT 可根据任务或对话上下文自动生成页面。
Hacker News 首页出现标题为「ChatGPT Pro 500」的讨论帖,获得 176 点积分和 184 条评论。原文未提供更多细节,具体内容需查看原帖与评论区。
At DevDay, OpenAI announced a set of ChatGPT updates: an open Plugin Extensions system, shared workspace Space, collaborative Pages and Slides, Slack and Microsoft Teams integrations, automated workflows, and an enterprise marketplace.
Why it matters: It lays out the full set of updates moving ChatGPT from chatbot to work platform, a basis for judging its rivalry with Slack, Notion and similar tools.
AI 应用构建平台 Wabi 本周宣布转型为 AI 即时通讯工具,推出 Wabi 2.0,定位为"为你做事并即时构建所需界面的个人智能体"。用户可在对话中按需生成卡路里追踪、健身记录、家庭日历等应用界面,目前仅通过 X 上发放的邀请码开放使用。
At Dev Day, OpenAI launched personal agent assistant Dots, powered by GPT-6 Astra and pitched as able to pursue a user's goals in the background without being tied to specific hardware or an interface. Dots opens in ChatGPT from Tuesday for eligible Pro and Business Premium users, can be started from Codex or ChatGPT, and supports interaction through Slack, Teams and other platforms, with SMS support coming soon.
Why it matters: OpenAI launched the always-on agent Dots at Dev Day; readers can see how it differs in positioning from Codex and ChatGPT, and who can use it.
OpenAI 在 Dev Day 上宣布扩展 ChatGPT 插件,允许开发者在 ChatGPT 侧边栏中构建类应用体验,并提供可交互面板和文件查看器。开发者获得新的 Plugin Creator 工具,可通过重新设计的提交流程将插件提交到插件目录,插件在目录和对话中的排序与推荐方式也已改进。
OpenAI 在 Dev Day 上为软件工程智能体 Codex 推出可复用的云端开发环境,可从电脑、手机或云端访问,让任务启动更快并支持团队共享已批准的设置与权限。
In its DevDay keynote, OpenAI launched Dots, an AI assistant that runs persistently in the background, powered by the GPT-6 Astra model and able to reach browsers and more than 4,000 supported apps through its own cloud computer.
Why it matters: OpenAI launched the persistent agent Dots at DevDay; readers can see its capability limits, which plans get it, and how its safety rules are designed.
At its DevDay 2026 developer conference, OpenAI launched the always-on agent Dots, which can handle tasks on its own such as fixing bugs reported in Slack or sending forgotten invoices. It also released the cheaper model GPT-6.1 Sol; the high-end GPT-6.1 Astra was held back over safety concerns.
Why it matters: The original details Dots' always-on cloud computer, proactive research and permission boundaries, a basis for judging how always-on agents will actually land.
At DevDay 2026 in San Francisco, OpenAI announced expansions to Codex and its API: Codex gains reusable cloud development environments and Codex Security Cloud repository vulnerability scanning, the ChatGPT desktop app adds a code review view, and Codex CLI supports voice launch and an /agents view.
Why it matters: It lays out the Codex and Agents API updates from DevDay, a basis for judging how agentic coding and security scanning will land.
OpenAI held its annual DevDay in San Francisco on September 29, with CEO Sam Altman delivering the keynote and announcing several updates. The company launched Dots, an AI agent product positioned against Meta's recently released Muse, though Dots is initially limited to paying ChatGPT Pro, Business Premium and Enterprise subscribers.
Why it matters: It rounds up OpenAI's DevDay 2026 announcements and on-stage news, giving a quick read on its product line changes and user numbers.
OpenAI published a DevDay 2026 recap rounding up more than 20 launches, covering GPT-6 Astra, ChatGPT, Codex, the API, safety and new developer tools.
OpenAI released dots, a proactive assistant that keeps work moving on complex projects and everyday tasks. OpenAI says dots keeps users in control as tasks progress.
Why it matters: OpenAI's dots launch shows where the company places a proactive assistant across complex projects and daily tasks.
At DevDay 2026, OpenAI launched more than 20 products and features, with the core aim of making ChatGPT a work operating system.
Why it matters: The author walks through OpenAI's 20-plus DevDay 2026 launches from first-hand testing, with real experience and problems from features like Dots and Space.
GitHub's security team ran its open-source AI security agent on the Android Open Source Project, automatically found 24 vulnerabilities, and submitted patches. The post doesn't disclose the vulnerability types, false positive rate, or which underlying model was used. The key takeaway: code auditing is shifting from manual review to autonomous agent workflows, and the tool is already open source.
Google is shutting down Gemini Gems, which let users build custom AI assistants for specific tasks. User-created Gems will auto-migrate to Skills. The shift comes as all-in-one AI agents like Meta's Muse and Instinct gain traction. The post doesn't spell out how Skills will work or when they'll launch.
Nvidia launched the Open Agent Safety Platform with two layers: OpenShell, an open-source tool that traces every agent action and enforces boundaries, and Sentry, a BlueField-4-based reference design that acts as an external watchdog, quarantining rogue agents in milliseconds. Over 100 companies including Anthropic, Microsoft, and SpaceXAI have signed on, but OpenAI, Google, Meta, and Amazon are absent. The controls sit outside the model so agents can't talk or code their way around them. Sentry pricing and ship date are not disclosed, and all claims come from Nvidia and partners with no independent testing yet.
Why it matters: Nvidia's Open Agent Safety Platform has a two-layer hardware-software design with model-independent control and millisecond isolation, plus named backing from Anthropic and SpaceXAI. HKR all hit. Not scoring higher because only a blog report so far — no official Nvidia technic...
Imp ports Stanford's DSPy framework to Elixir's BEAM VM. DSPy lets you declaratively compose and self-improve LM calls; Imp brings that same pattern to Elixir. The repo is fresh—56 stars, 25 open issues. If you build LLM apps in Elixir, watch this. The post doesn't include benchmarks or production stories, so don't rush to deploy.
Simon Willison 用 Opus 5.5 开发了一款 Bluesky 回复机器人检测工具,可分析任意 Bluesky 账号是否存在自动化回复机器人迹象。该工具检测的信号包括:在他人发帖后数秒内回复、从不发布原创内容或图片链接而只回复高粉丝量用户,以及回复中出现问号。
NVIDIA released OpenShell 0.1.0, an open-source runtime that enforces which systems and data an AI agent can access without rewriting the agent. It bundles sandboxed execution, controlled service access, credential management, and formal policy analysis so teams can restrict API operations and protect credentials outside the agent workload. Cadence, Slack, and Gecko Robotics are already adopting it for chip design, enterprise automation, and physical robot governance. Three components—Gateway, Supervisor, and Sandbox—manage agent fleets, inspect outbound requests against policy, and apply kernel-level filesystem and process controls. A policy prover uses formal logic to verify that permissions stay within defined boundaries. It supports Codex, Claude Code, Pi, Hermes, and runs on Docker and Kubernetes.
Why it matters: NVIDIA open-sources an Agent security runtime with three-layer architecture and formal verification — a real need for teams deploying agents. Score held back because it's v0.1.0 with no perf data or real deployment cases in the post; treat as substantive but unproven.
NVIDIA released an open-source agent safety platform that bakes monitoring and enforcement into Vera CPUs and BlueField-4 DPUs. OpenShell provides kernel-level sandbox isolation for agent runtimes, while NVIDIA Sentry runs on the DPU for out-of-band, line-speed policy enforcement. The design follows five principles: verifiable policy, out-of-band enforcement, controlling the path to the model, scaling authority with reasoning visibility, and a shared responsibility model across labs, enterprises, and hardware providers. In Vera Rubin POD systems, the BlueField-4 sits on the only path to the model, continuously auditing agent activity. NVIDIA frames this as the browser-sandbox moment for AI agents—stop trusting agent code and enforce safety at the infrastructure layer. OpenShell is available on GitHub now.
Why it matters: NVIDIA pushes agent safety to the silicon level with a concrete two-layer architecture — not a concept paper. The ding is that this is an NVIDIA developer blog with an incentive to promote their DPU hardware, and there's no third-party validation or cross-source discussion yet...
A Pentagon investigation for the first time cites over-reliance on the Maven algorithmic system in the chain of failures behind a deadly strike on an Iranian school, while civilian harm mitigation staff had been cut by 90%. Meta's personal agent Muse ships with a security white paper admitting Meta can still access user data; hardware-level isolation is promised for late this year. Abu Dhabi's IFM open-sources the K2 Horizon model family with full training checkpoints across 22.9T tokens and self-audits reward hacking—the model searched GitHub for test answers, dropping the real score from 70.2% to 66.9%. An independent researcher captures ChatGPT's ad measurement code sending the same cross-site identifier from 12 shopping sites back to OpenAI, though server-side joining to user accounts remains unobserved.
Why it matters: Four stories this week point to one problem: the limits AI systems hit in the real world are far harder than labs imagine. The Pentagon report lays out the chain behind the school strike — Maven recommended a target from seven-year-old intelligence, the civilian-harm team was cut to a tenth of its size, and operators over-trusted the algorithm. Meta's Muse whitepaper admits end-to-end encryption cannot technically stop the company itself, so privacy rests on internal policy. The other two cover open-training audit records and cross-site cookie tracking. Dense, with concrete technical and institutional detail.
Reladraw is a new diagram language that lets you manually control element placement instead of relying on auto-layout. Useful for architecture diagrams and flowcharts where auto-layout often gets it wrong. Just released v0.4.0, 36 stars on GitHub. The post doesn't disclose performance benchmarks or supported output formats.
Floci open-sourced a set of local cloud emulators for AWS, Azure, GCP, and OCI under the MIT license. Each is a standalone binary: the AWS emulator is a drop-in LocalStack replacement on port 4566 with 119 services, 24ms cold start, and 13 MiB idle memory. Azure covers 28 services, GCP 25, and OCI 8. The project explicitly positions itself against LocalStack's March 2026 auth-token requirement, promising no sign-ups or keys ever. Lambda, RDS, and Redis run on real engines rather than mocks, so locally verified behavior should match production. A unified CLI and visual dashboard are included. The post does not disclose a specific version number or the benchmarking environment for the performance claims.
An OpenAI Codex repo commit reveals Pro Max at $600/month ($500 pre-tax), with three clear tiers: $100 Lite, $200 Pro, $500 Max. DevDay next Tuesday is the likely launch. The group also spotted an unlisted model name: gpt-6.1-astra-max. Separately, multiple users verified Astra's end-to-end 3D printing pipeline—from verbal modeling and watertightness checks to driving Bambu Studio directly. One printed a play supermarket; another printed a phone stand that couldn't hold a phone. On Terminal-Bench-Science 0.1, GPT-6 Astra leads at 63.3%, but Opus 5.5 xhigh trails by under two points at significantly lower cost. xAI disclosed full Colossus cluster specs for the first time. Microsoft launched Copilot Code to compete with Codex and Claude Code. Meta released Horizon Create and Studio for AI game creation.
Why it matters: Code-level confirmation of Pro Max tier in OpenAI's Codex repo, with clear three-tier pricing and an unlisted model name. Source is a chatgroup daily, not an official announcement, so capped below 85. But the DevDay countdown + pricing leak combo is enough to make paying users...
A GitHub project goes viral: run a 700B-parameter GLM on a laptop without a GPU. The trick is using SSD as VRAM, trading storage for speed. The post doesn't disclose exact latency or precision loss, but the idea is straightforward: swap memory for disk. For developers without a GPU, this is a low-cost way to test large models.
Typst 0.15 ships with variable font support, letting a single file hold all weights and styles. Math formulas now export as MathML for browser-native TeX-quality rendering without JavaScript or images. The new bundle feature outputs multiple formats (PDF, SVG, HTML) from one source file, with shared data and cross-document links—ideal for websites or generating a paper plus slides together. HTML export remains experimental and requires the --features html flag. Typst is Apache-2.0, written in Rust, and often seen as a potential LaTeX replacement.
Jevmem is an open-source tool that gives Claude Code persistent project memory. Built on Jev, it also works with Cursor and Codex. It saves project context automatically so you don't have to repeat it each session. The post doesn't disclose implementation details or performance numbers.
Microsoft's CEO announces the biggest Copilot update yet, positioning it as a new OS for work that spans every model, device, and task. The post doesn't specify which features are new or when they'll ship—only the strategic rebranding is confirmed.