Skip to content

#其他

0 today

Jul 9Thursday

OpenAI News

OpenAI turns its bio bug bounty into an ongoing program, doubling rewards to $50K starting with GPT-5.6

OpenAI is turning its GPT-5.5 Bio Bug Bounty into an ongoing private program, now called the OpenAI Bio Bounty Program. The focus stays on universal jailbreaks that beat its biosafety challenges. Rewards jump from $25,000 to $50,000 for both GPT-5.6 and GPT-5.5, with smaller payouts possible for partial wins. GPT-5.5 testing ends July 27, 2026; after that only GPT-5.6 is in scope. Applicants need a ChatGPT account, must sign an NDA, and past GPT-5.5 applicants don't need to reapply.

Why it matters: OpenAI upgraded its bio-safety bounty from a one-off to a permanent program with doubled rewards and clearer rules — a substantive safety-mechanism update. But the audience fit is narrow: bio-jailbreak testing is far from most practitioners' daily work, so resonance is weak, k...

Jun 25Thursday

OpenAI News

OpenAI publishes economic research paper on how Codex is reshaping work

OpenAI released an economic research paper on June 25, using internal and external usage data to track Codex adoption over the past year. By May 2026, 80.6% of sampled individual users had run at least one Codex task estimated to exceed 30 minutes of human work, and 25.6% had run tasks exceeding eight hours. Inside OpenAI, Codex now accounts for 99.8% of weekly output tokens; Legal and Recruiting switched their primary AI tool from ChatGPT to Codex around April 2026. Non-developer users grew fastest—137x for individuals, 189x for organizations. The paper does not disclose Codex pricing or external enterprise conversion rates.

Why it matters: OpenAI's economic research team published a paper quantifying Codex's shift from chat to long-horizon agent tasks, with 80.6% and 25.6% penetration as the core hooks. It's a self-published promotional study, not independent research, so the score stays below 85.

Google Research Blog

How reasoning unlocks parametric knowledge in LLMs

Google Research shows that letting models think before answering sharply improves their ability to recall facts from training data. On Natural Questions, Gemini 2.5 Pro jumps from ~40% accuracy without reasoning to over 70% with it. The gain comes from the model connecting fuzzy memories into verifiable chains, not from external retrieval. The reasoning traces often include self-questioning and fact-checking steps. The post only covers QA tasks so far.

Why it matters: Google Research published a mechanism study with concrete numbers showing how reasoning helps models retrieve parametric knowledge, with a clear 40%→70% jump. Missing generalization evidence beyond Natural Questions keeps the score from going higher. Useful for RAG and eval pr...

Jun 22Monday

OpenAI News

OpenAI launches Patch the Planet to help open source maintainers patch bugs, not just find them

OpenAI's Daybreak initiative partners with Trail of Bits to pair GPT‑5.5‑Cyber and Codex Security with human review, finding and patching vulnerabilities across 19 critical open source projects including cURL, Go, and Python. Security engineers filter false positives and develop patches before handing off to maintainers. The initial sprint found hundreds of issues, merged dozens of patches, and built reusable fuzzing and variant-analysis pipelines.

Why it matters: OpenAI deployed security models against real open-source infrastructure with named projects and merged patches—not a concept piece. Hits all three HKR axes, but it's a one-off initiative rather than a product-line update, so capped at 78 in the featured tier.

OpenAI News

Samsung Electronics rolls out ChatGPT and Codex to employees in one of OpenAI's largest enterprise deals

Samsung Electronics is deploying ChatGPT Enterprise and Codex to all employees in Korea and its DX division worldwide, covering R&D, manufacturing, marketing, and more. OpenAI calls it one of its largest enterprise launches ever. Codex now has over 5 million weekly active users; weekly actives in Korea grew nearly 800% since Feb 1, 2026. The post does not disclose deal value or rollout timeline.

Why it matters: One of OpenAI's largest enterprise deployments ever, with a concrete 800% Codex WAU spike in Korea. No deal size or timeline disclosed, so it stays at 78 rather than the 85+ band.

Jun 18Thursday

OpenAI News

OpenAI o3 Deep Research reanalyzed 376 unsolved pediatric cases and surfaced leads for 18 rare-disease diagnoses

Boston Children’s, Harvard, and OpenAI used o3 Deep Research to reanalyze 376 previously unsolved pediatric rare-disease cases. The model proposed evidence-linked hypotheses; after expert review and lab confirmation, physicians established 18 new diagnoses—an additional yield of 4.8%. The model never made clinical decisions. All confirmed diagnoses went through CLIA-certified lab validation. The study appears in NEJM AI and the authors note it is a retrospective analysis, not yet a routine clinical tool.

Why it matters: NEJM AI-published study: o3 deep research reanalyzed 376 unsolved pediatric rare-disease cases and surfaced 18 new diagnoses (4.8%). Has a paper, concrete numbers, and a CLIA validation pipeline — not a PR fluff piece. Held at 78 rather than 85+ because it's a single study, no...

Jun 17Wednesday

Hugging Face Blog

Z.AI releases GLM-5.2: first open-source model with solid 1M-token context, built for long-horizon coding tasks

Z.AI open-sourced GLM-5.2, a model built for long-horizon coding tasks. It delivers a genuinely usable 1M-token context—not just accepting more tokens, but maintaining quality across long agent trajectories. IndexShare reuses one indexer across every four sparse attention layers, cutting per-token FLOPs by 2.9× at 1M context; MTP acceptance length improved by up to 20%. On FrontierSWE it beats GPT-5.5 by 1%, and on PostTrainBench it outranks both GPT-5.5 and Opus 4.7, placing second. It's the top open-source model across all three long-horizon coding benchmarks. MIT license, no regional restrictions.

Why it matters: Z.AI open-sources GLM-5.2 with a 1M-token context window and two new architectural components, explicitly targeting long-horizon agent tasks. Domestic flagship model release gets full weight per policy, but the body excerpt lacks full benchmarks, capping it below 85.

Jun 16Tuesday

OpenAI News

OpenAI simulates real-world deployment to catch undesired model behavior before release

OpenAI replays recent real conversations through a candidate model before release, then checks for new undesired behaviors. Across GPT‑5‑Thinking deployments, this Deployment Simulation gave more accurate frequency estimates than traditional evals, surfaced novel misalignment, and reduced the chance models could tell they were being tested. It also works for agentic rollouts with tool use. The method can’t catch behaviors rarer than 1 in 200,000 messages.

Why it matters: OpenAI published a concrete safety-testing method with a paper and reproducible workflow ahead of the GPT-5-Thinking release — not just a vague 'we did safety testing.' The method carries real information gain and hits the concerns of alignment practitioners. Not scored higher...

Jun 10Wednesday

OpenAI News

OpenAI banned PRC-linked ChatGPT accounts running covert influence ops on US AI debates

OpenAI published a threat report on June 10 detailing two clusters of ChatGPT accounts likely originating from China, both banned for covert influence operations. One cluster, named 'Data Center Bandwagon,' generated posts claiming AI data centers were raising household electricity prices. The other, 'Tech and Tariffs,' criticized US tariffs as tech competition tactics and instructed outputs to mention only President Trump, not Xi Jinping. That second cluster also spread false claims of a ChatGPT user data breach, which OpenAI calls entirely fabricated. OpenAI found no evidence the operations shifted public opinion, but sees them as testing narratives against US AI infrastructure. The post does not disclose account counts, target platforms, or reach metrics.

Why it matters: OpenAI's official threat report with concrete operational details and account clusters. Hits all three HKR axes, but as a security incident disclosure rather than a product/tech breakthrough, it lands in the 78-84 'good quality' band. Not scored higher because it doesn't resha...

NVIDIA Blog

NVIDIA confidential computing will help Apple expand Private Cloud Compute

Apple is bringing NVIDIA's confidential computing into Private Cloud Compute, running AI inference inside encrypted GPU environments. The setup uses H100 GPUs and Hopper architecture with hardware-level trusted execution environments, so data stays encrypted during processing and even the cloud provider can't access it. Apple previously ran private cloud inference only on its own silicon; this deal signals a shift of some workloads to NVIDIA while keeping the same security isolation. The post doesn't give a launch date or scale numbers, but confirms deployment will start in Apple's own data centers.

Why it matters: Apple is moving Private Cloud Compute workloads to NVIDIA H100 for the first time, using hardware-level TEEs to keep inference data encrypted while switching the compute substrate. Score isn't higher because the post gives no launch timeline or deployment scale — it's a direct...

Jun 1Monday

OpenAI News

OpenAI banned a likely PRC-origin cluster using ChatGPT to generate anti-US-data-center social media content

OpenAI's June threat report details a banned cluster of ChatGPT accounts likely originating in China. The operators used Simplified Chinese prompts to generate English posts and images on X, posing as ordinary Americans and claiming data centers and AI are driving up electricity costs for households. They also used ChatGPT for image editing, automation scripts, and harassing overseas dissidents. An internal work report they uploaded outlined tactics for building credible personas on Facebook and evading platform detection.

Why it matters: OpenAI's official threat report details a likely PRC-linked AI influence op with concrete tradecraft and a topic — AI driving up living costs — that's already a public flashpoint. Hits all three HKR axes, but as a security incident report rather than a product or research brea...

OpenAI News

OpenAI bans PRC-linked accounts using ChatGPT to generate comments on US tech policy and tariffs

OpenAI's June 2026 threat report details a banned cluster of ChatGPT accounts likely originating in China. The operators used Simplified Chinese prompts and VPNs to generate English comments and political cartoons criticizing US tariffs, rare earths, AI, and 5G policy. They instructed the model to depict only Trump, not Xi Jinping or China. The same cluster produced Chinese-language water army content attacking the US and Israel, amplifying anti-Jewish tropes, and harassing dissidents. OpenAI also linked these accounts to a separate X network that falsely claimed ChatGPT user data was compromised. The post does not disclose the exact number of banned accounts or the operators' specific institutional affiliation.

Why it matters: OpenAI's official threat report names a PRC-origin AI influence operation with concrete details and high topic sensitivity. Hits all three HKR axes, but it's a security incident report rather than a product/tech breakthrough, placing it in the 78-84 band per policy.

Apr 25Saturday

OpenAI News

OpenAI's jobs framework maps 921 occupations into four transition paths

OpenAI's AI Jobs Transition Framework sorts ~148M US jobs across 921 occupations into four paths: 18% at higher automation risk, 24% likely to reorganize, 12% could grow with AI-driven demand, and 46% face less immediate change. It goes beyond task exposure by asking whether a person remains central to delivery and whether lower costs expand demand. ChatGPT usage is roughly 3x higher in the most at-risk occupations, but recent unemployment shifts don't line up neatly with technical exposure. The report stresses that capability doesn't equal instant displacement—employers still need to rework workflows and weigh costs.

Why it matters: OpenAI drops a jobs-transition framework classifying 921 occupations and 148M US jobs into four automation-risk paths — 18% high risk, 24% restructured. The framework has analytical substance, not pure PR. Docked slightly because it's a policy-advocacy report rather than a pro...

Apr 16Thursday

OpenAI News

Codex for (almost) everything

OpenAI published a post titled "Codex for (almost) everything." The provided content has no body text, so the only confirmed facts are the mention of Codex and the phrase "almost everything," which is not enough to verify features, timing, or scope.

Why it matters: Major OpenAI product release for a huge installed base: Codex moves from coding assist toward a computer-using, memory-bearing agent across the dev lifecycle. HKR-H/K/R all pass, but the excerpt is truncated; pricing, rollout, and permission details are still missing, so it lands

Feb 1Sunday

OpenAI News

OpenAI banned accounts tied to a likely Russia-linked op using ChatGPT to write Africa-focused criticism of the US and allies

OpenAI banned a ChatGPT account linked to a previously unreported operation it calls 'No Bell,' likely originating from Russia. The account generated long-form articles and social media posts about sub-Saharan Africa, praising Russia and criticizing the US, UK, and Ukraine. The user prompted in English with occasional Russian instructions, asked the model to avoid em-dashes and write like a human journalist to hide AI involvement. Content appeared under the fake byline 'Dr. Manuel Godsin,' whose academic credentials could not be verified. Meta banned the associated Facebook Pages, some of which showed admin locations in Russia and Ukraine.

OpenAI News

OpenAI disrupts Cambodia-based romance-task scam network using ChatGPT

OpenAI banned a cluster of ChatGPT accounts and one API customer likely operating out of Cambodia. They ran a semi-automated romance-task scam targeting Indonesian men, using AI for ad copy, flirtatious chat, and translation. The workflow moved victims from social media ads to Telegram, then pressured them into paying escalating 'mission' fees. Internal scammer reports pasted into ChatGPT claimed hundreds of targets and thousands of dollars a day, though OpenAI could not independently verify those figures.

OpenAI News

OpenAI banned an account linked to Chinese law enforcement that used ChatGPT to plan a covert influence operation targeting Japan's prime minister

OpenAI 在 2026 年 2 月的安全报告里披露了一个案例。一个关联中国执法人员的 ChatGPT 账号,试图让模型帮忙设计一套搞臭日本首相高市早苗的方案,模型没答应。但用户后来上传了行动进展报告让模型润色,说明这事绕过 ChatGPT 照样推进了。报告里提到他们动用了 #右翼共生者 这类标签,在 X、Pixiv 等平台用小号发帖,内容集中在攻击...

Why it matters: OpenAI's own safety report case study with concrete op details — not a generic warning. Downside: it's one case in a larger report, and the model refused the request, so the news hook is softer than an active breach. But the geopolitics + AI misuse combo is strong signal for p...

OpenAI News

OpenAI banned Russia-linked ChatGPT accounts used to mass-produce multilingual influence content

OpenAI disclosed 'Fish Food' on Feb 1, banning ChatGPT accounts tied to the Russia-linked Rybar network. The accounts used Russian prompts to mass-generate content in English, Spanish, and other languages, posted via unaffiliated Telegram and X accounts. One batch of 7 tweets hit over 150K views for the top post from an account with 600K+ followers. The actor also drafted Africa election-interference proposals, with the priciest plan budgeted at $600K a year.

Why it matters: An official OpenAI disclosure of a malicious-use case with a named operation and adversary entity — high information density. The deduction is because this is a February re-publication, so timeliness is weakened, and it's an excerpt from a broader safety report rather than a s...

OpenAI News

OpenAI bans Cambodia-based accounts using ChatGPT to impersonate lawyers and the FBI in recovery scams

OpenAI's February safety report details 'Operation False Witness': a cluster very likely based in Cambodia used ChatGPT to create content for at least six fake law firms and to impersonate the FBI's IC3 unit. They lured fraud victims via social media ads, then used ChatGPT to draft lawyer-style messages on Telegram, demanding upfront crypto payments—such as a 15% service fee—before any recovery. The scammers also generated fake New York State Bar membership cards and bogus confidentiality agreements. OpenAI says individual victims may have lost thousands of dollars but cannot independently verify the amounts. The FBI and at least one impersonated firm have issued public alerts.

Why it matters: An official safety disclosure from OpenAI with concrete tactics and law enforcement tie-in, not a generic safety statement. But it's an excerpt from a broader report, not a standalone product launch — lands right at the featured threshold.

Oct 1, 2025Wednesday

OpenAI News

OpenAI banned accounts using ChatGPT to build Russian-speaking malware tooling

OpenAI's October threat report details banned ChatGPT accounts linked to Russian-speaking criminal groups. The operators used the model to prototype malware loaders, credential stealers, RAT components, and evasion layers. Direct malicious requests were refused, so they elicited building-block code and assembled it into attack workflows offline. OpenAI found no evidence the models provided novel capabilities and shared indicators with industry partners.

Why it matters: An official OpenAI threat report with concrete TTPs on how Russian-speaking groups bypassed safety filters by breaking malware construction into small code requests. It's a routine threat disclosure, not a new model or capability, so it lands at the featured threshold. Securit...

OpenAI News

OpenAI bans accounts using ChatGPT for phishing and scripting support tied to PRC intelligence requirements

OpenAI's October threat report details banned ChatGPT accounts whose activity overlapped with publicly tracked groups UNK_DROPPITCH and UTA0388. The actors used the model to draft phishing emails in Chinese, English, and Japanese, and to assist Go and PowerShell malware development—including encrypted C2 and process enumeration. The model introduced no novel offensive capabilities; the operators sought incremental speed and localization, yet left giveaway errors like implausible contact details in email signatures. They also explored DeepSeek for automating mass phishing, though OpenAI cannot confirm whether that work proceeded.

Why it matters: OpenAI's official threat report names specific actor overlaps (UNK_DROPPITCH, UTA0388) and concrete TTPs. Strong cross-source signal, but it's threat intel rather than AI capability news—direct value for product/research readers is limited, so it lands right at the featured th...

OpenAI News

OpenAI banned a covert PRC-linked influence network using its models to mass-produce posts on the South China Sea and Hong Kong

OpenAI's October 2025 report details a covert influence network, dubbed 'Nine-emdash Line,' that used ChatGPT to generate English and Cantonese social media posts. The content attacked Vietnam's environmental record in the South China Sea, smeared Philippine President Marcos, and discredited Hong Kong pro-democracy figures. The operators also used the models for recon—finding niche forums and generating Tibetan names for fake accounts—and sought advice on launching TikTok challenges to hijack trending hashtags. OpenAI banned the accounts; many X accounts were suspended. No technical link to the known 'Spamouflage' operation was found.

Why it matters: OpenAI's first public naming and dissection of a PRC-linked influence op, with concrete recon and cross-language generation details. Capped at 78 because it's a threat intel summary, not a product/model update — less pull for readers who only track technical progress.

Jun 1, 2025Sunday

OpenAI News

OpenAI banned accounts using ChatGPT for social engineering and fake personas

OpenAI's June threat report details 'VAGue Focus': a cluster of ChatGPT accounts used to generate social media posts, translate phishing-style messages, and pose as Europe- and Turkey-based consultancies for intelligence collection. The accounts operated mostly during mainland China business hours with Chinese prompts. They impersonated 'Focus Lens News,' 'Visionary Advisory Group,' and others, cold-messaging journalists and researchers on X, and claimed to offer $2,000 per hour for interviews. Public engagement was near zero; the one account with 17K followers was likely compromised and repurposed. OpenAI banned the network. The post does not identify who was behind it.

Why it matters: OpenAI's official threat report names a specific operation with entities, tactics, and geopolitical markers — dense enough. Score capped below 78 because the excerpt is a summary; the full report's technical depth isn't surfaced here.

OpenAI News

OpenAI bans accounts tied to AI-generated fake remote-job applications

OpenAI's June threat intel report details banned ChatGPT accounts linked to fraudulent remote-job campaigns. Operators used the models to mass-generate tailored résumés, answer interview questions, and research tools like Tailscale and OBS to mask remote laptop locations. The behavior matches publicly attributed North Korean IT worker schemes, with some collaborators possibly receiving corporate laptops inside the US. The post doesn't quantify how many companies were affected or the financial damage, but OpenAI shared the findings with industry peers and authorities.

Why it matters: Official threat intel from OpenAI with concrete tactics and an operational chain, not a generic safety reminder. Hits all three HKR axes, but as a security incident report rather than a product/model update, it caps in the 78-84 band.

OpenAI News

OpenAI banned accounts using ChatGPT to mass-produce Philippine political comments

OpenAI's June report details a takedown named 'High Five.' A Philippine marketing firm, Comm&Sense Inc, used ChatGPT to mass-produce short English and Taglish comments praising President Marcos and attacking VP Duterte on TikTok and Facebook. The workflow included analyzing political posts, generating sub-10-word comments, and running five coordinated TikTok channels. OpenAI banned the accounts; the actor tried to return multiple times. Thousands of comments were posted, but none got more than single-digit engagement.

Why it matters: An official OpenAI disclosure with a named company and specific TTPs — not a vague 'we banned some accounts' statement. Score capped here because it's one case study inside a monthly report, not a standalone policy or product update, and the post doesn't disclose ban volume or...

Feb 1, 2025Saturday

OpenAI News

OpenAI bans China-linked accounts that used ChatGPT to plant anti-US articles in Latin American media

OpenAI banned ChatGPT accounts likely tied to China that generated English posts attacking dissident Cai Xia and Spanish-language articles criticizing the US. The Spanish articles appeared on news sites in Peru, Mexico, and Ecuador, some labeled as sponsored content, with bylines pointing to a Jilin-based company. OpenAI says this is the first observed case of a China-origin influence operation successfully placing long-form articles in Latin American mainstream media, rating it Category 4 on the Breakout Scale. Social media engagement was minimal; the paid articles may have reached a wider audience.

Why it matters: OpenAI's official disclosure names a real company and provides operational details, denser than routine transparency reports. Score capped because this is a Feb 2025 re-run—would be 82-84 if fresh.

OpenAI News

OpenAI banned a Cambodia-based cluster using ChatGPT for pig-butchering scams

OpenAI banned a cluster of ChatGPT accounts originating in Cambodia that were used to translate and generate romance-investment scam conversations in Japanese, Chinese, and English. The scammers targeted men over 40 on Facebook, X, and Instagram using stolen influencer photos, then moved chats to LINE or WhatsApp within days. OpenAI reconstructed a six-step workflow from public engagement to fraudulent investment, noting the actors provided the model with detailed fake personas and used it mainly for translation and flirty replies.

Why it matters: An official OpenAI threat intel case study reconstructing a Cambodia-based scam ring's full AI-assisted pig-butchering pipeline, with concrete victim profiles and platform paths. The ding is that this is a Feb 2025 report — timeliness takes a hit — and it's a security ops disc...

OpenAI News

OpenAI banned China-linked accounts using ChatGPT for surveillance-tool pitches and document analysis

OpenAI disclosed in Feb 2025 that it banned a cluster of ChatGPT accounts likely from China, dubbed “Peer Review.” The operators used the models to analyze English document screenshots, draft sales pitches for a “Qianyue Overseas Public Opinion AI Assistant,” and debug related code. The tool claimed to scrape X, Facebook, and other platforms to spot China-related protest calls and report them. Code debugging primarily invoked Meta’s Llama 3.1 8B, with references to Alibaba’s Qwen and an unspecified DeepSeek model. OpenAI found no evidence the generated content was posted publicly and said impact assessment requires input from other model providers.

Why it matters: Official threat intel from OpenAI with a named operation and adversary TTPs — solid policy/safety crossover. Downside: it's a Feb 2025 re-run with no new angle, and it's a single-source narrative without third-party corroboration.