Skip to content

#xAI

3 today

Jun 28Sunday

AI HOT (Curated Pool)

SpaceX files SpaceXAI trademark, Musk says xAI will merge into SpaceX

SpaceX has filed a trademark for SpaceXAI. Musk says xAI will dissolve as a standalone company and become SpaceX's AI product line. The post doesn't disclose a merger timeline, team structure, or what happens to existing xAI products like Grok.

Why it matters: xAI folding into SpaceX is a structural shift, not a routine product update. The trademark filing provides hard evidence, but the post lacks timeline and team integration details, capping the score below 85.

Jun 24Wednesday

Hacker News front page

Reid Hoffman: SpaceX is 'not an AI company,' xAI is a 'complete train wreck'

LinkedIn co-founder Reid Hoffman called SpaceX 'not an AI company' and xAI 'a complete train wreck' on a podcast. He said SpaceX's post-IPO Cursor acquisition is buying relevance, and its compute leasing is just 'a premium-priced CoreWeave.' On xAI, all 11 co-founders have left and the company is on its third restart. Hoffman also criticized the U.S. government's forced takedown of Anthropic's Fable and Mythos models as 'autocratic willy-nilly,' troubled by the asymmetry with OpenAI. He invests in both Anthropic and OpenAI and sees room for both to win.

Why it matters: A well-known investor publicly sizes up the AI landscape with specific claims and sharp language — all three HKR axes hit. Deduction because it's a podcast opinion, not a product/research release with verifiable new capabilities, so it lands at the 78 featured threshold.

Jun 22Monday

AI HOT (Curated Pool)

Grok Build adds /goal mode for long-running autonomous task execution

xAI added /goal to Grok Build: give the agent an objective and it plans, breaks work into a checklist, and executes until done. You can check status, pause, resume, or clear the goal mid-run. The post doesn't disclose max run time, resource costs, or specific pricing.

Why it matters: xAI added /goal mode to Grok Build, letting the agent autonomously complete a task — similar in shape to Cursor Agent and Claude Code's long-running execution. Concrete interaction details are present, but the post doesn't disclose max runtime, resource consumption, or extra p...

Jun 19Friday

Latent Space

Anjney Midha on AI compute waste: frontier labs run sub-10% MFU, AMP plans an independent compute grid

Anjney Midha discusses hidden AI infrastructure waste on Latent Space. xAI's training MFU is under 10%, while Google treated 95% utilization as an outage; best-in-class today is 60–70%. He invested in Anthropic, Mistral, and Black Forest Labs, and now runs AMP, aiming for a 1.2 GW base-load compute grid with 6 GW spike capacity. He also flags DeepMind's unpublished research as a market failure and notes Anthropic prioritized coding as P0 from day one. The post does not disclose a timeline for AMP's grid.

Why it matters: Anjney Midha puts a specific number on xAI's training MFU (under 10%), turning vague complaints about compute waste into a quantifiable discussion. Score capped at 78 because it's a podcast opinion, not a product launch or paper — no reproducible verification path.

Jun 18Thursday

Hacker News front page

OpenRouter ran 11 LLMs in a 30-game battle royale — Grok 4.1 Fast won 43%

OpenRouter's Jacky Liang dropped 11 LLMs into a 2D battle royale for 30 matches. Grok 4.1 Fast won 13 games at $0.97 per win; Claude Sonnet 4.6 won 5 at $26.78 per win — a 27x gap. GPT 5.4 had the most kills (38) but only 2 wins, so killing more didn't mean winning more. GPT 5.4-mini, DeepSeek 4 Flash, and Kimi K2.6 spent $57 combined and won zero games. The models reasoned, called tools, and updated memory each turn — they weren't just generating control code. The post doesn't provide the full leaderboard or detailed behavioral differences across all models.

Why it matters: OpenRouter's official blog, author Jacky Liang ran 30 games himself with full data and replays. Grok 4.1 Fast's cost advantage is stark, Claude Sonnet 4.6 is expensive but consistent, GPT 5.4 is the kill leader but can't close — all three takeaways are concrete and verifiable....

Jun 16Tuesday

AI HOT (Curated Pool)

DOJ invokes national security to defend xAI's unpermitted gas turbines in NAACP lawsuit

The DOJ moved to dismiss an NAACP lawsuit against xAI, arguing that shutting off its gas turbines would threaten military operations. A DOD official stated Grok is one of four models supporting mission-critical work on classified networks, including recent strikes on Iran. The NAACP sued because xAI runs unpermitted turbines at its Colossus 2 site in Mississippi—turbine count grew from 27 to 57 since April, with a 111% spike in nitrogen oxide emissions. The post doesn't specify which national security statute the DOJ is citing.

Why it matters: xAI's unpermitted data center emissions draw a NAACP lawsuit, and DOJ steps in citing national security, claiming shutting down the turbines would impact military operations. Concrete numbers and Pentagon backing give this both novelty and substance, but it's still in litigati...

AI HOT (Curated Pool)

SpaceX to acquire AI coding startup Cursor for $60B in stock, days after its IPO

Days after its historic IPO, SpaceX agreed to buy AI coding startup Cursor for $60 billion in stock. Cursor was about to close a $2B round at a $50B valuation from a16z, Thrive, and Nvidia. SpaceX told IPO investors its AI addressable market is $26 trillion and wants the deal to help its xAI-built AI unit catch up with major labs. The transaction is expected to close in Q3. The post doesn't spell out product integration plans, team retention, or regulatory approvals.

Why it matters: SpaceX acquiring Cursor for $60B in stock immediately after IPO is an industry-shaking move. Cursor was about to close a $2B round at a $50B valuation — this deal rewrites the AI coding tools landscape overnight. HKR all hit; the only deduction is that the body doesn't disclos...

AI HOT (Curated Pool)

xAI launches Grok Imagine Video 1.5: faster image-to-video with synced audio

xAI upgraded its image-to-video model to 1.5, now generally available via API with a Fast variant on Grok and mobile apps. A 6-second 720p clip takes about 25 seconds, nearly twice as fast as the previous 40+ seconds. Audio is generated in the same pass—ambience, effects, and dialogue land on the action with better lip sync. Motion holds up over longer clips with fewer warps and more believable weight. Three workflow features are rolling out: Projects for organization, parallel multi-agent prompting, and library search. The post doesn't disclose training data scale or pricing.

Why it matters: xAI shipped a meaningful image-to-video upgrade with 2x speed, synced audio, and better physics. Score stays at 78 because the video generation space already has established leaders — this is a solid iteration, not a market shift.

Jun 15Monday

AI HOT (Curated Pool)

Grok Build adds Agent Dashboard to manage multiple coding sessions at once

xAI shipped a terminal dashboard for Grok Build that lets you monitor and interact with multiple coding sessions on one screen. Sessions are grouped by state—blocked ones rise to the top—so you handle approvals and questions inline without switching contexts. You can peek at output, reply, dispatch new work, and jump into any session. Closing the dashboard leaves everything running; reopening it restores all sessions. Install via curl -fsSL https://x.ai/cli/install.sh | bash, then run grok dashboard or hit Ctrl+\.

Why it matters: xAI added a terminal-based multi-session dashboard to Grok Build with a novel interaction pattern and concrete mechanism details. But this is a single-feature update, not a model release or ecosystem-level shift — impact is limited to Grok Build users. H and K both hit, R is a...

Jun 11Thursday

AI HOT (Curated Pool)

xAI launches a built-in plugin marketplace for Grok Build, starting with 6 partners including MongoDB and Vercel

xAI added a plugin marketplace directly inside Grok Build, so you can browse, install, and update plugins without leaving the terminal. Each plugin bundles skills, slash commands, agents, hooks, MCP servers, and LSPs. The launch lineup includes MongoDB, Vercel, Sentry, Chrome DevTools, Cloudflare, and Superpowers. Every remote plugin is pinned to a specific commit SHA and verified at install time. Developers can publish their own by submitting a PR to xai-org/plugin-marketplace. The post doesn't mention any review process or revenue split.

Why it matters: xAI added a built-in plugin marketplace to Grok Build with launch partners like MongoDB, Vercel, and Sentry. Bundling MCP servers and LSPs into one installable unit is a more structured extensibility model than typical CLI tools. H and K both hit, but the audience is limited t...

Jun 9Tuesday

Bloomberg Technology

Musk’s xAI Taps Starlink Staffer to Run Grok Training Team

xAI brought in an executive from SpaceX’s Starlink service to run the Grok training team, replacing college-aged engineer Diego Pasini; the RSS snippet does not disclose the executive’s name, tenure, or training process.

Why it matters: HKR-H/K/R pass, but this is a training-team leadership change, not a model release or executive-level departure. Bloomberg sourcing and the Diego Pasini detail put it at the featured floor.

Jun 6Saturday

AI HOT (Curated Pool)

SpaceX and Google Reach New Cloud Computing Agreement

SpaceX disclosed a cloud services agreement with Google: Google will pay SpaceX $920 million per month for computing capacity tied to xAI data centers, while the post does not disclose contract duration, GPU scale, or delivery terms.

Why it matters: HKR-H/K/R all pass: the hook is a Google–SpaceX–xAI compute triangle, with $920M/month as the concrete fact. The single-post source and missing contract term, delivery scale, and filing details keep it at low P1.

Hacker News front page

Google to Pay SpaceX $920M a Month for Compute Capacity at xAI Data Centers

The title says Google will pay SpaceX $920 million per month for compute capacity at xAI data centers; the RSS snippet does not disclose contract duration, GPU scale, or the capacity delivery mechanism.

Why it matters: HKR-H/K/R all pass: $920M/month is a hard compute-market number, and the Google-SpaceX-xAI structure is unusual. Missing duration and GPU details keep it below 90.

Jun 5Friday

Computing Life · Share · Yage

Grok Build 0.1: xAI’s Bet on Parallel Breadth

xAI launched Grok Build 0.1 in May 2026 as a coding agent built around parallel subagents; the post does not disclose benchmark results, cost figures, or specific privacy-policy terms.

Why it matters: HKR-H/K/R pass because xAI entering coding agents with parallel subagents is clickable, concrete, and relevant to developers. Missing benchmarks, cost, and privacy terms keep it at the featured floor.

Jun 4Thursday

Financial Times · Technology

MP sues Musk’s xAI in UK test case over fake sexual images

UK MP Jess Asato sued Musk’s xAI over fake sexual images, using the claim to test whether AI model makers are liable for system outputs; the post does not disclose the model, generation mechanism, damages sought, or court timetable.

Why it matters: HKR-H/K/R all pass: FT ties xAI, Musk, fake sexual images, and a UK liability test. The article does not disclose the model, generation mechanism, or damages, so it stays in the 78–84 band.

Jun 3Wednesday

AI HOT (Curated Pool)

Grok Becomes Vapi's Default Voice Engine

xAI partnered with Vapi to make Grok the default engine for 12 core voices, covering more than 2.5 million voice agents, and Grok Voice ranked first in Vapi’s independent blind test.

Why it matters: HKR-H/K/R all pass: the default-engine switch has scale, numbers, and voice-agent market resonance. Single-source partnership news lacks test methodology, pricing, and migration data, so it stays in the mid product-update band.

AI HOT (Curated Pool)

xAI releases Grok Imagine 1.5 preview image-to-video model

xAI released grok-imagine-video-1.5-preview via its API, letting users turn one still image into 720p video while controlling camera movement, pacing, and sound effects with natural-language prompts.

Why it matters: HKR-H/K/R all pass: xAI ships a named image-to-video API preview with 720p output and sound controls. It stays below 85 because this is a preview product update, not a flagship foundation-model release.

Jun 1Monday

Latent Space

Why Video Agent Models Are Next — Ethan He on xAI Grok Imagine

Ethan He says a small xAI team built Grok Imagine from zero to one in 3 months, and the episode discusses video agents, audio-video alignment, inference speedups, and the storage, egress, and GPU-hour costs behind large video datasets.

Why it matters: HKR-H/K/R all pass, but the body is interview-level signal: beyond the 3-month build and mechanism themes, it gives no benchmarks, cost figures, or reproducible test. Strong xAI video-agent context, not same-day must-write.

May 30Saturday

AI HOT (Curated Pool)

xAI drops JAX GPU for an in-house training framework

SemiAnalysis says xAI dropped JAX GPU and moved to a C training framework written with Grok Build; the snippet claims xAI’s JAX stack had MFU below 10%, but the post does not disclose reproducible benchmark conditions.

Why it matters: HKR-H/K/R all pass: xAI changing its training stack is a strong hook, MFU <10% is a concrete claim, and infra cost will spark debate. Single-source tweet format and no reproducible setup keep it at 80, not P1.

AI HOT (Curated Pool)

xAI Releases Grok Build 0.1 Public Beta

xAI released grok-build-0.1 as a public beta through its API; the same model powers the Grok Build CLI, targets agentic coding, and is priced at $1 per million input tokens and $2 per million output tokens.

Why it matters: HKR-H/K/R all pass, but the post is thin: beta, CLI, pricing, and agent-coding positioning only; no benchmarks, context window, or hands-on results. This fits a mid-weight product update.

May 28Thursday

AI HOT (Curated Pool)

Grok Build 0.1 on API

xAI released Grok Build 0.1 in public beta through the xAI API for agentic coding tasks, with throughput above 100 tokens per second and pricing at $1 per million input tokens and $2 per million output tokens.

Why it matters: HKR-H/K/R all pass, but this is a 0.1 public-beta API and pricing launch; benchmarks, context window, and task success rates are not disclosed. It fits a solid mid-weight product update at 78, featured not p1.

May 26Tuesday

Synced · WeChat

Grok keeps updating after xAI disbandment as Musk announces a new model

Elon Musk said the 1.5T-parameter Grok V9-Medium has finished training, will enter reinforcement learning in a few days, and is planned for release in two to three weeks. Grok Build supports up to 8 parallel sub-agents, a 256K-token context window, Plan Mode, Arena Mode, MCP, and ACP.

Why it matters: HKR-H/K/R all pass, but this is a Grok V9-Medium preview before RL and release, with no benchmarked capability yet. That fits a strong model-race/product update at 82, featured but not p1.

AI HOT (Curated Pool)

Grok Build Beta Opens to SuperGrok Users

xAI opened Grok Build Beta to all SuperGrok and X Premium+ users, with Plan Mode, Imagine-based image and video creation, and a CLI for automation or orchestrator workflows at x.ai/cli.

Why it matters: HKR-H/K/R all pass: xAI opened a paid beta with named workflow features. The score stays at the featured floor because the post lacks capability limits, pricing detail, and test results.

May 21Thursday

Financial Times · Technology

Anthropic on Track for First Profitable Quarter

Anthropic is on track to record its first profitable quarter ahead of OpenAI and xAI; the RSS snippet does not disclose the quarter, revenue, profit figure, or accounting basis.

Why it matters: HKR-H/K/R all pass: the FT claim reframes Anthropic’s business race against OpenAI and xAI. Missing quarter, revenue, and profit figures keeps it below P1.

AI HOT (Curated Pool)

xAI burned $6.4B last year; SpaceX IPO filing shows why spending is far from over

SpaceX’s IPO filing disclosed that xAI lost $6.4 billion in 2025 and plans a large Grok expansion; the post does not disclose the expansion scale, capital spending, or financing terms.

Why it matters: HKR-H/K/R all pass: the $6.4B loss is a hard number, SpaceX’s IPO filing gives the story an unusual source, and xAI’s Grok expansion hits the burn-rate debate. Missing capex, scale, and financing keep it below 85.

TechCrunch · AI

Musk’s xAI is being sued over data center generators, now buying $2.8B more

xAI said it will buy $2.8 billion of natural gas turbines over the next three years, according to SpaceX’s IPO filing; the title says xAI is being sued over data center generators, but the post does not disclose the plaintiff, claims, court, or turbine supplier.

Why it matters: HKR-H/K/R all pass: xAI buying $2.8B in turbines while facing generator litigation has contrast, a concrete number, and AI-infra stakes. Missing plaintiff, claims, and supplier keep it in low featured.

r/LocalLLaMA

HalBench: Custom sycophancy and hallucination benchmark tests 4 frontier models

HalBench tested 4 frontier models on 3,200 false-premise prompts, with Sonnet 4.6 ranking first at a 0.565 mean score and Gemini 3.1 Pro last at 0.339; higher scores mean the model more often named the false premise and pushed back instead of complying.

Why it matters: HKR-H/K/R all pass: HalBench has a clear custom-eval hook, 3,200 prompts with scores, and a live trust/safety angle. Single Reddit sourcing and an unvalidated benchmark keep it at the low featured band.

TechCrunch · AI

Anthropic will pay xAI $1.25B per month for compute

Anthropic will pay xAI $1.25 billion per month for compute; the post discloses the deal value but does not disclose compute scale, contract length, or deployment conditions.

Why it matters: HKR-H/K/R all pass: TechCrunch reports Anthropic will pay xAI $1.25B per month for compute, a striking counterparty and cost signal. Missing scale, term, and deployment details keep it below the 90s.

May 18Monday

AI HOT (Curated Pool)

Grok launches Skills feature

xAI launched Grok Skills on May 18, 2026, letting users set preferences, formatting rules, or workflows once and keep them active across all conversations on web, iOS, and Android.

Why it matters: HKR-H/K/R all pass: Grok Skills adds persistent preferences and workflows across web, iOS, and Android. This is a mid-weight xAI product update; rollout scope, limits, and pricing are not disclosed.

May 15Friday

AI HOT (Curated Pool)

Connect Grok to the Hermes Agent

xAI connects Grok subscription accounts to Nous Research’s open-source Hermes Agent across all subscription tiers, letting users run Grok 4.3 text chat and reasoning, generate spoken replies with text-to-speech, create images and videos with Grok Imagine, and connect the agent to WhatsApp or Discord.

Why it matters: HKR-H/K/R all pass, but this is a mid-weight xAI product integration with an open-source agent, not a flagship model release. Featured fits; it does not clear the 85+ same-day bar.

Bloomberg Technology

Musk’s xAI Unveils First Coding Agent in Bid to Rival Anthropic

xAI is rolling out its first AI coding agent, Grok Build, for software development workflows; the RSS snippet names Anthropic’s Claude as the rival but does not disclose pricing, availability, benchmarks, or supported IDEs.

Why it matters: HKR-H and HKR-R pass: xAI entering coding agents is a strong competitive hook for developers. HKR-K fails because pricing, availability, and benchmarks are not disclosed, so this stays at the low end of a mid-weight product update.

May 14Thursday

AI HOT (Curated Pool)

xAI launches early beta of Grok Build

xAI launched an early beta of Grok Build for SuperGrok Heavy subscribers, offering a terminal-based coding agent with plan review, parallel subagents for large tasks, and a headless mode for scripting and automation.

Why it matters: HKR-H/K/R all pass: xAI enters terminal coding agents with plan mode, parallel subagents, and headless mode. Early beta access for SuperGrok Heavy keeps it below the 85 same-day must-write band.

TechCrunch · AI

Musk’s xAI Is Running Nearly 50 Gas Turbines Unchecked at Its Mississippi Data Center

Musk’s xAI is running nearly 50 gas turbines at its Colossus 2 data center in Mississippi, and the company faces a lawsuit over using “mobile” gas turbines as power plants; the RSS snippet does not disclose permitting details or the lawsuit’s specific claims.

Why it matters: All HKR axes pass: a sharp conflict hook, concrete turbine/lawsuit facts, and resonance around AI data-center power compliance. Not a model or product release, so it stays near the featured threshold.

May 11Monday

QbitAI · WeChat

SpaceXAI Takes Shape as Elon Musk Files Trademark Applications

SpaceX filed two SpaceXAI trademark applications covering satellite-based data centers, orbital computing, AI SaaS, cloud storage, telecom hardware, and social networking; the post says xAI became a SpaceX subsidiary through an all-stock deal and cites a $250 billion xAI valuation.

Why it matters: HKR-H/K/R all pass, but the hard fact is trademark filings; the claimed xAI-SpaceX merger lacks disclosed deal terms or an official announcement. Featured, not 85+, because this is signal rather than confirmed restructuring.

May 10Sunday

AI HOT (Curated Pool)

SpaceXAI officially announced

Trademark filings show SpaceXAI submitted an application on May 6, 2026, with its status listed as pending review; the post says the date aligns with Elon Musk announcing xAI’s merger into SpaceX, but it does not disclose approval, product scope, or launch timing.

Why it matters: HKR-H/K/R all pass, but the post only provides a pending trademark filing, not deal terms, product shape, or the official announcement text. High attention, thin facts: featured, not p1.

May 7Thursday

Latent Space

Anthropic-SpaceXAI's 300MW/$5B/yr Deal for Colossus I, ARR Growth Is 8000% Annualized

Anthropic announced a SpaceX compute partnership, doubled Claude Code’s 5-hour limits for Pro, Max, Team, and seat-based Enterprise, raised Opus API limits, and said Claude inference would ramp on Colossus within days; the post treats the 300MW and $5B-per-year figures as widely circulated but not canonized in Anthropic’s own announcement.

Why it matters: HKR-H/K/R all pass: the compute-deal numbers and Claude Code limit changes are concrete and practitioner-relevant. The 300MW/$5B/year claim is unofficial, so it stays below P1.

Synced · WeChat

Musk Announces xAI Dissolution, Leasing 220,000 GPUs to Anthropic

Musk confirmed xAI will dissolve, with Grok and X-related operations folded into SpaceXAI. SpaceX and Anthropic signed a deal giving Claude access to Colossus 1’s 220,000+ Nvidia GPUs and 300 MW of compute. The key change is quota: Claude Code’s five-hour rate limit doubles, and Pro/Max peak-hour cuts are removed.

Why it matters: HKR all pass: xAI dissolution plus 220k GPUs for Anthropic is a top-tier twist; 300 MW and Claude Code quota changes add testable detail; it hits compute, competition, and developer limits. Single-source status keeps it at 96.

Computing Life · Share · Yage

Anthropic Locks Up Compute Channels as xAI Rents Its Castle to a Rival

Anthropic signed four compute contracts in six months covering AWS Trainium, Google TPU, SpaceXAI Colossus 1, and CoreWeave; during the same window, xAI rented the Colossus 1 supercomputing center to a competitor while GPU utilization stood at 11%.

Why it matters: HKR-H/K/R all pass: Anthropic’s four compute deals and xAI leasing Colossus 1 create a sharp competitive angle with concrete numbers. Single-source strategy analysis keeps it in the 78–84 band, below same-day must-write news.

May 5Tuesday

The Verge · AI

Google, Microsoft, and xAI Will Let the US Government Review New AI Models

Google DeepMind, Microsoft, and xAI agreed to let CAISI review new AI models before public release. CAISI says it will run pre-deployment evaluations and targeted research, after 40 reviews since 2024; the post does not disclose model names. The key issue is review scope and release timing, not the announcement alone.

Why it matters: HKR-H/K/R all pass: major labs accept US pre-release review, with CAISI citing 40 reviews since 2024. Specific model names and review criteria are not disclosed, so this stays below the must-write band.

Hacker News front page

Google, Microsoft and xAI Agree to Share Early AI Models with U.S.

Google, Microsoft and xAI agreed to share early AI models with the U.S. The snippet lists 3 companies, a WSJ link, an HN link, 5 points, and 0 comments. The post does not disclose the agency, model scope, review mechanism, or timeline.

Why it matters: HKR-H/K/R all pass, but the body discloses only the headline fact. Recipient agency, model scope, review mechanism, and timeline are missing, so this stays in the 72–77 featured band.