explainx / blog / topics
AI Chips and Infrastructure
Every model runs on chips, data centers, and power that are in short supply. NVIDIA results, new accelerators, data-center fights, and memory and power constraints decide how fast AI can grow.
This page tracks the hardware and infrastructure stories we covered.
115 stories · latest Oct 9, 2026
Start here
Gigabyte W775-V10-L01: A GB300 DGX Station-Class Workstation for Your Desk
ServeTheHome reviewed Gigabyte's W775-V10-L01, an NVIDIA DGX Station-class desktop built on the GB300 superchip. It has 252GB of HBM3E, 496GB of LPDDR5X, two 400Gbps ports and a 1.6kW power limit. explainx.ai explains what it is, who should buy one, and how it compares with DGX Spark, Mac Studio and RTX PRO towers.
What Is TPS? Tokens Per Second Explained for AI Models
Tokens per second, or TPS, is the headline speed number for AI models, and it is easy to misread. This explainer covers what a token is, how TPS is measured, what counts as fast, and which other numbers matter more for the way you actually use a model.
NVIDIA Inception: Free Cloud Credits and VC Intros for AI Startups
NVIDIA's Inception program offers AI startups free technical courses, SDK access, cloud credits, preferred hardware pricing, and investor exposure on Capital Connect — with no fees, no equity, and rolling applications. But the eligibility bar (an incorporated company with at least one developer) locks out solo, pre-incorporation builders, which is exactly what founders pushed back on when the program went viral on X this week.
Is Nvidia the "Central Bank of AI"? What ~$300B in Backstops Means
A widely discussed Economist briefing lays out roughly $300 billion in guarantees, backstops, and equity stakes Nvidia has extended to its own customers — neoclouds, AI labs, even Wall Street funds — to keep demand for its chips growing. Here's what the actual numbers show, why critics compare it to Cisco's dot-com-era vendor financing, and what it means for anyone paying for compute or choosing which model to build on.
Top 10 Neural Rendering Use Cases Beyond DLSS 5's Beauty Filter
The leaked DLSS 5 library turned Cyberpunk, GTA 5 and Dark Souls 3 into uncanny valley clips, and the discourse collapsed into "AI slop or not." That framing buries the actual question: what is a learned post-render stage genuinely good at? Ten use cases, ranked by how well the technique's strengths match the job.
AI Chip Architectures Explained: GPU vs TPU vs Trainium vs Cerebras vs Groq
AI accelerators all multiply matrices, but they disagree about nearly everything around the multiplication. This guide turns Jacob Peake's deep architecture survey into a decision framework: where data lives, who schedules it, how chips connect, and which workloads each design favors.
Timeline
October 2026
Oct 9
AI Chip Supply Chain This Week: TSMC Record, Samsung Phone Cut, GlobalFoundries InterposersThree stories broke on October 7-8, 2026, and they share one cause: AI demand is reshaping the chip supply chain. TSMC posted record revenue, Samsung is reportedly cutting phone parts as AI memory eats supply, and GlobalFoundries signed a $2 billion deal for US-made interposers. explainx.ai checks each against primary sources.
Oct 9
Atomic Machines Matter Compiler: AI-Native Fab for Micro-Machines and the PrimeSwitch RelayAtomic Machines came out of six years in stealth on October 7, 2026 with a manufacturing system it calls the Matter Compiler and its first device, the PrimeSwitch power relay. The company says the system builds working micro-machines from digital code with no molds or masks. The New York Times profiled it. No public demo of the fab exists yet.
Oct 9
Gigabyte W775-V10-L01: A GB300 DGX Station-Class Workstation for Your DeskOct 9
Long-WAM: NVIDIA Scales the Context of World-Action Models to Reach 78.7% on RoboCasaNVIDIA researchers released Long-WAM on October 7, 2026, a framework that lets a world-action model use up to 19.2 seconds of visual history while still running in real time. On RoboCasa GR-1 the success rate rises from 63.3% to 78.7%. The paper's central point is that having history is not the same as using it.
Oct 9
Lumentum Says AI Optical Parts Are Sold Out Through Early 2029On October 9, 2026, Lumentum CEO Michael Hurlston told Bloomberg that the company's optoelectronic capacity is sold out to nearly 2029, a year past what he said in April. This post explains the parts, the numbers, and the limits of what is confirmed.
Oct 9
Natura’s $99 Interface Ring Wants to Be Your Remote Control for AI AgentsNatura, founded by Carlo Edoardo Ferraris, announced Interface, a $99 smart ring that lets you ask AI agents to do tasks with a press of the finger. It also tracks heart rate, sleep and activity. Preorders open next month. A $9 monthly fee follows an initial free period.
Oct 9
Yandex Sasovo Data Center Drone Strike: 2,700 GPUs Claimed, What Is VerifiedUkrainian drones struck Yandex's Sasovo data center on the night of October 7-8, 2026. Headlines say 2,700 Nvidia GPUs were destroyed. Reuters, CNBC and Yandex confirm a fire and a shutdown, but not that count. explainx.ai separates the claim from the record.
Oct 7
Cloudflare Open-Sources Its Security Audit Skill: How the Six-Phase Agent Workflow WorksCloudflare published the single-repo skill that seeded its fleet-wide vulnerability discovery harness. It turns a coding agent into a security auditor across six phases with independent verification. Here is how it works, how to install it, and what developers say about cost and noise.
Oct 7
NVIDIA Nemotron Hits Gold-Level at IOI and IMO 2026: The Fine-Tuning Recipe and What Is OpenNVIDIA says one model family reached gold-medal level at both IOI 2026 and IMO 2026. This post covers the scores, the four-part recipe, the data sizes, and which checkpoints, datasets, and code are published. It also lists what NVIDIA has not shown.
Oct 5
What Is TPS? Tokens Per Second Explained for AI ModelsOct 3
Cloudflare Sandbox SDK 1.0: Your Durable Object Owns the ContainerOn September 30, 2026 Cloudflare shipped Sandbox SDK 1.0 and rebuilt Containers for agent sandboxes: your Durable Object picks image and size at start time, median cold start drops to 648 ms, and filesystem snapshots land in public beta. explainx.ai covers setup, pricing, migration from 0.x, and how this sits next to Artifacts and @cloudflare/computer.
Oct 3
Cloudflare Web Search API: Add Live Search to Any Agent Through AI GatewayOn October 2, 2026 Cloudflare put web search behind AI Gateway, with Ceramic.ai, Exa and Linkup as providers. One call from a Worker or the REST API returns titles, URLs and descriptions, billed at partner list prices with your normal gateway logging.
Oct 3
Prime Inference: Prime Intellect Opens Production Open-Model ServingOn October 2, 2026 Prime Intellect opened Prime Inference — the production serving stack behind its RL and agent workloads — as public serverless and reserved capacity for frontier open-weight models. First hosted model: GLM-5.3. Here is what builders should actually use it for this week.
Oct 1
Cloudflare Artifacts Open Beta: A Real Git Remote for AI AgentsOn October 1, 2026 Cloudflare put Artifacts into open beta on Workers Paid: programmable Git remotes for agents, Workers Builds and Previews on push, and a contest to build the next Git platform. explainx.ai covers pricing, limits, the docs lag, and what to wire into a coding-agent loop this week.
Oct 1
NVIDIA VSS 3.3: One Prompt Built a Juice-Line Vision AgentA September 30, 2026 NVIDIA tutorial walks a coding agent through vss-build-vision-ai: one English prompt, a 20-service stack, Cosmos watching 10-second camera chunks, Nemotron writing the incident report. The $3 is compose cost, not inference. Adaptive EVS is the runtime story.
September 2026
Sep 29
Nvidia–Anthropic $180B Contracted Value: What Builders PayOn September 28, 2026, Nvidia said contracted value with Anthropic exceeds $180 billion. That is a Nvidia-specific stack — GPU-backed compute, the November 2025 up-to-$10 billion equity arrangement, and IPO-anchor talks — not Anthropic's reported ~$517 billion multi-vendor compute total. This post maps the stack to Claude API capacity, rate limits, and vendor risk.
Sep 28
China MIIT May Clear ByteDance and Alibaba to Buy Nvidia RTX PRO 5500 Workstation GPUsOn September 27, 2026, Reuters published The Information''s reporting that China''s Ministry of Industry and Information Technology (MIIT) has signaled it may approve domestic purchases of Nvidia''s new RTX PRO 5500 Blackwell workstation GPU — with ByteDance and Alibaba named among the firms filing procurement plans. Reuters said it could not independently verify the story. explainx.ai maps the workstation-versus-datacenter export-control gap, what Nvidia lists on its product page (still "coming soon"), and what changes for teams serving open-weight inference if the orders land.
Sep 28
Jensen Huang Calls AI Distillation 'Competition' as Bessent Calls It TheftOn September 28, 2026, Jensen Huang told CNBC that training on rival models'' outputs is competition, contradicting Treasury Secretary Scott Bessent''s July theft framing and the same week''s CISA and Anthropic distillation warnings. explainx.ai maps the policy fight, Huang''s shovel-seller incentives, and practical defenses labs are already shipping.
Sep 28
NVIDIA Open Agent Safety Platform: OpenShell + Sentry for Agent TrustOn September 28, 2026, NVIDIA CEO Jensen Huang announced the Open Agent Safety Platform with 100-plus industry partners — OpenShell plus Sentry on BlueField-4. TechCrunch's Julie Bort (September 29) reports OpenAI is not a named public supporter even while a spokesperson said the lab is supportive and working on OpenShell. explainx.ai covers the reference design and what that absence actually is.
Sep 24
NVIDIA Nemotron 3 Diarization: An Open 100M-Parameter Model That Labels Up to Eight Overlapping Speakers[Speaker diarization](/dictionary/speaker-diarization), working out who spoke when, is the quiet piece behind every good meeting transcript and voice agent. NVIDIA's new Nemotron 3 Diarization is a 100-million-parameter open-weight model that handles up to eight speakers, overlapping speech and four streaming latency settings, and tops VoiceArena's diarization leaderboard at 14.72% DER.
Sep 23
Nebius Raises GPU Cloud Prices 16-20% in Its Second Hike This YearGPU cloud pricing has trended in exactly one direction relative to demand all year — up. Nebius's second price increase in 2026, this time 16-20% on its Token Factory offering, is a concrete data point in that trend worth understanding for anyone budgeting GPU rental costs, whether you're training your own models or just running inference at scale.
Sep 22
Cloudflare Worker Previews: A Production-Like Environment Per Git BranchOn September 22, 2026, Cloudflare shipped Worker Previews — a production-like environment for every git branch, created with one wrangler command. Each preview gets its own code, config, URL, Durable Objects and Containers state, and observability, so a PR can be tested in isolation before it ever touches production. explainx.ai breaks down what changed, how it differs from Cloudflare's existing preview tooling, and why it matters most for teams running coding agents.
Sep 21
Cloudflare Python Workers Reach GA: FastAPI, Django, and Postgres at the EdgeTwo years after beta, Cloudflare's Python Workers are generally available — meaning Python is now a first-class Workers language with native bindings, ASGI/WSGI support for FastAPI, Django, and Flask, Hyperdrive connections to Postgres and MySQL over real TCP sockets, and a newly-accepted Python packaging standard (PEP 783) for running native-extension packages in WebAssembly. A much older Quick Tunnels feature also went viral again this week as an ngrok alternative — here's what's actually new versus what's just resurfaced.
Sep 17
CoreWeave Deploys First Multi-Rack NVIDIA Vera Rubin ClusterCloud GPU provider CoreWeave deployed what it describes as the first multi-rack cluster built on NVIDIA's Vera Rubin platform, with hundreds of GPUs — a genuine early-production milestone that gives a first real signal of how quickly NVIDIA's newest hardware generation is moving from announcement to actual customer-facing capacity.
Sep 17
Fujitsu MONAKA: A "Made-in-Japan" AI CPU That TSMC FabricatesFujitsu's MONAKA is a genuinely interesting piece of engineering: a 3D-stacked Arm CPU claiming twice the AI inference throughput of other CPUs, air-cooled to 40C, sold as sovereign infrastructure for Japan and Europe. It is also fabricated by TSMC in Taiwan, which is the part the press release works hardest to avoid saying.
Sep 17
Nebius Opens Madrid AI Hub, Targets 4 Million H100-Equivalent GPUsNebius opened a new AI compute hub in Madrid, setting a target of 4 million H100-equivalent GPUs — a substantial European compute capacity expansion that lands the same week the company raised its GPU rental rates 20%, illustrating how demand and pricing can rise even as new supply comes online.
Sep 17
Nebius Raises GPU Rental Rates 20% in Second Hike Since MayCloud GPU provider Nebius raised its NVIDIA GPU rental rates by 20%, the second such price increase since May 2026 — a direct, quantifiable signal that the cost of renting AI compute keeps climbing, even as more supply comes online across the industry.
Sep 17
Neuralink Participant Speaks First Words via Brain Implant: "I Love You"Neuralink posted a video on September 17, 2026 showing a paralyzed clinical trial participant using a brain implant to produce speech, with reports saying the first words were "I love you." The device remains investigational and unapproved by the FDA, but the moment has become one of Neuralink's most emotionally resonant public updates.
Sep 17
NVIDIA CUDA Rust: Write GPU Kernels Natively in RustNVIDIA HPC Developer announced CUDA Rust on September 16, 2026 — two paths, cuda-oxide for SIMT kernels compiled to PTX and cutile-rs for tile-based programming on stable Rust, both designed to catch aliasing errors at compile time that CUDA C++ leaves to runtime debugging.
Sep 17
NVIDIA Rubin NVL72 Hits 67x Throughput-Per-Cost, Beating Own ClaimsIndependent analysis firm SemiAnalysis tested NVIDIA's Rubin NVL72 platform and found it delivers 67x throughput per dollar compared to a prior generation baseline — a result that exceeds NVIDIA's own published performance claims, a rare and notable outcome for vendor hardware benchmarking.
Sep 16
Micron Demonstrates 512GB DDR5 RDIMM, Enabling 12TB Per ServerMicron demonstrated a 512GB DDR5 RDIMM that packs 12TB of memory into a single 24-slot dual-socket server, cutting power draw more than 60% versus four 128GB modules doing the same job. AMD and Intel are validating it now, with volume production targeted for the second half of 2027.
Sep 16
Salesforce Koa: CRM Reasoning Built on NVIDIA NemotronSalesforce is developing a CRM reasoning model while bringing its sales workflows into Claude. Koa makes the enterprise build-versus-buy debate more specific: which parts of intelligence should a vendor own?
Sep 14
NVIDIA Inception: Free Cloud Credits and VC Intros for AI StartupsSep 13
Is Nvidia the "Central Bank of AI"? What ~$300B in Backstops MeansSep 11
NVIDIA BioNeMo Inference Runtime: Faster Boltz-2, OpenFold2, Protenix v2NVIDIA announced the public beta of BioNeMo Inference Runtime on September 10, 2026 — an open-source, PyTorch-native library for speeding up biomolecular structure-prediction inference with specialized kernels, CUDA Graphs, and Ray-powered GPU replica scaling.
Sep 10
Analog Devices Buys Alif Semiconductor for $1.35B to Expand Edge AIAnalog Devices acquired Alif Semiconductor for $1.35 billion, aimed at expanding low-power AI chip capability for edge devices. explainx.ai covers what Alif's chips actually do, why low-power edge AI silicon is a distinct and increasingly important category separate from data-center AI chips, and what it means for anyone building on-device AI products.
Sep 10
Pamir AI's Lapis One: A $359 Linux Computer Built for Your AI AgentsPamir AI launched the Lapis One on September 9, 2026 — a coaster-sized Debian Linux computer with a built-in NPU, designed to run AI agents around the clock without tying up your laptop. explainx.ai breaks down the real specs, the "Agent KVM" feature that's driving the buzz, the $359-vs-Raspberry-Pi debate already playing out in the replies, and who this device is actually for.
Sep 8
Nvidia Sol-H3 Reportedly Generates AI Video Faster Than It Plays BackReports on September 8, 2026 say Nvidia's Sol-H3 inference stack generates AI video at roughly 3x real-time speed — meaning a 10-second clip reportedly renders in about 3 seconds, faster than the video itself plays. Here's what that threshold actually unlocks and how it fits Nvidia's broader inference push.
Sep 7
Jensen Huang Says "AGI Has Arrived" With GPT-6 Astra — Is He Right?Hours after GPT-6 Astra's launch week wrapped, NVIDIA CEO Jensen Huang posted that "AGI has arrived" — crediting Astra's training run to 100,000-plus Grace Blackwell NVLink72 GPUs and previewing 400,000 more. The claim isn't new, "AGI" has no agreed definition, and practitioners actually using Astra are pushing back hard. Here's the full picture.
Sep 7
NVIDIA's $99B Equity Book and Berkshire's Alphabet Bet, ClaimedOne widely shared X essay by physicist Dr. Alex Wissner-Gross claims NVIDIA's direct equity stakes in AI companies have grown tenfold to $99 billion, that $105 billion in credit is tied to OpenAI's Ohio site, and that Berkshire Hathaway's Greg Abel bought $10 billion of an Alphabet capital raise at a discount. explainx.ai walks through what's plausible, what's already independently reported elsewhere, and what a circular compute-financing web means for anyone paying for AI compute.
Sep 5
Extropic Z1T: Transformers Built for Thermodynamic Chips, Not GPUsSep 4
The Largest Electric Aircraft Just Flew. It's Not Fully Electric.Sep 4
PAIR Gets Its First Real Partner: One-Click Local Models via Hermes DesktopSep 4
NVIDIA PAIR: Turn Idle Home PCs Into a Personal AI ClusterSep 3
Mostik Wants Models to Talk Without Words. The Numbers on X Aren't in Its Own Paper.Sep 3
NVIDIA Is Buying Hugging Face for $12.9 Billion. What Changes for You?Sep 3
Nvidia AI Infra Summit 2026: What to Expect From Ian Buck's KeynoteSep 3
Sam Altman's Almond-vs-ChatGPT Water Claim, Fact-CheckedSep 2
Top 10 Neural Rendering Use Cases Beyond DLSS 5's Beauty FilterSep 1
NVIDIA BioNeMo Agent Toolkit Comes to Claude Science: Protein Prediction by Prompt
August 2026
Aug 30
15 GW of AI Compute Sits Dark in 2027 — Power Is the Real BottleneckAug 27
How Cloudflare Saved 100 Terabytes of Memory Optimizing DNS CacheAug 27
Nvidia Reportedly Agrees to Buy Hugging Face for $12.9B — What Builders Should KnowAug 26
NVIDIA Dynamo Shadow Engine: 7s LLM Failover vs 5-Min Cold StartAug 25
Taiwan Indicts 9 Over 74 Nvidia B300 Servers Smuggled to ChinaAug 24
AI Chip Architectures Explained: GPU vs TPU vs Trainium vs Cerebras vs GroqAug 24
NVIDIA ACES: 27% of Agent Skill Runs Don’t Beat BaselineAug 24
Groq 3 LPX Hits 3,400 tok/s — Nebius First Cloud AdopterAug 21
NVIDIA AVO Hits 100% on ARC-AGI-3 — But Read the Fine PrintAug 20
Mojo Is Now Fully Open Source — What It Means for AI Kernel WritingAug 19
Cerebras CS-4: The Wafer-Scale Chip Claiming 30x Faster AI InferenceAug 18
NVIDIA Guarantees OpenAI's Ohio AI Factory: The PORTS-Pike LPS DealAug 17
Nvidia's $21 Billion SpaceX Stake Came From the xAI Merger, Not a New BetAug 16
How to Start a Small Data Center in 2026: Step-by-Step GuideAug 13
The Next AI Bottleneck Is Pumps, Coolant, and GasAug 12
Nvidia's $500B Plan to Make GPUs an Asset Class — and Why Its CDS DoubledAug 11
NVIDIA Nemotron 3.5 Lightning: A 30B Open MoE Built for Always-On AgentsAug 10
AI Server Buyers Are Paying Up to 3x Market Price for MLCCsAug 10
Samsung Hits 80% HBM4 Yield — Four Months Ahead of ScheduleAug 9
CMF Phone 1 as a Home Server: Termux Replaced a Hetzner VPSAug 7
AMD Acquires Taalas: The Chip That Etches Model Weights Into SiliconAug 7
Cloudflare Kitesurf: The Agent-First Browser Running in V8 IsolatesAug 6
NVIDIA Vera CPU: Great Silicon, Misleading WhitepaperAug 5
Cloudflare OS: An Open-Source Platform for Agents, Apps, and WorkAug 5
Cloudflare Wallets: Programmable Payments for AI Agents ExplainedAug 5
NVIDIA Alpamayo 2 Super: Open Reasoning Model for RobotaxisAug 3
Cloudflare Computer: Agents Need Isolates, Not Just Containers
July 2026
Jul 31
FastGen-PDD: NVIDIA's 4-8 Step Distillation for Video and Image ModelsJul 30
AI Companies Hiring Electricians and Carpenters by the ThousandsJul 29
NVIDIA × SSI: Ilya Sutskever’s Lab Gets Vera Rubin and a 10× Compute BetJul 27
Japan’s $9,000 “Human Fridge” for Extreme HeatJul 27
Nvidia’s First U.S.-Made GB300 Chips — Arizona Reality CheckJul 26
The AI Data Center Backlash, Mapped: What Was Actually Blocked?Jul 26
Can AI Solve Global Warming? What the Evidence SaysJul 26
Cloudflare AI Traffic: Search, Agent, TrainingJul 26
Data Centers, Water, and the Lawsuits Nobody Is TrackingJul 26
Every Hyperscaler Has a Nuclear Deal—Here Is What Each Actually BoughtJul 26
Nvidia–OpenAI $250B Backstop: Ohio’s 10GW Data CenterJul 23
AI Giants Carry $1.65 Trillion in Off-Balance-Sheet Debt — Is It Another Enron?Jul 23
Augmental MouthPad: A Tongue-Controlled Touchpad Goes on Public SaleJul 21
Google's Frozen v2 Chip and the Start of Gemini 4 Pre-TrainingJul 21
NVIDIA DLSS 5: Neural Rendering or an AI Slop Filter?Jul 21
NVIDIA MotionBricks: Real-Time Motion From Characters to Unitree G1Jul 21
NVIDIA SIGGRAPH 2026: Cosmos 3 Edge, MCP Creative Agents, and the Physical AI StackJul 16
GeForce NOW India Launch: RTX 5080 Pricing, Tiers, and What Changes July 15, 2026Jul 14
Japan Recovers 90% of Lithium From EV Batteries — What It Means for Supply ChainsJul 9
Cloudflare Drop: Deploy a Folder to the Edge in Seconds — No AccountJul 7
Cloudflare Monetization Gateway: x402 Micropayments for APIs, MCP Tools, and the Agent WebJul 7
Gaming and AI Hardware Costs in 2027: Research Forecasts, Charts, and Buy WindowsJul 2
NVIDIA Nemotron-Labs-TwoTower: Split a 30B Model in Two for 2.42× Faster Diffusion Generation
June 2026
Jun 29
Stanford MemoryDAX: 65 Years of DRAM, HBM, and NAND Flash Prices — What the Data Actually ShowsJun 27
What is the real environmental impact of AI data centers? Water, power, and why local fights are not about percentagesJun 24
NVIDIA BioNeMo Agent Toolkit: AI Agents for Drug Discovery [2026]Jun 21
RuView: See Through Walls With WiFi — ESP32 Spatial Intelligence PlatformJun 20
Cloudflare Temporary Accounts: How AI Agents Deploy Workers Without Signup (2026)Jun 16
GeForce Now and Cloud Gaming: The Complete Guide for 2026Jun 14
Zvec: Alibaba's Open-Source In-Process Vector Database (2026)Jun 6
NVIDIA Nemotron 3 Ultra: 550B Open-Weight MoE Model Redefines Agentic AI PerformanceJun 4
NVIDIA Cosmos 3: Open Physical AI World Models for Robots and Autonomous SystemsJun 1
NVIDIA Computex 2026: Complete Recap - Nemotron 3 Ultra, Cosmos 3, RTX Spark & Everything Announced
May 2026
May 30
NVIDIA's N1X ARM Chip: The 'New Era of PC' That Could End Intel and AMD's 40-Year ReignMay 17
60% of PC gamers shelve build plans as AI crunch drives component prices up 300%+May 15
NVIDIA's Video Search and Summarization: Building GPU-Accelerated Vision AgentsMay 8
Top 10 AI Tech Gadget & Hardware Directories (2026)