Infrastructure & Hardware
Chips, silicon, data centers, and the compute powering the models.
Chips, silicon, data centers, and the compute powering the models.
Huawei is accelerating the launch of its next Ascend AI chip by two quarters, even as its own numbers show the giant computing cluster meant to hold thousands of those chips has shrunk from its…
NVIDIA introduced DSX on September 15, 2026, a platform meant to manage an AI data center's power, cooling, networking, and compute as one integrated system rather than as separate components tuned…
NVIDIA is backing a major expansion of Australia's AI data center capacity, partnering with eight local infrastructure providers on a buildout that could more than double the country's existing…
OpenAI has published a detailed account of how it scaled Habitat, the internal storage platform behind ChatGPT, from a simple client library into a distributed system now handling tens of millions of…
NVIDIA has published deployment guidance and Day 0 performance numbers for running Alibaba's Qwen3.8-2.4T-A95B — a 2.4-trillion-parameter mixture-of-experts model — on its GB300 NVL72 rack-scale…
Anthropic has agreed to rent roughly $45 billion in AI computing capacity from Nscale, a British AI infrastructure company, in a six-year deal that will draw on Nscale's flagship data center…
Cerebras used its Hot Chips 2026 presentation to lay out the next two generations of its wafer-scale AI hardware roadmap, detailing a rack-scale "Nexus" platform architecture and a future CS-6 system…
NVIDIA reported revenue of $96.2 billion for its second fiscal quarter of 2027 (ended July 26, 2026), up 18% from the prior quarter and up 106% from a year earlier, the company said in its August 26…
Hugging Face has moved further into physical AI hardware with Microduck, a small duck-shaped robot designed to be taught new physical skills through open-source reinforcement-learning tools rather…
AWS and NVIDIA announced on August 26, 2026 an expansion of their GPU supply agreement, committing to deploy 2 million additional NVIDIA GPUs across AWS infrastructure in 2027 and 2028 — on top of…
Chris Malone, who led OpenAI's data center organization, has left the company, adding to a wave of senior executive departures at OpenAI this year. What's new Per TechCrunch's reporting, "Chris…
NVIDIA on August 25, 2026 announced the Jetson Orin Nano 2, a new robotics computer the company says brings frontier-class generative AI performance to entry-level edge devices. "NVIDIA today…
OpenAI has published the first performance results for Jalapeño, the custom inference chip it developed with Broadcom, and the company says the chip beats comparable Nvidia GB200 and GB300 rack…
NVIDIA published new efficiency benchmarks on August 24, 2026 for its Vera Rubin NVL72 rack-scale system, claiming a dramatic jump in performance-per-watt and per-token cost over its…
NVIDIA has published new details on NVLink Fusion, the program that lets hyperscalers and AI-native companies plug custom-built processors (XPUs) into NVIDIA's rack-scale interconnect fabric…
NVIDIA and SpaceXAI (Elon Musk's combined AI and space venture, formerly xAI) announced that SpaceXAI will deploy NVIDIA's new Vera CPU across its AI infrastructure, with plans to scale toward…
NVIDIA announced on August 24, 2026 that Groq 3 LPX, its dedicated interactive AI inference accelerator, is now in full production. The chip is built to speed up the "decode" phase of inference — the…
Marvell Technology has granted Google a warrant to buy up to $12.2 billion worth of Marvell stock, tying the size of Google's potential stake directly to how much custom AI silicon Google buys from…
Cerebras Systems unveiled the CS-4, the fourth generation of its wafer-scale AI computer, on August 18, 2026. The company's announcement calls it "the fastest AI accelerator in the industry, and a…
Cerebras Systems is the hardware behind OpenAI's newly announced Ultrafast mode for GPT-5.6 Sol, running the model at speeds the two companies say reach 750 output tokens per second with no quality…
OpenAI has agreed to lease roughly 8 gigawatts of data center capacity at a new campus in Pike County, Ohio, with NVIDIA committing up to $105 billion in financing support and agreeing to supply the…
NVIDIA is positioning its compute infrastructure as a distinct, investable asset class, and has partnered with six of the world's largest capital providers to build financing platforms aimed at…
Anthropic has signed a 20-year compute lease with Riot Platforms worth $9.1 billion, the latest in a string of multibillion-dollar power and data-center agreements the AI lab has struck this year to…
NVIDIA, Google, and Microsoft have jointly developed an 800-volt direct-current (VDC) power architecture for AI data centers, aiming to remove power-delivery bottlenecks that are increasingly…
Mistral AI has unveiled a three-part push to control the infrastructure Europe's AI runs on, anchored by a new financing structure called European Compute Units designed to lock in up to 1 gigawatt…
NVIDIA has signed memorandums of understanding with six of the world's largest asset managers — Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR — to build independent financing…
Armenian AI cloud company Firebird has switched on what NVIDIA calls the largest AI factory in the CIS region, a data center in Hrazdan, Armenia built with NVIDIA hardware and support from Dell…
Anthropic confirmed on August 5, 2026 that it is assembling an in-house team to design custom AI chips for Claude, marking the first time the company has publicly acknowledged the effort. The…
OpenAI says it used its own GPT-5.6 Sol model, working autonomously inside Codex, to redesign large parts of its inference stack and agentic harness, reducing end-to-end serving costs by 20% and…
Meta and BlackRock, the world's largest asset manager, announced a venture on July 28, 2026, to develop and operate a $14 billion data center campus in El Paso, Texas, with BlackRock-managed funds…
NVIDIA says it is now using its own Vera CPU internally to accelerate the electronic design automation (EDA) work behind its next generations of CPUs and GPUs, delivering measurable speedups on the…
NVIDIA's Vera Rubin NVL72 rack-scale AI system is now ramping into production, with early cloud partners publishing the first live-hardware benchmarks and NVIDIA detailing a wave of global…
AMD unveiled its next-generation AI accelerator, the Instinct MI455X GPU, alongside the Helios rack-scale system built around it, positioning the pairing as a direct competitor to NVIDIA's Vera Rubin…
Bristol Myers Squibb is building what NVIDIA calls the life science industry's most advanced AI factory, deploying eight DGX Vera Rubin NVL72 systems as a second DGX SuperPOD the pharmaceutical…
Alphabet reported second-quarter 2026 revenue of $119.8 billion, up 24% year over year, with Google Cloud posting its strongest growth in recent quarters as enterprises scale up AI workloads. CEO…
NVIDIA unveiled Spectrum-6, the next generation of its Spectrum-X Ethernet networking platform, on July 21, 2026, as the interconnect layer for its upcoming Vera Rubin AI platform. The…
NVIDIA CEO Jensen Huang commissioned a DGX GB300 supercomputer at the Naval Postgraduate School (NPS) in Monterey, California, on July 22, 2026, giving the school's students and faculty on-premises…
OpenAI has published details of Project Camellia, a long-term data center it is designing and developing in Effingham County, Georgia, backed by a 3.2-gigawatt power contract with Georgia Power…
Wistron opened its first U.S. manufacturing facility on July 21, 2026, a 324,000-square-foot plant in Fort Worth, Texas built to produce NVIDIA's most advanced AI superchips on American soil — part…
Microsoft announced three new Azure virtual machine families built on AMD's latest silicon and rackscale platform, expanding Azure's AI and high-performance computing infrastructure to cover data…
At SIGGRAPH 2026 (running through July 23 in Los Angeles), NVIDIA unveiled DGX Station, a desktop supercomputer built around the GB300 Grace Blackwell Ultra Desktop Superchip, paired with Nemotron 3…
Perplexity has published a technical account of SPACE, the sandboxing platform it built to run every code execution and file operation its AI agents perform, revealing that the system already handles…
Japan's government, a new industrial consortium called Noetra, and NVIDIA announced the launch of what NVIDIA calls the world's first state-tendered national AI infrastructure — a 140-megawatt AI…
OpenAI released its first branded hardware product on July 15, 2026: Codex Micro, a $230 keyboard built to control fleets of Codex coding agents, made in collaboration with specialty keyboard…
NVIDIA introduced two new Jetson Thor compute modules, the T3000 and T2000, extending its Blackwell GPU architecture down into compact, power-efficient systems built for robotics and edge AI. The…
NVIDIA published new performance-per-watt figures on July 14, 2026, arguing the metric — not raw throughput — is now the deciding factor in AI infrastructure economics, and backing that claim with…
Meta said in a blog post Monday, July 13, 2026, that its Hyperion data center supercluster in Richland Parish, Louisiana will be a 5 GW facility costing more than $50 billion — up from the $27…
Cerebras Systems is building out a major European data center footprint, with CEO Andrew Feldman announcing plans for 200 megawatts of AI compute capacity across the continent by the end of 2027, a…
Microsoft's newly published 2026 Environmental Sustainability Report shows the company's total greenhouse gas emissions rose 25% year over year in fiscal 2025, a jump the company attributes directly…
Nvidia unveiled its Vera Rubin platform at ISC High Performance 2026 in Hamburg, positioning it as a system that can pack the performance of a TOP500-class supercomputer into a single rack, aimed…
Research firm SemiAnalysis reported on July 6, 2026 that Nvidia's next-generation Kyber rack-scale architecture, designed to house the 2027 Vera Rubin Ultra chips, has slipped by more than a year to…
NVIDIA published benchmark results on July 7 showing its new Vera CPU, built for agentic AI workloads, delivering large per-core performance gains over x86 in real-world testing by Perplexity and…
Anthropic has signed a 20-year lease with data-center operator TeraWulf for an AI infrastructure campus in Hawesville, Kentucky, roughly an hour southwest of Louisville. The deal, announced July 6,…
Google's 2026 Environmental Report, published June 30, shows electricity consumption rose 37% year over year in 2025 — the company's largest single-year increase on record — driven primarily by its…
Google restricted how much Gemini model capacity Meta could purchase starting around March 2026, after Meta's requested usage outstripped what Google was able to supply, according to a Financial…
A new SemiAnalysis report, based on conversations with enterprises across industries, finds that the era of unlimited employee AI token consumption is ending — companies that spent early 2026…
Anthropic has held early-stage discussions with Samsung Electronics about developing a custom AI chip, with Samsung's 2-nanometer manufacturing process and advanced packaging facilities under…
NVIDIA introduced a new business model on July 1, 2026, that lets AI cloud providers procure its infrastructure through revenue-sharing and credit-support arrangements rather than upfront capital…
Amazon Web Services has expanded its Bedrock model catalog in AWS GovCloud (US) to include two OpenAI open-weight models and NVIDIA's Nemotron 3 series — extending access to frontier-class AI to…
NVIDIA announced on July 1, 2026 that it and its network of manufacturing and infrastructure partners are on track to produce up to $500 billion in AI infrastructure domestically, spanning 43 states,…
AI chip startup Etched reported on June 30, 2026, that it has booked $1 billion in contract orders for its purpose-built transformer inference chip, following a successful manufacturing run at TSMC.…
NVIDIA published a detailed breakdown on June 30, 2026 of how its layered inference software stack reduced token costs for DeepSeek V4 on Blackwell hardware by up to 5x within one month of…
Anthropic's Claude models became generally available on NVIDIA GB300 NVL72 systems through Microsoft Foundry on June 29, 2026, giving Azure-based enterprises a high-performance foundation for…
NVIDIA and Firefly Aerospace announced on June 29 that Blue Ghost Mission 2, targeted for late 2026, will carry the first operational NVIDIA Jetson edge AI deployment in lunar orbit. The mission will…
NVIDIA unveiled the Rubin platform, a next-generation AI computing architecture built from six co-designed chips that the company says delivers 10x lower inference token cost and requires 4x fewer…
OpenAI and Broadcom on June 24, 2026 jointly announced Jalapeño, a custom AI accelerator built specifically for large language model inference. Developed in nine months — what the companies call one…
NVIDIA and AWS announced a set of coordinated infrastructure expansions this week, including new Amazon EC2 G7 instances powered by NVIDIA's RTX PRO 4500 Blackwell Server Edition GPUs,…
Open-source AI startup Reflection AI has signed a compute agreement with SpaceX worth up to $6.3 billion, gaining exclusive access to NVIDIA's latest GB300 chips at the Colossus 2 data center near…
At Google Cloud Next '26 in Las Vegas on April 22, 2026, Google announced its eighth-generation Tensor Processing Units: the TPU 8t, built for massive-scale model training, and the TPU 8i,…
NVIDIA has published the cooling architecture for its next-generation AI factory infrastructure — and the headline spec is a coolant temperature of 45°C, warmer than most hot tubs, running through a…
Europe's first exascale supercomputer, JUPITER, is producing science at a scale previously impossible. Located at Forschungszentrum Jülich in Germany and powered by NVIDIA Grace Hopper Superchips,…
Midjourney, known until now exclusively for AI-generated images, unveiled its first hardware product on June 18, 2026: a full-body ultrasonic scanner capable of producing a complete 3D body map in…
Cerebras Systems made its public market debut on May 14, 2026, raising approximately $5.55 billion in the largest technology IPO of the year and the first major pureplay AI chip company to go public.…
Hewlett Packard Enterprise and NVIDIA announced a significant expansion of the HPE AI Factory program on June 16, 2026, targeting enterprise deployments of agentic AI workloads. The update introduces…
Cohere on June 15 announced a significant expansion of its London presence, relocating to a 14,000-square-foot office at 100 New Oxford Street — nearly three times the size of its previous space —…
Amazon Web Services announced the availability of Google DeepMind's Gemma 4 open-weight model family on Amazon Bedrock on June 10, 2026. Three model variants are now available through Bedrock's…
OpenAI and Oracle have formed a new enterprise distribution partnership that allows Oracle Cloud Infrastructure (OCI) customers to apply existing Universal Credits toward OpenAI's frontier models and…
On June 7, 2026, NAVER and NVIDIA announced a major sovereign AI infrastructure expansion: NAVER will build AI factories on the NVIDIA DSX platform at its GAK Sejong data center in South Korea,…
Microsoft on June 2, 2026, announced Majorana 2, its next-generation topological quantum chip, alongside the general availability of Microsoft Discovery, the company's agentic AI platform for…
Together AI announced on June 10, 2026, that it has achieved ISO 27001:2022 certification for its information security management system, certified by A-LIGN Compliance and Security, an…
AWS has released Neuron Agentic Development, a suite of specialized AI agents and skills that enable machine learning engineers to develop, debug, and optimize custom kernels for AWS Trainium and…
Google has agreed to pay SpaceX $920 million per month from October 2026 through June 2029, securing access to approximately 110,000 NVIDIA GPUs housed in a SpaceX data center facility, according to…
Meta has signed its first built-to-suit AI data center agreement in India, partnering with Reliance Industries to build a 168-megawatt facility in Jamnagar, Gujarat. The announcement deepens a…
At Apple's WWDC 2026 developer conference, NVIDIA announced that its Blackwell GPUs with Confidential Computing are now integrated into Apple's Private Cloud Compute (PCC) infrastructure. The…