Infrastructure & Hardware
Chips, silicon, data centers, and the compute powering the models.
Chips, silicon, data centers, and the compute powering the models.
OpenAI says it used its own GPT-5.6 Sol model, working autonomously inside Codex, to redesign large parts of its inference stack and agentic harness, reducing end-to-end serving costs by 20% and…
Meta and BlackRock, the world's largest asset manager, announced a venture on July 28, 2026, to develop and operate a $14 billion data center campus in El Paso, Texas, with BlackRock-managed funds…
NVIDIA says it is now using its own Vera CPU internally to accelerate the electronic design automation (EDA) work behind its next generations of CPUs and GPUs, delivering measurable speedups on the…
NVIDIA's Vera Rubin NVL72 rack-scale AI system is now ramping into production, with early cloud partners publishing the first live-hardware benchmarks and NVIDIA detailing a wave of global…
AMD unveiled its next-generation AI accelerator, the Instinct MI455X GPU, alongside the Helios rack-scale system built around it, positioning the pairing as a direct competitor to NVIDIA's Vera Rubin…
Bristol Myers Squibb is building what NVIDIA calls the life science industry's most advanced AI factory, deploying eight DGX Vera Rubin NVL72 systems as a second DGX SuperPOD the pharmaceutical…
Alphabet reported second-quarter 2026 revenue of $119.8 billion, up 24% year over year, with Google Cloud posting its strongest growth in recent quarters as enterprises scale up AI workloads. CEO…
NVIDIA unveiled Spectrum-6, the next generation of its Spectrum-X Ethernet networking platform, on July 21, 2026, as the interconnect layer for its upcoming Vera Rubin AI platform. The…
NVIDIA CEO Jensen Huang commissioned a DGX GB300 supercomputer at the Naval Postgraduate School (NPS) in Monterey, California, on July 22, 2026, giving the school's students and faculty on-premises…
OpenAI has published details of Project Camellia, a long-term data center it is designing and developing in Effingham County, Georgia, backed by a 3.2-gigawatt power contract with Georgia Power…
Wistron opened its first U.S. manufacturing facility on July 21, 2026, a 324,000-square-foot plant in Fort Worth, Texas built to produce NVIDIA's most advanced AI superchips on American soil — part…
Microsoft announced three new Azure virtual machine families built on AMD's latest silicon and rackscale platform, expanding Azure's AI and high-performance computing infrastructure to cover data…
At SIGGRAPH 2026 (running through July 23 in Los Angeles), NVIDIA unveiled DGX Station, a desktop supercomputer built around the GB300 Grace Blackwell Ultra Desktop Superchip, paired with Nemotron 3…
Perplexity has published a technical account of SPACE, the sandboxing platform it built to run every code execution and file operation its AI agents perform, revealing that the system already handles…
Japan's government, a new industrial consortium called Noetra, and NVIDIA announced the launch of what NVIDIA calls the world's first state-tendered national AI infrastructure — a 140-megawatt AI…
OpenAI released its first branded hardware product on July 15, 2026: Codex Micro, a $230 keyboard built to control fleets of Codex coding agents, made in collaboration with specialty keyboard…
NVIDIA introduced two new Jetson Thor compute modules, the T3000 and T2000, extending its Blackwell GPU architecture down into compact, power-efficient systems built for robotics and edge AI. The…
NVIDIA published new performance-per-watt figures on July 14, 2026, arguing the metric — not raw throughput — is now the deciding factor in AI infrastructure economics, and backing that claim with…
Meta said in a blog post Monday, July 13, 2026, that its Hyperion data center supercluster in Richland Parish, Louisiana will be a 5 GW facility costing more than $50 billion — up from the $27…
Cerebras Systems is building out a major European data center footprint, with CEO Andrew Feldman announcing plans for 200 megawatts of AI compute capacity across the continent by the end of 2027, a…
Microsoft's newly published 2026 Environmental Sustainability Report shows the company's total greenhouse gas emissions rose 25% year over year in fiscal 2025, a jump the company attributes directly…
Nvidia unveiled its Vera Rubin platform at ISC High Performance 2026 in Hamburg, positioning it as a system that can pack the performance of a TOP500-class supercomputer into a single rack, aimed…
Research firm SemiAnalysis reported on July 6, 2026 that Nvidia's next-generation Kyber rack-scale architecture, designed to house the 2027 Vera Rubin Ultra chips, has slipped by more than a year to…
NVIDIA published benchmark results on July 7 showing its new Vera CPU, built for agentic AI workloads, delivering large per-core performance gains over x86 in real-world testing by Perplexity and…
Anthropic has signed a 20-year lease with data-center operator TeraWulf for an AI infrastructure campus in Hawesville, Kentucky, roughly an hour southwest of Louisville. The deal, announced July 6,…
Google's 2026 Environmental Report, published June 30, shows electricity consumption rose 37% year over year in 2025 — the company's largest single-year increase on record — driven primarily by its…
Google restricted how much Gemini model capacity Meta could purchase starting around March 2026, after Meta's requested usage outstripped what Google was able to supply, according to a Financial…
A new SemiAnalysis report, based on conversations with enterprises across industries, finds that the era of unlimited employee AI token consumption is ending — companies that spent early 2026…
Anthropic has held early-stage discussions with Samsung Electronics about developing a custom AI chip, with Samsung's 2-nanometer manufacturing process and advanced packaging facilities under…
NVIDIA introduced a new business model on July 1, 2026, that lets AI cloud providers procure its infrastructure through revenue-sharing and credit-support arrangements rather than upfront capital…
Amazon Web Services has expanded its Bedrock model catalog in AWS GovCloud (US) to include two OpenAI open-weight models and NVIDIA's Nemotron 3 series — extending access to frontier-class AI to…
NVIDIA announced on July 1, 2026 that it and its network of manufacturing and infrastructure partners are on track to produce up to $500 billion in AI infrastructure domestically, spanning 43 states,…
AI chip startup Etched reported on June 30, 2026, that it has booked $1 billion in contract orders for its purpose-built transformer inference chip, following a successful manufacturing run at TSMC.…
NVIDIA published a detailed breakdown on June 30, 2026 of how its layered inference software stack reduced token costs for DeepSeek V4 on Blackwell hardware by up to 5x within one month of…
Anthropic's Claude models became generally available on NVIDIA GB300 NVL72 systems through Microsoft Foundry on June 29, 2026, giving Azure-based enterprises a high-performance foundation for…
NVIDIA and Firefly Aerospace announced on June 29 that Blue Ghost Mission 2, targeted for late 2026, will carry the first operational NVIDIA Jetson edge AI deployment in lunar orbit. The mission will…
NVIDIA unveiled the Rubin platform, a next-generation AI computing architecture built from six co-designed chips that the company says delivers 10x lower inference token cost and requires 4x fewer…
OpenAI and Broadcom on June 24, 2026 jointly announced Jalapeño, a custom AI accelerator built specifically for large language model inference. Developed in nine months — what the companies call one…
NVIDIA and AWS announced a set of coordinated infrastructure expansions this week, including new Amazon EC2 G7 instances powered by NVIDIA's RTX PRO 4500 Blackwell Server Edition GPUs,…
Open-source AI startup Reflection AI has signed a compute agreement with SpaceX worth up to $6.3 billion, gaining exclusive access to NVIDIA's latest GB300 chips at the Colossus 2 data center near…
At Google Cloud Next '26 in Las Vegas on April 22, 2026, Google announced its eighth-generation Tensor Processing Units: the TPU 8t, built for massive-scale model training, and the TPU 8i,…
NVIDIA has published the cooling architecture for its next-generation AI factory infrastructure — and the headline spec is a coolant temperature of 45°C, warmer than most hot tubs, running through a…
Europe's first exascale supercomputer, JUPITER, is producing science at a scale previously impossible. Located at Forschungszentrum Jülich in Germany and powered by NVIDIA Grace Hopper Superchips,…
Midjourney, known until now exclusively for AI-generated images, unveiled its first hardware product on June 18, 2026: a full-body ultrasonic scanner capable of producing a complete 3D body map in…
Cerebras Systems made its public market debut on May 14, 2026, raising approximately $5.55 billion in the largest technology IPO of the year and the first major pureplay AI chip company to go public.…
Hewlett Packard Enterprise and NVIDIA announced a significant expansion of the HPE AI Factory program on June 16, 2026, targeting enterprise deployments of agentic AI workloads. The update introduces…
Cohere on June 15 announced a significant expansion of its London presence, relocating to a 14,000-square-foot office at 100 New Oxford Street — nearly three times the size of its previous space —…
Amazon Web Services announced the availability of Google DeepMind's Gemma 4 open-weight model family on Amazon Bedrock on June 10, 2026. Three model variants are now available through Bedrock's…
OpenAI and Oracle have formed a new enterprise distribution partnership that allows Oracle Cloud Infrastructure (OCI) customers to apply existing Universal Credits toward OpenAI's frontier models and…
On June 7, 2026, NAVER and NVIDIA announced a major sovereign AI infrastructure expansion: NAVER will build AI factories on the NVIDIA DSX platform at its GAK Sejong data center in South Korea,…
Microsoft on June 2, 2026, announced Majorana 2, its next-generation topological quantum chip, alongside the general availability of Microsoft Discovery, the company's agentic AI platform for…
Together AI announced on June 10, 2026, that it has achieved ISO 27001:2022 certification for its information security management system, certified by A-LIGN Compliance and Security, an…
AWS has released Neuron Agentic Development, a suite of specialized AI agents and skills that enable machine learning engineers to develop, debug, and optimize custom kernels for AWS Trainium and…
Google has agreed to pay SpaceX $920 million per month from October 2026 through June 2029, securing access to approximately 110,000 NVIDIA GPUs housed in a SpaceX data center facility, according to…
Meta has signed its first built-to-suit AI data center agreement in India, partnering with Reliance Industries to build a 168-megawatt facility in Jamnagar, Gujarat. The announcement deepens a…
At Apple's WWDC 2026 developer conference, NVIDIA announced that its Blackwell GPUs with Confidential Computing are now integrated into Apple's Private Cloud Compute (PCC) infrastructure. The…