Graphics Processing Unit (GPU) stories
The deal could let enterprises serve more AI models on fewer GPUs by shifting cache data into DRAM, SSDs and remote storage.
Customers facing soaring AI compute demands will get access to one of the first cloud offerings of AMD's latest large-memory GPU and rackscale system.
Helios racks and new Instinct chips put AMD deeper into AI infrastructure as it courts major customers and challenges Nvidia.
Enterprises could see better GPU use as the partnership aims to cut data delays that slow AI training, inference and analytics.
Developers can now run accelerator-heavy AI workloads on managed GKE Autopilot without handling node setup or low-level network allocation.
Enterprises can now run governed AI workloads faster, as the validated stack aims to cut integration delays and simplify deployment across environments.
Demand for production-ready AI agents is pushing Google to reframe Cloud Run as a platform for long-running, data-driven workloads.
The new model could ease access to scarce AI computing for startups and cloud providers, while giving Nvidia a recurring revenue stream.
AI-driven demand could overwhelm available capacity by 2030, with spending on servers and GPUs pushing supply short of need across key markets.
Demand for sovereign AI compute has forced the Gagarin site to expand within weeks, with capacity set to rise to 5MW by year-end.
Software improvements have slashed the cost of serving DeepSeek V4 on Blackwell, underscoring the race to make AI deployments economical.
Memory makers are set to spend more than USD $50 billion on 300mm fab equipment in 2026 as AI demand drives HBM and DDR5 capacity.
The deployment could speed up incident response across Nebius's GPU-heavy AI cloud, where outages can leave costly compute idle and affect customers.
The deal gives Qualcomm a stronger software layer for developers as AI workloads spread from edge devices into data centres.
AI and HPC users could cut storage costs as WD's new designs shift colder data to hard drives while keeping active workloads on NVMe.
Australian buyers are weighing convenience against control as desktop PC demand rises, with local warranties and faster access swaying choices.
The Australian AI infrastructure provider is betting on onshore data control as it seeks AUD $40 million from investors in a fully underwritten float.
AI agents could run faster and keep graphics processors busier as Nvidia's Vera chip targets the bottlenecks that slow data centre workloads.
Research papers at ICML 2026 increasingly leaned on Nvidia's chips and open models, underlining its reach across robotics, biomedicine and AI.
NVIDIA says US AI demand will add USD $485 billion to GDP in 2026 as it expands chip, systems and data centre manufacturing.