Browsing Tag
AI Compute
3 posts
GPU capacity, AI accelerator access, model serving infrastructure, and compute economics for artificial intelligence workloads.
Micron’s Hiroshima HBM Expansion Shows AI Memory Is the Next Supply Fight
Micron has broken ground on a roughly $9.3 billion Hiroshima expansion that will produce high-bandwidth memory for AI processors, with shipments expected around summer 2028. The timing shows why memory, not just GPUs, has become a strategic bottleneck for AI infrastructure buyers.
SoftBank SB Neo Turns AI Cloud Capacity Into a 10-Gigawatt Race
SoftBank has formed SB Neo, a U.S.-based neocloud company meant to supply AI chips and cloud services to model developers and large enterprises. The plan, tied to SoftBank's 10-gigawatt AI infrastructure target by 2030, shows how AI compute is shifting from scarce GPU rental toward vertically managed infrastructure businesses built around power, chips, networking, and operations.
Meta Compute Would Turn AI Oversupply Into a Cloud Business
Meta is reportedly developing a cloud infrastructure business that would sell AI compute and hosted model access. The plan is not final, but it shows how Big Tech’s AI data-center spending is starting to look like a market of its own.