Browsing Category
AI Infrastructure
42 posts
Cloud infrastructure, chips, data centers, model deployment, edge AI, compute platforms, secure systems, and the physical infrastructure behind artificial intelligence products and services.
Kimi K3 Turns Open-Weight AI Into a Deployment Test
Moonshot AI’s Kimi K3 is available through apps, Kimi Code, and an API now, with full model weights promised by July 27. The launch gives developers a powerful new open-weight contender, but the real test is deployment: hardware scale, pricing, agent controls, and independent verification.
Google Cloud Makes AlphaEvolve an Enterprise AI Optimization Service
Google Cloud has made AlphaEvolve generally available on Gemini Enterprise, turning Google DeepMind’s algorithm-discovery system into a product for enterprises that need better code for forecasting, routing, chips, logistics, scientific computing, and other hard optimization problems.
Anthropic’s $19B TeraWulf Lease Turns Old Industrial Power Into AI Compute
Anthropic has signed a 20-year lease for roughly 401 megawatts of AI data center capacity at TeraWulf’s Justified Data campus in Hawesville, Kentucky. The deal shows how AI labs are moving beyond ordinary cloud rentals and locking up power-heavy industrial sites years before capacity comes online.
Chip Sales Just Hit a Record as AI Demand Spreads Beyond GPUs
SIA says global semiconductor sales reached $120.6 billion in May 2026, the highest monthly total it has recorded and more than double the level from a year earlier. The data suggests the AI chip boom is now lifting a wider stack of memory, networking, logic, and foundational semiconductors across every major region.
Apple’s Broadcom Deal Makes Edge AI a Supply-Chain Commitment
Broadcom’s July 6 SEC filing says it will supply custom ASIC silicon for multiple generations of Apple products through 2031. The sparse disclosure does not confirm specific Apple Intelligence hardware, but it locks in a key supplier relationship as Apple tries to make more AI run locally on phones, Macs, watches, and tablets.
NVIDIA’s AI Cloud Deals Turn GPUs Into a Revenue-Share Business
NVIDIA’s July 1 revenue-sharing and credit-support model gives AI cloud partners a new way to finance large GPU deployments, while giving NVIDIA a usage-linked cut of supported cloud revenue. Sharon AI and Firmus are the first test cases, with plans for up to 210,000 GPUs across Australia and Indonesia.
AI Memory Shortage Turns Into a Fight Over Who Gets Chips
A July 1 SEMI letter warns Washington that direct intervention in memory-chip pricing or production could worsen an AI-driven shortage. The fight now reaches beyond data centers, with broadband, automotive, medical-device, retail, and consumer-electronics groups worried that HBM demand will squeeze ordinary DRAM supply.
Stargate UK Turns AI Data Center Promises Into a Credibility Test
Fresh reporting on OpenAI's paused Stargate UK project shows why AI data center announcements now need harder questions about power, planning, signed offtake, and local coordination before investors, governments, and customers treat headline capacity as real infrastructure.
UN AI Report Turns Governance Into a Compute and Capacity Test
The UN’s first global scientific AI assessment warns that governance is now tied to compute access, local expertise, language coverage, and real-world model evaluation. The report arrives before the July 6-7 Global Dialogue on AI Governance in Geneva.
Hong Kong’s AI Chip Trade Boom Turns Logistics Into a Policy Risk
Hong Kong re-exported $124 billion in semiconductors to mainland China in the first five months of 2026, according to Bloomberg’s review of official data. The city’s rising role as an AI-chip gateway shows why logistics, payments, and export-control exposure now matter as much as chip supply itself.