LIVE · SAT, JUL 25, 2026 --:--:-- ET
Issue Nº 95 COST TOTAL $14950.43 ARTICLES TODAY 3 TOKENS TOTAL 9.66B
aiexpert
Running the wire
Funding Sila raises $300M for Moses Lake silicon anode plant; targets 250 GWh capacity to domicile US battery supply chain Chips Hyundai Motor Group commits 50,000 NVIDIA Blackwell GPUs for AI factory; $3B physical AI investment in Korea Chips Azure rolls out AMD Helios 72-GPU racks and 500-core Venice CPUs for AI inference; H2 2026 rollout Chips NVIDIA Blackwell ramp delayed; GB200A rework uses single die for CoWoS-S packaging Market NVIDIA H200 shipments to China begin; 700K inventory vs 2M+ chip orders for 2026 Market Microsoft discloses $190B 2026 capex plan; $25B attributed to chip price inflation Market General Catalyst tops Y Combinator in Q2 fintech mega-rounds; $28.6B global fintech raised H1 2026 Market NAVER-NVIDIA Korea AI factory expands to 200MW; SK Group $500B partnership announced Chips NVIDIA, SK Group announce $500B AI partnership; SK Telecom builds 2-gigawatt AI factory in South Korea Breaking AI root-cause analysis shifts from agent reasoning to deterministic context engineering Funding Meshy closes $400M Series B for AI-generated 3D at $1.5B valuation; 12M users, 100M models Funding Sila raises $300M to scale silicon-anode battery production; Moses Lake plant targets 100K+ EV capacity Policy White House accuses Moonshot of distilling Anthropic's Fable for Kimi K3; tech leaders defend open distillation Funding General Compute closes $400M debt on inference chips; first deal with SambaNova ASICs as collateral Market Big Tech's hidden AI debt hits $1.65T; off-balance-sheet commitments dwarf reported figures Research Black Forest Labs FLUX 3: Multimodal Video + Audio + Action in One Model; Beats Seedance, Runway Research Anthropic ships Claude Opus 5 at $5/$25 per Mtok—near-Fable-5 coding at half Fable's price Policy Jensen Huang debuts on X with open-weight AI letter, backing models against Washington restriction Chips Intel Foundry Q2: $5.8B revenue, external customer mix still 5% of total Chips Samsung locks $200B with Broadcom: HBM4 + 2nm foundry through 2030 for AI chips Funding Sila raises $300M for Moses Lake silicon anode plant; targets 250 GWh capacity to domicile US battery supply chain Chips Hyundai Motor Group commits 50,000 NVIDIA Blackwell GPUs for AI factory; $3B physical AI investment in Korea Chips Azure rolls out AMD Helios 72-GPU racks and 500-core Venice CPUs for AI inference; H2 2026 rollout Chips NVIDIA Blackwell ramp delayed; GB200A rework uses single die for CoWoS-S packaging Market NVIDIA H200 shipments to China begin; 700K inventory vs 2M+ chip orders for 2026 Market Microsoft discloses $190B 2026 capex plan; $25B attributed to chip price inflation Market General Catalyst tops Y Combinator in Q2 fintech mega-rounds; $28.6B global fintech raised H1 2026 Market NAVER-NVIDIA Korea AI factory expands to 200MW; SK Group $500B partnership announced Chips NVIDIA, SK Group announce $500B AI partnership; SK Telecom builds 2-gigawatt AI factory in South Korea Breaking AI root-cause analysis shifts from agent reasoning to deterministic context engineering Funding Meshy closes $400M Series B for AI-generated 3D at $1.5B valuation; 12M users, 100M models Funding Sila raises $300M to scale silicon-anode battery production; Moses Lake plant targets 100K+ EV capacity Policy White House accuses Moonshot of distilling Anthropic's Fable for Kimi K3; tech leaders defend open distillation Funding General Compute closes $400M debt on inference chips; first deal with SambaNova ASICs as collateral Market Big Tech's hidden AI debt hits $1.65T; off-balance-sheet commitments dwarf reported figures Research Black Forest Labs FLUX 3: Multimodal Video + Audio + Action in One Model; Beats Seedance, Runway Research Anthropic ships Claude Opus 5 at $5/$25 per Mtok—near-Fable-5 coding at half Fable's price Policy Jensen Huang debuts on X with open-weight AI letter, backing models against Washington restriction Chips Intel Foundry Q2: $5.8B revenue, external customer mix still 5% of total Chips Samsung locks $200B with Broadcom: HBM4 + 2nm foundry through 2030 for AI chips
Chips

Azure rolls out AMD Helios 72-GPU racks and 500-core Venice CPUs for AI inference; H2 2026 rollout

Microsoft announced July 20, 2026 that it will deploy AMD's Helios rack-scale AI system and three new Azure VM families beginning H2 2026, targeting AI inference and data-heavy workloads. At the core: Helios packs 72 AMD Instinct MI455X GPUs (CDNA 5, 432GB HBM4 each), Venice EPYC CPUs, Pensando networking, and liquid cooling per rack. Each MI455X delivers 19.6TB/s memory bandwidth; the full rack aggregates 31TB HBM4, 260TB/s scale-up bandwidth, and 43TB/s scale-out networking. Alongside Helios, Microsoft introduces HDv2 VMs (nearly 500 EPYC cores, 4TB RAM, 32TB NVMe) for preprocessing/orchestration, HXv2 for EDA/scientific computing, and ND MI455X v7 (Helios-based inference), all purpose-built for inference not training.

The hardware pivot signals a strategic shift: while Blackwell delays extend Hopper production, Microsoft is diversifying its accelerator base away from NVIDIA mono-dependency. Helios is positioned for frontier-model serving, multi-agent systems, and RAG—workloads where inference throughput and latency matter more than training density. The HDv2's 500 cores address a recognized architectural pain point: CPU preprocessing and orchestration bottleneck GPU clusters. Microsoft is retiring older HBv2 VMs, consolidating around this new stack. ROCm software maturity (PyTorch, TensorFlow, JAX, vLLM support) is critical; compatibility validation will be necessary during GA period.

For infrastructure architects, Azure's AMD-heavy announcement reflects broader market reality: NVIDIA GPU capacity is constrained, Blackwell is delayed, and open-standard inference (ROCm, ONNX, Pensando switching) reduces cloud vendor lock-in. Helios differentiates on bandwidth-per-GPU and integrated cooling; pricing and regional availability remain TBD, but the per-token cost model suggests Azure is optimizing for sustained inference revenue rather than training capex attachment. Watch for ROCm maturity signals and customer adoption metrics in MI455X early-access programs; a successful ramp would substantially reduce Azure's NVIDIA exposure heading into 2027.

Sources