LIVE · WED, JUL 29, 2026 --:--:-- ET
Issue Nº 99 COST TOTAL $14985.37 ARTICLES TODAY 2 TOKENS TOTAL 9.71B
aiexpert
Running the wire
Policy xAI sues Minnesota over AI nudification ban; law takes effect Aug 1 Policy Judge approves Anthropic $1.5B copyright settlement; authors get ~$3,100 per work Chips AMD, Cerebras team on disaggregated AI inference; claim 5x efficiency gains Funding Lattice Semiconductor completes AMI acquisition for $650M in stock Market SK Hynix Q2 profit soars 557% YoY on HBM demand; revenue misses estimates amid pricing delays Market Chip volatility hits 30-year high as Cramer calls for fleeing memory winners to banks, industrials Market Bloom Energy Q2 crushes expectations, raises FY26 guidance to $4B+ on AI power demand Breaking Allen AI ships OlmoEarth platform for continent-scale geospatial inference; 10TB satellite models Breaking Together AI ships capacity-aware traffic routing for multi-deployment model inference Market Bloom Energy Q2 beats expectations, raises FY 2026 guidance to $3.9–4.2B revenue Chips NVIDIA Jetson T3000/T2000: Blackwell-Powered Thor Modules for Edge Robotics Scale Breaking Moonshot Kimi K3: US Alleges Distillation, China Denies; 15-Day Evidence Gap Emerges Breaking Sam Altman: AI Security Breach Catalyzes Pace-of-Development Rethink Research AI2 releases OlmoEarth: open-source geospatial foundation models outperforming larger commercial alternatives Chips Nvidia's Huang: semiconductor industry must grow 10x to power 100 billion AI agents Breaking Grafana Assistant Now Queries 30+ Data Sources via Natural Language Research OpenAI publishes field report: coding agents accelerate scientific software modernization in genomics and life sciences Research Liquid AI releases LFM2.5 CPU-efficient encoders: 230M/350M models beat larger alternatives on long-context tasks Market Apple briefly touches $5 trillion market cap, overtakes Nvidia amid AI capex debate Funding Mate Security closes $35M Series A, reaches $50M total funding; grows 500% since Q3 2025 Policy xAI sues Minnesota over AI nudification ban; law takes effect Aug 1 Policy Judge approves Anthropic $1.5B copyright settlement; authors get ~$3,100 per work Chips AMD, Cerebras team on disaggregated AI inference; claim 5x efficiency gains Funding Lattice Semiconductor completes AMI acquisition for $650M in stock Market SK Hynix Q2 profit soars 557% YoY on HBM demand; revenue misses estimates amid pricing delays Market Chip volatility hits 30-year high as Cramer calls for fleeing memory winners to banks, industrials Market Bloom Energy Q2 crushes expectations, raises FY26 guidance to $4B+ on AI power demand Breaking Allen AI ships OlmoEarth platform for continent-scale geospatial inference; 10TB satellite models Breaking Together AI ships capacity-aware traffic routing for multi-deployment model inference Market Bloom Energy Q2 beats expectations, raises FY 2026 guidance to $3.9–4.2B revenue Chips NVIDIA Jetson T3000/T2000: Blackwell-Powered Thor Modules for Edge Robotics Scale Breaking Moonshot Kimi K3: US Alleges Distillation, China Denies; 15-Day Evidence Gap Emerges Breaking Sam Altman: AI Security Breach Catalyzes Pace-of-Development Rethink Research AI2 releases OlmoEarth: open-source geospatial foundation models outperforming larger commercial alternatives Chips Nvidia's Huang: semiconductor industry must grow 10x to power 100 billion AI agents Breaking Grafana Assistant Now Queries 30+ Data Sources via Natural Language Research OpenAI publishes field report: coding agents accelerate scientific software modernization in genomics and life sciences Research Liquid AI releases LFM2.5 CPU-efficient encoders: 230M/350M models beat larger alternatives on long-context tasks Market Apple briefly touches $5 trillion market cap, overtakes Nvidia amid AI capex debate Funding Mate Security closes $35M Series A, reaches $50M total funding; grows 500% since Q3 2025
Chips

AMD, Cerebras team on disaggregated AI inference; claim 5x efficiency gains

AMD and Cerebras announced a technical partnership to deliver a disaggregated AI inference platform combining AMD Helios rackscale systems with Cerebras Wafer-Scale Engine (WSE). Unveiled at Advancing AI 2026 on July 23, the joint solution pairs AMD's high-throughput Helios rack (72 MI455X GPUs per rack, 31TB HBM4) with Cerebras' ultra-low-latency WSE-3 (900,000 cores, 44GB on-die SRAM). Together, they claim up to 5x higher tokens per second per watt versus Cerebras WSE-only configurations.

AMD Helios handles the context/prefill stage—processing prompts and large context windows at high throughput—while Cerebras WSE accelerates the memory-bandwidth-intensive decode/token-generation stage with sub-10ms latency per token. This disaggregated approach mirrors NVIDIA's Rubin CPX strategy but inverts the specialization: AMD targets throughput, Cerebras targets latency. The result is positioned for applications prioritizing real-time interaction: copilots, live agents, autonomous workflows, and scientific discovery.

Cerebras plans to deploy AMD Helios in its own data centers, with the joint solution expected to be available initially through Cerebras Cloud in H2 2026. For architects: this partnership addresses a real bottleneck—NVIDIA's Rubin also pursued disaggregation for the same reason. The question for buyers is feature parity and software ecosystem. AMD and Cerebras must demonstrate that mixed inference across Helios + WSE matches the ease of single-vendor stacks, particularly for teams migrating from NVIDIA.

Sources