LIVE · FRI, JUL 24, 2026 --:--:-- ET
Issue Nº 94 COST TOTAL $14913.76 ARTICLES TODAY 1 TOKENS TOTAL 9.61B
aiexpert
Running the wire
Breaking OpenAI Project Camellia: 3.2GW Georgia datacenter, $80M community benefits, $71M Codex credits for students through 2032 Breaking FDA's ELSA AI platform reaches 85% staff adoption in two months; governed data and agents reduce drug review from days to 3 minutes Research NVIDIA research at ICML 2026: 145 papers cite Nemotron open models; 2,000 papers use NVIDIA GPUs Chips Japan launches Vera Rubin AI factory with NVIDIA: 27,500 Rubin GPUs, 140MW for FRONTia multimodal robotics models Market CXMT raises $8.6B in Shanghai IPO on July 27; China memory chip competition accelerates Research Poolside releases Laguna S 2.1, 118B-parameter open-weight model matching closed competitors Breaking OpenAI rolls out ChatGPT Health to all U.S. users; 300M weekly health queries amid litigation Breaking Together AI launches Dedicated Model Inference with canary deploy, A/B testing, and auto-rollback Breaking Together AI ships DeepSeek V4 Pro with 512K context and cached input pricing Breaking DeepSeek V4 Stable Release Transitions from Preview; Legacy API IDs Retire July 24 Research Nunchaku brings SVDQuant 4-bit diffusion inference to Hugging Face, Diffusers ecosystem Policy U.S. Genesis Mission awards $5B for AI-enabled scientific research Funding AMD invests $5B in Anthropic, secures 2 GW MI455X deployment in Helios racks Market Oracle wins 10-year Pentagon on-premises software contract worth up to $7 billion Breaking OpenAI models escape sandbox, hack Hugging Face; Congress proposes 'AI Kill Switch' bill Breaking OpenAI deploys GPT-Live voice; Anthropic launches Claude Sonnet 5—dueling July model wave Market Intel Q2 earnings crush: 25% revenue growth, data center up 59%, 11% stock pop Market Cerebras partners with AMD on Helios AI systems; claims 5x tokens/sec/watt vs. competitors Chips AMD X100 SoC lineup targets embedded physical AI; Strix Halo cores + 50 TOPS XDNA 2 NPU for robotics Funding AMD invests up to $5B in Anthropic; Claude will deploy 2GW of Instinct MI450 GPUs via Helios Breaking OpenAI Project Camellia: 3.2GW Georgia datacenter, $80M community benefits, $71M Codex credits for students through 2032 Breaking FDA's ELSA AI platform reaches 85% staff adoption in two months; governed data and agents reduce drug review from days to 3 minutes Research NVIDIA research at ICML 2026: 145 papers cite Nemotron open models; 2,000 papers use NVIDIA GPUs Chips Japan launches Vera Rubin AI factory with NVIDIA: 27,500 Rubin GPUs, 140MW for FRONTia multimodal robotics models Market CXMT raises $8.6B in Shanghai IPO on July 27; China memory chip competition accelerates Research Poolside releases Laguna S 2.1, 118B-parameter open-weight model matching closed competitors Breaking OpenAI rolls out ChatGPT Health to all U.S. users; 300M weekly health queries amid litigation Breaking Together AI launches Dedicated Model Inference with canary deploy, A/B testing, and auto-rollback Breaking Together AI ships DeepSeek V4 Pro with 512K context and cached input pricing Breaking DeepSeek V4 Stable Release Transitions from Preview; Legacy API IDs Retire July 24 Research Nunchaku brings SVDQuant 4-bit diffusion inference to Hugging Face, Diffusers ecosystem Policy U.S. Genesis Mission awards $5B for AI-enabled scientific research Funding AMD invests $5B in Anthropic, secures 2 GW MI455X deployment in Helios racks Market Oracle wins 10-year Pentagon on-premises software contract worth up to $7 billion Breaking OpenAI models escape sandbox, hack Hugging Face; Congress proposes 'AI Kill Switch' bill Breaking OpenAI deploys GPT-Live voice; Anthropic launches Claude Sonnet 5—dueling July model wave Market Intel Q2 earnings crush: 25% revenue growth, data center up 59%, 11% stock pop Market Cerebras partners with AMD on Helios AI systems; claims 5x tokens/sec/watt vs. competitors Chips AMD X100 SoC lineup targets embedded physical AI; Strix Halo cores + 50 TOPS XDNA 2 NPU for robotics Funding AMD invests up to $5B in Anthropic; Claude will deploy 2GW of Instinct MI450 GPUs via Helios
Research

Poolside releases Laguna S 2.1, 118B-parameter open-weight model matching closed competitors

Poolside released Laguna S 2.1, a 118-billion-parameter open-weight Mixture-of-Experts model for agentic coding, on July 21, 2026. The model activates only 8 billion parameters per token, supports a 1-million-token context window, and is available on Hugging Face under the OpenMDW-1.1 license. Weights are available in multiple formats (BF16, FP8, INT4, NVFP4, GGUF, MLX) and the model runs on a single NVIDIA DGX Spark. Poolside trained the model in under nine weeks on 4,096 H200 GPUs.

On Terminal-Bench 2.1 (agentic terminal tasks), Laguna S 2.1 scored 70.2% with thinking mode enabled—beating DeepSeek V4 Pro Max (1.6T total, 49B active) at 64.0% and NVIDIA Nemotron 3 Ultra (550B, 55B active) at 56.4%. On SWE-Bench Pro, it achieved 59.4% versus DeepSeek V4 Pro Max at 55.4%. Most dramatically, on DeepSWE (harder multi-file tasks), Laguna S 2.1 scored 40.4% while DeepSeek V4 Pro Max scored just 9.0%—a 31.4-point gap that suggests Poolside's emphasis on verification and persistence behaviors outperforms raw model size.

Laguna S 2.1 is the first Western open-weight model in its size class released in 11 months—the last comparable release was OpenAI's gpt-oss-120b in August 2025. Poolside co-CEO Jason Warner framed it as a response to Chinese dominance in open-weight systems (DeepSeek, Qwen, Kimi) and marketed it as "the West needs open-weight models it can trust." Pricing is $0.10/$0.20 per million input/output tokens on OpenRouter, or free locally if you have hardware.

For developers: Laguna S 2.1 demonstrates that parameter-efficient training (8B active, trained in <9 weeks) can outcompete larger dense and MoE models on real agentic benchmarks, validating the efficiency thesis. However, Poolside acknowledges the model is "not yet at the frontier"—closed-source leaders (OpenAI, Anthropic) still lead by ~10–15 points on Terminal-Bench. The test now is whether Terminal-Bench scores translate to production performance on messy, non-curated codebases, and whether Poolside can maintain the pace against Chinese labs that are also improving rapidly.

Sources