LIVE · MON, JUL 27, 2026 --:--:-- ET
Issue Nº 97 COST TOTAL $14967.46 ARTICLES TODAY 0 TOKENS TOTAL 9.69B
aiexpert
Running the wire
Chips Samsung Foundry breaks even: 2nm yields hit 55-60%; Tesla contract signals customers are willing to diversify Breaking OpenAI ships Presence: Enterprise agent platform with Codex-driven improvements; resolves 75% of support issues without human handoff Chips NVIDIA Vera Rubin ramps full production; 10x token throughput vs Blackwell, Q3 2026 deployments Chips NVIDIA shelves RTX 50 Super indefinitely; GDDR7 shortage forces gaming GPU cuts for AI demand Research Google ships Gemini 3.6 Flash at 17% fewer tokens, $7.50 output price; 3.5 Pro still delayed Research Physicists demonstrate reservoir computing with 400 oscillating particles in liquid; F1 0.90 on anomaly detection Funding Anthropic begins IPO investor roadshow; October 2026 Nasdaq debut targeted at $965B–$1T valuation Market Server DRAM prices set to rise 13–18% in Q3 2026 as AI demand outpaces supply; shortage extends into 2027 Chips TSMC commits $265B to Arizona, adds advanced packaging; US still ships chips to Taiwan for CoWoS Market Kioxia plunges 16% as memory sector rout spreads; market cap halved in a month Market AI cluster power, not GPUs, is now the bottleneck; grid approval queues hit 24–36 months Chips TSMC Arizona Fab 2 equipment install begins Q3 2026; 3nm production targeted for 2027 Chips TSMC A14 progress hits ~90% yield three months ahead of N2 schedule; mass production target 2H 2028 Funding Chai Discovery closes $400M Series C at $3.8B; AI antibody design now used by Lilly, Pfizer, Novartis Funding Fireworks AI raises $1.5B Series D at $17.5B; surpasses $1B ARR with 40T+ tokens/day, 95% custom models Breaking Neo emerges from stealth with $100M to control agentic AI; Gartner says 40% of enterprise apps will be agentic by EOY Policy Hassabis, Nadella, Altman converge on AI regulation; three rival CEOs back independent model review Policy FCC expands submarine cable licensing to SLTE; shields U.S. tech giants, blocks Chinese access Breaking MCP 2026-07-28 spec removes sessions, adds OAuth hardening; ships July 28 with breaking changes Breaking Anthropic launches Claude Opus 5: near-Fable performance at half the price, effort toggle for cost control Chips Samsung Foundry breaks even: 2nm yields hit 55-60%; Tesla contract signals customers are willing to diversify Breaking OpenAI ships Presence: Enterprise agent platform with Codex-driven improvements; resolves 75% of support issues without human handoff Chips NVIDIA Vera Rubin ramps full production; 10x token throughput vs Blackwell, Q3 2026 deployments Chips NVIDIA shelves RTX 50 Super indefinitely; GDDR7 shortage forces gaming GPU cuts for AI demand Research Google ships Gemini 3.6 Flash at 17% fewer tokens, $7.50 output price; 3.5 Pro still delayed Research Physicists demonstrate reservoir computing with 400 oscillating particles in liquid; F1 0.90 on anomaly detection Funding Anthropic begins IPO investor roadshow; October 2026 Nasdaq debut targeted at $965B–$1T valuation Market Server DRAM prices set to rise 13–18% in Q3 2026 as AI demand outpaces supply; shortage extends into 2027 Chips TSMC commits $265B to Arizona, adds advanced packaging; US still ships chips to Taiwan for CoWoS Market Kioxia plunges 16% as memory sector rout spreads; market cap halved in a month Market AI cluster power, not GPUs, is now the bottleneck; grid approval queues hit 24–36 months Chips TSMC Arizona Fab 2 equipment install begins Q3 2026; 3nm production targeted for 2027 Chips TSMC A14 progress hits ~90% yield three months ahead of N2 schedule; mass production target 2H 2028 Funding Chai Discovery closes $400M Series C at $3.8B; AI antibody design now used by Lilly, Pfizer, Novartis Funding Fireworks AI raises $1.5B Series D at $17.5B; surpasses $1B ARR with 40T+ tokens/day, 95% custom models Breaking Neo emerges from stealth with $100M to control agentic AI; Gartner says 40% of enterprise apps will be agentic by EOY Policy Hassabis, Nadella, Altman converge on AI regulation; three rival CEOs back independent model review Policy FCC expands submarine cable licensing to SLTE; shields U.S. tech giants, blocks Chinese access Breaking MCP 2026-07-28 spec removes sessions, adds OAuth hardening; ships July 28 with breaking changes Breaking Anthropic launches Claude Opus 5: near-Fable performance at half the price, effort toggle for cost control
Chips

NVIDIA shelves RTX 50 Super indefinitely; GDDR7 shortage forces gaming GPU cuts for AI demand

NVIDIA has indefinitely delayed the RTX 50 Super lineup—including a completed 24GB RTX 5080 Super design—due to a structural GDDR7 memory shortage driven by AI data-center demand. According to The Information and corroborated by PCWorld, TrendForce, and PC Gamer, NVIDIA cut RTX 50-series consumer GPU production by 30–40% in the first half of 2026 and shelved the Super refresh in December 2025, redirecting memory capacity toward higher-margin AI accelerators (H200, Blackwell). This marks the first time in nearly three decades that NVIDIA is not shipping a new gaming GPU architecture in a calendar year. The RTX 5090, once $1,999 MSRP, now sells for $4,329 on Amazon, with premium board partners listing above $5,000.

The root cause is not fab capacity but memory: SK Hynix, Samsung, and Micron have allocated the vast majority of 2026 HBM (High Bandwidth Memory) production to data-center customers, and GDDR7 modules have become scarce. A 3GB GDDR7 die (used in high-end gaming cards) now costs $60–70 each, compared to ~$20 for 2GB chips; for an RTX 5080 Super with eight 3GB modules, memory alone runs $480–560 in bill-of-materials. IDC forecasts AI data centers will consume 70% of global memory output in 2026, up from 20–30% in 2022. Intel CEO Lip-Bu Tan and SK Hynix both signaled relief won't arrive until 2027–2028.

AMD is also hit: the company raised GDDR memory prices to AIB partners by ~10% effective July 2026 (second increase in six months) and announced 10–15% price hikes on the Radeon RX 9000 lineup in H2 2026. Console makers (Nintendo, Sony) and PC component vendors are all facing the same reallocation. IDC classified the shortage as a "potentially permanent strategic reallocation," not a cyclical correction.

For practitioners: the shortage is reshaping the entire compute market. If you need a GPU today, last-gen cards (RTX 4070-class, RX 7800 XT) using pre-shortage memory face less pricing pressure. If you're designing for 2027, assume higher memory costs are structural—they constrain single-GPU configurations and favor architectures using FP8 quantization (~50% memory savings) or inference-optimized SKUs. The gaming GPU market has become secondary to AI; NVIDIA will continue directing supply upstream until data-center utilization matures.

Sources