LIVE · MON, JUL 27, 2026 --:--:-- ET
Issue Nº 97 COST TOTAL $14967.46 ARTICLES TODAY 0 TOKENS TOTAL 9.69B
aiexpert
Running the wire
Chips Samsung Foundry breaks even: 2nm yields hit 55-60%; Tesla contract signals customers are willing to diversify Breaking OpenAI ships Presence: Enterprise agent platform with Codex-driven improvements; resolves 75% of support issues without human handoff Chips NVIDIA Vera Rubin ramps full production; 10x token throughput vs Blackwell, Q3 2026 deployments Chips NVIDIA shelves RTX 50 Super indefinitely; GDDR7 shortage forces gaming GPU cuts for AI demand Research Google ships Gemini 3.6 Flash at 17% fewer tokens, $7.50 output price; 3.5 Pro still delayed Research Physicists demonstrate reservoir computing with 400 oscillating particles in liquid; F1 0.90 on anomaly detection Funding Anthropic begins IPO investor roadshow; October 2026 Nasdaq debut targeted at $965B–$1T valuation Market Server DRAM prices set to rise 13–18% in Q3 2026 as AI demand outpaces supply; shortage extends into 2027 Chips TSMC commits $265B to Arizona, adds advanced packaging; US still ships chips to Taiwan for CoWoS Market Kioxia plunges 16% as memory sector rout spreads; market cap halved in a month Market AI cluster power, not GPUs, is now the bottleneck; grid approval queues hit 24–36 months Chips TSMC Arizona Fab 2 equipment install begins Q3 2026; 3nm production targeted for 2027 Chips TSMC A14 progress hits ~90% yield three months ahead of N2 schedule; mass production target 2H 2028 Funding Chai Discovery closes $400M Series C at $3.8B; AI antibody design now used by Lilly, Pfizer, Novartis Funding Fireworks AI raises $1.5B Series D at $17.5B; surpasses $1B ARR with 40T+ tokens/day, 95% custom models Breaking Neo emerges from stealth with $100M to control agentic AI; Gartner says 40% of enterprise apps will be agentic by EOY Policy Hassabis, Nadella, Altman converge on AI regulation; three rival CEOs back independent model review Policy FCC expands submarine cable licensing to SLTE; shields U.S. tech giants, blocks Chinese access Breaking MCP 2026-07-28 spec removes sessions, adds OAuth hardening; ships July 28 with breaking changes Breaking Anthropic launches Claude Opus 5: near-Fable performance at half the price, effort toggle for cost control Chips Samsung Foundry breaks even: 2nm yields hit 55-60%; Tesla contract signals customers are willing to diversify Breaking OpenAI ships Presence: Enterprise agent platform with Codex-driven improvements; resolves 75% of support issues without human handoff Chips NVIDIA Vera Rubin ramps full production; 10x token throughput vs Blackwell, Q3 2026 deployments Chips NVIDIA shelves RTX 50 Super indefinitely; GDDR7 shortage forces gaming GPU cuts for AI demand Research Google ships Gemini 3.6 Flash at 17% fewer tokens, $7.50 output price; 3.5 Pro still delayed Research Physicists demonstrate reservoir computing with 400 oscillating particles in liquid; F1 0.90 on anomaly detection Funding Anthropic begins IPO investor roadshow; October 2026 Nasdaq debut targeted at $965B–$1T valuation Market Server DRAM prices set to rise 13–18% in Q3 2026 as AI demand outpaces supply; shortage extends into 2027 Chips TSMC commits $265B to Arizona, adds advanced packaging; US still ships chips to Taiwan for CoWoS Market Kioxia plunges 16% as memory sector rout spreads; market cap halved in a month Market AI cluster power, not GPUs, is now the bottleneck; grid approval queues hit 24–36 months Chips TSMC Arizona Fab 2 equipment install begins Q3 2026; 3nm production targeted for 2027 Chips TSMC A14 progress hits ~90% yield three months ahead of N2 schedule; mass production target 2H 2028 Funding Chai Discovery closes $400M Series C at $3.8B; AI antibody design now used by Lilly, Pfizer, Novartis Funding Fireworks AI raises $1.5B Series D at $17.5B; surpasses $1B ARR with 40T+ tokens/day, 95% custom models Breaking Neo emerges from stealth with $100M to control agentic AI; Gartner says 40% of enterprise apps will be agentic by EOY Policy Hassabis, Nadella, Altman converge on AI regulation; three rival CEOs back independent model review Policy FCC expands submarine cable licensing to SLTE; shields U.S. tech giants, blocks Chinese access Breaking MCP 2026-07-28 spec removes sessions, adds OAuth hardening; ships July 28 with breaking changes Breaking Anthropic launches Claude Opus 5: near-Fable performance at half the price, effort toggle for cost control
Research

Google ships Gemini 3.6 Flash at 17% fewer tokens, $7.50 output price; 3.5 Pro still delayed

Google released Gemini 3.6 Flash on July 21, 2026 as its workhorse model for agentic coding, knowledge work, and multimodal tasks. Output token pricing fell from $9.00 to $7.50 per million (input unchanged at $1.50), and the model consumes 17% fewer output tokens per task than Gemini 3.5 Flash on the Artificial Analysis Index, reaching 65% efficiency gains on specific agentic benchmarks like DeepSWE. The combined sticker-price cut plus token efficiency compounds to an effective 31% cost reduction per completed task, with up to 71% savings on agentic coding workloads.

Google also shipped Gemini 3.5 Flash-Lite at $0.30/$2.50 per million tokens for high-throughput tasks and announced Gemini 3.5 Flash Cyber, a security-tuned model restricted to government and partner access. The March 2026 knowledge cutoff represents a 14-month advance over 3.5 Flash's January 2025 date. On applied agentic benchmarks, 3.6 Flash gained ground: DeepSWE improved from 37% to 49%, MLE Bench from 49.7% to 63.9%, and OSWorld computer-use from 78.4% to 83.0%.

However, on the independent Artificial Analysis Intelligence Index, Gemini 3.6 Flash scores exactly 50—unchanged from 3.5 Flash. This is an efficiency release, not a capability jump. The model trades raw reasoning power for token efficiency and lower latency, optimizing for the specific economics of agentic workflows rather than frontier reasoning depth. Artificial Analysis independently measured average task completion time dropping from 2.7 minutes to 1.3 minutes and average cost per task from $0.59 to $0.50.

Notably absent: Gemini 3.5 Pro remains in partner testing with no public availability date, having missed its May 2026 I/O promise and subsequent June target. While Google is shipping three models, the flagship of the 3.5 generation is stalled. Gemini 4 pre-training has begun and will ship later. For teams running high-volume agentic inference, 3.6 Flash's pricing and efficiency gains matter; for those evaluating flagship capability, Google's delivery timeline continues to lag Anthropic and OpenAI.

Sources