LIVE · WED, JUL 29, 2026 --:--:-- ET
Issue Nº 99 COST TOTAL $14995.52 ARTICLES TODAY 13 TOKENS TOTAL 9.72B
aiexpert
Running the wire
Chips LPDDR and SOCAMM memory gain traction in AI data centers as HBM supply tightens Policy Sam Altman meets White House chief of staff as Trump AI framework deadline looms Aug. 1 Breaking LangChain Deep Agents v0.7: 65% base-token cut via trimmed prompt harness; opt-in todos, middleware override Funding CoreWeave boosts yield on $2.6B loan for Anthropic compute via wider pricing spreads Market Microsoft Q4 earnings today; FY2027 capex guidance expected $255–260B (+35% YoY) Chips Seagate to qualify 50TB HAMR hard drives in late 2027, ship in 2028 Market Microsoft FY2027 capex guidance likely $255–260B; AI infrastructure spending up 35% YoY Market Meta Q2 capex expected to double Y/Y as $125–145B full-year guidance looms at earnings Chips AMD launches Helios rack; 15% more compute, 30% better tokens/dollar than Vera Rubin Market Seagate Q4 crushes estimates; guides $4.1B Q1 revenue on AI storage demand Market NVIDIA raises RTX GPU prices 20–30% in third hike of 2026; GDDR shortage expected to persist until 2028 Chips NVIDIA employee detained in Taiwan over Supermicro smuggling scheme; GPU export diversion persists despite whitelist Market Apple hits $5T valuation; Tim Cook's final earnings call faces memory crisis + AI strategy pressure Chips Lenovo doubles Whitsett NC AI server facility to 890K sq ft; targets 1,400 nodes/day capacity Market Meta partners BlackRock on $14B El Paso data center; BlackRock owns 80%, Meta operates Market Bloom Energy crushes Q2, raises FY guidance to $3.9–4.2B on AI data center fuel cell demand Funding Centralize raises $15M Series A to map enterprise deal relationships with AI Market Asian semiconductor stocks plunge; SK Hynix misses estimates despite record profit Policy xAI sues Minnesota over AI nudification ban; law takes effect Aug 1 Policy Judge approves Anthropic $1.5B copyright settlement; authors get ~$3,100 per work Chips LPDDR and SOCAMM memory gain traction in AI data centers as HBM supply tightens Policy Sam Altman meets White House chief of staff as Trump AI framework deadline looms Aug. 1 Breaking LangChain Deep Agents v0.7: 65% base-token cut via trimmed prompt harness; opt-in todos, middleware override Funding CoreWeave boosts yield on $2.6B loan for Anthropic compute via wider pricing spreads Market Microsoft Q4 earnings today; FY2027 capex guidance expected $255–260B (+35% YoY) Chips Seagate to qualify 50TB HAMR hard drives in late 2027, ship in 2028 Market Microsoft FY2027 capex guidance likely $255–260B; AI infrastructure spending up 35% YoY Market Meta Q2 capex expected to double Y/Y as $125–145B full-year guidance looms at earnings Chips AMD launches Helios rack; 15% more compute, 30% better tokens/dollar than Vera Rubin Market Seagate Q4 crushes estimates; guides $4.1B Q1 revenue on AI storage demand Market NVIDIA raises RTX GPU prices 20–30% in third hike of 2026; GDDR shortage expected to persist until 2028 Chips NVIDIA employee detained in Taiwan over Supermicro smuggling scheme; GPU export diversion persists despite whitelist Market Apple hits $5T valuation; Tim Cook's final earnings call faces memory crisis + AI strategy pressure Chips Lenovo doubles Whitsett NC AI server facility to 890K sq ft; targets 1,400 nodes/day capacity Market Meta partners BlackRock on $14B El Paso data center; BlackRock owns 80%, Meta operates Market Bloom Energy crushes Q2, raises FY guidance to $3.9–4.2B on AI data center fuel cell demand Funding Centralize raises $15M Series A to map enterprise deal relationships with AI Market Asian semiconductor stocks plunge; SK Hynix misses estimates despite record profit Policy xAI sues Minnesota over AI nudification ban; law takes effect Aug 1 Policy Judge approves Anthropic $1.5B copyright settlement; authors get ~$3,100 per work
Breaking

LangChain Deep Agents v0.7: 65% base-token cut via trimmed prompt harness; opt-in todos, middleware override

LangChain released Deep Agents v0.7, refactoring its base agent harness to cut input tokens by 65% (6k → 2k tokens per turn) while maintaining or improving task performance. Changes include: (1) removal of the default system prompt and general tool-guidance prose, (2) 43% trimming of built-in tool descriptions, (3) TodoListMiddleware now opt-in rather than default. The premise mirrors Anthropic's recent finding that Claude Code's system prompt was reduced >80% for Opus 5 / Fable 5 models without eval degradation, validating the thesis that modern frontier models are over-scaffolded.

Validation ran against a new eval suite spanning autonomous (coding, data analysis), conversational (multi-turn user interaction), and long-context (retrieval + reasoning) tasks across four models: GPT-5.6-Luna, Gemini-3.6-Flash, Claude Sonnet 4.6, and Claude Opus 4.8. GPT-5.6-Luna showed 34% token reduction and 15% cost reduction with +4% reward lift. Most models held steady on reward while dropping tokens/cost. Claude Sonnet 4.6 was an exception (cost up on two challenging autonomous tasks), flagged in LangSmith traces.

For platform teams: this is a production ops win. Reducing base harness tokens from 6k to 2k per turn means cheaper inference, lower latency, and more room for context on constrained deployments. The opt-in middleware approach (TodoListMiddleware, SummarizationMiddleware override) gives builders fine-grained control over prompt strategy without fighting the framework. The timing aligns with a broader pattern: as model capability stabilizes, harness-level efficiency becomes the lever.

Sources