LIVE · TUE, JUL 21, 2026 --:--:-- ET
Issue Nº 91 COST TOTAL $14871.87 ARTICLES TODAY 10 TOKENS TOTAL 9.57B
aiexpert
Running the wire
Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Chips NVIDIA Vera CPU ships with 88 Olympus cores, 1.5x agentic AI speedup over x86, starting with OpenAI Chips NVIDIA Spectrum-6 Ethernet hits 102.4 Tbps, deployed by CoreWeave, Microsoft, Nebius for gigascale AI Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Chips NVIDIA Vera CPU ships with 88 Olympus cores, 1.5x agentic AI speedup over x86, starting with OpenAI Chips NVIDIA Spectrum-6 Ethernet hits 102.4 Tbps, deployed by CoreWeave, Microsoft, Nebius for gigascale AI
Chips

NVIDIA Vera CPU ships with 88 Olympus cores, 1.5x agentic AI speedup over x86, starting with OpenAI

NVIDIA detailed its Vera CPU at GTC on Tuesday, revealing an 88-core Arm-based design with 176 threads using proprietary Spatial Multithreading. The chip is engineered for agentic AI workloads—where CPUs orchestrate tool calls, code execution, and sandbox concurrency—and claims 1.5x better performance per sandbox and 1.8x higher agent throughput than x86 competitors (Intel Xeon, AMD EPYC). Vera is entering full production with deliveries already shipped to OpenAI, Anthropic, and SpaceX in June.

Unlike Intel and AMD, which have focused CPU designs on raw core counts for parallel workloads, NVIDIA optimized Vera for single-core speed, memory bandwidth, and latency. The chip consumes 250–450 watts and supports up to 1.5 terabytes of low-power memory per socket. Phoronix benchmarks showed Vera 11% faster than AMD's latest EPYC and 55% faster than Intel's best single-socket Xeon on a curated workload subset. Wolfe Research estimates $5,000 average selling price and 1.3 million unit shipments in 2026.

Vera opens a $200 billion server CPU market that NVIDIA has never addressed. NVIDIA CFO Colette Cress said the company expects $20 billion in Grace and Vera CPU revenue in FY2027, potentially making NVIDIA the world's largest CPU supplier by revenue. For architects, Vera reduces the CPU layer bottleneck that stalls expensive GPUs waiting for agent orchestration, making it critical for reinforcement learning, agentic inference, and code generation stacks where token density depends on CPU latency.

Sources