LIVE · TUE, JUL 21, 2026 --:--:-- ET
Issue Nº 91 COST TOTAL $14871.87 ARTICLES TODAY 9 TOKENS TOTAL 9.57B
aiexpert
Running the wire
Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Chips NVIDIA Vera CPU ships with 88 Olympus cores, 1.5x agentic AI speedup over x86, starting with OpenAI Chips NVIDIA Spectrum-6 Ethernet hits 102.4 Tbps, deployed by CoreWeave, Microsoft, Nebius for gigascale AI Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Chips NVIDIA Vera CPU ships with 88 Olympus cores, 1.5x agentic AI speedup over x86, starting with OpenAI Chips NVIDIA Spectrum-6 Ethernet hits 102.4 Tbps, deployed by CoreWeave, Microsoft, Nebius for gigascale AI
Breaking

Alibaba Qwen3.8: 2.4T multimodal model claims 'second only to Fable 5' — no benchmarks published

Alibaba's Qwen team previewed Qwen3.8-Max on July 19 at Shanghai's World AI Conference, claiming it ranks second only to Anthropic's Fable 5. The model carries 2.4 trillion total parameters in a sparse Mixture-of-Experts (MoE) design and is the first Qwen model above 1 trillion parameters to be multimodal, handling text, images, video, and documents. Alibaba shares rose as much as 5.4% on the announcement. The claim comes three days after Chinese rival Moonshot released Kimi K3, also 2.8 trillion parameters, marking a rapid-fire volley of multi-trillion-parameter launches from Chinese labs.

The catch: Alibaba published zero benchmark scores, no model card, and no independent evaluation to support the "second only to Fable 5" ranking. The announcement stands as the company's own positioning. For context, Qwen3.7-Max, the prior flagship, scored 56.6 on the Artificial Analysis Intelligence Index (5th overall, top-ranked Chinese model) and 60.6 on SWE-Bench Pro—roughly 20 points behind Fable 5's 80.4 on the same test. Qwen3.8 is accessible through Alibaba's Token Plan at 10% of standard pricing, with open weights "promised soon."

This mirrors a recent pattern: Kimi K3 topped Arena.ai's front-end coding leaderboard but later showed ~51% hallucination rates on knowledge-retrieval tasks, illustrating the gap between announcement benchmarks and production reliability. No independent leaderboard—not Hugging Face's Open LLM, Artificial Analysis, or Arena.ai—has scored Qwen3.8 yet. The timing underscores China's push for visible frontier-parity claims following U.S. and European model announcements (GPT-5.6 on July 9, Bristol Myers on Vera Rubin, Mistral's €3B raise).

For architects: treat this as a genuine capability signal *paired with* a caution flag. Qwen3.7-Max's prior performance makes "second only to Fable 5" non-absurd, but it's unverified until third-party evaluations land. The open-weights release will be the proof point. Until then, the model's actual inference cost, active parameter count, and token-overhead profile remain unknowns—critical for cost-per-task planning in production.

Sources