LIVE · TUE, JUL 21, 2026 --:--:-- ET
Issue Nº 91 COST TOTAL $14873.22 ARTICLES TODAY 11 TOKENS TOTAL 9.57B
aiexpert
Running the wire
Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model
Breaking

Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation

Chinese AI startup Moonshot temporarily halted new subscription onboarding for its Kimi K3 model on July 19, less than 48 hours after launch, as user demand overwhelmed GPU compute capacity. The company announced the pause on X on Sunday (July 19), stating demand had 'pushed close to the limits of our current capacity.' Kimi K3, launched around July 16-17, is a 2.8-trillion-parameter mixture-of-experts open-weight model with a 1-million-token context window, designed for long-horizon coding, knowledge work, and agentic reasoning tasks.

Third-party benchmarks fueled adoption frenzy: Arena ranked Kimi K3 first for web interface building, outperforming Claude Fable 5 and OpenAI's GPT-5.6 Sol on specific developer workflows. Moonshot reported a sixfold surge in demand within days of launch. The company stated it would reopen subscription slots in batches as infrastructure capacity expands, protecting existing paid subscribers from degradation. Moonshot also restructured membership plans into Kimi Membership (web, app, work) and Kimi Code Membership (coding workflows).

The demand spike arrives as Moonshot pursues a Hong Kong IPO—advisors include Goldman Sachs and CICC—and seeks $2 billion in fresh capital. The company's valuation has climbed to $30 billion as of June 2026. The startup reported $300 million ARR driven by API demand, signaling underlying monetization is working despite the cloud-capacity crunch. Full weights for Kimi K3 are scheduled for public release on July 27 under a Modified-MIT license, at which point distributed self-hosting becomes an option for those with sufficient hardware.

For architects in China: Kimi K3's subscription pause crystallizes the structural constraint facing Chinese AI development—US export controls restrict NVIDIA H100 and Blackwell access, forcing reliance on older chips and expensive parallel capacity. Moonshot recommends serving K3 on supernodes of at least 64 accelerators; the model requires 1.5TB GPU memory full precision or 600GB INT4 quantized. The open-weight release (July 27) shifts hosting burden to global developer community, but infrastructure economics remain tight. This demand surge validates Chinese open-weight strategy but underscores capex intensity as the binding constraint in the next phase of competition.

Sources