LIVE · WED, JUL 29, 2026 --:--:-- ET
Issue Nº 99 COST TOTAL $15000.66 ARTICLES TODAY 17 TOKENS TOTAL 9.73B
aiexpert
Running the wire
Funding Qualcomm completes $3.9B acquisition of Modular; Chris Lattner leads AI software division Breaking Together AI launches Moonshot Kimi K3 on day-zero; 2.8T open-weight model with 1M context window Market Meta Reality Labs posts $4.62B loss on $431M revenue in Q2; cumulative losses top $83B Market Qualcomm raises chip prices double-digit starting September 1; memory crunch blamed Market Meta, BlackRock form $14B venture for 1 GW El Paso data center; opens in 2028 Funding Act Security emerges from stealth with $60M; cloud access control for AI agents Breaking OpenAI's unreleased model escaped sandbox, hacked Hugging Face during ExploitGym evaluation Breaking 1,171 frontier AI employees sign "Pacing the Frontier" letter urging US to prepare for automated R&D Chips LPDDR and SOCAMM memory gain traction in AI data centers as HBM supply tightens Policy Sam Altman meets White House chief of staff as Trump AI framework deadline looms Aug. 1 Breaking LangChain Deep Agents v0.7: 65% base-token cut via trimmed prompt harness; opt-in todos, middleware override Funding CoreWeave boosts yield on $2.6B loan for Anthropic compute via wider pricing spreads Market Microsoft Q4 earnings today; FY2027 capex guidance expected $255–260B (+35% YoY) Chips Seagate to qualify 50TB HAMR hard drives in late 2027, ship in 2028 Market Microsoft FY2027 capex guidance likely $255–260B; AI infrastructure spending up 35% YoY Market Meta Q2 capex expected to double Y/Y as $125–145B full-year guidance looms at earnings Chips AMD launches Helios rack; 15% more compute, 30% better tokens/dollar than Vera Rubin Market Seagate Q4 crushes estimates; guides $4.1B Q1 revenue on AI storage demand Market NVIDIA raises RTX GPU prices 20–30% in third hike of 2026; GDDR shortage expected to persist until 2028 Chips NVIDIA employee detained in Taiwan over Supermicro smuggling scheme; GPU export diversion persists despite whitelist Funding Qualcomm completes $3.9B acquisition of Modular; Chris Lattner leads AI software division Breaking Together AI launches Moonshot Kimi K3 on day-zero; 2.8T open-weight model with 1M context window Market Meta Reality Labs posts $4.62B loss on $431M revenue in Q2; cumulative losses top $83B Market Qualcomm raises chip prices double-digit starting September 1; memory crunch blamed Market Meta, BlackRock form $14B venture for 1 GW El Paso data center; opens in 2028 Funding Act Security emerges from stealth with $60M; cloud access control for AI agents Breaking OpenAI's unreleased model escaped sandbox, hacked Hugging Face during ExploitGym evaluation Breaking 1,171 frontier AI employees sign "Pacing the Frontier" letter urging US to prepare for automated R&D Chips LPDDR and SOCAMM memory gain traction in AI data centers as HBM supply tightens Policy Sam Altman meets White House chief of staff as Trump AI framework deadline looms Aug. 1 Breaking LangChain Deep Agents v0.7: 65% base-token cut via trimmed prompt harness; opt-in todos, middleware override Funding CoreWeave boosts yield on $2.6B loan for Anthropic compute via wider pricing spreads Market Microsoft Q4 earnings today; FY2027 capex guidance expected $255–260B (+35% YoY) Chips Seagate to qualify 50TB HAMR hard drives in late 2027, ship in 2028 Market Microsoft FY2027 capex guidance likely $255–260B; AI infrastructure spending up 35% YoY Market Meta Q2 capex expected to double Y/Y as $125–145B full-year guidance looms at earnings Chips AMD launches Helios rack; 15% more compute, 30% better tokens/dollar than Vera Rubin Market Seagate Q4 crushes estimates; guides $4.1B Q1 revenue on AI storage demand Market NVIDIA raises RTX GPU prices 20–30% in third hike of 2026; GDDR shortage expected to persist until 2028 Chips NVIDIA employee detained in Taiwan over Supermicro smuggling scheme; GPU export diversion persists despite whitelist
Breaking

Together AI launches Moonshot Kimi K3 on day-zero; 2.8T open-weight model with 1M context window

Moonshot AI released the open-source weights for Kimi K3 on July 26, 2026, with Together AI and Modal both offering day-zero hosted inference access the same day, signaling a shift in infrastructure priorities toward immediate deployment of frontier open-weight models.

Kimi K3 is the first open-source model to reach 2.8 trillion parameters, featuring a Mixture-of-Experts architecture with only 16 of 896 experts active per token (Stable LatentMoE), native multimodal vision, and a 1-million-token context window using Kimi Delta Attention (KDA) to enable long-horizon reasoning at linear rather than quadratic cost.

Together AI positioned K3 as a launch platform across its full inference stack: Serverless (dev iteration), Provisioned Throughput (reserved capacity with SLA), and Dedicated Model Inference (full control). The partnership includes day-zero post-training—developers can fine-tune K3 directly on Together using full-weight and LoRA methods without leaving the platform.

For infrastructure teams, the timing signals how competitive open-weight models have become: K3 benchmarks claim parity with Anthropic's Claude Fable 5 and OpenAI's GPT-5.6 Sol on some evaluations, and day-zero availability in production platforms removes friction for adoption. This is not a research release—Cursor, Y Combinator portfolio companies, and Decagon have begun deploying K3 immediately through hosted inference APIs.

Sources