LIVE · FRI, JUL 24, 2026 --:--:-- ET
Issue Nº 94 COST TOTAL $14918.63 ARTICLES TODAY 2 TOKENS TOTAL 9.62B
aiexpert
Running the wire
Breaking Anthropic upgrades Claude Voice to Opus/Sonnet models; turn-based architecture competes with OpenAI's full-duplex GPT-Live on tool access, not speech naturalness Research LangChain publishes Deep Agents benchmark framework; three eval suites (Harbor-Index, τ³-bench, ContextBench) set long-horizon agent autonomy standard Research Black Forest Labs FLUX 3: unified multimodal model generates video, audio, and robot actions from single architecture Breaking NVIDIA, KAIST launch first joint AI lab in Korea; SK memory co-dev expanded Policy US Genesis Mission deploys $5B+ for 278 AI-for-science projects across 50 states Funding NVIDIA commits $300M ($50M/year) to KAIST agentic AI lab; first joint lab with Korean university Breaking Expedia's STAR platform uses LLMs for incident root-cause analysis; deterministic workflows reduce mean-time-to-recover Breaking Siemens Fuse EDA AI system automates chip design workflows; transforms hours-long tasks into seconds Policy UK dismantles DSIT, elevates AI minister to cabinet; AI brief moves closer to prime minister Funding AMD secures 2GW Anthropic deal, invests $5B in Claude maker to rival Nvidia in AI chips Research Nvidia launches DNA genomics model; learns what token prediction misses in biological data Breaking Black Forest Labs launches FLUX 3 unified multimodal model for image, video, audio, and robot action Breaking Microsoft Project Perception: multi-model security routing cuts Anthropic Mythos cost by 50% Research Moonshot Kimi K3: Chinese open-weight model tops Arena benchmark, outranks Claude on code Funding Anthropic in early talks with Meta for $10B compute deal; third major infrastructure partnership Funding Anthropic Files for IPO; Claude's ARR Hit $47B in May, Overtakes OpenAI Valuation at $965B Series H Policy 21 APEC Economies Endorse Open-Source AI with 'Strong Security Assurance' at Chengdu Summit Breaking OpenAI Project Camellia: 3.2GW Georgia datacenter, $80M community benefits, $71M Codex credits for students through 2032 Breaking FDA's ELSA AI platform reaches 85% staff adoption in two months; governed data and agents reduce drug review from days to 3 minutes Research NVIDIA research at ICML 2026: 145 papers cite Nemotron open models; 2,000 papers use NVIDIA GPUs Breaking Anthropic upgrades Claude Voice to Opus/Sonnet models; turn-based architecture competes with OpenAI's full-duplex GPT-Live on tool access, not speech naturalness Research LangChain publishes Deep Agents benchmark framework; three eval suites (Harbor-Index, τ³-bench, ContextBench) set long-horizon agent autonomy standard Research Black Forest Labs FLUX 3: unified multimodal model generates video, audio, and robot actions from single architecture Breaking NVIDIA, KAIST launch first joint AI lab in Korea; SK memory co-dev expanded Policy US Genesis Mission deploys $5B+ for 278 AI-for-science projects across 50 states Funding NVIDIA commits $300M ($50M/year) to KAIST agentic AI lab; first joint lab with Korean university Breaking Expedia's STAR platform uses LLMs for incident root-cause analysis; deterministic workflows reduce mean-time-to-recover Breaking Siemens Fuse EDA AI system automates chip design workflows; transforms hours-long tasks into seconds Policy UK dismantles DSIT, elevates AI minister to cabinet; AI brief moves closer to prime minister Funding AMD secures 2GW Anthropic deal, invests $5B in Claude maker to rival Nvidia in AI chips Research Nvidia launches DNA genomics model; learns what token prediction misses in biological data Breaking Black Forest Labs launches FLUX 3 unified multimodal model for image, video, audio, and robot action Breaking Microsoft Project Perception: multi-model security routing cuts Anthropic Mythos cost by 50% Research Moonshot Kimi K3: Chinese open-weight model tops Arena benchmark, outranks Claude on code Funding Anthropic in early talks with Meta for $10B compute deal; third major infrastructure partnership Funding Anthropic Files for IPO; Claude's ARR Hit $47B in May, Overtakes OpenAI Valuation at $965B Series H Policy 21 APEC Economies Endorse Open-Source AI with 'Strong Security Assurance' at Chengdu Summit Breaking OpenAI Project Camellia: 3.2GW Georgia datacenter, $80M community benefits, $71M Codex credits for students through 2032 Breaking FDA's ELSA AI platform reaches 85% staff adoption in two months; governed data and agents reduce drug review from days to 3 minutes Research NVIDIA research at ICML 2026: 145 papers cite Nemotron open models; 2,000 papers use NVIDIA GPUs
Research

Black Forest Labs FLUX 3: unified multimodal model generates video, audio, and robot actions from single architecture

Black Forest Labs unveiled FLUX 3 on July 23, a multimodal frontier model jointly trained on images, video, audio, and action prediction within a single unified architecture built on the company's Self-Flow approach. FLUX 3 Video generates up to 20-second clips with native, synced audio from text prompts, images, or video references; it supports video continuation, keyframe transitions, multilingual dialogue, and chaining multiple clips. In early internal evaluations, BFL claims FLUX 3 Video outperforms Runway Gen-4.5 in 77% of head-to-head comparisons and Luma Ray 3.2 in 93%, though these are preliminary, self-reported benchmarks with full methodology to follow at broader availability.

The architecture's key insight is that video generation and action prediction do not require separate foundations—the same learned representations of spatial structure, temporal dynamics, and causality serve both. Alongside FLUX 3 Video, BFL and robotics partner mimic announced FLUX-mimic, a video-action model in early deployment with Audi. FLUX-mimic can fine-tune new robotic manipulation tasks from as little as 30 minutes of robot data, compared to hours with prior approaches. BFL is staging the rollout: FLUX 3 Image (still-image successor to FLUX.2) ships in coming weeks, FLUX 3 Action for commercial partners in parallel, and an open-weight FLUX 3 Dev planned for later in 2026.

For practitioners, this unifies a long-standing architectural bet: that content generation and physical AI share enough underlying representation that one model can serve both. The multimodal training story—learning images teach structure, video teaches motion, audio teaches causality, each constraining the others—echoes recent work on world models. However, execution risks remain: BFL's video numbers are preliminary, FLUX 3 Image is not yet available for evaluation, and open weights are a later-2026 promise. If BFL delivers open weights on schedule, FLUX 3 Dev will be the first open multimodal model spanning video, image, audio, and action in one architecture, reshaping the stack for builders choosing between specialist tools and unified platforms.

Sources