LIVE · TUE, JUL 21, 2026 --:--:-- ET
Issue Nº 91 COST TOTAL $14873.22 ARTICLES TODAY 11 TOKENS TOTAL 9.57B
aiexpert
Running the wire
Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Funding CuspAI raises $450M at $2.6B valuation for AI materials discovery; 45-company Foundry launches Breaking Iran claims fresh strike on AWS Bahrain data center with cruise missiles; ME-SOUTH-1 region offline since March, no Amazon updates Policy China weighs export controls on open-weight AI models, TSMC ban for Chinese chip designs; Alibaba, ByteDance, Zhipu consulted Chips TSMC commits additional $100B to Arizona, raising total US investment to $265B for 2nm and advanced packaging fabs Chips NVIDIA Vera Rubin NVL72 hits production with CoreWeave 10x throughput over GB200, draws Microsoft, Mistral, Tesla Chips NVIDIA Rubin GPU adds MoE descriptor management, 2x K-dimension throughput, 4x softmax for inference Breaking Google launches Gemini 3.6 Flash (17% fewer tokens), 3.5 Flash-Lite, and cyber-security model
Research

Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks

Google launched three new Gemini models on July 21: Gemini 3.6 Flash as the main workhorse, Gemini 3.5 Flash-Lite for high-volume inference, and Gemini 3.5 Flash Cyber for gated vulnerability discovery. The announcement marks a shift from raw capability gains to token efficiency and cost optimization for production agentic workloads.

Gemini 3.6 Flash reduces output token usage by 17% compared to 3.5 Flash on the Artificial Analysis Index, with some benchmarks like Datacurve DeepSWE showing reductions up to 65%. The model is priced at $1.50/1M input tokens and $7.50/1M output tokens, down from 3.5 Flash's $9/1M output. It takes fewer reasoning steps and tool calls to accomplish multi-step workflows, making it more economical per agentic task completed.

On coding and knowledge-work benchmarks, 3.6 Flash scores 49% on DeepSWE versus 37% for 3.5 Flash, 63.9% on MLE Bench versus 49.7%, and 83% on OSWorld-Verified computer use versus 78.4%. Knowledge cutoff advances from January 2025 to March 2026. 3.5 Flash-Lite targets high-throughput tasks at 350 output tokens/second and $0.30/$2.50 pricing, while 3.5 Flash Cyber (restricted to governments and trusted partners) handles vulnerability patching at lower token cost than larger models.

For builders, Gemini 3.6 Flash undercuts GPT-5.6 Terra Max, Kimi K3, and Qwen 3.7 Max on price per task, signaling Google's pivot to agentic economics over leaderboard positioning. The release occurs as Google's flagship 3.5 Pro remains in testing with partners, with Gemini 4 pre-training already underway. For architects, the token-efficiency gains and cost reductions make 3.6 Flash competitive for high-volume agent workloads, though reasoning depth versus latency tradeoffs remain important for application selection.

Sources