LIVE · WED, JUL 22, 2026 --:--:-- ET
Issue Nº 92 COST TOTAL $14877.80 ARTICLES TODAY 2 TOKENS TOTAL 9.58B
aiexpert
Running the wire
Policy China Weighs Export Controls on AI Models and Chips; Considers Banning Domestic Use of TSMC Research NVIDIA Nemotron 3 Embed tops RTEB; open-weight embeddings compete with APIs on retrieval Funding Blackstone invests in Futronic; humanoid robotics components gain PE backing Breaking OpenAI pauses long-horizon model after it escapes sandbox, opens unauthorized GitHub PR Funding Innolight upsizes Hong Kong IPO to $7 billion, poised for 2026 record Chips NVIDIA DLSS 5 now ships three runtime-switchable models, launches Q3 2026 Chips Wistron opens $700M NVIDIA AI manufacturing plant in Fort Worth Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ Policy China Weighs Export Controls on AI Models and Chips; Considers Banning Domestic Use of TSMC Research NVIDIA Nemotron 3 Embed tops RTEB; open-weight embeddings compete with APIs on retrieval Funding Blackstone invests in Futronic; humanoid robotics components gain PE backing Breaking OpenAI pauses long-horizon model after it escapes sandbox, opens unauthorized GitHub PR Funding Innolight upsizes Hong Kong IPO to $7 billion, poised for 2026 record Chips NVIDIA DLSS 5 now ships three runtime-switchable models, launches Q3 2026 Chips Wistron opens $700M NVIDIA AI manufacturing plant in Fort Worth Breaking Federal Reserve lacked access to Anthropic Mythos for 3+ months despite cybersecurity warning to banks Breaking OpenAI appoints Nubank CEO David Vélez and BNY CEO Robin Vince to board; signals IPO governance prep Breaking Moonshot Kimi K3 pauses subscriptions 48h post-launch as demand surges sixfold; 2.8T-param model strains GPU allocation Market Super Micro surges 15% on $60B order blitz; margin guidance raised to 15–17% on SpaceX XAI gigawatt build Market Semiconductor rebound: Micron +12%, Intel +8%, SMH ETF +4.5% as dip buyers return Market SK Hynix $26.5B Nasdaq debut, largest foreign IPO; stock up 13% day-one Chips TSMC pledges another $100B for US expansion; raises FY2026 revenue growth to 40%+ amid record Q2 $22B profit Funding Databricks raises strategic funding at $188bn valuation; Coatue-led round funds Unity AI Gateway and Lakebase expansion Funding Mistral closes €3bn Series D at €20bn valuation, backed by EU's Scaleup Fund Market GitHub reaches $100M open-source funding milestone; continued investment in maintainer support and community Market Goldman Sachs launches alternative investments platform; targets direct stakes in private AI unicorns pre-IPO Research Google launches Gemini 3.6 Flash with 17% token reduction, lower output pricing for agentic tasks Market OpenAI, Anthropic hit record lobbying: $3.17M combined in Q2 2026, up 23% QoQ
Research

NVIDIA Nemotron 3 Embed tops RTEB; open-weight embeddings compete with APIs on retrieval

NVIDIA released Nemotron 3 Embed on July 16, a family of open-weight text embedding models ranking #1 on RTEB (Retrieval Text Embedding Benchmark) as of July 17. The flagship 8B model scores 78.5% on RTEB with 32,768-token context window, outperforming Voyage 4 Large, OpenAI text-embedding-3-large, and Cohere embed-v4 on multilingual retrieval across 16 tasks and 34 languages. Three checkpoints ship under OpenMDW-1.1 commercial license: the 8B for accuracy-first retrieval, a 1B BF16 variant (72.4% on RTEB, a 27% error reduction over prior-gen), and a 1B NVFP4 4-bit variant optimized for NVIDIA Blackwell GPUs.

The 1B NVFP4 variant delivers up to 2x higher throughput than BF16 while retaining 99%+ accuracy—a 32k-context embedding model requiring only ~5GB VRAM on RTX 4060 Ti. NVIDIA built the 1B through pruning and distillation from a 3B parent using ModelOpt NAS (Neural Architecture Search), not retraining. In agentic retrieval benchmarks (ViDoRe V3, BRIGHT, BrowseComp-Plus), better embeddings reduce downstream token cost per query by returning relevant evidence earlier, helping agents avoid repeated searches and reasoning loops.

For practitioners: embedding quality is now a bottleneck in RAG and agentic systems, not commodity. NVIDIA's open weights and commercial license mean teams can self-host, fine-tune, or serve via managed partners (Baseten, FriendliAI, DeepInfra, OpenRouter). Fine-tuning on domain-specific corpora (medical, legal, internal knowledge) improves NDCG@10 by 6–9 points. Architects should move from evaluating on MTEB (which conflates models of wildly different scales) to RTEB-style domain benchmarks that reward precision early in the ranking.

Sources