Key Takeaways
- NVIDIA emphasizes data center water efficiency and new AI agent toolkits. Liquid cooling can reduce water usage to near zero, and the Agent Toolkit enables domain-specific AI agents.
- OpenAI expands Daybreak cybersecurity initiative with GPT-5.5-Cyber and partnerships. Focus on automating vulnerability patching and security fixes.
- xAI integrates Grok with Interactive Brokers for portfolio insights. A practical AI fintech application.
- Menlo Fund raises $3B for AI bets, including Anthropic and Suno. Long-term commitment with lean partnership model.
- GLM-5.2 open-source model now available via Perplexity. Shows strong performance in long-context coding and agent workflows.
- AI engineer workforce polarizing into 'lazy' and 'craftsman' archetypes. Deedy warns of burnout and quality decline.
1. AI Infrastructure and Efficiency
- NVIDIA data centers use only 0.2% of US daily water. Liquid cooling at 45°C can cut water consumption from ~2.6 million gallons/MW/year to near zero, while improving energy efficiency and enabling heat recovery. — via 1
- NVIDIA and SGLang achieved 5x throughput improvement serving DeepSeek-V4 on GB300 using W4A4 MegaMoE quantization and MTP acceptance rate increase from 0.57 to 0.70. — via 1
- DFlash open-source lightweight block diffusion model for speculative decoding delivers up to 15x inference throughput on NVIDIA Blackwell, integrated with SGLang, TensorRT-LLM, and vLLM. — via 1
- NVIDIA Agent Toolkit released, integrating Nemotron models, tools, skills, and safety runtime for domain-specific AI agents. — via 1
2. AI Safety and Cybersecurity
- OpenAI expands Daybreak to democratize patching vulnerable software at machine speed, including Codex security plugin, GPT-5.5-Cyber model, cyber partner program, and Patch the Planet project. — via 1 2
- Sam Altman: GPT-5.5-Cyber achieves state-of-the-art on CyberGym; Patching The Planet and Codex Security will help fix security issues rather than just finding them. — via 1
- Ethan Mollick warns that all Mythos-level models carry similar risks, and open-sourcing such models (if China allows) in 6–12 months could increase danger; governments lack clear risk perception. — via 1
3. Model Releases and Integrations
- GLM-5.2 open-source model now available via Perplexity Agent API, strong in long-context coding and agent workflows, combining frontier reasoning with real-time programming search. — via 1
- Aravind Srinivas notes GLM model rekindles interest in open-source AI, matching frontier models on mid-level knowledge worker tasks in blind tests, with low deployment cost and sub-trillion parameters. More trillion-parameter open-source models coming, beneficial for token pricing and Jevons effect. — via 1 2
- Mistral OCR 4: creates structure with bounding boxes, block classification, and confidence scores in 170 languages. — via 1
- xAI integrates Grok with Interactive Brokers, allowing AI to view portfolio information. — via 1
4. Industry and Investment
- Menlo Fund raises $3B on its 50th anniversary, investing in Anthropic, Suno, Modal, etc. Each partner makes ~2 deals per year, covering seed to Series X, with deep industry experience. — via 1
- Cred founder Kunal Shah appointed as WhatsApp global head; Cred valued at $4.5B, $333M annual revenue, first profitable quarter, ~17M users, but Indian credit card penetration only 3-4%. — via 1
- Deedy warns of AI engineer workforce polarization: 'lazy' engineers rely on AI while 'craftsman' engineers face burnout from increased review pressure and quality decline, common in large companies with 10+ year experience. — via 1
