Top Topic
Daily AI intelligence
Daily AI Briefing — May 23, 2026
1089 current signals analyzed across AI news, research, social media, and open-source projects.
Daily synthesis
Executive Summary
Top Story
Trump abruptly canceled a planned executive order empowering federal pre-release testing of frontier AI models after overnight lobbying from Elon Musk, Mark Zuckerberg, and David Sacks, with top AI CEOs declining to attend the signing event.
Key Developments
- Microsoft: Released Fara1.5 browser computer-use agents (4B/9B/27B), with the 27B variant scoring 72% on Online-Mind2Web — beating OpenAI Operator (58.3%) and Gemini 2.5 Computer Use (57.3%).
- OpenAI: Opened its first Applied AI Lab outside the US in Singapore, backed by S$300M+ and 200+ technical roles, as global affairs chief Chris Lehane pushes favorable state-level laws.
- DeepSeek: Closed a $10.29B financing round with founder Liang Wenfeng committing to continued open-source releases over commercialization, alongside a permanent 75% API price cut.
- AI Infrastructure: Modal raised $355M at $4.7B, Exa raised $250M at $2.2B, and Turbopuffer hit $100M ARR profitably; NVIDIA quietly removed its gaming revenue category from financial reports.
- Qwen 3.7 Max: Drew attention for running unsupervised for 35 hours, making 1,158 tool calls to achieve a 10x speedup on a GPU kernel.
Safety & Regulation
- Microsoft reportedly cancelled internal Anthropic licenses as token-based billing exceeded budgets, with leaked data showing AI tooling now costs more than human employees in many roles.
- Palantir publicly attacked London Mayor Sadiq Khan after he blocked a £50m Met Police AI contract.
- Standard Chartered CEO Bill Winters apologized for calling the 7,800 staff being cut due to AI "lower-value human capital."
- Perplexity open-sourced Bumblebee, a read-only macOS/Linux scanner that audits developer machines for risky packages, extensions, and MCP configurations.
Research Highlights
- The AI Industrial Explosion (Part 3): Models post-AGI growth by reoptimizing production recipes for cheap labor and capital-intensive factor prices.
- AI2 ArtifactLinker: Found that older open models like DeBERTa often beat newer LLMs on under-evaluated benchmarks, challenging assumptions about progress.
- Tencent Z-Image 6B: Released open-source pixel-space image generation (no VAE, 1k resolution), alongside NVIDIA's LongLive-2.0 (NVFP4 long video) and AI-Q open-source research agent.
- Zvi's review: Positions Gemini 3.5 Flash as best-at-speed but inferior to Claude Opus 4.7 and GPT-5.5 for serious work.
Looking Ahead
With the federal pre-release testing order shelved and OpenAI establishing offshore labs in friendlier jurisdictions, US AI governance is shifting from federal oversight toward an industry-shaped state-level patchwork — even as cost backlash from Microsoft's Anthropic cancellation suggests the economics of agentic deployment may constrain adoption faster than regulation ever could.
Cross-category signals
Top Topics
Top Topic
Coding Agents and the Harness Era
Top Topic
AI Economics and Cost Backlash
Top Topic
Open-Source AI Momentum
Top Topic
AI Infrastructure Buildout
Top Topic
AI Cybersecurity and Supply Chain
Current evidence
AI News
AI policy and industry power dynamics dominated the cycle:
- Trump abruptly canceled an executive order that would have empowered government pre-release testing of frontier AI models, after Elon Musk, Mark Zuckerberg, and David Sacks lobbied against it overnight, citing China-competition concerns.
- OpenAI opened its first Applied AI Lab outside the US in Singapore, backed by S$300M+ and 200+ technical roles, while global affairs chief Chris Lehane pushes favorable state-level laws.
- Palantir publicly attacked London Mayor Sadiq Khan after he blocked a £50m Met Police AI contract.
Model and infrastructure progress:
- Microsoft released Fara1.5 (4B/9B/27B) browser computer-use agents; the 27B model hits 72% on Online-Mind2Web, beating OpenAI Operator (58.3%) and Gemini 2.5 Computer Use (57.3%).
- AI infrastructure unicorns multiplied: Modal ($355M at $4.7B), Exa ($250M at $2.2B), and Turbopuffer ($100M ARR, profitable).
- China completed AI mapping of its national renewable grid, a strategic edge as AI power demand strains Western grids.
Labor and society: Standard Chartered confirmed 7,800 AI-driven job cuts, with CEO Bill Winters apologizing for calling affected staff 'lower-value human capital.'
Musk and Zuckerberg convinced Trump to scrap AI executive order
By Dashveenjit Kaur
Detailed reporting on how Musk, Zuckerberg, and David Sacks lobbied Trump overnight to scrap the AI safety testing executive order, with Trump citing China competition concerns.
Trump abruptly cancels EO signing event after top AI firm CEOs declined to go
By Ashley Belanger
Trump abruptly canceled an executive order signing that would have given the government power to test frontier AI models pre-release, after CEOs declined to attend on 24 hours' notice. Musk and Zuckerberg reportedly lobbied against the order.
AI infrastructure newsletter notes new unicorns: Turbopuffer reached $100M ARR profitably, Exa raised $250M at $2.2B Series C, and Modal raised $355M at $4.7B Series C.
OpenAI is opening its first Applied AI Lab outside the US in Singapore, with over S$300M committed and 200+ technical roles, partnering with Singapore's Ministry of Digital Development.
China’s AI just mapped its entire renewable energy grid. Here’s why the rest of the world should pay attention
By Dashveenjit Kaur
China deployed AI to map its entire renewable energy grid at national scale, addressing the coordination challenge of matching AI's massive electricity demand with renewables. The achievement contrasts with US grid strain.
Current evidence
Research
Today's research and commentary cluster around post-AGI economics, AI safety methodology, and timely model evaluations, with limited novel technical contributions.
Transformative AI Economics
- *The AI Industrial Explosion (Part 3)* models post-AGI growth by reoptimizing production recipes for cheap labor and capital-intensive factor prices.
- *Will we really put data centers in space?* offers concrete cost/thermal analysis of orbital data centers, pushing back on speculative Musk-era claims.
AI Safety & Alignment
- *Which technical AI safety fields are automated first?* uses feedback quality and economic incentives to predict which subfields frontier labs will automate.
- *Counting Arguments in AI Safety* critiques goal-space counting arguments by drawing parallels to Bertrand's Paradox.
- *We made a map of the doom debate* releases an interactive probabilistic tree of AI threat pathways from AI Safety Camp.
- *AI is Not Normal Technology* rebuts Narayanan & Kapoor, citing biosecurity as a concrete disanalogy.
Model Reviews & Practice
- Zvi reviews Gemini 3.5 Flash as best-at-speed but inferior to Claude Opus 4.7 and GPT-5.5 for serious work.
- *Notes on Collaborating with Claude Opus* documents prompting patterns showing reasoned instructions improve compliance.
Strategy & Timelines
- The proposed DIAL distribution (decision importance adjusted for leverage) reframes timelines reasoning around decision-relevant probability mass.
- A philosophical defense of strong longtermism rounds out the field.
Third installment in a series modeling post-AGI economic growth, examining how reoptimizing production recipes for post-AGI factor prices (cheap labor, fast capital reproduction) would accelerate economic doubling beyond what fixed-recipe models suggest. Uses 2017 US input-output tables as baseline.
Analysis of the technical and economic feasibility of orbital data centers (ODCs) for AI compute, examining whether claims by Musk and others about space-based AI are realistic. Concludes that cost-competitiveness depends almost entirely on Starship reusability achieving Falcon-like economics (~$250/kg to orbit).
Zvi's review of Google's recently-released Gemini 3.5 Flash, arguing it's the best at its speed point but not preferable to Opus 4.7 or GPT-5.5 for most uses. Covers other Google I/O announcements.
Which technical AI safety fields are going to be automated first?
By Chamod Kalupahana
Analysis of which technical AI safety subfields are most likely to be automated first by frontier labs, using feedback quality and economic incentive as the two key factors. Notes Anthropic's use of Mythos and UKAISI evaluations as early signals.
Examines the structure of 'counting arguments' in AI doom reasoning (vast goal space → most goals are unfriendly), drawing parallels to Bertrand's Paradox to question whether the chosen measure is principled. Engages with prior LessWrong critiques.
Current evidence
Social Media
OpenAI's Codex Thursday and a strategic pivot dominated discussions, with Greg Brockman declaring 'the model alone is no longer the product' as Codex gained remote Mac control from phones. The framing of harnesses and integrations as the new battleground resonated widely.
- Anthropic's Project Glasswing reported 10,000+ critical vulnerabilities found, with hints that Claude Mythos Preview will scale this further—raising industry adaptation concerns
- Andrew Ng went viral criticizing the White House green card policy as harmful to US AI competitiveness, while separately opposing Harvard's grade caps
- Perplexity open-sourced Bumblebee, a developer machine scanner targeting AI supply chain risks (MCP configs, extensions, packages)
- Ethan Mollick showcased Gemini Omni's native multimodal video editing on the 1896 train film, highlighting Google's distinctive approach
- Alibaba's Qwen 3.7 Max drew attention for running unsupervised 35 hours and 10x'ing a GPU kernel via 1,158 tool calls
- NVIDIA Research shipped LongLive-2.0 (NVFP4 long video) and AI-Q (open-source deep research agent skill)
- AI2 introduced ArtifactLinker, finding that older models like DeBERTa often beat newer LLMs on under-evaluated benchmarks
Last month we launched Project Glasswing, our collaborative AI cybersecurity initiative. Since then,...
By @AnthropicAI
Anthropic announces Project Glasswing has found 10,000+ high/critical-severity vulnerabilities in essential software since launch.
Greg Brockman declares 'the model alone is no longer the product' - signaling shift to integrated AI systems/harnesses.
The new White House policy requiring green card applicants to apply from outside the US is a caprici...
By @AndrewYNg
Andrew Ng criticizes White House green card policy as harmful to AI competitiveness.
Highlights from today’s Codex Thursday launches: 1️⃣ Codex can now securely use apps on your Mac fr...
By @OpenAI
Following yesterday's Social teaser from Sam Altman, OpenAI Codex Thursday launch: Codex can securely use apps on your Mac from your phone even when locked, plus other features.
Today we're open-sourcing Bumblebee, a read-only scanner for macOS and Linux. It checks developer m...
By @perplexity_ai
Perplexity open-sources Bumblebee, a read-only scanner for macOS/Linux developer machines that checks for risky packages, extensions, and AI tool configs.