2. Unleash: We must unleash AI technology and establish an Agentic Government. For ex, we should use AI to improve vet healthcare, improve IRS fraud detection, and agency efficiency. We could also require agencies to establish flagship AI programs.
@alexandr_wang
-
America Must Support AI Innovation With Smart Regulation and Worker Preparation
By
–
3. Innovate: America must support innovation from companies. Best framework allows innovation while still creating guardrails. Congress should take a use-case based reg approach, set federal AI governance standards, and prepare American workers for an AI-driven future.
-
National AI Data Reserve Strategy for Data Dominance
By
–
1. Dominate: We need to establish Data Dominance. We can do this by setting up a National AI Data Reserve, making all govt data AI-ready, and investing to position data dominance as a national priority.
-
Congressional Testimony on AI Competition with China Strategy
By
–
I testified at today's @HouseCommerce hearing on AI.
— Alexandr Wang (@alexandr_wang) 10 avril 2025
The CCP has an AI master plan that’s working.
I told Congress that America must do the following to win on AI: Dominate, Unleash, Innovate, Promote.
🧵 pic.twitter.com/ZfEpFhA7GiI testified at today's @HouseCommerce hearing on AI. The CCP has an AI master plan that’s working. I told Congress that America must do the following to win on AI: Dominate, Unleash, Innovate, Promote.
-

DeepSeek V3 Ranked 8th and 12th on SEAL Leaderboards
By
–
Narrative Violation—DeepSeek V3 is a competitive, but NOT top model. SEAL leaderboards have been updated with DeepSeek V3 (Mar 2025). – 8th on Humanity’s Last Exam (text-only).
– 12th on MultiChallenge (multi-turn). View the full rankings: http://
scale.com/leaderboard -
AI Improvement Through Precision-Targeted Fixes and SEAL Research
By
–
AI improvement at the labs has evolved from a guessing game to one driven by precision-targeted fixes to identified failure modes. @scale_AI
's SEAL research benchmarks and evaluation platform is making this possible. Check out coverage by @willknight -

Gemini 2.5 Pro Exp Tops SEAL Leaderboards Across Multiple Benchmarks
By
–
Gemini 2.5 Pro Exp dropped and it's now #1 across SEAL leaderboards: Humanity’s Last Exam VISTA (multimodal) (tie) Tool Use (tie) MultiChallenge (multi-turn) (tie) Enigma (puzzles) Congrats to @demishassabis @sundarpichai & team! https://
scale.com/leaderboard -

MASK Benchmark Tests AI Honesty Under Pressure
By
–
Part of AI alignment is staying honest under pressure. Can models hold the line when pushed to lie? @scale_ai & @cais release MASK—1,000+ real-world scenarios designed to test AI honesty under pressure we'll release SEAL rankings on a private set https://
mask-benchmark.ai -
xAI restricts model evaluation without consent
By
–
xAI doesn’t allow their models to be evaluated without their consent. we are working on it
-

Leading DoD’s Thunderforge AI Military Planning Program
By
–
Excited to lead one of the DoD's flagship AI programs, Thunderforge, with @DIU_x
. It will be the flagship program within the DoD for AI-based military planning & operations. We'll be working alongside our partners from @microsoft @anduriltech & @google
.