Research
Curated daily papers, lab research blogs and national AI research programs across the US, China, EU, UK, Japan and beyond.
Curated daily papers, lab research blogs and national AI research programs across the US, China, EU, UK, Japan and beyond.
Latest in Research
30 storiesExplaining Attention with Program Synthesis
A longstanding goal of research on interpretable deep learning is to replace opaque neural computations with human-meaningful symbolic descriptions. In this pap…
Data Intelligence Agents: Interpreting, Modeling, and Querying Enterprise Data via Autonomous Coding Agents
Production data integration is bottlenecked by repeated, lossy handoffs between data owners, engineers, and analysts who must collaboratively discover, structur…
Reference-Driven Multi-Speaker Audio Scene Generation from In-the-Wild Priors
Existing multi-speaker dialogue systems bind speakers to utterances through structured supervision: per-turn tags, multi-stream transcriptions, or learnable spe…
Rethinking Reward Supervision: Rubric-Conditioned Self-Distillation
Post-training of reasoning language models is commonly driven by supervised distillation and reinforcement learning with verifiable rewards. Distillation often …
The Chandra-Gaia Catalog of Counterparts: Resolving ambiguous Gaia matches to X-ray sources in the Chandra Source Catalog using Machine Learning
We present a framework to cross-match sources from the Chandra Source Catalog (CSC v2.1) with optical sources from Gaia Data Release 3. Unlike purely spatial ap…

State of the blog, mid-2026
About 3 years since I started writing weekly.…

Frontier post-training recipe review with Finbarr Timbers
"Interview" #18…

Import AI 461: "Alignment is not on track"; FrontierCode; and synthetic research interns
Where are your agents right now?…

Welcome to the AGI era of AI governance
It's a one-way door and we weren't ready for it.…

Ire identifies another LOTUSLITE specimen
Project Ire examined a timely malware sample and determined its intent through reverse engineering—identifying LOTUSLITE characteristics even as most major EDR …

Scaling Up Reinforcement Learning for Traffic Smoothing: A 100-AV Highway Deployment
Training Diffusion Models with Reinforcement Learning We deployed 100 reinforcement learning (RL)-controlled cars into rush-hour highway traffic to smooth conge…

Repurposing Protein Folding Models for Generation with Latent Diffusion
PLAID is a multimodal generative model that simultaneously generates protein 1D sequence and 3D structure, by learning the latent space of protein folding model…

Defending against Prompt Injection with Structured Queries (StruQ) and Preference Optimization (SecAlign)
Recent advances in Large Language Models (LLMs) enable exciting LLM-integrated applications. However, as LLMs have improved, so have the attacks against them. P…

Whole-Body Conditioned Egocentric Video Prediction
.modal { display: none; position: fixed; z-index: 9999; padding-top: 50px; left: 0; top: 0; width: 100%; height: 100%; overflow: auto; background-color: rgba(0,…

Achieving 10,000x training data reduction with high-fidelity labels
Human-Computer Interaction and Visualization…

What exactly does word2vec learn?
What exactly does word2vec learn, and how? Answering this question amounts to understanding representation learning in a minimal yet interesting language modeli…

Sensible Agent: A framework for unobtrusive interaction with proactive AR agents
Human-Computer Interaction and Visualization…
Introducing interactive on-device segmentation in Snapseed
Human-Computer Interaction and Visualization…

RL without TD learning
In this post, I’ll introduce a reinforcement learning (RL) algorithm based on an “alternative” paradigm: divide and conquer. Unlike traditional methods, this al…

Information-Driven Design of Imaging Systems
An encoder (optical system) maps objects to noiseless images, which noise corrupts into measurements. Our information estimator uses only these noisy measuremen…

Import AI 441: My agents are working. Are yours?
Plus: Corrupting AI systems with a poison fountain…

Import AI 442: Winners and losers in the AI economy; math proof automation; and industrialization of cyber espionage
Is superintelligence a phase change or a gradual shift?…

Import AI 443: Into the mist: Moltbook, agent ecologies, and the internet in transition
Plus, a story about agents corrupting other agents…

Import AI 444: LLM societies; Huawei makes kernels with AI; ChipBench
How can you quantify creativity?…

Beyond one-on-one: Authoring, simulating, and testing dynamic human-AI group conversations
Human-Computer Interaction and Visualization…

Import AI 445: Timing superintelligence; AIs solve frontier math proofs; a new ML research benchmark
Will 2026 be looked back on as the pivotal year for making decisions about the singularity?…

Import AI 446: Nuclear LLMs; China's big AI benchmark; measurement and AI policy
Will AIs be jealous of one another?…

Import AI 447: The AGI economy; testing AIs with generated games; and agent ecologies
What might a superintelligence arcology be like?…

Olmo Hybrid and future LLM architectures
The latest Olmo model and discussions at the frontier of open-source post training tools.…

Dean Ball on open models and government control
Subtle precedents on the future of open models set by the unfolding Anthropic v. Department of War case.…