SIGNAL
Tracking the global AI frontier — labs · research · agents · policy
Frontier Signal

Briefing

The fast feed: breaking AI news from the global tech press, deduplicated and time-ordered.

The fast feed: breaking AI news from the global tech press, deduplicated and time-ordered.

Latest in Briefing

30 stories
Deepseek's DSpark boosts AI speed by up to 85 percent, a strategic win under tightening US export controls
Briefing

Deepseek's DSpark boosts AI speed by up to 85 percent, a strategic win under tightening US export controls

Deepseek's new DSpark framework boosts per-user response speed by 60 to 85 percent. A small model proposes token candidates that the larger model checks in batc…

Benchmarking Geospatial Foundation Models for Agriculture Applications
Research

Benchmarking Geospatial Foundation Models for Agriculture Applications

Geospatial foundation models pretrained on satellite imagery promise broad generalization across remote sensing tasks and regions, but their geographic transfer…

UniVAD v2: Unified Visual Anomaly Detection via Support-Conditioned Boundary Construction
Research

UniVAD v2: Unified Visual Anomaly Detection via Support-Conditioned Boundary Construction

Unified visual anomaly detection seeks to train a single detector that can be deployed across categories, domains, and application scenarios. In the few-shot tr…

DeepTrans Studio: Turning Expert Interventions into Shared Team Knowledge in Agentic Translation Workflows
Research

DeepTrans Studio: Turning Expert Interventions into Shared Team Knowledge in Agentic Translation Workflows

Professional translation is often a team-based process: translators, reviewers, and project managers must coordinate terminology, legal force, and accountabilit…

How Far Do On-Prem Open LLMs Get on Text-to-SQL? A Cross-Family Size x Technique Frontier on BIRD
Research

How Far Do On-Prem Open LLMs Get on Text-to-SQL? A Cross-Family Size x Technique Frontier on BIRD

Organizations that cannot send data to a cloud API increasingly ask: how good is Text-to-SQL if the model must run on-premises on open weights, and which popula…

HTC-SGA Former: A Hybrid Transformer-CNN Network with Self-Guided Attention and a New Boundary-Weighted Adaptive Loss for Coronary DSA Vessel Segmentation
Research

HTC-SGA Former: A Hybrid Transformer-CNN Network with Self-Guided Attention and a New Boundary-Weighted Adaptive Loss for Coronary DSA Vessel Segmentation

Accurate coronary Digital Subtraction Angiography (DSA) vessel segmentation is essential for computer-aided diagnosis and treatment planning of coronary artery …

Towards Generalizable and Evidential Nuclear Magnetic Resonance-Based Molecular Structure Elucidation via Large Language Model Agent
Research

Towards Generalizable and Evidential Nuclear Magnetic Resonance-Based Molecular Structure Elucidation via Large Language Model Agent

Nuclear Magnetic Resonance (NMR) spectroscopy is the gold standard for molecular structure elucidation, yet interpreting complex spectra for unknown molecules r…

Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data
Research

Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data

Demand for personalized financial advising is growing, but consistent advisor expertise is difficult to obtain, scale, and encode in LLM systems. Simple persona…

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering
Research

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires continuous en…

The Forgetting-Retention Dilemma: Certified Unlearning Theory in Continual Learning
Research

The Forgetting-Retention Dilemma: Certified Unlearning Theory in Continual Learning

Machine unlearning aims to eliminate the influence of specific data from trained models to safeguard privacy. However, this presents a significant challenge in …

Robust Trajectory Distillation: Hybrid Reweighting Meets Teacher-Inspired Targets
Research

Robust Trajectory Distillation: Hybrid Reweighting Meets Teacher-Inspired Targets

Dataset distillation (DD) condenses large corpora into compact, information-rich subsets for efficient training and reuse. However, under noisy supervision, DD …

Theory of Continual Learning Against Data Poisoning Attacks
Research

Theory of Continual Learning Against Data Poisoning Attacks

Continual learning (CL), where a model is trained on a sequence of data tasks, is increasingly being adopted across key fields such as large language models and…

Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding
Research

Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding

Flicker-banding (FB), arises from temporal aliasing between a camera's rolling shutter and a display's brightness modulation, degrading screen-captured image re…

ARKD: Adaptive Reinforcement Learning-Guided Bidirectional KL Divergence Distillation for Text Generation
Research

ARKD: Adaptive Reinforcement Learning-Guided Bidirectional KL Divergence Distillation for Text Generation

Knowledge distillation (KD) is a key technique for compressing Large Language Models (LLMs), yet methods relying on a single KL objective often fail to balance …

Decision-Value Attribution in Predict-then-Optimize Systems
Research

Decision-Value Attribution in Predict-then-Optimize Systems

Predictive models are increasingly embedded in operational decision-making, yet standard explanation methods typically explain forecasts rather than the decisio…

Exploiting Local Flatness for Efficient Out-of-Distribution Detection
Research

Exploiting Local Flatness for Efficient Out-of-Distribution Detection

Detecting out-of-distribution (OOD) data is crucial for reliable machine learning deployment. Among detection strategies, post-hoc methods are particularly attr…

SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows
Research

SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows

Spreadsheets are widely used for business analysis, financial modeling, reporting, and decision-making. However, most existing spreadsheet benchmarks evaluate i…

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D
Research

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view: it is h…

A multi-architecture study of specificity refinement and false-positive mechanism analysis in prostate MRI
Research

A multi-architecture study of specificity refinement and false-positive mechanism analysis in prostate MRI

Objectives: To characterize residual false positives in prostate MRI detection, and to evaluate a lightweight post-hoc refinement head for case-level specificit…

Rigel: Self-Distilled Score Adaptation for Image and Video Captioning Evaluation
Research

Rigel: Self-Distilled Score Adaptation for Image and Video Captioning Evaluation

Automatic evaluation of image and video captioning is essential for benchmarking multimodal systems, although standard evaluation metrics show limited alignment…

Parametric Skills
Research

Parametric Skills

Since intelligence fundamentally relies on efficient skill acquisition (Chollet, 2019), the ability to leverage skills is critical. For LLMs, skills, manually a…

Monte Carlo Energy Aggregation for Mobile 3D Gaussian Splatting
Research

Monte Carlo Energy Aggregation for Mobile 3D Gaussian Splatting

Recent advances in 3D Gaussian Splatting have demonstrated unprecedented success in novel view synthesis. However, the substantial inference and storage overhea…

OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data
Research

OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data

Music-driven dance video generation aims to synthesize expressive human motion that is temporally aligned with music while maintaining high visual fidelity. Des…

Data-Driven Energy-Based Learning via Gibbs Measures on Hierarchical Structures
Research

Data-Driven Energy-Based Learning via Gibbs Measures on Hierarchical Structures

We introduce a data-driven probabilistic framework for learning systems based on Gibbs measures on hierarchical structures. Unlike standard empirical risk minim…

Neural Subspace Reallocation: Continual Learning as Retrieval-Based Subspace Memory Management
Research

Neural Subspace Reallocation: Continual Learning as Retrieval-Based Subspace Memory Management

We introduce Neural Subspace Reallocation (NSR), which reframes continual learning as memory management over parameter subspaces. Instead of treating Low-Rank A…

Information Dynamics of Language Communication
Research

Information Dynamics of Language Communication

Quantifying how meaning propagates through communicative exchanges remains underdeveloped in computational linguistics. Here we introduce an information-theoret…

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation
Research

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

While Text-to-Image (T2I) models have shown remarkable success in generating photorealistic visual content, they still struggle with the rigorous semantic align…

A Dual-domain Refinement Network with FBP-based Jacobian Learning for Sparse-view Dual-Energy CT Material Decomposition
Research

A Dual-domain Refinement Network with FBP-based Jacobian Learning for Sparse-view Dual-Energy CT Material Decomposition

Dual-energy CT (DECT) exploits attenuation differences across different X-ray spectra to provide richer material information and has been widely used in medical…

The Many-Body Problem of the Data Centre
Research

The Many-Body Problem of the Data Centre

Modern Artificial Intelligence is often framed as limited by its own disembodiment, as if giving it a body would unlock its true potential. We argue to the cont…

Grounding LLM Reasoning under Incomplete Graph Evidence
Research

Grounding LLM Reasoning under Incomplete Graph Evidence

Knowledge graphs can guide large language models (LLMs) reasoning, but the graph seen by a system is usually a retrieved, linked, temporally scoped, and incompl…