arXiv cs.LG
arXiv cs.LG

arXiv cs.LG

arXiv cs.LG publishes articles covering LLM, AI, analysis, data. A trusted source for AI and technology insights.

Profile generated by AI for Anything

PhantomFill: When the Form Demands an Answer, Language Models Invent One

PhantomFill: When the Form Demands an Answer, Language Models Invent One

21h

Adaptive Depth in Looped Transformers: Diagnosing Learned Halting Gates and Trajectory Readouts

Adaptive Depth in Looped Transformers: Diagnosing Learned Halting Gates and Trajectory Readouts

21h

Scaling Closed-Loop Feature Channel Configuration with LLMs

Scaling Closed-Loop Feature Channel Configuration with LLMs

21h

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

21h

Generative Bayesian Filtering for State Estimation

Generative Bayesian Filtering for State Estimation

21h

Do Active SAE Feature Planes Carry More Holonomy? A Preregistered Reversal in Gemma

Do Active SAE Feature Planes Carry More Holonomy? A Preregistered Reversal in Gemma

21h

Multimodal CoLRAG-TF: Triple-Filtered Retrieval for Complex PDFs

Multimodal CoLRAG-TF: Triple-Filtered Retrieval for Complex PDFs

21h

The original title is "DataPrep-Bench: Benchmarking LLMs as Training Data Preparators"

The original title is "DataPrep-Bench: Benchmarking LLMs as Training Data Preparators"

21h

The Active Ingredient in Muon's Grokking

The Active Ingredient in Muon's Grokking

21h

CLOE: Christoffel Loss Autoencoder for Anomaly Detection

CLOE: Christoffel Loss Autoencoder for Anomaly Detection

21h

Bayesian Wind Tunnels for Model Selection

Bayesian Wind Tunnels for Model Selection

1d

CruiseBench: A Real-Flight-Aligned N-CMAPSS Benchmark for Engine RUL Prediction

CruiseBench: A Real-Flight-Aligned N-CMAPSS Benchmark for Engine RUL Prediction

1d

Building Fast, Evaluating Slow: Pipeline Choices Dominate Autointerpretability Score Variance

Building Fast, Evaluating Slow: Pipeline Choices Dominate Autointerpretability Score Variance

1d

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions

1d

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

1d

Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses

Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses

1d

Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation

Neural Operator Surrogates for Two-Dimensional Neutron Flux Estimation

1d

SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning

SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning

1d

Challenges of Explainability in Continual Learning for Time Series Forecasting

Challenges of Explainability in Continual Learning for Time Series Forecasting

1d

Air Quality Arena: A Large-Scale Multi-Region Ground Monitoring Dataset and Benchmark for Air Quality Forecasting with Time-Series Foundation Models

Air Quality Arena: A Large-Scale Multi-Region Ground Monitoring Dataset and Benchmark for Air Quality Forecasting with Time-Series Foundation Models

1d

FALCON-Discover: Discovering Concentrated False-Confidence Regions for Calibration

FALCON-Discover: Discovering Concentrated False-Confidence Regions for Calibration

2d

Beyond Output-Space Calibration: Spectral Evidence Bundling for Selective Reliability Estimation in Time-Series Classification

Beyond Output-Space Calibration: Spectral Evidence Bundling for Selective Reliability Estimation in Time-Series Classification

2d

FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images

FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images

2d

BearingNAS: Obtaining In-Sensor Intelligent Fault Diagnosis Systems for Bearings Using a Laptop

BearingNAS: Obtaining In-Sensor Intelligent Fault Diagnosis Systems for Bearings Using a Laptop

2d