arXiv:2609.24324v1 Announce Type: new Abstract: Electroencephalography (EEG) provides a non-invasive window into dynamic brain activity, yet modeling…
Category: AI
When and How Should an Agent Clarify? CIGAsk: Teaching LLMs to Clarify via Counterfactual Information Gain
arXiv:2609.24290v1 Announce Type: new Abstract: Instruction-tuned LLMs faced with underspecified queries often commit to a single interpretation rather…
How Many Pixels Is a Digit Worth? Place-Aware Coordinate Entropy for GUI Agent Confidence Estimation
arXiv:2609.24277v1 Announce Type: new Abstract: GUI agents predict click coordinates as digit-token sequences, but standard text-LLM confidence estimation…
Taming CoT Obfuscation in VLMs: From Mechanistic Evidence to Activation Enforcement
arXiv:2609.24243v1 Announce Type: new Abstract: Reinforcement learning (RL) improves reasoning in vision-language models (VLMs) but can induce…
High-Performance Data Processing with Polars: A KDnuggets Cheat Sheet
Polars is a DataFrame library written in Rust on the Apache Arrow memory format, and the speed comes less from the language than from the model. The…
Unsupervised Brain Anomaly Detection as a Bayesian Inverse Problem with Diffusion Prior
arXiv:2609.24265v1 Announce Type: new Abstract: Unsupervised anomaly detection (UAD) aims to localize abnormal regions in medical scans without…
APEXA: Execution-Integrity Enforcement for Multi-Agent LLM Automation of Synchrotron Data Reduction
arXiv:2609.24165v1 Announce Type: new Abstract: Synchrotron data reduction, detector calibration followed by azimuthal integration of terabyte-scale…
Recovering Lost Details: Multi-Scale Frequency Compensation for Long-Term Time Series Forecasting
arXiv:2609.24229v1 Announce Type: new Abstract: Long-term time series forecasting has made significant progress by leveraging multi-scale information to…
LIMIT: Less Is More for Instruction Tuning in Text-to-SQL
arXiv:2609.24186v1 Announce Type: new Abstract: Large language models have achieved remarkable progress on Text-to-SQL through reasoning-enhanced…
CREDO: Variance-Guided Rubric Evolution for Replay-Corrected Credit Assignment
arXiv:2609.24174v1 Announce Type: new Abstract: Long-horizon language agents receive sparse terminal feedback, while intermediate rubrics provide…
