arXiv:2609.19006v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Large Reasoning Models (LRMs) are typically evaluated on challenging…
Tag: cs.AI updates on arXiv.org
BadQubits: An LLM-Based Framework for Static Pre-Execution Detection of Structurally Harmful Quantum Circuits
arXiv:2609.18965v1 Announce Type: cross Abstract: This paper presents BadQubits, an LLM-based framework for static pre-execution detection of structurally…
StableEval Arena: A Cost-Aware Agentic Benchmark for Stablecoin Price Stability Prediction
arXiv:2609.18949v1 Announce Type: cross Abstract: We introduce StableEval Arena, a cost-aware benchmark framework for evaluating agentic AI systems on…
Transcribe, Then Reason: Two-Pass Decomposition for Multimodal Review
arXiv:2609.18958v1 Announce Type: cross Abstract: The natural way to review a long recording or document with a multimodal model is to hand it the raw…
Automated Dental Caries Segmentation in Panoramic Radiographs Using Dual-Stage Deep Learning
arXiv:2609.18952v1 Announce Type: cross Abstract: Early detection of dental caries remains challenging due to limitations in traditional diagnostic…
NeuroECG: ECGFounder-Based Deep ECG Representation for EEG-Free Neurological Prognostication After Cardiac Arrest
arXiv:2609.18891v1 Announce Type: cross Abstract: Neurological prognostication after cardiac arrest commonly relies on electroencephalography (EEG).…
Dose-Aware Cold Diffusion with Physics Consistency for Generalizable Low-Dose CT Reconstruction
arXiv:2609.18943v1 Announce Type: cross Abstract: Reducing radiation dose in computed tomography significantly degrades image quality and poses challenges…
Beyond Outcomes: Dual-View Relational Learning for Efficient Agent Benchmarking
arXiv:2609.18909v1 Announce Type: cross Abstract: Agent benchmarks are substantially more costly to evaluate than conventional LLM benchmarks. Benchmark…
Higher-order pruning of experts in mixture-of-experts language models
arXiv:2609.18916v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models suffer from large parameter counts, which create a significant…
Social Laws for Multi-agent Coordination in Stochastic Environments
arXiv:2609.18929v1 Announce Type: cross Abstract: In multi-agent environments, coordinating agents to prevent interference and ensure robust individual…
