arXiv:2609.11109v1 Announce Type: cross Abstract: The utility of AI in multi-coder qualitative coding has been widely discussed, yet little empirical…
Category: cs.AI updates on arXiv.org
A Fragility Spectrum for Recursive Language-Model Training
arXiv:2609.11149v1 Announce Type: cross Abstract: Model-generated text is finding its way back into training corpora, and there is plenty of evidence that…
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
arXiv:2609.11042v1 Announce Type: cross Abstract: Agent usage is shifting toward long-horizon tasks such as coding and scientific discovery, among which…
Less can be More: What Aspects of Speech Drive End-of-Turn Detection
arXiv:2609.11066v1 Announce Type: cross Abstract: In conversational AI, detecting when a speaker has finished talking is crucial for natural turn taking.…
DeFiFusion: Combining Transaction Events with Smart Contracts to Detect Price Manipulation Attacks
arXiv:2609.11008v1 Announce Type: cross Abstract: Decentralized Finance (DeFi) has emerged as a rapidly growing blockchain-based financial service, where…
Toward Interpretable Multimodal Fusion: Heat Conduction Modeling for Hyperspectral and LiDAR Joint Classification
arXiv:2609.11040v1 Announce Type: cross Abstract: The fusion of hyperspectral (HS) and Light Detection and Ranging (LiDAR) data plays a crucial role in…
Topological Necessities: Mechanism-Invariant Strategic Subgoals for Cross-Embodiment Goal-Conditioned Control
arXiv:2609.11014v1 Announce Type: cross Abstract: Long-horizon goal-conditioned reinforcement learning delegates control to a high-level module that…
New Evidence, Same Choice: Testing Physical Experiment Selection in Vision Language Models
arXiv:2609.11022v1 Announce Type: cross Abstract: A model first sees an image from one physical measurement experiment, such as how far a block coasted,…
BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evaluation Infrastructure
arXiv:2609.11028v1 Announce Type: cross Abstract: LM-agent benchmarks increasingly function as interactive evaluation infrastructure. Agents observe…
What a Random Draw from the MCP Registry Contains, and What Tool-Use Benchmarks Contain Instead
arXiv:2609.10962v1 Announce Type: cross Abstract: Studies of the Model Context Protocol (MCP) server ecosystem draw their samples in ways that quietly…
