arXiv:2609.28026v1 Announce Type: cross Abstract: We investigate whether state-of-the-art large language models (LLMs) generate feedback that reflects the…
Tag: cs.AI updates on arXiv.org
TEMPS: Temporal Sentence Embeddings for Temporal Information Retrieval
arXiv:2609.28048v1 Announce Type: cross Abstract: Modern information retrieval (IR) systems rarely represent time, yet many information needs depend on…
PISCES: Physics-Informed Solar-wind Convolutional autoEncoder for Space-weather Anomaly Detection and Early Warning
arXiv:2609.28022v1 Announce Type: cross Abstract: Space weather early warning depends on detecting solar wind transients in in-situ measurements at the…
Controlled Attribute-Specific Summarization of Interrogative Dialogues
arXiv:2609.28004v1 Announce Type: cross Abstract: Effective summarization of interrogative dialogues is a critical task in forensic and investigative…
Prompt, Probe, Train, or Annotate? Single-camera sports video understanding in amateur settings
arXiv:2609.28049v1 Announce Type: cross Abstract: Video understanding is usually benchmarked on curated, single-actor, or professionally filmed clips, and…
Compliant with Local Controls, Collectively Discriminatory. A Governance Architecture for Multi-Agent AI in Regulated Finance
arXiv:2609.27994v1 Announce Type: cross Abstract: Financial institutions are beginning to deploy agentic workflows in credit, fraud, collections,…
RelCheck: Dual-Evidence Spatial Grounding for VLM Hallucination Correction
arXiv:2609.27890v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) fre- quently generate text that is inconsistent with the input…
Riemannian Structure and Optimization for a Class of Low-Parametric Orthogonal Matrices
arXiv:2609.27982v1 Announce Type: cross Abstract: In this paper, we are concerned with matrices formed by block-diagonal factors interleaved with fixed…
Spread and Scale: What Determines Whether Test-Time Budget Allocation Pays
arXiv:2609.27917v1 Announce Type: cross Abstract: Neural combinatorial optimization solvers generate many candidate solutions per instance and report the…
Schr\”odinger’s Code Repository: Have LLMs Learned SWE-bench or Memorized It?
arXiv:2609.27891v1 Announce Type: cross Abstract: Repository-level coding benchmarks have become the standard for evaluating coding agents, yet they…
