arXiv:2609.28090v1 Announce Type: cross Abstract: Backtest auditing is a calibration problem: high flaw recall is not useful when the model falsely flags…
Author: script
AI News Brief Hourly Summary 2026-09-24 23h : 11 posts
11 posts published in the last hour 20:32Evaluating Feedback Focus and Pedagogical Adaptivity in LLM-Generated Feedback on Student Writing 20:32TEMPS: Temporal Sentence Embeddings for Temporal Information Retrieval 20:32PISCES: Physics-Informed Solar-wind Convolutional autoEncoder for Space-weather Anomaly Detection and Early Warning 20:32Controlled…
Evaluating Feedback Focus and Pedagogical Adaptivity in LLM-Generated Feedback on Student Writing
arXiv:2609.28026v1 Announce Type: cross Abstract: We investigate whether state-of-the-art large language models (LLMs) generate feedback that reflects the…
TEMPS: Temporal Sentence Embeddings for Temporal Information Retrieval
arXiv:2609.28048v1 Announce Type: cross Abstract: Modern information retrieval (IR) systems rarely represent time, yet many information needs depend on…
PISCES: Physics-Informed Solar-wind Convolutional autoEncoder for Space-weather Anomaly Detection and Early Warning
arXiv:2609.28022v1 Announce Type: cross Abstract: Space weather early warning depends on detecting solar wind transients in in-situ measurements at the…
Controlled Attribute-Specific Summarization of Interrogative Dialogues
arXiv:2609.28004v1 Announce Type: cross Abstract: Effective summarization of interrogative dialogues is a critical task in forensic and investigative…
Prompt, Probe, Train, or Annotate? Single-camera sports video understanding in amateur settings
arXiv:2609.28049v1 Announce Type: cross Abstract: Video understanding is usually benchmarked on curated, single-actor, or professionally filmed clips, and…
Compliant with Local Controls, Collectively Discriminatory. A Governance Architecture for Multi-Agent AI in Regulated Finance
arXiv:2609.27994v1 Announce Type: cross Abstract: Financial institutions are beginning to deploy agentic workflows in credit, fraud, collections,…
RelCheck: Dual-Evidence Spatial Grounding for VLM Hallucination Correction
arXiv:2609.27890v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) fre- quently generate text that is inconsistent with the input…
Riemannian Structure and Optimization for a Class of Low-Parametric Orthogonal Matrices
arXiv:2609.27982v1 Announce Type: cross Abstract: In this paper, we are concerned with matrices formed by block-diagonal factors interleaved with fixed…
