arXiv:2608.20083v1 Announce Type: cross Abstract: Question Answering over Temporal Knowledge Graphs (TKGQA) requires reasoning over time-sensitive facts,…
Tag: AI
Towards Quantifying Benchmark Optimization in ASR Models
arXiv:2608.19936v1 Announce Type: cross Abstract: Public benchmarks are important measures of Automatic Speech Recognition (ASR) model capabilities.…
Designing Human-mediated AI Guidance: Ready Together for Personalized Family Emergency Preparedness
arXiv:2608.19950v1 Announce Type: cross Abstract: Artificial intelligence (AI) systems are increasingly used across domains to provide personalized…
An Inclusive and Lightweight Approach to Federated Continual Learning for Cultural Heritage
arXiv:2608.20038v1 Announce Type: cross Abstract: Artificial intelligence can support cultural heritage and digital humanities through large-scale…
EchoCoT: Extracting Hidden Chain-of-Thought from Large Reasoning Models
arXiv:2608.20055v1 Announce Type: cross Abstract: Hidden chain-of-thought (CoT) traces, especially those from frontier proprietary large reasoning models…
Open-Vocabulary 3D Object Detection with Co-Distillation Discovery and Dual Guidance Robust Training
arXiv:2608.19973v1 Announce Type: cross Abstract: Recently, open-vocabulary 3D object detection (3D-OVD) has gained increasing attention for its ability…
A knowledge-guided agentic framework for mitigating patient-context ambiguity in health queries
arXiv:2608.19875v1 Announce Type: cross Abstract: Patients often submit short, underspecified queries to healthcare chatbots that lack the…
Evidence Before Expansion: Reuse, Spawn, or Defer in Lifelong Expert Pools
arXiv:2608.19888v1 Announce Type: cross Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an…
Separating Covariate Shift from Mechanism Change with Two Discriminators: CJSD, a Conditional Discrepancy with an Exact Covariate-Concept Decomposition
arXiv:2608.19885v1 Announce Type: cross Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an…
Interrupting the Loop: Periodic Subject Changes Raise Judged Surprise and Connection in Base Language Models
arXiv:2608.19893v1 Announce Type: cross Abstract: Where does the novelty a base language model produces with no task come from, and what can an LLM judge…
