arXiv:2609.22283v1 Announce Type: cross Abstract: Understanding the design space of streaming video diffusion is essential to exploring its potential for…
Tag: cs.AI updates on arXiv.org
ORDER: A Fictitious-World Benchmark for Domain-Adaptive Embodied AI
arXiv:2609.22285v1 Announce Type: cross Abstract: Adapting language models to new domains via continual pre-training raises a basic evaluation problem: if…
Complementary rPPG-Derived and Lip-Region Frequency Cues for Talking-Face Deepfake Detection
arXiv:2609.22284v1 Announce Type: cross Abstract: Talking-face (TF) deepfakes are detected unevenly by rPPG-based methods across generators. We study two…
Performance vs Consistency: Evaluating a Foundation Model in Lung-RADS Screening
arXiv:2609.22281v1 Announce Type: cross Abstract: Foundation models have recently demonstrated strong capabilities across a wide range of medical imaging…
Brain-to-Image Generation: Reconstructing Visual Stimuli from EEG using Generative Adversarial Networks
arXiv:2609.22282v1 Announce Type: cross Abstract: Reconstructing visual stimuli from electroencephalography (EEG) is difficult because scalp measurements…
Large language models in medical time series analysis
arXiv:2609.22262v1 Announce Type: cross Abstract: Medical time series (MedTS), including electrocardiograms (ECG), electroencephalograms (EEG),…
Used, Mentioned, or Condemned? A Controlled Contrast-Set Diagnostic for the Use-Mention Distinction in Code-Mixed Hinglish Misogyny Detection
arXiv:2609.22261v1 Announce Type: cross Abstract: Lexicon-driven misogyny detectors cannot, by construction, distinguish a slur used against a woman from…
Which Part of the Context Layer Does the Work? Separating Semantic Content from Retrieval Scaffolding in Text-to-SQL Agents
arXiv:2609.22259v1 Announce Type: cross Abstract: Context layers, curated documentation that an analytics agent fetches at query time, produce large…
Enabling Vision and Cross-Modal Learning for Multimodal Stroke Recurrence Prediction: An Interpretable Two-Step Framework
arXiv:2609.22271v1 Announce Type: cross Abstract: Multimodal stroke recurrence prediction requires effective integration of heterogeneous clinical and…
Hi-Singers: A Comprehensive High-Quality Dataset for Expressive Audio-Driven Singing Head Synthesis
arXiv:2609.22264v1 Announce Type: cross Abstract: State-of-the-art models for audio-driven digital human generation have achieved photo-realistic results…
