arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores…
Tag: AI
Assessing AI-generated music detection in real-world broadcast monitoring
arXiv:2608.07359v1 Announce Type: cross Abstract: The proliferation of AI-generated music in broadcast media raises concerns about transparency and fair…
Measurements Automatically Extracted from Zero Echo Time MRI Using Deep Learning Image Segmentation and Geometric Modeling Agree with Expert Manual Readings
arXiv:2608.07368v1 Announce Type: cross Abstract: Computed tomography (CT) remains the reference for 3D osseous morphometry in femoroacetabular…
Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding
arXiv:2608.07353v1 Announce Type: cross Abstract: Understanding concepts is fundamental to generalization. Despite their impressive performance on a wide…
H2AL: Hyperbolic Hierarchy-aware Aggregative Learning for Registration-based Few-shot Medical Image Segmentation
arXiv:2608.07340v1 Announce Type: cross Abstract: Registration-based Few-shot medical image segmentation (RFMIS) aims to generate pseudo-labels for…
EliSeg: Verified Target Construction for Report-Grounded Abnormality Segmentation
arXiv:2608.07299v1 Announce Type: cross Abstract: Radiology reports describe clinical observations but do not specify executable segmentation targets.…
Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallelized Q-Networks
arXiv:2608.07335v1 Announce Type: cross Abstract: Recent advancements in deep reinforcement learning have increasingly favored simplified, highly…
Towards Assurance Closure in AI-Native Large-Scale Agile Software Development
arXiv:2608.07317v1 Announce Type: cross Abstract: The AI-Native Manifesto envisions large-scale agile software development in which humans increasingly…
Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination
arXiv:2608.07302v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) often suffer from object hallucination, generating objects that are…
Natural Language Processing Psychometrics
arXiv:2608.07316v1 Announce Type: cross Abstract: Natural Language Processing (NLP) models predicting mental health outcomes rarely specify what they…