arXiv:2609.29358v1 Announce Type: cross Abstract: Vision-language models such as CLIP achieve strong zero-shot classification, yet under distribution…
Category: cs.AI updates on arXiv.org
From Policy Documents to Structured Survey Responses: Evaluating Large Language Models for Policy Monitoring
arXiv:2609.29370v1 Announce Type: cross Abstract: Science, technology, and innovation policies are crucial for competitiveness, yet their diversity and…
Hyperbolic Multimodal Continual Learning: A Closest-Admissible Solution
arXiv:2609.29329v1 Announce Type: cross Abstract: Existing continual-learning methods protect parameters, replayed examples, or Euclidean feature…
Where LLM Graders Succeed and Break: Evidence from Two Computer-Science Exams
arXiv:2609.29333v1 Announce Type: cross Abstract: One long-form exam in a large course costs hundreds of grader-hours, and qualified graders are scarce;…
ArGuard Shared Task: Harmful Content Detection in Arabic Memes and LLM Prompts
arXiv:2609.29349v1 Announce Type: cross Abstract: ArGuard is a shared task on harmful content detection in Arabic memes and LLM prompts. It includes two…
Neuralized Multi-Wavelet Decomposition for Time Series Classification and Forecasting
arXiv:2609.29317v1 Announce Type: cross Abstract: Time series analysis is fundamental in domains such as finance, healthcare, and meteorology. Real-world…
TP-CRIV: A Framework for Third-Party Challenge-Response Identity Verification of AI Models
arXiv:2609.29264v1 Announce Type: cross Abstract: Artificial intelligence (AI) models are increasingly deployed through remote services, making model…
DocuTeam: Mixed-Initiative Multi-Agent Discussions around Evolving Documents
arXiv:2609.29309v1 Announce Type: cross Abstract: In open-ended problem solving, collaborators often rely on discussion to surface concerns, challenge…
Deep learning of longitudinal visual fields predicts glaucoma progression rate and identifies fast progressors
arXiv:2609.29256v1 Announce Type: cross Abstract: Glaucoma is the leading cause of irreversible blindness, and timely identification of fast progressors…
Reasoning Instructions Can Break Answer Decoding in Vision–Language Models
arXiv:2609.29278v1 Announce Type: cross Abstract: Chain-of-thought (CoT) instructions can distort multiple-choice VLM evaluation when a scorer appends a…
