arXiv:2609.19818v2 Announce Type: replace-cross Abstract: Generalizing to unseen attacks remains challenging for audio deepfake detectors, and collecting…
Tag: cs.AI updates on arXiv.org
Large Language Model Agents for Evidence Based Genetic Disease Severity Classification
arXiv:2609.19569v2 Announce Type: replace-cross Abstract: Disease severity classification for genetic conditions is subjective and labor-intensive,…
PACE: Precise AI Cinematic Expression
arXiv:2609.19853v2 Announce Type: replace-cross Abstract: Between a screenplay and a film sits a planning problem that is spatial first: who stands where,…
ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions
arXiv:2609.18864v2 Announce Type: replace-cross Abstract: Privacy evaluations of tool-using LLM agents often inspect a designated action, final response,…
A primer on evaluation methods for large language models in healthcare
arXiv:2609.14819v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have a growing range of applications in medicine, and their…
Knowledge-Graph Based Augmentation versus Retrieval Augmented Generation for Cultural-Related Question Answering
arXiv:2609.18317v2 Announce Type: replace-cross Abstract: Large language models (LLMs) suffer from a long-tail deficit: culturally specific facts,…
From Momentary Emotion Inference to Sustained Emotion Support: Evaluating a Companion Agent in a Longitudinal Study
arXiv:2609.16344v2 Announce Type: replace-cross Abstract: Sustained emotional support is a long-horizon interaction task closely tied to human well-being.…
CPR: Combining global composing, local performing and full-sequence refining in piano rendering with continuous autoregressive modelling
arXiv:2609.18216v2 Announce Type: replace-cross Abstract: Prompt-conditioned piano MIDI-to-Music rendering aims to faithfully render target notes while…
PentestChain: A Cost-Aware, MCP-Orchestrated Framework for Automated Penetration Testing with Free-Tier LLMs
arXiv:2609.18120v2 Announce Type: replace-cross Abstract: AI-driven penetration testing has been demonstrated with premium frontier models such as GPT-4,…
Data-free On-policy Distillation
arXiv:2609.14193v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) has become a standard component of frontier post-training…
