arXiv:2609.04711v1 Announce Type: cross Abstract: Generative AI and coding agents can accelerate research software development, but they also increase the…
Category: cs.AI updates on arXiv.org
When Does an Interpretation Count as Established? The Formation, Evaluation, and Responsibility of Interpretation in Generative AI
arXiv:2609.04766v1 Announce Type: cross Abstract: Generative AI research has increasingly evaluated factuality, citation, coverage, and report structure.…
Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing False Refusals in Language Models
arXiv:2609.04714v1 Announce Type: cross Abstract: Striking a balance between helpfulness and safety remains a fundamental challenge in aligning large…
Knowing What Not to Answer: Selective Non-Compliance in Vision-Language Models
arXiv:2609.04720v1 Announce Type: cross Abstract: Vision-language models (VLMs) are expected to respond helpfully to appropriate requests while…
SCAPES: Semantically Conditioned Autoregressive Prior for Environmental Sounds
arXiv:2609.04634v1 Announce Type: cross Abstract: As generative audio models grow in complexity, the computational and ecological costs of synthesizing…
Tracing Audio Grounding and Answer Selection in Audio LLMs
arXiv:2609.04637v1 Announce Type: cross Abstract: Audio Large Language Models (Audio LLMs) have advanced in audio understanding, yet they can still…
Beyond Code Generation: Reliability, Verification, and Cost Economics in the Agentic Software Development Lifecycle
arXiv:2609.04681v1 Announce Type: cross Abstract: AI coding systems are moving from autocomplete and chat toward agents that can inspect repositories,…
Wireless Foundation Models: State-of-the-Art and Open Challenges
arXiv:2609.04707v1 Announce Type: cross Abstract: Wireless foundation models (WFMs) have emerged as a promising approach for learning reusable…
Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion
arXiv:2609.04690v1 Announce Type: cross Abstract: Multimodal emotion recognition has attracted growing interest due to its importance in human-computer…
PetQA: Benchmarking Veterinary Knowledge and Clinical Reasoning
arXiv:2609.04598v1 Announce Type: cross Abstract: We introduce PetQA, a Korean long-form question-answering (QA) benchmark for evaluating veterinary…
