arXiv:2609.04714v1 Announce Type: cross Abstract: Striking a balance between helpfulness and safety remains a fundamental challenge in aligning large…
Category: AI
XDOF, just 3 months out of stealth, is in talks for a Series B at a $1.2B valuation
The round is being raised just months after the robot data startup exited from stealth.
Knowing What Not to Answer: Selective Non-Compliance in Vision-Language Models
arXiv:2609.04720v1 Announce Type: cross Abstract: Vision-language models (VLMs) are expected to respond helpfully to appropriate requests while…
SCAPES: Semantically Conditioned Autoregressive Prior for Environmental Sounds
arXiv:2609.04634v1 Announce Type: cross Abstract: As generative audio models grow in complexity, the computational and ecological costs of synthesizing…
Tracing Audio Grounding and Answer Selection in Audio LLMs
arXiv:2609.04637v1 Announce Type: cross Abstract: Audio Large Language Models (Audio LLMs) have advanced in audio understanding, yet they can still…
Beyond Code Generation: Reliability, Verification, and Cost Economics in the Agentic Software Development Lifecycle
arXiv:2609.04681v1 Announce Type: cross Abstract: AI coding systems are moving from autocomplete and chat toward agents that can inspect repositories,…
Wireless Foundation Models: State-of-the-Art and Open Challenges
arXiv:2609.04707v1 Announce Type: cross Abstract: Wireless foundation models (WFMs) have emerged as a promising approach for learning reusable…
GPT-6 Astra beat Portal start to finish without human help in under 24 hours
GPT-6 Astra beat the puzzle game Portal entirely on its own in about 24 hours, with zero human help after the initial goal was set. Developer cozyblaze…
Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion
arXiv:2609.04690v1 Announce Type: cross Abstract: Multimodal emotion recognition has attracted growing interest due to its importance in human-computer…
PetQA: Benchmarking Veterinary Knowledge and Clinical Reasoning
arXiv:2609.04598v1 Announce Type: cross Abstract: We introduce PetQA, a Korean long-form question-answering (QA) benchmark for evaluating veterinary…
