arXiv:2608.26656v1 Announce Type: cross Abstract: Multi-object removal in 3D scenes is challenging due to severe occlusions, semantic entanglement, and…
Category: cs.AI updates on arXiv.org
Do LLMs Understand Personality? Rethinking Persona Fidelity Evaluation through Structured Behavioral Inference
arXiv:2608.26674v1 Announce Type: cross Abstract: As large language models are increasingly deployed to simulate diverse human characters, ensuring…
PailitaoGR: Latent Think-with-Images for Generative Image Retrieval
arXiv:2608.26658v1 Announce Type: cross Abstract: Generative retrieval has demonstrated strong performance by directly generating product semantic…
RTNav: Towards Real-Time Zero-Shot Object Navigation
arXiv:2608.26496v1 Announce Type: cross Abstract: Navigation in unknown environments to find unforeseen objects has become increasingly feasible with…
Physics-Informed Stochastic Configuration Machine: A Backpropagation-Free Neural Network with Fast Training for Nonlinear Differential Equations
arXiv:2608.26549v1 Announce Type: cross Abstract: While Physics-Informed Neural Networks (PINNs) have emerged as a transformative paradigm for solving…
J-Zero: Unified Challenger–Solver–Judge Co-Evolution from Zero Data
arXiv:2608.26582v1 Announce Type: cross Abstract: Self-evolving language models have recently emerged as a promising path toward superintelligence, with…
Zero-Shot Self-Orchestration with Ledger-Based Control for Improved LLM Coding Performance
arXiv:2608.26480v1 Announce Type: cross Abstract: Multi-agent large language model systems are widely reported to beat single-model baselines, but the…
Risks and Controls for Multi-Agent Systems: an analytical framework for deployment of AI agents across organisational boundaries
arXiv:2608.26626v1 Announce Type: cross Abstract: This report presents a framework to help organisations, policymakers and researchers reason about the…
SpeechGym: An Audio-Native Gym for Training Voice Agents via Reinforcement Learning
arXiv:2608.26432v1 Announce Type: cross Abstract: Voice agents must call tools and hold multi-turn dialogue entirely through speech, yet the dominant…
The Latent Diagnostic Taxonomy: A Framework for Constructing Classifiers and Diagnosing Their Decisions, Applied to Prompt Injection Detection
arXiv:2608.26423v1 Announce Type: cross Abstract: This paper proposes a framework for constructing a classifier as a safeguard layer, and for developing a…
