arXiv:2608.26846v1 Announce Type: cross Abstract: Interactive language-model agents use confidence signals to decide whether to answer immediately,…
Category: cs.AI updates on arXiv.org
Behavior2Trip: Towards Personalized Travel Planning via User Behavior Trajectory
arXiv:2608.26807v1 Announce Type: cross Abstract: Travel planning agents assist users in generating personalized travel plans by modeling their individual…
MedFG-VQA: Low-Frequency Memory and Graph Attention for Lightweight Medical VQA
arXiv:2608.26848v1 Announce Type: cross Abstract: Medical Visual Question Answering (Med-VQA) holds significant promise for clinical decision support, yet…
LiveVVT: High-Fidelity Video Virtual Try-On in Real Time
arXiv:2608.26714v1 Announce Type: cross Abstract: Diffusion-based Video Virtual Try-On (VVT) achieves high visual fidelity through bidirectional…
Rethinking Message Passing as Retrieval for Text-Attributed Graph Learning
arXiv:2608.26732v1 Announce Type: cross Abstract: Graph neural networks (GNNs) are typically conceptualized as message-passing neural networks, yet it…
Beyond Execution: Auditing Experimental Fidelity in LLM-Driven Scientific Research
arXiv:2608.26753v1 Announce Type: cross Abstract: LLM agents used for scientific experimentation must do more than generate executable code: they must…
FaultLens: Learning Compact Behavioral Test Suites for Generated Operational Programs
arXiv:2608.26746v1 Announce Type: cross Abstract: Generated operational programs are often validated with either a few hand-written examples or exhaustive…
Daydreaming: Stealing Hidden Agent Skills through Black-Box Task Interaction
arXiv:2608.26733v1 Announce Type: cross Abstract: Agent skills bundle instructions, reference data, and executable helpers that let a general agent…
AesCanvas: A Large-Scale Dataset and Benchmark for Aesthetic Critique and Contextual Suitability
arXiv:2608.26713v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have extended Image Aesthetic Assessment…
FOCUS & RePAIR: Mitigating Text Degeneration via Token-Level Guidance for Pruned Large Language Models
arXiv:2608.26676v1 Announce Type: cross Abstract: Pruning is a practical approach to compress large language models (LLMs), but it can amplify text…
