arXiv:2608.22852v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in investment decision-making, yet prior work shows…
Category: cs.AI updates on arXiv.org
FinixDoc: Rethinking Financial Document Parsing Beyond Saturated Benchmarks
arXiv:2608.22842v1 Announce Type: new Abstract: Financial document parsing requires accuracy, structural consistency, and verifiability that current…
GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data Synthesis
arXiv:2608.22847v1 Announce Type: new Abstract: Vision-Language Models (VLMs) based GUI agents stand to benefit significantly from online reinforcement…
The Retriever Should Remember: Experience-Amortized Reranking for Long-Term Agent Memory
arXiv:2608.22767v1 Announce Type: new Abstract: Long-term language-model agents accumulate memories across interactions, but their retrievers typically do…
The Compaction Cliff in Long-Running AI Agent Memory
arXiv:2608.22752v1 Announce Type: new Abstract: A safety rule and an episodic log compete for the same tokens in an AI agent’s context. When the budget…
Compositional Chain-of-Relations for Faithful Knowledge Graph Question Answering with Large Language Models
arXiv:2608.22762v1 Announce Type: new Abstract: Knowledge graph question answering (KGQA) is a key task for evaluating KG-augmented Large Language Models…
TailSieve: Partial-Rollout-Guided Tail Routing for LLM Rollouts
arXiv:2608.22788v1 Announce Type: new Abstract: Large-scale rollouts have become a core component of modern LLM systems, spanning reinforcement learning…
Performance of a domain-specific large language model in answering patient questions in psychiatry
arXiv:2608.22797v1 Announce Type: new Abstract: Background This study was designed to evaluate whether a domain-specific large language model (LLM)…
Does Rank Still Matter? Position Bias When AI Agents Shop on Our Behalf
arXiv:2608.22697v1 Announce Type: new Abstract: Search rankings are valuable because human attention is scarce and sequential. Higher-placed alternatives…
LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans
arXiv:2608.22731v1 Announce Type: new Abstract: Nonverbal behavior generation systems for virtual agents often take an utterance as input and generate…
