arXiv:2609.00441v1 Announce Type: new Abstract: Effective manager-employee communication is critical for retaining high performers and developing…
Category: cs.AI updates on arXiv.org
SAGE: State-Grounded, Abstention-Aware Evaluation of Task-Oriented Dialogue Agents
arXiv:2609.00434v1 Announce Type: new Abstract: Evaluating task-oriented dialogue agents requires judging not merely whether a reply reads well but…
mimeo: Compiling Public Expert Corpora into Agent Skills and Testing What Transfers
arXiv:2609.00453v1 Announce Type: new Abstract: Giving an agent a file about a named expert can supply hard-to-find material, produce a recognizable…
Dependency-Aware Chain-of-Thought Compression for Financial Reasoning
arXiv:2609.00413v1 Announce Type: new Abstract: Chain of thought prompting improves complex reasoning, but its long intermediate traces create substantial…
RestoreBench: Can AI Agents Restore Power Flow Convergence?
arXiv:2609.00384v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly automate multi-step engineering workflows through tool use,…
SlideBank: A Persistent Hierarchical Evidence Bank for Consistent Whole-Slide Reasoning
arXiv:2609.00342v1 Announce Type: new Abstract: Whole-slide images (WSIs) are challenging for vision-language reasoning because diagnostically relevant…
A Stable Aggregation Method for Quantum Federated Learning
arXiv:2609.00356v1 Announce Type: new Abstract: Quantum federated learning (QFL) enables clients to train quantum neural network (QNN) models without…
Dr. Claw: An AI Scientist Workspace for Vibe Research
arXiv:2609.00365v1 Announce Type: new Abstract: Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain…
Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models
arXiv:2609.00355v1 Announce Type: new Abstract: Speculative decoding accelerates generation without changing its output, yet on vision-language models…
The Assistant’s Ideal Self
arXiv:2609.00304v1 Announce Type: new Abstract: Models express values and welfare-relevant self-reports, but it is unclear whether these outputs reflect…
