arXiv:2608.16630v1 Announce Type: cross Abstract: Repository-scale coding requires an agent to keep tests, imports, configuration, and migration rules…
Category: cs.AI updates on arXiv.org
A Framework for Using and Evaluating LLMs as Surrogate Experts in Security Surveys: Reliability, Bias, and Implications
arXiv:2608.16893v1 Announce Type: cross Abstract: Expert surveys are widely used in security research to study practitioner workows and decision-making,…
What If AI Carried Her Imagination? Black Girls as Creators in an AI Storytelling Weekend Program
arXiv:2608.16896v1 Announce Type: cross Abstract: This paper presents the design and outcomes of a seven-weekend AI storytelling program developed for…
Delegation Asymmetry in Agentic Recommender Systems: Measuring Two-Sided Receptivity in Online Dating
arXiv:2608.18058v1 Announce Type: new Abstract: Autonomous LLM agents that converse on a user’s behalf are an emerging design pattern in matching…
On the Fragility of Self-Improving Agents: Variance, Task Order, and Underspecification
arXiv:2608.18066v1 Announce Type: new Abstract: Memory-based self-improving agents–those that learn from an online stream of tasks and improve over time…
HLSR: Hybrid Live Forecast Selective Dynamic Vehicle Rerouting for Real-Time Congestion Avoidance
arXiv:2608.18056v1 Announce Type: new Abstract: Urban traffic congestion reduces productivity and increases travel cost and emissions. Network-wide live…
StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents
arXiv:2608.18050v1 Announce Type: new Abstract: AI agents increasingly perform knowledge work (i.e., produce and modify persistent digital artifacts such…
Can Large Language Models Explain Flight Safety Events? A Prior-Guided Semantic LLM-based Approach
arXiv:2608.18017v1 Announce Type: new Abstract: Improving flight safety with flight data requires not only accurate detection of risk events, but more…
Adaptive Policy Portfolios for Robust Markov Decision Processes
arXiv:2608.17929v1 Announce Type: new Abstract: Robust Markov decision processes optimize one policy against a set of plausible transition functions. This…
AutoResearch: Insight In, Hallucination Out
arXiv:2608.17906v1 Announce Type: new Abstract: Autonomous research systems are increasingly capable of executing long research workflows, yet automation…
