arXiv:2609.28322v1 Announce Type: new Abstract: Benchmarking and routing platforms increasingly act as intermediaries connecting large language model…
Category: AI
A Resilience Recovery Method for Complex Traffic Network Security Based on Trend Forecasting
arXiv:2609.27903v1 Announce Type: new Abstract: Due to the rapid development of information technology, a huge and complex traffic network has been…
PASTABench: Proactive Assessment of Sequential Trajectories for Agent Safety
arXiv:2609.28197v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents that alter real-world states, ensuring…
Finite-Sample Probabilistic Safety Certification for AI-Based Grid-Edge Coordination
arXiv:2609.28182v1 Announce Type: new Abstract: Coordinating large population of flexible grid-edge devices can alleviate the need for time-consuming and…
Discovery of fully efficient fault indicators along a data-based diagnosis process
arXiv:2609.28087v1 Announce Type: new Abstract: The integration of model-based and data-driven paradigms provides a powerful framework for fault diagnosis…
SlackDrive: Reclaiming Runtime Slack for Adaptive Driving Inference
arXiv:2609.28064v1 Announce Type: new Abstract: Driving world-action models improve planning by coupling multimodal reasoning with future prediction, but…
Agentic Governance and Adversarial Verification for Policy-Constrained LLM Healthcare Appeal Generation
arXiv:2609.27844v1 Announce Type: new Abstract: Claim denial management costs U.S. healthcare approximately $260 billion annually in administrative…
Learning What to Activate: Combinatorial Capability Allocation for Long-Horizon Multimodal Agents
arXiv:2609.27869v1 Announce Type: new Abstract: Long-horizon multimodal agents rely on specialized capabilities for perception, retrieval, reasoning,…
Reachable Global Optimization in AI Systems: How Global Is Global?
arXiv:2609.27855v1 Announce Type: new Abstract: AI systems increasingly claim to optimize prompts, policies, architectures, plans, tool-use trajectories,…
A hierarchy of faithfulness criteria for knowledge base completion
arXiv:2609.27863v1 Announce Type: new Abstract: Knowledge graph completion is evaluated by ranking observed triples above randomly corrupted ones, which…
