arXiv:2608.08677v1 Announce Type: new Abstract: Skill evolution improves agent skills through feedback over time, with failed trajectories often providing…
Tag: cs.AI updates on arXiv.org
A Structural Dynamics Graph World Model: Unified Modeling, Constrained Rollout, and Interpretable Calibration
arXiv:2608.08689v1 Announce Type: new Abstract: The state evolution of a complex system arises jointly from object laws, relational propagation, domain…
MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning
arXiv:2608.08623v1 Announce Type: new Abstract: In Reinforcement Learning with Verifiable Rewards (RLVR) frameworks for mathematical reasoning tasks,…
Can Open-Weight Models Compete on Financial Text Comprehension?
arXiv:2608.08634v1 Announce Type: new Abstract: Open-weight language models from Chinese AI labs caught up on benchmarks relative to proprietary frontier…
Smart Compaction: Predicting Compaction Utility from Lakehouse Table Metadata
arXiv:2608.08639v1 Announce Type: new Abstract: Open lakehouse table formats accumulate small data files over time, which degrades query performance.…
UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models
arXiv:2608.08627v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers expand recommendation capacity through conditional computation, yet…
A QUBO-Inspired Computational Framework for Airport Landside Bottleneck Diagnosis and Dynamic Dispatch Optimization
arXiv:2608.08632v1 Announce Type: new Abstract: Airport landside traffic centers connect terminal arrivals with taxis, ride-hailing vehicles, private…
Walking through Discussions: A Mobile Visual Analytics System for In-Situ Group Discussion Analysis
arXiv:2608.08617v1 Announce Type: new Abstract: Group discussion-based teaching is widely used to foster collaborative learning, yet teachers in physical…
ForestBench: A Unified Graph Framework for Evaluating Multi-Agent Collaboration
arXiv:2608.08605v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on Large Language Models (LLMs) are proliferating rapidly, but their…
Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents
arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them.…
