arXiv:2609.05284v1 Announce Type: new Abstract: Recent years have witnessed great advances in the reasoning ability of Large Language Models (LLMs).…
Tag: AI
Beyond Aggregate Scores: Behavioral Correctness Assumptions for Assessing Reference-Based Automatic Evaluation Methods
arXiv:2609.05289v1 Announce Type: new Abstract: Automated reference-based evaluation methods play a critical role in assessing natural language generation…
Testing Interchangeability in LLM Agent Teams
arXiv:2609.05279v1 Announce Type: new Abstract: Production multi-agent systems replace agents constantly, on the assumption that an agent filling a role…
Don’t Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference
arXiv:2609.05275v1 Announce Type: new Abstract: Layer dropout (a.k.a. stochastic depth) has been shown to enable faster training, higher accuracy, and…
MG Ship adds AI route optimisation as logistics returns accelerate
MG Ship has introduced an AI route optimisation and carrier selection module as logistics deployments demonstrate rapid cost and time returns. The…
AI for Computational Design Science: A Responsible Human-AI Framework and Case Study on Short-Form Video Safety Surveillance
arXiv:2609.05270v1 Announce Type: new Abstract: Artificial intelligence (AI) is transforming not only what information systems researchers design, but…
Commonsense Reasoning in Computer Vision: Foundations, Recent Advancements, and Future Directions
arXiv:2609.05257v1 Announce Type: new Abstract: Commonsense reasoning in computer vision encompasses integrating visual data and contextual knowledge,…
A Unified Physics-Aware Quantum Machine Learning Framework across Power GaN HEMTs and Logic Nanowire FETs: Predicting Unseen Process Splits and Held-Out Geometry Combinations with Lower Error and Tighter Split-to-Split Variability
arXiv:2609.05251v1 Announce Type: new Abstract: We present a unified reinforcement-learning (RL) framework that discovers compact parametrized quantum…
Uncensored Open-weight Models: Redistribution as the Persistence Layer
arXiv:2609.05241v1 Announce Type: new Abstract: A rapidly expanding ecosystem of actors is removing built-in safety guardrails from open-weight AI models.…
Trace2Tower: Transition-Aware EigenTrace Induction of Multi-Level Skills for LLM Agents
arXiv:2609.05261v1 Announce Type: new Abstract: Large language model agents increasingly rely on execution traces to master complex interactive tasks.…
