arXiv:2609.06914v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated strong capabilities across diverse domains, showing…
Tag: cs.AI updates on arXiv.org
Fraglingo: Molecular Design via Attachment-Aware Autoregressive Fragment Generation
arXiv:2609.13519v2 Announce Type: replace Abstract: We introduce Fraglingo, an autoregressive molecular generator that constructs molecules step by step…
Balance of Benchmarks: Semantic Density Reweighting for Task-Conditioned Model Comparison
arXiv:2608.30044v3 Announce Type: replace Abstract: Model comparison increasingly relies on large collections of publicly reported benchmark scores, yet…
MOSCOPT: Mixture-of-Skills Collective Optimization for LLM Agents
arXiv:2609.14399v2 Announce Type: replace Abstract: Natural language prompts and skills serve as the strategic backbone of LLM-based agents. Recent…
Intent-Governed Tool Authorization for AI Agents
arXiv:2606.22916v4 Announce Type: replace Abstract: Tool-using AI agents commonly operate under integration credentials whose static permissions exceed a…
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
arXiv:2605.09134v4 Announce Type: replace Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse…
On the Limitations of Large Language Models for Conceptual Database Modeling
arXiv:2605.11986v2 Announce Type: replace Abstract: This article analyzes the use of Large Language Models (LLMs) as support for the conceptual modeling…
Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales
arXiv:2607.25364v3 Announce Type: replace Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales…
A Forced-Structure Reduction and Verifiable Bounds for Conway’s 99-Graph
arXiv:2608.11211v2 Announce Type: replace Abstract: Conway’s 99-graph problem asks whether a strongly regular graph with parameters…
MemeLens: Multilingual Multitask VLMs for Memes
arXiv:2601.12539v4 Announce Type: replace Abstract: Memes are a dominant medium for online communication and manipulation because meaning emerges from…
