arXiv:2605.09134v4 Announce Type: replace Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse…
Category: cs.AI updates on arXiv.org
On the Limitations of Large Language Models for Conceptual Database Modeling
arXiv:2605.11986v2 Announce Type: replace Abstract: This article analyzes the use of Large Language Models (LLMs) as support for the conceptual modeling…
Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales
arXiv:2607.25364v3 Announce Type: replace Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales…
A Forced-Structure Reduction and Verifiable Bounds for Conway’s 99-Graph
arXiv:2608.11211v2 Announce Type: replace Abstract: Conway’s 99-graph problem asks whether a strongly regular graph with parameters…
MemeLens: Multilingual Multitask VLMs for Memes
arXiv:2601.12539v4 Announce Type: replace Abstract: Memes are a dominant medium for online communication and manipulation because meaning emerges from…
Fact Grounded Attention: Eliminating Hallucination in Large Language Models Through Attention Level Knowledge Integration
arXiv:2509.25252v3 Announce Type: replace Abstract: “The greatest enemy of knowledge is not ignorance, it is the illusion of knowledge.” Large Language…
Beyond Final Answers: CRYSTAL Benchmark for Transparent Multimodal Reasoning Evaluation
arXiv:2603.13099v3 Announce Type: replace Abstract: We introduce CRYSTAL (Clear Reasoning via Yielded Steps, Traceability, and Logic), a diagnostic…
Collab-Solver: Collaborative Solving Policy Learning for Mixed-Integer Linear Programming
arXiv:2508.03030v3 Announce Type: replace Abstract: Mixed-integer linear programming (MILP) has been a fundamental problem in combinatorial optimization.…
Transferable knowledge graphs with executable learned operators for algorithm design
arXiv:2603.27922v2 Announce Type: replace Abstract: Procedural knowledge in algorithm design is embedded in source code and rebuilt for each new domain.…
DiaVLo: Diagnosing Behaviours of Vision-Language Models
arXiv:2609.22008v1 Announce Type: cross Abstract: Vision-language models (VLMs) rely on storing and transferring appropriate information across their…
