arXiv:2509.25252v3 Announce Type: replace Abstract: “The greatest enemy of knowledge is not ignorance, it is the illusion of knowledge.” Large Language…
Tag: cs.AI updates on arXiv.org
Beyond Final Answers: CRYSTAL Benchmark for Transparent Multimodal Reasoning Evaluation
arXiv:2603.13099v3 Announce Type: replace Abstract: We introduce CRYSTAL (Clear Reasoning via Yielded Steps, Traceability, and Logic), a diagnostic…
Collab-Solver: Collaborative Solving Policy Learning for Mixed-Integer Linear Programming
arXiv:2508.03030v3 Announce Type: replace Abstract: Mixed-integer linear programming (MILP) has been a fundamental problem in combinatorial optimization.…
Transferable knowledge graphs with executable learned operators for algorithm design
arXiv:2603.27922v2 Announce Type: replace Abstract: Procedural knowledge in algorithm design is embedded in source code and rebuilt for each new domain.…
DiaVLo: Diagnosing Behaviours of Vision-Language Models
arXiv:2609.22008v1 Announce Type: cross Abstract: Vision-language models (VLMs) rely on storing and transferring appropriate information across their…
Gricea: An Open Science Platform for Conversational AI Research
arXiv:2609.22039v1 Announce Type: cross Abstract: We need studies on conversational AI (CAI) at scale to understand human behavior and shape CAI design.…
Bayesian Belief Layer for Controllable Opinion Dynamics in LLM Agents
arXiv:2609.21997v1 Announce Type: cross Abstract: LLM agents in social simulation revise their opinions implicitly, in context: how open an agent is to…
NemotronLabs VoiceChat: An Open Full-duplex Speech-to-Speech Model with Tool Calling Capabilities
arXiv:2609.21967v1 Announce Type: cross Abstract: We introduce NemotronLabs VoiceChat, an open full-duplex speech-to-speech model with native tool-calling…
Value-Sensitive Delegation in Everyday AI Agent Use: Evidence from OpenClaw
arXiv:2609.22067v1 Announce Type: cross Abstract: Users increasingly delegate work to autonomous AI agents, yet evaluations typically measure task…
Neural Cellular Automata Learn General Features in their Hidden Channels
arXiv:2609.21870v1 Announce Type: cross Abstract: Modern deep learning models achieve impressive generalization through over-parameterization, but this…
