arXiv:2608.16628v1 Announce Type: new Abstract: Modern Multimodal Retrieval-Augmented Generation (M-RAG) systems are fundamentally limited by the binary…
Category: AI
CUBICS: Situation-aware performance estimation for safety-relevant ML components
arXiv:2608.16564v1 Announce Type: new Abstract: Machine learning (ML) is a key technology driving innovation today, but ensuring ML safety remains a major…
Physics of Agents: Statistical Mechanics Predicts Collective Behavior of AI Agents
arXiv:2608.16578v1 Announce Type: new Abstract: AI agents increasingly operate as part of interacting systems rather than in isolation. As agents exchange…
DeepInsight II: One Trace from Benchmark to Robot
arXiv:2608.16556v1 Announce Type: new Abstract: Across a Physical AI stack, evaluation maturity is inversely aligned with deployment risk: foundation…
NVIDIA Releases TensorRT Model Connect in Public Preview: Hugging Face Checkpoint to Native C++ Inference in Two Commands
NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to…
Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I)
arXiv:2608.16565v1 Announce Type: new Abstract: This cumulative habilitation thesis studies probabilistic circuits (PCs) as a powerful and tractable…
DeepSeek AI Releases DeepSeek Harness in Developer Preview: An MIT-Licensed Agent Harness Where Everything is a Plugin
DeepSeek Harness v0.1 is an MIT-licensed agent harness where every capability is a Cordis plugin. Four runtime modes, append-only session logs, and…
Large language models as synthetic clinical experts to inform longitudinal rare-disease modeling
arXiv:2608.16507v1 Announce Type: new Abstract: Due to the limited amount of information, modeling longitudinal rare-disease data can benefit from…
The Value of a Prompt: An LLM-Relative Kolmogorov-Complexity Approach
arXiv:2608.16438v1 Announce Type: new Abstract: In a world where valuable artifacts are increasingly created, completed, or processed by LLMs, the central…
HaReCAP: Habitual-action Grounding for Recursive Large Language Model Agents
arXiv:2608.16447v1 Announce Type: new Abstract: Long-horizon embodied tasks require LLM agents to iteratively decompose high-level goals, revise plans in…
