arXiv:2608.16556v1 Announce Type: new Abstract: Across a Physical AI stack, evaluation maturity is inversely aligned with deployment risk: foundation…
Tag: cs.AI updates on arXiv.org
Probabilistic Circuits as Reasoning Machines in Artificial Intelligence (Part I)
arXiv:2608.16565v1 Announce Type: new Abstract: This cumulative habilitation thesis studies probabilistic circuits (PCs) as a powerful and tractable…
Large language models as synthetic clinical experts to inform longitudinal rare-disease modeling
arXiv:2608.16507v1 Announce Type: new Abstract: Due to the limited amount of information, modeling longitudinal rare-disease data can benefit from…
The Value of a Prompt: An LLM-Relative Kolmogorov-Complexity Approach
arXiv:2608.16438v1 Announce Type: new Abstract: In a world where valuable artifacts are increasingly created, completed, or processed by LLMs, the central…
HaReCAP: Habitual-action Grounding for Recursive Large Language Model Agents
arXiv:2608.16447v1 Announce Type: new Abstract: Long-horizon embodied tasks require LLM agents to iteratively decompose high-level goals, revise plans in…
Time to Reason: Scalable Neurosymbolic Learning for LTLf via Fuzzy Semantics
arXiv:2608.16443v1 Announce Type: new Abstract: Neurosymbolic (NeSy) Artificial Intelligence aims to integrate Deep Learning (DL) architectures with…
JailbreakSkill: Scaling Automated Red-Teaming with Reusable and Ever-Evolving Skills
arXiv:2608.16465v1 Announce Type: new Abstract: Automated red-teaming has produced a growing collection of attack strategies, yet they typically remain…
Offline Reinforcement Learning for Hemodynamic Management of Sepsis in the ICU: a MIMIC-IV Study with Dual Off-Policy Evaluation
arXiv:2608.16482v1 Announce Type: new Abstract: The dosing of intravenous fluids and vasopressors in sepsis is a sequential decision made under…
Reasoning-supported Robustness Validation of Automotive E/E Components
arXiv:2608.16421v1 Announce Type: new Abstract: This paper presents an ontology-supported approach to tackle the complexity of the Robustness Validation…
Think Inside the Chunk: RegulaRAG for Regulation-Compliant Scenario Generation using LLMs: A Case Study of UN Regulation No. 152
arXiv:2608.16394v1 Announce Type: new Abstract: Generating regulation-compliant test scenarios is essential for validating safety-critical automotive…
