arXiv:2608.18089v1 Announce Type: cross Abstract: Instruction-tuned models often refuse harmful requests in English but comply with the same requests in…
Category: cs.AI updates on arXiv.org
Beyond the Transcript: Detecting Covert Co ordination in Latent Multi-Agent Communication
arXiv:2608.19161v1 Announce Type: new Abstract: Language-model agents can communicate through continuous hidden states that are invisible in public…
SuTRA : Structurally-Unified Tokenization with Root Awareness
arXiv:2608.18087v1 Announce Type: cross Abstract: Existing subword tokenizers optimize statistical compression but ignore morphological structure,…
Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk
arXiv:2608.19073v1 Announce Type: new Abstract: An agent still learning its environment should be cautious while ignorant and bold once confident. The…
Tuning the Stochastic Machine: A Systems Engineer’s Operating Model for Human-AI Engineering
arXiv:2608.19125v1 Announce Type: new Abstract: When an expert corrects an LLM assistant’s error, the correction usually dies with the session, and the…
Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering
arXiv:2608.19029v1 Announce Type: new Abstract: Accurate and responsible medical question answering (QA) is important in healthcare, where complex cases…
Eureka: Task-Conditioned Meta-Agent Orchestration for Scientific Discovery
arXiv:2608.19047v1 Announce Type: new Abstract: We present Eureka, a task-conditioned Meta-Agent architecture that compiles long-horizon tasks into…
What is Missing from AI Post-Training AI: An Empirical Analysis
arXiv:2608.19072v1 Announce Type: new Abstract: Large language model (LLM) agents can now post-train an LLM end-to-end. They can write code, launch…
\textsc{TestifAI}: Tomography-Based Testing for Deep Learning Systems
arXiv:2608.18900v1 Announce Type: new Abstract: As AI systems are increasingly deployed in safety-critical application domains (e.g., autonomous driving),…
Breaking the weakest link to evade vision language models
arXiv:2608.18938v1 Announce Type: new Abstract: Vision Language Models (VLMs) have recently emerged as a critical component of multimodal AI systems,…
