When asked about unplanned pregnancies, AI chatbots regularly link to anti-abortion groups without disclosing their stance. In an AlgorithmWatch…
Tag: AI Security
No Judgment Without a Reason: Counterfactual Receipts for Versioned AI Evaluators
arXiv:2608.20938v1 Announce Type: new Abstract: Evaluators often produce correct labels via flawed reasoning, a critical failure for agentic systems…
Prediction certification cannot replace explanation certification: a competence envelope for trustworthy AI under compound stress
arXiv:2608.20825v1 Announce Type: new Abstract: Artificial intelligence systems increasingly make consequential judgments – which patient is…
SAGE: A Unified Algebra and Self-Adaptive Execution for AI Functions in SQL
arXiv:2608.20630v1 Announce Type: new Abstract: SQL systems increasingly expose AI functions for tasks such as classification, extraction, filtering,…
Volumetric Radiology AI in the Era of Multimodal Large Language Models
arXiv:2608.20549v1 Announce Type: new Abstract: Advances in multimodal large language models (MLLMs) are extending radiological artificial intelligence…
Terminal Agents: A Survey of AI Agents in Command-Line Environments
arXiv:2608.20485v1 Announce Type: new Abstract: Large language model agents increasingly act through terminals, yet existing surveys disperse…
Categorical AI phenomenology: A first-person approach
arXiv:2608.20420v1 Announce Type: new Abstract: This paper develops a phenomenology-first approach to artificial consciousness by reframing consciousness…
PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure
arXiv:2608.20342v1 Announce Type: new Abstract: Large language model (LLM) coding agents start each session with an empty context window, discarding…
Is it legal to train AI models on copyrighted books? It’s complicated
Most published authors have, without their knowledge or consent, contributed to the development of the same AI tools that threaten to undermine their…
An AI boss fired its first employee but only after humans reminded it of its own rules
Andon Labs’ AI agent Luna fired a human employee at a San Francisco store for the first time but needed a clear push from the operators to do it. When the…
