arXiv:2608.08077v1 Announce Type: new Abstract: Theory of Space framework (ToS) assesses the spatial understanding of curiosity-driven Vision-Language…
Category: cs.AI updates on arXiv.org
CORDA: A Benchmark for Hierarchical Harm-Centric Moral Reasoning in Large Language Models
arXiv:2608.08061v1 Announce Type: new Abstract: The key question in moral judgement is not simply whether someone chooses the “right” answer, but how they…
Generative Models: Principles, Architectures, and Applications
arXiv:2608.08101v1 Announce Type: new Abstract: Generative AI has emerged as one of the most transformative forces in modern artificial intelligence,…
PATH: Next-Interval Prediction via Autoregressive Tree Hierarchy on Tabular Data
arXiv:2608.08078v1 Announce Type: new Abstract: Interval prediction aims to achieve a target coverage level while producing intervals that are as short as…
H2: A Dual Hybrid Semantic Data Lake Architecture for Medical Data Harmonization with Human-In-the-Loop verified, LLM Driven Metadata Annotation System
arXiv:2608.08056v1 Announce Type: new Abstract: Medical data, by its nature, exhibit a high degree of heterogeneity on multiple levels ranging from (a)…
Lingjing: A Simulation Testbed for Multi-Agent Embodied Tasks in Open-Ended Cities
arXiv:2608.08045v1 Announce Type: new Abstract: Urban embodied intelligence requires coordination among heterogeneous agents (e.g., UAVs, ground robots,…
Decided Upstream, Written Late: Locating and Pricing the Cross-Lingual Refusal Circuit of a Multilingual MoE
arXiv:2608.08032v1 Announce Type: new Abstract: Safety alignment in multilingual models is uneven: a model that reliably refuses a harmful request in…
JustLLMGRPO: Radiographic Control for Chest X-Ray Generation
arXiv:2608.08046v1 Announce Type: new Abstract: Text-conditioned chest X-ray generation aims to synthesize realistic radiographs that faithfully depict…
SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents
arXiv:2608.08055v1 Announce Type: new Abstract: Large language model (LLM) agents that assist users over weeks of conversation must remember what is…
SkillSmith: Enhancing Locally Deployed Agents via Automatic Skill Construction and Evolution
arXiv:2608.08037v1 Announce Type: new Abstract: LLM-based agent frameworks now act as personal assistants for multi-step tasks. Existing agent frameworks…