arXiv:2608.14227v2 Announce Type: replace Abstract: Preprocessing invariance is an appealing goal for spectral foundation models: a frozen model should…
Tag: AI
Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic Intelligent Interfaces
arXiv:2608.11354v2 Announce Type: replace Abstract: Modern recommender systems treat observed actions as reliable proxies for user preferences, yet…
TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Personalized Generation
arXiv:2608.08446v2 Announce Type: replace Abstract: Personalized generation systems retrieve user history by request–memory relevance and inject it into…
MemWM: Memory-Augmented Text-Based World Model
arXiv:2608.07107v2 Announce Type: replace Abstract: World models are increasingly used to support planning in agents by predicting how environment states…
Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability
arXiv:2608.11238v2 Announce Type: replace Abstract: Retrieval-augmented generation improves the factuality of large language models by grounding responses…
Fragility of Value under Imperfect Alignment
arXiv:2607.28881v4 Announce Type: replace Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that…
The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs
arXiv:2607.08734v2 Announce Type: replace Abstract: Post-Training Quantization has become widely used to compress large language models to make them…
Share the Judge, Learn the Deferral: Where Specialization Helps LLM Evaluation
arXiv:2607.27984v2 Announce Type: replace Abstract: Agentic systems generate outputs faster than human review. We contrast two LLM evaluator…
When Words Are Safe But Actions Kill: Probing Physical Jailbreak Beyond Textual Jailbreak in Hidden-State Risk Space
arXiv:2607.15218v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly serve as high-level planners for embodied agents, where…
WorldLines: Benchmarking and Modeling Long-Horizon Stateful Embodied Agents
arXiv:2606.18847v2 Announce Type: replace Abstract: To assist humans over extended periods in real homes, embodied agents must remember user routines,…
