arXiv:2601.08146v4 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting…
Category: cs.AI updates on arXiv.org
Shiva-DiT: Residual-Based Differentiable Top-$k$ Selection for Efficient Diffusion Transformers
arXiv:2602.05605v2 Announce Type: replace-cross Abstract: Diffusion Transformers (DiTs) are costly at high resolution because self-attention scales…
ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs
arXiv:2603.18579v2 Announce Type: replace-cross Abstract: Evaluating whether explanations faithfully reflect a model’s reasoning remains an open problem.…
Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks
arXiv:2601.22396v3 Announce Type: replace-cross Abstract: Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the…
ContextAnyone: Context-Aware Diffusion for Character-Consistent Text-to-Video Generation
arXiv:2512.07328v2 Announce Type: replace-cross Abstract: Text-to-video generation has advanced rapidly, yet preserving a character’s holistic appearance…
SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning
arXiv:2511.09681v3 Announce Type: replace-cross Abstract: Visual reinforcement learning has achieved remarkable progress in visual control and robotics,…
Inference-Time Optimization of Prompt Embeddings in Diffusion Models: A Comparison of sep-CMA-ES and Adam
arXiv:2511.03913v3 Announce Type: replace-cross Abstract: Deep diffusion models have revolutionized image generation by producing high-quality outputs.…
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
arXiv:2512.16310v4 Announce Type: replace-cross Abstract: LLM agents can combine individually non-revealing tool returns and disclose a sensitive…
An Energy-Based Mechanism for Compositional Behavior
arXiv:2512.04745v4 Announce Type: replace-cross Abstract: Flexible intelligence relies on the ability to reuse previously acquired behaviors and combine…
Multimodal Language Models as Text-to-Image Model Evaluators
arXiv:2505.00759v3 Announce Type: replace-cross Abstract: The steady improvements of text-to-image (T2I) generative models lead to slow deprecation of…
