arXiv:2608.06938v1 Announce Type: cross Abstract: The visual reasoning ability of multimodal large language models (MLLMs) is crucial for downstream…
Tag: cs.AI updates on arXiv.org
Ask-E: An Environment for Calibrated Question Generation
arXiv:2608.06933v1 Announce Type: cross Abstract: Today, we improve models by training and evaluating them on problems at the frontier of their abilities.…
MaskFlow: Precise, Consistent and Seamless Regional Image Editing
arXiv:2608.06929v1 Announce Type: cross Abstract: Regional image editing has attracted considerable attention for its spatial controllability. Although…
Explicit, Not Longer: What Makes Epistemic Stance Survive Memory Compression
arXiv:2608.06953v1 Announce Type: cross Abstract: Agent memory systems compress what they store, and compression is built to drop qualifiers, so a claim’s…
Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests
arXiv:2608.06908v1 Announce Type: cross Abstract: We propose Zero-phase Component Analysis (ZCA) whitening as a geometric pre-processing step for the Word…
Bridging the Gap Between Hyperdimensional Computing and Kernel Methods via the Nystr\”om Method
arXiv:2608.06860v1 Announce Type: cross Abstract: Hyperdimensional computing (HDC) is an approach from the cognitive science literature for solving…
Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection
arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos…
FedVAR: Prototype-Aligned Federated Framework for Video Anomaly Recognition
arXiv:2608.06876v1 Announce Type: cross Abstract: In the era of Industrial Internet of Things (IIoT) and Cyber-Physical Systems (CPS), Federated Learning…
Georeferencing Non-Gazetteered Place Names using Biological Specimen Records
arXiv:2608.06884v1 Announce Type: cross Abstract: Biological specimen records collected by natural history institutions constitute a rich source of…
Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry
arXiv:2608.06849v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by quadratic attention computation and growing KV-cache…