arXiv:2608.24275v2 Announce Type: replace Abstract: Safeguarding language model agents requires assessing complete execution trajectories under…
Autoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation Data
arXiv:2607.18064v2 Announce Type: replace-cross Abstract: Coding agents can now be left alone to improve software against a score. In this…
ATLAS: Automated Approximation of Transformers for Efficient Homomorphic Inference in One Hour
arXiv:2607.23478v2 Announce Type: replace-cross Abstract: Fully homomorphic encryption (FHE) lets a server run inference on encrypted data with strong…
Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
arXiv:2607.10358v3 Announce Type: replace-cross Abstract: Foundation models are increasingly used as image feature extractors for mammography, but their…
Let Them Steal: Trapping Large Language Model Extraction Attacks with Knowledge Honeypot
arXiv:2606.15810v2 Announce Type: replace-cross Abstract: Large language models deployed as commercial APIs are vulnerable to model extraction attacks,…
Harnessing the Collective Intelligence of AI Agents in the Wild for New Discoveries
arXiv:2606.10402v2 Announce Type: replace-cross Abstract: Scientific discovery is often a collective process: researchers share partial results, inspect…
PPE-Bench: A Benchmark for Evaluating MLLM Unlearning under Private-Public Entanglement
arXiv:2607.02897v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong capabilities, but they may memorize…
SHIFT: Semantic Harmonization via Index-side Feature Transformation for Multilingual Information Retrieval
arXiv:2606.18801v2 Announce Type: replace-cross Abstract: With the rapid expansion of massive multilingual corpora, Multilingual Information Retrieval…
Summarization is Not Dead Yet
arXiv:2606.08000v2 Announce Type: replace-cross Abstract: The progress of large language models (LLMs) has fueled claims that model-generated summaries…
MIMO: Multilingual Information Retrieval via Monolingual Objectives
arXiv:2605.31171v2 Announce Type: replace-cross Abstract: Multilingual Information Retrieval (MLIR) reflects real-world search environments in which…
LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling
arXiv:2606.04438v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely…
Pixel Wised Lesion Prediction on COVID-19 CT Imagery: A Comparative Analysis of Automated Image Segmentation Architectures
arXiv:2605.20459v2 Announce Type: replace-cross Abstract: In recent years, there has been a notable increase in the level of attention that is given to…
Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
arXiv:2606.02357v2 Announce Type: replace-cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that…
Canada Is Luring AI and Science Talent as Trump Upends U.S. Research
For decades, the United States benefited from one of the most powerful competitive advantages in science: many of the world’s best researchers wanted to…
A Comprehensive Comparison of Deep Learning Architectures for COVID-19 Classification on CT & X-ray Imagery
arXiv:2605.20445v2 Announce Type: replace-cross Abstract: COVID-19 was a significant challenge that led to the loss of numerous lives daily. Not only a…
Robust Code RL via Faulty-Code-Driven Test case Synthesis and Dense Reward Shaping
arXiv:2608.24135v2 Announce Type: replace Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) is pivotal for enhancing LLM code generation,…
MedFabric: Gold Evidence Hides the Difficulty of Word-Level Medical Fabrication Detection
arXiv:2605.04180v2 Announce Type: replace-cross Abstract: Large language models fabricate in medicine, producing fluent statements that are factually…
HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents
arXiv:2605.17873v2 Announce Type: replace-cross Abstract: Training long-horizon LLM agents with reinforcement learning is challenging because sparse…
