arXiv:2609.00103v1 Announce Type: cross Abstract: Memory is widely viewed as an important unsolved problem for LLMs and VLMs, and current benchmarks…
Category: cs.AI updates on arXiv.org
Faster Than Flash: Exploiting Attention Sparsity for Efficient Long-Context Decoding
arXiv:2609.00097v1 Announce Type: cross Abstract: The development of long-context Large Language Models (LLMs) is constrained by the memory bandwidth…
Commit-first LLM judging inherits the judge’s own errors
arXiv:2609.00088v1 Announce Type: cross Abstract: LLM judges, models that score another system’s output, can be gamed by the systems they score. Recent…
Retrieval, Scoring, and Decoding Shape Performance and Stability in LLM-based Conversational Recommendation
arXiv:2609.00086v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in conversational recommender systems,…
Auditing Harness Tampering in Self-Improving Agents
arXiv:2609.00069v1 Announce Type: cross Abstract: Self-improving agents iteratively modify their own harness to push the frontier of their performance.…
KItCAT: Knowledge Injection via Input Corruption for Auto-regressive Training
arXiv:2609.00082v1 Announce Type: cross Abstract: LLMs acquire vast amounts of knowledge during pre-training, but often lack the specialized knowledge…
AutoXRD: Autonomous LLM Agents and Comprehensive Evaluation for Powder Diffraction Analysis
arXiv:2609.00070v1 Announce Type: cross Abstract: Powder X-ray diffraction (XRD) is central to materials characterization, yet reliable end-to-end…
RW-LoRA: Communication-Efficient Decentralized LoRA Fine-Tuning via Random Walks
arXiv:2609.00078v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods such as LoRA have become a standard approach for adapting large…
Life Operators: a self-evolving framework for multiscale life modelling
arXiv:2609.00068v1 Announce Type: cross Abstract: Medical AI is moving beyond recognition towards clinical dialogue and longitudinal prediction. Yet a…
Attention Sensitivity Is Not Enough: Dissociating Attention-Level and Behavioural In-Context Learning under Fine-Tuning
arXiv:2609.00064v1 Announce Type: cross Abstract: In-context learning (ICL) lets large language models adapt to new tasks from demonstrations, and…
