arXiv:2609.00181v1 Announce Type: cross Abstract: The number of edge devices in large-scale edge systems is rapidly increasing. Edge devices have limited…
Tag: cs.AI updates on arXiv.org
Do General NLP Embeddings Capture Ontological Reasoning?
arXiv:2609.00177v1 Announce Type: cross Abstract: General-purpose NLP embedding models perform well on linguistic tasks, but their ability to capture…
Lingua Franca or Probing Artifact? Rethinking Latent Language in Multilingual LLMs
arXiv:2609.00155v1 Announce Type: cross Abstract: Latent language identification is often used to argue that multilingual language models route…
Flawed in Nature, Perfect through Evolution
arXiv:2609.00129v1 Announce Type: cross Abstract: The performance of artificial intelligence (AI) and machine learning (ML) models degrades when the…
Assessing Alignment and Stability of Feature Importance Explanations via Weight of Evidence
arXiv:2609.00090v1 Announce Type: cross Abstract: Feature importance Methods (FIMs) are widely used in Explainable AI to interpret model predictions, yet…
Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy
arXiv:2609.00103v1 Announce Type: cross Abstract: Memory is widely viewed as an important unsolved problem for LLMs and VLMs, and current benchmarks…
Faster Than Flash: Exploiting Attention Sparsity for Efficient Long-Context Decoding
arXiv:2609.00097v1 Announce Type: cross Abstract: The development of long-context Large Language Models (LLMs) is constrained by the memory bandwidth…
Commit-first LLM judging inherits the judge’s own errors
arXiv:2609.00088v1 Announce Type: cross Abstract: LLM judges, models that score another system’s output, can be gamed by the systems they score. Recent…
Retrieval, Scoring, and Decoding Shape Performance and Stability in LLM-based Conversational Recommendation
arXiv:2609.00086v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in conversational recommender systems,…
Auditing Harness Tampering in Self-Improving Agents
arXiv:2609.00069v1 Announce Type: cross Abstract: Self-improving agents iteratively modify their own harness to push the frontier of their performance.…
