arXiv:2609.04086v1 Announce Type: cross Abstract: For every coherent and sufficiently expressive finite syntactic system S, we prove the existence of at…
Category: cs.AI updates on arXiv.org
PatchBench: Evaluating AI Agents for Vulnerability Patching
arXiv:2609.04075v1 Announce Type: cross Abstract: AI agents have recently demonstrated strong performance in automated vulnerability patching. However,…
Subspace Inference Enables Efficient Active Reward Learning from Preferences
arXiv:2609.04066v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has emerged as a powerful yet sample-inefficient…
TAP-Path: Task-Adaptive Structural and Token Pruning for Efficient and Trustworthy Pathology Foundation Models
arXiv:2609.04071v1 Announce Type: cross Abstract: Pathology foundation models improve transferable representation learning for histopathology, but recent…
Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-Informed Deep Learning-FEM Approach
arXiv:2609.04028v1 Announce Type: cross Abstract: The geometric morphology of deposited filaments can significantly influence the structural performance…
Representational alignment yields generalizable safety in language models
arXiv:2609.04022v1 Announce Type: cross Abstract: Aligning large language models (LLMs) is essential for their safe deployment. Current alignment methods…
The Blind Spot in 2D Infants’ Pose Estimation:Robust Learning from Noisy Annotations
arXiv:2609.04009v1 Announce Type: cross Abstract: Noisy annotations pose a significant challenge for supervised deep learning, as neural networks rely on…
Translation as a Decision Space: A Multi-Agent Perspective on Low-Resource Dialect Generation
arXiv:2609.04048v1 Announce Type: cross Abstract: Neural machine translation (NMT) systems typically produce a single output per input, obscuring the…
When Models Edit Too Much: On the Fidelity of Minimal Code Edits
arXiv:2609.04061v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to edit existing code, but correctness alone is not…
Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes
arXiv:2609.03967v1 Announce Type: cross Abstract: Several studies have evaluated the ability of Large Language Models (LLMs) for meal planning, yielding…
