arXiv:2608.13517v1 Announce Type: cross Abstract: Current large language model development relies on massive, often non-permissible datasets, creating a…
Tag: cs.AI updates on arXiv.org
Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries and Referent Specificity
arXiv:2608.13484v1 Announce Type: cross Abstract: When asked about entities outside their knowledge boundary, LLMs routinely fabricate plausible-sounding…
CAPRI: Contract-Aware Proof Repair for Isabelle
arXiv:2608.13459v1 Announce Type: cross Abstract: We address the use of large language models (LLMs) to help discover Isabelle proofs. An Isabelle build…
ContactGuard: Pre-Contact Execution Monitoring with Action-Conditioned Latent World Models
arXiv:2608.13438v1 Announce Type: cross Abstract: Contact-rich manipulation failures are often detected only after the robot has committed to contact.…
MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification
arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often…
UniTexture: Cross-Task Universal Adversarial Textures for Vision-Language-Action Models
arXiv:2608.13453v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as generalist robotic policies capable of following…
Concept Drift Detection and Adaptive Retraining of Malware Classification Models
arXiv:2608.13465v1 Announce Type: cross Abstract: Concept drift refers to changes over time in the statistical properties of data, as compared to the data…
Heterogeneity-Aware Belief Synchronization for Semantic Communication in AI-Native 6G Networks
arXiv:2608.13394v1 Announce Type: cross Abstract: 6G networks will not be serving as communication infrastructures only; rather, they are expected to…
Reduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM Inference
arXiv:2608.13426v1 Announce Type: cross Abstract: Transformer-based language models achieve strong performance but incur substantial inference cost due to…
Are You Sure You’re Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity
arXiv:2608.13430v1 Announce Type: cross Abstract: Instruction-tuned language models achieve strong performance across a range of generation tasks, but…
