arXiv:2609.10750v1 Announce Type: cross Abstract: LLM agents increasingly rely on external skills retrieved at runtime, making skill selection from large…
Tag: cs.AI updates on arXiv.org
Beyond Static Guarantees: Measuring the Static-Pass Dynamic-Fail Gap in Security-Sensitive and LLM-Generated Python Code
arXiv:2609.10762v1 Announce Type: cross Abstract: Advances in large language models (LLMs) fuel the quest for scalable methods to assess the security of…
Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu
arXiv:2609.10758v1 Announce Type: cross Abstract: Multilingual large language models (LLMs) are increasingly used for open-ended text generation, yet…
Adaptive Margin Ordinal Loss: Penalizing Center-Class Hedging in Ordinal Classification
arXiv:2609.10752v1 Announce Type: cross Abstract: Standard cross-entropy loss causes neural networks trained on ordinal classification tasks to hedge…
AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activation Flow
arXiv:2609.10723v1 Announce Type: cross Abstract: Text-to-image diffusion transformers (DiTs) are powerful generators, yet direct prompting provides…
Data-Efficient Language Modeling: From Frontier Advancement to Principle-Guided Model Improvement
arXiv:2609.10702v1 Announce Type: cross Abstract: Learning from limited text requires models to use context, generalize to new inputs, and retain useful…
Architecting the Secure AI-SOC: A Neurosymbolic Framework for Pipeline Integrity and Threat Mitigation
arXiv:2609.10707v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Security Operations Centers (SOCs) streamlines…
CARTS: Contextual Autoregressive Rank Transcoding Steganography for Full-Capacity Keyed Text Encoding
arXiv:2609.10744v1 Announce Type: cross Abstract: Autoregressive language models can be used to transform a payload text into a stegotext of identical…
The Truth Was Never Gone: Perfect Aliasing in Compliant-Context Truth Probes
arXiv:2609.10739v1 Announce Type: cross Abstract: A truth probe fitted where truthful reporting and a task’s prescribed action coincide cannot distinguish…
Beyond Verified Answers: Solver-Informed Self-Distillation for Bootstrapping Operations Research Language Models
arXiv:2609.09957v1 Announce Type: cross Abstract: Modern large language models (LLMs) can translate natural-language descriptions into operations research…
