arXiv:2609.09883v1 Announce Type: cross Abstract: Depth pruning reduces large language model (LLM) inference cost by removing complete Transformer blocks.…
Category: cs.AI updates on arXiv.org
Can AI Agents Detect and Repair Artifact Drift in Network Experiments?
arXiv:2609.09849v1 Announce Type: cross Abstract: In recent years, AI agents have evolved into capable assistants that carry out multi-step tasks in…
uFlowCSP: Crystal Structure Prediction using Mean flow generative models
arXiv:2609.09799v1 Announce Type: cross Abstract: Crystal structure prediction (CSP) is fundamental to computational materials discovery. Generative…
Subgroup Membership Inference Audits of Differentially Private Synthetic Text
arXiv:2609.09848v1 Announce Type: cross Abstract: Synthetic data releases are increasingly proposed in the literature as a means of sharing realistic data…
CS-Guard: Benchmarking LLM Guardrails for Code Generation Security
arXiv:2609.09798v1 Announce Type: cross Abstract: Large language models (LLMs) have been ex- ploited to generate malware, but the effective- ness of…
LogiScope-VQA: Benchmarking Vision-Language Models for Logistics Hazard Identification in Industrial Scenarios
arXiv:2609.09790v1 Announce Type: cross Abstract: Large Multimodal Models (LMMs) large-scale deployment in industrial warehouse settings specifically…
Pairit: A Platform for Live Experiments on Human-AI Collaboration
arXiv:2609.09789v1 Announce Type: cross Abstract: Organizational design in the era of artificial intelligence requires experimental methods that can test…
BRACE: Anchored Bellman-Residual Correction for Stale Critics in Asynchronous RL
arXiv:2609.09783v1 Announce Type: cross Abstract: Asynchronous reinforcement learning has become the standard way to scale training for language models,…
How Fragile Is Safety Alignment at Frontier Scale? A Single-Direction Attack on a 320B MoE
arXiv:2609.09793v1 Announce Type: cross Abstract: Directional ablation removes an aligned language model’s ability to refuse by projecting a single…
Fine-Tuning a KV Cache Concatenation-Aware Model or Recomputing KV Caches? Why Not Both?
arXiv:2609.09768v1 Announce Type: cross Abstract: In Retrieval-Augmented Generation (RAG) systems, a large number of retrieved chunks are concatenated to…
