arXiv:2509.12626v4 Announce Type: replace-cross Abstract: Aligning agentic AI with user intent is critical for delegating complex, socially embedded…
Category: cs.AI updates on arXiv.org
Generating Individual Travel Diaries Using Large Language Models Informed by Census and Land-Use Data
arXiv:2509.09710v3 Announce Type: replace-cross Abstract: This study introduces a Large Language Model (LLM) scheme for generating key attributes of…
Post-Training Large Language Models via Reinforcement Learning from Self-Feedback
arXiv:2507.21931v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often produce plausible but poorly-calibrated answers, limiting…
R3: Robust Rubric-Agnostic Reward Models
arXiv:2505.13388v4 Announce Type: replace-cross Abstract: Reward models are essential for aligning language model outputs with human preferences, yet…
CLIP Embeddings for AI-Generated Image Detection: A Few-Shot Study with Lightweight Classifier
arXiv:2505.10664v2 Announce Type: replace-cross Abstract: Verifying the authenticity of AI-generated images presents a growing challenge on social media…
CBW: Towards Dataset Ownership Verification for Speaker Verification via Clustering-based Backdoor Watermarking
arXiv:2503.05794v4 Announce Type: replace-cross Abstract: Speaker verification models are trained on large-scale public datasets whose licenses usually…
Attention is All You Need Until You Need Retention
arXiv:2501.09166v2 Announce Type: replace-cross Abstract: Pretrained Transformers keep what they learned in their weights and lose what they observe once…
Measuring Human Contribution in AI-Assisted Content Generation
arXiv:2408.14792v4 Announce Type: replace-cross Abstract: With the growing prevalence of generative artificial intelligence (AI), an increasing amount of…
When majority rules, minority loses: bias amplification of gradient descent
arXiv:2505.13122v3 Announce Type: replace-cross Abstract: Despite growing empirical evidence of bias amplification in machine learning, its theoretical…
Who Teaches Which Token? Verifier-Gated Multi-Expert On-Policy Distillation for Scientific Reasoning
arXiv:2609.15404v2 Announce Type: replace Abstract: Multi-teacher on-policy distillation (OPD) is becoming the standard way to integrate specialist…
