arXiv:2608.26410v1 Announce Type: cross Abstract: Recent work in fair division has focused on either simultaneously satisfying closely related fairness…
Tag: cs.AI updates on arXiv.org
Diff Mining: Logit Differences Reveal Finetuning Objectives
arXiv:2608.26462v1 Announce Type: cross Abstract: Finetuning has become the gold standard for refining existing behaviors and inducing new ones in…
Redwood: A Frontier AI Accelerator Designed, Verified, and Deployed from Scratch in 2 Weeks by AI
arXiv:2608.26418v1 Announce Type: cross Abstract: Modern AI workloads and the hardware that runs them evolve on different timescales: architectural…
Decay-Region Group Delay as a Forensic Cue for AI-Generated Impulsive Sounds
arXiv:2608.26346v1 Announce Type: cross Abstract: We investigate whether AI-generated impulsive sounds can be distinguished from real ones through group…
Co-Evolving Structured Knowledge and Reasoning in Language Models
arXiv:2608.26386v1 Announce Type: cross Abstract: Retrieval-augmented methods improve factual accuracy by grounding language models in external knowledge,…
CG4AI: A Column Generation Framework for Training AI Models Under Constraints
arXiv:2608.26375v1 Announce Type: cross Abstract: Standard machine-learning training minimizes a loss function over a dataset, but does not guarantee that…
Why RAGs Hallucinate: Penalty-Aware Evaluation of Retrieval-Augmented Generation Systems with Knowledge-Gap Canaries
arXiv:2608.26385v1 Announce Type: cross Abstract: Volume-based accuracy rewards retrieval-augmented generation (RAG) systems for guessing: a system that…
Knowledge-Verified Emergent Deception in LLM Agents Under Conflicting Incentives
arXiv:2608.26372v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous agents serving users on behalf of…
MemToC: Benchmarking Memory-Tool Conflict Resolution in Large Language Models
arXiv:2608.26295v1 Announce Type: cross Abstract: Tool-augmented LLMs must arbitrate between two fallible sources when a tool return conflicts with their…
How Unlikely Is “Unlikely”? Assessing Verbal Probability Perception Across Large Language Models
arXiv:2608.26327v1 Announce Type: cross Abstract: Large language models increasingly produce and interpret verbal probability expressions, yet whether…
