arXiv:2609.02996v1 Announce Type: cross Abstract: Graph neural networks (GNNs) are a class of neural networks suitable for learning on graph-structured…
Tag: cs.AI updates on arXiv.org
When Optimization Becomes Manipulation: Defending Generative Search against Malicious Generative Engine Optimization
arXiv:2609.02964v1 Announce Type: cross Abstract: This paper focuses on defending generative search engines against malicious Generative Engine…
Toward Collective-Centric Evaluation of Preference Inference for Participatory Democracy
arXiv:2609.02990v1 Announce Type: cross Abstract: To scale up collective decision-making, participatory democracy platforms such as Polis and Remesh…
Verify Before You Distill: Prompt-Level Teacher Gating for On-Policy Distillation
arXiv:2609.02998v1 Announce Type: cross Abstract: On-policy distillation (OPD) accelerates post-training by providing dense token-level supervision from a…
Judging LLM-as-a-Judge: Concerning Rubric Artifacts in LLM-based Automated Text Generation Evaluation
arXiv:2609.02942v1 Announce Type: cross Abstract: LLM-as-a-Judge pipelines are increasingly used to evaluate AI-generated text, based on the assumption…
The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
arXiv:2609.02959v1 Announce Type: cross Abstract: What does a language model predict when it has few clues? The answer lurks in its unembedding geometry:…
Privacy-Preserving Heterogeneous Multi-LLM Federated Inference for Cognitive Diagnosis
arXiv:2609.02947v1 Announce Type: cross Abstract: Significant challenges remain in AI-driven educational systems in balancing privacy preservation with…
Reflect-SQL: A Self-Reflection Based Framework for Text-to-SQL
arXiv:2609.02944v1 Announce Type: cross Abstract: Democratizing data access through natural language is a crucial goal for modern enterprises, but the…
PrivateHub: Contrastive Diffusion Model for Private Sensor-Intensive Environment Data Generation
arXiv:2609.02958v1 Announce Type: cross Abstract: Sensor-intensive environments enable many intelligent services by inferring user applications from…
Counterexamples as Feedback for Agent Self-Correction
arXiv:2609.02892v1 Announce Type: cross Abstract: Single-turn code-generation metrics understate a central property of deployed agents: whether they can…
