arXiv:2608.17270v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for scientific hypothesis generation. However,…
Category: cs.AI updates on arXiv.org
DeAR: Decentralized Agentic Reasoning via Capability Grounding and Collaborative Thought Navigation
arXiv:2608.17282v1 Announce Type: new Abstract: Existing agentic reasoning systems typically rely on centralized protocols. This design introduces routing…
Explicit State Elicitation Is Not Enough: A Controlled Audit of Memory-Policy Classification
arXiv:2608.17247v1 Announce Type: new Abstract: Personalized agents must decide whether retrieved user memory should be used, ignored, updated, or queried…
Fool’s Gold: Defensive Deception Against Safety-Removal Attacks on Open-Weight Models
arXiv:2608.17202v1 Announce Type: new Abstract: Safety alignment in open-weight language models is trivially removable: abliteration projects a…
Benchmarking the Benchmarks: Evaluating Automated Safety Benchmarks for Small Language Models
arXiv:2608.17183v1 Announce Type: new Abstract: Small Language Models (SLMs) are increasingly deployed in resource-constrained, privacy-sensitive…
KnowSim: Evaluating Information Calibration in LLM Assistants with User Simulators that Learn
arXiv:2608.17150v1 Announce Type: new Abstract: To effectively collaborate with users on knowledge-intensive tasks, Large Language Models (LLMs) must…
Toward Personal Intelligence Through Cooperative Observation
arXiv:2608.17128v1 Announce Type: new Abstract: A personal AI system needs a model of the user’s goals, constraints, and ongoing commitments to plan and…
Synthesizing Feature Extractors: An Agentic Approach for Algorithm Selection
arXiv:2608.17170v1 Announce Type: new Abstract: Algorithm selection for constraint satisfaction problems requires extracting features that capture problem…
A decodability criterion predicts when hidden-state selection beats majority voting in large language models
arXiv:2608.17124v1 Announce Type: new Abstract: Combining the answers a large language model (LLM) samples for a question into one decision is a test-time…
DiSCO: Defending text-to-image generation through distribution-guided contrastive prompt optimization
arXiv:2608.17067v1 Announce Type: new Abstract: As text-to-image generative models advance, they raise critical safety concerns, particularly the…
