arXiv:2609.10221v2 Announce Type: new Abstract: Reinforcement learning over a frozen reasoner has become a common recipe for teaching a policy which…
Category: cs.AI updates on arXiv.org
Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather Alerts
arXiv:2609.10135v1 Announce Type: new Abstract: To address insufficient contextualization, weak generalization, and poor scenario adaptation in tourism…
Kernel-Managed Shared Memory for System-Wide Personalization
arXiv:2609.10144v1 Announce Type: new Abstract: AI systems become more useful when they can adapt to the people using them, but in multi-agent systems,…
What Should an Agent Forget? Separating What Is Stored from What Is Used
arXiv:2609.10263v1 Announce Type: new Abstract: Persistent language agents need stored experience to remain available across time, while each answer…
OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology Normalization
arXiv:2609.10055v1 Announce Type: new Abstract: Biomedical ontology normalization maps free-text expressions to standardized concepts, enabling consistent…
RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition Biases
arXiv:2609.10092v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as research agents, yet their ability to track shifts in…
Reference-Based Bias Detection in LLMs via Relative Representations of Hidden States
arXiv:2609.10060v1 Announce Type: new Abstract: Existing bias auditing methods typically rely on model outputs, requiring costly benchmarks or judge…
Belief-State Engine: Augmenting LLMs for Principled Planning Under Partial Observability
arXiv:2609.10036v1 Announce Type: new Abstract: Large language model agents produce fluent action sequences across a wide range of tasks, yet they fail in…
Structural Process Supervision for Latent Chain-of-Thought Reasoning
arXiv:2609.09928v1 Announce Type: new Abstract: Latent reasoning approaches enhance token-level efficiency and robustness by replacing verbose, explicit…
Grounded Evaluation and Repair for NL-to-PDDL Problem Generation
arXiv:2609.09898v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promise for translating Natural Language (NL) planning…
