arXiv:2609.10315v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has advanced language-model reasoning in domains…
Author: script
Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context Learning
arXiv:2609.10177v1 Announce Type: new Abstract: In-context learning (ICL) is widely used in multimodal large language models (MLLMs) and achieves strong…
Why Sample What You Can Enumerate? Exact Policy Optimization for Genomic Tool Selection
arXiv:2609.10221v2 Announce Type: new Abstract: Reinforcement learning over a frozen reasoner has become a common recipe for teaching a policy which…
Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather Alerts
arXiv:2609.10135v1 Announce Type: new Abstract: To address insufficient contextualization, weak generalization, and poor scenario adaptation in tourism…
Kernel-Managed Shared Memory for System-Wide Personalization
arXiv:2609.10144v1 Announce Type: new Abstract: AI systems become more useful when they can adapt to the people using them, but in multi-agent systems,…
What Should an Agent Forget? Separating What Is Stored from What Is Used
arXiv:2609.10263v1 Announce Type: new Abstract: Persistent language agents need stored experience to remain available across time, while each answer…
AI News Brief Hourly Summary 2026-09-11 10h : 12 posts
12 posts published in the last hour 07:33OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology Normalization 07:33RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition Biases 07:33Reference-Based Bias Detection in LLMs via Relative Representations of Hidden States…
OntologyAligner: Ontology-Aligned Retrieval and Hierarchy-Guided Large Language Model Reranking for Biomedical Ontology Normalization
arXiv:2609.10055v1 Announce Type: new Abstract: Biomedical ontology normalization maps free-text expressions to standardized concepts, enabling consistent…
RAP: Research Attention Prediction Reveals Target-Conditioned Evidence Acquisition Biases
arXiv:2609.10092v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as research agents, yet their ability to track shifts in…
Reference-Based Bias Detection in LLMs via Relative Representations of Hidden States
arXiv:2609.10060v1 Announce Type: new Abstract: Existing bias auditing methods typically rely on model outputs, requiring costly benchmarks or judge…
