arXiv:2502.01555v3 Announce Type: replace-cross Abstract: Associating user search queries with the correct brand entity is critical for e-commerce product…
Category: cs.AI updates on arXiv.org
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning
arXiv:2410.20487v5 Announce Type: replace-cross Abstract: Experience replay is widely used to improve learning efficiency in reinforcement learning by…
Predicting Estimated Times of Restoration for Electrical Outages Using Longitudinal Tabular Transformers
arXiv:2505.00225v2 Announce Type: replace-cross Abstract: Utilities publish Estimated Times of Restoration (ETRs) for customer-facing storm outages, and…
Safe Learning Under Irreversible Dynamics via Asking for Help
arXiv:2502.14043v3 Announce Type: replace-cross Abstract: Most learning algorithms with formal regret guarantees essentially rely on trying all possible…
Equity Promotion in Online Resource Allocation
arXiv:2112.04169v3 Announce Type: replace-cross Abstract: We consider online resource allocation under a typical non-profit setting, where limited or even…
A Taxonomy of Architecture Options for Foundation Model-based Agents: Analysis and Decision Model
arXiv:2408.02920v2 Announce Type: replace-cross Abstract: The rapid advancement of AI technology has led to widespread applications of agent systems…
Influence-Oriented Personalized Federated Learning
arXiv:2410.03315v2 Announce Type: replace-cross Abstract: Federated learning (FL) is a machine learning paradigm where clients with different behaviors…
BTBR: A Bayesian-Theory-Driven Probabilistic-Fuzzy Framework for Implicit Bias Removal in Large Language Models
arXiv:2408.10608v2 Announce Type: replace-cross Abstract: Large language models (LLMs) may encode biased associations from heterogeneous training corpora…
Incentives to Offer Algorithmic Recourse
arXiv:2301.12884v2 Announce Type: replace-cross Abstract: Algorithmic recourse promises to help applicants rejected by automated systems by explaining the…
Beyond Prompts: Measuring and Optimizing LLM Tool-Agent Harnesses
arXiv:2609.05736v2 Announce Type: replace Abstract: LLM tool agents can be improved without retraining by modifying the runtime harness around a fixed…
