arXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite…
Tag: cs.AI updates on arXiv.org
Energy and Performance Benchmarking of Deep Learning Models for Breast Cancer Detection
arXiv:2608.09996v1 Announce Type: cross Abstract: Recent advances in machine learning have greatly improved breast cancer detection, enabling more…
Evidence-Based Scientific Question Discovery: A Framework with Historical Backtesting
arXiv:2608.09968v1 Announce Type: cross Abstract: Current AI systems are optimized for answering questions; the scientific enterprise is bottlenecked…
Do AI weather models miss extremes?
arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in…
Eleven Years of BRACIS: A Meta-Scientific Study of the Brazilian Conference on Intelligent Systems
arXiv:2608.09964v1 Announce Type: cross Abstract: The Brazilian Conference on Intelligent Systems (BRACIS) is the main national venue for Artificial…
Rescene: band-limited stochastic forcing turns a frozen neural weather operator into a climate emulator
arXiv:2608.09971v1 Announce Type: cross Abstract: Over the past few years, the rapid development of machine learning (ML) models for weather forecasting…
HoosierHelp: Benchmarking LLM Agents for Social Service Navigation
arXiv:2608.09946v1 Announce Type: cross Abstract: Social service navigation requires connecting help-seeking individuals to resources that satisfy their…
How to Dogfood Your AI Chat Agent: A Three-Layer Evaluation Framework with Goal-Directed NPC Simulation
arXiv:2608.09939v1 Announce Type: cross Abstract: Production teams deploying LLM chat agents face a specific quality assurance gap: existing evaluation…
Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents
arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages…
LLM Agents Factory: Retrieval of Domain-Specific LLM Agents
arXiv:2608.09934v1 Announce Type: cross Abstract: Large language model (LLM) agents improve task performance by decomposing problems into role-specialized…
