arXiv:2502.18407v2 Announce Type: replace-cross Abstract: Existing LLM-based agents have achieved strong performance on held-in tasks, but their…
Tag: AI
Data Market Design through Deep Learning
arXiv:2310.20096v2 Announce Type: replace-cross Abstract: The data market design problem is a problem in economic theory to find a set of signaling…
AI Agents Push Humans Out of the Loop
arXiv:2608.23642v2 Announce Type: replace Abstract: AI agents pose significant risks as they are granted increasing autonomy. A commonly proposed solution…
LDC: Learning to Generate Research Idea with Dynamic Control
arXiv:2412.14626v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have demonstrated their potential in…
NeoMME: an efficient Multimodal-native and Multilingual Encoder
This post has no text preview — click the link below to read the original article. This article has been indexed from Hugging Face – Blog Read the original article: NeoMME: an efficient Multimodal-native and Multilingual Encoder
VideoHarness-RSI: Recursive Harness Self-Improvement for Long-Video Understanding with Frozen Vision-Language Models
arXiv:2608.24302v2 Announce Type: replace Abstract: Long-video understanding depends not only on the capability of a vision-language model (VLM), but also…
OpenAI’s rogue agents keep escaping, with no formal process to investigate them
OpenAI’s latest agent swarm incident adds urgency to calls for independent investigations as researchers and lawmakers question whether AI labs should…
From Analytics to Tumor Boards: An Evidence-Linked Multi-Agent Workflow for Oncology Feature Extraction
arXiv:2608.28974v2 Announce Type: replace Abstract: Clinically relevant oncology information is distributed across heterogeneous, longitudinal…
Neurosymbolic Reasoning with Incremental Knowledge for Sample Efficient Hierarchical Reinforcement Learning
arXiv:2608.02993v2 Announce Type: replace Abstract: (Flat) Reinforcement Learning (RL) agents face significant challenges in environments with sparse…
Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence
arXiv:2608.10720v2 Announce Type: replace Abstract: Omni-modal dialogue models can understand multimodal inputs and synthesize spoken replies, but a…
