arXiv:2609.24620v1 Announce Type: new Abstract: Answering epidemiological questions from real-world clinical data requires medical coding, schema-aware…
Tag: AI
The Endless Exam: Mathematical Constructions from Today’s Models toward Superintelligence
arXiv:2609.24555v1 Announce Type: new Abstract: We introduce the Endless Exam, a benchmark for measuring mathematical progress from today’s models toward…
Not All Task Vectors Need Equal Rank: Energy-Proportional Allocation for Model Merging
arXiv:2609.24517v1 Announce Type: new Abstract: Model merging aims to combine multiple fine-tuned models derived from a common pretrained model into a…
Spotify’s is giving you the keys to its recommendation algorithm with US launch of ‘Taste Profile’
Spotify is rolling out Taste Profile to Premium users in the U.S., letting listeners see how the streamer understands their tastes and use natural…
DUMA-Bench: A Dual-Control Multi-Agent Benchmark for Evaluating LLM Agent Security
arXiv:2609.24662v1 Announce Type: new Abstract: LLM-based agents increasingly operate in environments where they interact with users, tools, and external…
Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent
Alibaba’s AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), text-to-speech (TTS), and real-time interaction.…
Custom Named Entity Recognition and Topic Classification for Global Health Publications
arXiv:2609.24625v1 Announce Type: new Abstract: How should natural language processing models be selected and adapted for global health literature in…
Predicting Postprandial Glycemic Response from Meal Images, Clinical Variables, and Gut Microbiome Information
arXiv:2609.24453v1 Announce Type: new Abstract: Predicting postprandial glycemic response (PPGR) is fundamental to personalized nutrition and type 2…
LADDER: Graph-Guided Diffusion Language Models for Efficient Multi-Hop Reasoning
arXiv:2609.24346v1 Announce Type: new Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has remarkably enhanced large language models on complex…
Fathom-Vaidya: Advancing Medical Reasoning with Rubric-Based Rewards
arXiv:2609.24480v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in healthcare requires robust performance across two complementary…
