ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels…
Category: AI
TENET: One Step Toward Test-Driven Development for Repository-Level Code Generation
arXiv:2509.24148v4 Announce Type: replace-cross Abstract: Test-Driven Development (TDD) is a widely adopted practice that requires developers to create…
Edge Case Detection in Automated Driving: Methods, Challenges, and Future Directions
arXiv:2410.08491v3 Announce Type: replace-cross Abstract: Automated vehicles (AVs) promise to enhance transportation safety and efficiency. However,…
Musical Agent Systems: MACAT and MACataRT
arXiv:2502.00023v2 Announce Type: replace-cross Abstract: Our research explores the development and application of musical agents, human-in-the-loop…
BAT: Learning to Reason about Spatial Sounds with Large Language Models
arXiv:2402.01591v4 Announce Type: replace-cross Abstract: Spatial sound reasoning is a fundamental human skill, enabling us to navigate and interpret our…
OTIS: Learning High-Quality Time Series Features With Tiny Encoders
arXiv:2410.07299v3 Announce Type: replace-cross Abstract: We introduce OTIS, an open time series encoder that yields high-quality time series features for…
Leveraging Few-Shot Learning and Large Language Models for Analyzing Blood Pressure Variations Across Biological Sex from Scientific Literature
arXiv:2402.01826v2 Announce Type: replace-cross Abstract: Current blood pressure (BP) technologies and standards were established decades ago, and these…
Making AI-Generated Feedback Matter: A Large-Scale Study of Feedback Workflows and Student Enactment
arXiv:2608.11625v2 Announce Type: replace Abstract: Feedback processes strongly influence student learning, yet their educational value depends on…
LLM-Guided Graph Generation for Structure-Based Local Improvement Methods
arXiv:2608.13333v2 Announce Type: replace Abstract: Large neighborhood search normally selects a random subset of decision variables for iterative…
SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models
arXiv:2608.10538v2 Announce Type: replace Abstract: Agent skills represent a standardized format for packaging procedural knowledge and domain expertise,…
