arXiv:2608.19674v1 Announce Type: cross Abstract: Computing has been an astonishing success – but the accumulated technical debt exposes us all to huge…
Category: cs.AI updates on arXiv.org
PEA-DPO: Perception-Enhanced Alignment Direct Preference Optimization for MLLMs Alignment
arXiv:2608.19598v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has emerged as an effective approach for aligning large language…
DeltaML-Bench: Evaluating Machine Learning Agents on Real-World Research Repositories
arXiv:2608.19653v1 Announce Type: cross Abstract: Autonomous agents for machine learning experimentation must navigate heterogeneous repositories, repair…
When Machines Speak: A Unified Generative Framework for Integrating Machine-Native Symbols into Pretrained Large Language Models
arXiv:2608.19529v1 Announce Type: cross Abstract: Many real-world AI systems represent entities, behaviors, and structured information using discrete…
DraftFM: A FoundationModel for Day-Zero Drafting in Magic: The Gathering
arXiv:2608.19568v1 Announce Type: cross Abstract: Drafting a new Magic: The Gathering expansion begins before any pick from it has been observed: the…
Automated Summarization of Financial News Using Large Language Models and Retrieval-Augmented Generation: An Early Empirical Study (Fall 2023)
arXiv:2608.19526v1 Announce Type: cross Abstract: Stock market analysts and investors face a daily challenge: too much financial news, too little time.…
Stream4D: 4D-Consistency for Streaming Autoregressive Diffusion Video Models
arXiv:2608.19556v1 Announce Type: cross Abstract: Streaming autoregressive diffusion models enable real-time, long-horizon video generation, but their…
CVSD-Reg: Cross-Modal Visual Semantic Prior Distillation for Robust LiDAR Registration
arXiv:2608.19536v1 Announce Type: cross Abstract: Learning-based global point cloud registration has achieved remarkable progress, yet its reliance on…
Measuring What a Specification Determines: A Formal Semantic-Block Model and an Execution-Judged Benchmark
arXiv:2608.19475v1 Announce Type: cross Abstract: This work introduces a formal semantic-block model for specifications and an execution-judged benchmark…
Are LLMs becoming similarly creative? Evidence from three years of models
arXiv:2608.19437v1 Announce Type: cross Abstract: Many benchmarks track Large Language Model (LLM) performance on tasks with verifiable answers, but less…
