arXiv:2608.19759v1 Announce Type: cross Abstract: Multifingered grasping is a crucial robotic skill, but current deep-learning grasp planners often…
Category: cs.AI updates on arXiv.org
Learning to Beat: Phenotype-Guided Latent Flow with Regional Motion Priors for Biventricular Motion Synthesis
arXiv:2608.19738v1 Announce Type: cross Abstract: Full-cycle biventricular geometry is essential for characterizing cardiac function. However, dense and…
Truncate Bad, Upweight Good: BoN-Style Distillation via Rank-Based Classification
arXiv:2608.19748v1 Announce Type: cross Abstract: Inference-time selection methods, such as Best-of-N, improve generation by sampling a pool of candidates…
Loreley: Repository-Scale Program Evolution with Quality-Diversity Search
arXiv:2608.19703v1 Announce Type: cross Abstract: Sequential agent search accumulates changes from its current champion but discards alternative branches;…
Scale-Separated Conditioning for Style-Encoder-Free Diffusion Stylization
arXiv:2608.19719v1 Announce Type: cross Abstract: Reference-based diffusion stylization requires separating target geometry from transferable appearance.…
A Locally Tokenized Generative Model for Robust Time-Series Watermarking
arXiv:2608.19727v1 Announce Type: cross Abstract: Watermarking is a central tool for provenance in generative models, yet its application to multivariate…
Robust Cross-Modal Foundation Model Perception for Underwater Robots under Degraded Visual Conditions
arXiv:2608.19710v1 Announce Type: cross Abstract: Reliable underwater robotic perception remains difficult because optical imagery degrades under…
TempJail: Temporal Jailbreak Attack against Large Vision-Language Models via Subtitle Scheduling
arXiv:2608.19737v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have achieved remarkable progress in video understanding and…
VGI-BENCH: Probing Visual Intelligence in Video Generation Models
arXiv:2608.19583v1 Announce Type: cross Abstract: Recent studies suggest that video generation models can exhibit certain forms of zero-shot visual…
Forking Fast: Efficiently Estimating Uncertainty Dynamics in Text Generation
arXiv:2608.19611v1 Announce Type: cross Abstract: LLM reasoning is stochastic, and so understanding a model requires grappling with the distribution of…
