arXiv:2608.21487v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) exhibit strong zero-shot capabilities, making them an attractive solution…
Category: cs.AI updates on arXiv.org
KAN-Robust-Bench: A Benchmark for Evaluating the Robustness of Kolmogorov-Arnold Networks
arXiv:2608.21488v1 Announce Type: cross Abstract: While machine learning models have demonstrated strong performance in many domains, these models have…
presto: Efficient, Training-free, and Open-world Object Placement via Imaginary Search
arXiv:2608.21543v1 Announce Type: cross Abstract: Object placement is critical in image composition, requiring spatially and semantically coherent…
Reliability- and Anatomy-Consistency-Aware Multimodal Learning for Robust Fracture Classification from Bangladeshi Radiographs
arXiv:2608.21482v1 Announce Type: cross Abstract: Background: Multimodal fracture classifiers may benefit from patient and anatomical metadata, but they…
FigmaTrace: Capturing Creative Nuances in Human Figma Design Workflows
arXiv:2608.21460v1 Announce Type: cross Abstract: Vision Language Models have recently shown improvements in several objective and verifiable domains such…
Complexity Induction: Compositional Generalization via Structured Label Distortion
arXiv:2608.21464v1 Announce Type: cross Abstract: We demonstrate that structured distortion of training data – which we term complexity induction – can…
Constructing Predictive Surgical Path for AI-based Capsulorhexis Skill Transfer
arXiv:2608.21441v1 Announce Type: cross Abstract: Automated training of surgeons is one of the most crucial factors that significantly minimize surgical…
CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance
arXiv:2608.21462v1 Announce Type: cross Abstract: Due to the selection of their training data, large language models (LLMs) perform best on…
Mitigating Bias in Large Vision-Language Models via Counterfactual Ensemble Decoding
arXiv:2608.21415v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable performance across a wide range of tasks;…
Agentic Security: A Systematization of Tools, Failure Modes, and Design Laws for LLM-Driven Penetration Testing
arXiv:2608.21423v1 Announce Type: cross Abstract: Agentic security uses large-language-model (LLM) agents to plan, dispatch, and interpret security tools.…
