arXiv:2608.21499v1 Announce Type: cross Abstract: Cardiac auscultation remains the most cost-effective screening procedure for cardiovascular diseases,…
SecOPD: Mitigating Adaptive Prompt Injections by On-Policy Distillation
arXiv:2608.21500v1 Announce Type: cross Abstract: Prompt injection is listed as the \#1 threat to AI agents. When an agent accesses external data from…
TASSO: TAsk-Specific Subspace Optimization for Continual Learning of Vision-Language Models
arXiv:2608.21487v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) exhibit strong zero-shot capabilities, making them an attractive solution…
KAN-Robust-Bench: A Benchmark for Evaluating the Robustness of Kolmogorov-Arnold Networks
arXiv:2608.21488v1 Announce Type: cross Abstract: While machine learning models have demonstrated strong performance in many domains, these models have…
presto: Efficient, Training-free, and Open-world Object Placement via Imaginary Search
arXiv:2608.21543v1 Announce Type: cross Abstract: Object placement is critical in image composition, requiring spatially and semantically coherent…
Reliability- and Anatomy-Consistency-Aware Multimodal Learning for Robust Fracture Classification from Bangladeshi Radiographs
arXiv:2608.21482v1 Announce Type: cross Abstract: Background: Multimodal fracture classifiers may benefit from patient and anatomical metadata, but they…
FigmaTrace: Capturing Creative Nuances in Human Figma Design Workflows
arXiv:2608.21460v1 Announce Type: cross Abstract: Vision Language Models have recently shown improvements in several objective and verifiable domains such…
Complexity Induction: Compositional Generalization via Structured Label Distortion
arXiv:2608.21464v1 Announce Type: cross Abstract: We demonstrate that structured distortion of training data – which we term complexity induction – can…
Constructing Predictive Surgical Path for AI-based Capsulorhexis Skill Transfer
arXiv:2608.21441v1 Announce Type: cross Abstract: Automated training of surgeons is one of the most crucial factors that significantly minimize surgical…
CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance
arXiv:2608.21462v1 Announce Type: cross Abstract: Due to the selection of their training data, large language models (LLMs) perform best on…
AI News Brief Hourly Summary 2026-08-26 01h : 12 posts
12 posts published in the last hour 22:32Mitigating Bias in Large Vision-Language Models via Counterfactual Ensemble Decoding 22:32Agentic Security: A Systematization of Tools, Failure Modes, and Design Laws for LLM-Driven Penetration Testing 22:32Operational digital twin clinics enable task-based evaluation of…
Mitigating Bias in Large Vision-Language Models via Counterfactual Ensemble Decoding
arXiv:2608.21415v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable performance across a wide range of tasks;…
Agentic Security: A Systematization of Tools, Failure Modes, and Design Laws for LLM-Driven Penetration Testing
arXiv:2608.21423v1 Announce Type: cross Abstract: Agentic security uses large-language-model (LLM) agents to plan, dispatch, and interpret security tools.…
Operational digital twin clinics enable task-based evaluation of embodied AI
arXiv:2608.21416v1 Announce Type: cross Abstract: Embodied artificial intelligence (AI) must be tested in the clinical environments where it will operate,…
Aligning Human Sense: Calibrated Distributional Reward Learning for Video Generation
arXiv:2608.21425v1 Announce Type: cross Abstract: Video generation is central to AI-powered content creation. Aligning generated videos with human…
AI Method Reveals What Genomic Models Learn From DNA and Exposes Hidden Experimental Bias
Researchers at the Stowers Institute for Medical Research have built an interpretation method that shows, base by base, what a deep-learning model has…
Geo-VLA: Geometry-Aware Vision-Language-Action Planning via Internalization of Map Semantics
arXiv:2608.21440v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have advanced end-to-end autonomous driving by leveraging foundation…
Mamba-based Selective State Space Modeling Improves the Accuracy-Complexity Tradeoff of SmolVLA Vision-Language-Action Experts
arXiv:2608.21407v1 Announce Type: cross Abstract: Vision-language-action (VLA) models face a crucial tradeoff between their task success rate and the…
