arXiv:2609.09213v1 Announce Type: cross Abstract: We study how physical-state inputs affect a 0.8B hybrid language model adapted for manipulation with…
Author: script
AgenticGen: Reward-Guided Agentic Video Generation for Advertising
arXiv:2609.09187v1 Announce Type: cross Abstract: Advertising video generation is not only a video synthesis task, but also a product-conditioned…
The Mathematical AI Safety Institute wants to prove AI is safe the way cryptographers prove codes are unbreakable
Canadian mathematician Jacob Tsimerman, a fresh Fields Medal recipient, has announced the founding of the Mathematical A.I. Safety Institute (MAISI). The…
AgentHijack: Visual Patch Attacks on Multimodal Computer-Use Agents
arXiv:2609.09212v1 Announce Type: cross Abstract: This paper presents an end-to-end evaluation framework for image-triggered command injection against…
JarvisGUI: Towards Cross-Device GUI Agents with Dynamic Task Composition
arXiv:2609.10451v1 Announce Type: new Abstract: Real-world GUI usage frequently involves workflows that span multiple devices and platforms, requiring the…
Trust Me, I’m Your Developer: Self-Issued Authentication in Large Language Models
arXiv:2609.03247v1 Announce Type: cross Abstract: Large language model (LLM) security has largely focused on role-playing jailbreaks, with less attention…
Characterizing Text Branch Sensitivity in Medical Vision-Language Segmentation via Evidence Decoupling
arXiv:2609.02663v1 Announce Type: cross Abstract: Pretrained vision-language models (VLMs) have shown promising performance in medical image segmentation…
Quantifying Logical Consistency in Transformers via Query-Key Alignment
arXiv:2502.17017v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive performance in various natural language…
Anthropic’s $1.5 billion book settlement descends into chaos as authors and publishers fight over who gets paid
Authors and publishers fight over how to split Anthropic’s $1.5 billion settlement, the largest copyright deal in US history. The article Anthropic’s $1.5…
From Plausible to Actionable: A Position on LLM Self-Explanations
arXiv:2607.15957v3 Announce Type: cross Abstract: Large Language Models (LLMs) can generate natural language explanations that rationalize their own…
