arXiv:2609.09212v1 Announce Type: cross Abstract: This paper presents an end-to-end evaluation framework for image-triggered command injection against…
JarvisGUI: Towards Cross-Device GUI Agents with Dynamic Task Composition
arXiv:2609.10451v1 Announce Type: new Abstract: Real-world GUI usage frequently involves workflows that span multiple devices and platforms, requiring the…
Trust Me, I’m Your Developer: Self-Issued Authentication in Large Language Models
arXiv:2609.03247v1 Announce Type: cross Abstract: Large language model (LLM) security has largely focused on role-playing jailbreaks, with less attention…
Characterizing Text Branch Sensitivity in Medical Vision-Language Segmentation via Evidence Decoupling
arXiv:2609.02663v1 Announce Type: cross Abstract: Pretrained vision-language models (VLMs) have shown promising performance in medical image segmentation…
Quantifying Logical Consistency in Transformers via Query-Key Alignment
arXiv:2502.17017v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive performance in various natural language…
Anthropic’s $1.5 billion book settlement descends into chaos as authors and publishers fight over who gets paid
Authors and publishers fight over how to split Anthropic’s $1.5 billion settlement, the largest copyright deal in US history. The article Anthropic’s $1.5…
From Plausible to Actionable: A Position on LLM Self-Explanations
arXiv:2607.15957v3 Announce Type: cross Abstract: Large Language Models (LLMs) can generate natural language explanations that rationalize their own…
AI News Brief Hourly Summary 2026-09-11 11h : 12 posts
12 posts published in the last hour 08:34ConvMem: Convolutional Memory for Long-Context Reasoning 08:34From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric Reasoning 08:33Cyber-Financial Contagion: Modeling the Propagation of an AI Vendor Compromise Through the…
ConvMem: Convolutional Memory for Long-Context Reasoning
arXiv:2609.10441v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated impressive capabilities, they often struggle with…
From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric Reasoning
arXiv:2609.10335v1 Announce Type: new Abstract: Plane geometry remains a significant challenge in AI, requiring the integration of visual perception and…
Cyber-Financial Contagion: Modeling the Propagation of an AI Vendor Compromise Through the Banking System
arXiv:2609.10350v1 Announce Type: new Abstract: The banking system now depends on a small set of shared artificial intelligence vendors for fraud…
Fortunate Recall: Ontology-Driven Memory Lifecycle Management for Persistent Coherence in LLMs
arXiv:2609.10413v1 Announce Type: new Abstract: Current LLM memory systems treat all personal facts identically, so stores grow without bound while…
OpenAI’s new Agents API gives developers the infrastructure behind Codex and ChatGPT
OpenAI is releasing the Agents API as a public beta. It lets developers build cloud agents that run autonomously for hours, execute code, and hand off…
TRACE: Training Reasoning Agents for Causal Exploration with Synthesized Rewards
arXiv:2609.10315v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has advanced language-model reasoning in domains…
Beyond Surface Imitation: Contrastive Modeling for Reasoning Path Alignment in Multimodal In-Context Learning
arXiv:2609.10177v1 Announce Type: new Abstract: In-context learning (ICL) is widely used in multimodal large language models (MLLMs) and achieves strong…
Why Sample What You Can Enumerate? Exact Policy Optimization for Genomic Tool Selection
arXiv:2609.10221v2 Announce Type: new Abstract: Reinforcement learning over a frozen reasoner has become a common recipe for teaching a policy which…
Agent-Based ML-LLM Fusion with Self-Optimizing Prompts for Plateau Weather Alerts
arXiv:2609.10135v1 Announce Type: new Abstract: To address insufficient contextualization, weak generalization, and poor scenario adaptation in tourism…
Kernel-Managed Shared Memory for System-Wide Personalization
arXiv:2609.10144v1 Announce Type: new Abstract: AI systems become more useful when they can adapt to the people using them, but in multi-agent systems,…
