Chinese state-backed hacking groups have more than doubled their attacks since they started using AI models like DeepSeek to write exploit code and scan…
Repo2Skill-Evo: Repository Skills Go Stale in Silence
arXiv:2608.21964v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate over evolving software repositories, where success…
AI News Brief Hourly Summary 2026-08-25 11h : 12 posts
12 posts published in the last hour 08:32ESCRAG-R1: Retrieval-Augmented Reinforcement Learning for Emotional Support Conversation 08:32GuardianBench: A Same-Scene Instruction-Contrastive Benchmark for Latent Contextual Risk in Embodied AI 08:32Training Needs Trustworthy Worlds: Verified Synthetic Web Environments for Agent Learning 08:32Consistency Is…
ESCRAG-R1: Retrieval-Augmented Reinforcement Learning for Emotional Support Conversation
arXiv:2608.21925v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) systems aim to provide holistic support by balancing professional…
GuardianBench: A Same-Scene Instruction-Contrastive Benchmark for Latent Contextual Risk in Embodied AI
arXiv:2608.21928v1 Announce Type: new Abstract: In embodied AI, safety risk can be latent: a benign instruction and a safe scene become hazardous only…
Training Needs Trustworthy Worlds: Verified Synthetic Web Environments for Agent Learning
arXiv:2608.21898v1 Announce Type: new Abstract: Web agents promise to automate complex digital workflows, but their training remains limited by synthetic…
Consistency Is Not Coherence: Orientation Search for Certified Alignments Between 4D Defence Upper Ontologies
arXiv:2608.21914v1 Announce Type: new Abstract: We align three upper ontologies that sit under UK and NATO defence data infrastructure: the Information…
Sentante’s Endovascular Robot Enters Commercial Use in Vascular Surgery
Sentante has begun commercial deployment of its CE-marked endovascular robotic platform, the medical robotics company announced on August 20, 2026, with…
Multimodal Prompt Learning with Irregular EHRs for Robust Monitoring of Critical Care Patients
arXiv:2608.21941v1 Announce Type: new Abstract: Accurate assessment of patients in intensive care units (ICUs) is essential for timely clinical…
MemGuard: Persisting Verifier Signals for LLM-Agent Memory Governance
arXiv:2608.21867v1 Announce Type: new Abstract: LLM agents are moving from single-prompt use to long task streams in which reusable memory becomes a core…
AI Watchdog: Agent Interfaces for Detecting and Defending Against Manipulative Dark Patterns in AI Conversations
arXiv:2608.21841v1 Announce Type: new Abstract: Conversational AI increasingly shapes consequential decisions, yet users have limited support for…
HiMA-MDD: A Hierarchical Multi-Agent Harness for Interpretable Multimodal Depression Detection in Clinical Interviews
arXiv:2608.21868v1 Announce Type: new Abstract: Depression assessment from multimodal clinical interviews requires integrating dispersed evidence from…
From Solver Feedback to Faithful Plans: Multi-Role Reinforcement Learning for Symbolic Planning
arXiv:2608.21897v1 Announce Type: new Abstract: Reliable planning requires converting natural-language instructions into executable symbolic…
LLM4LLM: Bridging Kernel Benchmarks and Real Deployment via Closed-Loop Agentic Optimization
arXiv:2608.21836v1 Announce Type: new Abstract: Large language models have become increasingly capable agents for low-level code and kernel optimization,…
AI News Brief Hourly Summary 2026-08-25 10h : 11 posts
11 posts published in the last hour 07:33VisAdj: Learning Adjacency Matrices from Node-Link Images 07:33GameXpert-Bench: How Far Are Coding Agents from Expert Game Development? 07:33Beyond Success and Failure: Length-Aware Contrastive Learning for GUI Agents 07:32HIRA: A Human-in-the-Loop Retrieval-Augmented Cascade for…
VisAdj: Learning Adjacency Matrices from Node-Link Images
arXiv:2608.21825v1 Announce Type: new Abstract: Learning adjacency matrices from node-link images is a fundamental problem for recovering structured graph…
GameXpert-Bench: How Far Are Coding Agents from Expert Game Development?
arXiv:2608.21833v1 Announce Type: new Abstract: Recent large language models (LLMs) can operate as coding agents that build complete games from natural…
Beyond Success and Failure: Length-Aware Contrastive Learning for GUI Agents
arXiv:2608.21830v1 Announce Type: new Abstract: Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) have shown…
