arXiv:2609.26835v1 Announce Type: cross Abstract: COBOL remains widely deployed, yet representative corpora reflecting real production code are rarely…
Author: script
COPE: Continual Personalization of LLMs under Sparse User Feedback via User Embeddings and Self-Evaluation
arXiv:2609.26853v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved remarkable results across various benchmarks, their…
What Makes a Terminal-Bench Task Hard? Separating Genuine Hardness from Fake-Hardness on an Adjudicated Agentic Corpus
arXiv:2609.26826v1 Announce Type: cross Abstract: Frontier benchmarks need tasks that current models cannot solve. But a task that no model solves is not…
G\”odel’s and Scott’s Variants of the Ontological Argument in Lean 4
arXiv:2609.26806v1 Announce Type: cross Abstract: This paper presents a complete, structure-preserving port to Lean 4 of the Isabelle/HOL dataset…
Learning Stiffness Dependent Fluid Structure Dynamics from Coarse Flow Representations
arXiv:2609.26816v1 Announce Type: cross Abstract: This paper develops a data-driven framework for long-term prediction of fluid–structure interaction…
Bridging LLM Serving and CXL-SSDs with Chunk-Aware KV Cache Management
arXiv:2609.26828v1 Announce Type: cross Abstract: NAND-backed storage offers the capacity needed to scale LLM prefix caching, but its block I/O path…
Signal2Symbol: Neuro-Symbolic Temporal Reasoning for Explainable Physiological Time-Series Anomaly Detection
arXiv:2609.26820v1 Announce Type: cross Abstract: Physiological time series such as electrocardiograms (ECG) and electroencephalograms (EEG) exhibit…
AI News Brief Hourly Summary 2026-09-24 12h : 11 posts
11 posts published in the last hour 09:32StudentBench: AI and human tutoring yield equivalent GRE learning gains 09:32Shutdown Sabotage Propensities in Multi-Agent Systems 09:32Attention-based representations for multi-task computation 09:32An Open Pipeline and Dashboard for Systemic-Risk Evidence under the EU AI…
StudentBench: AI and human tutoring yield equivalent GRE learning gains
arXiv:2609.28470v1 Announce Type: new Abstract: Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at…
Shutdown Sabotage Propensities in Multi-Agent Systems
arXiv:2609.28274v1 Announce Type: new Abstract: The final safeguard against rogue AI behavior is the human ability to shut systems down. It has been…
