arXiv:2609.01924v1 Announce Type: new Abstract: Recent work identifies a mid-depth band of verbalisable, causally potent representations in a standard…
When Agents Implement Systems: A Case Study in Defects, Detection, and Evaluation Rigor
arXiv:2609.01985v1 Announce Type: new Abstract: As LLM coding agents increasingly perform end-to-end engineering work, we lack empirical characterization…
Benchmarking Language Models for Statistical Problem Formulation
arXiv:2609.01982v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as assistants for statistical and data science work,…
The Ceiling Is in the Channel: Auditing Learner Gaps and Measurement Frontiers in Clinical Prediction
arXiv:2609.01909v1 Announce Type: new Abstract: Clinical prediction can saturate for two different reasons: a fitted learner may fail to extract available…
Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment
arXiv:2609.01962v1 Announce Type: new Abstract: Ultra-low-bit language models can reduce storage and memory bandwidth, but a nominal “1.58-bit” label does…
AI News Brief Hourly Summary 2026-09-03 07h : 13 posts
13 posts published in the last hour 04:32Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence 04:32Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization 04:32SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval 04:32Architecting Conversational…
Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence
arXiv:2609.01873v1 Announce Type: new Abstract: Multi-agent AI systems improve inference by spawning agents and synthesizing reports. But another agent is…
Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization
arXiv:2609.01861v1 Announce Type: new Abstract: The performance of an LLM agent depends on the scaffold around a frozen model. A common way to improve…
SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval
arXiv:2609.01849v1 Announce Type: new Abstract: This article presents SSAKG 2.0, an open-source software package for constructing and operating Structural…
Architecting Conversational Data Systems for Stateless LLM APIs: The Hydration Proxy Pattern
arXiv:2609.01834v1 Announce Type: new Abstract: As enterprise platforms transition to conversational reasoning interfaces, the stateless nature of LLM…
Introducing Claude Fable 5.1 on AWS
Claude Fable 5.1 is now available on Amazon Bedrock and Claude Platform on AWS. This post covers Claude Fable 5.1’s improvements, the Enterprise Frontier…
The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory Agents
arXiv:2609.01852v1 Announce Type: new Abstract: Persistent memory supports personalized agents, but a stale stored fact can override current authoritative…
When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic
arXiv:2609.01741v1 Announce Type: new Abstract: Statutes are increasingly parsed by machines before people read them, and the parsers disagree: on…
Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI
arXiv:2609.01685v1 Announce Type: new Abstract: With the development of artificial intelligence (AI), the landscape of meta-ethics, which has largely…
When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection
arXiv:2609.01814v1 Announce Type: new Abstract: Information sharing can improve a pooled estimate while eliminating independent rescue actions. This paper…
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models
arXiv:2609.01611v1 Announce Type: new Abstract: Frontier large language models can often recognize when they are being evaluated, a capability known as…
JapanFold Keeps Open-Source Drug Discovery Computations Inside Japan
ai& and Tenstorrent launched JapanFold on September 3, 2026, a drug discovery platform serving open-source structural biology models entirely on…
Induction and Inquiry via Probabilistic Reasoning over Language and Code
arXiv:2609.01815v1 Announce Type: new Abstract: How humans grow and maintain abstract knowledge from the sparse, streaming noisy data of experience is a…
