arXiv:2608.08220v1 Announce Type: new Abstract: The overlapping disciplines of machine ethics and value alignment are concerned with designing artificial…
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
arXiv:2608.08210v1 Announce Type: new Abstract: Collaborative dialogue can end with apparent agreement while participants still differ on goals,…
LatticeMind: A Conflict-Aware Memory Primitive for Multi-Agent Systems
arXiv:2608.08236v1 Announce Type: new Abstract: Multi-agent LLM systems often fail not for lack of candidate answers, but because they have no persistent…
Anthropic says it will watermark text generated by its AI models
Anthropic will extend support for watermarking AI generations for older models as well.
A Fair Objective for Human-Empowerment-Preserving AI: Desiderata, Design, and Likely Behavioral Consequences
arXiv:2608.08240v1 Announce Type: new Abstract: This paper explores the idea of promoting well-being and safety in human-AI interactions by forcing AI…
AI’s Memory Problem: Why Efficient Long Context Changes Everything, and What It Takes to Get It Right
A colleague who forgot every prior conversation the moment it ended would not last long in most jobs. Yet a lot of business AI works exactly that way. The…
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment
arXiv:2608.08212v1 Announce Type: new Abstract: In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter…
A Minimal $\kappa$–$\tau$ Logic for Risk-Sensitive Abduction
arXiv:2608.08192v1 Announce Type: new Abstract: Standard approaches to abductive reasoning can retain multiple candidate explanations, but they do not…
Janus: An Algorithm-Evaluator Co-Evolution Framework for LLM-Driven Discovery under Expensive Evaluation Budgets
arXiv:2608.08189v1 Announce Type: new Abstract: LLM-driven program discovery relies on rapid evaluator feedback, but many scientific and engineering tasks…
3 Visual Proofs of the Central Limit Theorem to Build Your Intuition
To build your intuition, this article shows three visual proofs that the classic bell curve appears in myriad situations.
Quantization Degradation in Large Language Models: A Signal-Noise Perspective
arXiv:2608.08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a…
Anthropic signs $9.1 billion data center deal with Bitcoin miner Riot Platforms
Anthropic is leasing $9.1 billion worth of data center capacity from Bitcoin miner Riot Platforms in Texas, according to Bloomberg. The deal covers 191…
Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language Models
arXiv:2608.08199v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly involved in group decision-making with other LLMs and…
OpenAI introduces $125 Premium Seats for ChatGPT Business as agentic AI burns through more tokens
OpenAI is rolling out “Premium Seats” for ChatGPT Business customers at $125 per user per month, five times the price of the existing Standard Seats. In…
Large Multimodal Agents for Intelligent Transportation Systems: Architectures, Evidence, and Deployment Challenges
arXiv:2608.08184v1 Announce Type: new Abstract: Large multimodal agents (LMAs) are increasingly proposed for intelligent transportation systems (ITS), but…
AI News Brief Hourly Summary 2026-08-11 14h : 13 posts
13 posts were published in the last hour 11:33 : When Is a Steerable Concept Representation Real? Measurement Confounds in a Cross-Family Audit of Neuroscience Parallels in LLMs 11:33 : A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning…
When Is a Steerable Concept Representation Real? Measurement Confounds in a Cross-Family Audit of Neuroscience Parallels in LLMs
arXiv:2608.08159v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly reported to exhibit human-like neural and cognitive…
A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning
arXiv:2608.08158v1 Announce Type: new Abstract: Sparse, delayed, and weakly informative rewards remain central obstacles to efficient reinforcement…
