arXiv:2609.20732v1 Announce Type: new Abstract: Semantic cell annotation improves chunking interpretability for spreadsheets in LLM-driven RAG systems,…
Optimal Transport Metric Learning for Feature Alignment in Partially Supervised Segmentation
arXiv:2609.19176v1 Announce Type: cross Abstract: Multi-organ segmentation is often challenged by partially annotated datasets and domain shifts across…
Deep Noir: Autonomous Steering Discovery via Architectural Chronometry in Transformer Models
arXiv:2609.20722v1 Announce Type: new Abstract: Activation steering modifies LLM behavior at inference time, but identifying where and how strongly to…
Decades-Old Anonymized Medical Data May Cause AI Misdiagnoses Now
According to a new research collaboration between Germany and the UK, anonymized patient records that were included in medical datasets even decades ago…
RAFT: A Stateful Retrieval-Augmented Framework for Troubleshooting Agents
arXiv:2609.20754v1 Announce Type: new Abstract: Effective troubleshooting agents in enterprise customer support depend on retrieving actionable guidance…
Snap tries to make the case again for its $2,200 smart glasses
Since Specs’ debut earlier this year, Snap has clearly been looking for an opportunity to explain why the smart glasses deserve to exist.
An Empirical Study of Harness Design for Coding Agents
arXiv:2609.20804v1 Announce Type: new Abstract: Coding harnesses shape how autonomous coding agents translate model capabilities into long-horizon…
Ownership in AI-Assisted Everyday Tasks
arXiv:2609.20658v1 Announce Type: new Abstract: When does work done with AI still feel like ours? As AI becomes woven into everyday tasks, we must examine…
5 Prompt Optimization Strategies That Actually Improve LLM Output
This article covers five prompt optimization strategies such as: prompt optimization, prompt engineering, LLM output quality, few-shot prompting,…
Language-model groups overstate consensus when replaying human deliberation on a reasoning task
arXiv:2609.20543v1 Announce Type: new Abstract: Full-consensus rates are often treated as indicators of collective cognition, yet depend on how…
Could AI really kill us all? Your questions, answered.
On Wednesday, MIT Technology Review hosted a live Roundtables event for subscribers that asked the question everyone’s asking right now: Could AI really…
PAA: The Probabilistic Allen Algebra: A Generative and Complete Probabilistic Extension of Allen’s Interval Relations
arXiv:2609.20634v1 Announce Type: new Abstract: Allen’s interval algebra is a qualitative calculus for temporal relations, but its thirteen base relations…
42 leading mathematicians warn that AI existential risk is real and urgent
42 Fellows of the Royal Society, including Fields Medal winners Martin Hairer and Peter Scholze, warn of existential AI risks in an open letter. Leading…
Limits of Confidence in Diffusion
arXiv:2609.20581v1 Announce Type: new Abstract: Discrete diffusion, including remasking and uniform-state samplers, generate a sequence by writing…
SK Hynix Debuts Ventures CVC Brand at Inaugural Silicon Valley Event
SK hynix announced the launch of SK hynix Ventures on September 18, 2026, a corporate venture capital brand that the South Korean memory maker said is…
Refuse, Decompose, Refresh: A Claim-Safe Protocol for Closed-Loop AI Evaluation
arXiv:2609.20538v1 Announce Type: new Abstract: An AI evaluation can be perfectly reproducible and still support the wrong claim. This risk is acute in…
AI News Brief Hourly Summary 2026-09-18 14h : 14 posts
14 posts published in the last hour 11:33The Organization of Inference: Information, Resource Constraints, and AI Production 11:33SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness 11:33How Do Agent Harnesses Create Value? Planning Information and Release Control in Stateful LLM…
The Organization of Inference: Information, Resource Constraints, and AI Production
arXiv:2609.20449v1 Announce Type: new Abstract: The economic value of inference depends on how capacity and task information are distributed across stages…
