arXiv:2609.29773v1 Announce Type: new Abstract: Many real-world tasks (e.g., office workflows, scientific experimentation) require LLM agents to interact…
Category: AI
The Gold in Bias: Maturing the AI Design Process through Verification
arXiv:2609.29730v1 Announce Type: new Abstract: Bias in AI systems is typically framed as a flaw to be minimized, yet it also serves as a critical…
A General Framework for Budgeted Threshold Incentives on Request
arXiv:2609.29724v1 Announce Type: new Abstract: On-demand delivery platforms pay riders through incentive activities whose tiers are set from recent…
C3M: Cross-Session Multimodal Memory Maintenance for Long-Horizon Tasks
arXiv:2609.29735v1 Announce Type: new Abstract: Long-horizon tasks require preserving and later recovering cross-session evidence under a bounded,…
Fair Like Us? Auditing LLM Alignment in Resource Allocation
arXiv:2609.29692v1 Announce Type: new Abstract: Fair allocation of scarce, indivisible resources is an important challenge in many societal problems.…
Anthropic says its biology lab has already found something big
But maybe the biggest reveal is that Anthropic has not let Claude run loose in its biology lab. Humans are still, so far, in the loop.
To Think or Not to Think: Allocating Reasoning Where It Helps
arXiv:2609.29664v1 Announce Type: new Abstract: Reinforcement learning (RL) has proven effective in enhancing the reasoning performance of large language…
Sequential knowledge editing breaks a model’s ability to tell good evidence from bad, without costing it accuracy
arXiv:2609.29587v1 Announce Type: new Abstract: Knowledge editing is evaluated on whether the edited fact changed, whether paraphrases follow, and whether…
PEEL: Physics-Enabled Evidential Learning for Identifiable Uncertainty in CT Imaging
arXiv:2609.29599v1 Announce Type: new Abstract: Normal-inverse-gamma (NIG) regression is not uniquely identifiable from its marginal Student-t likelihood:…
Ingest-Time Fact Compilation for Cost-Efficient and Reliable Question Answering over Revised Corpora
arXiv:2609.29661v1 Announce Type: new Abstract: Most agentic question answering (QA) systems do an important part of their semantic work at the worst…
