14 posts published in the last hour 05:32Text, Pixels, or Both? Evaluating Input Representations for Multimodal Document QA 05:32Generative Embodied Multiple Behavior Control Systems for Human-like Agents 05:32Toward Auditable and Calibrated AI for Dementia-Related Crash Severity Prediction: A Selective Deferral…
Author: script
Text, Pixels, or Both? Evaluating Input Representations for Multimodal Document QA
arXiv:2609.22628v1 Announce Type: new Abstract: Every document QA system begins with a choice that is rarely studied on its own: whether to feed the model…
Generative Embodied Multiple Behavior Control Systems for Human-like Agents
arXiv:2609.22691v1 Announce Type: new Abstract: An enduring and richly elaborated dichotomy in cognitive neuroscience is that of human behavior control…
Toward Auditable and Calibrated AI for Dementia-Related Crash Severity Prediction: A Selective Deferral Framework to Support Human Review
arXiv:2609.22694v1 Announce Type: new Abstract: Public crash databases increasingly support automated safety analysis, but crash severity prediction…
OpenAI Releases GPT-6 Sol and Luna: 50% Cheaper API Pricing and Benchmarks
OpenAI has released GPT-6 Sol and GPT-6 Luna, 2 lower-cost models trained with methods similar to GPT-6 Astra. Sol costs $2/$10 and Luna $0.10/$0.50 per…
Self-Organizing Agent Teams Learn to Reason Together
arXiv:2609.22682v1 Announce Type: new Abstract: Collective intelligence depends not only on what team members know, but also on how they organize their…
‘We’re already fighting yesterday’s battle’: Greece’s prime minister gets candid about AI
Most leaders on a trade mission stick to the pitch, but when I interviewed Greek Prime Minister Kyriakos Mitsotakis this week, he also admitted that no…
A Survey on the Linear Representation Hypothesis
arXiv:2609.22695v1 Announce Type: new Abstract: The term “linear representation hypothesis” (LRH) has appeared across diverse subfields of artificial…
AutoGym: Blueprint-First Generation of Verifiable Agent Gyms
arXiv:2609.22592v1 Announce Type: new Abstract: Training agents with reinforcement learning requires a gym, comprising a task, an executable environment…
Splitting Documents at Lower Cost: Multi-Split Boundary Decisions for LLM-Based Page Stream Segmentation
arXiv:2609.22620v1 Announce Type: new Abstract: Scanned mail, uploaded PDFs, and consolidated attachments often arrive as page streams that must be split…
