arXiv:2608.21107v1 Announce Type: new Abstract: Large Language Models (LLMs) are moving from code completion toward repository-scale agents that retrieve…
Category: AI
Kids outlearn AI—and we still don’t know why
People have been talking to each other for at least 100,000 years, as best we can tell. And in all that time, there has been only one thing in the world…
ReFrame: Evidence-Guided Test-Time Safety Alignment in Multimodal Large Language Models
arXiv:2608.21100v1 Announce Type: new Abstract: While multimodal large language models (MLLMs) extend model capabilities beyond text, they also make…
Socialized Division and Collaboration: Rethinking Class-Incremental Learning under Optimization Conflicts
arXiv:2608.21044v1 Announce Type: new Abstract: Class-incremental learning is commonly instantiated as a single-model paradigm, where a unified model…
Don’t Solve, Just Compare: Tiny Advisors for Runtime Intervention in LLM Agents
arXiv:2608.21027v1 Announce Type: new Abstract: LLM agents are emerging as an important paradigm for real-world tasks that require reasoning, tool use,…
Belief Without Behavior: Measuring the Translation of Theory of Mind into Coordinated Social Action in Vision-Language Models
arXiv:2608.20975v1 Announce Type: new Abstract: Effective social interaction requires agents to translate mental state inferences into coordinated…
When AI Reads Between the Lines: OCR vs. VLMs
Can machines truly understand documents, or have they simply become more effective at extracting information from them? With traditional OCR, an error can…
Evaluating Large Language Model Performance on International Maritime Dangerous Goods Code Compliance
arXiv:2608.21036v1 Announce Type: new Abstract: The transport of dangerous goods by sea is a high-consequence activity governed by the International…
Cerebras unveils CS-4 with double the performance on the same chip
Cerebras has introduced its CS-4 AI accelerator, which CEO Andrew Feldman calls the fastest system in the industry. The article Cerebras unveils CS-4 with…
The Cost of a Physics Prior Is Bounded by the Ablation Gap
arXiv:2608.21059v1 Announce Type: new Abstract: Shape-constrained and physics-informed learning reports an accuracy cost of enforcing a prior and treats…
