arXiv:2608.14107v1 Announce Type: new Abstract: Reasoning-intensive retrieval requires text representations to capture not only semantic similarity, but…
Category: AI
A Pathway to General-Purpose Scientific AI: Multimodal Comprehension of Scientific Images
arXiv:2608.14075v1 Announce Type: new Abstract: Scientific figures and tables encode essential experimental evidence, yet remain difficult for digital…
Scaling Domain Data Repetition in LLM Pretraining
arXiv:2608.14071v1 Announce Type: new Abstract: As large language models scale, their training-token budgets must also increase to maintain an appropriate…
Mandato: Protocol-Level Enforcement of Digitally Signed Mandates on AI Agent Actions with Cryptographically Chained Audit Trails
arXiv:2608.14074v1 Announce Type: new Abstract: AI agents increasingly act on external systems through standardized tool-calling protocols such as the…
Get closer to the game with Gemini and Pixel
Low-angle view of a soccer player kicking a ball mid-air against a bright blue sky, with grass flying from their cleats.
Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers
arXiv:2608.14089v1 Announce Type: new Abstract: Safety classifiers deployed with large language models often fail for two reasons: their decisions reflect…
Agent-Orchestration in Autonomous Chip Design
arXiv:2608.14035v1 Announce Type: new Abstract: Recent developments in large language models (LLMs) and tool-using agents encourage people to explore the…
Buy the Rumor, Sell the News: When Is News Priced In?
arXiv:2608.14014v1 Announce Type: new Abstract: Two old market sayings hold that news is already priced in by the time it is published, and that the rumor…
Benchmarking data-driven material models on the classic Treloar dataset
arXiv:2608.14063v1 Announce Type: new Abstract: Machine learning is rapidly reshaping constitutive modeling, offers new ways to learn material behavior…
Demystifying Agent Skills: Why They Work-Until They Don’t
arXiv:2608.14036v1 Announce Type: new Abstract: Skills have emerged as a practical and effective approach for enhancing LLM agents at inference time…
