arXiv:2608.13581v1 Announce Type: cross Abstract: Personalized glucose regulation remains a central yet unresolved challenge in precision nutrition, as…
Tag: cs.AI updates on arXiv.org
Jais 2: A Family of Arabic-Centric Open Large Language Models
arXiv:2608.13580v1 Announce Type: cross Abstract: Jais 2 is a family of Arabic-centric large language models developed jointly by MBZUAI, Cerebras, and…
BCMT: Blockwise Causal Memory Transformer
arXiv:2608.13578v1 Announce Type: cross Abstract: Transformer architectures rely on dense self-attention to model long-range dependencies, but this…
Think in Latent, Explain in Language: Self-Explainable Latent Reasoning
arXiv:2608.13570v1 Announce Type: cross Abstract: Latent reasoning has emerged as a powerful alternative to text-based Chain-of-Thought (CoT), offering…
Not All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM Systems
arXiv:2608.13571v1 Announce Type: cross Abstract: When a language model fails to answer a query on the first attempt, an agentic system retries, consuming…
The Architect: Interactive Visualization of Deep Learning Mathematics Directly in Microsoft Excel
arXiv:2608.13572v1 Announce Type: cross Abstract: We present The Architect, a system that turns Microsoft Excel into an interactive view of deep learning…
Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study
arXiv:2608.13568v1 Announce Type: cross Abstract: Coding agents spend most of their context budget on retrieval. Lexical retrieval (grep) is universal,…
Don’t Claim Benchmark-Oriented Optimization Improves General Coding Capability — Diverse Evaluation Is Required
arXiv:2608.13566v1 Announce Type: cross Abstract: Post-training papers, model cards, and blog posts often treat scores on a small set of coding benchmarks…
Split the Labor: Separating Evidence Interpretation from Decision Aggregation
arXiv:2608.14509v1 Announce Type: new Abstract: Systems that ask a language model to reach a conclusion from many sources usually concatenate them into…
Twin: Playing an Unknown Game with a Test-Time Digital Twin
arXiv:2608.14490v1 Announce Type: new Abstract: We present a Test-time World-model Inference (Twin) system, in which a frontier coding agent writes an…
