arXiv:2609.11319v1 Announce Type: new Abstract: Most of mathematical knowledge has been communicated through so-called informal use of mathematics and…
Category: cs.AI updates on arXiv.org
Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents
arXiv:2609.11318v1 Announce Type: new Abstract: Deep research agents are increasingly capable of web search, tool use, multimodal evidence analysis, and…
AI Exposure and AI Resilience: A Two-Dimensional Assessment Framework for Software and Software-Based Business Model
arXiv:2609.11321v1 Announce Type: new Abstract: Artificial intelligence is changing both software production and the economics of software-based business…
Routing by Reasoning Need: Trajectory-Aware Decoding Control for Diffusion Vision-Language Models
arXiv:2609.11315v1 Announce Type: new Abstract: Diffusion vision-language models generate answers through iterative refinement, exposing intermediate…
Bio-inspired Learning and Decision-Making with Probabilistic In-Memory Computing Hardware: Part 1
arXiv:2609.11281v1 Announce Type: new Abstract: Learning and decision-making in animals are often modeled as Bayesian processes, where sensory evidence is…
Off-Target Effects of Response-Style Alignment in a Korean 27B Language Model
arXiv:2609.11291v1 Announce Type: new Abstract: We post-train Qwen3.8-27B for Korean response style — verbosity, list and markdown usage, discourse…
When Does Text Inform? Benchmarking Information-Theoretic Metrics for Multimodal Time-Series Forecasting
arXiv:2609.11282v1 Announce Type: new Abstract: Multimodal forecasting models that combine time series with text annotations promise richer prediction…
Memory Compression for High-Fanout Agent Sandboxes
arXiv:2609.11294v1 Announce Type: new Abstract: High-fanout agent workloads create a growing memory bottleneck because a single task may spawn many…
Generating a Consistent Enterprise: Synthesis and Reference-Free Evaluation of Multi-System Business Data
arXiv:2609.11286v1 Announce Type: new Abstract: Synthetic relational data is normally produced by a model trained on a real dataset, and its quality is…
Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Agents
arXiv:2609.11243v1 Announce Type: new Abstract: Autonomous research agents are increasingly expected to search the literature, analyze experimental…
