arXiv:2608.23834v1 Announce Type: new Abstract: The key-value (KV) cache is a primary capacity and bandwidth bottleneck in long-context LLM serving. We…
Tag: AI
AWS Backs Agentic Resource Discovery as Federation Layer for Agent Registry
Amazon Web Services has put its weight behind the Agentic Resource Discovery specification, publishing a detailed account on August 24, 2026 of how the…
SyPS: Measuring Sycophancy Prompt Sensitivity in Large Language Models
arXiv:2608.23837v1 Announce Type: new Abstract: Large language models (LLMs) are known to exhibit social sycophancy, often validating or agreeing with…
IBM Releases Granite 4.2: Bringing Native Reasoning and Agentic RL to Open Enterprise Models
IBM has released Granite 4.2, a family of open reasoning language models in 3B, 8B, and 30B sizes, all under Apache 2.0. Every model exposes a thinking /…
Exploit More, Explore Smarter for Budget-Constrained Agentic Search
arXiv:2608.23848v1 Announce Type: new Abstract: Budget-constrained agentic search arises when an LLM agent must refine candidates under a small evaluation…
Lancium and NVIDIA Partner to Deploy Gigawatt-Scale AI Factories
Lancium, the Texas energy-infrastructure company behind the Abilene campus that anchors the Stargate buildout, has signed a strategic collaboration with…
Granite.Trust Policy Tools: Shareable, Actionable Policies for Generative AI Applications
arXiv:2608.23870v1 Announce Type: new Abstract: When it comes to safety policies for generative AI, one size does not fit all. Each organization and use…
Oana Jinga, Co-Founder, Chief Commercial and Product Officer of Dexory – Interview Series
Oana Jinga, Co-Founder, Chief Commercial and Product Officer of Dexory, is an experienced technology and commercial leader with a background spanning…
In-Context Inpainting for Time Series Forecasting
arXiv:2608.23855v1 Announce Type: new Abstract: We propose ICI-Time, a novel framework that reframes time series forecasting as a visual inpainting task,…
Do LLMs Understand Limit Order Book Dynamics?
arXiv:2608.23706v1 Announce Type: new Abstract: A large language model (LLM) trained on synthetic limit order book (LOB) data achieves near perfect scores…
