Google began rolling out Google Pics, its AI image creation and editing tool built on the company’s Nano Banana model, to Google AI Pro and Ultra…
Category: AI
SkillZip Pro: Execution-Aware Dynamic Compression of Progressively Loaded Skills for Self-Evolving Agents
arXiv:2608.30785v1 Announce Type: new Abstract: Production agent skills are directory bundles, not isolated prompts. The root is loaded at activation;…
MedAgent-R1: Faithfulness-Aware Reinforcement Learning for Evidence-Grounded Medical Reasoning
arXiv:2608.30676v1 Announce Type: new Abstract: When medical AI systems hallucinate clinical reasoning, the consequences extend beyond incorrect answers:…
PyKEEN-NSX: A Modular Framework for Static, Dynamic and Schema-Aware Negative Sampling in PyKEEN
arXiv:2608.30652v1 Announce Type: new Abstract: Embedding methods have become popular due to their scalability on link prediction and/or triple…
ATLAS: Dual-Horizon Diagnostic Evaluation for Industrial Tool-Use Agents
arXiv:2608.30685v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed in user-facing services that require iterative…
Introducing Claude Fable 5.1 on AWS
Claude Fable 5.1 is now available on Amazon Bedrock and Claude Platform on AWS. This post covers Claude Fable 5.1’s improvements, the Enterprise Frontier…
HiRS-Agent: A Hierarchical Multi-Agent System for Reliable Long-Horizon Remote Sensing Task Solving
arXiv:2608.30672v1 Announce Type: new Abstract: Recent advances in large language models and multimodal models have pushed remote sensing (RS) processing…
You.com Web Search Highlights Reaches 95.17% on SimpleQA
You.com on September 1, 2026 introduced a new extraction mode for its Web Search API called Highlights, reporting a 95.17% score on the SimpleQA benchmark…
Geometry of Divergence: Tracking Hidden-State Trajectories for Adaptive Multi-Turn Reasoning
arXiv:2608.30650v1 Announce Type: new Abstract: LLM agents need to sustain goal-consistent reasoning across long multi-turn interactions under strict…
Automated Testing of LLM-Based Post Hoc Explainers Using Model Checking as an Oracle
arXiv:2608.30581v1 Announce Type: new Abstract: Large language models (LLMs) are used as post hoc explainers of sequential decision-making policies,…
