AI is hitting a fixed benchmark performance level at a rapidly falling cost. Epoch AI measures a price decline of about 13x per year. After stripping out…
Category: AI
Quantum Reinforcement Learning for Cost and Delay Tradeoffs in Quantum Cloud Orchestration
arXiv:2609.27446v1 Announce Type: cross Abstract: Quantum cloud computing, delivered through the quantum-as-a-service (QaaS) model, provides access to…
Right-size generative AI endpoints with concurrency sweeps on Amazon SageMaker AI
Concurrency sweeps help you right-size a generative AI endpoint on Amazon SageMaker AI by systematically benchmarking it at increasing load levels. This…
Issuer-Sovereign Agentic Payments
arXiv:2609.27452v1 Announce Type: cross Abstract: AI agents are beginning to make real payments. Current approaches let an agent pay by relying on a…
Claude Opus 5.5 is now available on AWS
Claude Opus 5.5, Anthropic’s most capable Opus model for agentic coding, knowledge work, and long-running tasks, is now available on Amazon Bedrock and…
Forget who you Forgot: Speaker Unlearning to Prevent Re-Identification in Zero-Shot Text-to-Speech
arXiv:2609.27399v1 Announce Type: cross Abstract: Recent zero-shot text-to-speech (ZS-TTS) systems can reproduce a speaker’s voice with high fidelity from…
Evaluate skill-equipped agents with Strands Evals and Amazon Bedrock AgentCore
Skills let you encode domain-specific procedures as reusable, portable instructions for agents, but a fluent answer doesn’t prove the agent picked the…
What Looks Like a Capability Limit in Vision-Language Models Is a Readout Limit
arXiv:2609.27408v1 Announce Type: cross Abstract: Benchmarks for vision-language models offer their answer choices in some convention: a letter, a color…
Bravely AI Browsing with Leo
Learn about private AI browsing for data professionals.
Forecast Workflow Bench: Evaluating Language-Model Decisions with Budgeted Forecast Tools
arXiv:2609.27385v1 Announce Type: cross Abstract: Time-series foundation models (TSFMs) provide forecasts for operational decisions, but accuracy alone…
