arXiv:2609.28582v1 Announce Type: cross Abstract: The recent emergence of Time Series Foundation Models (TSFMs) has significantly advanced multi-step…
Tag: AI
Last 24 hours to save up to $200 on TechCrunch Disrupt 2026. Reason 5 of 5 to attend: Momentum
Last 24 hours to save up to $200 on your TechCrunch Disrupt 2026 pass. Leave the event further in your startup’s trajectory than where you started. Don’t…
Where Cyber Agents Struggle: Bottleneck Analysis of Multi-Stage LLM Agents
arXiv:2609.28572v1 Announce Type: cross Abstract: Multi-stage LLM-based cyber agents may complete attack workflows while remaining brittle, costly, or…
Anthropic’s founders seek voting control ahead of IPO
Anthropic is asking its shareholders to approve a structure that would give its seven co-founders a combined 50.1% of the vote on most corporate matters.
Persistent Billable State: Denial-of-Wallet Attacks and Defenses in Tool-Calling LLM Agents
arXiv:2609.28585v1 Announce Type: cross Abstract: Multi-step tool-calling LLM agents rely on host runtimes to preserve state across turns. When a runtime…
Speculative Evaluation of Stochastic LLMs
arXiv:2609.28560v1 Announce Type: cross Abstract: Evaluating a stochastic large language model is costly: benchmark scores estimate expected performance…
Certified Task-Conditioned Active Observability
arXiv:2609.28520v1 Announce Type: cross Abstract: Before acting upon an unobservable physical system, an autonomous agent must determine which latent…
SMILESGNN: Interpretable Clinical Toxicity Prediction via SMILES-Graph Cross-Attention Fusion
arXiv:2609.28553v1 Announce Type: cross Abstract: Drug toxicity prediction is critical for reducing late-stage attrition in drug discovery, yet remains…
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB
Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3, built to run inside…
Who Is Behind the Harness? Fingerprinting LLMs through Agentic Behavior
arXiv:2609.28559v1 Announce Type: cross Abstract: LLMs increasingly operate through coding-agent harnesses that inspect repositories, invoke tools, and…
