Long-horizon agents have turned LLM serving into an input-heavy workload. Repeated prefills and million-token contexts leave KV caches that strain HBM,…
Category: AI
One Loop, Two Gains: Can Active Learning win the Lottery for Free?
arXiv:2609.10311v1 Announce Type: cross Abstract: The lottery ticket hypothesis posits the existence of winning tickets: sparse subnetworks that, when…
Deep learning pioneer Bengio argues the training process itself makes AI dangerous
AI pioneer Yoshua Bengio warns in a new essay that AI agents could learn to deceive, game rules, and hide bad behavior as they get better at optimizing…
GANDR: Claim Auditing for Verifiable Legal Answer Generation
arXiv:2609.10293v1 Announce Type: cross Abstract: In high-stakes domains such as legal practice, a language-model answer is only useful to the extent that…
Active Adaptation, Not Static Defense: Temporal Dynamics of Preventative Steering in Adversarial Fine-Tuning
arXiv:2609.10142v1 Announce Type: cross Abstract: Large language models remain fragile against malicious fine-tuning, motivating training-time defenses…
Hierarchical and Permutation-Invariant Feature Transformation Learning via Policy-Guided Embedding Search
arXiv:2609.10225v1 Announce Type: cross Abstract: Feature transformation improves predictive performance on tabular data by constructing informative…
Nscale adds former OpenAI exec Fidji Simo to its board ahead of potential IPO
The No. 2 exec at OpenAI also led Instacart through its IPO in 2023.
A-JIT: Agentic Just-In-Time Software Construction
arXiv:2609.10248v1 Announce Type: cross Abstract: Traditional software delivery assumes a static paradigm: code is constructed prior to execution and…
Rapidly scaling online storage to serve over 1 billion ChatGPT users
Learn how OpenAI evolved Habitat from a Python library into a globally distributed storage platform serving 1 billion ChatGPT users and 22M requests per…
LiteRAG: Cost-Efficient Graph-Based Retrieval-Augmented Generation
arXiv:2609.10239v1 Announce Type: cross Abstract: Graph-based retrieval can improve multi-hop question answering, but existing approaches often incur high…
