arXiv:2608.27391v1 Announce Type: new Abstract: LLMs are increasingly able to answer complex questions about enterprise-scale document collections. But…
Category: AI
Anthropic gets its first court win over the Pentagon’s supply chain risk label
A federal judge ruled the Trump administration illegally labeled Anthropic a supply chain risk, handing the AI company a victory as its second Pentagon…
Learning a Continuous Sepsis Severity Score Without Hour-by-Hour Supervision: A Two-Site Retrospective Study
arXiv:2608.27421v1 Announce Type: new Abstract: Currently used sepsis severity indices rely on fixed variables and weights established decades ago, which…
What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents
arXiv:2608.27260v1 Announce Type: new Abstract: LLM agents increasingly rely on generated interaction data to learn how to interact with external…
LLMs Can Design Near-Optimal OR Algorithms
arXiv:2608.27296v1 Announce Type: new Abstract: We ask whether large language models (LLMs) can design effective algorithms for well-specified operations…
Meta executive leaves for OpenAI as the social media giant faces growing scrutiny in India
Sandhya Devanathan will oversee some OpenAI operations across Southeast Asia and Australia in her new role.
Verify Smarter, Evolve Further: Efficient Harness Evolution through Behavior-Aware Verification
arXiv:2608.27311v1 Announce Type: new Abstract: Agent harnesses shape how language-model agents use instructions, tools, and runtime components, but…
Quantization and Pruning Methods to Make Your LLM Leaner
This article walks through what each technique actually does, why skipping them costs real money and real latency, and then gets hands-on with five…
BrailleBench: Investigating Multi-Criteria Braille Comprehension in Large Language Models
arXiv:2608.27268v1 Announce Type: new Abstract: Although Large language models (LLMs) mediate access to knowledge and computational assistance, their…
Agents Are Always Day-One Hires. It’s Time We Designed for It.
By 2027, 74% of companies are expected to use agents in some capacity, according to a recent Deloitte study. For years, we’ve designed and built software…
