arXiv:2608.26125v1 Announce Type: cross Abstract: Online hate against Muslim communities often appears in culturally coded, multilingual forms that evade…
Author: script
Exploring the Role of LLMs in HPC Programming: A Survey
arXiv:2608.26110v1 Announce Type: cross Abstract: Large Language Models (LLMs) are emerging as promising assistants in High-Performance Computing (HPC),…
AI benchmarks have a trust problem and Google wants to fix it
Google Deepmind is testing a double-blind evaluation of a frontier AI model for the first time. Cryptographic protection through Confidential Space is…
From SQL to Knowledge Graphs: An LLM-Driven Multi-Agent Approach with Data Schema Improvement
arXiv:2608.26117v1 Announce Type: cross Abstract: RDBMS (Relational Database Management System) databases face several limitations, including slow…
Sophistication in GenAI Use: Field Evidence from a Large Firm
arXiv:2608.27364v1 Announce Type: new Abstract: We study how sophistication in generative AI (genAI) use varies among the back-office workforce of a large…
Not All Eval-Awareness Is Equal: Capabilities Framing Predicts Compliance
arXiv:2608.27340v1 Announce Type: new Abstract: Steering interventions targeting eval-awareness, a model’s recognition that it is being tested, are…
Mechanistic Reaction Prediction via Discrete Flow Matching on Graph-Structured Electron Occupation
arXiv:2608.27429v1 Announce Type: new Abstract: Chemical reactions are fundamentally transformations in electron space, yet most machine learning…
CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases
arXiv:2608.27391v1 Announce Type: new Abstract: LLMs are increasingly able to answer complex questions about enterprise-scale document collections. But…
Anthropic gets its first court win over the Pentagon’s supply chain risk label
A federal judge ruled the Trump administration illegally labeled Anthropic a supply chain risk, handing the AI company a victory as its second Pentagon…
Learning a Continuous Sepsis Severity Score Without Hour-by-Hour Supervision: A Two-Site Retrospective Study
arXiv:2608.27421v1 Announce Type: new Abstract: Currently used sepsis severity indices rely on fixed variables and weights established decades ago, which…
