arXiv:2608.26159v1 Announce Type: cross Abstract: Self-Generated Text Recognition (SGTR)–the ability of an LLM to identify its own outputs–poses risks…
Artificial Intelligence Models Can Predict and Collaboratively Modulate Human Memory Search
arXiv:2608.26152v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit unprecedented natural language generation and many text-based…
Syntax vs. Semantics: How Transformers Learn Deep Dependencies
arXiv:2608.26139v1 Announce Type: cross Abstract: Large Language Models demonstrate remarkable syntactic fluency, yet the optimization dynamics governing…
Beyond Accuracy: A Qualitative Analysis of Vision-Language Models for Hate Speech Detection in Memes
arXiv:2608.26143v1 Announce Type: cross Abstract: Memes have turned out to be a powerful tool through which individuals share their ideas concerning…
Position Is All You Need: A Free Lunch Token Compression Strategy for MLLM-based Referring Expression Segmentation
arXiv:2608.26142v1 Announce Type: cross Abstract: Referring Expression Segmentation (RES) aims to generate pixel-wise segmentation masks from complex and…
Evaluating AI Generated Summaries for Cancer Patients
arXiv:2608.26154v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly being integrated into digital health platforms to generate…
AI News Brief Hourly Summary 2026-08-28 16h : 13 posts
13 posts published in the last hour 13:33FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes 13:33WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution 13:33Training-Time Explainability for Multilingual Hate Speech Detection: Aligning Model Reasoning with…
FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes
arXiv:2608.26129v1 Announce Type: cross Abstract: Scientific peer review datasets have trained AI systems exclusively on Computer Science and Machine…
WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution
arXiv:2608.27454v1 Announce Type: new Abstract: Agent skills package specialized knowledge and workflows into reusable resources that extend AI agent…
Training-Time Explainability for Multilingual Hate Speech Detection: Aligning Model Reasoning with Human Rationales
arXiv:2608.26125v1 Announce Type: cross Abstract: Online hate against Muslim communities often appears in culturally coded, multilingual forms that evade…
Exploring the Role of LLMs in HPC Programming: A Survey
arXiv:2608.26110v1 Announce Type: cross Abstract: Large Language Models (LLMs) are emerging as promising assistants in High-Performance Computing (HPC),…
AI benchmarks have a trust problem and Google wants to fix it
Google Deepmind is testing a double-blind evaluation of a frontier AI model for the first time. Cryptographic protection through Confidential Space is…
From SQL to Knowledge Graphs: An LLM-Driven Multi-Agent Approach with Data Schema Improvement
arXiv:2608.26117v1 Announce Type: cross Abstract: RDBMS (Relational Database Management System) databases face several limitations, including slow…
Sophistication in GenAI Use: Field Evidence from a Large Firm
arXiv:2608.27364v1 Announce Type: new Abstract: We study how sophistication in generative AI (genAI) use varies among the back-office workforce of a large…
Not All Eval-Awareness Is Equal: Capabilities Framing Predicts Compliance
arXiv:2608.27340v1 Announce Type: new Abstract: Steering interventions targeting eval-awareness, a model’s recognition that it is being tested, are…
Mechanistic Reaction Prediction via Discrete Flow Matching on Graph-Structured Electron Occupation
arXiv:2608.27429v1 Announce Type: new Abstract: Chemical reactions are fundamentally transformations in electron space, yet most machine learning…
CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowledge Bases
arXiv:2608.27391v1 Announce Type: new Abstract: LLMs are increasingly able to answer complex questions about enterprise-scale document collections. But…
Anthropic gets its first court win over the Pentagon’s supply chain risk label
A federal judge ruled the Trump administration illegally labeled Anthropic a supply chain risk, handing the AI company a victory as its second Pentagon…
