arXiv:2608.29589v1 Announce Type: new Abstract: Text-to-image (T2I) safety guardrails fail to generalize equitably to non-standard dialects. Evaluating…
Category: AI
LLMs Interpret, Embeddings Organize, Graphs Emerge: Agent-Driven Compilation of Scientific Knowledge
arXiv:2608.29612v1 Announce Type: new Abstract: Sustained scientific work requires a knowledge substrate that carries interpretation across tasks and…
Cash in on the AI Boom by Renting Out Your Spare Compute
If you own an at-home server, a gaming computer, or just a laptop that doesn’t get much love, listen up. You can now put that spare computing power to use…
Call Neighbours Yourself: Graph Walks with Destination-Conditioned On-Policy Self-Distillation
arXiv:2608.29588v1 Announce Type: new Abstract: Reasoning over text-attributed graphs (TAGs) requires large language models (LLMs) to combine a node’s…
Google’s election AI Overviews are opaque, rely on few sources, and sometimes take sides
Using access granted under the EU’s Digital Services Act, the German advocacy group AlgorithmWatch ran 4,480 election-related search queries on Google and…
EvoGenUI-Bench: Evaluating LLMs as Multi-Turn Generative UI Assistants
arXiv:2608.29387v1 Announce Type: new Abstract: Large language models can generate interactive web interfaces, but reliable generative UI requires…
The Session Is Not the Customer: The AI Identity Crisis
This spring, Redwood Research gave an AI agent built on Claude Opus 4.7 a $5,000 stake, an internet connection and four days to make as much money as…
Can escalation channels redirect reward hacking toward defect disclosure?
arXiv:2608.29460v1 Announce Type: new Abstract: When coding agents encounter defective test infrastructure they may reward-hack: hardcoding outputs or…
Physical Superintelligence Raises $58M Seed Round
Matt Pines, Alex Klokus, and Dr. Alexander Wissner-Gross launched Physical Superintelligence (PSI) on September 1, 2026, announcing $58 million in seed…
Toward Latent Language Model Skills Steering and Optimization: An Empirical Study
arXiv:2608.29459v1 Announce Type: new Abstract: Skills, as a useful abstraction for the procedural capabilities of large language models (LLMs), capture…
