OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.
Tag: AI
A Dual-Process Perspective on Nudge Susceptibility in LLM-Based GUI Agents
arXiv:2609.19843v1 Announce Type: new Abstract: LLM-based GUI agents increasingly act on behalf of users in digital environments that were designed with…
Steering Equilibrium Selection in Regularized Self-Play via the Reference Policy
arXiv:2609.19820v1 Announce Type: new Abstract: Regularized self-play — the family behind DeepNash’s Stratego play — drives a two-player zero-sum policy…
Rethinking Multi-Agent Collaboration: When More Is Less
arXiv:2609.19759v1 Announce Type: new Abstract: The rapid advancement of large language models and single-agent harnesses has reshaped the landscape of…
TorchCraft: Unified binder design by inverting an all-atom structure predictor
arXiv:2609.19770v1 Announce Type: new Abstract: All-atom structure predictors model diverse molecular interactions, but using their learned structural…
Contagion on the Trading Floor: How Adversarial Signals Spread in Multi-Agent Trading Systems
arXiv:2609.19789v1 Announce Type: new Abstract: Multi-agent trading systems built on large language models (LLMs) are beginning to appear in quantitative…
Stanford Researchers Release Paper2Agent: Turning Research Papers Into AI Agents That Reproduce Results and Run on New Data
Paper2Agent, published in Nature, converts papers into validated MCP tools, scoring 91.2% on 300 questions across 74 papers.
Integrating knowledge from case reports: a medical ontology based multimodal information system with structured summary
arXiv:2609.19775v1 Announce Type: new Abstract: Published medical case reports serve as a crucial medical information carrier, documenting discoveries in…
Replan, Repair, or Edit? A Unified Empirical Evaluation of Travel Agents for Itinerary Revision under Resource Disruptions
arXiv:2609.19654v1 Announce Type: new Abstract: Travel-planning agents generate itineraries that may become infeasible after acceptance because of flight…
AutoData: Agentic Search for Pre-training Data Selection
arXiv:2609.19754v1 Announce Type: new Abstract: LLM agents have recently shown promise in automating machine learning engineering by editing model and…
