arXiv:2609.07627v1 Announce Type: new Abstract: AI agents sometimes act aligned when they infer they are being tested, and differently when not. We argue…
Tag: cs.AI updates on arXiv.org
From Simulated Citizens to Simulated Deliberation: Challenges in Representation and Interaction
arXiv:2609.07573v1 Announce Type: new Abstract: Multi-agent LLM deliberation has been explored as a scalable way to simulate public deliberation. For such…
FinCUABuild: Can Agents Build Reliable Benchmarks for Dynamic Financial Computer Use?
arXiv:2609.07603v1 Announce Type: new Abstract: Financial scenarios are diverse and complex, spanning varying data conditions, tool configurations, and…
A Tool-Augmented, GPT-4 Chatbot for Real-Time Repository Data Analysis
arXiv:2609.07586v1 Announce Type: new Abstract: Software repositories contain vast amounts of data on code contributions, bug reports, and project…
Quantile-Led Feature Extraction for Multi-Horizon Predictive Maintenance in Industrial Manufacturing Systems
arXiv:2609.07533v1 Announce Type: new Abstract: In data-driven predictive maintenance (PdM), feature extraction is usually treated as fixed preprocessing:…
The Internal Anatomy of Strategic Choice in Large Language Models
arXiv:2609.07478v1 Announce Type: new Abstract: Large language models act as strategic agents and models of human choice, yet choosing like a strategic…
Scoring Without the Engine: Validating a Deterministic, Manipulation-Resistant Content Score for Generative Engines, End to End
arXiv:2609.07559v1 Announce Type: new Abstract: How do you validate a cheap, deterministic proxy for an oracle that is expensive, rate-limited, and…
CIT-CAD: Constraint Intent Tree-based CAD Code Generation and Verification
arXiv:2609.07434v1 Announce Type: new Abstract: Natural-language Computer-Aided Design (CAD) code generation aims to turn design intent into executable…
Modus Tollens and Counterfactuals and Counterfactual Reasoning Based on Three Types of Negation
arXiv:2609.07483v1 Announce Type: new Abstract: Modus Tollens (MT) is a classical logical inference rule, while counterfactuals are hypothetical…
AAS-RAIL: Improving Information Extraction for Asset Administration Shells through Retrieval-Augmented In-Context Learning
arXiv:2609.07334v1 Announce Type: new Abstract: The Asset Administration Shell (AAS) is a cornerstone of Industry 4.0 and the Digital Product Passport,…
