arXiv:2608.28646v1 Announce Type: new Abstract: Large language models (LLMs) can generate plausible-sounding ETF portfolios while silently violating basic…
Category: AI
From Extraction to Governed Memory: Multi-Agent Knowledge Graph Construction with Domain-Expert Review
arXiv:2608.28642v1 Announce Type: new Abstract: Knowledge graphs used by agentic systems are often treated as flat stores of extracted triples, with…
CrossAudit: A Git-Native, Cross-Vendor Audit Loop for Agentic Science
arXiv:2608.28631v1 Announce Type: new Abstract: An AI scientist should not grade its own homework. Yet in the systems we examined, the agent that reviews…
AI Scientist Mission Control (AIMC): Visual Analytics for Human Oversight of Autonomous Scientific Discovery
arXiv:2608.28637v1 Announce Type: new Abstract: Autonomous scientific discovery systems can generate large numbers of research ideas, experiments, and…
Self-Evolving Skills via Surrogate-Guided Solve-and-Reproduce
arXiv:2608.28638v1 Announce Type: new Abstract: Agent skills are portable packages of instructions and resources an agent consults at deployment.…
Reward-Oracle MCTS for Formal Theorem Proving: Sample-Efficient Search and the Need for Kernel-Level Proof Auditing
arXiv:2608.28639v1 Announce Type: new Abstract: Formal theorem proving with large language models remains challenging due to the difficulty of navigating…
AutoScientist-Quant: Self-Evolving Coding Agents for Automatic Research in Quantitative Investment
arXiv:2608.28632v1 Announce Type: new Abstract: Large language model agents can discover alphas, yet current methods have three weaknesses. The search…
CDEP Agent: Connecting Meteorologically Detected Temporal Compound Events to Real-World Documentary Evidence
arXiv:2608.28628v1 Announce Type: new Abstract: Compound drought-to-extreme-precipitation (CDEP) events are recognized in climate science as a growing…
TPvG: A Moral Decision Framework for Large Language Models from One-Shot to Sequential Feedback
arXiv:2608.28610v1 Announce Type: new Abstract: Existing LLM moral evaluations typically present models with isolated moral vignettes and elicit a…
Machine Learning-Enhanced Tabu Search for Tactical Wireless Network Design
arXiv:2608.28627v1 Announce Type: new Abstract: Designing high-performance tactical wireless networks under realistic operational constraints gives rise…
