arXiv:2609.05079v1 Announce Type: new Abstract: Autonomous coding agents are increasingly proposed as AI-scientist systems that conduct analyses and write…
Tag: AI
Constructing and Evaluating Clinical Reasoning Trajectories for Medical Agent
arXiv:2609.05090v1 Announce Type: new Abstract: Evaluation of medical artificial intelligence agents remains predominantly answer-centric, assessing only…
LLM-Guided Program Evolution for Circle Packing: Breaking 10 Packomania Records for $28
arXiv:2609.05093v1 Announce Type: new Abstract: We present Discovery Loop, a lightweight system that uses a large language model (LLM) to iteratively…
Measuring AI Accountability Through Argumentation Analysis: Can Model Reasoning Withstand Scrutiny?
arXiv:2609.05088v1 Announce Type: new Abstract: AI oversight methods rely on ground truth for validation, but what constitutes appropriate AI behavior is…
MePo++: Unifying Representation Refinement and Reconciliation for General Continual Learning
arXiv:2609.05075v1 Announce Type: new Abstract: General continual learning (GCL) aims to learn from evolving data streams without task identities,…
Language models judge war differently when tested for alignment
arXiv:2609.05009v1 Announce Type: new Abstract: Safety evaluations can mischaracterize deployed behaviour if artificial-intelligence systems respond to…
Towards Efficient Evaluation of Evolutionary Transfer Optimization: Case Studies on Task-Parameterized Applications
arXiv:2609.05040v1 Announce Type: new Abstract: As evolutionary transfer optimization (ETO) scales to larger collections of related tasks, problem…
Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent Alignment
arXiv:2609.05036v1 Announce Type: new Abstract: AI alignment requires AI systems to adhere to human norms, values, or intentions. Under value pluralism…
A Tree-based RAG Framework for Evidence-Intensive QA via Adaptive Planning and Topology-Aware Evidence Gathering
arXiv:2609.04981v1 Announce Type: new Abstract: Recent structured RAG methods leverage tree- or graph-based reasoning structures to improve multi-hop QA.…
Proteomic Aging Clocks Track Biological Age Reversal in Rentosertib Trial
Researchers from Insilico Medicine and an international group of academic collaborators have applied six independently developed proteomic aging clocks to…
