arXiv:2609.22254v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a promising approach for transferring knowledge between language models,…
Category: cs.AI updates on arXiv.org
Strategy Accumulation and Guided Execution for Automated LLM Fine-Tuning
arXiv:2609.22257v1 Announce Type: cross Abstract: Producing task-specific large language models requires discovering effective training strategies through…
Deep Persona: A Psychologically Grounded Architecture and Evaluation Framework for Role-Playing Agents and Simulations
arXiv:2609.22255v1 Announce Type: cross Abstract: Existing approaches to persona simulation with Large Language Models (LLMs) mostly rely on shallow…
RS-Claw-Evolution: Environment-Feedback-Driven Evolution for Lightweight Remote Sensing Agents in Long-Horizon Tasks
arXiv:2609.22258v1 Announce Type: cross Abstract: Large language model-driven remote sensing (RS) agents offer a promising approach to automating…
DIPLOMAT: Dialogue-Span-Aware Direct Preference Optimization for Polite Persuasive Workplace Negotiation Dialogues
arXiv:2609.22256v1 Announce Type: cross Abstract: Effective workplace negotiation requires balancing multiple objectives, including achieving task goals,…
Predictors and Orchestrators: Parsimonious Machine Learning within an Agentic AI Harness for Multi-Horizon Karst Aquifer Forecasting
arXiv:2609.22251v1 Announce Type: cross Abstract: Forecasting karst aquifer dynamics is difficult because recharge responses are nonlinear, event-driven,…
CAMFT: Conflict-Aware Mergeable Fine-Tuning for Large Language Models
arXiv:2609.22253v1 Announce Type: cross Abstract: Model merging has emerged as a promising paradigm for integrating multiple task-specific capabilities…
Checkpoints Are Not Enough: Trust Calibration in CoSLR, a Human-AI System for Systematic Literature Reviews
arXiv:2609.22248v1 Announce Type: cross Abstract: Systematic Literature Reviews (SLRs) are essential for evidence-based research but remain…
A Tutorial on Prompt Engineering: From Messy Thoughts to AI Workflows
arXiv:2609.22249v1 Announce Type: cross Abstract: This paper treats prompt engineering as a discipline for turning informal human intent into structured…
CALM: A Calibrated LLM Choice Network Framework for Activity-Based Traveler Simulation
arXiv:2609.22252v1 Announce Type: cross Abstract: We present CALM, a reproducible hybrid framework that integrates an optional large language model (LLM)…
