13 posts published in the last hour 09:32IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot Conversations 09:32DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training 09:32Epistemic Warrant for LLM Recommendations: Characterizing the Basis for Reliance…
Author: script
IRWOZ 2.0: A Large Language Model-driven Dialogue Dataset for Industrial Robot Conversations
arXiv:2609.04030v1 Announce Type: new Abstract: IRWOZ has improved industrial human-robot interaction (HRI) dialogue systems through domain-specific…
DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training
arXiv:2609.04094v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards works well when a task has a programmatic checker, but most…
Epistemic Warrant for LLM Recommendations: Characterizing the Basis for Reliance When Ground Truth Is Unavailable
arXiv:2609.04127v1 Announce Type: new Abstract: Large language models are increasingly used to support organizational decisions, yet users often lack a…
Spurious Advantage Hidden in GRPO
arXiv:2609.04063v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is widely studied for reinforcement learning with verifiable…
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM
arXiv:2609.04098v1 Announce Type: new Abstract: Hybrid LLMs pair softmax attention with linear-attention layers such as Gated DeltaNet (GDN), whose…
LLM4CKD: Large Language Models for Early Stage Chronic Kidney Disease Screening
arXiv:2609.04013v1 Announce Type: new Abstract: Early screening of chronic kidney disease (CKD) is critical for timely intervention, yet most machine…
Instruction Duplication as an Inference-Time Control Primitive
arXiv:2609.04024v1 Announce Type: new Abstract: Procedural instruction following is a basic requirement for controllable language-model systems,…
InSituMeasure: Probing Situated Measurement Grounding in Industrial Scenes with Multimodal Large Language Models
arXiv:2609.04014v1 Announce Type: new Abstract: For trained operators, gauge reading requires little specialized knowledge, low cognitive effort, and high…
ATV Big Air Tour turned 3 days of work into 3 hours with ChatGPT
ATV Big Air Tour uses ChatGPT Work to speed up marketing, merchandising, and more. It even turned merchandise photos into an inventory website in 15…
