arXiv:2609.24346v1 Announce Type: new Abstract: Graph Retrieval-Augmented Generation (GraphRAG) has remarkably enhanced large language models on complex…
Category: AI
Fathom-Vaidya: Advancing Medical Reasoning with Rubric-Based Rewards
arXiv:2609.24480v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in healthcare requires robust performance across two complementary…
Ema raises $77M as AI starts eating into enterprise software and services
Ema has raised $140 million to date and has more than 50 enterprise customers, including Google and Microsoft.
Few-Shot Demonstrations Elicit the Use of In-Context World Representations in LLMs
arXiv:2609.24352v1 Announce Type: new Abstract: Large language models (LLMs), when acting as agents, are expected to take observed data in context, infer…
OpenAI extends cyber access to Ukraine for civilian defense
OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.
VLM-in-Sandbox: Visual Workspaces for Agentic Visual Reasoning
arXiv:2609.24362v1 Announce Type: new Abstract: Sandboxed computer environments support multi-step reasoning with tools, executable programs, and…
Brain-Token Learning: Microstate-Based Tokenization and Multi-Scale Interaction for Long-Horizon EEG Sequence Modeling
arXiv:2609.24324v1 Announce Type: new Abstract: Electroencephalography (EEG) provides a non-invasive window into dynamic brain activity, yet modeling…
When and How Should an Agent Clarify? CIGAsk: Teaching LLMs to Clarify via Counterfactual Information Gain
arXiv:2609.24290v1 Announce Type: new Abstract: Instruction-tuned LLMs faced with underspecified queries often commit to a single interpretation rather…
How Many Pixels Is a Digit Worth? Place-Aware Coordinate Entropy for GUI Agent Confidence Estimation
arXiv:2609.24277v1 Announce Type: new Abstract: GUI agents predict click coordinates as digit-token sequences, but standard text-LLM confidence estimation…
Taming CoT Obfuscation in VLMs: From Mechanistic Evidence to Activation Enforcement
arXiv:2609.24243v1 Announce Type: new Abstract: Reinforcement learning (RL) improves reasoning in vision-language models (VLMs) but can induce…
