arXiv:2604.11557v3 Announce Type: replace Abstract: Tool-use capability is a fundamental component of LLM agents, enabling them to interact with external…
Category: AI
This Python Library Can Run Pandas Workloads Up to 20x Faster
Discover how FireDucks can speed up pandas workloads with lazy execution, compiler optimization, and multithreaded processing, delivering up to 20x faster…
From High-Dimensional Spaces to Verifiable ODD Coverage for Safety-Critical AI-based Systems
arXiv:2604.02198v2 Announce Type: replace Abstract: While Artificial Intelligence (AI) offers transformative potential for operational performance, its…
Real-Time Intelligence with IBM Time Series Models on Confluent
This post has no text preview — click the link below to read the original article. This article has been indexed from Hugging Face – Blog Read the original article: Real-Time Intelligence with IBM Time Series Models on Confluent
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
arXiv:2603.03072v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist scientists across diverse workflows. A…
Safety overview: GPT-6 Astra
GPT-6 Astra is our most capable broadly deployed model and our first to reach the Critical level of cybersecurity capability under our Preparedness…
BUZZY: Contrastive Scoring to Mitigate Text-Induced Bias in Multimodal Multiple-Choice QA
arXiv:2603.28026v3 Announce Type: replace Abstract: Multimodal multiple-choice question answering (MCQA) provides a standardized and objectively…
Edit Knowledge, Not Just Facts via Multi-Step Reasoning over Background Stories
arXiv:2602.02028v3 Announce Type: replace Abstract: Enabling artificial intelligence systems, particularly large language models, to update knowledge and…
Stepwise Think-Critique: Interleaved Reasoning and Self-Critique in a Single LLM
arXiv:2512.15662v4 Announce Type: replace Abstract: Human beings solve complex problems through critical thinking, where reasoning and evaluation are…
Achieving Olympiad-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning
arXiv:2512.10534v4 Announce Type: replace Abstract: Large language model (LLM) agents exhibit strong mathematical problem-solving abilities and can even…
