arXiv:2609.17226v1 Announce Type: cross Abstract: An agent that learns from rewards has to trust whatever reports those rewards. When the reports suddenly…
Category: AI
Vroom-Vroom at SHROOM-Visions: A Multi-Judge Committee for Detecting Hallucinated Spans in Vision-Language Outputs
arXiv:2609.17327v1 Announce Type: cross Abstract: This paper describes our submission to the SHROOM-Visions shared task on detecting and classifying…
Knowledgator Releases GLiFormer: A 575M-Parameter Encoder That Hits 91.10 F1 on Nested JSON Extraction Without Generating Tokens
GLiFormer Large scores 91.10 F1 on nested JSON, near GPT-5.6-luna’s 91.96, while grounding every value in source spans.
Mo’ Models, Mo’ Problems: How to best select model pools when designing Multi-Agent Systems
arXiv:2609.17306v1 Announce Type: cross Abstract: Multi-agent Systems (MAS) combine multiple model outputs to solve complex reasoning tasks. However,…
Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
Anthropic and OpenAI want to embed independent safety evaluators inside their AI labs. Researchers welcome the unprecedented access, but warn meaningful…
After the Party: Governing What a Viral Agent-Skill Ecosystem Left Behind
arXiv:2609.17274v1 Announce Type: cross Abstract: AI agents increasingly act through agent skills, i.e., natural-language instructions, that direct a host…
5 Free Microsoft GitHub Courses to Learn Data Science and Artificial Intelligence
Explore five free Microsoft GitHub courses covering data science, machine learning, artificial intelligence, generative AI, LLMs, RAG, fine-tuning, and AI…
FROD: Feature Matching Residual Denoising Oracle Bone Decipher
arXiv:2609.17227v1 Announce Type: cross Abstract: Oracle bone script (OBS), one of the earliest Chinese writing systems, plays an important role in the…
MUMINS: Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis
arXiv:2609.17169v1 Announce Type: cross Abstract: Forecasting anatomical changes such as tumor growth and neurodegeneration is a challenging generative…
A unified framework for global and local interpretability using adaptive derivative-ordered random explanation
arXiv:2609.17171v1 Announce Type: cross Abstract: The interpretability of complex machine learning models is of paramount importance, especially in…
