arXiv:2609.29014v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems can automate alpha factor mining, but their reliance…
Author: script
AI News Brief Hourly Summary 2026-09-25 08h : 13 posts
13 posts published in the last hour 05:32PFArena: Benchmarking Language Models for Protein Modification 05:32Control the Harness, Control the Cost: Routing and Governing AI Coding Agents in the Enterprise 05:32Human-AI-Powered Hypothesis Testing: Cost-Aware Selective AI Scoring and Sequential Human Escalation…
PFArena: Benchmarking Language Models for Protein Modification
arXiv:2609.28921v1 Announce Type: new Abstract: Protein modification requires navigating an immense sequence space, yet wet-lab validation remains…
Control the Harness, Control the Cost: Routing and Governing AI Coding Agents in the Enterprise
arXiv:2609.28919v1 Announce Type: new Abstract: Harnesses, the products that run AI coding agents, are multiplying, and enterprises are rolling them out…
Human-AI-Powered Hypothesis Testing: Cost-Aware Selective AI Scoring and Sequential Human Escalation
arXiv:2609.28859v1 Announce Type: new Abstract: Large language models are increasingly used as inexpensive judges to evaluate outputs, label data, and…
RECLAIM: Can Agents Reproduce the Claims of Machine Learning Papers?
arXiv:2609.28850v1 Announce Type: new Abstract: Reproducing a machine learning paper involves most research steps, from installing software and debugging…
Lightspeed targets $250M for new India fund, focusing on early-stage AI
The Silicon Valley firm is aligning its India fundraising cycle with its global funds for the first time, as it shifts to a shorter investment period.
Forecast-Dojo: Replayable Environments for Benchmarking and Training LLM Forecasting Agents
arXiv:2609.28876v1 Announce Type: new Abstract: We introduce Forecast-Dojo, a replayable environment for benchmarking and training LLM forecasting agents.…
Reinforcement Learning with Verifiable Rewards for Small Search Agents
arXiv:2609.28765v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) performs well on problems with clear rewards, such…
Learned Cross-Task Relationships in Multi-Task Models
arXiv:2609.28776v1 Announce Type: new Abstract: We propose a framework that learns cross-task relationships in multi-task models by approximating the…
