arXiv:2609.30199v1 Announce Type: new Abstract: Scientific discovery begins where known problems end. There, AI systems must engage in exploration:…
Tag: AI
Affected by layoffs? Don’t miss this $75 deal for your TechCrunch Disrupt 2026 Expo+ Pass
Your next opportunity could be one conversation away. Get your Expo+ Pass for just $75. Limited to the first 100 qualifying people.
Jev-Mobile: Jev as an Executor for Mobile GUI Agents
arXiv:2609.30186v1 Announce Type: new Abstract: Vision-language models (VLMs) have become a common foundation for autonomous mobile GUI agents, but most…
Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation
Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessions, including failed ones. The…
SAGE: Mitigating Long-Horizon Reasoning Biases via Topological Guidance
arXiv:2609.30192v1 Announce Type: new Abstract: Long-horizon reasoning remains a central challenge for large language models (LLMs) under sparse-reward…
PrivDrift: Auditing User-Secret Leakage Under Topic Drift in Active LLM Conversations
arXiv:2609.30094v1 Announce Type: new Abstract: Large language models increasingly operate as persistent assistants in user-facing, shared-session, and…
Everything new coming to Meta’s AI agent Muse
CEO Mark Zuckerberg kicked off the company’s annual Connect event in Menlo Park on Wednesday with a keynote that made one thing clear: Meta is going…
Self-Play Pretraining with Zero Data
arXiv:2609.30063v1 Announce Type: new Abstract: Advances in language modeling have been driven by scaling pretraining on ever more data. Yet, the training…
Batching by Length Instead of Looping Item by Item for SLM Optimization
We finish off our short series on SLM optimization with the third entry, focused on batching by length instead of looping item by item.
HEXIS: Compiling Skills into Extended Finite State Machines
arXiv:2609.30123v1 Announce Type: new Abstract: Agent skills provide reusable knowledge and instructions, yet agents must repeatedly infer how to apply…
