arXiv:2608.12841v1 Announce Type: cross Abstract: We study recursive self-improvement at the level of quantitative-investment research: whether an…
Tag: AI
From Atomic Evidence to Logical Composition: Structured Compositional Reasoning over Compound Answer Options
arXiv:2608.12836v1 Announce Type: cross Abstract: Large language models often fail when answer options require combining atomic judgments under explicit…
Erase but Preserve: Controllable Removal of Copyrighted Animation Characters via Optimized Semantic Anchors
arXiv:2608.12806v1 Announce Type: cross Abstract: The exceptional generation capabilities of text-to-image diffusion models have raised copyright…
CRAFT: LLM-Based Iterative Refinement for Temporal Reasoning over Clinical Narratives
arXiv:2608.12779v1 Announce Type: cross Abstract: Understanding the temporal progression of symptoms in clinical narratives is critical for disease…
Fast A/B/n Testing: Exact Multi-Policy Comparison via Tree-Coupled Feedback Sharing
arXiv:2608.12831v1 Announce Type: cross Abstract: Online platforms increasingly compare many adaptive decision policies—ranking systems, recommendation…
PIPES: Securing Agent Perception with Provenance and Priors
arXiv:2608.12789v1 Announce Type: cross Abstract: Tool-using agents consume external data from sources with different levels of trust, yet tool responses…
ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval
arXiv:2608.12720v1 Announce Type: cross Abstract: While Large Language Model (LLM) agents increasingly rely on long-term memory for persistent…
Gambit Security’s “AI Across the Intrusion Lifecycle” Shows How AI Is Moving Deeper Into Real-World Cyberattacks
Gambit Security’s new report, AI Across the Intrusion Lifecycle offers a detailed look at how artificial intelligence is moving beyond a supporting role…
SynAct: A Reasoning-Acting Large Language Model Agent for Adaptive Synthesis Optimization
arXiv:2608.12751v1 Announce Type: cross Abstract: Logic synthesis transforms RTL designs into gate-level netlists, where PPA results are highly sensitive…
SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned for Long-Running Agents, Coding, and Knowledge Work
SpaceXAI released Grok 4.6 on August 12, 2026 — a post-training upgrade over Grok 4.5, not a larger base model. It ties GPT-5.6 Sol Max at 61 on the…
