200 posts published today
- 21:32DIET: Deletion-response Expert Trimming for Video Diffusion Transformers
- 21:32One year in: How Microsoft Research Asia – Singapore is advancing research, partnership and talent for real-world impact
- 21:32A neural network that maintains and retrieves memories based on context
- 21:32Valor, Atreides, and Sequoia back AI startup Flow Engineering at $750M valuation
- 21:32Adam under Generalized Smoothness with Second-Moment-Type Stochastic Gradients
- 21:32OpenAI Releases GPT-6.1 Sol: Near-Astra Coding and Computer Use at One-Fifth of Astra’s Token Price
- 21:32OmniVCBench: Benchmarking Evidence-Grounded Multimodal Reasoning Towards AI Virtual Cells
- 21:32Google DeepMind Unveils Gemini 4 Argon with 1M Output Tokens for Coding, Knowledge Work and Cyber Defense
- 21:32Can a Cacheable Decision Model Follow Rules?
- 21:03Cross-Entropy Guided Routing in Mixture-of-Experts Large Language Models
- 21:02Spatiotemporal Hyperedges for EEG Seizure Detection and Prediction
- 21:02Context Language Models
- 21:02ContextRender: From Execution Dependencies to Agent Context
- 21:02Generative Interactions: Weaving Multiparty Human Motion with Bilevel Latent Dynamics
- 21:00AI News Brief Hourly Summary 2026-09-30 23h : 13 posts
- 20:32EngiWorld: What Can Frontier Agents Deliver in Professional Engineering Environments?
- 20:32KUPAS MASTER: Distilling the Tacit Expertise of Master Practitioners into Agent-Ready Experience Corpora
- 20:32Learning from Shared-Control Overrides: Context-Driven Acceleration Profile Prediction for Personalized Overtaking
- 20:32WISE-ATTA: When to Ask for Labels in Budgeted Active Test-Time Adaptation
- 20:32Locating Answer-Correctness Signals in Frozen Large Language Models
- 20:03XU-RS: Explaining Credal Width in Random-Set Language Models
- 20:03EnterpriseBench: Benchmarking LLM Agents on Enterprise-Level Strategic Reasoning and Decision-Making
- 20:03Flattening the Connectome Spectrum: A Spectral Filter for FC Induces a Pretraining Target for fMRI Encoders
- 20:03Gemini 4 Argon: our next era of frontier intelligence
- 20:03MeanFlowAdvantage: Stable Reward Fine-Tuning for Few-Step Average-Velocity Generators
- 20:03Watch the winning trailer from the Future Vision XPRIZE, The Gifted.
- 20:03Beyond a single latent space: a dual-latent world model for long-horizon planning
- 20:00AI News Brief Hourly Summary 2026-09-30 22h : 14 posts
- 19:32SkillGym: Training Skill-Use Agents with Automatic Verifiable Environment Generation
- 19:32FOCUS: Training-Free Decision-Preserving Context Compression for LLM Agents
- 19:32Rational Clarification by Assistive Agents via Value-of-Information Reasoning
- 19:32Introducing Claude Sonnet 5.5 on AWS
- 19:32How Can Recommendation Feedback Evolve Agent Memory?
- 19:32OpenAI and Synopsys team up to build an AI model that designs chips like a seasoned engineer
- 19:32Boundary-State Control for Tool-Using Language-Model Agents: Commit-Time Consistency under State Drift
- 19:03VeriWeave Govern: Evidence-Gated Deterministic Runtime Governance for Enterprise AI Agents
- 19:03Authority Before Utility: Non-Compensatory Control for Persistent LLM Memory
- 19:03Commitment Hierarchies under Intent Revision: A Belief-Revision Account of Salvage in Tool-Use Agents
- 19:03Governing the Edge: Automating Commercial Property and Casualty Insurance Underwriting via a Hybrid Local-Cloud Multi-Agent Framework
- 19:03OpenAI’s Jev clone could help the frontier lab stop its swarming agents
- 19:03CRASM-Gate: Deterministic-First Constraint- and Role-Aware Semantic Mapping with Selective Model Assistance Across Heterogeneous Industrial Standards
- 19:00AI News Brief Hourly Summary 2026-09-30 21h : 17 posts
- 18:33Demistifying Data and Simulator Assumptions in Supervised Causal Discovery
- 18:33The Lenfest Institute grows landmark program with expanded OpenAI support
- 18:33Beyond Prompt Count: How Data Shapes Transfer in On-Policy Distillation
- 18:33Reddit is killing RSS feeds and ending public API access because of AI bots
- 18:33Pretrain Once, Route Anywhere: Towards a Foundation Model for LLM Routing
- 18:33Google drops Gems for Skills, joining OpenAI and Anthropic in the shift to agent-ready prompt formats
- 18:33Direct Experience World-Model Optimization: Learning the World Beyond Action Imitation
- 18:33AI voice startup ElevenLabs doubles valuation to $22B
- 18:32Routing Should Pay for Itself: Sparse Supervision for Economical LLM Routing
- 18:03VISTA: Value-Informed Event Appraisal for Multimodal Emotion Conflict
- 18:03Solving Without Stopping: On-Policy Distillation at Small Scale
- 18:03Mubric: Mutation Testing-Guided Rubric Generation for LLM Evaluation
- 18:03Disrupting a coordinated model-distillation campaign
- 18:03Teaching LLMs to Generate Challenging MILP Instances via Solver Feedback
- 18:03When can we say AI made a scientific discovery?
- 18:03Seek Before You Move: Evidence Seeking for Progress Grounding in Vision-Language Navigation
- 18:01AI News Brief Hourly Summary 2026-09-30 20h : 14 posts
- 17:32ReMem: Rethinking Perception and Memory in Long-Context Recommendation Agents
- 17:32MetaCtrl: Your Large Language Models Can Reason Better and More Concisely with a Metacognitive Controller
- 17:32Transolver-$\sigma$: Joint Spectral-Physical Subspace Modeling for Neural PDE Solving
- 17:32The ugly economics of consumer AI
- 17:32Task-Relevant Null-Space Residuals for Non-Injective Neural Mappings
- 17:32Meta dodges billions in US taxes by calling its AI data centers experiments
- 17:32AssayRouter: Historical Utility Priors for Frozen Molecular Predictor Routing
- 17:03Learning to Prove, Not Just to Answer: Reinforcement Learning from Formal Verification for Natural-Language Logical Reasoning
- 17:03OptiCom : A Unified Framework for State-Conditioned Composition in LLM-Driven Optimization
- 17:03Information Bottleneck-Guided Adaptive Hypergraph Transformer for Brain Disease Diagnosis
- 17:03Generate images and video with vLLM-Omni on SageMaker AI – Part 2
- 17:03Foundations of Proactive Agents: Principles, Technical Layers, and Proactivity-Gym
- 17:03Build real-time voice applications with vLLM-Omni on SageMaker AI – Part 1
- 17:03Asking for What Was Never Requested: Horizontal and Vertical Proactivity in Agents
- 17:00AI News Brief Hourly Summary 2026-09-30 19h : 21 posts
- 16:33Destro AI’s secret sauce is getting robots and humans on the same page
- 16:33V-Engram: Trigger-Indexed External Memory for Modular Text-to-Image Personalization
- 16:33Implementing synthetic monitoring using Amazon Nova Act
- 16:33SimpleEvol: An Agent-Loop Framework for LLM-Driven Automated Heuristic Design with Minimal Human Priors
- 16:33FTC launches sweeping probe into OpenAI, Anthropic, and other AI labs over consumer protection concerns
- 16:33From Learner Behavior to Reusable Skills for Effective and Efficient Learner Simulation
- 16:33Meta disputes claim that Muse read a user’s private messages without permission
- 16:32Absorbed in Inertia: Activation Analysis for Computer-Use Agents
- 16:32Automating Amazon Textract adapter lifecycle management across accounts
- 16:32Accelerated surrogate dynamics for dynamical, stochastic system evolution
- 16:04Train Ahead, Distill Back: Bootstrapping On-Policy Self-Distillation for Large Language Models
- 16:04Forecasting space weather risks on power grids
- 16:03DoorDash launches an AI agent you can text to order food
- 16:03From Judgment Quality to Downstream Utility: Rethinking LLM-as-a-Judge for Open-Ended Tasks
- 16:03Helping small businesses put AI to work
- 16:03When Tools Silently Lie: Evaluating and Mitigating Blind Compliance in Tool-Augmented Data Agents
- 16:03Query claims in natural language with Amazon Bedrock Knowledge Bases
- 16:03SkillCome: Group Contrast Skill Optimization with Dual Memory
- 16:03Instinct’s new product recommendations are giving some users the ick
- 16:03When Should Agents Check External State? Budgeting Observations for Stored Intentions
- 16:00AI News Brief Hourly Summary 2026-09-30 18h : 16 posts
- 15:33actr: aligning thoughts and responses for multilingual safety in reasoning llms
- 15:33Watch-Think-Interact: Bootstrapping Long-Horizon Multi-Turn Streaming Video Reasoning with Reinforcement Learning
- 15:33Learning from Viable Failure Prefixes: Milestone Viability Potential Policy Optimization for Long-Horizon LLM Agents
- 15:33Gemini 3.5 Transcribe vs OpenAI’s GPT-Transcribe
- 15:33MatToolBench: Benchmarking Multimodal Agents in Real-World Materials Science Workflows
- 15:33Build a multi-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances
- 15:32Breaking the Illusion of Review Reliability under Static Evaluation: SCOPE Fuzzing for LLM-based Scientific Reviewers
- 15:04FedLAFP: Low-Rank Aggregation Meets Full-Rank Personalization in Federated Fine-Tuning
- 15:04Physics-Informed Multi-Agent Coordination for Hospital Patient Flow Optimization
- 15:03China’s AI industry closes ranks as Deepseek ships open-source software for Huawei’s Ascend chips
- 15:03Beyond Low-Rank Parameterization: Narrowing the Gap Between LoRA and Full Fine-Tuning via Gradient Decomposition
- 15:03Introducing SynthID Bio
- 15:03AnyAct: Universal Action for Self-Evolving Agents
- 15:03Open TTS Leaderboard: Scalable Evaluation for Multilingual Text-to-Speech and Voice Cloning
- 15:03Language as the Interface: Foundation-Model Contrastive Learning Links Transcriptomes and Electrophysiology
- 15:00AI News Brief Hourly Summary 2026-09-30 17h : 15 posts
- 14:33Dual-Channel Robust Group-Relative Policy Optimization via Advantage and Sequence-Weight Estimation
- 14:33Cerebras Systems’ Andrew Feldman on whether AI can keep scaling at TechCrunch Disrupt 2026
- 14:33CADOC: Cache-Aware Dynamic Object Context for Long-Horizon Agents
- 14:33Did AI Just Solve One of Mathematics’ Biggest Problems?
- 14:33CF-LoRA: Decoupled Factor Aggregation and Adaptation-Aware Client Clustering for Federated LoRA Fine-Tuning
- 14:33Restate lands $20M as the need for durable infrastructure increases with AI agents
- 14:33ImbalancE: Inference-Time Latent Search Against Degree Imbalance in Link Prediction
- 14:333 days left to exhibit: Turn visibility into your next opportunity at TechCrunch Disrupt 2026
- 14:33REALHOP: Rethinking Multi-Hop Reasoning Evaluation via Behavioral Auditing
- 14:03Neuro-Symbolic Computer Use: Learning Reusable Policies for Reliable and Efficient Execution
- 14:03VLALight: A Vision-Language-Action Model for Traffic Signal Control
- 14:03CoEM: Empowering Long-Context Reasoning with Commit-on-Evidence Memory
- 14:03Learn from the Gap: Differential-Aware Advantage Pruning with Adaptive Rollout Sampling for GRPO
- 14:03SCA: Spatial Credit Assignment for Reinforcement Learning of GUI Agents
- 14:00AI News Brief Hourly Summary 2026-09-30 16h : 11 posts
- 13:33STAR-GRPO: Canonical Anchoring and Reliability-First Advantages against Representation-Dependent Reward Hacking
- 13:33Beyond Sub-Gaussian Detector Scores: Robust Weighted Profile-Loss Change Point Detection for Human-LLM Text Segmentation
- 13:33HorizonFlow: Variable-Length Planning for Offline Goal-Conditioned RL
- 13:33PrecogUI: Proactive GUI Agents via Pre-cognitive Simulation and Experience Retrieval
- 13:33Harness Evolution as Learning: Approximation, Generalization, and Optimization Limits of Self-Improving Personal Agents
- 13:04State Trace Rationale As Auxiliary Task in Reinforcement Learning
- 13:04IronLLM: Forging Compact Edge-Native Language Models for Real-Time Embodied Intelligence
- 13:03WEFT: Scaling Tool-Use Post-Training for General-Purpose Agents
- 13:03When Upstream Messages Override Correct Answers: A Controlled Study of Multi-Agent LLM Collaboration
- 13:03Automated Screw Planning for Reduced Pelvic Fractures Based on Statistical Shape Models and Deep Learning
- 13:00AI News Brief Hourly Summary 2026-09-30 15h : 14 posts
- 12:33Geometry-Conditioned Fixed-Scaffold Encoders for Time-Warp Robust Sequence Retrieval
- 12:33Where Does Staleness Accumulate? Pool Aware Effective Staleness Control for Asynchronous RL in LLM Post-Training
- 12:33The Default Trap: Rethinking Plan Evaluation in Tool-Using LLM Agents
- 12:333 Numba Tricks for Python Runtime Optimization
- 12:32Calibrate the Decisions That Change the Future: On-Policy Post-Training Quantization for Multimodal Large Language Models
- 12:32Ollama for Managing Local Language Models: A KDnuggets Cheat Sheet
- 12:32ARC-KV: Amortizing Anchor Search for Reconstruction-Based KV Cache Compaction
- 12:04CAD-Native Transformer Operators for AI-Aided Engineering
- 12:04Aperture: Merge-Consistent Rotary States for Compressed Tokens
- 12:04AI as a Compiler: Compiling Triton kernels without the Triton compiler
- 12:04Code4Scene: Benchmarking Coding Agents for Constructing and Editing 3D Scenes
- 12:04Airbnb adds AI search, more social features
- 12:03UpliftMem: Learning Set-Level Uplift for Agent Memory Retrieval
- 12:00AI News Brief Hourly Summary 2026-09-30 14h : 15 posts
- 11:33SIPO: Unifying Reinforcement Learning with On-Policy Self-Distillation
- 11:33Generalizable Lifelong Model Editing via Preference Optimization
- 11:33EASE: Behavior-Adaptive Skill Curation for Self-Evolving Agents
- 11:33Generative AI Gives Spacecraft the Autonomy Engineers Once Feared
- 11:33Distinguish or Homogenize: Last-Chance Policy Identification and Risk-Budgeted Recovery under Irreversible Resource Depletion
- 11:33Anthropic says Zhipu’s open-weight GLM-5.3 nearly matches Claude Mythos Preview at building exploits
- 11:33Emergent Specialization in Populations of Self-Supervised Collaborative Vision Experts Without a Shared Gate or Cross-Agent Gradients
- 11:04Can AI Scientists Change Their Minds? Prior-Evidence Conflict in Synthetic Universes
- 11:03JudgeProfile: Understanding and Steering Subjectivity in LLM Judges
- 11:03BiFE: Search-Efficient Discovery of CPU-Only Branching Policies via LLM-based Bi-Fidelity Evolution
- 11:03Holo4: powering generalist computer-use agents
- 11:03Can Agents Design Libraries for Agents?
- 11:03“We’re not going to shoot ourselves in the foot” over hack fallout, says OpenAI’s chief research officer
- 11:03MLToolBench: Learning Tool-Augmented Agents for Machine Learning Development
- 11:00AI News Brief Hourly Summary 2026-09-30 13h : 13 posts
- 10:33Semantic Projection for Continual Self-Evolution of Language Agents
- 10:33FineSID: Scalable and Efficient Semantic Identifier Learning for Generative Recommendation
- 10:33FairDiff: Mitigating the Self-Reinforcing Matthew Effect in Diffusion Recommender Models
- 10:33RankBuffer: Efficient Ranking-Based Rewards for Open-Ended Generation
- 10:33Google is paying almost no publishers almost nothing for content used in AI answers
- 10:33Distilling Agentic Systems: A Roadmap across Models, Artifacts, and Harnesses
- 10:03Multi-Channel Mitigation of Source-Trust Shortcuts in Fact-Checking RL Agents
- 10:03Neural Structural Reasoner: A Brain-inspired Architecture for Reasoning over Structured Knowledge
- 10:03SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation
- 10:03Who’s liable when AI agents go rogue?
- 10:03Transformers Stop Thinking Too Early, and a Tiny LoRA Fixes It
- 10:03Trump and tech CEOs sign an AI code of conduct that’s only “morally binding”
- 10:03DualTrack: Synchronized speech-gesture generation via symmetric coupling of pretrained priors
- 10:00AI News Brief Hourly Summary 2026-09-30 12h : 11 posts
- 09:32SafeCoEvo: Co-Evolving Safety Harnesses and Guards for LLM Agents at Test-Time
- 09:32Visual sensitivity is not claim retractability: persistence-aware credit assignment for multimodal reinforcement learning
- 09:32MemEvo: Automatic Discovery of Streaming Video Memory Mechanisms
- 09:32Divide and Inject: Can Agents Reconstruct an Indirect Prompt Injection from Fragments?
- 09:32MAADBench: The Refreshable Paradigm for Anomaly Detection in Multi-Agent Systems
- 09:03Human-AI Collaboration: From Paradoxes to Patterns
- 09:03AVIO: Learning to Add and Remove Sounding Objects in Audiovisual Scenes
- 09:03Learning to Harvest Without Collapse in a Regenerative Commons: A Lagrangian Framework
- 09:03Going Beyond State-Reaching: Learning Abstractions for Intrinsically Motivated Option Discovery
- 09:03BRIDGE: Bilevel Retrieval-Credit-Aware Agentic Reinforcement Learning
- 09:00AI News Brief Hourly Summary 2026-09-30 11h : 13 posts
- 08:32Bits Under ZK-LLM: Evaluating Zero-Knowledge-Friendly Quantization for Verifiable Private LLM Inference
- 08:32Rethinking Reasoning Paths as Phase-Structured Trajectories
- 08:32Support-Set Target Leakage in Relational Foundation Models during In-Context Learning: Model Dependence and Evaluation Reliability
- 08:32From Retrieval to Reasoning: Agentic Mechanism Prediction from Cell Painting Profiles
- 08:32Perplexity Introduces Photon: A Rust-Based Retrieval Engine That Cuts p99 Latency From 800 ms to 65 ms
- 08:32The Safety Operator: Modulating the Expression of Safety Instructions via Spectral Optimization
- 08:03Persona Dosing: Calibrated Activation Steering for Graded Trait Control
- 08:03ARCagent: An Adaptive Retrieval Calibration Agent for Clinical Question Answering
- 08:03Engineering Simplicity: Simple Mechanism Interfaces Steer LLM Agents
