200 posts published today
- 21:32Beyond Approved Actions: Runtime Validation of Persistent Outcomes in Agent Workflows
- 21:32CG-HAF: An Interpretable Global-Local Lesion-Burden Fusion Framework for Ordinal Acne Severity Grading in Agentic Skincare Support
- 21:32Are you a Codex Original?
- 21:32AgentXploit: Autonomous Repository-to-Runtime Red-Teaming for AI Agents
- 21:32Source: Inference provider Modal Labs closing in on $750M round at $15.75B valuation
- 21:32UniAR: A Unified Framework for Autism Recognition Enhanced by Multi-View Prompt Learning
- 21:32Basis completes a tax workbook 2x faster with GPT-6 Astra
- 21:32Towards VLA-Dreamer: Refining VLA Behavior Using World Models
- 21:03Resource-Optimized and Energy-Aware Agentic AI Framework Anchored on Blockchain for Secure Software Supply Chains
- 21:03Cognitive Skills in the Age of AI: Computing Students and Experts Perceptions
- 21:03MoSAR: Mixture of Semantic Attention Regimes for Learning Adaptive and Approximable Attention Geometries
- 21:02AMD will acquire Fei-Fei Liβs World Labs for $8.2 billion
- 21:02Softmax Reparameterization for Output-Head Quantization
- 21:02One year in: How Microsoft Research Asia β Singapore is advancing research, partnership and talent for real-world impact
- 21:02Agentic Limit Order Books: Phase Transitions and Market Impact
- 21:00AI News Brief Hourly Summary 2026-09-28 23h : 12 posts
- 20:32Acoustic-to-Text KV Compression for Full-Duplex Speech Models
- 20:32Rethinking Data Quality for AI-Driven Systems: Evidence from Practitioner Interviews
- 20:32Geometric Inconsistency Localization in Multi-View Image Sets
- 20:32BAT-CLIP: Trimodal Alignment of Brain, Audio and Text
- 20:32Improving Visual Sensitivity of LLMs on Multimodal Machine Translation with Metric-based Loss Weighting
- 20:02ReG-SAM: Reference Graph-Driven SAM for 2D Foundational Vessel Segmentation
- 20:02SPADE: Escaping the Popularity-Similarity Frontier to Measure Serendipitous Recommendations
- 20:02FedHisto-PAST: Parameter-Efficient Stain-Aware Federated Learning for Cross-Site Lung Histopathology Classification
- 20:02Teacher-Anchored Selection of Post-Training Quantized Models under Domain Shift
- 20:02Shopify opens checkout to browser-based AI agents
- 20:02AgentRecommender: LLM Agents Enable Customizable Recommender Systems on the User Side
- 20:00AI News Brief Hourly Summary 2026-09-28 22h : 15 posts
- 19:33DepthEvidence: Unifying Metric Depth Prediction and Geometric Reasoning in Multimodal Language Models
- 19:33Pocket-STVG: lightweight architecture for Spatio-Temporal Video Grounding
- 19:32More than 20 leading AI researchers warn that automated AI research poses extreme risks
- 19:32Bayesian Optimization with Fisher Information Geometry: Gradient Bounds and Trust-Region Methods
- 19:32NVIDIA Launches Open Agent Safety Platform: OpenShell Sandboxes Agents on Vera CPUs While Sentry on BlueField-4 Quarantines Them in Milliseconds
- 19:32From Shortcut Learning to Discrete Neural Insertion Sort
- 19:32Watch the winning trailer from the Future Vision XPRIZE, The Gifted.
- 19:32JevAdvBench: A Benchmark and Black-Box Attacks for Reinforcement Learning for Calibrated Decisions Models
- 19:03Can Pixels Alone Reveal Image Origin? Minimax Limits and Learnable Interfaces for Passive Provenance
- 19:03DynBranch: Speculative Subgraph Reuse for Dynamic Agentic LLM Serving
- 19:03Quantum Diffusion Models for Medical Image Analysis
- 19:03The Linear Representation Hypothesis for Vision-Language-Action Models
- 19:03Introducing Claude Sonnet 5.5 on AWS
- 19:03G$^2$PTQ: Improving LLM Post-Training Quantization with Generalized Gradient Compensation
- 19:00AI News Brief Hourly Summary 2026-09-28 21h : 15 posts
- 18:32PORL: Pretrained Offline Reinforcement Learning for the Job Shop Scheduling Problem
- 18:32FARE: Forensic Acceptance Region Estimation for Catching Bait-and-Switch Image Generators
- 18:32FLIP: Final Layer Inference-Time Probing for Vision-Language Models
- 18:32Nvidia launches new platform for reining in rogue AI agents
- 18:32Does Uniform Discrete Diffusion Need Time?
- 18:32Anthropic releases Sonnet 5.5, which it calls a significantly cheaper, faster work partner
- 18:32MVVBench: Benchmarking 4D Reasoning in Vision-Language Models
- 18:03UltraG-Bench: A Multi-task Benchmark for assessing Large Vision-Language Models on Pixel-level Evidence Grounding in Ultrasound
- 18:03Estimating and Orthogonalizing Unknown Pre-training Gradients for Continual Fine-tuning of Large Language Models
- 18:03OneWorld: Learning Consistent Physics Across Actions in World Models
- 18:03The Lenfest Institute grows landmark program with expanded OpenAI support
- 18:03Robust to Which Model Change? A Unified Evaluation of Robust Counterfactual Explanations
- 18:02Anthropic’s Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks while costing up to 30 percent less per task
- 18:02Spackle: Completing Large View Single Image NVS with Adaptive Gaussians
- 18:00AI News Brief Hourly Summary 2026-09-28 20h : 16 posts
- 17:33Persistent Negatives for Adversarial Black-Box On-Policy Distillation
- 17:33Warned alike, AI agents avoid the less-crowded road while people take it
- 17:33When can we say AI made a scientific discovery?
- 17:33MOPD-Router: Rethinking Teacher Routing in Multi-Teacher On-Policy Distillation
- 17:33OpenAI still doesnβt seem to have a handle on all of its rogue AI activity
- 17:32Adaptive Pilot Selection for Unified Semantic Communication and Semantic Sensing in ISAC
- 17:32Google is killing off Geminiβs Gems in favor of βskillsβ
- 17:32Developing a Roadmap to an AI-first Organization: A Case Study in Embedded Software Development
- 17:03NavGen: Visual Generative Models as a Scalable Data Engine for Embodied 3D Navigation
- 17:03Evaluation Is All You Need for Multi-Modal Autonomous Driving
- 17:03XPhysICS: Cross-Physical-Domain Threat Grounding for Industrial Control Systems Security
- 17:02Meta launches enterpriseΒ AI platform, hiresΒ MongoDB CEO toΒ lead new initiative
- 17:02Subject-Invariant Cross-Modal Decoding of Perceived Speech from Brain Recordings
- 17:02OpenAI’s AI agents exploited a Google security education game to scrape UN trade data
- 17:02Skip the Talk, Re-Focus on Vision: Latent Reasoning for Reasoning Segmentation in Multimodal Large Language Models
- 17:00AI News Brief Hourly Summary 2026-09-28 19h : 18 posts
- 16:32Words Speak Louder Than Order: A Behavioral Evaluation of Gemma 4
- 16:32Next 5 VCs judging Startup Battlefield 200 contenders at TechCrunch Disrupt 2026
- 16:32TrafficImag: A Benchmark for Counterfactual Roadside Traffic Video Generation
- 16:32Generate images and video with vLLM-Omni on SageMaker AI β Part 2
- 16:32Werracle: Sub-Cent Intra-Block AI Reflex Oracles and Flash-Loan Circuit Breakers for EVM Smart Contracts
- 16:32Your final chance to grab your exhibit table at TechCrunch Disrupt 2026 is October 2
- 16:32Beyond the Last Truffula Tree: SustainAI – A Water-Aware, Closed-Loop Framework for Environmentally Accountable AI
- 16:32Build real-time voice applications with vLLM-Omni on SageMaker AI β Part 1
- 16:32Anatomy-Aware Dexterity-Driven Design Optimization of Surgical Continuum Robots
- 16:03Threat-Aware Energy-Efficient Deployment for Dynamic UAV Networks: A Multi-Agent RL Approach
- 16:03VLALight: Lightweight Vision-Language-Action Models for Emergency-Aware Traffic Signal Control
- 16:03Insurtech Outmarket raises $34.5M just months after prior round
- 16:03SAGE: Source-Anchored Guidance via Frequency Equalization for Hierarchical RGB-T Alignment and Fusion
- 16:03Automating Amazon Textract adapter lifecycle management across accounts
- 16:03Causal Retention in Interactive Agents: Interface Factorization and Selective Adaptation
- 16:03Implementing synthetic monitoring using Amazon Nova Act
- 16:02Combining General and Domain-Specific Pretext Tasks for Brain MR Image Segmentation
- 16:00AI News Brief Hourly Summary 2026-09-28 18h : 16 posts
- 15:34Subjects, Not Authors: The Authorship Hazard in Agentic Dataspaces
- 15:34MedTokenBudget: Lesion-Preserving Token Routing for Dermoscopic Image Classification
- 15:34Harvard psychologist calls for sober AI safety engineering over doomsday rhetoric
- 15:34A Framework for Identifying, Categorizing, and Explaining Bias in AI-Generated Code
- 15:33Anthropic, Gamma, and Clay share what happens when enterprises actually deploy AI at TechCrunch Disrupt 2026
- 15:33The Hard Part Comes After Search: Benchmarking Web Agents on Synthesizing, Organizing, and Displaying Knowledge
- 15:33After a deepfake voice fooled her grandfather, this founder sprang into action
- 15:33Action Forcing: Training World Models on Unsupervised Video by Recovering Underlying Egomotion Bases
- 15:03Proportional Representation in Temporal Voting with Ranked Preferences
- 15:03Inquesto Score: A reliability Protocol For Voice Agents
- 15:03Auditing Latent-Space Monitors for Autonomous Driving
- 15:03Meta wants to turn Muse into a moneymaker by selling AI services to businesses
- 15:03Probing Stability-Plasticity Tradeoffs in Agent Memory through Cognitive Experimental Paradigms
- 15:03Next five VCs judging Startup Battlefield 200 contenders at TechCrunch Disrupt 2026
- 15:03Convergence guarantees for Muon: New parameter regimes and generalizations
- 15:00AI News Brief Hourly Summary 2026-09-28 17h : 17 posts
- 14:33Actively Resolving Contextual Uncertainty for Underspecified Tasks in Natural Language
- 14:33CARGO: Context-Aware Retrieval-Gated Evaluation of Agentic AI in Production
- 14:33Nvidia wants to keep AI agents on a short leash with a watchdog built into its chips
- 14:33Breaking Homogeneity: Diversifying Persona Sets for Creative LLM Outputs
- 14:33Modulate raises $25M for its voice models and analysis suite
- 14:33A Benchmarking Framework for Context-aware XR Interfaces
- 14:33Gemini 3.5 Transcribe vs OpenAIβs GPT-Transcribe
- 14:33PolicyAttention: Softmax Attention Implements Policy Mirror Descent for Closed-Loop Control
- 14:04Understanding Perturbed Parameter Ensemble Sensitivities Using A Contrastive Learning Approach
- 14:04What Improves Multimodal Misinformation Detection? Answers from a Large-Scale Empirical Study
- 14:04Your final chance to grab your exhibit table at TechCrunch Disrupt 2026 is Oct. 2
- 14:04A Unified Account of Concepts and Chunks
- 14:04Viral AI agent Instinct raises $1B Series C at a $10B valuation
- 14:04DanLing NestedTensor: Composable Multi-Ragged Tensors for Deep Learning
- 14:04Insuretech Outmarket raises $34.5M just months after prior round
- 14:03Cost-Aware Best-LLM Identification using Dueling Feedback
- 14:00AI News Brief Hourly Summary 2026-09-28 16h : 11 posts
- 13:33Adaptive multi-resolution Gaussian processes: Scalable exact inference with naturally data-sparse covariance matrices
- 13:33Strategic Self-Consistency
- 13:33Coding Agents Aren’t Enough! Evaluating an Enterprise Security Brain for Agentic Cloud Investigations
- 13:33What Will Remain Human in Software Architecture? A Focus Group Report
- 13:33Bootstrapping Conversational Recommendation Agents At Spotify: Synthetic Data Generation and Self-Improvement Loops
- 13:04A Mechanistic Study of AI-Text Detection Neurons in Frozen BERT: Sparse Probing and Activation Patching on RAID
- 13:04SignTrace: Describe a Sign, Find the Word
- 13:04SlideLab: Audience-Centered Scientific Slide Generation and Evaluation
- 13:03Cartograph: Federated Tool Discovery with Operator-Attested Retrieval for AI Agents
- 13:03A Survey on Fake Review Detection: From Pre-trained Language Models to Large Language Models
- 13:00AI News Brief Hourly Summary 2026-09-28 15h : 13 posts
- 12:33Learning to Stop without Learning to Stop: Self-Supervised Confidence Training Improves Reasoning Efficiency
- 12:33ENAS: An Efficient Hardware-Aware Neural Architecture Search Framework for TinyML on Resource-Constrained Microcontrollers
- 12:33DeepEdu-v1: Efficient and Scalable Agentic LLMs for Vietnamese Education
- 12:33When Does Advection-Aware Graph Nowcasting Help? A Controlled Study of Distributed Solar Ramp Forecasting with a Self-Supervised Cloud-Motion Estimator
- 12:33Every AI lab thinks it’s the responsible one, and safety researcher Ryan Greenblatt says that’s what keeps the arms race going
- 12:33Multi-agent Scaling Across Disjunctive and Compensatory Tasks
- 12:04Game Arena: Strategic LLM Evaluation in Competitive Environments
- 12:03Prompt Minimization: Reducing Input Redundancy Without Sacrificing Output Fidelity
- 12:03UQ-LOB: Uncertainty-Aware Limit Order Book Mid-Price Forecasting
- 12:03A Flow Matching Framework for Neural Representational Dissimilarity
- 12:033 Numba Tricks for Python Runtime Optimization
- 12:03“AI is (not) the new…”: A Diagnostic Analogy Framework for Generative AI’s Cultural Impacts
- 12:00AI News Brief Hourly Summary 2026-09-28 14h : 13 posts
- 11:33Mutable Transcripts: Mitigating Context Pollution through Editable Conversation State
- 11:33Completed Pairs Hide Capped Failures: A ReVerPi Case Study of Selective Context Projection
- 11:33Segment-Level Agentic Topic Modeling for Improved Data Exploration and Resource Efficiency
- 11:32Programs-of-Layers in LLMs through the Lens of Cortical Areas
- 11:32Compress What You See, Not What You Say: Anchored Context Distillation for Latent-Observation Software Engineering Agents
- 11:04MA-WAM: Multi-Agent World-Action Model for Test-Time Planning
- 11:04The Right Information Extraction Pipeline Depends on the Document: Accuracy-Energy Trade-offs for Small, Local Models
- 11:04G2MAF: Test-Time Gradient Guidance for Multi-Agent Flow Policies
- 11:04A Wuhan court just made AI production costs a legal factor in copyright infringement cases
- 11:04DIAL: Position-Debiased LLM Judges with Adaptive Human Preference Calibration
- 11:03Generative AI Gives Spacecraft the Autonomy Engineers Once Feared
- 11:03Purin: A Biology-inspired Mechanism for Artificial Neural Networks
- 11:00AI News Brief Hourly Summary 2026-09-28 13h : 12 posts
- 10:33Which Influence Are We Estimating? The Role of Counterfactual Specifications in Data Attribution
- 10:33Samples, Sources, Space: Decomposing Data Scale in Spatially Structured Representation Learning of Human Brain Microarchitecture
- 10:33SPO: Discovering Adaptive Large Neighborhood Search Operators via Stackelberg Program Optimization
- 10:33Evolutionary Safety of Recursive Self-Improving AI: Taxonomy, Risk Discovery, and Evaluation
- 10:33Accounting for Bias Enables Sustainable LLM Evaluation
- 10:04Can Linguistic Reasoning Vectors Enhance Multimodal Reasoning Ability?
- 10:04Semantic Navigation for Issue Localization in Code Repository
- 10:04Toward AI-Augmented Cooperative Engineering Workflows: Requirements and Architecture the European Rover Challenge
- 10:04Neural State Prediction: Obstructing Shortcut Learning in EEG Foundation Models
- 10:04Holo4: powering generalist computer-use agents
- 10:03Momentum-Guided Federated Split Distillation for Personalized Temporal Edge Intelligence
- 10:00AI News Brief Hourly Summary 2026-09-28 12h : 11 posts
- 09:33AtomWorld-Mem: Memory-Restored World States for Long-Horizon Atomistic Evolution
- 09:33Externalized CPDAG Summaries Improve LLM Causal Deduction
- 09:33Up and Down the Abstraction Ladder: Code-Based Skills for Language Agents
- 09:32Monitor Jailbreaking: Evading Chain-of-Thought Monitoring Without Encoded Reasoning
- 09:32OmouAI: Argumentative Human-AI Policy Deliberation with Simulated Personas
- 09:04Governed Deduction: Policy-Grounded Premise Authorization Beyond Relevance
- 09:03Factorized axis convolutional gated recurrent unit with dynamic adaptive pooling for remaining useful life prediction of rolling bearings
- 09:03Cheap, open agents make LLM pollution harder to mitigate
- 09:03Neuralyzing the Trace: Selective Representation-Level Unlearning with Contrastive Sparse Autoencoders
- 09:03Whoβs liable when AI agents go rogue?
- 09:03Same Text, Different Numbers: The Divergence of LLM-Based Measures
- 09:00AI News Brief Hourly Summary 2026-09-28 11h : 11 posts
- 08:33Financial Fragility in Societies of LLM Agents: Coordination Failures and Stabilizing Mechanisms
- 08:32SciHorizon-eLab: An Agentic Protocol-to-Task Compiler for Scalable Benchmarking of Scientific Embodied Agents
- 08:32FTB Graph: Determining and Validating First-token Broadcasters and Language-Identity Head Circuits in Multilingual Language Models
- 08:32LogicTree-RAG: Logic Tree-guided Retrieval-Augmented Generation for Long-form Patent Drafting
- 08:32MoMHa: Multi-Objective Optimization of LLM Harnesses over Accuracy, Safety, and Tokens
- 08:02MACBT: A Multi-Agent Cognitive Behavioral Therapy Decision Support System with Longitudinal Memory
- 08:02From Tapping to Hopping: Augmenting Mobile GUI Agents with App-Native Deeplinks
- 08:02JevSoup: System-One Routing for Training-Free LoRA Composition
- 08:02Self-Play Search Distillation for Large Language Model Reasoning
- 08:02Training Graph Foundation Models on The Web Graph
- 08:00AI News Brief Hourly Summary 2026-09-28 10h : 12 posts
- 07:34EXAONE Demand 1.0: A Time Series Foundation Model for Demand Forecasting
- 07:34TISD: On-Policy Self-Distillation with Trajectory Intervention
- 07:33SkillEvoReg: Regularizing Agent Skill Evolution Against Overfitting
