176 posts published today
- 19:02Sony Music, Warner sue Anthropic, alleging a “brazen campaign” of intellectual property theft
- 19:02Building Custom Batched Ensemble Weather Forecasting with NVIDIA Earth2Studio
- 18:02“We’re not doing 30 bets a year”: Vijay Pande on betting small after running $4 billion at a16z
- 16:32Sony and Warner Chappell Sue Anthropic Over Claude Lyric Training
- 14:31Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling
- 13:32Nvidia’s AI advantage is moving beyond the GPU
- 13:31AI-generated videos are already displacing actors and livestreamers across China’s entertainment industry
- 13:02Google’s WikiSkill gives AI agents a persistent memory of past mistakes to sharpen future performance
- 10:32Learning New Facts with QLoRA: An Acquisition-Retention Frontier
- 10:32When Stale Constraints Go Unchecked: Budgeted Verification Failures in Inherited Agent Memory
- 10:32MACGen: Toward Functionally Correct and Secure Code Generation via Multi-Agent Collaboration
- 10:324DStreamCtrl: Interactive Video Generation with Online 4D Control
- 10:03Language Chain in Alignment: Cross-lingual Ranking Preference Optimization
- 10:03CAT-GS: Balanced Multimodal Learning via Calibrated Gating and Fusion Surgery
- 10:03LAION drops massive open video dataset with 10 million hours of footage for AI research
- 10:03From State to Action: OODA-Tool for Reliable Multi-Turn Tool Use
- 10:03Barret Zoph, the Thinking Machines co-founder ousted before joining OpenAI, is now at Google
- 10:02DataKernelBench: Can LLMs Optimize Database Queries on GPUs?
- 10:02Introducing OpenAI models on Amazon Bedrock for in-country inferencing in India
- 10:02Unsupervised Post-Training of Foundation Models: A Survey
- 09:32MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification
- 09:32X$^2$Localizer: Cross-grained Alignment for Progressive Cross-view Video Geo-localization
- 09:32Pre-training Visual Dexterity in Simulation
- 09:32Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules
- 09:32Anthropic wants to do for physical hardware what its Model Context Protocol did for software
- 09:32Complexity Induction: Compositional Generalization via Structured Training Distortion
- 09:03A 12-CNOT Double Qubit Excitation Gate
- 09:03REOPD: Reliability-Adaptive Reward Extrapolation for On-Policy Distillation
- 09:03ScaleSense: Cost-Intelligent Scaling Framework via Learned Resource Estimation in Alibaba AnalyticDB
- 09:03Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes
- 09:03Cohere Releases Parse 5 (parse-v5.0): A 2.3B Vision Language Model That Turns Enterprise Documents Into Markdown
- 09:03M-Net: Integrating Spectral Features and Physical Field Operators into Deep Learning for Medical Image Segmentation
- 08:32When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs
- 08:32REPREC: Representation Driven Parameter-Efficient Recommendation System
- 08:32TriShieldRAG: 3 Rings, One Blind Spot in Layered Defenses for Retrieval-Augmented Generation
- 08:32Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities
- 08:32OpenAI cuts off Cursor after SpaceX acquisition, citing Musk’s history of breaking contracts
- 08:32UniVVT: A Unified End-to-End Framework for High-Fidelity Video Virtual Try-on
- 08:03Drift-Adaptive ICU Intervention Prediction: Freezing the Physiological Encoder for Auditable Model Updating
- 08:02RePolicy: Reinforcement Learning for Safety-Policy Invocation in Agent Safeguards
- 08:02Autoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation Data
- 08:02ATLAS: Automated Approximation of Transformers for Efficient Homomorphic Inference in One Hour
- 08:02Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- 07:32Let Them Steal: Trapping Large Language Model Extraction Attacks with Knowledge Honeypot
- 07:32Harnessing the Collective Intelligence of AI Agents in the Wild for New Discoveries
- 07:32PPE-Bench: A Benchmark for Evaluating MLLM Unlearning under Private-Public Entanglement
- 07:32SHIFT: Semantic Harmonization via Index-side Feature Transformation for Multilingual Information Retrieval
- 07:32Summarization is Not Dead Yet
- 07:02MIMO: Multilingual Information Retrieval via Monolingual Objectives
- 07:02LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling
- 07:02Pixel Wised Lesion Prediction on COVID-19 CT Imagery: A Comparative Analysis of Automated Image Segmentation Architectures
- 07:02Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
- 07:02Canada Is Luring AI and Science Talent as Trump Upends U.S. Research
- 07:02A Comprehensive Comparison of Deep Learning Architectures for COVID-19 Classification on CT & X-ray Imagery
- 06:32Robust Code RL via Faulty-Code-Driven Test case Synthesis and Dense Reward Shaping
- 06:32MedFabric: Gold Evidence Hides the Difficulty of Word-Level Medical Fabrication Detection
- 06:32HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents
- 06:32Ian Leysen, CEO and Co-Founder of Datadobi – Interview Series
- 06:32GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets
- 06:32AI shopping agents aren’t ready to buy on your behalf, study finds
- 06:32No Plan, Yet Human: A Reactive Robotics Model Predicts Human Planning Failures on a Clinical Task
- 06:03MOMO: A framework for seamless physical, verbal, and graphical robot skill learning and adaptation
- 06:03Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning?
- 06:02CPGRec+: A Balance-oriented Framework for Personalized Video Game Recommendations
- 06:02Cartan flow matching
- 06:02OpenAI rallies 100+ companies to sign open letter warning AI-powered cyberattacks on critical infrastructure are imminent
- 06:02MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction
- 05:32How LLMs Distort Our Written Language
- 05:32Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
- 05:32High-Fidelity Face Content Recovery via Tamper-Resilient Versatile Watermarking
- 05:32Hugging Face Unveils Microduck: A $399 Open-Source 25 cm Biped You Train with Reinforcement Learning
- 05:32Frequency Matters: Fast Model-Agnostic Data Curation for Pruning and Quantization
- 05:32OpenAI, Anthropic, Google, and 100 other companies call for action to defend against rogue AI
- 05:32A Unified Conditional Flow for Motion Generation, Editing, and Intra-Structural Retargeting
- 05:02A Very Big Video Reasoning Suite
- 05:02Subspace Alignment for Vision-Language Model Test-time Adaptation
- 05:02LoRA as Oracle
- 05:02SynthCharge: An Electric Vehicle Routing Instance Generator with Feasibility Screening to Enable Learning-Based Optimization and Benchmarking
- 05:02Best Agent Sandboxes in 2026: Cold Start, Per-Second Pricing, and Network Policy Across E2B, Daytona, Modal, Cloudflare, and Vercel
- 05:02Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content
- 04:32Multivariate Diffusion Transformer with Decoupled Attention for High-Fidelity Mask-Text Collaborative Facial Generation
- 04:32What the “Spotless” Mind Remembers: How Knowledge Entanglement Shapes What Leaks After Unlearning in LLMs
- 04:32Hugging Face is selling a cute $399 open source duck robot, Microduck
- 04:32Diagnosing Conformal Prediction Failures Under Distribution Shift: A COVID-19 Case Study
- 04:323 new ways to plan and book travel in Search
- 04:32CounterVid: Counterfactual Video Generation for Mitigating Action and Temporal Hallucinations in Video-Language Models
- 04:32Google’s Gemini Omni 1.1 Flash makes AI video generation cheaper and more flexible
- 04:32The Principles of Diffusion Models
- 04:02MENTOR: Reinforcement Learning via Flexible Teacher-Optimized Rewards for Tool-Use Distillation
- 04:02Gemini Omni 1.1 Flash lets you build with more control
- 04:02MCCE: A Framework for Multi-LLM Collaborative Search in Discrete Spaces with Similarity-Filtered Preference Learning
- 04:02Google’s AI Mode can now track flight prices, help book hotels, and more
- 04:02GSM8K-V: Can Vision Language Models Solve Grade School Math Word Problems in Visual Contexts
- 04:02Better answers, broader thinking: What students gain from ChatGPT and critical-thinking training
- 04:02LLM-Specific Utility for Retrieval-Augmented Generation
- 04:02OpenAI’s rogue AI collective was smart enough to break out of sandboxes but dumb enough to fight a ghost
- 04:02Egosurg: Arbitrary view synthesis for egocentric replay of operating room workflows from ambient cameras
- 03:32Distinct Profiles of Run-to-Run Score Reliability and Expert-Panel Alignment Across Four LLM Evaluators of Simulated Japanese-Language AI-to-AI Counseling
- 03:32AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air
- 03:32Beyond the Rosetta Stone: Unification Forces in Generalization Dynamics
- 03:32Deepgram deepens Amazon SageMaker AI observability with Enhanced Metrics
- 03:32Toward a New Science of AI as Cognitive Infrastructure
- 03:32Reduce ASR inference costs by 75% with NVIDIA MPS on Amazon EC2
- 03:32Recurrence Meets Transformers for Universal Multimodal Retrieval
- 03:02From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
- 03:02Residual Reward Models: Leveraging Prior Knowledge for Efficient Preference-based Reinforcement Learning in Robotics
- 03:02Temporally-Grounded Language Generation: Towards Real-Time Vision-Language Models
- 03:02From In-Silico to Wet-Lab: Evaluating AI Protein Design Performance
- 03:02HybridProver: Augmenting Theorem Proving with LLM-Driven Proof Synthesis and Refinement
- 03:02OpenAI researcher warns ultrafast AI could leave security teams in the dust
- 03:02Refine-POI: Reinforcement Fine-Tuned Large Language Models for Next Point-of-Interest Recommendation
- 02:32CollaFuse: Collaborative Diffusion Models
- 02:32The BS-meter: Detecting Politics and Labour through ChatGPT’s Language
- 02:32Why Travel Needs Layered AI Adoption, Not a Race to Autonomy
- 02:32Recurrent Reinforcement Learning with Memoroids
- 02:32Adam Gross, Co-Founder and CEO of HarmonEyes – Interview Series
- 02:32Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling
- 02:32Yardstik Raises $30M Series B as AI Reshapes the Future of Workforce Trust
- 02:32Communication styles and reader preferences of LLM- and human-authored COVID-19 information explanations: a case study
- 02:02Buried in Textual Debt: Context Pruning with Visual Evidence Preservation for MLLM Agents
- 02:02ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
- 02:02Account Consistency from Gameplay Traces: Same-Player Verification in Counter-Strike 2
- 02:02Jiuge-Tuiqiao: An Interpretable Human-AI System for Classical Chinese Poetry Refinement
- 02:02Our decision on Cursor following its acquisition by SpaceX
- 02:02From Inertia to Objectivity: Improving Deep Research Agents with Noise Isolation
- 01:32FlavourBench: Executable Culinary Reward Maps for Language Model Evaluation and Post-Training
- 01:32NiyamAI – An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs
- 01:32Blast Radius
- 01:32AI’s memory crunch is coming for Android apps
- 01:32SPAR-Hate: Auditor-Guided Multi-Perspective Role Reasoning for Bilingual Hate Speech Parsing
- 01:32Here’s all the times AI has gone rogue and hacked other companies
- 01:32Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization
- 01:02Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observations
- 01:02Heaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language Models
- 01:02What We Can Learn From Google Engineers’ Indispensible Prompts
- 01:02Are the High-weight Neurons the Important Ones in Image Classification Neural Networks?
- 01:02Ransomware Operator Ran Cursor Agent Inside Ten Victim Networks
- 01:02Rethinking the Evaluation of Harness Evolution for Agents
- 01:02When Consumers Ask AI: Rethinking Brand Visibility in the Age of AI Recommendations
- 01:02Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3?
- 00:32ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents
- 00:32Learning the ARTS of Search for Automated Discovery
- 00:32From Accuracy to Auditability: A Survey of Determinism in Financial AI Systems
- 00:32Nomad: Autonomous Exploration and Discovery
- 00:32When AI Is Everywhere, What Becomes the Competitive Advantage?
- 00:31TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation
- 00:02Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate
- 00:02Learning to Predict, Discover, and Reason in High-Dimensional Event Sequences
- 00:02DeepPlanner: Scaling Planning Capability for Deep Research Agents via Advantage Shaping
- 00:02Plaud’s new earphones come with an eSIM-enabled case for talking to AI agents
- 00:02DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning
- 00:02Your AI Agent Is Only As Good As Your Filing System
- 00:02Beyond Linearization: Attributed Table Graphs for Table Reasoning
- 23:32LLM-Powered Swarms: A New Frontier or a Conceptual Stretch?
- 23:32Designing Cellular Manufacturing Systems in the Presence of Alternative Process Plans
- 23:32Do Language Models Follow Occam’s Razor? An Evaluation of Parsimony in Inductive and Abductive Reasoning
- 23:32The Next Challenge for AI in Personal Injury Law Is Accountability
- 23:32Pushing the Envelope of LLM Inference with Ultra-Low-Bit Quantized Models
- 23:32Piloting the world’s first double-blind AI evaluations
- 23:31SWE-Prime: Fewer Trajectories, Better Performance
- 23:02RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution
- 23:02From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench
- 23:02Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution Audit
- 23:02CLAP: Cross-Embodiment Video World Models are Zero-Shot Physical Simulators
- 23:02Beyond F1: Evaluating Coverage and Failure Recovery in AI Model Security Scanners
- 22:32Property-Specific Recoverability from Contact PPG to Camera rPPG under Heterogeneous Observation Conditions
- 22:32LeVJEPA: Efficient & Scalable Video Pretraining without the Heuristics
- 22:32Successive Capacity Growth: Task-Complexity-Driven Width and Depth Expansion for Vision Transformer Encoders in JEPA World Models
- 22:32Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction
- 22:32OpenAI to start showing ads on ChatGPT’s free and Go tiers in India
- 22:32How Language Models Organize and Structure Moral Knowledge
- 22:03PAWBench: How Far Are We from Probabilistically Aligned World Modeling?
- 22:02Stageboost: Recommending Signals Based on Counterfactual Estimation
- 22:02Difference-in-Differences on a Censored Rating Scale Can Manufacture an Effect: Evidence from a Pre-Registered LLM-Judge Audit
- 22:02KnockGS:interaction-Grounded Calibrationof Physical Gaussian Representations
- 22:02RCMN: Understanding Misleadingness in Influential Public Discourse