200 posts published today
- 21:32Leveraging Low-Level Symbolic Competences for Unsupervised Grounding in Hallucination Detection
- 21:32How do LLMs Evaluate Perceived Moral Agency? Investigating Moral Decision-Making in Human-Artificial Agents Interactions
- 21:32Amortizing Scaling Law Construction Costs
- 21:32VICAL: Vicinal Consistency Alignment for Long-Tailed Visual Recognition
- 21:32How a Chatbot’s Response Style Shapes a Classroom: A Multi-Agent Simulation of Students Consulting AI
- 21:02ARIA – An Agentic Framework for Autonomous Testing of Infotainment Systems
- 21:02Better Understanding, Better Fixes? A Study of Hallucination in LLM-based Automated Program Repair
- 21:02MCPO: Modality-Contrastive Preference Optimization for Multimodal Chain-of-Thought Compression
- 21:02One Diffusion Model, Two Roles: Guided Trajectory Planning and Safety-Critical Scenario Generation in Closed-Loop Simulation
- 21:02TreeFI: Value-Aware Statistical Fault Injection for Deep Neural Networks
- 21:00AI News Brief Hourly Summary 2026-09-07 23h : 11 posts
- 20:32RefactorPlatform: An Open-Source Harness for Controlled Evaluation of Repository-Scale Refactoring Agents
- 20:32Methane Detection On Board Satellites from Unorthorectified Imagery
- 20:32Sound-based Multi-Person 3D Pose Estimation
- 20:32Attention-guided super-resolution of 4D flow MRI in carotid arteries
- 20:32Adaptation Interfaces for In-Context Tabular Foundation Models in Time-to-Event Prediction
- 20:03PRISM-Bench: An Audio-Centric Diagnostic Benchmark for Text-to-Audio-Video Generation
- 20:02Forgetting Without Restarting: Execution-State Unlearning for Stateful LLM Agents
- 20:02SimFuse3D: Source-Guided Target Simulation and Confidence-Guided Multi-Stage Localization Reweighting for Cross-Platform 3D Object Detection
- 20:02Mitigating Performance Discrepancy in Cross-Domain 3D Class-Incremental Learning
- 20:02ReCAST: Restoration-aware Cascaded Stage-wise Training for Obfuscated SMS Risk Classification
- 20:00AI News Brief Hourly Summary 2026-09-07 22h : 14 posts
- 19:32CC-Mediation: Evaluating Large Language Models for Cross-Cultural Conflict Mediation
- 19:32MABPD: Multi-Agent Bias Probing & Detection via Structured Argument Debate
- 19:32Cost-Aware Hierarchical Multi-Agent Ransomware Detection and Family Attribution
- 19:32OpenBMB Releases MiniCPM5-2B: A 2.52B Dense Model Averaging 53.9 Across 34 Benchmarks and Built to Run On Device
- 19:32MMTClinic: Multimodal, Multilingual Time Series Question Answering and Reasoning Benchmark for Clinical Domain
- 19:32Opaque recurrence, and other AI terms that you should probably know
- 19:32Reinforcement Learning for improving Large Language Models’ Catalan text simplification capabilities
- 19:03Linguistic Trajectory Encoding for Efficient Long-Horizon Spatial Memory in Embodied Agents
- 19:03Recurrence Is Not Enough: Causally Validating Multilingual SAE Translation Features in Gemma 2 and 3
- 19:03Persistent Teacher Anchoring for Tool-Using Agents
- 19:03Can Activation Steering Capture Multidimensional Authorship Style?
- 19:03Axis Robotics Releases AXIS: A Browser-Based Data Engine With 207 Robot Manipulation Tasks and 50,129 Trajectories
- 19:03Dynamic Heterogeneous Graph Representation Learning: A Survey
- 19:00AI News Brief Hourly Summary 2026-09-07 21h : 13 posts
- 18:32Simulation-free Unbalanced Dynamic Optimal Transport with General Growth Penalty
- 18:32Building a research-software catalog with a coding agent: from hackathon prototype to public deployment
- 18:32When Does an Interpretation Count as Established? The Formation, Evaluation, and Responsibility of Interpretation in Generative AI
- 18:32Anthropic reportedly signs $517 billion in compute deals after Dario Amodei warned rivals about reckless risk
- 18:32Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing False Refusals in Language Models
- 18:32XDOF, just 3 months out of stealth, is in talks for a Series B at a $1.2B valuation
- 18:32Knowing What Not to Answer: Selective Non-Compliance in Vision-Language Models
- 18:03SCAPES: Semantically Conditioned Autoregressive Prior for Environmental Sounds
- 18:03Tracing Audio Grounding and Answer Selection in Audio LLMs
- 18:03Beyond Code Generation: Reliability, Verification, and Cost Economics in the Agentic Software Development Lifecycle
- 18:03Wireless Foundation Models: State-of-the-Art and Open Challenges
- 18:02GPT-6 Astra beat Portal start to finish without human help in under 24 hours
- 18:02Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion
- 18:00AI News Brief Hourly Summary 2026-09-07 20h : 13 posts
- 17:32PetQA: Benchmarking Veterinary Knowledge and Clinical Reasoning
- 17:32Dynamic Adaptation of the LLM Context for Generating Routines with Coupled Semantics
- 17:32Dual-Part Multi-Lateral Branched Network for Multi-Class Segmentation in Cardiovascular Catheterization Angiograms
- 17:32Training-Free Halving of Activated Experts in Fine-Grained Mixture-of-Experts Models
- 17:31When Do Internal Probes Beat Reading the Answer? Miscalibrated Readouts and Behavior-Concealed Knowledge in Language Models
- 17:02Pitch-class Steering for Diffusion-based Music Generation via Latent-space Probes
- 17:02Atlas: Optimizing Deployment of Compound AI Workflows on Heterogeneous Clusters
- 17:02A Semantic Model of Genetic Evidence: A Step Toward Bridging the Basic-Science-Clinic Gap
- 17:02ChatGPT claws back web traffic share to 55.5 percent as Gemini’s brief comeback fades
- 17:02Continual Field-Adaptive Models (CFAMs) for Post-Deployment Physical AI
- 17:02AI-designed drug appears to turn back the body’s biological clock in early trial
- 17:02Repeat-After-Me: Black-Box Adaptive Visual Prompt Injection
- 17:00AI News Brief Hourly Summary 2026-09-07 19h : 13 posts
- 16:54Daily Roundup
- 16:32Towards Understanding Pause Token Fine-Tuning Dynamics: A Mode Retention Perspective
- 16:32Patterns of Priming in Production: Lexical, Semantic and Structural Alignment in Language Model Generation
- 16:32Shared circuits predict whether LLMs generalize across formats in arithmetic reasoning
- 16:32Cultural Misalignment in Large Language Models: Detection, Measurement, and Mitigation Through Targeted Fine-Tuning
- 16:32Hakken: Predicting future discoveries to fill the gaps in today’s knowledge
- 16:03A Systematic Evaluation of Cross-Lingual Consistency Enhancement Methods in Multilingual Language Models
- 16:03GRACE: Graph-Grounded Reflective Agent Copilot Engine for Expert-in-the-Loop Knowledge Expansion
- 16:03REFINE: LLM Refinement over Budgeted Text-Attributed Graphs for Personalized Medical Concept Representation
- 16:03When Load-Balancing Goes Too Far: Expert Pruning in Over-Dispersed Mixture-of-Experts Models
- 16:03Matt Clifford Steps Down as ARIA Chair After Anthropic Move
- 16:03A Roadmap for MEG Foundation Models
- 16:00AI News Brief Hourly Summary 2026-09-07 18h : 12 posts
- 15:33Cross-modal triage network: a multimodal deep learning framework for severity-based triage and visual explainability in chest radiographs
- 15:33You Really Didn’t Get That? Benchmarking Social Pragmatic Inference for Indirect and Playful Chinese Online Comments
- 15:33Where Appearance Fails, Geometry Recognizes: A CAD-Free 3D Shape Prior That Complements Vision Foundation Models
- 15:32Ultrasound-Based Prediction of Cirrhosis Decompensation Using Large-Scale Computer Vision Models
- 15:32Grupo Financiero Inbursa Adopts Harvey Across Its Legal Organization
- 15:32What Moves? Localized Motion Representations for Compositional Scene Control
- 15:03Data-Driven Learning of Unknown Nonlinear Differential Equations Using Functional Analysis
- 15:03Abstraction Agent
- 15:03Adapting from Downturns: Prediction of Long-Term Conversational-Skill Development in Mental-Health Crisis Counselors
- 15:03VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models
- 15:03Blockchain-Enabled Secure Logging for Fiscal Electronic Mechanisms: Evaluation of the Greek eSEND and myDATA Tax Systems
- 15:00AI News Brief Hourly Summary 2026-09-07 17h : 11 posts
- 14:32A Deep Generative Model for Synthesizing Labeled Wireless Signals
- 14:32Scalable Context Orchestration for Serving LLMs Over Voice
- 14:32Evidence Integration in Large Language Models
- 14:32When Seeing Overrides Knowing: Visual Dominance and Deferral-Based Method for Personalized Safety in VLMs
- 14:32AlcaTRAz – Anchored Tree-Rule Defense Against Jailbreaks
- 14:02Molecular D\’ej\`a Vu: Digit-Level Retrieval of Published Values in Frontier Language Models
- 14:02CUA-Universe: A Scalable and Dynamic Environment for Hybrid GUI+CLI Agents
- 14:02Multi-Step Tool-Calling over Korean Open Public APIs: A Benchmark and a Data-Synthesis Recipe
- 14:02Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence
- 14:02Who Should Grade My Work? Student Perspectives on Transparent AI-Assisted Writing Assessment in Higher Education
- 14:00AI News Brief Hourly Summary 2026-09-07 16h : 17 posts
- 13:339kV Solid-State Transformer Anchors Hyosung’s U.S. AI Grid Push
- 13:33LLM-Driven Algorithm Design for Quantum Circuit Synthesis based on Binary Decision Diagrams
- 13:33New York City bans AI tools from public schools through eighth grade
- 13:33Technical Manual for a Toolkit for Measuring Contextual Individuation in Transformer Language Models
- 13:33At UBS, AI skills are now a condition for landing a job
- 13:33Does Your Agent’s Memory Survive a Model Upgrade? A Controlled Study of Memory Portability
- 13:33OpenAI reports AI “research interns” and warns about its own pace at the same time
- 13:32Large Language Models for HVAC Operations in Building Energy Systems: A Critical Review of Methods, Applications, and Deployment Readiness
- 13:32How AI wiped out an entire industry in Nairobi
- 13:32RISE: Recursive Improvement via Self-Extrapolating Policy Distillation
- 13:03GUT: Quantifying and Optimizing the Reasoning Uncertainty of LLMs via Graph Complexity
- 13:03Beyond Aggregate Scores: Behavioral Correctness Assumptions for Assessing Reference-Based Automatic Evaluation Methods
- 13:03Testing Interchangeability in LLM Agent Teams
- 13:03Don’t Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference
- 13:03MG Ship adds AI route optimisation as logistics returns accelerate
- 13:03AI for Computational Design Science: A Responsible Human-AI Framework and Case Study on Short-Form Video Safety Surveillance
- 13:00AI News Brief Hourly Summary 2026-09-07 15h : 12 posts
- 12:33Commonsense Reasoning in Computer Vision: Foundations, Recent Advancements, and Future Directions
- 12:33A Unified Physics-Aware Quantum Machine Learning Framework across Power GaN HEMTs and Logic Nanowire FETs: Predicting Unseen Process Splits and Held-Out Geometry Combinations with Lower Error and Tighter Split-to-Split Variability
- 12:33Uncensored Open-weight Models: Redistribution as the Persistence Layer
- 12:33Trace2Tower: Transition-Aware EigenTrace Induction of Multi-Level Skills for LLM Agents
- 12:33Qwen-Drive 1.0 tells you why it brakes, just don’t expect the explanation to match the maneuver
- 12:33Do LLMs Exhibit Coherent Knowledge Structures in Mathematical Reasoning? A Perspective from Knowledge Space Theory
- 12:04Substrate-Aware AI Agents: Execution Context as a First-Class Input
- 12:04The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior
- 12:04CABAL: Multi-Agent Simulacra for Tracing the Effects of Collusive Bidding in Peer Review
- 12:04ACE: Adaptive Calibration-Free Expert Skipping for MoE-based LLMs
- 12:04What Matters in On-Policy Distillation? A Perspective on Data Efficiency and Data Selection
- 12:00AI News Brief Hourly Summary 2026-09-07 14h : 10 posts
- 11:33SciDocBench: A Workflow-Centered Benchmark and Data Pipeline for Scientific Document Understanding
- 11:33A Hybrid Predictive Ensemble of Machine Learning and Deep Neural Networks for Early Cardiovascular Disease Risk Assessment
- 11:33ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding
- 11:33Compact Bellman-Grounded Cognitive Maps for Cost-Aware Navigation
- 11:33Unifying ICL, SFT, KL-Regularized RL Through a Bayesian Lens
- 11:03TruthInsightBench: An Evidence-Grounded Benchmark for Automated Evaluation of Open-Ended Scientific Discovery Agents
- 11:03Constructing and Evaluating Clinical Reasoning Trajectories for Medical Agent
- 11:03LLM-Guided Program Evolution for Circle Packing: Breaking 10 Packomania Records for $28
- 11:03Measuring AI Accountability Through Argumentation Analysis: Can Model Reasoning Withstand Scrutiny?
- 11:03MePo++: Unifying Representation Refinement and Reconciliation for General Continual Learning
- 11:00AI News Brief Hourly Summary 2026-09-07 13h : 13 posts
- 10:33Language models judge war differently when tested for alignment
- 10:33Towards Efficient Evaluation of Evolutionary Transfer Optimization: Case Studies on Task-Parameterized Applications
- 10:33Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent Alignment
- 10:33A Tree-based RAG Framework for Evidence-Intensive QA via Adaptive Planning and Topology-Aware Evidence Gathering
- 10:33Proteomic Aging Clocks Track Biological Age Reversal in Rentosertib Trial
- 10:33TROVE: Adaptive Agent Skill Orchestration via Trace-Grounded Route Validation and Editing
- 10:03Compact-Memory LLM Agents via Online Max-Member Clustering and Atom-Aware Packing
- 10:03Artificial Intelligence in Equity and Crypto Markets: Progress, Profitability Evidence, and the Limits of Automated Investing
- 10:03Why We Care About Understanding: Competence through Predictive Compression
- 10:03Global to Local: Topology-Preserving Adaptive Graph Pooling via Granular-Ball
- 10:02UST Completes Majority Stake Acquisition of Italdesign From Audi
- 10:02Solving Hard XAI Queries Based on a Compiled Dual-Rail Encoding
- 10:00AI News Brief Hourly Summary 2026-09-07 12h : 11 posts
- 09:33AutoLR: Automating the Path from Research to Launch Review in Industrial Recommender Systems
- 09:33From Language Models to World-Acting Systems: Progress and Limits of Agentic AI across Digital, Social, Virtual, and Physical Environments
- 09:33Reinforcement Learning for Sequential Solar PV Policy Design under Uncertainty: An Agent-Based Approach
- 09:33CHAMP: Cross-domain Hybrid Architecture for Matchmaking and Prediction in Online Multi-Player Games
- 09:33MARLA: A Conceptual Scaffold for Regulatory Learning under the EU AI Act
- 09:03MZ-Rain: Moisture-Budget-Guided Zero-Inflated Model for Station-Level Precipitation Nowcasting
- 09:03CoSkill: Joint Reinforcement Learning of Reasoning and Meta-Skill Agents for Hierarchical Skill Evolution
- 09:03MM-IFEval-Pro: A Multilingual and Attack-Resistant Benchmark for Instruction-Following in Vision-Language Models
- 09:03LLM-Assisted Behavioural and Scenario Augmentation for Agent-Based Energy Adoption Models
- 09:03From Interaction Traces to Persistent Skills: Online Evolution for Computer-Use Agents
- 09:00AI News Brief Hourly Summary 2026-09-07 11h : 13 posts
- 08:34MedFlow: Class-Aware Multi-Scale Generation for Medical Time-Series Synthesis
- 08:33Long Horizon Transformer Quantile Fault Prediction for Multi Site Industrial Predictive Maintenance
- 08:33When Financial Fine-tuning Fails: A Three-Level Detectability Analysis of Numerical Hallucination in Domain-Adapted Language Models
- 08:33ElderBench: Benchmarking Autonomous Mobile Agents for Older Adults
- 08:33Google, Cathay Pacific Expand Contrail Avoidance Trials in Asia-Pacific
- 08:33CPR-IE:A Compression-Prediction-Resource Intelligence Efficiency Metric
- 08:04Diffusion Language Models for Mobile Edge Agentic AI: Foundations, Applications, and Challenges
- 08:04Whose record is this? Diagnosing and authorizing record use in personalized multimodal models
- 08:04ProtLingo: Efficient Protein Language Modeling via Conditional Memory and Expert Routing
- 08:04Hierarchical Possession-Aware Graph Pointer Network for Pass Receiver Selection
- 08:04Supporting independent journalism in Ukraine
- 08:04DODR: Deterministic Operator-Driven Reasoning in Latent Space
- 08:00AI News Brief Hourly Summary 2026-09-07 10h : 11 posts
- 07:33PLUME: Parameter-Efficient Personalization of Large Language Models via Low-Rank User Modulation in Shared Subspaces
- 07:33Shadow Queries for Private Retrieval in Vector Databases
- 07:33Aplaud: Adaptive Personalized Low-Rank Decomposition for User-Specific LLM
- 07:33DCFA: Dual-view Causal-inspired Attribution for Failure Reasoning in LLM-based Multi-agent Systems
- 07:33FinalityBench: An Effect-Level Benchmark for Agent Decisions Under Delayed and Conflicting Financial Finality
- 07:03Model Retirement Creates Reproducibility Risk in Biomedical AI Publications
- 07:03Predicting Spatiotemporal Mobile Sensing-Based PM2.5 Concentrations Using Low-Rank Adapted Spatially Attentive Graph Neural Network
- 07:03Train What You Deploy:Token-Faithful Post-Training of a Production Coding
- 07:03SQL-Zero: Self-Evolving Text-to-SQL
- 07:03ERPBench: Evaluating LLM Agents for Enterprise Decision-Making Across Competitive Market Ecologies
- 07:00AI News Brief Hourly Summary 2026-09-07 09h : 11 posts
- 06:32Leveraging Imperfect Restoration for Data Availability Attack
- 06:32Harness-agnostic detection and immunization of reward hacking in self-evolving language models
- 06:32A Cost-Aware Agentic Architecture for NL-to-SQL over Nested Enterprise Schemas, with a New Benchmark
- 06:32SiLR: Structure-Preserving Admission and Process Reward for LLM Tool Agents
- 06:32NEURA and SECO Partner on Robot Compute Modules Built in Europe
- 06:32Continual Graph Memory for Adaptive Recommendation under Intent Drift
- 06:03Extremely Sparse Supervision Incentivizes Reasoning Ability
- 06:03Reducing Hallucinated Transcripts in Whisper via Hallucination Space Projection
- 06:03Does the Selected Object Reach the Reader? Auditing Identity Handoffs in Grounded Language-Model Pipelines
- 06:03La Agente \’Optima: Towards Agentic Self-Driving Laboratories
- 06:03$\tau^\tau$-Bench: An Environment for End-To-End, Realistic Agent Construction
- 06:00AI News Brief Hourly Summary 2026-09-07 08h : 12 posts
- 05:32MaxKernel: Agentic Kernel Generation for TPUs
