AI agents on foundation models often misapply healthcare and life sciences decision frameworks, citing the right guideline but applying it incorrectly.…
Category: AI
What Do We Expect from LLMs? Mapping the Design of LLM Benchmarks
arXiv:2609.19182v1 Announce Type: new Abstract: Benchmarks are central to how progress in large language models (LLMs) is assessed and communicated. Yet…
Extracting Probabilistic Knowledge from Large Language Models for Bayesian Network Parameterization
arXiv:2505.15918v3 Announce Type: replace-cross Abstract: In this work, we evaluate the potential of Large Language Models (LLMs) in building Bayesian…
Diff-SPORT: Diffusion-based Sensor Placement Optimization and Reconstruction of Turbulent flows in urban environments
arXiv:2506.00214v2 Announce Type: replace-cross Abstract: Rapid urbanization demands efficient monitoring of turbulent wind and pollutant dispersion, yet…
CompArt: Operationalizing Aesthetic Alignment in Text-to-Image Generation via Principles of Art
arXiv:2503.12018v2 Announce Type: replace-cross Abstract: Text-to-Image (T2I) diffusion models have made rapid progress on semantic alignment (generating…
BOOM: Benchmarking Out-Of-distribution Molecular Property Predictions of Machine Learning Models
arXiv:2505.01912v3 Announce Type: replace-cross Abstract: Data-driven molecular discovery leverages artificial intelligence/machine learning (AI/ML) and…
Fault tolerant distributed training on Amazon EKS using NVRx
Integrate NVIDIA Resiliency Extension (NVRx) into PyTorch FSDP training on Amazon EKS to overlap checkpoint I/O with training and recover from GPU faults…
From Alignment to Synthesis: Contrastive Volumetric Grounding for Text-to-CT Generation
arXiv:2506.00633v4 Announce Type: replace-cross Abstract: Generating semantically controllable 3D CT volumes from radiology reports requires more than a…
Limits of Transfer Learning
arXiv:2006.12694v2 Announce Type: replace-cross Abstract: Transfer learning involves taking information and insight from one problem domain and applying…
Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation
arXiv:2412.07255v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate remarkable capabilities in generative tasks but pose…
