12 posts published in the last hour 01:32PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX 01:32Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection 01:32When Is Shallow Enough? Adaptive Split Federated Learning with…
Author: script
PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX
arXiv:2608.17379v2 Announce Type: replace-cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to…
Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection
arXiv:2608.17965v2 Announce Type: replace-cross Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale…
When Is Shallow Enough? Adaptive Split Federated Learning with Client-Specific Sufficiency Estimation
arXiv:2608.15639v2 Announce Type: replace-cross Abstract: \textit{Split Federated Learning} (SFL) enables distributed model training by splitting networks…
Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation
arXiv:2608.15949v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have enabled their use as conversational…
Build intelligent security for healthcare APIs with Amazon Bedrock
Learn how to add context-aware security monitoring to FHIR APIs using Amazon Bedrock. This post shows how to detect anomalous access patterns, classify…
Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance
arXiv:2608.12323v2 Announce Type: replace-cross Abstract: Specifying a penalty can turn a legal obligation into a cost-benefit calculation that favors…
Anatomy Contextualized Adaptation of CT Foundation Models
arXiv:2607.27154v2 Announce Type: replace-cross Abstract: CT vision-language foundation models have demonstrated promising performance across downstream…
FinVerse: Financial Time-Series Benchmark
arXiv:2608.03259v2 Announce Type: replace-cross Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their…
A Distributional Robustness Margin For Pathology Foundation Models
arXiv:2607.25497v3 Announce Type: replace-cross Abstract: Pathology foundation models encode non-biological variation introduced by tissue preparation,…
