arXiv:2608.20384v1 Announce Type: new Abstract: Multimodal affect and behaviour classifiers that fuse heterogeneous text, audio, and visual streams must…
Tag: cs.AI updates on arXiv.org
A Survey on Foundations and Frontiers of Multimodal Agentic Frameworks: Techniques and Applications
arXiv:2608.20379v1 Announce Type: new Abstract: Advances in large language models (LLMs) have fueled a wave of research into agency: the ability to…
Truth Lies Deep: Countering Semantic Camouflage via Latent Intent Verification
arXiv:2608.20378v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) is often superficial, relying on refusal mechanisms that…
PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX
arXiv:2608.17379v2 Announce Type: replace-cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to…
Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection
arXiv:2608.17965v2 Announce Type: replace-cross Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale…
When Is Shallow Enough? Adaptive Split Federated Learning with Client-Specific Sufficiency Estimation
arXiv:2608.15639v2 Announce Type: replace-cross Abstract: \textit{Split Federated Learning} (SFL) enables distributed model training by splitting networks…
Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation
arXiv:2608.15949v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have enabled their use as conversational…
Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance
arXiv:2608.12323v2 Announce Type: replace-cross Abstract: Specifying a penalty can turn a legal obligation into a cost-benefit calculation that favors…
Anatomy Contextualized Adaptation of CT Foundation Models
arXiv:2607.27154v2 Announce Type: replace-cross Abstract: CT vision-language foundation models have demonstrated promising performance across downstream…
FinVerse: Financial Time-Series Benchmark
arXiv:2608.03259v2 Announce Type: replace-cross Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their…
