arXiv:2609.03734v1 Announce Type: cross Abstract: BLEU-4 is the standard metric for evaluating sign language translation (SLT), but spoken-language…
Category: cs.AI updates on arXiv.org
Can LLMs Extract Architectural Design Decisions from Source Code Commits? – A Preliminary Exploratory Study
arXiv:2609.03721v1 Announce Type: cross Abstract: Context: Architectural Design Decisions (ADDs) capture the rationale behind the structure and evolution…
Symmetries and Causality: Causal Effect Identification Beyond IID Data
arXiv:2609.03697v1 Announce Type: cross Abstract: In the natural sciences, symmetries and cause-effect relationships are ubiquitous. Yet for complex…
IndicSafeEval: Safety Robustness of Large Language Models under Multilingual Persuasive Jailbreak Attacks
arXiv:2609.03781v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in multilingual settings, yet their safety is still…
Cross-Dataset Transfer and Reliability of Explainable Artificial Intelligence for RhythmFormer Remote Photoplethysmography
arXiv:2609.03663v1 Announce Type: cross Abstract: Background. Remote photoplethysmography estimates the cardiovascular pulse from facial video, and its…
Enhancing Financial Question Answering: A Novel Benchmark Dataset of Banks’ financial statements
arXiv:2609.03654v1 Announce Type: cross Abstract: The comparative analysis of banks’ financial statements poses significant challenges for automated…
Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning
arXiv:2609.03667v1 Announce Type: cross Abstract: Generalising to unseen tasks remains a fundamental challenge in offline multi-agent reinforcement…
Local Updates, Global Learning (LUGL): Playing Games with non-incremental Learners
arXiv:2609.03660v1 Announce Type: cross Abstract: The dominance of Neural Networks (NNs) in RL is partially due to their incremental learning capability,…
Doesn’t Stop Reasoning: Analysis of Spurious CoT Termination
arXiv:2609.03633v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning improves large reasoning models (LRMs) on complex tasks but often…
Test-time adaptation for speech enhancement with an autoregressive speech prior
arXiv:2609.03622v1 Announce Type: cross Abstract: Test-time adaptation (TTA) offers a promising direction for improving speech enhancement models under…
