arXiv:2608.20400v1 Announce Type: new Abstract: Agentic memory under a fixed budget involves two stages: retention and retrieval. Existing…
Category: cs.AI updates on arXiv.org
Environmental Slow AI: Design Principles for Generative Systems
arXiv:2608.20398v1 Announce Type: new Abstract: Generative AI (genAI) systems produce cultural artefacts at scale, but they also reflect embedded cultural…
Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness
arXiv:2608.20389v1 Announce Type: new Abstract: A production agent harness must discover and rank, from a growing library of skills, the one most…
PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure
arXiv:2608.20342v1 Announce Type: new Abstract: Large language model (LLM) coding agents start each session with an empty context window, discarding…
SDAD: Spec-Driven Agentic Development for the AI-Native SDLC
arXiv:2608.20341v1 Announce Type: new Abstract: Frontier coding agents backed by large language models with context windows from hundreds of thousands to…
Interpretable Multimodal Classification with Linear Discriminant Tree Ensembles
arXiv:2608.20384v1 Announce Type: new Abstract: Multimodal affect and behaviour classifiers that fuse heterogeneous text, audio, and visual streams must…
A Survey on Foundations and Frontiers of Multimodal Agentic Frameworks: Techniques and Applications
arXiv:2608.20379v1 Announce Type: new Abstract: Advances in large language models (LLMs) have fueled a wave of research into agency: the ability to…
Truth Lies Deep: Countering Semantic Camouflage via Latent Intent Verification
arXiv:2608.20378v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) is often superficial, relying on refusal mechanisms that…
PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX
arXiv:2608.17379v2 Announce Type: replace-cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to…
Too Sure to Be Safe: Model Calibration for Reliable Log Anomaly Detection
arXiv:2608.17965v2 Announce Type: replace-cross Abstract: Online log anomaly detection is critical for maintaining the reliability of large-scale…
