arXiv:2608.13712v1 Announce Type: cross Abstract: Deposition training requires attorneys to manage dynamic witness behavior, yet legal-AI evaluations…
The Hidden Cost of AI: How “Black Box” Models Are Eroding Trust, Budgets, and the Environment
The AI industry has a model design problem. Enterprises are starting to feel it in cost, performance, and reliability. AI can talk – but can it listen?…
TeachMateGPT: A Multi-Agent Knowledge-Grounded Framework for Pedagogical Assessment Generation from Science Curriculum Materials
arXiv:2608.13708v1 Announce Type: cross Abstract: Automatically generating textbook-grounded assessment items can reduce science teachers’ workload, but…
Axiom Math’s AI Verifies the 246 Prime-Gaps Theorem in Lean
Axiom Math says its AxiomProver system has produced a machine-checked Lean 4 proof of the strongest known result on gaps between prime numbers: the…
Does ISO-Grounded NFR Specification Improve LLM Code Generation? A Comparison of Rich and Structured Interventions against a Natural-Language Baseline
arXiv:2608.13742v1 Announce Type: cross Abstract: In LLM-based code generation, Non-Functional Requirements (NFRs) are often specified as terse one-line…
AI News Brief Hourly Summary 2026-08-17 17h : 18 posts
18 posts published in the last hour 14:33Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT 14:33From BERT to Frontier Agents: Eight Years of Language-Model Progress, the Collapse of the Capability-Cost Curve, and…
Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT
arXiv:2608.13681v1 Announce Type: cross Abstract: Translating C code into safe, idiomatic Rust is a longstanding software-engineering goal because it can…
From BERT to Frontier Agents: Eight Years of Language-Model Progress, the Collapse of the Capability-Cost Curve, and the Rise of Task-Targeted Models
arXiv:2608.13675v1 Announce Type: cross Abstract: Between October 2018 and July 2026 AI models progressed from simple systems like BERT to massive agents…
MedPlex: Deep Vision-Language Co-Adaptation for Clinically Grounded Medical Segmentation
arXiv:2608.13690v1 Announce Type: cross Abstract: Medical image segmentation is still largely treated as a vision-only problem, although clinical…
Monetizing AI: A Modern Framework for How Companies Can Maximize Business Growth
AI has rapidly moved from a competitive differentiator to a baseline expectation for software companies. However, it has also become the most expensive…
CLAIR-Fin: An Adversarial Multi-Agent Framework for Claim-Level Verification and Adaptive Debate in Cross-Modal Financial QA
arXiv:2608.13706v1 Announce Type: cross Abstract: Existing defenses against hallucination in retrieval-augmented and multi-agent pipelines remain partial:…
OpenAI signs record Ohio data center lease with Nvidia backing up to $105 billion
OpenAI has signed a 20-year lease for an 8-gigawatt data center in Ohio. Nvidia is guaranteeing up to $105 billion for the residual value of the…
SAGE: Surrogate-gradient Adaptation via Attention-Guided Entropy for Spiking Transformers
arXiv:2608.13702v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer an energy-efficient alternative to conventional deep neural…
What Can I Actually Do with a Small Language Model?
But by keeping these limits in mind, and planning for them, we can effectively use these small, local models for the following broad operations scenarios.
Measuring Fairness in Large Audio Language Models via Semantic-Aware Bias Estimation
arXiv:2608.13624v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) have seen increasing use for audio understanding tasks such as…
From AI Copilots to Agent Swarms
The impact of AI on software development has been both profound and ever-evolving. Last year, I wrote about AMD’s plans to use AI not just for generating…
IterCOMP: Reasoning-aware Adaptive Prompt Compression for Multi-hop Question Answering
arXiv:2608.13588v1 Announce Type: cross Abstract: Multi-hop question answering requires complex reasoning across multiple evidence segments, which often…
Changing Font Colors Can Hijack AI Reasoning
A new study finds that ordinary formatting can quietly steer AI reasoning, causing it to overlook words, misread meaning, and reach different conclusions…
