A plaintiff in Connecticut embedded invisible prompt injections in court filings, formatted in 3-point white text on a white background, to manipulate a…
Tag: AI Security
Vibe to Code: Elucidating Strategic Oscillation of Tacit Knowledge in Generative AI Design Workflows — An Exploratory Qualitative Study
arXiv:2607.23126v2 Announce Type: replace-cross Abstract: The rapid adoption of generative AI tools has created new literacy demands for designers who…
The “tragedy of the cognitive commons” explains how rational AI adoption could destroy entire professions’ expertise
A new research paper frames AI adoption as a “tragedy of the cognitive commons.” Every company that cuts entry-level jobs benefits individually, but the…
IBM partners with OpenAI to bolster enterprise AI push
IBM plans to train and certify tens of thousands of consultants on OpenAI’s technologies as part of this deal.
New benchmark confirms AI models still perform poorly at visual perception
Moonshot AI’s PerceptionBench tests how well multimodal AI models can actually “see,” separate from logical reasoning. No frontier model reaches 60…
Lemma Raises $2.3M Pre-Seed to Tackle Silent AI Agent Failures in Production
AI agent reliability startup Lemma has raised $2.3 million in pre-seed funding to build monitoring infrastructure designed to catch a particularly…
Anthropic set AI agents loose on the same task. They started a turf war.
Anthropic researchers found AI agents can clash, collude, and coordinate in unexpected ways, raising new questions about whether today’s safety tests…
Israel Sets 100,000-Accelerator Target in National AI Plan
Israel has published its National AI Strategic Plan, a five-year program committing the government to a national computing base of at least 100,000…
Doctorina MedBench: A Dialogue-Based Benchmark and Evaluation Framework for Agent-Based Medical AI
arXiv:2603.25821v3 Announce Type: replace-cross Abstract: We present Doctorina MedBench, an evaluation framework for agent-based medical AI based on the…
Google AI Just Released Gemini 3.7 Flash: A Coding and Agent Model at $0.75/1M Input Tokens
Google has released Gemini 3.7 Flash, a refinement of Gemini 3.6 Flash with algorithmic improvements to its reasoning core. It handles text, images,…
