AI security for December 4, 2025 centers on a massive $130M round for AI-agent SOC startup 7AI, fresh evidence of Chinese-backed hackers using AI to automate campaigns, a study showing major AI companies falling short of global safety standards, new analysis of AI-driven software supply chain attacks, and a malicious npm package that embeds prompts to trick AI-based security tools.
Model Safety
AI Security Daily Briefing — December 2, 2025
AI security on Dec 2 centers on a serious Codex CLI command-injection flaw, new data showing layered AI defenses still buckle under targeted attacks, Anthropic’s agents successfully exploiting real DeFi contracts, Android zero-days hitting AI-enabled mobile endpoints, and a major investment push into explainable AI-driven investigations for national security
Prompt Injection and LLM Jailbreaking — Operational Playbook for Defense
Prompt injection and jailbreaks exploit LLMs by embedding malicious instructions in user inputs or retrieved content. This playbook outlines real-world cases and practical defenses including sanitization, least-privilege design, and red-team testing.
AI Security Daily Briefing — September 30, 2025
Today’s briefing spotlights serious vulnerabilities in Google’s Gemini assistant, Microsoft’s new unified AI security stack, and a CAISI evaluation of DeepSeek’s security weaknesses. Broader context from WEF emphasizes how AI is increasingly integral to both offense and defense in cybersecurity.