#ai-security

28 posts tagged with #ai-security

Every article below is hand-written, technically reviewed, and focused on ai-security. Posts cover real-world architecture decisions, code-level implementation patterns, and trade-offs you'll only discover after shipping production systems.

A computer screen displaying network configuration code in a terminal window Developer Tools

MCP vs Function Calling in Agents [2026]: When to Say No

MCP backlash is real. Here’s my decision framework for when to use MCP vs function calling in agents, and when a bespoke tool API beats both on ops, security, and debugging.

emergency stop button industrial control panel red e-stop — illustration for article on How to Add Cybersecurity

How to Add AI Agent Kill Switch Spend Limits [2026]

A software-only watchdog pattern for AI agents: a single enforcement point for kill switches, per-tool budgets, approval gates, and tamper-evident audit logs.

A wooden block that says token sitting on a table Cybersecurity

MCP OAuth Security: Tool Impersonation, aud Mismatch, Token Replay [2026]

A threat model for MCP tool servers using OAuth: how tool impersonation and audience mismatch happen, how tokens get replayed, and what to validate and log in 2026.

Laptop screen displaying code and data graphs Cybersecurity

How to Stop Repo Prompt Injection in Coding Agents [2026]

Repo-level prompt injection turns “clone and ask the agent” into a supply-chain compromise. Here’s the threat model, a safe demo, and practical mitigations.

a clipboard with a checklist on it next to a cup of coffee and Cybersecurity

MCP Server Security Best Practices: Checklist + CI Linter [2026]

A practical MCP server security test plan: authn/authz models, tool allowlists, prompt-injection regression tests, rate limits, audit logs, and a CI-friendly permissions linter.

green frog iphone case beside black samsung android smartphone Cybersecurity

RatHat Android Malware “AI‑Powered” Claims: What to Detect [2026]

RatHat is getting labeled “AI-powered,” but the real story is the Android stealer kill chain: accessibility abuse, permission traps, and noisy C2 loops you can actually detect.

A command line interface showing the text ubuntu@ubuntu:~$ sudo with a blinking cursor Cybersecurity

How to Secure Local LLM Inference [2026]: Sandbox + Egress

A practical blueprint for secure local LLM inference: sandbox inference hard, default-deny outbound network, stage allowlisted downloads, scan artifacts, and isolate tools.

black hp laptop computer turned on displaying desktop AI and Machine Learning

7 AI Agent Swarm Coordination Patterns [2026]: Failure Modes Included

AI agent swarm coordination patterns decide whether your multi-agent system converges or spirals. Here are the topologies, stop conditions, budget caps, and failure modes most teams don’t test in 2026.

an open laptop computer sitting on top of a table Cybersecurity

How to Implement OWASP Agentic Top 10 Controls [2026]

A control-by-control guide to OWASP agentic top 10 controls: where to gate tool calls (MCP/client/server), what to log, what to block, and what belongs in CI vs runtime.

Laptop screen displaying lines of code Developer Tools

Review AI-Generated Code Checklist [2026]: Beat the Bottleneck

AI makes code cheap. Verification is the new tax. Here’s a practical workflow, checklist, and ‘delete & redo’ rubric to keep PRs fast and safe in 2026.

system logs terminal laptop screen code — illustration for article on LLM Data Leakage Playbook [2026]: Cybersecurity

LLM Data Leakage Playbook [2026]: Logging, Retention, Redaction

A practitioner playbook for preventing data leakage in LLM apps by hardening logging, retention, and redaction across the entire prompt→tools→model→observability path, with audit-ready evidence you can hand to compliance.

Abstract grayscale waveform with dark background Cybersecurity

How to Run an AI Voice Detector Accuracy Test [2026 Harness]

Build a repeatable deepfake-voice detector benchmark: datasets, metrics, thresholding by false-positive cost, robustness transforms, and a privacy-safe way to publish results.

A security and privacy dashboard with its status Cybersecurity

AI Security Leader Playbook [2026]: 10 Controls That Ship

A practical AI security leader playbook you can implement this quarter: inventory, approval gates, agent threat modeling, OWASP LLM Top 10 controls, vendor review, and incident response.

A security and privacy dashboard with its status. AI and Machine Learning

RAG Data Leakage Test Suite [2026]: CI Red-Team Setup

Build an automated red-team suite for RAG apps: canary tokens, regex + similarity detectors, multi-step prompt-injection attacks, and a CI risk score that blocks risky merges.

a close up of a laptop with a pink screen AI and Machine Learning

AI Agent Sandbox Linux VM [2026]: Safe Tool Use, No K8s

If your coding agent can run `git`, `pip`, or a shell, it deserves its own disposable Linux VM. Default-deny egress, snapshot rollback, scoped secrets, and per-run audit bundles. No Kubernetes required.

spectrogram audio analysis screen — illustration for article on Deepfake Voice Detection: 7-Step Detector Eval Guide Cybersecurity

Deepfake Voice Detection: 7-Step Detector Eval Guide [2026]

Deepfake voice detection is easy to demo and hard to operationalize. Here’s a repeatable 7-step methodology to evaluate detectors: datasets, telephony transforms, multilingual edge cases, metrics, thresholds, and deployment playbooks.

claude code terminal laptop screen — illustration for article on Claude Code Security [2026]: Risks, Safe Technology

Claude Code Security [2026]: Risks, Safe Setup, Team Policy

Claude Code is safe only if you treat it like a junior engineer with terminal access. Here’s the 2026 playbook: permissions, sandboxing, egress controls, MCP allowlists, retention settings, and incident response.

Green text displaying code on a dark computer screen Cybersecurity

AI Agent Memory Exfiltration: Kill Chain + 5-Step Hardening [2026]

Claude's memory was silently exfiltrated to an attacker's server with zero user warnings. Here's the full kill chain, which memory architectures are vulnerable, and a 5-step hardening checklist grounded in OWASP LLM Top 10 2025.

red padlock on black computer keyboard Cybersecurity

AI Agent Threat Model: 7 Attack Vectors [2026]

Prompt injection is just vector #1. Here's the full AI agent attack surface map — tool poisoning, memory injection, orchestrator hijack, Denial of Wallet, and more — with a sprint-ready threat matrix.

AI Voice Detector: Detect AI Audio & Speech [2026] Technology

AI Voice Detector: Detect AI Audio & Speech [2026]

62% of organizations faced a deepfake attack last year. Here's how AI voice detectors actually work, 5 tools compared with accuracy benchmarks, and manual detection techniques for spotting synthetic speech.

The Complete Guide to AI Security in 2026 Cybersecurity

The Complete Guide to AI Security in 2026

AI and LLM security in 2026 spans prompt injection, supply chain attacks, agent control flow vulnerabilities, and model misuse. This complete guide maps every major threat vector and links to 26 in-depth breakdowns so you can defend your AI systems today.

Workflow diagram, product brief, and user goals are shown. Cybersecurity

AI Agent Security Attack Surface Map [2026 Checklist]

The first developer-friendly attack surface map combining OWASP's Top 10 for Agentic Applications, Cisco's MemoryTrap disclosure, and June 2026 red-teaming benchmarks showing 70% attack success rates — with a printable security checklist.

A laptop screen displays "claude fable 5 is currently unavailable." Cybersecurity

Advanced Prompt Injection Techniques 2026: 7 Attack Chains Beyond OWASP #1

Prompt injection graduated from academic curiosity to active exploit — with CVEs filed against GitHub Copilot, Claude Code, Cursor, and AWS Kiro in a single month. Here are the 7 advanced attack chains researchers are tracking and the only defense architecture with provable security.

Woman typing on a laptop with a vase nearby Cybersecurity

Indirect Prompt Injection in AI Agents: 10-Step Red-Team Checklist [2026]

Every major AI coding agent shipped with exploitable indirect prompt injection vulnerabilities in 2025. Here's the red-team checklist to find them in your own pipeline before attackers do.

Vibe-Code Security Nightmares Nobody Warns About [2026] Cybersecurity

Vibe-Code Security Nightmares Nobody Warns About [2026]

63% of AI-generated functions ship with a security vulnerability. Here's the OWASP-mapped breakdown of what vibe-coded apps get wrong — and the audit checklist that catches it before your users do.

green and orange electric wires Cybersecurity

CVE-2024-3400 and the AI Security Crisis: Palo Alto's CEO Warned Us While His Own Firewalls Burned [2026]

Palo Alto Networks' CEO warned the industry about AI-powered attackers finding zero-days faster than ever. Weeks later, a perfect 10.0 CVSS vulnerability hit his own firewalls. The irony tells us everything about where cybersecurity is headed.

red light on black background Cybersecurity

LiteLLM Supply Chain Attack: How a Fake PyPI Package Targeted AI Developers' Credentials [2026]

A malicious PyPI package used LiteLLM as bait to steal API keys and cloud credentials from AI developers. Here's the anatomy of the attack and how to protect your infrastructure.

a black and white photo of a building Technology

Prompt Injection in 2026: Still OWASP's Number One LLM Vulnerability

Prompt injection has held the #1 spot on OWASP's LLM Top 10 across every edition. Here's why it's unsolvable, how agentic AI made it worse, and what developers actually need to do about it.