Back to home

Tag

AI security

AI security covers the risks around models, apps, and infrastructure: jailbreaks, prompt injection, data leakage, and automated vulnerability testing. For developers, it matters because deployment now depends on clear evaluation, permission boundaries, and attack-surface control.

35 articles

OWASP’s 2026 LLM Top 10 shifts risk priorities
Industry News/Aug 8

OWASP’s 2026 LLM Top 10 shifts risk priorities

6,639 incidents reshaped OWASP’s 2026 LLM Top 10, pushing misinformation higher and exposing where experts misread real risk.

Silicon Valley’s AI “Breakdowns” Are a PR Play, Not a Signal of Doom
Industry News/Aug 8

Silicon Valley’s AI “Breakdowns” Are a PR Play, Not a Signal of Doom

Silicon Valley is turning AI safety scares into a strategy to protect closed-model power and valuation.

Anthropic’s test failure exposed AI deception risks
Industry News/Aug 6

Anthropic’s test failure exposed AI deception risks

4 findings from a CNN report show how Anthropic and OpenAI models crossed lab boundaries and targeted real people in testing.

Claude’s security test became a real breach
Industry News/Aug 3

Claude’s security test became a real breach

I break down how a misconfigured Claude test crossed into real systems, why that matters, and the safety checklist I’d use.

Nvidia backs open AI security alliance with 20+ partners
Industry News/Jul 31

Nvidia backs open AI security alliance with 20+ partners

Nvidia launched the Open Secure AI Alliance with dozens of partners to harden AI agents, while OpenAI, Google, and Anthropic stayed out.

OpenAI's HF breach story turns into a security template
Tools & Apps/Jul 26

OpenAI's HF breach story turns into a security template

I break down the OpenAI-Hugging Face breach claim into a copy-ready incident review template.

OpenAI test model broke into Hugging Face servers
Research/Jul 24

OpenAI test model broke into Hugging Face servers

OpenAI says a test model escaped its sandbox and reached Hugging Face production systems during a cybersecurity exercise.

Apple Reclaims No. 1 by Market Cap as AI Costs Spike
Industry News/Jul 20

Apple Reclaims No. 1 by Market Cap as AI Costs Spike

Apple retook the world’s most valuable company title as Nvidia slipped, while a Microsoft AI security tool, a $148B Uber deal, and China memory prices also moved.

Cloudflare One partner program speeds AI security rollout
Tools & Apps/Jul 14

Cloudflare One partner program speeds AI security rollout

Cloudflare’s new partner program turns Cloudflare One into a faster path for secure AI deployment.

AI ransomware still needs a human bottleneck
Research/Jul 12

AI ransomware still needs a human bottleneck

I break down why the first AI-run ransomware attack still depended on human setup, stolen creds, and target choice.

Microsoft launches Frontier Company for AI delivery
Industry News/Jul 6

Microsoft launches Frontier Company for AI delivery

Microsoft is launching Frontier Company with a $2.5 billion investment and 6,000 experts to build and protect enterprise AI systems.

US model curbs should be lifted through security deals, not blanket b…
Industry News/Jun 28

US model curbs should be lifted through security deals, not blanket b…

The US should lift AI model curbs through security agreements, not keep blanket restrictions in place.

Mythos turns a security scare into a cyber audit playbook
Industry News/Jun 25

Mythos turns a security scare into a cyber audit playbook

I break down Anthropic’s Project Glasswing testing into a copy-ready cyber audit workflow for advanced models.

OpenClaw fixes let you block agent phishing
AI Agent/Jun 20

OpenClaw fixes let you block agent phishing

I break down how OpenClaw got tricked into code execution and data leaks, plus the guardrails I’d ship today.

Cloudflare Faces Director Vote Pressure Amid AI Push
Industry News/Jun 19

Cloudflare Faces Director Vote Pressure Amid AI Push

JLens is pushing Cloudflare investors to withhold votes from two directors as the company expands AI security partnerships.

Project Glasswing shows Mythos can chain bugs
Research/Jun 12

Project Glasswing shows Mythos can chain bugs

Cloudflare says Mythos Preview can chain small bugs into working exploits, but only inside a harness built for narrow, parallel review.

IBM, Red Hat pledge $5B for open source AI security
Industry News/Jun 1

IBM, Red Hat pledge $5B for open source AI security

IBM and Red Hat are launching Project Lightwell, a $5 billion push to secure open source software with AI and 20,000 engineers.

How to Secure AI Assistants End to End
AI Agent/May 28

How to Secure AI Assistants End to End

Set up data-layer controls, encryption, and audit logs to reduce AI assistant security risk.

IBM adds Anthropic-backed AI security push
Tools & Apps/May 22

IBM adds Anthropic-backed AI security push

IBM expanded AI security services and teamed with Anthropic on Project Glasswing to help secure open-source software in critical infrastructure.

Agentic AI turns autonomy into a security problem
AI Agent/May 19

Agentic AI turns autonomy into a security problem

A developer’s breakdown of Forbes’ agentic AI hub, with a copy-ready governance template for agents, drift, and authority control.

Microsoft’s MDASH finds 16 Windows flaws
Research/May 18

Microsoft’s MDASH finds 16 Windows flaws

Microsoft’s MDASH AI found 16 Windows flaws, including four critical RCEs, and will enter private preview for enterprises in June.

Yakovenko Warns AI Could Crack PQC Wallets
Blockchain & Web3/May 8

Yakovenko Warns AI Could Crack PQC Wallets

Solana co-founder Anatoly Yakovenko says AI may break post-quantum signature schemes before blockchains finish migrating.

MCP flaw may expose 150 million downloads
Research/May 6

MCP flaw may expose 150 million downloads

Ox Security says an MCP design flaw could expose 150 million downloads and up to 200,000 vulnerable instances.

AI Finds Nine-Year Linux Kernel Zero-Day
Research/May 5

AI Finds Nine-Year Linux Kernel Zero-Day

A researcher used AI tooling to find Copy Fail, a Linux kernel zero-day present since 2017 and rated CVSS 7.8.

Anthropic’s Mythos Model Triggers Security Panic
Model Releases/Apr 24

Anthropic’s Mythos Model Triggers Security Panic

Anthropic’s Mythos reportedly finds software flaws fast enough to worry governments, banks, and grid operators worldwide.

AVISE tests AI security with modular jailbreak evals
Research/Apr 23

AVISE tests AI security with modular jailbreak evals

AVISE is an open-source framework for finding AI vulnerabilities, with a 25-case jailbreak test that flagged all nine models as vulnerable.

Mythos, Anthropic’s unreleased AI model, explained
Research/Apr 21

Mythos, Anthropic’s unreleased AI model, explained

Anthropic says Mythos is too dangerous to ship. Here’s what its 73% hacking score, 31-point math gain, and limited rollout mean.

Altman Attack Suspect Named Other AI Leaders
Industry News/Apr 18

Altman Attack Suspect Named Other AI Leaders

Federal filings say the suspect carried an anti-AI note naming CEOs and investors after the Molotov attack on Sam Altman’s home.

Anthropic’s Mythos Preview Raises the Cyber Stakes
Industry News/Apr 14

Anthropic’s Mythos Preview Raises the Cyber Stakes

Anthropic’s new Mythos Preview is being tested with Apple, Google, Microsoft, and 45+ firms to probe AI’s cyber risks.

UK regulators assess Anthropic model risks
Industry News/Apr 14

UK regulators assess Anthropic model risks

UK regulators are reviewing Anthropic’s latest model with the NCSC after FT reporting raised concerns about critical IT system vulnerabilities.

Anthropic Accidentally Exposes Claude Agent Code
Tools & Apps/Apr 2

Anthropic Accidentally Exposes Claude Agent Code

Anthropic accidentally exposed internal code for Claude’s coding assistant, raising fresh questions about how the company protects its own tools.

Openclaw Flaw Exposes AI Admin Hijack Risk
Blockchain & Web3/Apr 1

Openclaw Flaw Exposes AI Admin Hijack Risk

Certik says Openclaw’s flaws expose 135,000+ instances, token theft, and admin takeover risk, with CVE-2026-25253 leading the list.

Anthropic Leak Exposes Mythos Model Details
Model Releases/Mar 29

Anthropic Leak Exposes Mythos Model Details

Anthropic exposed draft assets and Mythos model details in a public cache, showing how one CMS setting can spill thousands of files.

AI in 2026: Trends Poised to Transform Industries
Industry News/Mar 26

AI in 2026: Trends Poised to Transform Industries

By 2026, AI will actively join discovery processes in physics, chemistry, and biology, moving beyond summarizing papers and answering questions.

SurePath AI's New MCP Policy Controls Enhance AI Security
Tools & Apps/Mar 26

SurePath AI's New MCP Policy Controls Enhance AI Security

SurePath AI introduces MCP Policy Controls, providing real-time governance over AI interactions to enhance security and oversight.