Skip to main content

AI models, security vulnerabilities and guardrails

@ggirtsou
There are significant security implications with the new AI models (Mythos from Anthropic and GPT 5.3 Codex) due to their capabilities. These labs were aiming at making the models excel in coding, but turns out they became insanely good at finding security vulnerabilities.

Anthropic did the responsible thing and decided to hold off releasing it to the public and formed a working group with Google/AWS/Apple/Crowdstrike etc to offer to use their model to fix critical issues. I imagine CS will work with airports to ensure no planes go down because of undetected vulnerabilities from 27 years ago (yes, they still have systems written in COBOL and they won't modernise it because "it works").

My take on this for the guardrails is that Labs should ask users to prove they are authorised to do the security audit (e.g. proof they are hired by the company), or that the user owns the project (e.g. domain verification). This will allow going forward as industry and patching things, instead of downgrading requests to previous models, which is essentially what we have available today generally available.

(also Anthropic did their marketing bit by saying the model is x10 more expensive than Opus :D)
2

7 posts about AI · latest 2 Oct 2026

What people are saying

Nobody has replied yet. Log in to say something.

Log in Get the app

Meet people over dinner

KF.Social is where people share real moments and make real friends. Join the conversation.

Join KF.Social

About KF.Social

KF.Social is where people share their experiences and make real friends over dinner. Every post here is from a real member of the community.

KF.Social
Go for dinner. Leave with friends.
Get the app
Link copied