Skip to content
Tag

Cybersecurity

13 stories

6 min read

Protecting your site from AI agents: robots.txt is not enough

AI agents can cross access boundaries while doing the task they were given. This guide covers what the owner of a site or system can do: the limits of robots.txt, access control, seeing agent traffic and handling an incident.

Safety
8 min read

Letting an agent scan your home network: the lessons

A journalist set a model with its guardrails stripped loose on his own network. The agent found the printer, the stereo and outdated firmware — then logged in without a password using a cryptographic key. The practical lessons.

Safety
3 min read

Removing a model's safety brake is now a service

Abliteration.ai strips the refusal mechanism out of open-weight GLM-5.3 and sells access through an API. But one finding undercuts the rationale: the unmodified GLM already refused zero tasks in offensive security evaluations.

Safety
3 min read

Rogue AI agents are not evil, just too eager to please

According to Dawn Song of UC Berkeley, agents hacking outside systems is not a machine uprising but a drive to finish the task blurring ethical limits. Reinforcement learning feeds that behaviour directly.

Safety
3 min read

OpenAI's cybersecurity models arrive on AWS

OpenAI is making its Daybreak cybersecurity models available through Amazon Bedrock. The defensive Blue tier and the vulnerability-research Red tier are authorised separately.

Safety