#ai
- Corv: my Hermes agent.
A tour of the AI agent that lives on my homelab: what runs it, what it can reach, and how it stays useful without being babysat.
- Rook, My Agent Workstation
How I run coding agents inside a Debian LXC, and keep their infrastructure changes behind OpenTofu plans, Ansible playbooks, and Forgejo review.
- The [model-dependent] Caveat, Executed
Last time I measured whether a model obeys an injection. I never let it act. So I ran the same chains against real servers, and half my numbers didn't survive.
- The [model-dependent] Caveat, Measured
My MCP audits proved a malicious instruction reaches the agent. They never checked whether it obeys. So I measured it across thirteen models.
- Auditing MCP Servers with an AI on a Short Leash
The methodology behind my MSc dissertation: depth-first security audits of MCP servers, using an LLM as an instrument I'm not allowed to trust.