Skip to content

AI

I've Been Thinking a Lot About Waiting

I've been thinking a lot about waiting lately. Not the dramatic kind. The ordinary kind. Waiting for a build to finish. Waiting for someone in procurement to decide whether a tool I needed last month can be approved this month. Waiting for a security review. Waiting for legal. Waiting for a meeting to start because everyone is still hunting for the HDMI cable. Waiting for a colleague to come back from holiday, because of course the one answer I need is locked inside their head and nowhere else. Waiting because somebody marked an email High Importance, apparently unaware that Outlook still has no button labelled "ignore until tomorrow, this is not actually important."

I Built a Local LLM Benchmark Harness, and It Mostly Started as an Argument With My GPU

For a while now I have had the same nagging question that I suspect a lot of people in security and IT have been quietly circling. Which local model is actually good enough for the work I do, and what does "good enough" even mean once you stop hand waving? Not the leaderboard scores, not the demos where someone asks a model to write a haiku about Kubernetes, but the actual workloads. Reading a log. Spotting a brute force that turns into a successful login. Writing an incident report without quietly inventing a threat actor that never existed.