Aidan Harding

@aidanharding.eurosky.social

Architect at Aquiva Labs on Salesforce and cloudy things. Also trail running and country living in N Devon

Last week, I tried open weights models for coding compared to my existing method of a Claude Team Premium account. And I'm missing something, or the cost per token difference makes open weights unusable for me. Reviewing one PR on GLM-5.3 cost more than a day of tokens with Claude!

AI can now detect, automate, and chain together exploits faster than any human hacker can. That is significant. Which means that the system vulnerabilities once considered not so low hanging fruit are actually very low hanging.

Chinese Room only demonstrates that the operator doesn't need to understand. It doesn't say the operator *cannot* understand. The operator may still understand. If the operator decides to take a night course in Chinese to make his job easier, the room still works exactly the same

Part 1 of realising we'd got too attached to Anthropic was checking with my immediate team on Opus 5, then wider team, then cross company. Everyone finds it slow, expensive and obtuse. Simply going back to 4.8 has helped loads. Long term, keeping options open with other models and harnesses

OMG, a while back, I added a section to Claude.md to look for some of the standard cognitive biases and it actually works. Sunk cost fallacy check in action: "We chose Rooms-in-scope on my incorrect claim that the data was already there. Worth asking whether we'd choose it again now."

A Claude trace saying, We chose Rooms-in-scope on my incorrect claim that the data was already there. Worth asking whether we'd choose it again now.

When I used to write for The Guardian, every time I wrote the word “meow” the editors would change it to “miaow”. It used to really get on my wick. Don’t try to make my cats sound like your upper middle class London cats, you fuckers. My cats did not go to your posh cat schools. They say “meow”.