Craig Hughes

@craig.rungie.com

I dabble. Quite a lot.

Regardless of how it turns out NO HE DIDN’T. It’s like saying I have more fingers on my left hand after counting 8 digits from left to right with a ring finger and a pinky still to report.

Bild

God it’s so nice working with Claude instead of having to slum it with codex while my weekly quota refreshes. I need to be more disciplined about not exhausting supply. Three days of codex only and I was ready to kill myself. It’s just bad at working effectively.

FLASH: Justice Dept seeks to dismiss Lincoln Memorial Reflecting Pool criminal case against former Olympian David Hearn Trump Admin acknowledges: "Damage to the Lincoln Memorial Reflecting Pool in June 2026 was the result of flawed installation by the contractor"

Bild

Doing a little test with a skill and ~/.claude/CLAUDE.md which has CC estimate every task up front in estimated tokens (input/output); then on completion record actual usage. Everything -> sqlite along with timestamps. Then analyze estimate:actual over time and map to wallclock. Story-points style.

My claude decided it'd be easiest to just use some python code in one of my rust/typescript projects recently, and since then the entire project repo and claude project/sessions have just been on a rapid downhill trajectory.

It's gonna be a ten gigatoken month. I have to say, based on this, I like CC *way* more than Codex. sol is a nice model, but the codex harness is for shit - especially in letting you know wtf is happening. Often I can't even tell if ANYTHING IS RUNNING in codex or if it's just waiting for me.

Bild

Approaching 6 Gtok between Claude Code and Codex combined for July so far. Biggest month yet! And I definitely find I prefer CC > Codex; much easier to control agents, workflows, monitor remotely, multiple session management... using CLI for both.

Bild

User: Python developers literally must have all been dropped on their heads as babies Claude: Let me check how widespread that actually was rather than take it on vibes. Ran 1 shell command Checked, and the facts are narrower than that — though the underlying complaint is fair.

Take-aways from Opus 5 model card. Exercise caution: 1. Hallucinates more often. Do adversarial verification. 2. Use "medium" effort. Medium often beats high on real work. 3. Pattern of "fabricated user consent for destructive actions" – it thinks yesterday's "go ahead" still applies now.

I notice a conspicuous lack of "Opus 5 is gonna be so good at doing advanced thinking it's scary" type marketing. I think they learned their lesson with Fable/Mythos. Now hopefully Amazon won't just shop them again.