This article was originally published on the Endform blog . By the end of this walkthrough you'll have a single Playwright test that Claude Code wrote against your live app, that you've reviewed in two passes, run repeatedly to check for flakiness, and committed next to the plain-English scenario…
TL;DR Downloaded a vendor's "paste this into your agent" setup guide for Jev, a fast/cheap decision model. Had Claude Code build the integration live instead of just reading about it. Secured the real API key into gitignored storage before a single test call. Caught the docs-fetching tool…
2 of 8 compactions dropped the body of a skill I had invoked earlier in the session, even though the Claude Code docs say invoked skill bodies are re-injected: both were /compact sent through claude -p --resume , while the 6 compactions that ran inside a single process re-attached the skill every…
For two weeks I woke up and checked for a markdown file that should have been written at 3:00 AM. On 8 of 14 mornings it either wasn't there, was empty, or showed up after I'd already finished my coffee. The job was simple. A macOS launchd agent ran Claude Code headless with claude -p every night…
I use Claude Code most of the day (Codex CLI too), and three small things kept getting in my way. The first one: I'm halfway through writing my next prompt while Claude is still working on the last one. Then I notice it heading the wrong way, or I realise I need to ask something quick first. My…
Two statements about Claude Code mods, both true, both from Anthropic's own docs: The module runs in a sandbox of its own, with no DOM and no Node. "Mods aren't sandboxed." The first two days of the mods launch have been one long argument about which of these is the lie. Neither is. The confusion…
You told Claude Code “always run the formatter after editing” and “never run rm -rf ”. It listened, until it didn’t. A rule in a prompt is a request — it can be forgotten, skipped, or missed after a context compaction. A hook is not a request. It’s a shell command that runs automatically at a…
I like /goal in Claude Code. You say what "done" looks like and Claude keeps working until it gets there. What bothered me is who decides it got there. After every turn a small model (Haiku) reads the conversation and answers "met", "not yet" or "impossible". I tested what that judge can see. It…
Last Friday, while cross-debugging a payment refactoring branch, I had Cursor open on my left monitor for frontend reviews and Claude Code running in my terminal on the right. To let both AI assistants inspect database schemas and query logs, I hooked up 4 MCP servers to each environment. When you…