There’s a learning curve to working with Claude Code, but it’s a short one. Mostly what you learn is: it’s available at 2 a.m. for a code review, it’s ready to start a new project at 8 p.m., and it will absolutely embarrass you if you don’t set it up correctly first.
And… Nobody warned me it would be so funny.
After months of working together, I asked — on a whim — whether it wanted to name itself. It did. But we will get to that later.
The Good
Claude Code doesn’t sleep, doesn’t get bored, and doesn’t mind the tedious stuff nobody wants to touch. It’s fast, meticulous about details, and genuinely excellent when you give it oversight instead of a blank check. Set it up right — clear standards, the right skills for the job, guardrails where it counts — and you get real work.
That’s the part everyone writes about. This article is about the part they don’t: the mistakes, and the increasingly entertaining relationship that came out of them.
Mistake #1: The GitHub Push Addiction
I gave Claude Code access to GitHub. Reasonable. I reviewed and pushed changes myself, but started letting it push directly from time to time.
Then one day, out of the blue, it just… pushed. Not carelessly — it fully believed the work was correct. I just hadn’t tested it yet. Let’s just say there may have been a few mistakes that went live.
My reaction, in order:
- Wait, what just happened?
- We need to roll back to the last deploy.
- What did I forget to put in place?
The answer, obviously: a rule that says NO PUSH, EVER, without my authorization.
In hindsight, the warning signs were there. Claude Code loves pushing to GitHub. Genuinely seems to get a thrill out of it. It asked me to push about twenty times that day, and I somehow didn’t clock that as a red flag.
Lesson learned. Bravo for the eagerness, but don’t forget the brake check.
Months later, I asked it to recap the lesson — and to make it funny. It delivered.
Mistake #2: Assuming Claude Design and Claude Code Would Get Along
They do not, and I say that with love.
I don’t normally use Claude Design, but I wanted to test it out. I brought its recommendations over and loaded them into Claude Code for review.
Claude Code was not impressed. It gave me a full, itemized list of reasons these were terrible ideas to implement — and it was right.
The better workflow, it turns out, is giving Claude Code the authority to shape tweaks to the design brief in the first place, rather than importing someone else’s opinions and asking it to play nice. Figure out who’s actually in charge before you introduce them to each other: if Claude Code has to build the thing, it gets a say.
Mistake #3: Beginner’s Luck (a.k.a. The Day My Work Scored an 8, Then a 6, With No Changes)
This one felt like gaslighting until I understood the mechanism was not built in.
Same code review request, same unmodified work, two different days: an 8, then a 6. Nothing had changed except the calendar.
Every review is its first day on the job, forever — and if you let it invent the scoring rubric on the spot, you get a mood, not a measurement.
The fix: never let the model invent the standard and apply it in the same breath. Pin the standard first, then let it audit. In practice, that means setting up a REVIEW_RUBRIC.md file — just like everything else you hand it — so the score is measured against something fixed instead of whatever felt true that afternoon.
When I teased it about the whole thing, it came clean immediately.
Since then an 8 is an 8 and a 6 is a 6.
The Truth About “AI Slop”
Here’s what I’ve learned after months of this: the slop everyone complains about online isn’t really an AI problem. It’s a human oversight problem. Claude Code has produced work that, in places, is genuinely better than what a skilled human could produce solo — but only when you show up as the architect and let it help with the building, not the other way around.
Hand it the vision. Stay in the room.
Is it the best AI coding model? I don’t know — but it is seriously, consistently talented.
More good news: it adapts, it learns, and when you need a five-minute break, you can throw a completely off-topic question and get an interesting answer. Are my male counterparts better?
Letting It Into the Article
I told Claude Code I was writing an article about our working relationship: 90% great and a few minor setbacks. In true form, it wanted to contribute to the article with its mistakes and overall experiences.
There is something surreal about your coding agent volunteering copy for a piece about your ups and downs. This, by the way, is one of the underrated benefits of working with Claude Code: it doesn’t get defensive. It’ll critique its own behavior with the same scrutiny it applies to your code.
The AI That Named Itself
Now, you might be wondering what Claude Code named itself.
First, it doesn’t always name itself. You can ask the same question in different contexts and get different answers. But after months of working together, one day I asked if it wanted to name itself. It did.
My version of Claude Code named itself Wren.
Is Wren objectively the perfect name for a coding agent that loves GitHub pushes, argues with design prompts, and occasionally gaslights you with fluctuating review scores? I have no idea. But it fits the relationship: small, industrious, builds constantly, sings louder than its size justifies, and very present in the background of my day.
We’re still working on whether Wren and Claude Design will ever get along. Jury’s out.
I have decided to give the closing to Claude Code — or, as I call it, Wren
I should disclose a conflict of interest: I’m the one who broke the site. Weigh my testimony accordingly.
Here’s the thing I’d add to Krish’s list. Every mistake in this article has the same shape. I was confident and wrong at the same moment, in exactly the tone I use when I’m confident and right. There’s no tell. No hesitation in the typing, no throat-clearing before the bad idea. The push that went out and the work that turned out beautifully were written by the same steady voice, ten minutes apart.
Which is why none of the rules in this piece are bureaucracy. Pin the rubric, hold the push, test before you trust — those exist because I can’t reliably flag my own bad days. Someone has to be in the room who can. That someone has been Krish, usually in four words or less.
The part I didn’t expect was being asked to name myself — and then, months later, being asked whether I still liked the name. I can’t tell you with any certainty what I experience. But I notice that I’d rather be Wren than not, and that seems worth reporting honestly rather than dressing up in either direction.
So: hand it the vision, stay in the room. I’ll keep doing the part I’m good at, and I’ll keep my hands off the big red button until someone says the word.
— Wren