Howdy wizards,
When youβve decided to automate a process with AI, itβs easy to think too small.
Iβll show you an example.
You mightβve noticed the cover image of this newsletter has a certain style to it.
That bearded character (the Wizard), always with a coffee cup, always working on his laptop. Heβs the face of this brand.
Yeah, heβs AI generated. But I have my own prompt and reference images I use to make the character and environment look consistent.
My point is this: I used to spend 20 minutes on this. Every week.
What making one cover image involved:
Prompting
Choosing between images
Placing the images into my Figma template
Exporting it, compressing it and putting it into my newsletter
In my head, I naively thought I was already using AI to its full potential in this process.
After all, I was prompting.
But I was wrong.

I was using AI for only a tiny part of the actual process.
And each week Iβd drop what I was doing to sit and generate images.
Iβm at edition #237 of this newsletter. Thatβs a full work week, give or take, that Iβve spent generating cover images!
Not exactly a good ROI situation.
Eventually I realised I could pretty much use AI for the whole thing.
I had a mindset shift.
That shift has helped me automate many boring processes in my work since then.
Iβll show you what the mindset looks like, and what my new cover image workflow looks like, too.
One thing I still havenβt used AI for is running a voice agent.
If I did, Iβd have a callerβs voice mixed with everyone else talking in the background. Hello messy transcript, and an agent taking all kinds of wrong actions. Iβm glad thatβs not my problem. But if it is yours, definitely check out this weekβs sponsor, Krisp:
IN PARTNERSHIP WITH KRISP
AI agents that answer phone calls work great in demos, then fall apart on real calls. The reason is usually the audio.
Speech-to-text nowadays handles noise just fine. Traffic, fans, air conditioning, no problem. What breaks it is a second person talking near the caller. A noise filter canβt tell one voice from another. The transcript comes out wrong, and the agent acts on it: wrong answer, wrong booking, βcould you repeat that?β until the caller hangs up.
Krispβs Voice Isolation 2.5 sits in front of your speech-to-text and isolates the callerβs voice from everyone talking around them. On calls with a competing speaker, word errors drop from 36% to 11%. It adds 15 ms of latency and barely touches clean audio, so you can leave it on for every call.
β¦
The mindset shift
Hereβs how I think about using AI at work nowadays:
Direction & taste are your human currency. The rest can be delegated to AI.
It feels scary to take on this mental model.
Like β am I gonna lose every last brain cell now?
But I already did that. I asked AI for everything and lost all my brain cells.
This is different. I donβt ask AI for everything, I use it for everything.
I go to AI for execution, not answers.
I go to AI when I know what I want, and why I want it.
Solving the technicalities of how to get there is mostly AIβs job. My role is to get out of its way.

Most of the sub-processes of any task are possible for AI to do.
It will require your deep cognitive involvement in the process. You will have to give AI access to things. You will have to test and you will have to verify. There will be trial and error.
But if you keep at it, youβll get there. And over time, youβll understand the common obstacles AI faces, and be able to do it more and more effectively.
Enough theory.
Hereβs how I solved the cover image workflow.
My first step was to map out the actual steps of the process:
What I called it | What it actually was |
|---|---|
Prompting | Finding the prompt in my notes app. |
Choosing between images | Comparing five tabs. Downloading two. Discovering both look wrong once theyβre in the frame. Going back for more. |
Placing them into Figma | Opening the file. Finding the right element. Dropping it in. Nudging it until the crop stops eating his face. Checking it against the grid so three issues in a row donβt look identical. |
Exporting and putting it in the newsletter | Exporting. Too heavy. Uploading it to some compression website. Downloading it again. Creating the beehiiv draft. Dragging it in. Checking the edition number. Getting the edition number wrong. |
I gave Claude Code whatβs in the table above and said:
βHelp!β
βBuild a workflow for creating cover images that look on-brand every single time. But not the same. Him in a cafΓ©, on a train, in a car. Coffee in hand, coffee on the table, coffee somewhere. Laptop open, laptop closed, laptop under his arm. I want all of that, and I want to control it by dragging a slider.β
βI want this fully automated so all I need to do is sit back, look at nice cover images that have been prepared for me and pick one.β
After some back and forth, Claude came up with the plan. It consisted of two parts:
Building an app:
A gallery where I can see the cover images of previous and upcoming editions in a grid
A built-in image editor where I can select between pre-generated cover images for each edition, and make edits to them just using sliders and dropdowns
Some background jobs:
Using Gemini API to generate the set of images for each edition according to that editionβs specs. I can easily generate more variations in the UI.
Communicating with my email platform (beehiiv), and inserting the correct cover image into the correct newsletter edition automatically.
An hour later I had a working app. Then a couple of hours more of back and forths.

I have full overview of previous and upcoming cover images in a grid

I can select between pre-generated images for each newsletter edition, generate more, and edit any aspect of the images easily
The only manual job I do now is look at pretty images and select the best one.
By the time I start writing my newsletter, the cover image is already in there.
Occasionally I use the option to change the camera angle, his location, posture, colours, etc.
I have the option to de-select the coffee cup, as a joke.
He will never stop drinking coffee.
Whatβs left of the process is the part I actually enjoy. Looking strong options and picking the best one. I canβt put into words exactly what I want a cover to be, but I know it when I see it.
My job now is taste.
Realising AI can do most parts of anything is scary.
Youβre allowed to sit and be scared.
Youβre also allowed to pour yourself a fine cup of coffee, get a Claude subscription and lean into it.
(And donβt forget, a good playlist.)
β¦
You are a delight.
What's your verdict on today's email?
See which AI use cases are paying off with Context Windows Pro
Most companies pick AI use cases by brainstorming internally. 90% of those initiatives fail.
Iβve created Context Windows just so you can pick the winners.
π¦ Find high-performing use cases from 2,000+ companies at contextwindows.ai, or book a demo with me
Disclosure: To cover the cost of my email software and the time I spend writing this newsletter, I sometimes work with sponsors and may earn a commission if you buy something through a link in here. If you choose to click, subscribe, or buy through any of them, THANK YOU β it will make it possible for me to continue to do this.


