Okay so you've got an AI system. You've spent some money. You've spent more hours than you want to count tweaking prompts and uploading documents and watching YouTube videos at 1.5x speed. And now you have this nagging feeling that maybe it's not doing what you hoped, but you don't want to admit it because you've sunk a lot into it.
I have a name for that feeling. It's called the "I'm still doing all the work" feeling. And there's a test I run on my own systems to figure out if I should keep them, fix them, or kill them. I call it the You-Replacement Test.
The premise is simple. If you handed your AI system to a smart assistant who has never met you, could that assistant produce most the work that comes out of your office without you in the room? Not 100. We're not trying to clone you. Eighty.
Here are the seven questions I run through.
Question 1: Does it sound like me on a bad day?
Not your best writing day. Your tired Tuesday writing day. The day where you cranked out an email at 4pm between calls.
Most people test their AI against their best work and then are disappointed when it doesn't match. Wrong test. Your AI should match your average day, which is the day that fills your inbox and content calendar. If the AI is too polished, it's wrong. If it's missing your weird little turns of phrase, it's wrong.
This is a Claude Projects strength, by the way. If you've fed the Project enough of your actual writing (the messy stuff, not only the published stuff), the output starts catching your rhythm. As of mid-2026, you can have persistent custom instructions plus an attached knowledge base of your own writing samples. That's the combination that gets you to "sounds like me on a bad day."
Question 2: Can someone else use it without calling me?
Imagine you go on vacation for a week. No phone. No Slack. (I know, terrifying. Stay with me.)
Your assistant or your VA opens up your AI system and tries to draft a follow-up email to a new client. Do they have everything they need? The pricing table, your standard objection responses, the way you sign off, the tone you use when someone is on the fence?
If they have to text you, the system failed. The whole point is portability of you.
Question 3: Does it stop me from re-explaining the same thing for the fifth time?
I have a coaching client, let's call her Dana. She's a financial coach with about 90 active clients. She kept telling me she "uses AI for content" but then I asked her how often she still types out her credentialing story or her three-step money framework into a new chat. Five times a week. Maybe more.
That's not a working AI system. That's an expensive autocomplete.
A real system stores the things you'd otherwise re-explain. Your origin story, your frameworks, your standard answers to objections, your case studies, your pricing logic. Once. Forever. Available to the AI on demand. If you're still copy-pasting your own bio into prompts, your knowledge layer is broken.
Question 4: Does it know when to NOT respond?
This one trips people up. Good AI systems include rules for what not to do. Don't reply to refund requests automatically. Don't draft anything political. Don't promise timelines you haven't approved. Don't quote prices that haven't been finalized.
If your AI is willing to answer anything you throw at it, that's not freedom, that's risk. The systems that scale well have guardrails written into the brand voice Skill or the Project instructions. Mine has about a dozen "do not" rules and they've saved me embarrassment more than once.
Question 5: Does the output need more than 15 minutes of editing per piece?
This is the practical one. Time it. Set an actual timer.
If a 600-word email draft needs 25 minutes of editing, your AI is not saving you time. You'd have been faster writing it yourself. The whole point of the system is to compress your editing time, not your typing time. (There's a difference. Typing was never the slow part for most of us. Decision-making and structure are the slow part.)
When my system is dialed in, an email goes from 35 minutes of writing to maybe 7 minutes of editing. That's the win. Not "I clicked a button and it published."
Question 6: Does it improve when I correct it?
Here's where I see entrepreneurs lose patience. They correct the AI once, the next output is wrong in the same way, they assume the AI is dumb, and they quit.
Wrong. The AI is doing exactly what you told it to do. If you didn't write the correction into a Skill or a Project instruction, the AI has no memory of your feedback. (That's not a bug. That's how the tools work, at least at the time I'm writing this.)
A working system has a feedback loop. Every time you correct something, you update the underlying Skill or the Project's custom instructions. The system gets sharper week over week instead of repeating the same mistakes.
Quick tangent. This is the part of AI nobody markets, because "you have to maintain it" is not a sexy selling point. But it's the truth. A garden you don't water dies.
Question 7: Does it free up your highest-value hours?
The final question, and maybe the most important. What did you do with the time the AI gave you back?
If the answer is "I'm doing the same volume of work, with more output," your AI is a treadmill. You're moving faster but going nowhere different. The point of the You-Replacement Test isn't to produce more content. It's to free up the hours you'd otherwise spend on production so you can do the things only you can do. Sales calls. Strategy. Recording the next round of source material so the system gets sharper. Going outside. Whatever.
If you can't point to a specific block of time the AI gave back to you, the system isn't replacing you. It's distracting you.
Scoring it
Run all seven on your current system. Be honest. (You don't have to tell anyone.)
Seven yesses means you have a real working system and you should stop tweaking it and go sell more.
Five or six means you have a system with one or two leaks. Fixable in a weekend.
Four or fewer means you don't have a system, you have a tool collection. That's a different problem and it requires the kind of foundational work I write about in the 5-layer AAOS piece.
What to do with the result
Don't beat yourself up if you scored low. Most people score low the first time. (I scored a three the first time I ran this on myself in early 2025. Three out of seven. Humbling.)
Then I rebuilt the brand voice layer, redocumented my knowledge layer, and the score climbed. It took me about four months of incremental work, not a heroic weekend.
If you want the full step-by-step on how to fix each of the seven failure points, that's the bulk of what I teach inside my Level Up training. It's where I walk through the exact diagnostic and the rebuild process I used on my own business. You can grab it at kristamashore.com/LevelUp.