The 3-metric framework we use to kill or scale AI ad creative every week
Most people testing AI-generated ad creative get stuck because they never define what "working" actually means. Here's the exact scoring framework we run every Friday inside the AI Ad Creative Sprint.
Score every variation on 3 numbers
Hook rate (3-second view rate / thumbstop) — did it stop the scroll?
CTR — did the angle make people want to click?
Cost per result — did it actually convert?
The decision rule
Kill anything below your account average on all three metrics
Scale anything above average on all three
Everything else goes on Watch for one more test cycle
Why most people skip this
Testing without a scoring ritual turns into an excuse to never make a decision — creative just piles up with no verdict. Forcing a weekly decision (kill/watch/scale) is what actually compounds performance over time.
Every hypothesis you test should also start from a one-sentence structure before you even generate the creative:
"If we show [audience] a [creative angle] about [pain/desire], they will [action] because [reason]."
That single sentence is the difference between "testing" and just generating random images and hoping.
If you want the full weekly system — creative generation, this scoring framework, checkout/trial setup, and a repeatable deliverables ritual — check out the AI Ad Creative Sprint.
