A creative test log you'll actually read
Tests without a written record get re-run every quarter at full cost. Log these fields per test, then review on a rhythm that feeds a hypothesis backlog.
Tests without a written record get re-run every quarter at full cost. Log these fields per test, then review on a rhythm that feeds a hypothesis backlog.
A diagnostic taxonomy for retention graphs: eight shapes, the look-alikes that get confused, and the discrimination test that tells them apart before you edit.
A viewer who leaves after getting the answer is not the same as one who leaves bored, but average percentage viewed scores them the same. What to use instead.
An inconclusive Test & Compare result is information, not a glitch. Why variants were too similar or impressions too thin, and what to change before you re-run.
Click-through and retention form a two-by-two that tells you whether to repackage or re-edit. Most creators re-edit videos that only needed a new title.
Stretching a six-minute idea to clear a runtime threshold leaves a visible seam in the retention graph. How to spot it and compare per-video return honestly.
Early data favours whichever variant posted first and often reverses with volume. Pre-commit a sample size and a stop rule before you look.
Lifetime views are dominated by later distribution, not by your hook. A fixed 24-hour checkpoint sheet that reads average watch time instead of view counts.
A two-point CTR gap is invisible until each variant has thousands of impressions. Here is the sizing math, and which videos are worth testing.
Shorts sit outside YouTube's native thumbnail test. A consecutive-post protocol, the metric to hold fixed, and the confound that wrecks most cover tests.
Chasing an absolute virality score is a dead end. Running the predictor on two cuts of the same video is a gate that decides which one ships.