When you kill, pause the ad, write one sentence on why, and harvest the lesson into the next 3 × 2 batch. Do not leave losers delivering 'just in case'. Do not dump six untested hooks into the only scaling ad set on a Monday. New creative is allowed to fail. Scaling creative is not. The test lane stays open; the scale lane only takes files that cleared the stack.
The log is the unglamorous part that saves the most money. Capture the hook family, the face, the spend, the hook rate, the CTR, the CPA, and the reason code — diagnostic kill, economic kill, policy, fatigue, or offer. Six weeks later nobody remembers whether the ad died from a broken landing page or a 14% open, and without that note the next brief is a guess. Teams that keep a simple retired-creative sheet routinely find that their best new ad of the quarter is a hook family they already killed for the wrong reason, or a face they discarded on preview taste.
Scale what survives gradually: 20–30% budget steps every two to three days rather than an overnight double that sends a healthy CPA back into learning. Keep replacements rendering, because frequency 2.5–3.5 arrives sooner as you spend. Copies of the same open are not diversification; they are frequency with extra filenames. Genuine next tests are a new first line, a new face, or a new proof beat — not twenty clones of the file that won sitting in twelve lookalike ad sets on the same afternoon. Scale spend, or scale genuine variants.