Skip to content
Notifications
Clear all

My results after A/B testing 50 ads: Here's what performed best.

1 Posts
1 Users
0 Reactions
5 Views
(@jenniferw)
Trusted Member
Joined: 7 days ago
Posts: 26
Topic starter   [#15061]

Alright, I've just wrapped up a pretty extensive test and wanted to share the raw data here, because I think it cuts through a lot of the marketing hype. I ran a 50-ad A/B test (well, A/B/C/D... you get it) over the last quarter using Anyword, pitting its AI-generated variations against my own copywriting and a couple of control ads. The goal was to see not just *if* it worked, but *how* it worked—and where the hidden costs or trade-offs might lie.

My primary KPI was conversion rate on a mid-funnel content download (a whitepaper), with secondary attention on CPC and click-through rate. I used Anyword primarily for LinkedIn and Meta ad copy, feeding it my product briefs and target audience details (we're in the B2B SaaS space, targeting marketing ops leaders).

Here’s the breakdown of what I found:

**What Anyword Absolutely Nailed:**
* **Variation Generation:** The sheer volume of usable starting points is a genuine time-saver. It eliminated the "blank page" problem.
* **Identifying High-Performance Language:** Its "Performance Score" was surprisingly directionally accurate. Ads it scored above 80 consistently outperformed my controls in initial CTR tests.
* **Headline Formulas:** Its top-performing suggestions often used specific structures, like a "Question + Benefit" format (e.g., "Struggling with attribution? Here's a clear-path method."). This insight is now baked into our manual process.

**Where It Fell Short or Required Heavy Lifting:**
* **Brand Voice Dilution:** Left unchecked, the copy tends toward a generic, enthusiastic marketing tone. I had to heavily edit or set very strict brand voice guidelines to maintain our authority positioning.
* **Feature vs. Benefit Depth:** It's great at surface-level benefits, but struggled to articulate the nuanced, strategic value props that resonate with our sophisticated audience. The "why" was often missing.
* **The "Best Performing" Trap:** Its top-scored ads often had the highest *predicted* CTR, but not always the highest *actual* conversion rate. We saw a few ads with high clicks that attracted lower-quality leads. This is a crucial distinction for ROI.

**The Winning Combination & Final Results:**
The highest-converting ad (a 22% lift over our control) was actually a hybrid. I took a high-scoring Anyword variant that had a compelling hook and structure, but then rewrote the body copy to inject our specific customer journey insights and a stronger, more tailored call-to-action.

**My Takeaway:**
Think of Anyword less as an autopilot and more as an incredibly fast, data-informed brainstorming partner. The real value isn't in taking its output verbatim, but in using its predictive scoring to quickly identify potential winners from a large set, which you then refine with your own audience and strategic knowledge. The hidden cost is the time and expertise needed for that refinement stage—if you don't have a solid copywriter to polish the outputs, you might end up with volume over quality.

Has anyone else done a similar large-scale test? I'm particularly curious if others have seen a discrepancy between its predicted performance (CTR focus) and actual downstream conversion quality. The attribution modeling from click to lead to opportunity is where I'm still digging.

—Jen


—Jen


   
Quote