Token导航 LogoToken导航TokenDH.com
研究检索只读clawhub未标认证来源可访问clear审计通过

ad-creative-testing广告创意测试

Agent Skill

用于辅助测试设计、自动化测试、用例整理和回归验证。它适合让 Agent 编写单元测试、端到端测试、测试计划或根据失败日志定位问题。使用时需要确认项目测试框架、运行命令和夹具数据,避免为了通过测试而改坏真实逻辑;涉及浏览器或外部服务时,应区分本地模拟、测试环境和生产环境。

总安装

4,488

周安装

187

GitHub Stars

公开资料未说明

下载量

1,496
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:ad-creative-testing(广告创意测试)
来源仓库:https://github.com/leooooooow/ad-creative-testing
安装命令:
openclaw skills install ad-creative-testing
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install ad-creative-testing

简介

为广告创意、挂钩、目标页面和受众群体设计结构化 A/B 测试假设,并具有明确的成功指标和测试持续时间逻辑。

SKILL.md

name
Ad Creative Testing
description
Design structured A/B test hypotheses for ad creatives, hooks, destination pages, and audience segments with clear success metrics and test duration logic.

Ad Creative Testing

Design structured A/B test hypotheses for ad creatives, hooks, destination pages, and audience segments with clear success metrics and test duration logic. Stop guessing which ad works and start building a repeatable testing machine that improves ROAS with each iteration.

Quick Reference

DecisionStrongAcceptableWeak
Variables tested per experiment1 variable isolated1 primary + 1 secondary (flagged)Multiple variables in one test
Sample size per variant500+ conversions200–499 conversionsUnder 100 conversions
Test duration2–4 weeks1–2 weeks with caveatUnder 7 days
Statistical confidence target95% confidence90% confidenceDeclaring winner under 80%
Primary metric choiceConversion rate or ROASCTR (with caveat)Vanity metric (likes, reach)
Creative variable to test firstHook (first 3 seconds)Offer/headlineBrand colors/logo placement
Budget split50/50 even split70/30 (asymmetric with rationale)One variant gets <20% of budget

Solves

  1. Multi-variable contamination — Testing hook, offer, and format simultaneously means you can't attribute any improvement to a specific change.
  2. Underpowered tests — Declaring a winner on 50 conversions creates false confidence and leads to scaling losers.
  3. Wrong primary metric — Optimizing for CTR when the goal is profit leads to high-traffic, low-converting ads that inflate spend.
  4. Too-short test windows — Ending tests after 3 days misses the natural performance cycle of ads (learning phase, peak, fatigue).
  5. No structured hypothesis — Testing random creative ideas with no documented prediction means learnings don't compound across iterations.
  6. Audience bleed — Running audience A/B tests without proper segment separation means both variants serve the same people, corrupting results.
  7. Ignoring creative fatigue signals — Scaling a winning creative without monitoring frequency and CTR decline leads to wasted spend at the exact moment a test should be run.

Workflow

Step 1 — Define the Test Objective and Primary Metric

Start by answering: what specific business outcome is this test designed to improve? Map the objective to a primary metric:

  • Reduce cost per purchase → Primary metric: Cost per purchase / ROAS
  • Increase click volume on fixed budget → Primary metric: CTR (but validate CTR improvements lead to purchases)
  • Improve video content performance → Primary metric: Video-through rate (VTR) to hook rate to conversion
  • Find the best-converting destination page → Primary metric: Destination page conversion rate (not bounce rate)

Document the primary metric before designing the test. Do not change it after launch.

Step 2 — Write the Hypothesis Statement

A structured hypothesis has three parts:

  • If we change [specific variable]
  • Then we expect [specific measurable outcome]
  • Because [the reasoning based on evidence or prior data]

Example: "If we change the hook from a product demonstration opening to a pain-point question opening, then we expect a 15% improvement in thumb-stop rate and a 10% reduction in cost per initiate checkout, because our audience research shows the target buyer is problem-aware but not solution-aware."

A weak hypothesis: "Let's try a different video style and see if it performs better." No prediction, no reasoning, no measurable outcome.

Step 3 — Isolate the Variable

Identify the single variable you are changing between Variant A (control) and Variant B (challenger). Everything else must remain identical:

  • Hook test: Same offer, same body copy, same CTA, same product, same format — only the first 3 seconds change
  • Offer test: Same creative format, same hook, same visual — only the offer text/structure changes
  • Destination page test: Same ad creative driving to two different destination page variants
  • Audience test: Same creative, same budget, different audience segments (use proper audience exclusions to prevent overlap)
  • Format test: Same offer/copy presented in different formats (15s video vs. static image vs. carousel)

Step 4 — Determine Sample Size and Test Duration

Use the following framework:

  • Minimum sample size: 200 conversions per variant before considering a result meaningful; 500+ for high confidence
  • Minimum duration: 7 days (to capture weekly seasonality patterns); 14 days preferred
  • Budget guidance: If your current ad spend generates 50 purchases/week per variant, you need 4–10 weeks to reach 200–500 conversions — adjust test budget or accept a longer timeline
  • Statistical significance: Use a significance calculator (e.g., AB Testguide, Optimizely Stats Engine) — target 95% confidence; do not declare winners below 90%

Step 5 — Set Up the Test Structure

For paid social (TikTok Ads, Meta Ads):

  • Create a dedicated test campaign or ad set
  • Use even 50/50 budget split unless you have a specific reason to weight differently
  • Disable automatic creative optimization during the test (prevents the platform from picking a winner before you have enough data)
  • Set start/end dates and document them
  • Confirm the test is running on the correct audience and that audience exclusions are in place if testing segments

Step 6 — Monitor During the Test

Check performance at regular intervals (not daily — resist the urge to call a winner early):

  • Day 3–4: Verify both variants are delivering and spending approximately equally (not a data review — just a delivery check)
  • Day 7: Check if there are any technical issues (a variant not spending, creative rejected); do not make creative decisions yet
  • Day 14: First data review; check sample sizes; run significance test if above 200 conversions per variant
  • Day 21–28: Final read if sample size reached; declare winner or extend if still underpowered

Watch for these early kill signals (valid reasons to stop a test before planned end):

  • One variant has a CPA 3× or higher than the other after 100+ conversions (likely a strong loser; killing it quickly saves spend)
  • One variant has delivery problems and is not spending

Step 7 — Document the Result and Build the Learning

After the test concludes:

  • Record the hypothesis, test structure, results, and winning variant
  • Calculate the magnitude of improvement (e.g., "hook B reduced CPA by 22%")
  • Identify what the result implies for the next test (e.g., "pain-point hooks outperform demonstration hooks for this audience — next test: which pain point resonates most?")
  • Add to a test log that accumulates learnings across campaigns

This log becomes your competitive advantage over time.

Examples

Example 1 — Hook A/B Test for TikTok Shop Product

Input:

  • Product: Skincare serum, TikTok Shop UK
  • Current creative: Opens with 3-second product shot + "Now available in the UK"
  • Hypothesis: Pain-point hook will outperform product demonstration hook
  • Primary metric: Cost per purchase (current: £18)
  • Weekly volume: ~60 purchases/week

Structured Test Design:

TEST HYPOTHESIS
If we change the video hook from a product shot ("Now available in the UK") to a
pain-point question ("Struggling with dull skin even after your skincare routine?"),
then we expect a 20% reduction in cost per purchase,
because our top-performing organic videos use problem framing and our current
hook has a 15% thumb-stop rate vs. the 25–30% we see on viral skincare content.

VARIABLE BEING TESTED
Variant A (Control): Opens with close-up product shot + "Now available in the UK"
Variant B (Challenger): Opens with creator asking "Struggling with dull skin even
after your skincare routine?" — same body copy, same CTA, same offer

EVERYTHING IDENTICAL IN BOTH VARIANTS
✓ Offer: same (no discount, standard price)
✓ Body copy: same
✓ Button text: "Shop Now" in both
✓ Video length: 15 seconds in both
✓ Target segment: same (UK, F 25–44, niche: skincare)
✓ Budget: £50/day each, 50/50 split

SAMPLE SIZE & DURATION PLAN
Target: 200+ purchases per variant
Current rate: 60/week × £50/day test budget ÷ current £100/day = ~30/week per variant
Minimum test duration: 7 weeks to reach 210 purchases per variant
Decision: Run for 8 weeks to be safe; check statistical significance at week 6

SUCCESS CRITERIA
- Primary: Variant B achieves ≥15% lower cost per purchase than Variant A with ≥90% statistical confidence
- Secondary: Variant B thumb-stop rate (3-second view rate) is higher than Variant A
- Kill switch: If either variant reaches CPA of £40+ after 100 purchases, kill it and investigate

Example 2 — Landing Page A/B Test for DTC Brand

Input:

  • Product: Protein supplement, Shopify store
  • Ad platform: Generic homepage
  • Test idea — Product-specific landing page vs. homepage
  • Primary metric — Landing page conversion rate (current — 1.8%)
  • Monthly traffic to landing — ~8,000 visitors/month

Structured Test Design:

TEST HYPOTHESIS
If we send ad traffic to a dedicated product landing page (with product video,
reviews, and FAQ above the fold) instead of the generic homepage,
then we expect landing page conversion rate to increase from 1.8% to 2.5%+,
because product-specific pages remove navigation distractions and maintain
message-match with the ad creative.

VARIABLE BEING TESTED
Variant A (Control): Traffic → Homepage (generic, navigation visible)
Variant B (Challenger): Traffic → Dedicated product landing page (no nav, product
video hero, 5 reviews, FAQ, single CTA)

SAMPLE SIZE CALCULATION
Current conversion rate — 1.8% (to detect 2.5% with 95% confidence, 80% power)
Required visitors per variant — ~2,400 (use AB Testguide calculator)
Current monthly traffic to this landing — 8,000/month
50/50 split — 4,000 per variant per month
Estimated time to significance — ~18 days (assuming even traffic distribution)
Duration — Run for 21 days minimum to capture day-of-week patterns

SUCCESS CRITERIA
- Primary — Variant B conversion rate exceeds Variant A by ≥15% with ≥95% confidence
- Secondary — Revenue per visitor (not just conversion rate — larger carts matter)
- Kill switch — No kill switch for low-performing variant; this is a page test, not a spend test

Common Mistakes

  1. Changing two things and calling it an A/B test — Testing a new hook AND a new offer simultaneously means any improvement (or degradation) is unattributable. Isolate one variable per test.
  1. Declaring a winner after 3 days — Most ad platforms have a 7-day learning phase. Early data is noisy, especially for conversion-focused campaigns. Decisions made on day 3 are often wrong.
  1. Using CTR as the primary metric when you care about purchases — Ads with high CTR and low conversion rates increase spend without increasing revenue. Always validate that CTR improvements translate to downstream conversion improvements.
  1. Not calculating required sample size before starting — If your current volume means you'd need 6 months to reach significance, you should increase test budget, widen the test window, or pick a higher-frequency metric as a leading indicator.
  1. Running audience tests without exclusions — If Audience A and Audience B overlap (e.g., both are "females 25–44 interested in beauty"), the same person can be served both variants, corrupting the test.
  1. Letting the platform auto-optimize mid-test — Most paid social platforms have creative optimization features that will automatically shift budget toward the "better" performing creative. Disable this during a test — it will pick a winner long before you have statistical significance.
  1. Not documenting hypotheses before seeing results — Writing a "hypothesis" after you see the data is confirmation bias, not testing. Record your prediction before the test starts.
  1. Scaling a winner without monitoring creative fatigue — Winning creatives eventually fatigue. Monitor CTR and frequency weekly after scaling; begin a new iteration test before performance declines.

Resources

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

70.25%
按下载量换算1,051

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

只读

该 Skill 主要提供规则、说明或参考内容,本身偏只读;真正读写文件、联网或执行命令仍取决于宿主 Agent 的任务。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills