Token导航 LogoToken导航TokenDH.com
研究检索只读clawhub未标认证来源可访问clear审计通过

cb-ab-testing-frameworkcb ab 测试框架

Agent Skill

用于辅助测试设计、自动化测试、用例整理和回归验证。它适合让 Agent 编写单元测试、端到端测试、测试计划或根据失败日志定位问题。使用时需要确认项目测试框架、运行命令和夹具数据,避免为了通过测试而改坏真实逻辑;涉及浏览器或外部服务时,应区分本地模拟、测试环境和生产环境。

总安装

2,246

周安装

90

GitHub Stars

公开资料未说明

下载量

727
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:cb-ab-testing-framework(cb ab 测试框架)
来源仓库:https://github.com/harrylabsj/cb-ab-testing-framework
安装命令:
openclaw skills install cb-ab-testing-framework
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install cb-ab-testing-framework

简介

提供文化敏感的海外 A/B 测试设计框架。

  • 适用于本地化实验规划与偏差控制场景。cb-ab-testing-framework 属于研究检索类 Skill,可作为该场景下的辅助能力补充。
  • 包含细分策略、先验设置与衡量指标建议。
  • 使用前需确认实验周期与样本量是否达标。
  • 建议结合本地用户行为特征调整变量组合。

SKILL.md

slug
cb-ab-testing-framework
version
1.0.0
type
descriptive
language
en

Overseas A/B Testing Design Framework

Overview

This skill provides a structured framework for designing culturally aware experiments when entering or operating in overseas markets. It recognizes that what works in your home market may not transfer, and that running experiments without accounting for cultural, seasonal, and segmentation differences can produce misleading results. The framework covers turning vague growth hypotheses into testable ones, mapping which localization variables to isolate, designing sample and segmentation logic appropriate for each market, building bias and seasonality checklists, creating a learning interpretation template, and maintaining a prioritized experiment backlog.

The framework is designed for growth teams, product managers, UX researchers, ecommerce operators, and data-informed marketers.

When to Use

  • You are running A/B tests in a new overseas market and want to ensure they are properly designed
  • Your home-market tests have been producing results that do not replicate when you apply them overseas
  • You want to build a systematic experiment backlog for international markets rather than running ad-hoc tests
  • You are planning localization changes and want to know which variables to test and how
  • You need to convince stakeholders that a result from one market should or should not be applied to another

Inputs to Collect

  1. Growth objective: what specific business outcome you are trying to move (conversion rate, signup rate, average order value, retention)
  2. Hypothesis or observation: what you have noticed that you believe could be improved (e.g., "our landing page conversion in Germany is lower than expected")
  3. Market(s): which markets are in scope for this experiment
  4. Current baseline metrics: existing conversion rates, traffic volumes, and seasonality patterns in each target market
  5. Localization changes under consideration: what specific changes you are planning (headline translation, visual adaptation, CTA button change, pricing display, trust badge placement)
  6. Traffic and sample availability: estimated weekly visitors per market, which determines how long tests need to run
  7. Team analytics capability: whether you have access to analytics support for statistical significance calculations and results interpretation

Workflow

  1. Turn the overseas growth question into a falsifiable hypothesis that states the market, audience, variable, expected behavior change, rationale, and decision threshold.
  2. Choose localization variables deliberately, separating language, creative, imagery, proof point, offer, price display, trust signal, onboarding step, payment message, and support promise.
  3. Design the experiment structure with market segmentation, sample-size caveats, traffic source control, timing rules, guardrail metrics, and a plan for qualitative interpretation when samples are small.
  4. Prepare a bias and validity checklist covering seasonality, translation quality, novelty effects, mixed audiences, device differences, paid-channel skew, and accidental cross-market averaging.
  5. Translate the result into market-specific next actions, including scale, retest, localize deeper, narrow the segment, or reject the assumption, without assuming a winner should be copied globally.

Output Modules

  1. Experiment Hypothesis Builder — template and three example hypotheses for the target market
  2. Localization Variable Map — categorized list of variables to consider, with the primary test variable identified
  3. Sample and Market Segmentation Logic — sample size calculator template, segmentation approach, and traffic quality checks
  4. Bias and Seasonality Checklist — pre-analysis checklist with documentation format
  5. Learning Interpretation Template — completed template structure for recording results and decisions
  6. Experiment Backlog Prioritization — scoring rubric and backlog format for managing experiments across markets

Example Prompts

  • "We ran a test in the US where changing our CTA button from gray to green increased conversions by 15%. Can we run the same test in Japan and expect the same result?"
  • "Our landing page has a different conversion rate in Germany versus Brazil even though we have not changed anything. Help us design an experiment to understand why."
  • "We want to test whether translating our testimonials into local language increases trust in Southeast Asia. How should we design this test?"
  • "We have a list of 20 localization changes we want to test. How do we prioritize which to run first?"

Safety and Limitations

This framework provides experiment design guidance, not statistical certification. High-stakes decisions (large budget reallocations, permanent product changes, market entry decisions) should not be made on the basis of low-sample or single-test results. Consult analytics or statistics experts for decisions with significant financial or strategic impact. Results from one market should not be automatically generalized to another market without explicit validation.

Acceptance Criteria

  • Turns vague growth observations into structured, falsifiable test hypotheses with a clear primary variable
  • Separates language, offer, creative, and trust-signal variables and identifies which to test independently
  • Includes risk controls for small sample sizes, seasonality, and external events in each target market
  • Provides a prioritization matrix for the experiment backlog using impact, effort, confidence, and learning criteria
  • Prevents overgeneralizing one-market results to other markets with explicit cross-market validation requirements

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

76.45%
按下载量换算556

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

只读

该 Skill 主要提供规则、说明或参考内容,本身偏只读;真正读写文件、联网或执行命令仍取决于宿主 Agent 的任务。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills