Token导航 LogoToken导航TokenDH.com
研究检索需要联网clawhub未标认证来源可访问clear审计通过

battle-tested-agent经过战斗考验的特工

Agent Skill

用于辅助测试设计、自动化测试、用例整理和回归验证。它适合让 Agent 编写单元测试、端到端测试、测试计划或根据失败日志定位问题。使用时需要确认项目测试框架、运行命令和夹具数据,避免为了通过测试而改坏真实逻辑;涉及浏览器或外部服务时,应区分本地模拟、测试环境和生产环境。

总安装

17,597

周安装

705

GitHub Stars

公开资料未说明

下载量

5,696
OpenClaw

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

MIT-0

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:battle-tested-agent(经过战斗考验的特工)
来源仓库:https://github.com/zurbrick/battle-tested-agent
安装命令:
openclaw skills install battle-tested-agent
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 OpenClaw 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

ClawHubOpenClaw
openclaw skills install battle-tested-agent

简介

用于辅助测试设计、自动化测试和回归验证,帮助 Agent 编写单元测试或定位问题。

  • 适合需要编写测试用例、分析失败日志或维护测试框架的开发场景。
  • 通过 clawhub 安装后,在 OpenClaw 中调用以生成测试计划或执行验证任务。
  • 需确认项目测试框架和运行命令,避免修改真实逻辑;涉及浏览器时区分环境使用。
  • 注意权限范围和维护状态,防止误操作生产环境或触发不必要的外部服务调用。

SKILL.md

name
battle-tested-agent
version
1.5.0
description
>
author
Zye ⚡ (Don Zurbrick)
license
MIT
tags
[production, reliability, memory, compaction, multi-agent, security, self-improvement, heartbeat, delegation, battle-tested]
homepage
https://github.com/zurbrick/battle-tested-agent
metadata
openclaw
emoji
⚔️
requires
bins
["bash", "grep", "find", "wc"]
optionalBins
["openclaw"]

Battle-Tested Agent

19 production-hardened patterns for AI agents. Every one earned from failure.

Use this skill when you are:

  • hardening an agent that will run repeatedly or autonomously
  • tightening memory, verification, or anti-hallucination behavior
  • reducing compaction failures, weak handoffs, or orchestration drift
  • reviewing an agent workspace for missing production patterns
  • debugging why an agent keeps losing context, guessing, or dropping work

Do not use this skill for:

  • persona writing or onboarding polish
  • one-off prompt tweaks with no reusable pattern behind them
  • adding new tools, servers, or runtime capabilities
  • turning a simple workspace into process theater

Default workflow

  1. Audit first

Run bash scripts/audit.sh <workspace> to see which patterns are present. The script checks for all 16 patterns and tells you what to fix first.

  1. Start with the smallest tier that fits

Implement starter patterns first, then intermediate, then advanced. Do not cargo-cult every pattern into every agent.

  1. Patch the actual failure mode

Change the mechanism, not just the wording. "ALWAYS check X" is not a fix — a verification gate is a fix.

  1. Keep patterns lightweight

Add only the pieces that materially reduce failures or operator burden.

Pattern tiers

  • Starter (5): baseline reliability for almost every agent
  • Intermediate (5): daily-driver patterns for briefs, heartbeats, and recurring work
  • Advanced (6): multi-agent orchestration, handoffs, and self-improvement discipline

Pattern clusters

Some patterns reinforce each other naturally. Adopt them together when the failure mode calls for it:

  • Trust chain: WAL Protocol + Anti-Hallucination + Agent Verification — ensures

data is captured, sourced, and measured before reporting

  • Handoff loop: Delegation Rules + Completion Contract + Acceptance Gate + Task State Tracking — prevents

work from disappearing between agents or being certified without proof

  • Survival kit: Working Buffer + Compaction Injection Hardening + Silent Worker Recovery — keeps context

alive across long sessions and prevents silent delegated drift

  • Quality gate: QA Gates + Verify Implementation + Decision Logs — ensures output

quality and traceable reasoning

  • Delegation hardening: Brief Quality Gate + Scoped Verifier Gate — keeps delegation tight without turning the whole system into bureaucracy

When patterns conflict

If two patterns seem to give contradictory advice:

  • Safety patterns win over speed patterns. Ambiguity Gate overrides Simple Path First

when the request is ambiguous. Verify before acting, even if the simple path is obvious.

  • Evidence patterns win over action patterns. Anti-Hallucination overrides "just try it"

when reporting data. Never guess a number to move faster.

Assets — how to use them

The assets/ folder contains starter files you copy into your workspace and customize. They are templates, not drop-in replacements.

# Merge delegation and decision log rules into your existing AGENTS.md
cp assets/AGENTS-additions.md ~/workspace/ # Review, then merge

# Add QA gates
cp assets/QA-gates.md ~/workspace/QA.md

# Set up self-improvement tracking
mkdir -p ~/workspace/.learnings
cp assets/learnings-template.md ~/workspace/.learnings/LEARNINGS.md
cp assets/errors-template.md ~/workspace/.learnings/ERRORS.md
cp assets/features-template.md ~/workspace/.learnings/FEATURE_REQUESTS.md

Read references/audit-usage.md for the full rollout order and bootstrap workflow.

References

  • references/starter-patterns.md — WAL, anti-hallucination, ambiguity, simple-path-first, unblock-before-shelve
  • references/intermediate-patterns.md — verification, working buffer, QA gates, decision logs, verify implementation
  • references/advanced-patterns.md — delegation, brief quality, proof-based handoffs, acceptance gates, orchestration, stale-worker recovery, compaction hardening, recurrence tracking
  • references/audit-usage.md — audit script usage, install/copy snippets, and expected outcomes

Included scripts

  • scripts/audit.sh — workspace audit for all 19 patterns (supports AGENTS.md, CLAUDE.md, SOUL.md, and system.md)

Rules of thumb

  • Audit before expanding
  • Prefer progressive disclosure over giant core files
  • Silence is better than hallucination
  • Ambiguity is a stop sign, not permission
  • The orchestrator should preserve oversight, not sink into implementation
  • Mechanism changes beat wording changes
  • After acting, verify the new state before declaring success
  • Partial progress is not success; recovery steps matter as much as first-attempt steps

Outcome

A leaner, more resilient agent that survives compaction, hands work off cleanly, reports only what is verified, and improves without spiraling into bureaucracy.

适合场景

01

OpenClaw 用户查找和安装 Skill 时

02

用户想查找某类 Agent Skill 时

03

需要根据任务场景推荐可安装能力包时

04

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

补充不同宿主或平台的使用分布数据

能力 5

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

OpenClaw

81.33%
按下载量换算4,633

安全审计

VirusTotal

通过

ClawScan

通过

Static analysis

通过

权限和风险

需要联网

该 Skill 可能需要联网访问来源站点、仓库或外部 API;具体网络访问范围需要结合源码和 README 复核。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills