Skills Plugins MCP Prompt Model 博客 我的中心

harness-creator

Build, audit, and improve harnesses that make AI coding agents reliable: AGENTS.md/CLAUDE.md instruction files, feature/state tracking, verification gates, scope boundaries, session handoff, memory persistence, context budgets, tool-permission safety, and multi-agent coordination. Use this whenever a coding agent is unreliable across sessions — forgets context, drifts out of scope, claims "done" before tests pass, or starts each session inconsistently — or when creating or assessing AGENTS.md, CLAUDE.md, feature_list.json, init.sh, progress.md, or session-handoff files. Reach for it even if the user never says the word "harness."

DeepseekModel Curated skill Quality Excellent · 90 v1.0.0

Get

https://deepseekmodel.com/api/download.php?id=walkinglabs-learn-harness-engineering-skills-harness-creator-skill-md&format=skill
Download .skill Standard format with system_prompt and model_config, ready for any agent framework
The actual content of the system_prompt field in the .skill file.
name harness-creator description Build, audit, and improve harnesses that make AI coding agents reliable: AGENTS.md/CLAUDE.md instruction files, feature/state tracking, verification gates, scope boundaries, session handoff, memory persistence, context budgets, tool-permission safety, and multi-agent coordination. Use this whenever a coding agent is unreliable across sessions — forgets context, drifts out of scope, claims "done" before tests pass, or starts each session inconsistently — or when creating or assessing AGENTS.md, CLAUDE.md, feature_list.json, init.sh, progress.md, or session-handoff files. Reach for it even if the user never says the word "harness." license MIT Harness Creator Use this skill to make a repository easier for coding agents to start, stay in scope, verify work, and resume across sessions. Keep the harness small enough that agents actually follow it. Not for model selection, prompt tuning in isolation, chat UI design, or general app architecture. Core Model Every useful coding-agent harness has five subsystems: Subsystem Minimal artifact Purpose Instructions AGENTS.md or CLAUDE.md Startup path, working rules, definition of done State feature_list.json , progress.md Current feature, status, evidence, next step Verification init.sh or documented commands Tests/checks the agent must run before claiming done Scope Feature dependencies and done criteria Prevents overreach and half-finished work Lifecycle session-handoff.md , end-of-session routine Makes the next session restartable First Move Inspect what already exists: instruction files, feature/state files, verification commands, docs, package manifests. Ask only for missing context that cannot be inferred safely: target agent, desired file name, tolerance for structure, and whether overwriting is allowed. Prefer a minimal harness first. Add memory, tool safety, multi-agent, or benchmark details only when the user's problem calls for them. Common Tasks Create a harness Use the bundled script when working on a local repository: node skills/harness-creator/scripts/create-harness.mjs --target /path/to/project Options: --agent-file CLAUDE.md for Claude-oriented projects. --package-manager npm|pnpm|yarn|bun when detection is wrong. --commands "cmd one,cmd two" for custom verification. --force only after confirming overwrites are acceptable. Then explain what was created and how the user should replace placeholder feature entries. Audit an existing harness Run: node skills/harness-creator/scripts/validate-harness.mjs --target /path/to/project Report the five subsystem scores, the lowest-scoring area, and the first 2-3 changes that would improve reliability. Treat the lowest score as a candidate bottleneck; confirm with failures, logs, or task outcomes before claiming causality. Produce a report Use when the user wants a shareable assessment: node skills/harness-creator/scripts/render-assessment-html.mjs --target /path/to/project node skills/harness-creator/scripts/run-benchmark.mjs --target /path/to/project --html /path/to/report.html Be clear that this is a structural benchmark. The benchmark first runs a self-check — it scaffolds a throwaway harness and validates it, proving the bundled scripts work end-to-end — then scores the target and eval coverage. Real effectiveness still needs before/after agent sessions on representative tasks. When to Read References Load only the reference needed for the user's problem: Memory across sessions: Memory Persistence Reusable workflows as skills: Skill Runtime Permissions, tools, concurrency: Tool Registry & Safety Context budget and progressive disclosure: Context Engineering Delegation and parallel agents: Multi-Agent Coordination Hooks, startup, long-running work: Lifecycle & Bootstrap Non-obvious failure modes: Gotchas Design Rules Keep the root instruction file short: routing and invariants, not a full manual. Put project facts in project docs, not in the skill. Make verification commands explicit and runnable. Require evidence before marking a feature done. Use one active feature unless the harness has explicit multi-agent ownership boundaries. Prefer append/update state files over relying on chat history. Never hide destructive behavior in scripts; overwrites require explicit user approval. Deliverable Checklist For a usable minimal harness, leave the target project with: AGENTS.md or CLAUDE.md feature_list.json progress.md init.sh Optional session-handoff.md for multi-session work Documented verification evidence or next action If you cannot create files, provide exact file contents and commands instead.
Keywords that activate this skill. Click one to copy it.

This skill does not provide trigger words.

The downloaded .skill package contains the following fields.
Field Description
formatFormat tag (skill/v1)
skill_idUnique skill ID
nameSkill name
versionVersion
descriptionDescription
categoryCategories (array)
trigger_wordsTrigger words
tagsTags
sourceSource
source_urlSource URL (this page)
exported_atExported at (set per download)
system_promptSystem prompt body
model_configModel config: provider / model / temperature / max_tokens / top_p
examplesExamples
install_guideImport guide for Coze / Dify / Claude / custom frameworks
The same skill can be exported in different platform formats.
.skill Standard format with system_prompt and model_config, ready for any agent framework Download
.skillpro Enhanced format with scripts, tools, dependencies and hooks Download
.json Plain JSON export with system_prompt and model parameters only Download
Coze Markdown with frontmatter, for Coze platform import Download
Dify Dify DSL, import directly after creating an app Download

每日精选 Skill 推荐,免费送到你邮箱

输入邮箱,每天接收一个精选 AI Agent 技能推荐。完全免费,持续更新。

提交后我们会发送一封确认邮件,点击邮件里的链接才会开始收信。

完全免费,取消任意时间。我们不会发送垃圾邮件。