Source profileQuality 87/100Review permissions

tamdogood/builder-essential-skills/skills/run-smoke-tests/SKILL.md

run-smoke-tests

Inspect an unfamiliar repository, interpret a broad Markdown user journey at runtime, operate the real product through its supported web, API, CLI, desktop, or mobile surface, and produce an auditable pass, fail, or blocked judgment with screenshots, logs, recordings, and a step timeline when available. Use when asked to run smoke tests, validate an end-to-end user journey, dogfood a product, test a release candidate, or execute a non-deterministic workflow that may include authentication or hum

Source repository stars
88
Declared platforms
0
Static risk flags
2
Last source update
2026-08-04
Source checked
2026-08-04

Decision brief

What it does—and where it fits

Execute a broad user journey against a real build and leave enough evidence for a person to audit the judgment. Understand the host project before starting the product or touching its state.

Best for

  • Use when asked to run smoke tests, validate an end-to-end user journey, dogfood a product, test a release candidate, or execute a non-deterministic workflow that may include authentication or hum

Not for

  • Tasks that require unconfirmed production actions or broad system permissions.
  • Environments where the pinned source and install steps cannot be inspected.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexNot declaredNo explicit evidencePortability before use
Claude CodeNot declaredNo explicit evidencePortability before use
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/tamdogood/builder-essential-skills --skill "skills/run-smoke-tests"
Safe inspection promptEditorial

Inspect the Agent Skill "run-smoke-tests" from https://github.com/tamdogood/builder-essential-skills/blob/3bbbfb668959af812f3892537f263c64cedc12c4/skills/run-smoke-tests/SKILL.md at commit 3bbbfb668959af812f3892537f263c64cedc12c4. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    Workflow

    Read the nearest repository instructions, product README, runbook, contribution guide, manifests, launch scripts, existing smoke or end-to-end tests, and the smallest relevant product documentation. Search for auth setup, seed data, test accounts, feature flags, observability, a…

    represents one coherent user goal;names its environment and allowed mutations;defines critical and noncritical oracles;
  2. 02

    Operating contract

    Treat the Markdown as runtime instructions, not source to compile into code.

    Treat the Markdown as runtime instructions, not source to compile into code.Test through a supported user-facing surface: web UI, API, CLI, desktop,Separate execution from repair. Do not edit the product, tests, or journey
  3. 03

    1. Pass the repository-understanding gate

    Read the nearest repository instructions, product README, runbook, contribution guide, manifests, launch scripts, existing smoke or end-to-end tests, and the smallest relevant product documentation. Search for auth setup, seed data, test accounts, feature flags, observability, a…

    Read the nearest repository instructions, product README, runbook, contribution guide, manifests, launch scripts, existing smoke or end-to-end tests, and the smallest relevant product documentation. Search for auth setu…Before launching anything, be able to state:Do not guess missing commands, URLs, credentials, or permissions. If a missing item blocks safe execution, report it as a preflight blocker.
  4. 04

    2. Validate the journey contract

    Read references/smoke-format.md. Normalize supplied prose without silently adding product behavior.

    represents one coherent user goal;names its environment and allowed mutations;defines critical and noncritical oracles;
  5. 05

    3. Preflight without changing product behavior

    Build or launch the product with documented commands. Confirm health, test-data availability, driver connectivity, recording support, and enough storage for artifacts.

    Build or launch the product with documented commands. Confirm health, test-data availability, driver connectivity, recording support, and enough storage for artifacts.Prefer the repository's existing artifact directory. Otherwise use .context/smoke-runs/ when .context already exists; if it does not, create a temporary directory and report its absolute path. Do not add generated evide…Create a run manifest containing:

Permission review

Static risk signals and limitations

Reads files

low · line 25

The documentation asks the agent to read local files, directories, or repositories.

Read the nearest repository instructions, product README, runbook,

Runs scripts

medium · line 166

The documentation asks the agent to run terminal commands or scripts.

[Developer CLI first run](examples/developer-cli-first-run.smoke.md), covering

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score87/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars88SourceRepository attention, not individual Skill quality
Compatibility0 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
tamdogood/builder-essential-skills
Skill path
skills/run-smoke-tests/SKILL.md
Commit
3bbbfb668959af812f3892537f263c64cedc12c4
License
MIT
Collected
2026-08-04
Default branch
main
View the original SKILL.md

Run Smoke Tests

Execute a broad user journey against a real build and leave enough evidence for a person to audit the judgment. Understand the host project before starting the product or touching its state.

Operating contract

  • Treat the Markdown as runtime instructions, not source to compile into code.
  • Test through a supported user-facing surface: web UI, API, CLI, desktop, mobile, or a documented combination.
  • Separate execution from repair. Do not edit the product, tests, or journey during a run. Finish the report before starting any requested fix.
  • Default to local, test, preview, or staging environments. Never mutate production, spend money, send real invitations, or change real user data without explicit authorization.
  • Make every pass or fail claim traceable to captured evidence.
  • Keep the first failed run. Do not erase evidence by retrying until green.

Workflow

1. Pass the repository-understanding gate

Read the nearest repository instructions, product README, runbook, contribution guide, manifests, launch scripts, existing smoke or end-to-end tests, and the smallest relevant product documentation. Search for auth setup, seed data, test accounts, feature flags, observability, artifact conventions, and cleanup commands. Trace the journey across the relevant UI, service, state, and external-system boundaries so failures can be attributed to the right part of the product.

Before launching anything, be able to state:

product and user surface:
authoritative journey source:
relevant architecture and system boundaries:
build and start commands:
target environment and data boundary:
interaction driver or available tools:
authentication and human checkpoints:
evidence capture capabilities:
safe cleanup path:
known abort conditions:

Do not guess missing commands, URLs, credentials, or permissions. If a missing item blocks safe execution, report it as a preflight blocker.

2. Validate the journey contract

Read references/smoke-format.md. Normalize supplied prose without silently adding product behavior.

Confirm that the journey:

  • represents one coherent user goal;
  • names its environment and allowed mutations;
  • defines critical and noncritical oracles;
  • identifies authentication, payment, security, or destructive checkpoints;
  • has a cleanup or quarantine plan;
  • names conditions that require an immediate abort.

Move small deterministic behaviors into scenario tests. Keep smoke coverage for the larger path and the seams between systems.

3. Preflight without changing product behavior

Build or launch the product with documented commands. Confirm health, test-data availability, driver connectivity, recording support, and enough storage for artifacts.

Prefer the repository's existing artifact directory. Otherwise use .context/smoke-runs/ when .context already exists; if it does not, create a temporary directory and report its absolute path. Do not add generated evidence to version control unless the user asks.

Create a run manifest containing:

journey ID and source path
commit or build identifier
environment and base URL or executable
start time and executor
human checkpoints
recording and log sources

Abort before step one when the build is unhealthy, the environment is wrong, the account or data boundary is unsafe, or required evidence capture is unavailable for a critical oracle.

4. Interpret and execute the Markdown

Follow the journey in order through the real product surface. For each step, record:

timestamp
instruction
action taken
observable result
evidence path or log reference
provisional judgment
deviation, latency, or uncertainty

Use normal user affordances. Do not call internal APIs to bypass UI steps unless the journey explicitly tests an API or uses setup APIs only for isolated fixtures.

Pause at a declared human checkpoint. State the exact action needed and resume from the same state after the human completes it. Never request that a user paste secrets into chat.

If the driver supports video, record the complete run and generate subtitles or a timestamped transcript from the step timeline. Otherwise capture screenshots, terminal output, structured logs, and state snapshots. Do not claim to have a video when the environment cannot produce one.

5. Judge once, retry carefully

Assign one primary result:

  • PASS: every critical oracle was observed and no abort condition occurred;
  • FAIL: a critical oracle was violated or an unexpected product error prevented the goal;
  • BLOCKED: an external prerequisite, permission, environment, or human checkpoint prevented a valid judgment.

Add FLAKY as a flag when the same build and inputs produce inconsistent results.

Do not convert missing evidence into a pass. A noncritical issue may preserve a pass only when the report calls it out separately.

Retry at most once, and only when evidence points to an environmental or driver failure rather than a product failure. Preserve both runs and explain why the retry was allowed.

6. Clean up and hand off

Run only the cleanup authorized by the journey. If cleanup fails, quarantine the test data, record identifiers, and report the residue. Do not hide a cleanup failure behind a passing product judgment.

Write a report with:

  1. result, confidence, build, environment, and duration;
  2. a timestamped step table;
  3. critical oracle results with evidence links;
  4. human checkpoints and deviations;
  5. console, network, log, screenshot, and recording locations;
  6. cleanup result and residual data;
  7. candidate deterministic tests exposed by the run.

The final user message must distinguish product failures, test-environment failures, and uncertain judgments.

Demonstration examples

Use the examples to demonstrate runtime interpretation. Replace every command, URL, account, and expected result with evidence from the host repository.

Failure handling

  • Product cannot start: capture build and launch output, mark BLOCKED, and do not improvise a different environment.
  • Authentication unavailable: pause at the checkpoint or mark BLOCKED; never bypass it with unapproved credentials.
  • Product behavior diverges from the Markdown: preserve evidence and mark FAIL unless the behavior source is genuinely contradictory.
  • Evidence capture fails mid-run: abort when it affects a critical oracle; otherwise continue with the limitation visible.
  • A defect is found: finish the report first. Diagnose or fix only under a separate user request.

Alternatives

Compare before choosing

Computed 9532,606

K-Dense-AI/scientific-agent-skills

simpy

Build, inspect, test, and analyze bounded process-based discrete-event simulations with SimPy, including events, resources, interrupts, monitoring, replications, warm-up, and reproducible output analysis.

Computed 9410,895

huggingface/skills

hf-cloud-sagemaker-production-defaults

Create a SageMaker endpoint (real-time, real-time scale-to-zero, or async) with autoscaling, CloudWatch alarms, and tagging enabled by default. Use this skill whenever about to create a SageMaker endpoint, write deployment code that calls `create_endpoint`, or finalize a deployment after the image URI and IAM role are known. Provides deploy.py for real-time endpoints, deploy_ic.py for real-time endpoints that scale to zero instances via inference components, and deploy_async.py for async endpoin

Computed 9482

aAAaqwq/AGI-Super-Team

trade-prediction-markets

Build and test Polymarket prediction market trading strategies for YES/NO token trading. Provides 6 tools: get_all_prediction_events (browse markets, $0.001), get_prediction_market_data (analyze price history, $0.001), create_prediction_market_strategy (generate code, $1-$4.50), run_prediction_market_backtest (test performance, $0.001). Trade on real-world events (politics, economics, sports, crypto). Currently simulation only (live deployment coming soon).

Computed 9337,425

github/awesome-copilot

flowstudio-power-automate-build

Build, scaffold, and deploy Power Automate cloud flows using the FlowStudio MCP server. Your agent constructs flow definitions, wires connections, deploys, and tests — all via MCP without opening the portal. Load this skill when asked to: create a flow, build a new flow, deploy a flow definition, scaffold a Power Automate workflow, construct a flow JSON, update an existing flow's actions, patch a flow definition, add actions to a flow, wire up connections, or generate a workflow definition from