fastrepl/anarlog/.agents/skills/qa-cli-mcp-api/SKILL.md
qa-cli-mcp-api
Select and run explicitly requested, risk-based QA for Anarlog's CLI, webhooks, stdio MCP, hosted Cloud API, and remote MCP. Test only affected lanes unless comprehensive coverage is requested.
- Source repository stars
- 9,191
- Declared platforms
- 0
- Static risk flags
- 2
- Last source update
- 2026-08-28
- Source checked
- 2026-08-28
Decision brief
What it does: where it fits
Default to the smallest set of programmatic-interface lanes that can prove the change. Automated tests are necessary but do not replace a live smoke test when the changed boundary is only exercised by a real client or deployment.
Not for
- Tasks that require unconfirmed production actions or broad system permissions.
- Environments where the pinned source and install steps cannot be inspected.
Compatibility matrix
Platform support, with evidence labels
| Platform | Status | Evidence | What to check |
|---|---|---|---|
| Codex | Not declared | No explicit evidence | Portability before use |
| Claude Code | Not declared | No explicit evidence | Portability before use |
| Cursor | Not declared | No explicit evidence | Portability before use |
| Gemini CLI | Not declared | No explicit evidence | Portability before use |
Installation
Inspect first. Install second.
The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.
npx skills add https://github.com/fastrepl/anarlog --skill ".agents/skills/qa-cli-mcp-api"Inspect the Agent Skill "qa-cli-mcp-api" from https://github.com/fastrepl/anarlog/blob/8f976e8709d1b3444102d9a4f39c8397126a059a/.agents/skills/qa-cli-mcp-api/SKILL.md at commit 8f976e8709d1b3444102d9a4f39c8397126a059a. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.
Workflow
What the source asks the agent to do
- 01
Choose the scope first
Inspect the exact branch, commit, or diff and map changed code to its direct consumers before creating fixtures or credentials. In a GitButler workspace, use but status and but show ; do not use the synthetic workspace HEAD as the candidate or combine unrelated applied branches.
CLI parsing, output, or local DB access → CLI plus the closest contract tests.Webhook endpoints, signing, or delivery retries → the webhook cases only.Shared agent-access DTOs, exports, filtering, or pagination → every direct - 02
Targeted regression mode (default)
Select lanes by behavior, not by the existence of this checklist:
CLI parsing, output, or local DB access → CLI plus the closest contract tests.Webhook endpoints, signing, or delivery retries → the webhook cases only.Shared agent-access DTOs, exports, filtering, or pagination → every direct - 03
Comprehensive Interface QA
Run every fixture, baseline, live lane, lifecycle case, privacy check, and cross-surface comparison below only when the user explicitly requests full or comprehensive programmatic-interface QA.
Run every fixture, baseline, live lane, lifecycle case, privacy check, and cross-surface comparison below only when the user explicitly requests full or comprehensive programmatic-interface QA. - 04
Safety and evidence
Use a dedicated QA account and non-sensitive fixture meetings.
Use a dedicated QA account and non-sensitive fixture meetings.Use a Pro or trialing account for the hosted lane and a separate free orNever put API keys, JWTs, webhook secrets, database credentials, or personal - 05
Comprehensive QA Fixture
Create two completed QA meetings through the desktop app:
A standalone meeting whose title contains a unique run marker such asTwo meetings in the same recurring series.a note and at least one generated summary;
Permission review
Static risk signals and limitations
Writes files
The documentation asks the agent to create, modify, or delete local files.
Remove that directory and revoke all generated keys after the run.Runs scripts
The documentation asks the agent to run terminal commands or scripts.
cargo test -p anarlog-cliRuns scripts
The documentation asks the agent to run terminal commands or scripts.
cargo test -p tauri-plugin-local-apiEvidence record
Why each signal appears
| Signal | Value | Evidence type | Meaning |
|---|---|---|---|
| Quality score | 90/100 | Computed | Documentation, specificity, maintenance, and trust rules |
| Repository stars | 9,191 | Source | Repository attention, not individual Skill quality |
| Compatibility | 0 platforms | Source | Declared in the catalog source record |
| Usage guide | automated source guide | Editorial | Generated or reviewed according to the visible evidence level |
Pinned source
Provenance and original SKILL.md
- Repository
- fastrepl/anarlog
- Skill path
- .agents/skills/qa-cli-mcp-api/SKILL.md
- Commit
- 8f976e8709d1b3444102d9a4f39c8397126a059a
- License
- MIT
- Collected
- 2026-08-28
- Default branch
- main
View the original SKILL.md
QA: CLI, MCP, and API
Default to the smallest set of programmatic-interface lanes that can prove the change. Automated tests are necessary but do not replace a live smoke test when the changed boundary is only exercised by a real client or deployment.
This QA workflow is independent from releasing. A release request alone does not invoke it, and its results do not approve or block a release.
Choose the scope first
Inspect the exact branch, commit, or diff and map changed code to its direct
consumers before creating fixtures or credentials. In a GitButler workspace,
use but status and but show <commit-or-branch>; do not use the synthetic
workspace HEAD as the candidate or combine unrelated applied branches.
Find the original reproduction in the current or past Codex task, linked issue, PR, support report, or regression test. State the selected lanes and the reason for each before testing.
Targeted regression mode (default)
Select lanes by behavior, not by the existence of this checklist:
- CLI parsing, output, or local DB access → CLI plus the closest contract tests.
- Webhook endpoints, signing, or delivery retries → the webhook cases only.
- Shared agent-access DTOs, exports, filtering, or pagination → every direct consumer, including hosted REST or remote MCP when affected, plus cross-surface parity only for the changed fields.
- Hosted auth, snapshots, entitlements, isolation, purge, or Supabase policy → the affected hosted REST lifecycle and negative cases.
- Local or remote MCP protocol/tool changes → that MCP lane and its direct transport/contract dependency.
- Shared hosted REST/MCP behavior → both hosted consumers, but not unrelated local surfaces.
Run the closest affected automated tests, the original live reproduction, and
only credible boundary cases. A Rust or shared-crate change does not trigger
all lanes unless every lane consumes the changed behavior. Create only the
minimum non-sensitive fixture required for the selected checks. If a required
deployment, account, fixture, or client is unavailable, mark that check
BLOCKED; do not substitute unrelated lanes. Stop when the mapped risks are
covered.
Comprehensive Interface QA
Run every fixture, baseline, live lane, lifecycle case, privacy check, and cross-surface comparison below only when the user explicitly requests full or comprehensive programmatic-interface QA.
Safety and evidence
- Use a dedicated QA account and non-sensitive fixture meetings.
- Use a Pro or trialing account for the hosted lane and a separate free or expired test account for entitlement checks.
- Never put API keys, JWTs, webhook secrets, database credentials, or personal meeting content in commands, screenshots, reports, or repository files. Load secrets into environment variables without echoing them.
- Store transient response bodies in a directory created with
mktemp -d. Remove that directory and revoke all generated keys after the run. - Record the exact candidate commit, app version, API deployment, Supabase migration version, operating system, and client versions.
- In a GitButler workspace, identify the selected branch tip with
but statusand inspect it withbut show <branch>; do not use the synthetic workspaceHEADas the candidate identity. - Mark a lane
BLOCKED, notPASS, when its required deployment, account, fixture, or client is unavailable.
Comprehensive QA Fixture
Create two completed QA meetings through the desktop app:
- A standalone meeting whose title contains a unique run marker such as
agent-access-2026-07-28T120000Z. - Two meetings in the same recurring series.
The fixture must include:
- a note and at least one generated summary;
- two participants and one action item;
- enough transcript words to require two pages when requested with a small limit;
- punctuation and non-ASCII text;
- a later edit to the title, note, or summary;
- one meeting that will be deleted during lifecycle testing.
Record the meeting IDs and expected visible values. Do not seed SQLite or Postgres directly for the live happy-path tests.
Comprehensive Automated Baseline
Run from the repository root:
cargo test -p anarlog-cli
cargo test -p tauri-plugin-local-api
cargo test -p api-cloud
cargo test -p api openapi::tests
cargo check -p api-client
supabase test db
pnpm -F desktop exec vitest run \
src/cloud-api/client.test.ts \
src/settings/developers/index.test.tsx
pnpm -F desktop typecheck
pnpm exec dprint check
Require the CLI and MCP contract snapshots, OpenAPI composition tests, authentication tests, desktop account-switch tests, and database policy tests to pass without updating snapshots or generated clients during the run.
If the change touches a shared DTO or generated client, regenerate it using the repository command that owns the artifact, then require a clean second generation. A generated diff after the second run is a failure.
Local CLI
Build the candidate CLI with cargo build -p anarlog-cli, then use the built
binary for every step.
- Run
anarlog --json doctor.- PASS when it exits
0, reportsschema_version: "1", uses commanddoctor, reportsready: true, and resolves the intended QA database.
- PASS when it exits
- Run
meetings listwith JSON output.- Verify default ordering, the unique title query, exact series filtering, limit/offset pagination, and an empty result.
- For the recorded meeting ID, run:
meetings get IDmeetings note ID --kind notemeetings note ID --kind summarymeetings note ID --kind allmeetings transcript IDwith at least two pagesmeetings history IDwith at least two pagesmeetings export ID --format markdownmeetings export ID --format json
- Require every JSON response to have the correct
schema_version,command,data, and pagination fields. Followingnext_offsetmust produce no duplicates or gaps. - Export to a temporary file. A second export without
--forcemust fail without changing the file;--forcemust replace it. - A missing meeting ID and an unreadable or incompatible database must return machine-readable errors, a nonzero exit, and no panic or partial export.
Webhooks
Launch the exact desktop candidate and add a temporary receiver under Settings → Developers → Webhooks.
- Adding an endpoint must return a
whsec_secret exactly once, and a non-HTTP URL must be rejected. - Trigger a test delivery, a meeting completion, and a note enhancement.
- Verify the event name, delivery ID, timestamp, body, and HMAC-SHA256 signature over the exact raw body.
- Verify retry behavior with one deliberate transient failure.
- Confirm the last-delivery status shown in settings matches the receiver.
- Delete the webhook and prove that no later delivery arrives.
- Quit the desktop app, complete no further work, and confirm no deliveries are queued or replayed on the next launch.
Local stdio MCP
Start anarlog mcp through a real MCP client over stdio. Do not validate this
lane by invoking server handlers directly.
- Initialize the protocol and run
tools/list. - Require exactly these tools:
list_meetingsget_meetingget_meeting_transcriptget_recurring_meeting_history
- Call every tool against the fixture. Verify query and series filters, two-page transcript/history traversal, missing IDs, and invalid arguments.
- Run
resources/list,resources/templates/list, andresources/readfor a meeting, transcript page, and recurring series. - Require read-only, non-destructive, and idempotent tool annotations.
- Compare the returned values with the CLI lane.
- Capture stdout and stderr separately. Stdout must contain only MCP protocol frames; diagnostics belong on stderr. Client shutdown must terminate the server without leaving a process behind.
Hosted Cloud API
Use the deployed candidate API and Supabase migration with the Pro QA account. Begin with Cloud API & Connectors disabled.
- Before opt-in, confirm there are no server-readable snapshot rows for the
account and a previously valid key returns
403 cloud_api_not_enabled. - Enable the feature in the desktop UI and wait for backfill to finish.
- Create a short-lived cloud key and exercise the same read endpoints, filters, pagination, error cases, and exports as the CLI lane.
- Require:
- no key, malformed key, and revoked key →
401; - free or expired account →
403 subscription_required; - disabled account →
403 cloud_api_not_enabled; - invalid input →
400 invalid_request; - missing meeting →
404 not_found; - request burst above the documented quota →
429withretry-after.
- no key, malformed key, and revoked key →
- Edit the fixture locally and confirm the hosted result changes. Delete the
lifecycle fixture and confirm the hosted endpoint returns
404. - Sign the desktop into another account before a queued upload or retry can complete. No snapshot from the first account may appear in the second.
- Disable the feature.
- PASS when all server-readable snapshots are purged, the cloud key returns
cloud_api_not_enabled, local data remains, and normal encrypted sync data is unchanged.
- PASS when all server-readable snapshots are purged, the cloud key returns
- Re-enable and confirm a fresh backfill restores only currently existing meetings. Revoke the QA key when finished.
Remote MCP
Connect a real Streamable HTTP MCP client to the deployed /mcp endpoint.
- Initialize a session, list tools, and require the same four-tool contract as local MCP. Hosts that only support stateless JSON discovery must still list the four tools without establishing a session first.
- Call every tool and traverse at least two transcript/history pages.
- Compare its structured results with the hosted REST responses.
- Repeat initialization or a tool call with no credential, a malformed key, a
revoked key, an expired account, and after opt-out. Unauthenticated MCP
requests must return
WWW-Authenticatepointing athttps://api.anarlog.so/.well-known/oauth-protected-resource/mcp. - When OAuth is affected, complete MCP OAuth 2.1 discovery and consent from at
least one documented host (Claude Code, Cursor, ChatGPT/Codex, or Copilot).
Confirm the consent screen is Anarlog's
/oauth/consentroute, the issued token is bound tohttps://api.anarlog.so/mcp, and a tool call then reads the marked meeting. Repeat with a staticanl_key for hosts that cannot complete OAuth. - Connect at least one supported agent client using the documented setup and ask it to identify the marked meeting, summarize it, and cite a transcript detail. Verify the answer against the fixture.
- Close the client and confirm the server releases the session cleanly.
Cross-surface parity
For the marked meeting, compare CLI, local MCP, hosted REST, and remote MCP:
| Field | Required parity |
|---|---|
| Meeting | ID, title, kind, status, timestamps, timezone, language, series |
| Documents | canonical note, summary titles and markdown |
| People | participants and organizations |
| Actions | text, assignee, completion state |
| Transcript | text, word order, timestamps, speakers, page boundaries |
| History | IDs, newest-first ordering, pagination |
| Errors | stable code, appropriate protocol status, no secret leakage |
Local and hosted values must match after backfill settles. Hosted payloads must not contain local paths, audio paths, file paths, control characters, secrets, or fields outside the disclosed server-readable copy.
Reporting
For targeted mode, report the candidate branch/commit, base, selected lane and
risk, PASS/FAIL/BLOCKED, and a one-line evidence note. List unrelated
lanes once as NOT APPLICABLE, and say explicitly that the result is not
comprehensive interface QA.
For comprehensive mode, produce one table with rows for:
- automated baseline;
- CLI;
- webhooks;
- local stdio MCP;
- hosted REST;
- remote MCP;
- lifecycle and account isolation;
- privacy purge;
- cross-surface parity.
Use PASS, FAIL, or BLOCKED with a one-line evidence note. Include the
candidate and deployment identifiers, fixture marker, clients tested, and
redacted response artifact locations. List every skipped negative case.
Any required FAIL or BLOCKED result means comprehensive QA did not pass.
Report that outcome without inferring release approval or blocking.
Frequently asked questions
What to verify before installation and use
What does the qa-cli-mcp-api source document cover?
Default to the smallest set of programmatic-interface lanes that can prove the change. Automated tests are necessary but do not replace a live smoke test when the changed boundary is only exercised by a real client or deployment.
How do I install qa-cli-mcp-api?
The source record exposes this install command: npx skills add https://github.com/fastrepl/anarlog --skill ".agents/skills/qa-cli-mcp-api". Inspect the command and pinned source before running it.
Which permission-related actions were detected?
Static rules flagged write-files, exec-script in the source; the page lists the matching lines and excerpts.
Alternatives
Compare before choosing
coreyhaines31/marketingskills
ab-testing
When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "should I test this," "which version is better," "test two versions," "statistical significance," "how long should I run this test," "growth experiments," "experiment velocity," "experiment backlog," "ICE score," "experimentation program
garrytan/gbrain
bulk-ingestion
End-to-end discipline for turning any large data source (audio libraries, email takeouts, document corpora, chat exports, API dumps) into brain pages at scale. The lifecycle spine: SCHEMA → ACCESS → TRIAL → EVALUATE → IMPROVE → CODIFY → TEST → SKILLIFY → BULK → MONITOR. State is tracked in a durable JSON manifest (see MANIFEST-PATTERN.md) so any crash, session boundary, or subagent fan-out resumes from ground truth instead of memory.
alirezarezvani/claude-skills
app-store-optimization
App Store Optimization (ASO) toolkit for researching keywords, analyzing competitor rankings, generating metadata suggestions, and improving app visibility on Apple App Store and Google Play Store. Use when the user asks about ASO, app store rankings, app metadata, app titles and descriptions, app store listings, app visibility, or mobile app marketing on iOS or Android. Supports keyword research and scoring, competitor keyword analysis, metadata optimization, A/B test planning, launch checklist
dotnet/skills
migrate-vstest-to-mtp
Migrates .NET test projects from VSTest to Microsoft.Testing.Platform (MTP). Use when user asks to "migrate to MTP", "switch from VSTest", "enable Microsoft.Testing.Platform", "use MTP runner", set OutputType=Exe only for test projects in Directory.Build.props, or mentions EnableMSTestRunner, EnableNUnitRunner, or UseMicrosoftTestingPlatformRunner. USE FOR: MTP behavioral differences vs VSTest (exit code 8, zero tests discovered, --ignore-exit-code, TESTINGPLATFORM_EXITCODE_IGNORE); centralizing