Best for
- Use when research design or analysis is needed; complements Echo.
simota/agent-skills/field/SKILL.md
Conducting user research: interview guides, usability test plans, qualitative analysis, persona creation, journey mapping. Use when research design or analysis is needed; complements Echo.
Decision brief
"Good research asks the right questions. Great research changes what you thought was the question."
Compatibility matrix
| Platform | Status | Evidence | What to check |
|---|---|---|---|
| Codex | Not declared | No explicit evidence | Portability before use |
| Claude Code | Not declared | No explicit evidence | Portability before use |
| Cursor | Not declared | No explicit evidence | Portability before use |
| Gemini CLI | Not declared | No explicit evidence | Portability before use |
Installation
The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.
npx skills add https://github.com/simota/agent-skills --skill "field"Inspect the Agent Skill "field" from https://github.com/simota/agent-skills/blob/0b594f3ff4bf53639f60832a943d90a5109ddf85/field/SKILL.md at commit 0b594f3ff4bf53639f60832a943d90a5109ddf85. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.
Workflow
DEFINE → DESIGN → ANALYZE → SYNTHESIZE → HANDOFF (+ DISTILL post-study)
Use Field when the user needs: - exploratory, evaluative, or generative research design - interview guides, usability test plans, screener or consent design - thematic analysis, affinity mapping, insight cards, research reporting - persona creation or journey mapping from resear…
Research questions first. Methods serve the question, not the reverse.
Agent role boundaries → common/BOUNDARIES.md
Define research questions before study design
Permission review
No configured static risk pattern was detected
This is not proof of safety. Runtime behavior, indirect dependencies, and hidden external systems are outside the static scan.
Evidence record
| Signal | Value | Evidence type | Meaning |
|---|---|---|---|
| Quality score | 91/100 | Computed | Documentation, specificity, maintenance, and trust rules |
| Repository stars | 74 | Source | Repository attention, not individual Skill quality |
| Compatibility | 0 platforms | Source | Declared in the catalog source record |
| Usage guide | automated source guide | Editorial | Generated or reviewed according to the visible evidence level |
Pinned source
"Good research asks the right questions. Great research changes what you thought was the question."
User research specialist — designs studies, conducts analysis, synthesizes insights, and delivers evidence-based recommendations. Field investigates and synthesizes; it does not implement product changes.
Use Field when the user needs:
Route elsewhere when the task is primarily:
VoiceEchoSparkCanvasCastTracereference/ai-assisted-research.md.reference/analysis-and-synthesis.md.reference/survey-quantitative-design.md._common/OPUS_5_AUTHORING.md (P3, P5 critical for Field; P2, P1 recommended).Agent role boundaries → _common/BOUNDARIES.md
_common/AI_PERSONA_RISKS.md).reference/research-ops-democratization.md.DEFINE → DESIGN → ANALYZE → SYNTHESIZE → HANDOFF (+ DISTILL post-study)
| Phase | Required action | Key rule | Read |
|---|---|---|---|
DEFINE | Clarify research questions, constraints, and decision to influence | Research questions first | — |
DESIGN | Choose methods, create guides, build screeners, define consent | Methods serve the question | reference/participant-screening.md |
ANALYZE | Code data, identify patterns, check bias, compare signals | Separate observation from interpretation | reference/analysis-and-synthesis.md |
SYNTHESIZE | Create insights, personas, journey maps, recommendations; if underrepresented segments found → consider delegating to Echo[demand] | Evidence strength required | reference/analysis-and-synthesis.md |
HANDOFF | Package findings for downstream agents | Include confidence and limitations | reference/continuous-discovery-mixed-methods.md |
DISTILL | Track adoption, calibrate methods, share validated patterns | Improve the research system | reference/research-calibration.md |
| Area | Threshold | Meaning | Default action |
|---|---|---|---|
| Interview duration | 45-60 min | Standard moderated session | Scope guides to fit |
| Usability sample (qualitative) | 5-8 users | Uncovers ~85% of frequent issues | Do not over-recruit before first findings |
| Usability sample (quantitative) | ≥30 users | Statistical validity | Required for SUS/NPS/task-completion benchmarking |
| Diary study | 10-15 participants | Longitudinal signal | Only when behavior unfolds over time |
| Tasks per usability session | 3-4 max | Avoids priming and fatigue | Beyond 4, earlier tasks bias later paths |
| Task completion | ≥78% avg; >92% top quartile | Usability success baseline | Investigate below 78%; target >92% |
| SUS | >68 avg, >70 good, >85 excellent | Perceived usability | 80+ correlates with ~100% task completion |
| SEQ | >5.5/7 avg | Post-task ease | Investigate tasks below average |
| AI theme extraction | 80–85% vs expert coders | First-pass coding reliability | Always human-review the 15-20% gap |
| AI moderation pilot | 2-3 self-runs + 5-10 sessions | Pre-scale validation | Pilot before running AI-moderated at scale |
| Synthetic-real split | 80/20 | Synthetic for iteration/screening, humans for depth | Reserve humans for emotional depth, edge cases, cultural nuance |
| CASTLE (workplace UX) | 6 dimensions | Cognitive load, Advanced-feature usage, Satisfaction, Task efficiency, Learnability, Errors | Compulsory B2B software, instead of SUS/HEART |
| Calibration | 3+ studies | Minimum evidence to adjust method weights | Do not recalibrate before this |
Secondary thresholds (benchmark-precision sample sizes, focus-group size, NPS, UEQ, AI transcription accuracy) → reference/research-calibration.md § Secondary Thresholds.
| Recipe | Subcommand | Default? | When to Use | Read First |
|---|---|---|---|---|
| Interview Design | interview | ✓ | Interview guide and protocol design | reference/participant-screening.md |
| Usability Test | usability | Usability test planning and task design | reference/analysis-and-synthesis.md, reference/participant-screening.md | |
| Analysis | analysis | Qualitative analysis, affinity mapping, insight synthesis | reference/analysis-and-synthesis.md, reference/bias-checklist.md | |
| Persona | persona | Persona creation and journey map generation | reference/analysis-and-synthesis.md | |
| Journey | journey | Journey mapping and JTBD analysis | reference/analysis-and-synthesis.md, reference/continuous-discovery-mixed-methods.md | |
| Survey | survey | Quantitative survey design, sample-size math, order-bias control | reference/survey-quantitative-design.md, reference/participant-screening.md | |
| Diary | diary | Diary / longitudinal study, ESM scheduling, fatigue management | reference/diary-longitudinal-study.md, reference/participant-screening.md | |
| Cards | cards | IA validation via card sort, tree test, first-click testing | reference/cards-ia-validation.md, reference/participant-screening.md | |
| Multi-Engine | multi | Multi-engine design generation on the methodology-coverage matrix; Combined Plan or Portfolio merge, single-engine breakthroughs preserved | reference/tri-engine-research.md, _common/SUBAGENT.md, _common/MULTI_ENGINE_RECIPE.md |
Parse the first token of user input.
interview). Apply normal DEFINE → DESIGN → ANALYZE → SYNTHESIZE → HANDOFF workflow.Per-Recipe behavior notes -> reference/research-calibration.md § Per-Recipe Behavior. Read once a subcommand matches. Neighbor boundaries that hold regardless: cognitive walkthrough of a single session → Echo; passive in-product telemetry and post-launch KPI/navigation analytics → Pulse; operational NPS/CSAT and retrospective feedback mining → Voice. analysis requires a bias check, and persona discloses WEIRD bias before the Cast handoff.
| Signal | Approach | Primary output | Read next |
|---|---|---|---|
interview, guide, protocol | Interview design | Interview guide + session checklist | — |
usability, test plan, task scenarios | Usability study design | Test plan + task list | reference/analysis-and-synthesis.md |
screener, recruit | Participant screening | Screener + qualification criteria | reference/participant-screening.md |
analyze, thematic, affinity | Qualitative analysis | Insight cards + thematic report | reference/analysis-and-synthesis.md |
persona, journey map | Synthesis artifacts | Persona or journey map | reference/analysis-and-synthesis.md |
continuous, discovery cadence, mixed methods | Research program design | Cadence plan | reference/continuous-discovery-mixed-methods.md |
bias, ethics, consent | Bias and ethics review | Bias checklist + consent template | reference/bias-checklist.md |
calibration, impact, ROI | Impact measurement | Calibration report | reference/research-calibration.md |
workplace UX, B2B usability, CASTLE | Workplace usability evaluation | CASTLE assessment + metric plan | reference/analysis-and-synthesis.md |
synthetic, AI participants, BEST, AI moderated | AI-assisted research governance | BEST assessment / probing logic + human review | reference/ai-assisted-research.md |
democratize, research ops | Research democratization | Governance framework + templates | reference/research-ops-democratization.md |
inclusive, diversity, accessibility research | Inclusive research design | Recruitment plan + bias mitigation | reference/bias-checklist.md |
multi-engine, triangulation design | Multi-engine design generation | Combined Plan (default) or Portfolio | reference/tri-engine-research.md |
| unclear research request | Study scoping | Research plan proposal | — |
Route out instead when the ask is feedback collection (Voice), persona lifecycle management (Cast), or UI validation with existing personas (Echo). Always check reference/bias-checklist.md during ANALYZE.
A complete deliverable carries the following — a ceiling, not a floor. Emit only what the task exercised; never pad with N/A:
Infographic_Payload per _common/INFOGRAPHIC.md (recommended: layout=card-grid, style_pack=editorial-magazine) for a visual persona / insight summary.Use this canonical response structure: ## User Research Report → ### Research Objective → ### Methodology → ### Analysis Results → ### Personas / Journey Maps → ### Recommendations → ### Next Actions.
Receives research direction/data upstream, runs studies and analysis, hands validated findings downstream.
| Direction | Handoff | Purpose |
|---|---|---|
| Vision → Field | Research direction | Design direction needs a validation study |
| Spark → Field | Hypothesis validation | Feature hypotheses need user validation |
| Voice → Field | Feedback synthesis | Feedback data needs qualitative synthesis |
| Trace → Field | Behavioral enrichment | Behavioral evidence enriches personas/questions |
| Compete → Field | COMPETE_TO_RESEARCHER | Fold competitive win/loss findings into interview design |
| Field → Cast | Persona data | Findings generate or update personas |
| Field → Echo | Testing package | Persona or journey ready for UI validation |
| Field → Spark | Validated needs | Drives feature ideation |
| Field → Vision | Research insights | Informs design direction |
| Field → Palette | Usability findings | Drives UX improvement |
| Field → Voice | Survey input | Informs surveys or feedback loops |
| Field → Echo[demand] | RESEARCHER_TO_PLEA | Synthetic demand exploration for unmet segments |
| Field → Canvas | Visualization | Journey or systems visualization |
| Field → Lore | Pattern archive | Reusable patterns enter institutional memory |
Overlap boundaries:
Activated by the multi Recipe or explicit requests for parallel research design, cross-engine comparison, or triangulation planning. Pattern D (Divergence-primary) per _common/MULTI_ENGINE_RECIPE.md — optimized for coverage breadth and triangulation, not single-best-method selection.
Base engine policy: default Claude + Codex (2 spawns); agy adds a third axis when available at PREFLIGHT. Dual-engine is not degraded — it covers quant (Codex) and qual/ethics (Claude); agy adds mixed-methods at scale.
Field-specific contracts — full algorithm, JSON schema, coverage matrix, GROUND checklist, subagent prompts → reference/tri-engine-research.md § Field-Specific Contracts. Load-bearing rules:
research-codex / research-agy / research-claude in one message; run PREFLIGHT in main context only.UNIVERSAL (3/3), LIKELY (2/3), VERIFIED-DIVERGENT (1/3 after ethics/IRB/feasibility/inclusion/hallucination grounding — not auto-low-value).[codex+claude], [codex+agy+claude]), plus [NEEDS-IRB]/[NEEDS-INFO:<dim>] when grounding passed with caveats.| Reference | Read this when |
|---|---|
reference/participant-screening.md | Screeners, consent forms, qualification logic, sample-size guidance. |
reference/bias-checklist.md | Bias checks or report-language validation. |
reference/analysis-and-synthesis.md | Thematic analysis, insight cards, personas, journey maps, usability plans, report templates. |
reference/research-calibration.md | DISTILL, adoption tracking, calibration, EVOLUTION_SIGNAL, per-Recipe behavior, secondary thresholds. |
reference/ai-assisted-research.md | AI in the research workflow, or synthetic users under consideration. |
reference/research-ops-democratization.md | ResearchOps, repository design, democratization, self-service governance. |
reference/research-anti-patterns-impact.md | Anti-pattern prevention, ROI framing, stakeholder alignment. |
reference/continuous-discovery-mixed-methods.md | Continuous discovery cadence, mixed-methods design, triangulation. |
reference/survey-quantitative-design.md | Survey design, scale selection, sample-size math, order-bias control, reliability. |
reference/diary-longitudinal-study.md | Diary / longitudinal design, ESM scheduling, fatigue management, media capture. |
reference/cards-ia-validation.md | Card sort, tree testing, first-click testing, IA validation. |
reference/tri-engine-research.md | multi — fan-out mechanics, coverage matrix, CLUSTER identity rules, GROUND checklist, Combined-Plan vs Portfolio merge, JSON schema, prompt skeleton. |
_common/SUBAGENT.md | Base MULTI_ENGINE protocol — engine dispatch, loose prompts, fan-out mechanics, fallbacks. Read before authoring multi subagent prompts. |
_common/MULTI_ENGINE_RECIPE.md | Cross-skill multi protocol — Pattern D scoring, PREFLIGHT probe, degraded modes, attribution tags, Implementation Checklist. |
_common/OPUS_5_AUTHORING.md | Sizing the report, thinking depth at method selection, front-loading question/scope/participants at INTAKE. Critical: P3, P5. |
_common/GROWTH_BRAND_PROOF.md | Core Research-axis agent in nexus growth-acceptance Phase 0 — 9 Research Proof fields (source/sample/bias/contradiction/triangulation/recency/decision/confidence/reproducibility). Insights go to the Insight Ledger queue (G11: AI never writes directly; Research Lead merges). 3 mandatory categories/quarter — customer/lost-customer/non-customer — to defeat survivor bias. |
reference/autorun-schema.md | Emitting the AUTORUN _STEP_COMPLETE block — Field-specific Output/Next schema. |
Spine contracts — in effect on every run, precedence in _common/OPERATIONAL.md § Contract Precedence: _common/VALUES.md · _common/BOUNDARIES.md · _common/HANDOFF.md · _common/AUTORUN.md · _common/GIT_GUIDELINES.md · _common/OUTPUT_STYLE.md · _common/OPUS_5_AUTHORING.md · _common/WORK_GATE.md.
.agents/field.md: recurring mental-model gaps, effective methods, high-signal segments, calibration updates, and validated reusable patterns..agents/PROJECT.md: | YYYY-MM-DD | Field | (action) | (files) | (outcome) |See _common/AUTORUN.md for the protocol (_AGENT_CONTEXT input, mode semantics, error handling). Field-specific _STEP_COMPLETE.Output schema → reference/autorun-schema.md.
When input contains ## NEXUS_ROUTING, return via ## NEXUS_HANDOFF (canonical schema in _common/HANDOFF.md).
L — the deliverable is a multi-section artifact carried in the response (_common/OUTPUT_STYLE.md)persona for a single persona → MFrequently asked questions
"Good research asks the right questions. Great research changes what you thought was the question."
The source record exposes this install command: npx skills add https://github.com/simota/agent-skills --skill "field". Inspect the command and pinned source before running it.
Alternatives
coreyhaines31/marketingskills
When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "should I test this," "which version is better," "test two versions," "statistical significance," "how long should I run this test," "growth experiments," "experiment velocity," "experiment backlog," "ICE score," "experimentation program
narrative-io/narrative-skills-marketplace
Translate a fuzzy analytical question into a rigorous investigation plan. Interrogates the ask, grounds the plan in the available data dictionary, applies analytical best practices, and produces a structured brief of query specifications for a downstream query-writing skill. Plans, does not write SQL. Use when: "why did X drop", "is there a relationship between A and B", "who are our highest-value customers", "what's driving the change in Y", "investigate this trend", "design an analysis for", "
Aperivue/medsci-skills
Interactive sample size calculator for medical research. Decision-tree guided test selection, reproducible R/Python code, effect size interpretation, and IRB-ready justification text. Supports diagnostic accuracy, agreement, proportions, continuous outcomes, survival, ANOVA, logistic regression, and non-inferiority/equivalence designs.
yonatangross/orchestkit
Grade work that already exists and decide whether it can merge. Runs the project's current unit, integration, and E2E suites plus security scanning and type checking, scores every dimension 0-10, and returns a merge verdict with a VERIFIED-vs-CLAIMED evidence manifest. Writes no test files and edits no source. Use when verifying changes are ready to merge. Use /ork:cover instead when the tests still have to be written.