Source profileQuality 95/100

vasilyu1983/AI-Agents-public/frameworks/shared-skills/skills/software-ux-research/SKILL.md

software-ux-research

Guides user research methods and research ops. Use when running interviews, usability tests, surveys, or A/B tests to de-risk product decisions.

Source repository stars
80
Declared platforms
2
Static risk flags
0
Last source update
2026-08-21
Source checked
2026-08-25

Decision brief

What it does: where it fits

Use this skill to reduce product and design risk with evidence. It owns research method choice, study design, findings synthesis, and research operations. It does not own UI implementation.

Best for

  • what user problem matters and for whom
  • whether a concept, flow, or prototype is understandable and usable
  • which research method is appropriate

Not for

  • Running surveys to answer why questions that need observed behavior or interviews.
  • Treating five usability sessions as statistically representative rather than as directional evidence about failure patterns.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexDeclaredSource recordInstall path and trigger
Claude CodeDeclaredSource recordInstall path and trigger
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/vasilyu1983/AI-Agents-public --skill "frameworks/shared-skills/skills/software-ux-research"
Safe inspection promptEditorial

Inspect the Agent Skill "software-ux-research" from https://github.com/vasilyu1983/AI-Agents-public/blob/53f6cb73ea53a2646e3e7d4665062ad66f3683ac/frameworks/shared-skills/skills/software-ux-research/SKILL.md at commit 53f6cb73ea53a2646e3e7d4665062ad66f3683ac. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    Workflow

    1. Define the decision and deadline. 2. Inventory existing evidence. 3. Choose the method and explain why weaker alternatives were rejected. 4. Produce one decision-ready output. 5. Tag confidence and data-handling constraints.

    Define the decision and deadline.Inventory existing evidence.Choose the method and explain why weaker alternatives were rejected.
  2. 02

    Stage Guidance

    Review the “Stage Guidance” section in the pinned source before continuing.

    Review and apply the “Stage Guidance” source section.
  3. 03

    Verification Checklist

    Before delivering any research output:

    [ ] Decision the study was designed to unblock is named explicitly[ ] Method justified: weaker alternatives were considered and rejected with reasons[ ] Participants match the target segment — not convenience, panel-only, or CS rolodex
  4. 04

    Quick Reference

    Review the “Quick Reference” section in the pinned source before continuing.

    Review and apply the “Quick Reference” source section.
  5. 05

    When to Use This Skill

    Use this skill when the main question is:

    what user problem matters and for whomwhether a concept, flow, or prototype is understandable and usablewhich research method is appropriate

Permission review

Static risk signals and limitations

No configured static risk pattern was detected

This is not proof of safety. Runtime behavior, indirect dependencies, and hidden external systems are outside the static scan.

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score95/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars80SourceRepository attention, not individual Skill quality
Compatibility2 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
vasilyu1983/AI-Agents-public
Skill path
frameworks/shared-skills/skills/software-ux-research/SKILL.md
Commit
53f6cb73ea53a2646e3e7d4665062ad66f3683ac
License
MIT
Collected
2026-08-25
Default branch
main
View the original SKILL.md

Software UX Research

Use this skill to reduce product and design risk with evidence. It owns research method choice, study design, findings synthesis, and research operations. It does not own UI implementation.

Quick Reference

NeedDefaultOutput
discovery and JTBDsemi-structured interviews with 5-8 participantsopportunity brief
usability evaluationmoderated usability test with 5-7 participantsfindings report with severity
quantification after qual insightsurvey or analytics reviewsegment or pattern readout
causal change validationcontrolled experiment or staged rolloutexperiment brief
research ops and repository designlightweight intake, taxonomy, and consent modelresearch-ops recommendation
accessibility or low-digital-literacy researchmoderated sessions with adapted materialsrisk and inclusion report

When to Use This Skill

Use this skill when the main question is:

  • what user problem matters and for whom
  • whether a concept, flow, or prototype is understandable and usable
  • which research method is appropriate
  • how to design a study and synthesize findings
  • how to run research ops, repository, and consent workflows

Route elsewhere when the main task is:

NeedUse Instead
UI design and interaction patterns../software-ui-ux-design/SKILL.md
code-level accessibility remediation../software-accessibility/SKILL.md
accessibility testing automation and CI gates../qa-testing-accessibility/SKILL.md
analytics instrumentation implementationmarketing-product-analytics and ../qa-observability/SKILL.md

Defaults

  • start from the decision to unblock
  • choose the smallest method mix that can answer the question
  • use qual for motives and friction, quant for scale and segmentation
  • treat synthetic participants as hypothesis generation only
  • require confidence level and evidence trail in every output
  • current standards and regulatory claims must be verified before final advice

Quality Lens

Consumer-grade research looks past task completion to whether the experience is efficient, considerate, and worth coming back to. Evaluate every research question and finding through four layers — methods that only cover the top layer will miss why people churn or never habit-form. See references/consumer-experience-quality.md for methods, instruments, and recipes.

LayerQuestionPrimary Methods
Taskcan users complete the job?usability testing, task success, SEQ
Frictionwhat slows, frustrates, or shames them?friction logging, diary studies, session replay paired with interview
Emotionhow does it feel — proud, calm, tense, ignored?PrEmo, AttrakDiff, Microsoft Desirability Toolkit, micro-interviews
Meaningdoes it earn a place in their life? does it cause harm?JTBD Switch interviews, Continuous Discovery (OST), longitudinal/diary, retention cohorts

A finding that names task pass-rate but not friction or emotion is incomplete. Discovery work without Meaning-layer questions tends to ship features people use once.

Workflow

  1. Define the decision and deadline.
  2. Inventory existing evidence.
  3. Choose the method and explain why weaker alternatives were rejected.
  4. Produce one decision-ready output.
  5. Tag confidence and data-handling constraints.

ASCII Flow

UX research task
  -> Define decision, audience, deadline, and risk
  -> Inventory existing evidence and data constraints
  -> Choose smallest method mix that answers the decision
  -> Run or design study with consent and evidence trail
  -> Synthesize findings with confidence level
  -> Deliver options, tradeoffs, and next decision

Output Types

Default outputs:

  • research plan
  • study protocol
  • findings report
  • decision brief

Every substantial output should include:

  • method justification
  • confidence level
  • evidence trail
  • consent and data-handling note
  • recommendation framed as options and tradeoffs

Method Chooser

NeedPrimary Methods
motives, needs, switching triggersinterviews, contextual inquiry, diary studies
usability and learnabilitymoderated usability testing, cognitive walkthroughs, heuristic review
scale, segments, or behavioral patternsanalytics review, surveys, feedback mining
causal effectcontrolled experiment, staged rollout, preference test

Use moderated testing by default when failure paths, assistive technology, or complex workflows matter.

Stage Guidance

StageTypical Research Focus
discoveryproblem selection, JTBD, forces of progress
concept or MVPconcept comprehension, prototype usability, onboarding risk
launchblocker identification, accessibility, and readiness
growthretention, friction, and segment behavior
maturityoptimization, simplification, or feature retirement

Verification Checklist

Before delivering any research output:

  • Decision the study was designed to unblock is named explicitly
  • Method justified: weaker alternatives were considered and rejected with reasons
  • Participants match the target segment — not convenience, panel-only, or CS rolodex
  • Sample size appropriate to method: ≥5 for usability, ≥8 for discovery interviews, power-calculated for experiments
  • Confidence level and evidence trail stated in the output
  • Synthetic participants labeled as hypothesis generation only — not cited as evidence
  • AI-assisted analysis audited (≥10-15% of AI tags verified against human coding)
  • Consent obtained; recordings, transcripts, and participant identity stored separately
  • EU/UK participant data: DPA in place before sending to AI-processing vendor; EU AI Act high-risk (Annex III) deployer obligations postponed from 2026-08-02 to 2027-12-02 under the Digital Omnibus — the European Parliament (16 June 2026) and Council (29 June 2026) have both given final approval; the act enters into force shortly after Official Journal publication (verify the exact effective date before citing it as settled law)
  • Disconfirming evidence documented, not only confirming clips
  • Agentic products: study ran multi-turn, exercised at least one interruption, and included seeded incorrect outputs if trust was measured

Research Ops Rules

  • capture the decision, audience, segment, and evidence links in intake
  • use one taxonomy across studies and atomic insights
  • separate participant identity from notes and recordings
  • redact broad-share artifacts
  • let non-researchers run only templated studies with review guardrails

AI and Accessibility Notes

For AI-powered product research (the thing being studied is AI-driven):

  • test trust calibration, failure recovery, explainability, tool-use disclosure, and approval gating
  • separate wrong output from unclear output and non-recoverable failure
  • run multi-turn sessions for agentic products — single-turn studies miss most of the failure surface
  • test steering explicitly: users change their mind mid-task, and addition/revision/retraction fail differently
  • measure trust calibration against seeded incorrect outputs; an all-correct study cannot distinguish good judgment from blind acceptance
  • see references/ai-in-research.md for the full dimension list and method mapping, and references/agentic-evaluation-methods.md for the multi-turn protocols

For AI in the research workflow (synthesis tools, AI moderators, synthetic users):

  • treat synthetic users as hypothesis generation only (NN/g position), never as evidence
  • start analysis from human-coded seed sample, then let AI extend; audit at least 10–15% of AI tags
  • AI moderators are appropriate only when the protocol is structured enough for a junior human to follow
  • inventory every AI tool that processes participant data for EU AI Act enforcement (high-risk deployer obligations postponed to 2 December 2027 under the Digital Omnibus, now approved by Parliament and Council as of June 2026 — verify current in-force date)

For accessibility-sensitive research:

  • recruit assistive-technology users when accessibility is in scope
  • distinguish accessibility usability findings from formal conformance findings

Known Traps

  • Starting with a preferred method before naming the actual decision the study needs to unblock.
  • Recruiting convenience participants whose context, literacy, or workflow is too far from the target segment.
  • Treating generated summaries, AI note clustering, or synthetic participants as evidence instead of support material.
  • Mixing discovery, usability, and causal-validation questions into one study and getting ambiguous output from all three.
  • Reporting severity or confidence without tying it to sample quality, task coverage, and evidence strength.
  • Storing recordings, transcripts, and participant identity with weaker controls than the sensitivity of the study requires.
  • Sending EU/UK participant recordings to a non-EU AI vendor (Dovetail, Marvin, Looppanel, or any foundation-model-backed service) without a current DPA and explicit AI processing disclosure in consent — Chapter V GDPR transfer rules apply now, and EU AI Act high-risk deployer obligations follow (postponed from 2 August 2026 to 2 December 2027 under the Digital Omnibus, approved by Parliament and Council in June 2026 — verify the current in-force date before relying on it).
  • Recruiting only from professional research panels (Prolific, UserTesting panel) for behavior studies, then generalising to product users — panel respondents are experienced participants whose behavior systematically diverges from first-time real users.

Common Anti-Patterns

  • Running surveys to answer why questions that need observed behavior or interviews.
  • Treating five usability sessions as statistically representative rather than as directional evidence about failure patterns.
  • Converting every insight into a roadmap request instead of separating evidence, interpretation, and action options.
  • Using heuristic review as a replacement for user research when task comprehension or domain literacy is the core risk.
  • Repeating studies without a repository, taxonomy, or decision log, so the team relearns the same lesson every quarter.
  • Democratising research as cover for cutting researcher headcount: non-researchers run uncontrolled studies, cherry-pick confirming insights, and quality silently degrades. Templated studies with reviewer guardrails are the supported pattern; "anyone can run any study" is not.
  • Letting an AI moderator handle generative or first-time discovery work — leading prompts produce leading follow-ups at scale.
  • Confirmation bias in moderation: the moderator unconsciously seeks confirming clips and discounts disconfirming ones. Mitigation: code clips before discussing, require double-coder agreement on findings above severity 2, and explicitly document disconfirming evidence in every report.
  • Decision-by-quote / champion-user-as-segment: shipping a feature because one passionate user wanted it. Single-N evidence is hypothesis, not finding.
  • Post-hoc segmentation hunting: slicing experiment results by 20 segments until one is significant. Pre-register segmentation analysis before the experiment reads out, or apply a correction (Bonferroni, FDR) when segments are exploratory.
  • Satisfaction theater: surveys conducted to put a number on a slide rather than to inform a decision. If the survey result would not change anything, do not run it.
  • Power-gaming experimentation: extending experiments until significance appears, hiding losing variants, or changing the primary metric mid-experiment to ship a desired outcome. Each of these invalidates the result.
  • Rating agent transcripts instead of having raters use the agent. Someone who did not have the conversation cannot judge trust, patience, or perceived competence — their scores track fluency instead. Multi-turn evaluation requires first-person experience.
  • Measuring trust in an AI product using only correct outputs. Without seeded errors you can measure acceptance, but you cannot distinguish good calibration from blind acceptance — and over-reliance is the failure that matters.
  • Reporting task completion for agentic tasks without elapsed time and cost. A task that completed after six minutes and four retries is not the same outcome as one that took twenty seconds; completion rate alone hides it.
  • Citing a model benchmark as a UX finding. Benchmarks tell you the capability ceiling, not whether your interface lets users reach it.
  • Recruiting the customer-success rolodex as a research panel: those users are atypically engaged, vocal, and cooperative. Generalizing from them is a top-of-funnel research failure — find disengaged, lapsed, and never-converted users too.

Navigation

References

Assets

Related Skills

Fact-Checking

  • Known bugs, regressions, framework/compiler/runtime footguns, and version-specific crash or workaround guidance must be verified against current primary web sources before being treated as current fact.
  • Verify current standards, legal deadlines, and external research-method claims before final advice.
  • Prefer ISO, W3C, regulator, and primary-method sources over summaries.
  • If live verification is unavailable, mark external claims as unverified.

Learnings Loop

Before applying this skill on a non-trivial task, read learnings.consolidated.md in this directory (and learnings.md if present).

After applying it, if you encountered a pattern worth remembering, a mistake worth preventing, or a domain fact that surprised you, append one dated bullet to learnings.md via agents-skills-feedback-loop/scripts/append_learning.py. Do not modify SKILL.md itself.

Frequently asked questions

What to verify before installation and use

What does the software-ux-research source document cover?

Use this skill to reduce product and design risk with evidence. It owns research method choice, study design, findings synthesis, and research operations. It does not own UI implementation.

How do I install software-ux-research?

The source record exposes this install command: npx skills add https://github.com/vasilyu1983/AI-Agents-public --skill "frameworks/shared-skills/skills/software-ux-research". Inspect the command and pinned source before running it.

Which Agent platforms does the source record declare?

The pinned source record declares support for: codex, claude code.

Alternatives

Compare before choosing

Computed 9420

upex-galaxy/agentic-qa-boilerplate

test-documentation

Analyze, prioritize, and document test cases in TMS (Jira/Xray), or repair an existing Story-ATS-ATP-ATR-TC cascade through a sealed explicit mode. Use for Test/ATP/ATR artifacts, ROI and automation verdicts, maintaining traceability, fix-traceability, or broken TMS links. The repair-traceability mode audits, plans, waits for explicit approval, applies, and verifies without launching the general documentation workflow. Do NOT use for writing test code (test-automation) or running suites (regress

Computed 9320

upex-galaxy/agentic-qa-boilerplate

sprint-testing

Orchestrates in-sprint manual QA per ticket across Stages 1 (Planning), 2 (Execution) and 3 (Reporting). Use for user-story testing, bug retesting, and batch-sprint QA loops. Creates the PBI folder, drives session-start, runs the triage + veto + risk-score decision tree on bugs, produces the ATP + ATR + TC artifacts in the TMS, executes smoke and trifuerza (UI/API/DB) exploration, and files the final QA comment + bug reports. Triggers on: test this ticket, QA this user story, retest this bug, ve

Computed 914,000

nyldn/claude-octopus

flow-discover

Multi-AI research using available external providers (Double Diamond Discover phase)

Computed 9024,921

alirezarezvani/claude-skills

helm-chart-builder

Helm chart development agent skill and plugin for Claude Code, Codex, Gemini CLI, Cursor, OpenClaw — chart scaffolding, values design, template patterns, dependency management, security hardening, and chart testing. Use when: user wants to create or improve Helm charts, design values.yaml files, implement template helpers, audit chart security (RBAC, network policies, pod security), manage subcharts, or run helm lint/test.