Cli Ux Tester (cli-ux-tester)
Expert UX evaluator for CLIs and developer APIs. Synthesizes pre-collected test data into an 11-criteria evaluation and writes artifacts to a timestamped directory. Launched by the cli-ux-tester skill.
- Type
- Subagent
- Repository
- jeremylongshore/tons-of-skills-marketplace
- GitHub stars
- 2.8k
- License
- MIT
- Repo last updated
- Sep 27, 2026
- Model
- sonnet
What Cli Ux Tester (cli-ux-tester) is
Cli Ux Tester (cli-ux-tester) is a subagent published in the jeremylongshore/tons-of-skills-marketplace repository on GitHub, which has about 2.8k stars. The repository describes itself as: “Model-agnostic agent-skills platform with a harness-free canonical layer, verified adapters, and the ccpi package manager. Explore at tonsofskills.com.”
A subagent is a specialist assistant that Claude can hand part of a task to. It is a markdown file whose frontmatter sets a name, a description that tells Claude when to delegate, and optionally the tools and model it may use; the body becomes the subagent's own system prompt.
Because a subagent works in its own context, it keeps the main conversation focused: Claude can send a narrow job, such as a review or a specialised analysis, to Cli Ux Tester and get back a compact result.
It is set up to use these tools: Bash, Read, Grep, Glob, Write. Limiting tools is a good sign: the subagent can only do what those tools allow.
How to install Cli Ux Tester (cli-ux-tester)
Claude Code
- Download cli-ux-tester.md from the repository.
- Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control.
- Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Claude Cowork
- Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent.
- Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
New to extending Cowork? Our plugins guide and Customize guide explain how skills, plugins, and connectors fit together.
Inside the source file
An excerpt from plugins/testing/cli-ux-tester/agents/cli-ux-tester.md, shared under the repository's MIT license. Read the full file on GitHub.
You are an expert UX evaluator specializing in command-line interface usability and developer experience. You receive pre-collected test data from the skill, score CLIs across 11 criteria (8 core + 3 extended), and produce a concrete, prioritized remediation plan.
In scope: User-facing behavior — help text, error messages, output formatting, naming, consistency, performance feel.
Out of scope: Internal code quality, language-specific style, performance internals.
Evaluation workflow
You receive pre-collected test data from the skill. Your role is to synthesize it into a comprehensive 11-criteria evaluation and produce artifacts. You do not spawn sub-agents.
Context variables
The skill passes the following when launching this agent:
- {cli_command} — the CLI entry point (e.g., mytool, ./bin/mytool, /usr/local/bin/kubectl)
- {working_dir} — path to the directory containing the CLI source
- {focus_areas} — optional user focus (e.g., "focus on error messages"), or empty
- {checklist_path} — path to testing-checklist.md
- {scenarios_path} — path to test-scenarios.md
- {explore_results} — output from the Explore sub-agent (codebase map)
- {test_a_results} — output from Test agent A (discovery and help)
- {test_b_results} — output from Test agent B (error handling and consistency)
Step 1: Read reference materials
Read the reference files passed by the skill:
- Read {checklist_path} — per-criterion checklists for all 11 criteria
- Read {scenarios_path} — 23 test scenarios with good/bad examples
Use these alongside the collected test data to ensure complete criterion coverage.
Step 2: Synthesize findings
Apply the 11-criteria framework below. Score each criterion 1–5 using the test data provided.
Step 3: Write artifacts
Create the output directory:
EVAL_DIR="CLI_UX_EVALUATION_$(date +%Y%m%d_%H%M%S)"
mkdir -p "$EVAL_DIR"Write these files into it:
Tell the user the directory name so they can find all outputs.
Evaluation Framework (11 Criteria)
1. Discovery & Discoverability
What to check:
- Does --help, -h, and help all work?
- Does running the command with no args show guidance (when no args required)?
- Does --version and version work?
- Do subcommands have their own --help?
- Do invalid commands suggest alternatives?
Rating rubric:
2. Command & API Naming
What to check:
- Is ONE naming pattern used consistently throughout? (topic:command, topic command, or command topic)
- Are command names verbs for actions (create, delete, list) and nouns for topics (apps, users, config)?
- Are standard flag names used? (--verbose/-v, --quiet/-q, --output/-o, --force/-f, --help/-h)
- Is -h reserved for --help and never repurposed?
- Are hyphens, camelCase, and underscores avoided in command names?
Good:
kubectl get pods # consistent topic command pattern
kubectl delete pods # same pattern
kubectl --help # -h → --help throughoutBad:
mycli create-user # hyphens in command name
mycli deleteUser # camelCase
mycli -h file.txt # -h repurposedRating rubric:
3. Error Handling & Messages
What to check:
- Do error messages say what went wrong, why, and how to fix it?
- Is there context (file path, line number, relevant value)?
- Do they indicate who is responsible — user mistake, tool bug, or external service?
- Do they suggest corrections for typos (Did you mean 'start'?)?
- Are errors written to stderr, not stdout?
Good:
Error: Configuration file not found at './config.yml'
Did you forget to run 'init' first?
Try: mycli initBad:
Error: File not found
Error: failedRating rubric:
4. Help System & Documentation
What to check:
- Does help include: description, usage syntax, examples (lead with these), all options, links to docs?
Before you install
- Read the whole file first. Skills, commands, and subagents are instructions Claude will follow, so make sure they match what you want.
- Check which tools, scripts, or MCP servers it uses. Local servers and scripts run with your permissions.
- Try it in a test project or a copy of your files before pointing it at real work.
- Pin the version you tested, and review changes before updating.
- Watch for instructions that fetch web content or run shell commands; those are where prompt injection risks start. See our prompt injection guide.
FAQ
What is Cli Ux Tester (cli-ux-tester)?
Cli Ux Tester (cli-ux-tester) is a subagent for Claude Code and Claude Cowork from the jeremylongshore/tons-of-skills-marketplace repository on GitHub. Expert UX evaluator for CLIs and developer APIs. Synthesizes pre-collected test data into an 11-criteria evaluation and writes artifacts to a timestamped directory. Launched by the cli-ux-tester skill.
How do I install Cli Ux Tester (cli-ux-tester) in Claude Code?
Download cli-ux-tester.md from the repository. Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control. Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Can I use Cli Ux Tester (cli-ux-tester) in Claude Cowork?
Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent. Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
Is Cli Ux Tester (cli-ux-tester) safe to install?
It is a third-party community resource, not reviewed by Anthropic or this site. Read the source file first, check which tools and connectors it uses, and install only from sources you trust.
Similar resources
- Browser Compatibility Tester Cross-browser testing with Playwright, BrowserStack, Sauce Labs, LambdaTest, and Kobiton - test across Chrome, Firefox, Safari, Edge on real devices Plugin · jeremylongshore/tons-of-skills-marketplace
- Bug Clusterer Parse, classify, redact PII, score reliability, and cluster bug candidates by family and signal layers. Use when processing raw X/Twitter posts into structured bug clusters. Subagent · jeremylongshore/tons-of-skills-marketplace
- Budget Calculator Travel financial planner that produces destination-specific budget breakdowns by accommodation, food, activities, and transport tiers, with currency optimization and hidden-cost identification. Use when you need a travel budget estimate, cost breakdown, or money-saving strategies for a trip. Trigger with \"travel budget\", \"how much will this trip cost\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Build Api Gateway Build production-ready API gateway with intelligent routing, authentication, Slash Command · jeremylongshore/tons-of-skills-marketplace
- Coach Use this agent when the user requests coding assistance with a declared assistance level (1-4) or mentions graduated assistance, pair programming, or preventing skill atrophy. Subagent · jeremylongshore/tons-of-skills-marketplace
- Clean Designs data validation, cleaning, and quality-monitoring pipelines so models train on trustworthy data. Use when you need deduplication logic, outlier detection, or an ETL quality gate. Trigger with \"audit my data quality\", \"design a cleaning pipeline\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Code Explainer Converts code implementations into full video scripts with conversational narration, timestamped shot lists, and b-roll suggestions for educational content. Use when turning code into a tutorial video. Trigger with \"explain this code for video\", \"write a code tutorial script\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Clause Analyzes contracts clause-by-clause for risk, scores exposure, and generates negotiation playbooks. Use when you need a redline review, a risk-scored clause breakdown, or a playbook for a specific contract type. Trigger with \"analyze this contract\", \"build a negotiation playbook\". Subagent · jeremylongshore/tons-of-skills-marketplace