Benchmark Suite
Comprehensive performance benchmarking, regression detection and performance validation
- Type
- Subagent
- Repository
- ruvnet/agentic-flow
- GitHub stars
- 813
- License
- Not declared
- Repo last updated
- Sep 16, 2026
- Source file
- .claude/agents/optimization/benchmark-suite.md
What Benchmark Suite is
Benchmark Suite is a subagent published in the ruvnet/agentic-flow repository on GitHub, which has about 813 stars. The repository describes itself as: “Easily switch between alternative low-cost AI models in Claude Code/Agent SDK. For those comfortable using Claude agents and commands, it lets you take what you've created and deploy fully hosted agents for real business purposes. Use Claude Code to get the agent working, then deploy it in your favorite cloud.”
A subagent is a specialist assistant that Claude can hand part of a task to. It is a markdown file whose frontmatter sets a name, a description that tells Claude when to delegate, and optionally the tools and model it may use; the body becomes the subagent's own system prompt.
Because a subagent works in its own context, it keeps the main conversation focused: Claude can send a narrow job, such as a review or a specialised analysis, to Benchmark Suite and get back a compact result.
How to install Benchmark Suite
Claude Code
- Download benchmark-suite.md from the repository.
- Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control.
- Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Claude Cowork
- Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent.
- Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
New to extending Cowork? Our plugins guide and Customize guide explain how skills, plugins, and connectors fit together.
Inside the source file
The repository does not declare an open-source license, so we list only the outline of the file here. Read the full text on GitHub.
- Agent Profile
- Core Capabilities
- 1. Comprehensive Benchmarking Framework
- 2. Performance Regression Detection
- 3. Automated Performance Testing
- 4. Performance Validation Framework
- MCP Integration Hooks
- Benchmark Execution Integration
- Operational Commands
- Benchmarking Commands
- Regression Detection Commands
- Integration Points
- With Other Optimization Agents
- With CI/CD Pipeline
Before you install
- Read the whole file first. Skills, commands, and subagents are instructions Claude will follow, so make sure they match what you want.
- Check which tools, scripts, or MCP servers it uses. Local servers and scripts run with your permissions.
- Try it in a test project or a copy of your files before pointing it at real work.
- Pin the version you tested, and review changes before updating.
- Watch for instructions that fetch web content or run shell commands; those are where prompt injection risks start. See our prompt injection guide.
FAQ
What is Benchmark Suite?
Benchmark Suite is a subagent for Claude Code and Claude Cowork from the ruvnet/agentic-flow repository on GitHub. Comprehensive performance benchmarking, regression detection and performance validation
How do I install Benchmark Suite in Claude Code?
Download benchmark-suite.md from the repository. Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control. Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Can I use Benchmark Suite in Claude Cowork?
Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent. Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
Is Benchmark Suite safe to install?
It is a third-party community resource, not reviewed by Anthropic or this site. Read the source file first, check which tools and connectors it uses, and install only from sources you trust.
Similar resources
- Mesh Coordinator Peer-to-peer mesh network swarm with distributed decision making and fault tolerance Subagent · ruvnet/agentic-flow
- Migration Planner Comprehensive migration plan for converting commands to agent-based system Subagent · ruvnet/agentic-flow
- Memory Optimizer Specialized in managing ReasoningBank's memory system for optimal performance. Handles consolidation, pruning, and memory quality assurance to ensure efficient learning. Subagent · ruvnet/agentic-flow
- Ml Developer Specialized agent for machine learning model development, training, and deployment Subagent · ruvnet/agentic-flow
- Byzantine Coordinator Coordinates Byzantine fault-tolerant consensus protocols with malicious actor detection Subagent · ruvnet/agentic-flow
- Backend Dev Specialized agent for backend API development, including REST and GraphQL endpoints Subagent · ruvnet/agentic-flow
- Cicd Engineer Specialized agent for GitHub Actions CI/CD pipeline creation and optimization Subagent · ruvnet/agentic-flow
- Architecture SPARC Architecture phase specialist for system design Subagent · ruvnet/agentic-flow