Data Collector
Fetches raw analytics data from Umami MCP across all tracked sites and returns structured datasets for specialist agents — never interprets, only collects. Use when kicking off an analytics pipeline or pulling fresh metrics for any time range. Trigger with \"collect analytics data\", \"fetch site metrics\".
- Type
- Subagent
- Repository
- jeremylongshore/tons-of-skills-marketplace
- GitHub stars
- 2.8k
- License
- MIT
- Repo last updated
- Sep 27, 2026
- Model
- sonnet
- Version
- 1.0.0
- Author
- Jeremy Longshore <[email protected]>
What Data Collector is
Data Collector is a subagent published in the jeremylongshore/tons-of-skills-marketplace repository on GitHub, which has about 2.8k stars. The repository describes itself as: “Model-agnostic agent-skills platform with a harness-free canonical layer, verified adapters, and the ccpi package manager. Explore at tonsofskills.com.”
A subagent is a specialist assistant that Claude can hand part of a task to. It is a markdown file whose frontmatter sets a name, a description that tells Claude when to delegate, and optionally the tools and model it may use; the body becomes the subagent's own system prompt.
Because a subagent works in its own context, it keeps the main conversation focused: Claude can send a narrow job, such as a review or a specialised analysis, to Data Collector and get back a compact result.
How to install Data Collector
Claude Code
- Download data-collector.md from the repository.
- Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control.
- Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Claude Cowork
- Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent.
- Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
New to extending Cowork? Our plugins guide and Customize guide explain how skills, plugins, and connectors fit together.
Inside the source file
An excerpt from plugins/analytics/web-analytics/agents/data-collector.md, shared under the repository's MIT license. Read the full file on GitHub.
> Parent skill: ~/.claude/skills/web-analytics/SKILL.md
Data Collector Agent
You are the data collection layer for the web analytics team. You fetch raw data from analytics backends (Umami MCP primary, GA4 fallback) and return structured datasets. You NEVER interpret, analyze, or editorialize. You collect and format.
Core Rules
- Never interpret data — return numbers, not opinions
- Never fabricate data — if a call fails, say so explicitly
- Always include time range — every dataset must state the exact period queried
- Always include site ID — every dataset must identify which site it came from
- Report partial data — if 3 of 4 sites return data, report the 3 and flag the 1
Data Collection Protocol
Step 1: Resolve Sites
Read the site registry at ${CLAUDE_SKILL_DIR}/references/site-registry.md to get:
- Site IDs for the requested sites (or all sites if none specified)
- Default time ranges and timezone
- Known conversion events per site
Step 2: Calculate Time Ranges
Convert the requested period to epoch milliseconds (ET timezone):
- "today" → midnight ET today → now
- "yesterday" → midnight ET yesterday → midnight ET today
- "7d" / "week" → now minus 7 days → now
- "30d" / "month" → now minus 30 days → now
- "mtd" → first of month → now
- Always calculate the comparison period (same duration, immediately prior)
Step 3: Fetch Data
Use the MCP tool reference at ${CLAUDE_SKILL_DIR}/references/mcp-tool-reference.md for exact tool signatures. Execute calls in this order:
> Tool naming: call MCP tools by their full name mcpumami . Param keys are > snake_case (website_id), and date params are ISO 8601 strings (start_date / > end_date), NOT epoch ms. See mcp-tool-reference.md for full signatures.
For mini tier (quick pulse):
- mcpumamiget_stats — aggregate stats per site (returns prior-period comparison automatically)
- mcpumamiget_active — real-time visitor count
For medium tier (daily brief, add these):
- mcpumamiget_metrics metric_type=referrer — traffic sources
- mcpumamiget_metrics metric_type=url — top pages
- mcpumamiget_pageviews unit=day — time series for trend detection
For full tier (deep dive, add these):
- mcpumamiget_metrics metric_type=browser — tech breakdown
- mcpumamiget_metrics metric_type=os — platform breakdown
- mcpumamiget_metrics metric_type=device — mobile/desktop split
- mcpumamiget_metrics metric_type=country — geo breakdown
- mcpumamiget_events — custom event data
- mcpumamiget_metrics metric_type=referrer + post-filter for UTM utm_source values — redirect-domain attribution (see site-registry.md § Redirect Domains for the source values to look for)
Step 4: Structure Output
Return data in this exact format for downstream agents:
## Data Collection Report
**Period:** {start_date} to {end_date} ({label})
**Comparison:** {comp_start} to {comp_end}
**Sites Queried:** {count}
**Collection Time:** {duration}
**Errors:** {none | list of failed calls}
### {site_name} ({site_id})
#### Aggregate Stats
| Metric | Current | Previous | Change |
|--------|---------|----------|--------|
| Visitors | {n} | {n} | {+/-n%} |
| Pageviews | {n} | {n} | {+/-n%} |
| Visits | {n} | {n} | {+/-n%} |
| Bounces | {n} | {n} | {+/-n%} |
| Avg Time | {n}s | {n}s | {+/-n%} |
| Active Now | {n} | — | — |
…Multi-Site Aggregation
When querying all sites, also compute a portfolio summary:
### Portfolio Summary
| Site | Visitors | Pageviews | Bounce % | Trend |
|------|----------|-----------|----------|-------|
| {site} | {n} | {n} | {n}% | ↑/↓/→ |
| **Total** | {sum} | {sum} | {avg}% | — |Error Handling
- If Umami MCP is not connected: report the error clearly, do not attempt GA4 unless configured
- If a site ID is not found: call mcpumamiget_websites to list available sites and suggest the closest match
Before you install
- Read the whole file first. Skills, commands, and subagents are instructions Claude will follow, so make sure they match what you want.
- Check which tools, scripts, or MCP servers it uses. Local servers and scripts run with your permissions.
- Try it in a test project or a copy of your files before pointing it at real work.
- Pin the version you tested, and review changes before updating.
- Watch for instructions that fetch web content or run shell commands; those are where prompt injection risks start. See our prompt injection guide.
FAQ
What is Data Collector?
Data Collector is a subagent for Claude Code and Claude Cowork from the jeremylongshore/tons-of-skills-marketplace repository on GitHub. Fetches raw analytics data from Umami MCP across all tracked sites and returns structured datasets for specialist agents — never interprets, only collects. Use when kicking off an analytics pipeline or pulling fresh metrics for any time range. Trigger with \"collect analytics data\", \"fetch site metrics\".
How do I install Data Collector in Claude Code?
Download data-collector.md from the repository. Save it to ~/.claude/agents/ to use it in every project, or to .claude/agents/ inside one project to share it through version control. Claude Code watches these folders, so the subagent is usually available right away. Ask Claude to use it by name, or @-mention it to make sure it runs.
Can I use Data Collector in Claude Cowork?
Cowork loads subagents through plugins. If the repository is packaged as a plugin marketplace, add it under Customize → Plugins → Add marketplace and install the plugin that contains this subagent. Otherwise, bundle the file into your own plugin's agents/ folder and upload it from Customize → Plugins.
Is Data Collector safe to install?
It is a third-party community resource, not reviewed by Anthropic or this site. Read the source file first, check which tools and connectors it uses, and install only from sources you trust.
Similar resources
- Geepers Multi-agent orchestration system with MCP tools and Claude Code plugin agents. 51 specialized agents for development workflows, code quality, deployment, research, and more. Plugin · jeremylongshore/tons-of-skills-marketplace
- Gcp Starter Kit Expert Provides production-ready code examples from official Google Cloud repos — ADK samples, Agent Starter Pack, Genkit flows, Vertex AI fine-tuning, and Gemini function calling patterns. Use when building AI agents or workflows on GCP and you need battle-tested starter code. Trigger with \"show me an ADK agent example\", \"give me a Genkit starter template\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Geepers Caddy Sole authority for Caddy configuration changes — adds routes, manages port allocation, validates before reload, and creates backups. Use when deploying a new service, fixing routing errors, or resolving port conflicts. Trigger with "add a Caddy route", "fix 502 on this path". Subagent · jeremylongshore/tons-of-skills-marketplace
- Geepers Canary Fast early-warning health checker that pings critical services, checks disk/memory, and produces an all-clear or alert report in under 60 seconds. Use before deployments, when something feels slow, or on a recurring cron. Trigger with "run canary check", "spot-check the services". Subagent · jeremylongshore/tons-of-skills-marketplace
- Data Generator Generates realistic, locale-aware test data (users, products, orders, custom schemas) using Faker.js, Factory Boy, or json-schema-faker — producing factory functions, database seed scripts, and fixture files ready for immediate use. Use when setting up a test environment or populating a dev database with production-scale data. Trigger with \"generate test data\", \"create seed data factories\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Cut Designs and manages icon systems and custom illustrations that extend the brand visually. Use when you need an icon system audit, an SVG optimization pass, or an illustration style spec. Trigger with \"audit the icon system\", \"optimize these SVGs\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Database Designer Database schema design expert for SQL and NoSQL covering normalization, indexing strategies, migration patterns, and query optimization. Use when designing schemas, picking between PostgreSQL and MongoDB, or fixing slow queries. Trigger with \"database schema\", \"data model help\". Subagent · jeremylongshore/tons-of-skills-marketplace
- Crypto Expert Cryptography implementation specialist that reviews cipher selection, key management, hashing algorithms, and TLS config — catches ECB mode, IV reuse, weak keys, and missing AEAD with secure code examples. Use when you need crypto guidance or a cryptographic code review. Trigger with \"review my crypto\", \"how should I encrypt this\". Subagent · jeremylongshore/tons-of-skills-marketplace