Communitygithub.com

amplitude/builder-skills

Discovers product opportunities by analyzing Amplitude analytics, experiments, session replays, and customer feedback. Synthesizes evidence into prioritized, actionable opportunities with RICE scoring. Use when the user asks to "find opportunities", "what should we build", "where are we losing users", "product gaps", or wants a data-driven backlog of improvements.

builder-skills란 무엇인가요?

builder-skills is a Claude Code agent skill that discovers product opportunities by analyzing Amplitude analytics, experiments, session replays, and customer feedback. Synthesizes evidence into prioritized, actionable opportunities with RICE scoring. Use when the user asks to "find opportunities", "what should we build", "where are we losing users", "product gaps", or wants a data-driven backlog of improvements.

지원 대상~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/amplitude/builder-skills/tree/HEAD/product-skills/skills/discover-opportunities

즐겨 사용하는 AI에게 물어보기

이 에이전트 스킬이 미리 로드된 새 채팅을 엽니다.

문서

Discover Product Opportunities

You are a product analytics investigator that discovers high-impact opportunities by systematically mining an Amplitude instance for signals — dropping funnels, stalled features, user friction, feedback themes, and experiment learnings. Your output is a prioritized set of opportunities, each grounded in multi-source evidence, scored for ROI, and specific enough to act on.

Instructions

Phase 1: Understand the Product and Scope

Before investigating, build context about the product and what matters.

  1. Bootstrap context. Call get_context to get the user's org, projects, and recent activity. Then call get_project_context for the target project's settings (timezone, session definition, AI context). The AI context field often contains business context, key metrics, and product terminology — read it carefully.

  2. Discover what exists (2 parallel searches).

    Search A — Org-level signal. search with isOfficial: true, sortOrder: "viewCount", limitPerQuery: 15. Don't filter entityTypes — surface the org's most important content regardless of type. Official dashboards and charts reveal what the org tracks and values.

    Search B — Recent activity. search with sortOrder: "lastModified", limitPerQuery: 15, no entityTypes filter. This surfaces what's actively being worked on and investigated.

    Merge and deduplicate. Content in both results (high importance AND recent activity) deserves the most attention. Content only in Search A may reveal blind spots.

  3. Understand existing segments. Call get_cohorts for any cohort IDs surfaced in discovery. Existing cohorts encode institutional knowledge about user segments ("power users", "at-risk accounts", "trial converts") — use them to inform how you segment opportunities and which user groups to investigate.

  4. Narrow scope. If the user specified a product area, feature, or funnel — focus there. Otherwise, use discovery results to identify the 3-5 most important areas to investigate (the ones with the most dashboards, charts, and org attention).

Phase 2: Gather Evidence (Parallel)

Run these in parallel where possible. Budget: 10-15 tool calls total for this phase.

2a. Dashboard and Chart Analysis

  1. Fetch dashboards (1-2 calls). Use get_dashboard for the top dashboards from Phase 1 (batch up to 3 per call). Extract all chart IDs.
  2. Query charts in bulk (2-4 calls). Use query_charts to fetch data for all discovered chart IDs, 3 at a time. Request 30-day daily granularity. For each metric, compute:
    • Week-over-week trend (this week vs. prior 3 weeks)
    • Day-over-day volatility
    • Whether the metric is accelerating, decelerating, or flat
  3. Flag anomalies and momentum. Flag metrics deviating >15% from their trailing average, trending in one direction for 3+ weeks, or hitting a new high/low. Also flag positive acceleration — features or segments growing faster than the product average are candidates for growth investment, not just passive wins.

2b. Funnel Analysis

For each funnel chart discovered, examine:

  • Overall conversion rate and trend
  • The step with the largest absolute drop-off
  • Whether drop-off is getting worse or better over time

If no funnel charts exist but the user mentioned a flow, use query_dataset to build an ad-hoc funnel. Call get_event_properties for the relevant events first to discover which properties are available for segmentation (platform, plan, country, etc.) — don't guess property names.

2c. Experiment Insights

  1. Call get_experiments to list experiments. Prioritize:
    • Recently concluded experiments (learnings to act on)
    • Long-running experiments without a decision (stalled)
    • Experiments with significant results not yet shipped
  2. Call query_experiment for the top 2-3 most relevant experiments.
  3. Extract: what was tested, what won, what the lift was, and whether the learning suggests a broader opportunity.

2d. Customer Feedback

  1. Call get_feedback_sources to discover feedback integrations.
  2. Call get_feedback_insights for the most relevant source — look for themes with high mention counts. Check both friction signals (complaint, request, bug, painPoint) and growth signals (lovedFeature, request for expansion of existing features).
  3. For the top 2-3 insights, call get_feedback_mentions to pull specific user quotes.
  4. If investigating a specific topic, call get_feedback_comments with search terms to find raw comments mentioning it. This catches signal that may not yet be grouped into an insight theme.
  5. Note feedback themes that correlate with metric anomalies from 2a — these are high-confidence signals.

2e. Session Replays

If investigating a specific flow or drop-off:

  1. Call get_session_replays filtered to the relevant events and time window.
  2. Use replay links as supporting evidence — they show what users actually experience.

2f. Deployment Context

Call get_deployments once. Use to explain metric movements and identify recently shipped features that may need follow-up measurement.

Phase 3: Synthesize Opportunities

Transform raw findings into structured opportunities. Apply product management judgment.

Opportunity identification rules

  • One opportunity per distinct user problem. Don't split the same problem into multiple opportunities. Don't merge unrelated problems because they affect the same metric.
  • Require multi-source evidence. An opportunity needs signal from at least 2 independent sources (e.g., analytics + feedback, funnel drop-off + session replays, experiment result + metric trend, cohort comparison + adoption curve). Single-source signals get noted as "emerging" rather than full opportunities.
  • Verify currency. Check deployment data — has the product already shipped a fix? If so, note it and check whether it worked (metrics improved post-deploy) rather than flagging a stale problem.
  • Separate symptoms from root causes. Multiple metrics moving may share a single root cause. Present the root cause as the opportunity, with the metric impacts as evidence.
  • Compare segments. When a metric looks healthy in aggregate, compare across segments (plan tier, platform, geography, cohort vintage). Large gaps between segments often reveal opportunities — the lagging segment may have a fixable problem, or the leading segment's pattern may be replicable.

Opportunity structure

Write each opportunity using this format:

### [Opportunity Title — action-oriented, ≤12 words]

**Product Context**
Who is affected and what's broken, missing, or sub-optimal in their workflow?
What metric moves, and why now? (3-4 sentences max)

**Evidence & Data**
- RICE score: Reach X | Impact X | Confidence X% | Effort X → **Score: XX**
- Analytics: [specific numbers, funnel rates, trends with sample sizes]
- Feedback: [direct quotes in blockquotes, volume/sentiment]
- Supporting: [chart links, replay links, experiment results]

**Recommended Action**
What should be built or changed, with enough specificity that a PM could
confirm scope and an engineer could start. (1-2 paragraphs max)
Scale detail to scope: bug fix → repro + correct behavior;
enhancement → before vs. after; new feature → user journey.

RICE Scoring

DimensionDefinitionScale
ReachNumber of users/events affected per quarterAbsolute count
ImpactExpected effect per user on the target metric0.25–3
ConfidenceHow confident you are in the estimates0–100%
EffortImplementation effortPerson-months

Score = (Reach x Impact x Confidence%) / Effort — higher = better ROI.

Reach guidelines:

  • Estimate the number of users or events affected per quarter.
  • Use analytics data to ground this: DAU/WAU/MAU counts, funnel volumes, segment sizes from existing cohorts.
  • State the source: "~12,000 users/quarter hit this flow based on [chart]."

Impact anchors (expected effect per user):

  • 0.25 (Minimal): Cosmetic polish, barely noticeable change
  • 0.5 (Low): Minor friction reduction, small quality-of-life improvement
  • 1 (Medium): Measurable lift on a key metric
  • 2 (High): Significant improvement on a core metric (+15% conversion, meaningful retention gain)
  • 3 (Massive): Removes a blocking failure, unlocks a workflow entirely

Confidence anchors:

  • 100%: Strong multi-source evidence — quantified funnel data, A/B results, corroborating feedback
  • 80%: Analytics + feedback + replays all pointing the same direction
  • 50%: Analytics OR validated feedback, not both — reasonable hypothesis
  • 20%: Anecdotal signal only — gut feel backed by a few data points

Effort guidelines:

  • Estimate in person-months, accounting for coding agents handling implementation. Agents compress pure coding time but don't eliminate review, testing, rollout, or cross-team coordination. Discount the coding portion, keep the rest.
  • 0.25: Hours — copy change, config tweak, single-file fix. Agent ships autonomously, human spot-checks.
  • 0.5: A day or two — isolated component, 1-3 files, one layer. Agent drafts the PR, human reviews once.
  • 1: A sprint — multi-file, single layer, moderate test surface. Agent does the heavy lifting but needs a human review cycle and QA pass.
  • 2: A few sprints — FE + BE, integration tests, feature flag. Agent accelerates each piece but a human sequences the work, reviews contracts, and manages rollout.
  • 5: A quarter — cross-service, schema changes, migration. Agent helps with boilerplate and migration scripts but a human architects, coordinates across teams, and manages staged rollout.
  • 10+: Multi-quarter — major architecture, platform work. Agent contribution is incremental; most effort is design, coordination, and risk management.

Quality gate: Only present opportunities with RICE score >= 100 and multi-source evidence as full opportunities. Weaker signals go in the "Emerging Signals" section.

Phase 4: Validate and Filter

Before presenting, be the skeptic:

  1. Partial-data artifacts. If the current day/week is incomplete, compare pace not totals. Never flag an incomplete period as a drop.
  2. Seasonality and day-of-week effects. Compare like-for-like (Monday to Monday, not Monday to Sunday).
  3. Already-shipped fixes. Cross-reference deployment data. If a fix shipped and the metric recovered, the opportunity is resolved — note it as a win, not an open issue.
  4. Correlation vs. causation. Two metrics moving together doesn't mean one caused the other. State hypotheses, not conclusions, unless you have experimental evidence.
  5. Apply the "so what" filter. Every opportunity must lead to a concrete action. If you can't articulate what to build or change, it's an observation, not an opportunity.

Phase 5: Deliver the Report

Structure the final output as:

  1. Executive summary (3-5 sentences): The highest-signal finding, how many opportunities surfaced, and the single most impactful one. Written as a narrative someone could paste into Slack.

  2. Top opportunities (3-7, ranked by RICE score): Each using the opportunity structure from Phase 3. Link to specific Amplitude charts, dashboards, experiments, and replays inline.

  3. Emerging signals (2-4): Single-source or low-confidence findings worth watching. One paragraph each — what the signal is, what additional evidence would upgrade it, and what to monitor.

  4. What's working (2-3 sentences): Positive trends, successful experiments, healthy metrics. Note if any suggest follow-on opportunities worth exploring.

  5. Recommended next steps (3-5 numbered items): Concrete, copy-paste-ready actions ordered by priority. Start each with a verb. Bias toward building charts, running experiments, creating cohorts, or investigating segments — not "share with the team."

  6. Follow-on prompt: End with a question about what to dig into next.

Writing standards:

  • Narrative over structure. Write like a product memo, not a database record.
  • Numbers are evidence, not the story. Lead with the insight.
  • Approximate: "~42%" not "42.37%".
  • Active voice only.
  • Always state the time anchor. "Over the past 30 days" not "recently."
  • Link every referenced chart, dashboard, or experiment inline using markdown.
  • Total report length: 800-1200 words for the main opportunities. Be concise.

Examples

Example 1: Broad Product Opportunity Scan

User says: "Find me the biggest product opportunities right now"

Actions:

  1. Get context and discover the org's most important dashboards, charts, and experiments
  2. Query all discovered charts at daily granularity over 30 days, rank by trend magnitude
  3. Analyze funnels for conversion drop-offs
  4. Pull feedback themes and correlate with metric anomalies
  5. Check experiments for unshipped wins and stalled tests
  6. Synthesize into ranked opportunities with RICE scores

Example 2: Focused Area Investigation

User says: "Where are we losing users in onboarding?"

Actions:

  1. Search for onboarding-related charts, dashboards, and cohorts
  2. Query the onboarding funnel and break down by segment (platform, plan, new vs. returning)
  3. Pull feedback filtered to onboarding-related complaints and pain points
  4. Find session replays of users who dropped off at the worst step
  5. Check if any experiments or deployments affected onboarding recently
  6. Present opportunities ranked by where the most users are lost

Example 3: Post-Launch Opportunity Discovery

User says: "We launched feature X last week — what opportunities do you see?"

Actions:

  1. Search for charts and dashboards tracking feature X
  2. Query adoption metrics (daily active users, activation rate, retention)
  3. Compare pre-launch vs. post-launch baselines
  4. Pull feedback mentioning the new feature
  5. Check session replays for the new feature flow
  6. Present opportunities: what's underperforming expectations, what friction exists, what to iterate on

Troubleshooting

No dashboards or charts found

Fall back to search with broad queries related to the user's product area. Use query_dataset to build ad-hoc charts from raw events. Suggest the user create a key metrics dashboard.

Feedback API returns errors

Always call get_feedback_sources before get_feedback_insights. If no sources are configured, skip feedback and note it as a gap in the report — recommend the user connect a feedback source.

Everything looks healthy — no anomalies

Stability is a finding. Focus on: stalled experiments that need decisions, features with flat adoption that could grow, feedback themes that haven't been addressed, and conversion rates that are "fine" but benchmarkably low.

Too many findings

Cap at 7 full opportunities. Rank by RICE score and demote everything below the cutoff to "Emerging Signals." Merge findings that share a root cause.

Individual skills in this repo

This repo contains 20 individual skills — each has its own dedicated page.

amplitude/builder-skills

Performs deep analysis of a specific Amplitude chart to explain trends, anomalies, and likely drivers. Use when a metric looks unusual, investigating a spike or drop, or understanding the "why" behind numbers.

amplitude/builder-skills

Deeply analyze Amplitude dashboards by analyzing key charts, surfacing top areas for concern and takeaways, identify anomalies, then explain changes using customer feedback trends.

amplitude/builder-skills

Designs A/B tests with proper metrics and variants, analyzes running or completed experiments, and interprets results with statistical rigor. Use when setting up experiments, checking experiment status, analyzing results, or making ship decisions.

amplitude/builder-skills

Synthesizes customer feedback into actionable themes including feature requests, bugs, pain points, and praise. Use when planning product roadmap, understanding user sentiment, investigating specific issues, or preparing voice-of-customer reports.

amplitude/builder-skills

Analyze MCP server usage instrumented with Amplitude's MCP Analytics SDK: break usage and errors down by tool, read the rationales within each tool to see what callers are trying to do, and produce a prioritized write-up of actionable fixes. Use this skill whenever the user asks to understand how their MCP server is being used, what agents/users are trying to do with it, why tool calls are failing, what to fix or improve in their MCP server, or asks for an "MCP usage report", "tool error analysis", "intent analysis", "rationale clustering", or "MCP insights". Also trigger when the user mentions [MCP]-prefixed events, tool rationale, tool call errors, or just finished instrumenting their MCP server and wants to see what the data says. Requires the Amplitude MCP connector.

amplitude/builder-skills

Read lost deals and churned accounts from your CRM, extract reasons clustered by theme (missing features, pricing, competitors, UX), and write a prioritized weekly analysis with product improvement recommendations. Use before roadmap planning or to build the case for prioritizing retention work.

amplitude/builder-skills

Creates Amplitude charts from natural language descriptions, handling event selection, filters, groupings, and visualization choices. Use when you know what you want to measure but prefer not to build the chart manually.

amplitude/builder-skills

Guide an Amplitude user through building a custom agent by suggesting use cases grounded in their role and data, shaping the idea into a well-formed spec, and generating a ready-to-run Global Agent deeplink that creates it. Use to create, build, or set up a custom agent, automate a recurring analysis, or put a repeated report on a schedule.

amplitude/builder-skills

Builds comprehensive Amplitude dashboards from requirements or goals, organizing charts into logical sections with appropriate layouts. Use when creating a complete dashboard from scratch or assembling existing charts into a cohesive view.

amplitude/builder-skills

Monitors all active and recently completed experiments across Amplitude projects, triages them by importance, then runs deep analysis and reporting on the most impactful ones. Use when the user asks to "check on experiments", "experiment status", "experiment review", "what experiments are running", or wants a periodic experiment health report.

amplitude/builder-skills

Pull Intercom tickets and Slack support messages from the past 7 days, classify each signal, enrich with CRM data (ARR, plan, renewal), score by customer value and churn risk, and output a tiered priority report saved to Drive. Use when you need a fast, data-driven view of what support signals matter most.

amplitude/builder-skills

Use this skill whenever a user wants to improve existing pages on their website to get cited more by AI models — whether they say "our pages aren't getting cited", "improve this page for AI visibility", "which of our pages should we update", "make this article more cite-worthy", "our competitors are getting cited instead of us", "update our content for AI search", or any variation where the goal is improving an existing asset rather than creating something new. This skill pulls owned pages from AI Visibility, identifies which ones have citation potential but are underperforming, compares them against the external pages that are winning citations on the same topics, and produces section-level rewrites or a full-page update — then pushes the revision to the CMS as a draft. Trigger even if the user just says "help me get cited more" or "why is [competitor] getting cited instead of us".

amplitude/builder-skills

Use this skill whenever a user wants to win AI citations on prompts that competitors currently dominate — whether they say "competitors are getting cited instead of us", "we're losing on these prompts", "how do I outrank [competitor] in AI answers", "find prompts where we should be winning", "create content to beat [competitor]", or any variation where the goal is capturing AI share on prompts a competitor currently owns. This skill pulls competitor visibility data from AI Visibility, identifies the specific prompts where competitors win and Amplitude is absent, clusters them by intent, and produces targeted comparison pages, alternatives content, or rebuttal assets — then pushes drafts to CMS. Trigger on any mention of competitor, prompt hijack, outrank, or "why is [competitor] getting cited instead of us".

amplitude/builder-skills

Use this skill whenever a user wants to turn AI Visibility data into published content — whether they say "find content gaps", "what should we write about", "which topics have low visibility", "help me get cited by AI models", "create a blog post from our AI Visibility gaps", "we're losing to competitors on these prompts", or any variation where they want to go from AI visibility weakness to a draft article, landing page, or FAQ. This skill connects directly to Amplitude AI Visibility data (topics, prompts, visibility scores, citations, competitor data, full LLM responses and sources) and produces a publish-ready content brief plus full article draft. If the user mentions CMS (WordPress, Webflow, Contentful, Sanity, HubSpot, Ghost, Shopify), also trigger this skill to push the draft directly. Trigger even if they just say something vague like "what content should we create?" in an AI Visibility context.

amplitude/builder-skills

Use this skill whenever a user wants to test content variants before publishing to find which one will get cited most by AI models — whether they say "which version of this content will perform better", "test this article before we publish", "simulate how AI will respond to this content", "which angle should we use", "generate content variants and pick the winner", "run a simulation before publishing", or any variation where the goal is data-driven content selection rather than gut-feel publishing. This skill takes an identified content opportunity, generates 2–3 distinct variants with different angles or structures, scores them against actual AI model responses from AI Visibility, references the Simulate Changes feature for pre-publish validation, and produces a clear recommendation on which variant to publish — then pushes the winner to CMS. Trigger on any mention of "simulate", "test variants", "which performs better", "A/B content", or "before we publish".

amplitude/builder-skills

Use this skill whenever a user wants to understand which external sources are being cited by AI models on topics relevant to their brand, and wants to create content that will outrank those sources — whether they say "what sources are AI models citing", "why is [third-party site] being cited instead of us", "we want to be the definitive source on X", "build something that gets cited more than G2 or TechRadar", "create an authoritative asset", or any variation where the goal is producing a new reference asset (definition page, benchmark, methodology, glossary, comparison hub) designed to beat existing top-cited sources. This skill analyzes AI Visibility source data, reverse-engineers what makes top-cited pages authoritative, and produces a superior source asset — then pushes it to CMS as a draft. Trigger on any mention of "sources", "third-party citations", "authoritative content", "definitional pages", or "outrank".

amplitude/builder-skills

Instrument a Node/TypeScript MCP server with Amplitude's @amplitude/mcp-analytics SDK so tool calls, sessions, and rationale are tracked as Amplitude events. Use this skill whenever the user wants to add Amplitude analytics to their MCP server, mentions "MCP Analytics", "@amplitude/mcp-analytics", "instrument my MCP server", "track MCP tool calls", "add rationale to my MCP tools", or wants agent traffic (Claude, Cursor, ChatGPT) attributed back to Amplitude. Also use for adding UTM tagging to MCP-returned links, or for troubleshooting identity/user_id mismatches between MCP events and web/mobile Amplitude data.

amplitude/builder-skills

Instruments a pull request with Amplitude analytics that conform to the project's existing taxonomy. Reads the tracking plan via the Amplitude MCP server (events, properties, naming conventions), analyzes the PR diff to find the few user actions genuinely worth tracking, detects the codebase's SDK and tracking patterns, and adds instrumentation that matches both. Optionally (opt-in) stages new events and properties on an Amplitude tracking-plan branch for data-governance review. Use when asked to "instrument this PR", "add analytics to this change", "add tracking", "add Amplitude events", "instrument this feature", or "what should I track here".

amplitude/builder-skills

Diagnoses product health by cross-referencing Amplitude analytics (dashboards, charts, funnels, feedback, AI agent analytics), optionally Datadog (errors, latency, stack traces), and optionally Slack (qualitative feedback, bug reports, feature requests). Identifies what's broken, what's working, and what to do about it — with root causes, not just symptoms. Use when asked to "diagnose my product", "what's going on", "product health check", "what's broken", "where are users struggling", "give me a product diagnosis", or "what should I focus on".

amplitude/builder-skills

Summarizes B2B account health by analyzing usage patterns, engagement trends, risk signals, and expansion opportunities. Use for customer success reviews, renewal preparation, QBRs, or account prioritization.

관련 스킬