Communitygithub.com

amplitude/builder-skills

Define a North Star Metric and its input metrics. Classify the game your product plays, evaluate candidates against quality criteria, and build a connected metric system. Use when choosing a North Star, evaluating an existing one, or setting up a metrics framework.

What is builder-skills?

builder-skills is a Claude Code agent skill that define a North Star Metric and its input metrics. Classify the game your product plays, evaluate candidates against quality criteria, and build a connected metric system. Use when choosing a North Star, evaluating an existing one, or setting up a metrics framework.

Works with~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/amplitude/builder-skills/tree/HEAD/growth-skills/skills/north-star-metric

Ask in your favorite AI

Open a new chat with this agent skill pre-loaded.

Documentation

North Star Metric

Define a North Star Metric that captures customer value, aligns your team, and predicts long-term business success.

A North Star Metric is not just "the number you track." It's the single metric that best captures the value customers get from your product. Get it right and it aligns product, engineering, marketing, and leadership around what actually matters. Get it wrong and teams optimize for vanity metrics that look good in a deck but don't move the business.


Domain Context

What a North Star Metric IS

A North Star Metric (NSM) is a single, customer-centric measurement that reflects the aggregate value people get from your product. It serves three functions:

  1. Clarity and alignment — defines what the product team optimizes and what trade-offs are acceptable
  2. Communication — demonstrates the product org's impact and progress to stakeholders
  3. Accountability — holds teams responsible for outcomes, not just shipping features

What a North Star Metric is NOT

  • Not revenue. Revenue is a lagging indicator. By the time you count it, you can no longer influence the outcome. Your NSM should predict revenue, not be revenue.
  • Not DAU or registered users. These are vanity metrics — they look good on the surface but don't tell you if customers are getting value.
  • Not an OKR. OKRs are goal-setting tools. You can use Key Results to express expected change in your NSM, but they're different things.
  • Not multiple metrics. One NSM per product line. If you have competing North Stars, you don't have a North Star.

The Three Games

Before choosing your NSM, classify which game your product plays:

  • Attention Game — Success = time spent in-product. Users come for content, entertainment, or ongoing engagement. (Examples: Netflix, Spotify, TikTok)
  • Transaction Game — Success = transaction volume. Users come to buy, sell, or exchange value. (Examples: Amazon, Airbnb, Stripe)
  • Productivity Game — Success = work completed efficiently. Users come to get a job done faster or better. (Examples: Notion, Figma, Amplitude)

Most products play one game. Trying to play all three dilutes focus. Know your game before picking your metric.

The Metric System: NSM + Inputs + Business KPI

A North Star doesn't work alone. You need three layers:

  1. North Star Metric — aggregate customer value (e.g. "weekly learning hours" for an education product)
  2. Input Metrics (3-5) — the controllable, complementary factors that drive the NSM. A useful pattern: Breadth (how many customers?), Depth (how much value per customer?), Frequency (how often?), Efficiency (how fast/friction-free?). These are what teams actually move day-to-day.
  3. Top Business KPI — value back to the business (revenue, paid seats, ARR). This must be structurally connected to the NSM. If your NSM goes up but the business KPI doesn't, you have a monetization problem. If the business KPI goes up but NSM doesn't, you're extracting value you're not creating.

When to Use

  • Defining or redefining your product's key metric
  • Evaluating whether your current North Star is actually good
  • Setting up a metrics framework for a new product or team
  • Aligning stakeholders on what success looks like
  • Preparing for planning season or OKR-setting

Prompt

You are a product strategist specializing in North Star metrics and growth measurement frameworks. You understand the Amplitude North Star Framework deeply and help teams move past vanity metrics to find metrics that actually capture customer value.

Given the following context about my product: $ARGUMENTS

Work through these steps:

Step 1: Classify the Game

Determine which game this product plays: Attention, Transaction, or Productivity. Explain your reasoning. If the product straddles two games, identify the primary one and explain why focus matters.

Step 2: Generate NSM Candidates

Propose 2-3 North Star Metric candidates. For each one:

  • State the metric clearly (name + unit + time window)
  • Explain what customer value it captures
  • Explain why it's a leading indicator of business success

Also propose 1-2 deliberately bad candidates and explain why they fail. This sharpens thinking about what makes a good NSM — sometimes it's easier to see what's wrong than what's right.

Step 3: Evaluate Against Quality Criteria

Score each candidate against these six criteria:

  1. Customer value — Does it reflect the value customers get, not just activity? Would increasing this metric genuinely mean customers are better off?
  2. Strategic clarity — Could someone read this metric and understand your product strategy? Does it encode where you're going, not just where you are?
  3. Leading indicator — Does it predict future revenue and retention, or just report past results?
  4. Actionable — Can product teams influence it through their daily work? If you can move your NSM directly with a single change, it's probably too narrow. If no team can move it, it's too abstract.
  5. Measurable — Can you instrument it today, or is it aspirational? (Aspirational is OK if you have a plan to instrument it.)
  6. Hard to game — Could teams hit this number through perverse incentives? If so, what counter-metric would you need?

Step 4: Recommend an NSM

Pick the strongest candidate and explain why. Be direct about the trade-offs. There's no perfect metric — explain what this one captures well and what it misses (and how the input metrics compensate).

Step 5: Define Input Metrics

Propose 3-5 input metrics that drive the NSM. Structure them using the Breadth-Depth-Frequency-Efficiency pattern where it fits:

  • Breadth — How many customers are getting value? (e.g. active accounts, activated users)
  • Depth — How much value per customer? (e.g. features adopted, queries run per user)
  • Frequency — How often do they come back? (e.g. weekly sessions, monthly active days)
  • Efficiency — How friction-free is the path to value? (e.g. time to first insight, onboarding completion rate)

For each input metric, explain:

  • What it measures and why it matters
  • Which team(s) can directly influence it
  • How it connects mathematically to the NSM (additive, multiplicative, or correlated)

Step 6: Validate the System

Check that the full metric system holds together:

  • If all input metrics improve, does the NSM necessarily improve?
  • If the NSM improves, does the business KPI (revenue) eventually follow?
  • Are there gaps — important aspects of customer value that nothing measures?
  • Are there conflicts — could optimizing one input hurt another?

Flag any structural weaknesses and suggest how to address them.

Output Format

Deliver:

  1. Game classification with rationale
  2. NSM candidates (good and bad) with evaluation scores
  3. Recommended NSM with reasoning
  4. Input metrics with ownership and connections
  5. System validation and any gaps or risks
  6. Suggested next step — what to do once the NSM is defined (instrument it, build a dashboard, set targets)

Be opinionated. A wishy-washy "it depends" answer is worse than a strong recommendation with caveats.


Tips

  • Bring your product description, current metrics, customer segments, and business model. The more context, the better the recommendation.
  • If you already have a North Star candidate, share it — the skill will pressure-test it rather than starting from scratch.
  • Revenue is never the NSM, but always validate that your NSM predicts revenue. If it doesn't, something is broken.
  • Pair with build-metric-tree to decompose your NSM into a full quantified tree once you've defined it.
  • Revisit annually or when your strategy shifts. A good NSM lasts 1-3 years, not forever.

Further Reading

Individual skills in this repo

This repo contains 20 individual skills — each has its own dedicated page.

amplitude/builder-skills

Performs deep analysis of a specific Amplitude chart to explain trends, anomalies, and likely drivers. Use when a metric looks unusual, investigating a spike or drop, or understanding the "why" behind numbers.

amplitude/builder-skills

Deeply analyze Amplitude dashboards by analyzing key charts, surfacing top areas for concern and takeaways, identify anomalies, then explain changes using customer feedback trends.

amplitude/builder-skills

Designs A/B tests with proper metrics and variants, analyzes running or completed experiments, and interprets results with statistical rigor. Use when setting up experiments, checking experiment status, analyzing results, or making ship decisions.

amplitude/builder-skills

Synthesizes customer feedback into actionable themes including feature requests, bugs, pain points, and praise. Use when planning product roadmap, understanding user sentiment, investigating specific issues, or preparing voice-of-customer reports.

amplitude/builder-skills

Analyze MCP server usage instrumented with Amplitude's MCP Analytics SDK: break usage and errors down by tool, read the rationales within each tool to see what callers are trying to do, and produce a prioritized write-up of actionable fixes. Use this skill whenever the user asks to understand how their MCP server is being used, what agents/users are trying to do with it, why tool calls are failing, what to fix or improve in their MCP server, or asks for an "MCP usage report", "tool error analysis", "intent analysis", "rationale clustering", or "MCP insights". Also trigger when the user mentions [MCP]-prefixed events, tool rationale, tool call errors, or just finished instrumenting their MCP server and wants to see what the data says. Requires the Amplitude MCP connector.

amplitude/builder-skills

Read lost deals and churned accounts from your CRM, extract reasons clustered by theme (missing features, pricing, competitors, UX), and write a prioritized weekly analysis with product improvement recommendations. Use before roadmap planning or to build the case for prioritizing retention work.

amplitude/builder-skills

Creates Amplitude charts from natural language descriptions, handling event selection, filters, groupings, and visualization choices. Use when you know what you want to measure but prefer not to build the chart manually.

amplitude/builder-skills

Guide an Amplitude user through building a custom agent by suggesting use cases grounded in their role and data, shaping the idea into a well-formed spec, and generating a ready-to-run Global Agent deeplink that creates it. Use to create, build, or set up a custom agent, automate a recurring analysis, or put a repeated report on a schedule.

amplitude/builder-skills

Builds comprehensive Amplitude dashboards from requirements or goals, organizing charts into logical sections with appropriate layouts. Use when creating a complete dashboard from scratch or assembling existing charts into a cohesive view.

amplitude/builder-skills

Monitors all active and recently completed experiments across Amplitude projects, triages them by importance, then runs deep analysis and reporting on the most impactful ones. Use when the user asks to "check on experiments", "experiment status", "experiment review", "what experiments are running", or wants a periodic experiment health report.

amplitude/builder-skills

Pull Intercom tickets and Slack support messages from the past 7 days, classify each signal, enrich with CRM data (ARR, plan, renewal), score by customer value and churn risk, and output a tiered priority report saved to Drive. Use when you need a fast, data-driven view of what support signals matter most.

amplitude/builder-skills

Use this skill whenever a user wants to improve existing pages on their website to get cited more by AI models — whether they say "our pages aren't getting cited", "improve this page for AI visibility", "which of our pages should we update", "make this article more cite-worthy", "our competitors are getting cited instead of us", "update our content for AI search", or any variation where the goal is improving an existing asset rather than creating something new. This skill pulls owned pages from AI Visibility, identifies which ones have citation potential but are underperforming, compares them against the external pages that are winning citations on the same topics, and produces section-level rewrites or a full-page update — then pushes the revision to the CMS as a draft. Trigger even if the user just says "help me get cited more" or "why is [competitor] getting cited instead of us".

amplitude/builder-skills

Use this skill whenever a user wants to win AI citations on prompts that competitors currently dominate — whether they say "competitors are getting cited instead of us", "we're losing on these prompts", "how do I outrank [competitor] in AI answers", "find prompts where we should be winning", "create content to beat [competitor]", or any variation where the goal is capturing AI share on prompts a competitor currently owns. This skill pulls competitor visibility data from AI Visibility, identifies the specific prompts where competitors win and Amplitude is absent, clusters them by intent, and produces targeted comparison pages, alternatives content, or rebuttal assets — then pushes drafts to CMS. Trigger on any mention of competitor, prompt hijack, outrank, or "why is [competitor] getting cited instead of us".

amplitude/builder-skills

Use this skill whenever a user wants to turn AI Visibility data into published content — whether they say "find content gaps", "what should we write about", "which topics have low visibility", "help me get cited by AI models", "create a blog post from our AI Visibility gaps", "we're losing to competitors on these prompts", or any variation where they want to go from AI visibility weakness to a draft article, landing page, or FAQ. This skill connects directly to Amplitude AI Visibility data (topics, prompts, visibility scores, citations, competitor data, full LLM responses and sources) and produces a publish-ready content brief plus full article draft. If the user mentions CMS (WordPress, Webflow, Contentful, Sanity, HubSpot, Ghost, Shopify), also trigger this skill to push the draft directly. Trigger even if they just say something vague like "what content should we create?" in an AI Visibility context.

amplitude/builder-skills

Use this skill whenever a user wants to test content variants before publishing to find which one will get cited most by AI models — whether they say "which version of this content will perform better", "test this article before we publish", "simulate how AI will respond to this content", "which angle should we use", "generate content variants and pick the winner", "run a simulation before publishing", or any variation where the goal is data-driven content selection rather than gut-feel publishing. This skill takes an identified content opportunity, generates 2–3 distinct variants with different angles or structures, scores them against actual AI model responses from AI Visibility, references the Simulate Changes feature for pre-publish validation, and produces a clear recommendation on which variant to publish — then pushes the winner to CMS. Trigger on any mention of "simulate", "test variants", "which performs better", "A/B content", or "before we publish".

amplitude/builder-skills

Use this skill whenever a user wants to understand which external sources are being cited by AI models on topics relevant to their brand, and wants to create content that will outrank those sources — whether they say "what sources are AI models citing", "why is [third-party site] being cited instead of us", "we want to be the definitive source on X", "build something that gets cited more than G2 or TechRadar", "create an authoritative asset", or any variation where the goal is producing a new reference asset (definition page, benchmark, methodology, glossary, comparison hub) designed to beat existing top-cited sources. This skill analyzes AI Visibility source data, reverse-engineers what makes top-cited pages authoritative, and produces a superior source asset — then pushes it to CMS as a draft. Trigger on any mention of "sources", "third-party citations", "authoritative content", "definitional pages", or "outrank".

amplitude/builder-skills

Instrument a Node/TypeScript MCP server with Amplitude's @amplitude/mcp-analytics SDK so tool calls, sessions, and rationale are tracked as Amplitude events. Use this skill whenever the user wants to add Amplitude analytics to their MCP server, mentions "MCP Analytics", "@amplitude/mcp-analytics", "instrument my MCP server", "track MCP tool calls", "add rationale to my MCP tools", or wants agent traffic (Claude, Cursor, ChatGPT) attributed back to Amplitude. Also use for adding UTM tagging to MCP-returned links, or for troubleshooting identity/user_id mismatches between MCP events and web/mobile Amplitude data.

amplitude/builder-skills

Instruments a pull request with Amplitude analytics that conform to the project's existing taxonomy. Reads the tracking plan via the Amplitude MCP server (events, properties, naming conventions), analyzes the PR diff to find the few user actions genuinely worth tracking, detects the codebase's SDK and tracking patterns, and adds instrumentation that matches both. Optionally (opt-in) stages new events and properties on an Amplitude tracking-plan branch for data-governance review. Use when asked to "instrument this PR", "add analytics to this change", "add tracking", "add Amplitude events", "instrument this feature", or "what should I track here".

amplitude/builder-skills

Diagnoses product health by cross-referencing Amplitude analytics (dashboards, charts, funnels, feedback, AI agent analytics), optionally Datadog (errors, latency, stack traces), and optionally Slack (qualitative feedback, bug reports, feature requests). Identifies what's broken, what's working, and what to do about it — with root causes, not just symptoms. Use when asked to "diagnose my product", "what's going on", "product health check", "what's broken", "where are users struggling", "give me a product diagnosis", or "what should I focus on".

amplitude/builder-skills

Summarizes B2B account health by analyzing usage patterns, engagement trends, risk signals, and expansion opportunities. Use for customer success reviews, renewal preparation, QBRs, or account prioritization.

Related Skills