Communitygithub.com

austin-starks/Public-Portfolio-Challenge

Run a multi-family strategy bakeoff — the SEARCH→CERTIFY funnel that screens many candidate mechanisms down to a certified deploy winner without letting the cheap search layer issue a verdict. Use when replaying the Episode 10 bakeoff, when exploring several distinct strategy families before certifying, when deciding whether \"no deployable winner\" is even a legal conclusion, or when building the per-family certification ledger. Enforces verdict-integrity: only certification can end a campaign.

O que é Public-Portfolio-Challenge?

Public-Portfolio-Challenge is a Claude Code agent skill that run a multi-family strategy bakeoff — the SEARCH→CERTIFY funnel that screens many candidate mechanisms down to a certified deploy winner without letting the cheap search layer issue a verdict. Use when replaying the Episode 10 bakeoff, when exploring several distinct strategy families before certifying, when deciding whether \"no deployable winner\" is even a legal conclusion, or when building the per-family certification ledger. Enforces verdict-integrity: only certification can end a campaign.

Funciona com~Claude Code~Codex CLI~Cursor
npx skills add https://github.com/austin-starks/Public-Portfolio-Challenge/tree/HEAD/skills/strategy-bakeoff

Perguntar na sua IA favorita

Abre um novo chat com esta habilidade de agente já pré-carregada.

Documentação

Multi-Family Strategy Bakeoff

A two-layer SEARCH → CERTIFY funnel with a mandated minimum body of work. The central rule: a verdict is FORBIDDEN from the search layer. Only walk-forward certification (walk-forward-oos) can end the campaign.

Two tools, two jobs

  • SEARCH layer (cheap): backtest_portfolio, create_portfolio_variant, optimize_portfolio, systematic_sweep. It may only KILL or PROMOTE an idea into certification. It outputs "a ranked set of SEEDS" and may never issue a verdict.
  • CERTIFY layer (workhorse): walk-forward sweep certification on promoted seeds. Only this layer can conclude PASS / FAIL or "no deployable winner."

systematic_sweep over the deployment × structure × DTE × delta grid is mandatory in the search layer before any neighborhood is declared dead.

The mandated minimum body

  • ≥ 3 distinct mechanism families certified before any "no winner" verdict. Genuinely different mechanisms, not three tweaks of one family. Family menu (pick ≥3): (a) momentum long calls/verticals; (b) regime-switch long/flat; (c) defined-credit / short-premium (earns in flat/down tape); (d) calendar / diagonal time-spreads; (e) protective-collar or put-spread-financed longs.
  • Up to 3 certification attempts per family (re-sweep from a CLEAN seed with refined gene_intents, deeper grid, or wider validation) before a family is "exhausted."
  • Realistic total: ~3–9 sweep certifications, plus the incumbent certified as a seed.
  • Each sweep non-trivial: ≥3 gene_intents axes × ≥3 values each (see sweep-reoptimization for the exact field mappings). A 1-axis or 2-value sweep does NOT qualify.

Per-family certification ledger (required)

Log this table in the run log — a family with no ledger row was NOT explored:

| Family | Seed id | get_sweep_surface ✓ | gene_intents axes | Study id | crossFoldRobustSelection key | Assembled book id | Per-fold OOS gate result | PROMOTE / KILL (reason) |

Per-seed certification loop

get_sweep_surface → run_walk_forward_study (preview_only first) → read aggregate.crossFoldRobustSelection → assemble the deploy-shape book → gate on the assembled object → write the ledger row. Certification floors (certification:true): participation ≥ 0.35, distinct names ≥ 9, validation return ≥ 0, validation Sortino ≥ 0.5.

Verdict-integrity rules (the campaign's central discipline)

  • Zero optimizing certifications = INVALID, not "an honest no-deploy." Issuing any verdict — especially "no deployable winner" / "structural infeasibility" — without optimizing certifications having run on candidate seeds makes the campaign INVALID. (A prior run declared infeasibility having run ZERO optimizing certifications.)
  • No infeasibility claim from N hand-built points — must cite systematic_sweep + the ≥3-family body and name the SINGLE binding gate per family with study-id evidence.
  • One counterexample in your own data forbids a blanket infeasibility claim — localize the binding constraint, don't generalize.
  • Binding-constraint attribution: every kill names the single failed gate + certified evidence. "It generally underperformed" is not a kill reason.
  • One mechanism change per candidate — so a kill/promote attributes cleanly.
  • Remedy-coherence: a proposed remedy must not contradict the killed candidate's measured state (don't propose "raise deployment" for a book that failed under the posture cap; don't propose "add convexity" for one that failed on drawdown).

See options-structure-rules for the affordability-ladder kill path and known losers, and lockbox-holdout for the baselines (A/B/C) and the final single-touch confirmation.

Individual skills in this repo

This repo contains 11 individual skills — each has its own dedicated page.

austin-starks/Public-Portfolio-Challenge

Build alternative-data custom indicators on NexusTrade (Reddit/WSB mentions, congressional disclosures, insider filings, news-flow) and wire them into a certified book as a rank/tilt/filter signal. Use when adding alt-data to a strategy, building a CustomIndicator via compute sessions, auditing lookahead safety or per-ticker coverage, or deciding whether a data series is dense enough to drive a rank. Covers the compute-session workflow, the sparse-series rule, and source recipes (see references/). Invoke with the NexusTrade MCP connected.

austin-starks/Public-Portfolio-Challenge

Measure a NexusTrade options book's TRUE participation at a fixed cold-start capital base, defeating the compounded-NAV breadth illusion. Use whenever a book's headline \"holds N/21 names\" needs validating, when a live book collapses to a single name (the OSCR trap), when checking whether a SelectTop / per-name-allocation / total-budget change degraded simultaneous participation, or when a certification needs a breadth gate. Uses audit_backtest_breadth at held-fixed $25k.

austin-starks/Public-Portfolio-Challenge

The \"loudly declare\" protocol for platform, engine, or data bugs during a NexusTrade certification campaign — bugs are a first-class deliverable, not something to route around. Use whenever a NexusTrade tool errors, hangs, or returns numbers that contradict the config; whenever you're tempted to write \"probably an engine quirk\"; or when a discovered bug invalidates prior backtests and results must be quarantined. Includes the bug hand-off doc template (see references/BUG_TEMPLATE.md).

austin-starks/Public-Portfolio-Challenge

The GATED deploy + cleanup flow for a NexusTrade live book — clone the finalist, preview a delta reconcile, stage UNAPPROVED orders, verify fills. Use ONLY after the human explicitly says \"deploy + clean up\" and names a finalist. Covers the clone-before-reconcile ordering, the single-tick reconcile expectation, stale-pending-order hygiene, the signal-freshness gate, and why you can never approve orders yourself. Nothing here runs during certification.

austin-starks/Public-Portfolio-Challenge

The mandatory pre-flight contract checks (Stage S0) that must pass before trusting NexusTrade's certification engine for any strategy work. Use at the start of a bakeoff or certification campaign to prove the walk-forward engine runs end-to-end, windows don't leak, fold winners persist, and the engine doesn't fabricate values for dead/not-yet-listed names. Each check has a STOP-and-report failure mode; these runs do NOT count toward any certification minimum.

austin-starks/Public-Portfolio-Challenge

The single-touch lockbox — a final anti-overfitting holdout run once, after design freeze, on a window held out from every fold, sweep, and search. Use when finalizing a bakeoff winner before deploy, setting up the A/B/C baselines as OOS bars, or running the S1.5 gate-coherence auto-relax. Covers why looking at the lockbox twice burns it, the lockbox pass conditions, and the three baselines. The lockbox is distinct from the walk-forward OOS folds.

austin-starks/Public-Portfolio-Challenge

The hard structural constraints for the Public Portfolio Challenge momentum-LEAP options book — the spread-shape rule, the take-profit convexity-cap footgun, the affordability ladder, and the known losers not to re-test. Use whenever building or auditing an options strategy structure, checking spread-shape compliance before a certification, choosing DTE/strike rungs, or deciding whether a proposed structure change is even allowed. A violation is an automatic certification FAIL.

austin-starks/Public-Portfolio-Challenge

Orchestrate an out-of-sample certification of a NexusTrade trading strategy or live book — the master discipline behind the Public Portfolio Challenge. Use whenever you must decide PASS/FAIL on whether a portfolio holds up out of sample before deploying real money, replaying the Episode 10 runbooks, or running a \"certify my book\" / \"prove it out of sample\" / \"re-certify the fix\" task with the NexusTrade MCP connected. Pulls in walk-forward-oos, breadth-audit, sweep-reoptimization, options-structure-rules, bug-protocol, and deploy-gate.

austin-starks/Public-Portfolio-Challenge

The single entry point that executes a Public Portfolio Challenge episode or addendum runbook end-to-end, delegating each stage to the functional skills. Use when asked to run/execute/replay Episode 10, its bakeoff, or its addendum with the NexusTrade MCP connected. Reads the target runbook, pins its real artifacts (IDs, the incumbent bar), sequences the stages, and stops at the gated deploy. It orchestrates; the functional skills do the work.

austin-starks/Public-Portfolio-Challenge

Re-optimize a NexusTrade strategy with a walk-forward SWEEP and label parameter provenance — the discipline that prevents deploying inherited knobs. Use whenever a structural change (sizing, rung depth, universe membership, DTE family, adding a rank signal) forces a re-sweep, when authoring gene_intents from get_sweep_surface, when choosing sweep over GA for a deploy cert, or when selecting the cross-fold-robust winner instead of the per-fold argmax. Invoke with the NexusTrade MCP connected.

austin-starks/Public-Portfolio-Challenge

Run and read a NexusTrade walk-forward out-of-sample study — the certification engine behind the Public Portfolio Challenge. Use when certifying a fixed portfolio (backtest_only) or re-optimizing one (sweep), when setting fold_count / anchored / validation / embargo params, when monitoring a run_walk_forward_study to completion, when reading per-fold OOS returns/Sortino/drawdown, or when a variant's fold calendar needs calendar-alignment against a base control. Invoke with the NexusTrade MCP connected.

Habilidades Relacionadas