Files
waggle-os/docs/decisions/2026-04-25-pm-pre-fill-decision-matrix-recommendations.md
Oleg Maslov 0c3e2ead3b
Some checks failed
Installer Smoke / installer-smoke (push) Has been cancelled
moving
2026-09-02 10:10:29 +02:00

330 lines
17 KiB
Markdown
Raw Blame History

This file contains ambiguous Unicode characters
This file contains Unicode characters that might be confused with other characters. If you think that this is intentional, you can safely ignore this warning. Use the Escape button to reveal them.
# PM Pre-Fill Recommendations — Launch Gate Decision Matrix
**Date**: 2026-04-25
**Author**: claude-opus-4-7 (PM Cowork)
**Companion**: `decisions/2026-04-25-launch-gate-reframe-decision-matrix.md`
**Purpose**: Reduce Marko-vov sutra-jutarnji decision time from 30-45 min na 10-15 min review-mode by providing PM-recommended verdict + rationale + confidence per cell. Marko accepts default ili overrides per cell.
**Confidence scale**:
- **HIGH** — strategic logic + repo/research data alignment + few legitimate alternatives → expect Marko ratification
- **MEDIUM** — defensible default but legitimate alternatives exist → expect ~70% Marko ratification, 30% override
- **LOW** — Marko-personal-preference-dependent (capital structure, hiring posture, public persona); PM defers but recommends conservative stance
**24 cells = 8 dimensions × 3 scenarios.**
---
## Dimension 1 — Claim narrative
### Scenario PASS
**PM RECOMMEND**: "Waggle hit [LOCOMO_SCORE]% on LoCoMo, beating Mem0's [BASELINE_REF]% reference. Built on hive-mind, the local-first Apache 2.0 cognitive substrate that follows you across Claude Code, Cursor, Hermes, and 3 more AI IDEs."
Add cross-IDE shim portfolio kao supporting evidence — strengthens narrative beyond benchmark singleton.
**Rationale**: Beat-Mem0 hook je strongest possible narrative. Cross-IDE shim mention extends story past one-day cycle. Honest caveats (n=400, agentic κ, Qwen subject) mandatory u body, not headline.
**Confidence**: HIGH
### Scenario PARTIAL
**PM RECOMMEND**: "First Apache 2.0 local-first memory substrate to publish a comparable LoCoMo result. We're [DELTA]pp behind cloud-shaped Mem0; that gap is the cost of local-first sovereignty. One memory file, every AI IDE you use."
Lead sa local-first-quadrant clean-win framing + cross-IDE positioning, NOT benchmark headline.
**Rationale**: PARTIAL band requires reframe from "we hit SOTA" to "we lead in our category." Honesty about gap signals integrity (Zep/Mem0 dispute precedent shows overstated claims unwind). Cross-IDE is the productized differentiator that doesn't depend on benchmark winning.
**Confidence**: HIGH
### Scenario FAIL
**PM RECOMMEND**: NOT a benchmark-led narrative. "hive-mind: the Apache 2.0 cognitive substrate that follows you across 6 AI IDEs. One memory file. Zero cloud. Honest evaluation included."
Methodology blog post (separate from launch announcement) discloses LoCoMo number neutral framing: "our pre-registered evaluation produced [LOCOMO_SCORE]%; we don't lead with this because it's not SOTA."
**Rationale**: Pre-registration honor mandatory — bands LOCKED, no post-hoc reframe attempt to rescue. Pivot to product narrative (cross-IDE shim portfolio) salvages launch momentum without dishonesty about benchmark.
**Confidence**: HIGH
---
## Dimension 2 — Launch coupling sequence
### Scenario PASS
**PM RECOMMEND**: Coupled simultaneous Day 0 launch — hive-mind core + 3 MVP shims + Waggle landing live + paid tiers Stripe checkout active. All press/community/email channels synchronized within 30-min window.
**Rationale**: SOTA momentum is renewable per asset (blog, LinkedIn, Twitter, HN, Reddit) but each asset only fires once. Coupling maximizes day-0 attention concentration.
**Confidence**: HIGH
### Scenario PARTIAL
**PM RECOMMEND**: Decoupled — hive-mind + 3 shims first (week 0), Waggle 2-4 weeks later. Build infrastructure adoption signal, then commercial product launches as "the fully featured cousin."
**Rationale**: PARTIAL benchmark + Waggle tier launch sa PARTIAL banner = competing for attention with weaker narrative; better to claim local-first infra category first, then commercial product on pre-built audience.
**Confidence**: MEDIUM (alternative: still couple to capture brand-narrative momentum)
### Scenario FAIL
**PM RECOMMEND**: Pivot launch — hive-mind + 3 shims as primary product day 0. Waggle delayed 4-8 weeks (revisit benchmark methodology + ship LongMemEval first to validate alternative-axis claim before re-attempt at Waggle launch).
**Rationale**: Without SOTA claim, Waggle u crowded "agent IDE" space lacks differentiator. Defer commercial launch until benchmark improvement OR adoption metrics from shim portfolio justify revisit.
**Confidence**: HIGH
---
## Dimension 3 — Pricing posture
### Scenario PASS
**PM RECOMMEND**: Hold ratified pricing per `project_locked_decisions`. Free / Pro $19/mo / Teams $49/seat/mo. SOTA premium justifies above-Mem0-Starter pricing. No discount, no early-bird.
**Rationale**: Pricing power is real with NEW_SOTA banner. Discounting signals weakness; hold position.
**Confidence**: HIGH
### Scenario PARTIAL
**PM RECOMMEND**: Hold pricing same as PASS, but emphasize Free tier value first. Pro/Teams targeted at compliance-driven buyers (local-first = mandate not preference for them).
**Rationale**: PARTIAL doesn't unlock SOTA premium pricing power, but doesn't crash it either. Free tier as adoption funnel; Pro/Teams targeted properly. No discount needed.
**Confidence**: MEDIUM (alternative: 6-month "early supporter" Pro $9/mo to drive adoption, revisit at $19 after retention proven)
### Scenario FAIL
**PM RECOMMEND**: Pricing pivot — Free tier extended (10 workspaces vs 5, all skills marketplace tiers free). Pro $9/mo introductory price for first 1,000 users (12-month commitment). Teams holds at $49 (enterprise sale not benchmark-dependent).
**Rationale**: Without SOTA, must compete on value-density and ergonomics. Aggressive Free tier seeds adoption; introductory Pro pricing hooks early-supporter cohort; Teams unchanged because enterprise compliance buyers weighed against KVARK already.
**Confidence**: MEDIUM (alternative: hold full pricing, accept slower adoption ramp)
---
## Dimension 4 — Audience targeting
### Scenario PASS
**PM RECOMMEND**: Three primary audiences day 0, prioritized:
1. **Privacy-conscious devs** (highest converting) — Cursor + Claude Code + Hermes audiences via shim distribution
2. **AI/ML researchers + tool builders** — methodology + pre-reg manifest deep-link
3. **Compliance-mandated enterprises** (banks, healthcare, EU regulated) — KVARK lead capture forms
**Rationale**: Privacy-conscious devs convert fastest sa local-first message + working shims. Researchers validate via methodology. Enterprise sale follows technical credibility.
**Confidence**: HIGH
### Scenario PARTIAL
**PM RECOMMEND**: Two primary audiences day 0:
1. **Privacy-conscious devs** (most receptive to local-first message)
2. **Compliance-mandated enterprises** (KVARK lead via outbound, not inbound campaign)
ML research audience deferred to second blog post 2-4 weeks later.
**Rationale**: PARTIAL doesn't strongly attract ML researcher audience (they index on SOTA numbers). Concentrate on dev + compliance buyers where local-first is clear value.
**Confidence**: HIGH
### Scenario FAIL
**PM RECOMMEND**: Single primary audience day 0 — privacy-conscious devs using AI agent IDEs. Lead with cross-IDE shim portfolio. Deprioritize ML research + enterprise audiences until later (sequel benchmark or adoption metrics inform).
**Rationale**: FAIL eliminates ML researcher audience entirely (dishonest to court). Enterprise sales typically need >12 month sales cycles + benchmarks for procurement; postpone until benchmark improvement.
**Confidence**: HIGH
---
## Dimension 5 — Press / PR strategy
### Scenario PASS
**PM RECOMMEND**: Synchronized launch with 48h press embargo. Pre-brief: TechCrunch, The Verge, Ars Technica, IEEE Spectrum, AI newsletters (TLDR AI, Ben's Bites, AI Tidbits). Day 0: simultaneous press + community + product release. Hacker News + Reddit submissions by Marko personally.
**Rationale**: NEW_SOTA banner deserves earned-media leverage. Press embargo aligns coverage to launch moment. Community submissions by founder accumulate organic goodwill (vs paid PR which devs distrust).
**Confidence**: MEDIUM (alternative: skip paid PR, community-only — risk: stories told without context if random journalist picks up)
### Scenario PARTIAL
**PM RECOMMEND**: Soft press strategy — community-first (HN, Reddit, Twitter, Discord). No paid PR, no embargo. Reach out to specific researchers/orgs (DAIR, Anthropic, Nous Research) post-launch.
**Rationale**: PARTIAL doesn't justify embargo machinery. Community-led discovery scales naturally for honest local-first claim; targeted research outreach builds methodology credibility.
**Confidence**: HIGH
### Scenario FAIL
**PM RECOMMEND**: No press push initially. Quiet launch via GitHub + Twitter. Methodology blog post (separate from launch) targets researchers 2-4 weeks later.
**Rationale**: FAIL + press attention = brand damage risk if "we launched memory thing that didn't beat SOTA" becomes the headline. Quiet launch lets shim portfolio adoption signal product-market fit before public scrutiny.
**Confidence**: HIGH
---
## Dimension 6 — Hires plan
### Scenario PASS
**PM RECOMMEND**: Aggressive — `jobs.egzakta.com` active day 0 sa 3-5 roles:
- **DevRel lead** (developer engagement, conference speaking, content)
- **Senior eng OSS maintenance** (hive-mind core + shim portfolio)
- **GTM / growth marketing lead** (paid acquisition, conversion optimization)
- **Enterprise sales lead** (KVARK pilots)
- **Optional: ML research engineer** (next-benchmark sequence)
Prominent "We're hiring" CTAs u blog post + LinkedIn.
**Rationale**: SOTA + funded competitors precedent (Mem0 $24M, Cognee $7.5M, Supermemory $3M) means org capacity is the bottleneck within 6-12 months; hire signal capitalizes attention moment.
**Confidence**: MEDIUM (LOW dimension because Marko personal preference re hiring pace, especially given Egzakta cash flow already 4.5M EBITDA = no immediate hire pressure; alternative: post 1-2 roles, hold rest until adoption metrics)
### Scenario PARTIAL
**PM RECOMMEND**: Moderate — 1-2 roles: OSS maintainer + DevRel. Hold enterprise sales hire until adoption metrics inform. KVARK enterprise via existing Egzakta consulting team initially.
**Rationale**: PARTIAL doesn't require aggressive scaling; OSS maintainer protects shim portfolio quality; DevRel covers community. Enterprise sale wait-and-see.
**Confidence**: HIGH
### Scenario FAIL
**PM RECOMMEND**: No hiring announcements at launch. Hold all roles until 4-8 weeks of adoption metrics inform org capacity needs.
**Rationale**: Hiring announcement on FAIL launch = "they're scaling without traction" perception damage. Hold + execute on existing team.
**Confidence**: HIGH
---
## Dimension 7 — Investor narrative
### Scenario PASS
**PM RECOMMEND**: Open conversations with seed/A funds aktivno week 1-2 post-launch.
Pitch deck v1 messaging: "First Apache 2.0 local-first cognitive substrate that beat cloud SOTA on LoCoMo. Cross-IDE shim portfolio in 6 AI agent platforms. Egzakta cash flow signal as bootstrap credential."
Targets: Initialized, BoxGroup, Founders Fund (Peter Thiel local-first sympathies), NEA (memory-tech adjacent), DCVC (deep tech), Acrew (developer tools focus).
Aim: $5-10M seed at $20-40M post-money valuation.
**Rationale**: SOTA + working shims + revenue path (Waggle paid tiers + KVARK enterprise) = strong seed narrative. Egzakta financial backbone reduces fund-raise pressure (Marko negotiates from strength).
**Confidence**: LOW (Marko-personal-preference: capital structure, dilution tolerance, growth speed vs control trade-off; PM defers heavily but provides default)
### Scenario PARTIAL
**PM RECOMMEND**: Silent investor conversations only — pitch deck v1.5 emphasizes "best local-first published result" + commercial trajectory + Egzakta financials.
Targets: 5-10 most aligned funds (focused on infra + dev tools + privacy-tech), no broad outreach.
Aim: $3-7M seed at $15-25M post-money.
**Rationale**: PARTIAL doesn't carry hype premium; silent conversations preserve optionality without waste-of-time discovery.
**Confidence**: LOW (same Marko-personal preference dependency)
### Scenario FAIL
**PM RECOMMEND**: Defer fundraise 6-12 months. Continue bootstrap from Egzakta. Use period to:
1. Ship LongMemEval + BEAM 1M + 10M results + at least one Claude config benchmark (per `research/00-SYNTHESIS.md` §3 missing-gap)
2. Drive shim portfolio adoption (target: 10K combined installs across 6 shims)
3. Validate enterprise sale via 3-5 KVARK pilots through Egzakta network
Re-engage investors in 6-12 months sa stronger numbers + adoption signal.
**Rationale**: FAIL fund-raise = forced down-round risk; defer until story ripens. Egzakta cash flow provides runway to repair narrative without external pressure.
**Confidence**: HIGH
---
## Dimension 8 — KVARK enterprise pivot timing
### Scenario PASS
**PM RECOMMEND**: KVARK pilots active immediately — Marko's network targets 5-10 named enterprise prospects week 1-4. EU compliance-driven, sovereign on-prem deployment value prop. Pricing custom per deployment ($50K-$250K typical).
Lean on Egzakta ~200 staff consulting backbone for delivery + advisory bundle.
**Rationale**: SOTA banner gives sales team explicit credibility; compliance + sovereignty are pre-existing enterprise demand drivers; Egzakta delivery capability eliminates capacity concern.
**Confidence**: HIGH
### Scenario PARTIAL
**PM RECOMMEND**: KVARK pilots active 4-8 weeks post-Waggle-launch. Build case from Waggle + shim adoption metrics first; pilot conversations open with "the local-first stack you've heard about, now sovereign-deployed for your hardware."
**Rationale**: Need adoption signal to land enterprise conversations confidently u PARTIAL band; 4-8 weeks builds enough story.
**Confidence**: HIGH
### Scenario FAIL
**PM RECOMMEND**: KVARK becomes the lead commercial track. Day 0 launch: hive-mind + shims (OSS) + KVARK enterprise pilots opened (consulting+software bundle).
Aim 3-5 paid pilots within 90 days. Pricing per deployment $100K-$300K bundled with Egzakta advisory + LM TEK H200 hardware integration.
Waggle commercial launch decoupled 4-8 weeks (informed by KVARK pilot signal).
**Rationale**: FAIL eliminates Waggle SOTA narrative as primary acquisition; KVARK enterprise compliance-mandated buyers don't depend on public benchmark — they buy for sovereignty + GDPR + EU AI Act + data residency. Egzakta consulting provides high-touch delivery moat that Waggle commercial channel can't replicate.
**Confidence**: HIGH
---
## Aggregate confidence summary
| Dimension | PASS conf | PARTIAL conf | FAIL conf |
|---|---|---|---|
| 1. Claim narrative | HIGH | HIGH | HIGH |
| 2. Launch coupling | HIGH | MEDIUM | HIGH |
| 3. Pricing posture | HIGH | MEDIUM | MEDIUM |
| 4. Audience targeting | HIGH | HIGH | HIGH |
| 5. Press/PR strategy | MEDIUM | HIGH | HIGH |
| 6. Hires plan | MEDIUM | HIGH | HIGH |
| 7. Investor narrative | LOW | LOW | HIGH |
| 8. KVARK timing | HIGH | HIGH | HIGH |
**Translation for Marko**:
- 18/24 cells HIGH confidence → likely accept PM defaults
- 4/24 cells MEDIUM → review and consider alternative noted
- 2/24 cells LOW (investor narrative under PASS/PARTIAL — depends on Marko-personal capital structure preference) → PM defers heavily, expect override
Sutra-jutarnji ratification time: review HIGH cells (~30s each = 9 min), MEDIUM cells (~60s each = 4 min), LOW cells (~3 min each = 6 min). **Total 19-20 min** vs original 30-45 min estimate.
---
## Marko-side override mechanism
Per cell, Marko može:
-**Accept** — default verdict applies
- ✏️ **Override** — write 1-line explicit override + rationale
- 🤔 **Defer** — punt na separate decision moment, e.g., investor narrative may be 30-day decision window post-launch
Override format example:
```
Dimension 7 (Investor narrative), Scenario PASS:
✏️ OVERRIDE: defer fund-raise indefinitely; bootstrap permanently.
Rationale: Egzakta cash flow sufficient for org plan; equity dilution avoided; control prioritized. Re-evaluate u 24 months.
```
---
## Cross-cutting decisions (locked, no pre-fill needed)
Already scenario-independent per Decision Matrix §10:
- Honest framing always
- Apache 2.0 commitment locked
- hive-mind core scope locked (substrate vs runtime split per `01-architecture.md`)
- EU AI Act + GDPR positioning preserved
- Domain locked (waggle-os.ai primary)
These don't need PM pre-fill ratification — they're already LOCKED.
---
## What this enables
When Phase 2 final halt ping arrives:
1. Marko reads halt ping (5 min)
2. Marko classifies scenario per pre-registered bands (1 min)
3. Marko opens this document, navigates to relevant scenario column across 8 dimensions
4. Marko ratifies HIGH cells (9 min), reviews MEDIUM (4 min), decides LOW (6 min)
5. Marko PM-RATIFY-V6-N400-COMPLETE + signs off Gate D exit
6. Triggered downstream sequences per scenario:
- PASS → paste apps/www brief u CC-1, populate launch comms placeholders sa actual numbers, schedule publish
- PARTIAL → softer comms variant, signal decoupled sequence
- FAIL → pivot strategy ratify, defer Waggle 4-8 weeks, KVARK lead
Total Marko time-to-decision: **~25 min** sa pre-fills vs ~45-60 min raw matrix-only.
---
## Authorization
PM (claude-opus-4-7) authoring 2026-04-25 dok agentic cell radi. Marko ratifies post-final-halt-ping. PM defaults are non-binding suggestions; Marko-vov decision authority preserved on every cell.