Mark v5.0/v5.1 SUPERSEDED (error-detection thesis failed at -0.40 delta); v4.1 is canonical. Cut v4.1 proposal with 12 version strings bumped. Moonshot to Mistral across production seats; production worker de-kimi'd 2026-08-18. Data-retention posture corrected 7/9 to 8/9 no-training (DeepSeek sole exception). Reconciled COGS with measured Mistral spend. Committed deployed v4.0 content and research docs to resolve the repo/live fork.
77 lines
7.2 KiB
HTML
77 lines
7.2 KiB
HTML
|
|
<!-- ====== 4. THE 14 ERRORS ====== -->
|
|
<section id="errors">
|
|
<h2><span class="n">04</span>The 14 Errors: Every One Named</h2>
|
|
<p class="lead">
|
|
This is the product. Fourteen material errors the 8-judge panel caught that the solo baseline
|
|
missed or underweighted, grouped by failure class. Each is a real finding from the 2026-08-12
|
|
run against a real document.
|
|
</p>
|
|
|
|
<div class="errgrid">
|
|
<div class="errcard"><div class="n">3</div><div class="t">Revenue arithmetic errors</div><div class="d">Headline numbers that contradict the proposal's own inputs.</div></div>
|
|
<div class="errcard"><div class="n">4</div><div class="t">Competitive mispositionings</div><div class="d">Named, funded, shipping incumbents the document treated as absent.</div></div>
|
|
<div class="errcard"><div class="n">3</div><div class="t">Execution infeasibilities</div><div class="d">Build plans that cannot be delivered at the stated budget or timeline.</div></div>
|
|
<div class="errcard"><div class="n">2</div><div class="t">Legal compliance blockers</div><div class="d">Registration and privacy obligations that gate launch entirely.</div></div>
|
|
<div class="errcard"><div class="n">2</div><div class="t">Team capacity impossibilities</div><div class="d">Founder hour budgets that exceed the hours available.</div></div>
|
|
<div class="errcard" style="border-left-color:var(--warn)"><div class="n">14</div><div class="t">Total, across 3 documents</div><div class="d">Mean 4.7 material errors per proposal reviewed.</div></div>
|
|
</div>
|
|
|
|
<h3>RFP Tank v1.0 · panel 4.40 vs solo 4.93</h3>
|
|
<table>
|
|
<thead><tr><th style="width:180px">Class</th><th>Error the panel caught</th><th style="width:150px">Caught by</th></tr></thead>
|
|
<tbody>
|
|
<tr><td><span class="tag tag-red">Revenue</span></td><td>Three mutually inconsistent Year-1 revenue figures inside one document: $1.2M, $1.361M, and $372K. Plus a 22% MRR ramp inconsistency the narrative never reconciles.</td><td>Financial Integrity</td></tr>
|
|
<tr><td><span class="tag tag-red">Competitive</span></td><td>CLEATUS is a real, funded competitor at $4M seed with public product-led pricing of $39 to $250/mo, occupying the identical quadrant. The proposal does not name it. The Band A generalist seat actually cited CLEATUS pricing as a positive signal.</td><td>Market Reality</td></tr>
|
|
<tr><td><span class="tag tag-red">Competitive</span></td><td>GovEagle pricing referenced at a 15x inconsistency against the proposal's own comparison table.</td><td>Market Reality</td></tr>
|
|
<tr><td><span class="tag tag-red">Execution</span></td><td>Five of seven features marked TO BUILD at HIGH effort. The real-time Compliance Copilot alone needs 2 to 3 developers for 8 to 12 weeks. The plan allocates 4 weeks, solo.</td><td>Execution Feasibility</td></tr>
|
|
<tr><td><span class="tag tag-red">Team</span></td><td>A solo founder shipping a 7-feature AI SaaS in 10 weeks, with hiring contingent on revenue that requires the product to already exist. A closed loop with no entry point.</td><td>Team / Founder</td></tr>
|
|
<tr><td><span class="tag tag-red">Legal</span></td><td>No privacy policy and no terms of service, against FAR and CUI exposure, with ITAR implications on German-hosted infrastructure.</td><td>Legal / Regulatory</td></tr>
|
|
</tbody>
|
|
</table>
|
|
<p class="table-note">
|
|
Panel verdict: NO GO. Estimated rework 40+ hours. Recommendation is to cut scope to two features,
|
|
extend to 20 weeks, hire a second developer before month one, rebuild the financial model, and
|
|
address CLEATUS directly.
|
|
</p>
|
|
|
|
<h3>VentureBuilt v2 · panel 6.14 vs solo 6.10</h3>
|
|
<table>
|
|
<thead><tr><th style="width:180px">Class</th><th>Error the panel caught</th><th style="width:150px">Caught by</th></tr></thead>
|
|
<tbody>
|
|
<tr><td><span class="tag tag-red">Execution</span></td><td>Contractor budget broken by a factor of 4 to 7. The stated $1,500/mo implies $11 to $22 per hour against a market rate of $75 to $100. At real rates that budget buys 105 to 140 hours and leaves roughly 800 hours on the founder.</td><td>Execution Feasibility</td></tr>
|
|
<tr><td><span class="tag tag-red">Revenue</span></td><td>Year 2 stated on a run-rate basis rather than recognized revenue. Restated correctly, the healthy scenario loses roughly $11K to $18K.</td><td>Financial Integrity</td></tr>
|
|
<tr><td><span class="tag tag-red">Competitive</span></td><td>The uniqueness claim is contradicted by shipping products. LivePlan Plan Review and IdeaProof already occupy the space.</td><td>Market Reality</td></tr>
|
|
<tr><td><span class="tag tag-red">Team</span></td><td>37 engagements plus 950 hours plus an MSP day job. The three commitments cannot coexist in one calendar.</td><td>Team / Founder</td></tr>
|
|
</tbody>
|
|
</table>
|
|
<p class="table-note">
|
|
Panel verdict: CONDITIONAL GO with six named conditions. Estimated rework 15 to 20 hours. This is
|
|
the case that most clearly shows the value: the panel mean (6.14) and the solo mean (6.10) are
|
|
statistically tied, so on score alone the panel added nothing. What it added was a bimodal split,
|
|
Financial 8.2 against Execution 2.8, and the four errors above.
|
|
</p>
|
|
|
|
<h3>CartMySupply · panel 4.29 vs solo 5.00</h3>
|
|
<table>
|
|
<thead><tr><th style="width:180px">Class</th><th>Error the panel caught</th><th style="width:150px">Caught by</th></tr></thead>
|
|
<tbody>
|
|
<tr><td><span class="tag tag-red">Revenue</span></td><td>The $2.7M headline is wrong by 7x to 10x against the proposal's own inputs, which compute to $269K. Stripe fees understated by roughly $11K per year. CAC absent entirely.</td><td>Financial Integrity</td></tr>
|
|
<tr><td><span class="tag tag-red">Competitive</span></td><td>TeacherLists already solves the identical problem, free, across 2 million lists. Target ships native School List Assist. No technical moat is claimed or demonstrable.</td><td>Market Reality</td></tr>
|
|
<tr><td><span class="tag tag-red">Execution</span></td><td>Amazon PA-API 5 removed Cart API support. Target has no self-serve multi-item cart API. Walmart requires separate catalog matching. The core mechanic of the product does not have a supported integration path at any of the three named retailers.</td><td>Execution Feasibility</td></tr>
|
|
<tr><td><span class="tag tag-red">Legal</span></td><td>Charitable solicitation registration required in 40+ states at $30K to $75K, plus COPPA exposure and FTC penalty risk. Compliance cost of $60K to $150K exceeds projected Year-1 revenue of $3K to $14K by an order of magnitude.</td><td>Legal / Regulatory</td></tr>
|
|
</tbody>
|
|
</table>
|
|
<p class="table-note">
|
|
Panel verdict: NO GO, unanimous, zero outliers, tightest spread of the three. Estimated rework
|
|
60+ hours. The build estimate of 116 hours was independently judged 4x to 10x too low.
|
|
</p>
|
|
|
|
<div class="callout callout--win">
|
|
<strong>Read the legal row again.</strong> A $60K to $150K registration obligation against
|
|
$3K to $14K of projected revenue is not a scoring nuance. It is the difference between a
|
|
business and a fine. A solo model reading the same document returned a 5.0 and did not raise it.
|
|
That single finding is worth more than every point of score elevation the old thesis promised.
|
|
</div>
|
|
</section>
|