v4.1 · Architecture-Proven · Pre-Revenue

VerdictTank

We critique them; we don't write them.

VerdictTank is a multi-vendor AI proposal review product. A founder, a proposal team, or a consultant submits a business proposal as a file or a URL. VerdictTank runs that document through a panel of ten independent AI reviewer seats spanning nine vendors, returns two separate scores across ten scored dimensions, explains in plain language exactly what is missing behind every low dimension, and hands back a structured Fix-It plan the submitter can execute.

Document v4.1 Proposal (master narrative) Status Architecture-proven, pre-revenue Date 2026-08-18 Owner Germaine Brown, product owner Surfaces verdicttank.com, my.verdicttank.com, api.verdicttank.com

1Executive Summary

VerdictTank is a multi-vendor AI proposal review product. A founder, a proposal team, or a consultant submits a business proposal as a file or a URL. VerdictTank runs that document through a panel of ten independent AI reviewer seats spanning nine vendors, returns two separate scores across ten scored dimensions, explains in plain language exactly what is missing behind every low dimension, and hands back a structured Fix-It plan the submitter can execute. The submitter revises, re-submits, and re-scores, and the delta is shown dimension by dimension. Every review lands in a queryable corpus, and every public artifact leaves the system through a sanitization gate.

VerdictTank does not write proposals. It critiques them. That boundary is the product. A writing tool is incentivized to tell you the draft it produced is good. A critique tool is only valuable if it is willing to tell you the draft is not ready, name the reason, and quantify how far off it is. Every design decision in v4.1 follows from that boundary: the pre-submit coach asks structuring questions and never composes paragraphs, the panel scores independently before any synthesis, and the synthesis seat never scores at all.

The two scores are the reason a submitter trusts the output. A single blended number hides the most useful signal in proposal review, which is the gap between a strong idea in a weak document and a weak idea in a polished document. VerdictTank separates them:

Ten dimensions feed those two scores, five to each. Every dimension carries a required, evidence-backed explanation sentence naming the specific missing artifact, not a grade with no reason attached.

The commercial model is five price points: Free at $0, One-Shot at $29, Pro at $119 per month, Enterprise at $699 per month, and White-Label at $3,000 per month. Standard-tier cost of goods sold is $0.36 per review at the ceiling, $0.43 loaded. White-Label, which seats four reserved premium models, is $3.00 per review at the ceiling, $3.60 loaded. Every paid tier clears 88 percent gross margin at its full included allotment, and the standard tiers clear 96 percent or better. Cost is not the binding constraint on this business; distribution is.

2The Full Journey

VerdictTank is one continuous loop, not a scoring endpoint. The loop is the product.

2.1 Pre-submit coach

Before a submitter pays for anything, the coach is open and unlimited. It reads the draft in progress and asks structuring questions: where is the total addressable market (TAM) derivation, which competitor pricing is cited, what the operating plan assumes about hiring, which regulatory regime applies. It surfaces gaps. It never emits a score, and it never writes a full paragraph on the submitter's behalf. The submitter arrives at the review with a better draft, and the review is worth more because of it.

The coach is available on every tier including Free, and it is the top of the funnel. A submitter who has spent twenty minutes being asked hard questions about their own document already understands why the panel is worth paying for.

2.2 Submit by file or URL

Intake accepts DOCX, PDF, and TXT uploads, and it accepts pasted text. URL-to-Review accepts a link, extracts the page text, and converts it into the same canonical intake format a file upload produces, so a public pitch page or a hosted memo runs through the identical pipeline. Text is normalized to UTF-8, and a PII sanitizer redacts emails, phone numbers, and identifier-shaped patterns from the stored working text before any model call is dispatched. The original binary is preserved intact and separately.

2.3 The ten-seat panel

The standard panel is ten seats across nine vendors: nine independent scoring seats plus one synthesis and integrity gate that never scores. Every seat is dispatched in parallel against a single-pass contract, with a per-seat timeout and a per-seat fallback binding, and a pre-dispatch health probe runs against the rostered models before any spend is committed.

The nine scoring seats each own a distinct error class:

SeatError class it is built to catch
Research Agentgrounding failures and context errors
Primary Reviewerfull-rubric anchor plus revenue arithmetic
Market-Realitycompetitive mispositioning and TAM overstatement
Financial Integrityrevenue arithmetic and financial-model errors
Legal and Compliancelegal blockers and compliance gaps
Execution-Feasibilityexecution infeasibility and timeline-scope errors
Team and Founderteam capacity and founder-fit gaps
Risk and Ethics Red-Teamsafety-washing and overstated risk claims
Live Groundinghallucinated facts and stale or uncited market data

The tenth seat is the Synthesis and Integrity Gate. It reads all nine scoring outputs, reconciles them, computes the panel mean, median, and standard deviation per dimension, flags outliers at 1.5 sigma, records the spread between the anchor score and the cross-check scores, and issues the verdict. It has no scoring authority of its own, which is what makes it a gate rather than a tenth opinion.

Nine vendors are represented so that no single vendor's blind spots become the panel's blind spots. Vendor and model identity is server-side configuration only. It never appears on a customer-facing surface.

The panel is built for error detection density: the number of distinct material error classes surfaced per review, not the number of comments generated. A reviewer that produces forty stylistic notes and misses a broken revenue calculation has scored zero on the only metric that matters.

2.4 Two scores across ten dimensions

Each of the ten dimensions is tagged idea-facing or proposal-facing at generation time, and the two groups aggregate separately into the two published scores.

Proposal Strength Score (proposal-facing, five dimensions)

DimensionWhat it measures
Problem and Solution Claritywhether the problem, the solution, and the causal link between them are stated without ambiguity
Evidence and Citation Qualitywhether every load-bearing claim has a source, a date, and a derivation
Financial Model Integritywhether the numbers reconcile, the unit economics close, and the assumptions are visible
Execution and Operating Planwhether the plan has sequencing, owners, dependencies, and honest timelines
Compliance and Legal Readinesswhether the applicable regime is identified and the blockers are addressed

Investor Readiness Score (idea-facing, five dimensions)

DimensionWhat it measures
Market Reality and Sizingwhether the market exists at the claimed size and the TAM is derived, not asserted
Competitive Differentiationwhether the moat survives contact with named, priced competitors
Go-to-Market and Tractionwhether there is a repeatable path to the first and hundredth customer
Team and Founder Fitwhether the team can actually execute this plan at this scale
Risk and Ethics Exposurewhether the material risks are named honestly rather than minimized

The two scores are published side by side with the disagreement delta between the anchor seat and the cross-check seats. The spread is signal, not noise: high panel agreement on a low dimension is a hard finding, and high disagreement is itself reported as a flag for the submitter to investigate.

2.5 Explain the low score

Every per-dimension score ships with a required explanation field. The field is evidence-backed and specific. A dimension score of 41 on Market Reality and Sizing does not return "market sizing is weak." It returns the concrete absence: no TAM calculation, no competitor pricing data, no source for the growth rate cited on page four. The explanation names the missing artifact, because a missing artifact is actionable and an adjective is not.

The explanation field is generated on the same pass as the score, so an explanation can never drift away from the number it justifies. Free tier receives the score summary; Pro and above receive the full per-dimension explanation set.

2.6 The Fix-It plan

After the verdict, VerdictTank generates a structured, prioritized, dimension-tagged remediation plan keyed to the lowest-scoring dimensions. Each item carries the finding, the specific fix, and where practical the instrument required to execute it: the formula to compute, the table template to fill, the citation target to obtain, the section to rewrite and what it must contain.

The plan is ordered by score impact, so a submitter with two hours works the top of the list rather than guessing. Pro and above receive the full structured plan. Free receives a one-paragraph summary, which is deliberate: the summary proves the plan exists and is specific, and the full plan is the upgrade.

2.7 Revise and re-score

The submitter revises against the Fix-It plan and re-submits through the re-score endpoint. A fresh panel runs, and the report shows a before-and-after delta per dimension along with both new scores. The original review is preserved as a historical version with full lineage, so the improvement trail is durable and auditable by the submitter. Re-scores are billed as reviews, which keeps the incentive honest: VerdictTank is paid to run panels, not to declare victory.

2.8 Queryable review corpus

Every review lands in a corpus record: both scores, all ten dimension scores, the explanation set, the findings, the conditions, the verdict, the remediation list, revision lineage, and the vertical classification. The corpus is the substrate for percentile context today and for outcome-calibrated scoring later.

Corpus handling is strict by construction. Raw proposal text is never corpus-eligible; the filter is enforced at the schema level, not by policy. Only structural and aggregate metadata is eligible, contribution is opt-in with the flag defaulting to false, identifiers are stripped, and any aggregate publication is gated behind a k-anonymity threshold. White-Label tenants are isolated: every corpus query filters by tenant, and a tenant-scoped credential cannot read across the boundary.

2.9 Share, export, integrate

The verdict leaves the system three ways, and all three pass the same gate.

2.10 The sanitization gate

One gate governs every path out of the system. It strips personally identifiable information and it strips model and vendor identity from every public surface, including the free-text explanation and remediation fields where such identity is most likely to appear. It blocks deploys and it blocks artifacts; it is not an advisory scan. The PDF pipeline verifies the rendered output by extracting text from the finished file and asserting zero vendor names, zero model identifiers, and the presence of the non-removable AI disclaimer before the report is released.

3Feature Set

3.1 Before the review

FeatureFreeOne-ShotProEnterpriseWhite-Label
Pre-submit coach (unlimited)YesYesYesYesYes
File upload intake (DOCX, PDF, TXT, paste)YesYesYesYesYes
URL-to-Review intakeNoYesYesYesYes

3.2 During the review

FeatureFreeOne-ShotProEnterpriseWhite-Label
Panel size4 scoring seats plus gate10 seats10 seats10 seats11 seats, premium models
Vendors represented49999
Two scores across ten dimensionsSingle summary scoreYesYesYesYes
Per-dimension explanationsScore summary onlyFullFullFullFull
Panel spread and outlier flagsNoYesYesYesYes
Vertical auto-classificationNoYesYesYesYes
Vertical templatesNoNoNoYesYes
Configurable Review Rules EngineNoNoNoYesYes

3.3 After the review

FeatureFreeOne-ShotProEnterpriseWhite-Label
Fix-It plan1-paragraph summaryFull structuredFull structuredFull structuredFull structured
Re-score with per-dimension deltaNoNo (re-purchase)Yes, billedYes, billedYes, billed
Branded PDF reportNoYesYesYesYes
Shareable report linkNoYesYesYesYes
Review-as-a-Service APINoNoNoYesYes
Reviewer accuracy track recordFeeds dataFeeds dataFeeds dataData plus dashboardData plus dashboard
White-Label track: domain, logo, email templates, portfolio consoleNoNoNoNoYes

3.4 Foundation, every tier

FeatureStatus
Review corpus with two-score schemaAll tiers
Sanitization gate on every public surfaceAll tiers, blocking
Automated PDF pipeline, zero manual stepsPaid tiers
Outcome-tracking cron at T+90, T+180, T+365All tiers, accumulating
Tenant isolation on corpus and credentialsWhite-Label
Non-removable, versioned AI disclaimer on every reportAll reports

3.5 Reviewer accuracy track record

Every seat accumulates a track record from its own scores against later recorded outcomes. VerdictTank surfaces that track record on Enterprise and White-Label dashboards and continues accumulating it on every tier. v4.1 does not weight live verdicts by accuracy. A weighting scheme applied before the corpus can support it would be a confidence claim the data cannot back, so the track record is published and the verdict stays unweighted.

4Pricing

Five price points. Monthly, with annual available on the three subscription tiers.

TierMonthlyAnnual (per month)Included reviewsPanelImplied per reviewOverage
Free$0n/a1 lifetime4 scoring seats plus gaten/anone
One-Shot$29none1Full 10-seat$29.00none
Pro$119$998 per monthFull 10-seat$14.88$18
Enterprise$699$58250 per monthFull 10-seat$13.98$16
White-Label$3,000$2,499100 per month11-seat premium$30.00$28
Free
$0/ lifetime
  • One lifetime review
  • 4 scoring seats plus gate
  • Single summary score
  • Unlimited pre-submit coach
  • Percentile context
  • One-paragraph Fix-It summary
A demonstration, not a workflow.
One-Shot
$29/ one review
  • Full 10-seat panel
  • Both scores across ten dimensions
  • Full per-dimension explanations
  • Full structured Fix-It plan
  • Branded PDF report
  • Shareable report link
No overage. No annual plan. Revise and re-score by buying again or moving to Pro.
Enterprise
$699/ month
  • 50 reviews per month
  • Vertical templates
  • Configurable Review Rules Engine
  • Review-as-a-Service API
  • Accuracy dashboard
  • Overage at $16
Annual: $582 per month. Implied per review: $13.98.
White-Label
$3,000/ month
  • 100 reviews per month
  • 11-seat premium panel, 4 reserved models
  • Custom domain, logo, email templates
  • Portfolio console
  • Tenant-isolated corpus
  • Overage at $28
Annual: $2,499 per month. Implied per review: $30.00.

4.1 What each tier is for

Free, $0, one lifetime review on the reduced panel. Four scoring seats plus the gate, single summary score, percentile context, unlimited coach, and a one-paragraph Fix-It summary. It proves the panel is real without giving away the full ten-seat output. One review is lifetime, not monthly, so Free is a demonstration rather than a workflow.

One-Shot, $29, one review on the full ten-seat panel. The bridge for the founder who needs one honest read and is not ready for a subscription. It runs the complete standard panel, both scores, the full per-dimension explanations, the full structured Fix-It plan, the branded PDF, and a shareable link. There is no overage, because a single purchase has nothing to exceed, and there is no annual plan, because it is not a subscription. The buyer who wants to revise and re-score buys again or moves to Pro.

Pro, $119 per month, eight reviews. Eight reviews is a real iteration cadence: two per week, propose, review, revise, re-review. At $14.88 implied per review it prices below the one-off, so the subscription reads as the better deal for anyone actually iterating. Overage at $18 sits just above the included rate, which nudges heavy solo users toward the bundle they already have or up to Enterprise. Annual is $99 per month.

Enterprise, $699 per month, fifty reviews. Roughly six times Pro's volume for roughly six times the price, so the ladder stays proportional and the upgrade is easy to justify. A team running multiple proposals and request-for-proposal (RFP) responses lands in the thirty to fifty range per month, so fifty is generous but bounded. Enterprise adds vertical templates, the Configurable Review Rules Engine, the API, and the accuracy dashboard. Overage at $16. Annual is $582 per month.

White-Label, $3,000 per month, one hundred reviews. The reseller and consultancy tier. Custom domain, custom logo, branded email templates, a portfolio console, and tenant-isolated corpus segments, so a consultancy runs the panel entirely under its own brand. It is the only tier that seats the four reserved premium models, across eleven seats rather than ten: a premium primary reviewer, a premium execution-feasibility seat, a premium red-team seat, a premium synthesis gate, and a retained cross-check seat that preserves vendor diversity against the premium anchor. Per-review cost is roughly ten times standard, and the price carries it. Overage at $28 reflects the premium roster. Annual is $2,499 per month.

4.2 Configurable Review Rules Engine

Enterprise and White-Label administrators define additive custom checks through a no-code builder: compliance rules, brand-voice guidelines, internal investment criteria, mandatory sections. Custom rules layer on top of the fixed ten-dimension rubric. They never replace it and they never suppress a dimension score, so a tenant cannot configure away a finding it does not want to see. That constraint is what keeps a white-labeled verdict worth the same as a first-party one.

5Roadmap

5.1 The v4.1 release track

v4.1 is planned as five gated phases over twenty-four weeks, built by a solo developer plus an AI-agent build pipeline. The review engine and pipeline are proven in operation today; the client portal, billing, and API are the build. Each phase closes on a hard acceptance gate, and no phase closes on a self-report.

PhaseWeeksDeliverableClosing gate
0. Foundation and corpus1 to 3Blocking sanitization gate covering explanation and remediation fields; scripted PDF pipeline; corpus schema with two-score columns and reserved outcome columns; outcome cron scheduledZero manual sanitization steps in any public path; one review runs end to end into the corpus and out as a PDF with no manual step; cron logs its first run
1. Intake and scoring4 to 8Coach tier-wide; URL-to-Review alongside file upload; two scores as the panel's default output contract; per-dimension explanation as a required field; accuracy tracking beginsEvery review emits two scores and a per-dimension explanation; coach and URL intake in production with zero gate failures; accuracy tracking records every seat on every review
2. Fix-It and reports9 to 13Structured Fix-It on Pro and above; shareable report links; vertical auto-classification and the first three vertical templatesFix-It plan on 100 percent of Pro-and-above reviews; a shareable report passes the gate end to end; two of three vertical templates validated against known outcomes
3. Enterprise controls and API14 to 18Rules Engine to Enterprise and White-Label; Review-as-a-Service API to Enterprise; vertical templates to fiveRules Engine live with three Enterprise accounts; the API completes 100 reviews with zero gate failures
4. White-Label and GA19 to 24White-Label track live with custom domain, logo, email templates, portfolio console, tenant isolation; API generally available; five price points live on both surfacesFirst White-Label pilot renews past month one; API generally available to Enterprise; marketing site and portal show the five price points with no stale pricing anywhere

5.2 What comes next

v4.1 and beyond, in order of expected value:

  1. Outcome-calibrated scoring. The outcome cron accumulates from launch day at T+90, T+180, and T+365, and the schema carries the outcome columns from day one, so no migration is required. Once the corpus clears a minimum-N threshold, the accuracy track record becomes a published dashboard and then a candidate weighting input.
  2. Competitive review comparisons. Full benchmarking of a new proposal against the corpus distribution, dimension by dimension and vertical by vertical, replacing the percentile context available today.
  3. Market simulation. Replaces the static financial table with a twelve-month trajectory model, adding a per-review cost that the Enterprise and White-Label price points absorb.
  4. Second-opinion audit agent. A dedicated blind-spot pass over the panel's own output, held until the corpus and accuracy data can measure its catch rate against a real baseline rather than an assumption.
  5. Adversarial red-team per vertical. Industry-specific attack vectors, sequenced after the vertical classifier has a proven accuracy record.

5.3 Explicit non-goals

Scope is bounded on purpose:

6Financial Model

6.1 Unit economics and cost of goods sold

Costing assumption: 12,000 input tokens per scoring seat for a twenty-page proposal, roughly 40,000 characters, with the worker truncating above that; the synthesis gate reads roughly 20,000 input tokens. Per-seat cost is input rate times input tokens plus output rate times output budget. Dispatch is single-pass chat completion with no tool calling, which is the regime that makes a ten-seat panel cost cents rather than dollars. All margin guarantees below use the ceiling figure, which assumes every seat burns its full upper-bound output budget. The base figure is the typical case and is never used for a margin claim.

RosterSeatsCOGS ceiling per reviewLoaded ceiling (x1.20)COGS base per review
Free reduced panel5 model calls$0.15$0.18$0.10
Standard panel10$0.36$0.43$0.25
White-Label premium panel11$3.00$3.60$1.52

The loaded figure applies a flat 20 percent infrastructure and overhead buffer covering the application host, the database, object storage, the PDF renderer, and email delivery.

6.2 Gross margin at full allotment consumption

Margin is computed at the pessimistic bound: every included review consumed, every seat at its ceiling output budget, loaded cost.

TierRevenueIncluded reviewsLoaded COGS at full consumptionGross margin
Free$01 lifetime$0.18 one timeloss leader
One-Shot$291$0.4398.5%
Pro$1198$3.4497.1%
Enterprise$69950$21.5096.9%
White-Label$3,000100$360.0088.0%

Overage is itself high margin by construction. Pro overage at $18 and Enterprise overage at $16 both sit far above the $0.43 loaded standard cost, so overage carries better than 95 percent margin while still reading as a nudge toward the next tier. White-Label overage at $28 against $3.60 loaded carries roughly 87 percent margin.

Stress case: if both premium reasoning seats in the White-Label roster burn a full 12,000-token output budget, per-review ceiling reaches $3.87, or $4.64 loaded. One hundred such reviews cost $464.40 against $3,000 revenue, which is 84.5 percent gross margin. The worst realistic case on the most expensive tier still clears 84 percent.

The structural conclusion is that cost of goods sold is not the constraint on this business. Even two hundred Enterprise reviews in a month cost roughly $86 loaded against $699 revenue. Review allotments are therefore set by value anchoring and ladder logic, not by cost recovery, and pricing pressure can be absorbed without touching the panel.

6.3 Revenue projection, floor and ceiling

Two scenarios at month twelve post-launch, measured as monthly recurring revenue (MRR). Both are stated as assumption sets, not forecasts. Both assume every included review is consumed, which overstates cost and understates margin.

Floor scenario, month 12

LineAccounts or volumeMonthly revenueMonthly loaded COGS
One-Shot40 purchases per month$1,160$17.20
Pro35 accounts$4,165$120.40
Enterprise3 accounts$2,097$64.50
White-Label0 accounts$0$0.00
Total$7,422 MRR$202.10

Floor gross margin: 97.3 percent. Annual run rate at month twelve: $89,064.

Ceiling scenario, month 12

LineAccounts or volumeMonthly revenueMonthly loaded COGS
One-Shot150 purchases per month$4,350$64.50
Pro180 accounts$21,420$619.20
Enterprise14 accounts$9,786$301.00
White-Label4 accounts$12,000$1,440.00
Total$47,556 MRR$2,424.70

Ceiling gross margin: 94.9 percent. Annual run rate at month twelve: $570,672.

Free tier cost is a one-time charge per account rather than recurring, since Free grants one lifetime review. At $0.18 loaded per Free review, one thousand two hundred cumulative Free reviews cost $216 in total and six thousand cost $1,080 in total. Free is affordable at any signup volume the funnel can realistically produce, which is why the reduced panel exists rather than a time-limited trial.

Annual billing at $99, $582, and $2,499 per month trades 16.7 percent of headline revenue, two months free, for twelve months of committed cash and materially lower churn exposure. At the ceiling scenario, a fifty percent annual mix on Pro and Enterprise reduces month-twelve MRR by roughly $2,619 and converts roughly $156,000 of annualized revenue into prepaid commitment.

6.4 What moves the model

Sensitivity, ranked:

  1. Pro account count. Pro is the volume tier and the largest single revenue line in both scenarios. It is the number to move.
  2. White-Label logos. Each White-Label account is worth roughly twenty-five Pro accounts. Landing one changes the shape of the revenue curve; landing four is the difference between the floor and the ceiling scenario.
  3. One-Shot to Pro conversion. One-Shot is priced as a bridge, and its value is mostly in what fraction of buyers subscribe after seeing the full ten-seat output once.
  4. Enterprise seat expansion. Enterprise is the highest-effort sale, and the Rules Engine and API are the features that make it defensible rather than a volume discount.
  5. Cost of goods sold. Last, and by a wide margin. A doubling of every model rate in the panel would still leave Pro above 94 percent gross margin.

7Risk Assessment

Ranked by expected impact on the product at launch, each with the control that is in place.

7.1 Reviewer availability High impact

The panel depends on nine vendors, and any one of them can rate-limit, exhaust credit, or return transport errors. Controls: a pre-dispatch health probe runs one cheap call per rostered model before spend is committed; every seat carries a named fallback binding; a review completes on a documented reduced panel rather than failing when a seat cannot be filled, and any review that ran reduced is flagged as such on the report and in the corpus record. The live-grounding seat is the single most availability-sensitive seat in the standard roster and is provisioned with a direct vendor credential rather than a shared route, plus a same-vendor-class fallback that is already live in the panel.

7.2 Model output integrity High impact

Reasoning models can truncate structured output at low token caps, and some models constrain sampling parameters. Controls: per-seat output budgets are sized above the truncation threshold for every reasoning seat rather than set to a global default; seats with sampling constraints carry an explicit per-seat parameter quirk in the roster configuration; every seat response is schema-validated before it enters synthesis, and a seat returning unparseable output is retried once and then fails over rather than silently contributing a null score. Models without a published per-token input cost are not seated in any cost-guaranteed tier at all.

7.3 Cost verification Medium impact

Two standard-roster seats bill through providers whose usage responses do not always carry a cost field, so their per-review contribution is computed from configured rates rather than reconciled against metered spend. Controls: those seats are provisioned with native vendor credentials so spend meters end to end; the margin guarantee uses the ceiling COGS across the whole roster, which absorbs a material rate change on any single seat; the combined ceiling contribution of the two affected seats is $0.0632 per review against a $0.36 total, so even a doubling of both moves loaded standard COGS from $0.43 to roughly $0.51 and leaves Pro above 96 percent margin. Base COGS figures are internal planning numbers and are not quoted externally.

7.4 Sanitization and confidentiality High impact if it fails

A critique product handles unreleased strategy documents, and a leak of raw text or reviewer identity is an existential trust failure rather than a bug. Controls: the sanitization gate blocks rather than warns, and it covers the free-text explanation and remediation fields alongside structured fields; raw proposal text is schema-level ineligible for the corpus; corpus contribution is opt-in and defaults to false; the PDF pipeline extracts text from the finished file and asserts zero vendor and model identifiers plus the verbatim presence of the AI disclaimer before release; share tokens are unguessable, expirable, and revocable, and a revoked token returns not-found; tenant isolation is enforced on every corpus query and every scoped credential.

VerdictTank does not use submitted proposals to build its own products, train models, or inform its own proposals. A submission is processed only to produce that submitter's review, and is retained only for the submitter's own reference and legal record. Review processing does run through third-party AI providers under their own data-handling terms.

7.5 Verdict liability Medium impact

A submitter can act on a verdict and attribute an outcome to it. Controls: a versioned, non-removable AI disclaimer renders on every report and cannot be templated away; inter-seat agreement and the panel spread are published beside every score so confidence is visible rather than implied; the liability cap is the greater of $100 or twelve months of fees; the status page distinguishes VerdictTank incidents from upstream provider incidents so an outage is not misread as a product defect.

7.6 Positioning drift Medium impact and slow

The most likely way this product degrades is by drifting toward writing. Customers will ask for it, and a generated paragraph is easier to deliver than an honest score. Controls: the coach is architecturally forbidden from composing paragraphs, the Rules Engine is additive only and cannot suppress a dimension score, and no roadmap item shifts VerdictTank toward authorship. We critique them; we don't write them, and that is a product constraint, not a slogan.

7.7 Concentration and capacity Medium impact

v4.1 is built and operated by a solo developer plus an AI-agent build pipeline, and the review worker is serial at launch. Controls: the review state machine is durable and replayable, with per-seat evidence stored so a partial panel resumes rather than restarting; the scale path is a depth cap plus parallel workers, which is a configuration change and not a redesign; the full customer-facing stack and the review engine sit on a single host with one state machine, so there is no cross-host coordination to debug under load. At the ceiling scenario, month-twelve volume is roughly 2,690 reviews per month, which a serial worker at a sixty-second critical path clears with substantial headroom.

7.8 Distribution Highest impact overall

The panel works and the margins are structurally excellent, which means the binding risk is that not enough submitters find the product. Controls: Free is a real full-loop demonstration on a reduced panel rather than a time-limited trial; One-Shot at $29 removes the subscription objection entirely; the coach is unlimited on every tier and is the widest part of the funnel; shareable report links put a branded verdict in front of the submitter's own investors and advisors; the API and the White-Label track make other people's distribution into ours.

8Summary

VerdictTank v4.1 ships the complete critique loop: coach the draft, submit a file or a URL, run ten seats across nine vendors, publish a Proposal Strength Score and an Investor Readiness Score across ten dimensions, explain every low dimension with the specific missing artifact, hand back a prioritized Fix-It plan, re-score the revision with a per-dimension delta, keep it all in a queryable corpus, and ship the verdict as a branded PDF, a shareable card, or an API response with a sanitization gate on every path out.

Five price points cover the range from a single honest read at $29 to a fully branded reseller platform at $3,000 per month, and every paid tier clears 88 percent gross margin at full consumption with the standard tiers above 96 percent. The economics are settled. The architecture is proven. What remains is the build, then distribution.

We critique them; we don't write them.