Deal Intelligence for Used Heavy-Duty Diesel Trucks
Buyers of used GM heavy-duty diesel pickups (2012-2015 Silverado/Sierra 2500HD/3500HD, 6.6L LML Duramax) face wide regional price dispersion, undisclosed rust exposure from geography, and known model-specific failure modes such as CP4 injection pump risk. Listings show price and mileage but not whether a given truck is a good deal or a good truck - two different questions that buyers conflate. The product scores both…
What an analysis cost to produce belongs beside it. A reader deciding whether to trust a verdict is entitled to know whether it came from twenty-six stages or one, and nothing else in this category will tell them.
The case against it
| Charge | Rebuttal | Ruling |
|---|---|---|
| Subscription model mismatched to one-time purchase behavior | CONCEDED — Conceded. No evidence of repeat-use hooks, alerts, or multi-purchase touchpoints exists in the record to counter the collapse to a single-month LTV. | upheld — Correct on the facts but the smallest_offer already abandons subscription for a $15 one-time report, so it kills the $19/mo architecture, not the candidate. |
| Scoring engine ships uncalibrated numbers as confident verdicts to a domain-expert audience | CONCEDED — Conceded. The record states outright that FMV and repair-reserve figures are guesses, and no hedging layer or human review is described anywhere in th | upheld — This is the real charge: an expert forum audience is the acquisition channel AND the audit panel, and uncalibrated FMV/repair-reserve numbers get caught in public on the first bad report. |
| Marketcheck cost structure breaks unit economics at realistic early subscriber counts | CONCEDED — Conceded. The cost figures and cap numbers in the charge match the thesis exactly, and no alternate data source or cost-mitigation plan is documented. | partial — True at scale, irrelevant at test scale — 500 free calls/month covers roughly 50-100 single-VIN manual reports, far more than the first-dollar test needs; the $749 wall is a month-6 problem, |
| Buyer population figure is a guess built on proxy data, not a counted market | CONCEDED — Conceded outright — the record self-labels this figure as unverified. | partial — Guessed, yes, but even the pessimistic floor (a few hundred active shoppers/month) supports a $15-49 one-time report business for one person; it caps the ceiling, not the launch. |
| Only high-intent channels are manual and unscalable for a solo operator | CONCEDED — Conceded. The channel list itself documents unpaid manual hours as the mechanism for the two high-intent channels and thin/high-CPC search as the only | partial — Manual is a feature at day 8 — DuramaxForum threads asking 'is this a good deal' are a named, findable, free channel with buyers self-identifying by VIN; the scaling ceiling is a later probl |
| Price anchor to VinAudit conflates two different buyer jobs | CONCEDED — Conceded. The channel evidence itself shows buyers already get this exact judgment for free in the target community, undermining transfer of VinAudit' | upheld — Decisive on price: the forum already answers 'is this a good deal' for free with human expertise the tool cannot match, so the product competes against free crowdsourced judgment in its own |
| Domain-specific mechanical claims (CP4 risk, rust exposure) generated by LLM without a verified domain data source | CONCEDED — Conceded. No specialized CP4/rust dataset is named anywhere in the mechanism — only Marketcheck listing metadata and an LLM call for prose. | upheld — CP4 risk and rust exposure are the entire differentiator and they are the least verified part of the stack; a generic LLM will produce forum-common-knowledge prose that this audience already |
| Data moat requires infrastructure that doesn't exist yet, leaving zero defensibility during the most vulnerable period | CONCEDED — Conceded verbatim by the candidate's own data_moat section. | dismissed — Moats are earned in year two. No moat at conception is baseline, not a defect. |
| Advice liability on high-stakes purchase decisions with no human check | CONCEDED — Conceded. The mechanism confirms zero human review on an uncalibrated engine issuing purchase verdicts on $25-50k transactions. | dismissed — A $15 disclaimered directional report is the same liability posture as every free deal-rating tool in the competitor list; ordinary risk, not severity. |
A separate agent argued against this idea, a second answered, a third ruled. 9 of 9 charges were conceded rather than defended. Published in full because a score with the objections removed is a advertisement, and because the objections are usually more useful than the verdict.
How it scored
| Dimension | Score | Reasoning |
|---|---|---|
| D1 | 58 | Real anxiety on a $30-45k purchase with a known $3-5k failure mode, but the buyer's actual pain is resolved for free in 20 minutes by posting the listing on DuramaxForum. |
| D2 | 42 | Adjacent paid markets exist (VinAudit ~$20/mo, Carfax ~$45/report) but those buy documented history records, not opinion; the direct product — deal scoring — has a hard $ |
| D3 | 72 | DuramaxForum, r/duramax and two named Facebook groups contain buyers posting VINs and asking exactly this question daily — 100 reachable this month is realistic and free; |
| D4 | 25 | Thin wrapper on commodity Marketcheck data with an LLM prose layer; delisting-outcome data could compound but doesn't exist yet. Only 5% weight, so barely matters. |
| D5 | 68 | Scoring engine already written, Marketcheck free tier covers manual single-lookup volume, 6 days to a form + PDF pipeline; the uncosted work is calibration, which is judg |
| D6 | 40 | One-time $15 report is honest to the buying behavior but yields ~$15 LTV against manual channel labor measured in hours per customer; the $19/mo subscription in the recor |
| D7 | 22 | One engine family, one four-year window, single-transaction buyers, guessed population of low thousands — this tops out in the low four figures monthly even executed well |
| D8 | 70 | Python/Flask/Postgres/Stripe is exactly the stack; the gap is diesel domain credibility in a forum that will test it, which the operator's technical fluency does not supp |
| D9 | 88 | 8 days is credible: the engine exists, a free-tier key is instant, and the first sale is a forum reply with a Stripe link — well under 14 days. |
Who already does this
| Competitor | Pricing | Funding | Launched | Overlap |
|---|---|---|---|---|
| DealJudge | app store listing, exact price not disclosed in search resul | not found | active as of May 2026 per | exact |
| CarGurus Deal Ratings | free to consumers (ad/dealer-lead funded) | public company (NASDAQ: CARG) | deal rating feature long-s | partial |
| CoPilot | free to consumers | VC-backed per Crunchbase profile, amount not d | active, iterating with AI | partial |
| CarEdge | free calculator; CarEdge Pro subscription for negotiation/in | not found | active, Pro tier reference | partial |
| VinAudit Market Value API | B2B API pricing, not public; consumer report typically ~$20- | not found | established, API-first pro | adjacent |
| Marketcheck | $749/mo tier referenced in candidate brief; free tier 500 ca | not found | established data vendor, n | adjacent |
Where the buyers actually are
| Channel | Why it reaches them |
|---|---|
| DuramaxForum.com (For Sale/Wanted + general discussion subforums) | |
| Facebook Groups (e.g. 'Duramax Diesel Truck Owners', '6.6 Duramax LML | |
| r/duramax (Reddit) | |
| Google Ads on 'CP4 Duramax buying guide' / '[year] Silverado 2500HD Du | |
| Cold DM to Craigslist/Marketplace diesel HD sellers |
Regulatory gates
| Gate | Finding |
|---|---|
| G1 | No identified legal violation. Vehicle data aggregation and resale of analysis is standard practice (Edmunds, KBB, Carvana all operate this model). No terms-of-service violation apparent in |
| G2 | Fully autonomous after deployment. Inventory ingestion is automated API pull, scoring is deterministic Python, verdict generation is LLM call. Zero per-listing human review required. No sale |
| G3 | Edmunds, KBB, TrueCar, Carvana, and niche players like Bring a Trailer all charge for vehicle intelligence and deal scoring. Market willingness to pay is proven. Specificity to one engine fa |
| G4 | Thin MVP is buildable in 30 days: ingest Marketcheck free tier via anchor-ZIP grid (workaround is tedious, not novel), run existing Python scoring logic, call OpenAI API for verdict prose, s |
| G5 | Buyers of used 2012-2015 Duramax trucks are reachable via: Duramax-specific forums (Duramax Diesel Forum, Chevy Truck Forum), Facebook groups (Duramax owners), Reddit (r/Trucks, r/Diesel), Y |
The pre-registered test
| Term | Value |
|---|---|
| days | 14 |
| offer | Post in DuramaxForum 'For Sale/Wanted' and 'LML' subforums plus the two named Facebook groups, replying to live 'is this a good deal?' threads with a FREE full sample verdict (Deal Score, Truck Score, repair reserve, rus |
| price | 29 |
| metric | Stripe payments completed for paid VIN reports from strangers (free proof reports do not count) |
| channel | DuramaxForum.com (For Sale/Wanted + LML subforums) and Facebook groups 'Duramax Diesel Truck Owners' / '6.6 Duramax LML Owners' |
| threshold | 5 paid reports at $29 within 14 days, AND zero public forum posts disputing a number in a delivered report — if 5+ paid but a number gets publicly contradicted, that is a fail requiring calibration before proceeding |
Recorded at the moment the verdict was issued and not editable afterwards. If this is launched, the result lands on the ledger whether it passes or fails.
Where this idea is
| Phase now | Analysed — Underwritten, with the argument against it on the record. |
| What you do here | Read the case against it first. An upheld charge you cannot answer is the verdict, whatever the score says. |
| To leave this phase | You have read the upheld charges and decided the idea survives them. |
| Gate status | This gate is a judgement, not a query. The system will not rule on it and will not pretend to — you decide, and the reason is recorded. |
| Next phase | Validating — A pre-registered test is live and running. |
This gate is a judgement rather than a query, so the system states it and refuses to rule on it. Pretending software can decide whether a business "can take money from somebody who is not you" would make every gate on this site meaningless. Advancing an idea needs its link — the one handed back when it was submitted. Founder-owned ideas are advanced from the console. See the whole pipeline.
| What to do in this phase | What it proves | From which part of the analysis |
|---|---|---|
| Answer the upheld charge: Subscription model mismatched to one-time purchase behavior | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the upheld charge: Scoring engine ships uncalibrated numbers as confident verdicts to a domain-expert audience | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the partial charge: Marketcheck cost structure breaks unit economics at realistic early subscriber counts | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the partial charge: Buyer population figure is a guess built on proxy data, not a counted market | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the partial charge: Only high-intent channels are manual and unscalable for a solo operator | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the upheld charge: Price anchor to VinAudit conflates two different buyer jobs | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
| Answer the upheld charge: Domain-specific mechanical claims (CP4 risk, rust exposure) generated by LLM without a verified domain data source | The verdict survives its strongest objection, or it does not and you have learned that before spending. | arbitration |
Every step traces to a field this idea's own underwriting produced — not generic best practice, which is free everywhere. 0 of 7 complete. Mark them off in the console.
How this was produced
| Measure | Value |
|---|---|
| Cost to produce | $0.01 |
| Wall clock | 12 minutes |
A verdict produced by 22 of 23 stages is not the same artefact as one produced by all of them, and which stages failed was recorded on every run and shown nowhere until now. If a stage that feeds a section died, the section came from somewhere else or nowhere — and you are entitled to know which is in front of you before you act on it.
The money
| price point | anchor: VinAudit's own $20/month unlimited-report dealer tier and the $9.99-$44.99 single-report price band buyers already pay; monthly: 19; rationale: The deal-score verdict itself has a $0 price floor because CarGurus/CoPilot give it away, so the product can't charge for the score alone. It can charge for the bundle (Deal Score + Truck Score + repair reserve + shipping estimate + rust risk) at roughly the price point buyers already accept for a single pre-purchase VIN check. Pricing at $19/month unlimited-during-search (undercutting VinAudit's <cite index="37-1">$20 per month · $1 per report Pay month-by-month, expect no other fees, and cancel at any time</cite> model) or $24 one-time per- |
| current spend | amount: $0 for deal-rating/verdict tools; $5-$45 for the adjacent vehicle history report most buyers actually pay for before a purchase decision; source: CARFAX listing pages; VinAudit/Carfax price comparison (epicvin.com); findthebestcarprice.com VinAudit review; vinaudit.com/dealers pricing page; on what: Deal Score/Truck Score type verdicts are already free: CarGurus and CoPilot give consumers deal ratings and value estimates at no charge, monetizing via dealer leads instead. The closest thing buyers currently pay cash for is a VIN history report. <cite index="35-2">VinAudit car reports cost significantly less than Carfax ($9.99 vs $44.99).</cite> Dealers/power-shoppers can instead pay <c |
| funding route | reinvest |
| revenue model | hybrid |
| churn monthly pct | why: This is not a recurring-need product. A buyer shops this narrow truck category for roughly one 2-8 week window per multi-year ownership cycle, then has zero further use for the tool. Any monthly subscription wrapper will see the bulk of subscribers cancel after their purchase closes or they walk away from the search - churn in the 40-60%/month range is the realistic expectation for an intent-driven, one-decision product, not a stable utility subscription. This is inferred from the nature of the buying cycle, not from a churn dataset for this specific product.; value: 50 |
| cash to first dollar | 150 |
| marginal cost per unit | value: 0.05; components: Two real cost lines per report: (1) Marketcheck API calls to build the national comparable set - free at 500 calls/month but capped at a 100-mile radius, so each report likely consumes several of the 15-20 anchor-ZIP calls needed for a national comp pull rather than one call per report, meaning the free tier exhausts after roughly 25-35 reports/month before the $749/month tier must be turned on; that fixed cost, not the marginal cost, is what dominates at low volume. (2) LLM verdict generation - using a low-cost model like GPT-4o-mini at <cite index="27-1">$0.15 per million input tokens, $0.60 per million output tokens</cite>, a ~1,500 input / 500 output token verdic |
What has to be built
| wedge | segment: LML Duramax (2011-16) truck buyers; evidence: DealJudge, CarGurus, CoPilot and CarEdge all do general price-vs-comps scoring; none encode CP4 pump risk timing or geographic corrosion exposure into a per-listing verdict, which is the exact gap this cohort's high repair costs make expensive to ignore.; why they switch: A buyer about to spend $30-45k on a 2012-2015 2500HD/3500HD needs to know two different things a generic tool conflates: is the price fair, and is this specific truck's drivetrain/rust exposure a $3-5k time bomb. A repair-reserve number and rust-risk flag are decision-changing information a Camry-style deal score can't produce, and nobody else in the results is scoring |
| data moat | If delisting events are tracked as a proxy for sold price/time-on-market against VIN-level condition data, this accumulates into a proprietary outcome dataset (what actually sold, at what discount to ask) that Marketcheck's raw listing API doesn't expose and a new entrant would need months of continuous tracking to rebuild. Otherwise the moat is thin — the underlying inventory data itself is a commodity anyone can buy from the same vendor. |
| components | Marketcheck ingestion pipeline (anchor-ZIP grid, pagination, VIN-based dedup): risk: med; units: 3; Scoring engine calibration against real market data: risk: high; units: 3; Deal Score refinement (comp-set selection, FMV curve): risk: high; units: 2; Truck Score refinement (CP4 risk, rust/geography table, mileage curve): risk: med; units: 2; Repair reserve model (known failure-mode cost table): risk: med; units: 1; Shipping-to-destination estimator (distance/haversine, no paid API needed): risk: low; units: 1; Negotiation range calculator (deterministic): risk: low; units: 1; LLM verdict prose generation + prompt tuning + fallback template: risk: med; units: 2; Postgres schema + listing ver |
| total units | 28 |
| smallest offer | what: Manual/semi-automated one-page 'Buy or Pass' report: buyer submits a listing URL/VIN + ZIP via a simple form, operator pulls comps on-demand from Marketcheck's free tier (no national anchor-ZIP crawl needed for single-lookup volume), runs the already-built deterministic scoring against that pull, and an LLM call writes the plain-English verdict, repair reserve, and rust-risk note — delivered as a PDF within a few hours, sold as a one-time report, no subscription, no calibration debt disclosed upfront ('directional estimate, not appraisal') until real transaction data validates the model.; price: 15; format: report; days to build: 6 |
| wedge strength | workable |
| hardest unknown | The deterministic scoring engine has never touched real Marketcheck data — its FMV comps and repair-reserve dollar figures are unvalidated guesses. Nobody knows if the anchor-ZIP workaround for the free tier produces a representative national comp set or a systematically biased one until an actual key is pulled and the output is checked against known recent sales. This isn't a build risk, it's a 'does the core product claim hold up' risk, and it's untested. |
| days to first dollar | 8 |
If you decide to do this
| Step | What it means | Where it happens |
|---|---|---|
| 1 · Read the case against it first | Charges the arbiter upheld are the ones to answer before committing. If an upheld charge is fatal for you, the verdict is not. | on this page |
| 2 · Commit the pre-registered test | The test is already written: Stripe payments completed for paid VIN reports from strangers (free proof reports do not count) at 5 paid reports at $29 within 14 days, AND zero public forum posts disputing a number in a delivered report — if 5+ paid but a number gets publicly contradicted, that is a fail requiring calibration before proceeding. Committing freezes it with a date, and it cannot be edited afterwards. | promote it → |
| 3 · Stand up the offer | A landing page, a price, and an instrumented link. Nothing is proven until somebody who does not know you is asked to pay. | ventures → |
| 4 · Run distribution and let it resolve | The test resolves mechanically on its deadline: actual against threshold, no judgement. A test never distributed resolves VOID rather than FAIL — inaction is not evidence. | automatic, daily |
| 5 · The outcome grades this verdict | Whatever happens is written back against this prediction and scored. That is what makes the next verdict better, and it is the only honest basis for ever claiming an accuracy. | the ledger → |
Worth testing before committing. The premise is plausible and at least one load-bearing assumption is unevidenced. Steps 2 and 3 open the operator console, which lives under this same domain at /account and requires a log-in — the public record is readable by anyone, and committing a prediction against it is not. Step 5 happens automatically: this prediction is already frozen with its score, its confidence, and every dimension as it stood, waiting for an outcome to grade it against.