The Brick Drop · State of the Project & Roadmap

Where we are & where we go

A full, plain-language tour of the whole system as it stands today — every layer, what's live, what's built-but-not-yet-visible — and a detailed map of what's left to do.

2026-07-12 · LIVE at thebrickdrop.com · main @ 31bd050 · 20 PRs merged · 556 tests green
The state in one breath

The Brick Drop is a live, installable web app that grades 5,483 LEGO sets — each through four buyer lenses, colored by its own photo, with a brand-voice take. The engine, the catalog, personalization, and the share loop are all built and shipped. The current frontier is trust and activation: a just-shipped value upgrade that isn't visible yet, an accuracy fix waiting to be committed, and the work to prove the grade is right.

5,483
Sets graded in the live database (of ~5,754 buildable, 2017–2026)
4
Buyer lenses, each a real grade: display · builder · parent · investor
5,883
Review references feeding grades — 3,916 blog · 1,711 YouTube · 256 forum
20
Pull requests merged; 556 automated tests, all green
6
Scoring clusters roll dozens of dimensions into one honest letter
LIVE
thebrickdrop.com — auto-deploys on every merge to main
Part I · Where we are

01The whole system, plainly

Four layers, each feeding the next — from raw reviews all the way to a card you can post to Instagram.

1 · The engine
Turns opinions into a grade. It reads real reviews (sarcasm-aware), scores each quality along dozens of dimensions, rolls them into six clusters, and lands one letter — re-weighted four ways for four kinds of buyer.
2 · The catalog
Grades every set, ahead of time. A background sweep graded 5,483 sets and stored them, so pages load instantly. Reviews come from blogs, YouTube, and forums.
3 · The app
Shows the grade beautifully. Search, a grade-sorted Browse grid, and an image-forward result card — grade colored by the photo, a brand-voice take, a one-tap "for you" lens switcher.
4 · The share loop
Sends it into the world. Every link unfurls as a branded card; a downloadable 4:5 card is ready to post. Both respect your chosen lens.

02The grading engine

The engine is the heart, and it's the most mature part of the project — built and hardened over weeks before the app existed.

It answers one question — how good is this set, really? — by distilling what actual reviewers said. It's deliberately honest: it detects sarcasm ("oh great, another grey spaceship"), it weights community consensus over lone hot-takes, and when there isn't enough evidence it says so ("Early read") rather than faking confidence.

How a letter gets made

  1. Dimensions.Each review is scored along specific qualities — build fun, display appeal, minifig desirability, whether it feels worth the money, and more. Fact-based dimensions (price, part-out value) come from data, not opinion.
  2. Clusters.Those dimensions roll up into six clusters — the big themes like "the build," "the look," "the value," and "does it hold value."
  3. Lenses.The clusters are weighted four ways. A display collector cares about the look; an investor cares about hold-value (weighted 30% for them); a parent cares about play. Same evidence, four grades.
  4. The letter.A final 0–10 score, banded to a letter (A+ … F), with a confidence read that's always shown — never a bare letter.
Why this matters

Most "best LEGO" lists are one person's taste. The Brick Drop's grade is distilled from many reviews, adjusted for hype and sarcasm, and shown with its receipts. That's the whole credibility of the product — which is why the roadmap leans so hard on proving it (see §08, "Trust the grade").

03The catalog & its reviews

A grading tool is only as good as its coverage. In a few days the catalog went from a demo to nearly complete.

A week ago

28 sets. Enough to prove the app worked, not enough to be useful. Browse was tiny.

→
Today

5,483 sets graded and stored — essentially the entire buildable catalog from 2017 to 2026. Browse is a real, grade-sorted shelf.

Where the opinions come from

A harvest of Brick Insights added 5,883 review references, and confirmed a hunch: there's more written coverage than video.

SourceReferencesWhat it adds
Blog / written3,916The biggest well — detailed, considered reviews the engine distills via article text
YouTube1,711Transcripts + top comments (community consensus, like-weighted)
Forum256Enthusiast back-and-forth, useful for niche/older sets

The sweep also logged ~271 sets it couldn't finish — almost all transient rate-limit blips from the review source, not real gaps. They're re-gradeable on a retry pass.

04The app people use

Everything above is invisible plumbing. This is the part a fan actually touches — and it's polished.

The surfaces
Home — search by name or number + a "Rated" shelf Browse — a grade-sorted grid (best first), each tile colored by its photo Result card — big grade, three reads (price / love / hold), a brand-voice take Lens switcher — one tap to see the grade "for you"
The craft
Color means something — grade tinted by a shade sampled from the set's own image Brand voice — honest, group-chat takes; no emoji, no price-per-piece, ever Real brand font — Avenir Next Condensed, even inside generated share images Installable + shareable — PWA, public pages, branded link previews, 4:5 download

05Just shipped: the investor-value redesign

The newest engine change — and a perfect example of the gap between "shipped" and "live."

The problem it fixes: the "does it hold value?" cluster was a frozen 5.0 for every set. Its two ingredients were both dead — one needed retirement/exclusive flags we never populated, the other needed resale prices we don't have. So the investor grade was uniformly mediocre and meant nothing.

The fix (PR #20): rebuild those two ingredients from data we actually have —

Value Floor

How much of your money is backed by the raw bricks — the set's part-out value ÷ its price. We have part prices for 98% of sets.

Appreciation Outlook

How likely it is to grow — how far into its life it is (age) × how collectible its theme is (a curated, tunable list of themes).

The catch — this is where we go

The engine now computes real hold-value, but the 5,483 grades in the live database were written ~30 minutes before this merged — under the old frozen-5.0 engine. So on the site right now, investor grades still don't reflect the upgrade. The feature is shipped in code but not yet activated in the data — the catalog needs a refresh pass to bring it to life. (Because these two ingredients are computed from stored facts, that refresh can likely be a fast offline recompute, not a full re-mine.)

06Live vs. built-but-not-visible

An honest ledger of what a visitor sees today versus what exists in the code but hasn't reached them yet.

CapabilityIn codeLive on site
Grades for ~5,483 setsYesYes
Four buyer lenses + switcherYesYes
Blog-review depth in gradesYesYes
Investor hold-value (new)YesNot yet — needs re-grade
Smart Price accuracy fixWritten, uncommittedNot yet
Per-lens takes (words match lens)Spec'd onlyNo
Part II · Where we need to go

07Immediate — activate & secure

Before new features, close the gap between what's built and what's live, and back up loose work.

  1. Re-grade (or recompute) the catalog for investor-value.Bring the shipped hold-value upgrade to life across all 5,483 sets. Likely a fast offline recompute from stored facts, not a full re-mine. Highest-value, lowest-effort win — a feature already paid for, just not switched on.
  2. Commit & ship the Smart Price fix.A tested one-file fix (a missing price guide for a promo set now falls back to retail instead of aborting the grade) is sitting uncommitted. Land it — it directly reduces grade failures.
  3. Put the sweep scripts under version control.The tools running the live grading sweep aren't committed. Back them up so they're reviewable and safe.
  4. Retry the ~271 unfinished sets & fix the "null score" save bug.Most failures were rate-limit blips — a retry pass reaches full coverage. And 46 sets that the engine honestly withheld can't save (a database rule rejects a blank score); a small fix lets "not enough to grade yet" persist properly.

08The tracks — everything that's left

DoneIn flightNextLater

A · Trust the grade — prove it's right

The grade's credibility is the whole product. This is the most important track.

In flightSmart Price fix — commit & ship the pricing fail-soft (see §07).
NextValidate the sentiment engine — hand-label a sample of reviews and measure the AI's read against them, so we can retire the "unvalidated" caveat. This is the single biggest un-started credibility item.
NextTune theme desirability — the investor "Appreciation Outlook" leans on a curated list of which themes hold value. Seth's eye on that list makes investor grades sharper.
LaterPrediction model — an honest, labeled "Early read" grade for sets before they release.

B · Cover the catalog — every set, always fresh

DoneFull sweep — 5,483 of ~5,754 buildable sets graded.
In flightCoverage cleanup — retry the ~271 rate-limited stragglers.
NextLiving grades — a routine that re-grades sets as fresh reviews land, so grades never drift and new releases sharpen over time.
NextFull-catalog search — find (and grade-on-demand) any set by name, not just the ones already graded.
LaterNew-release ingestion — auto-pull and grade sets as LEGO announces them; more review sources beyond today's three.

C · Make it personal — all the way down

DonePer-lens grades + switcher — live.
NextPer-lens takes (Slice 17, spec'd) — the words match the lens, not just the number: the investor view talks resale and cites the investor grade. Small, safe, high polish-per-effort.
LaterPer-set tweak — a light "I care more about play value here" adjustment without a full profile.
LaterProfile & a "For You" home — a saved identity (lens, history, watchlist) and curated picks that match it.

D · Beautiful & shareable — craft & reach

DoneShare loop — branded link previews + downloadable 4:5 card, both lens-aware.
NextCrop-perfect share cards — bring the Crop Curator framing into the generated 4:5 card so wide sets sit right (today it uses a generic center-crop).
NextTrue Value, expanded — grow "Is the price fair?" from one line into a real price/parts breakdown.
LaterOnboarding polish & discoverability (SEO + a share-driven growth loop).

09The recommended order

Activate what's paid for, secure loose work, prove the grade, then build depth.

  1. Activate investor-value + ship the Smart Price fix.Two features already built; get them in front of people. Fastest value on the board.
  2. Secure the tooling & reach full coverage.Commit the sweep scripts, retry the stragglers, fix the withheld-grade save bug. Clean foundation.
  3. Ship per-lens takes (Slice 17).Small and safe; closes the gap between a personalized number and personalized words.
  4. Validate the sentiment engine + tune theme desirability.The credibility work. Turn "an AI guessed" into "we checked," and sharpen investor grades with a curated eye.
  5. Then depth: living grades, True Value breakdown, crop-perfect cards.The features that make it richer and more shareable, once the grade underneath is proven and complete.
The headline

The build is done and live — engine, catalog, personalization, share loop. The immediate job isn't to build more, it's to activate what's already built (investor-value, the Smart Price fix) and then prove the grade (validate the sentiment engine). None of it is a rewrite. It's finishing well — and switching on the value that's already sitting in the code.


The Brick Drop · built with receipts, not vibes