Rederivation benchmark: can the engine re-derive our follows' best calls?
Rederivation benchmark: can the engine re-derive our follows' best calls?
Question: Could our own engine have derived the best calls our Twitter/Discord follows made, from our own data, at or before the date they published? 10 diversity-selected cases, graded LEAD / MATCH / PARTIAL / MISS. Verdict: assign-follow-up — the tape and fundamentals halves of the calls were consistently in our own data at or before publish; the misses cluster in four nameable gaps, each now a filed task.
Method
As-of tape replay: for each case, per-symbol change7d/30d/3m, acceleration
(7d pace vs 30d weekly pace, the insight-wire formula), distance from
trailing-252d high, and relative volume were computed at the trading day
before the source's publish date, from
research/market-engine/data/stocks/<SYM>/ohlc/12m.json (252 daily bars,
close/volume/sma20/sma50). Basket relative strength = basket median 30d minus
SPY 30d from the same files. Detector verdicts apply the live wire thresholds
(accel > 3, quiet-turn week >= 5 on a negative month, fresh-high within 3% of
the 252d high, strength >= 15%/30d). Caveat named honestly: this proves
derivability from data we carry, not that the wires ran on those dates —
the wires were built 2026-07-02/03.
Scorecard
| # | Case (source, date) | Grade | One-line receipt |
|---|---|---|---|
| 1 | Memorable-3 memory > Mag7 (@jiahanjimliu 06-28) | LEAD (tape+fundamentals) / PARTIAL (2027 profit model) | At 06-27: memory basket RS +13.2 vs SPY; MU +218.7%/3m vs NVDA -9.3%/30d; fundamentals-wire live-fired MU/STX/WDC revenue-accel |
| 2 | OUST industrial-lidar spine (@jiahanjimliu 06-03) | PARTIAL — coverage gap | At 06-02: OUST +74.0%/30d, at its 252d high, RS +68.6 — flaggable, but OUST was in no watchlist until robotics seeded 06-27; founder-positioning half is agent-read tier |
| 3 | Mag7 FCF collapse + rotation out (Apollo 06-28) | PARTIAL — stale fundamentals + missing detector | At 06-27: every Mag7 name negative 30d, basket RS -10.6 (rotation half = derivable); FCF-collapse half blocked by stale on-disk statements for mega-caps and no capex-ratio detector |
| 4 | Actuator/reducer humanoid chokepoint (MS 06-27) | MISS — structural | Harmonic Drive/Nabtesco are uncovered Japan listings; layer maps are research-trace tier, no trace had run on the actuator layer |
| 5 | Commercial electricity > residential 2027 (@ChurchillWw 05-19) | MISS (expected) | EIA series — no macro-data intake exists |
| 6 | Solar $54-82/MWh LCOE (@ChurchillWw 05-18) | MISS (expected) | IRENA/BNEF-class data — same macro gap |
| 7 | Memory +276%/3m, healthcare held (rotation-map 06-26, tweet later) | LEAD | At 06-25: MU and SNDK both accelerating + at fresh highs (accel +8.0/+8.3), memory RS +30.7; healthcare RS +9.5 with JNJ/XBI/IBB accelerating + fresh-high — both rotations mechanically flagged pre-tweet |
| 8 | HLIT hardware→software re-rate + breakout (@CKCapitalxx 05-23) | LEAD (tape) / PARTIAL (story) | At 05-22: HLIT +45.7%/30d, accel +10.5, at its high, 6.12x average volume — accelerating + fresh-high + volume-spike all fire; the segment-mix story needs a filing read (filing-diff exists now) |
| 9 | 8-layer HBM bottleneck map (blind test 05-25) | PARTIAL | At 05-24: KLIC fresh-high +23.3%/30d, BESIY fresh-high at 3.55x volume — 2 of 8 layer names tape-flagged; the layer MAP itself is physics/trace tier, not tape-derivable |
| 10 | Physical-AI optical supercycle (@crux_capital_ 05-27) | PARTIAL | At 05-26: CIEN accelerating + fresh-high +75.8%/3m, GLW accelerating, basket RS +8.4 — convergence visible; the 11-layer book structure is trace tier |
Totals: 3 LEAD-class, 4 PARTIAL, 3 MISS (2 by construction).
What the pattern says
- The tape half of nearly every follow call was in our data first. Cases 1, 7, 8: our detectors (as now built) would have fired before the tweet — including the single most "found" call in the set (HLIT: a 6x-volume breakout at a fresh high with +10.5 acceleration is a textbook wire hit). The follows' edge on these was surfacing, not information — consistent with the standing front-run finding.
- The story half is agent-read tier, and that's fine. Segment-mix re-rates (HLIT), founder positioning (OUST), layer books (crux) need a narrative read of filings/transcripts — this is exactly what the new deep-read queue routes: the wire flags the symbol, the agent does the read.
- The real gaps are nameable and small:
- Coverage/discovery (case 2): OUST screamed for 3 weeks before we covered it. Names outside watchlists are invisible to the wires; monster-discover exists but its cadence and hand-off didn't catch this.
- Mega-cap fundamentals freshness + one missing detector (case 3): hyperscaler statements were stale on disk, and no detector watches FCF-collapse / capex-as-%-of-OCF — the Apollo claim was free public data.
- Structural layer maps (cases 4, 9, 10): research-trace machinery exists but runs ad-hoc; the misses were un-run traces, not missing tooling.
- Macro series (cases 5, 6): EIA/IRENA-class data has no intake path; sized as its own feasibility-gated decision.
Follow-ups filed (see ops/tasks/TASKS-ENGINE.md + research/tasks/)
- capex-ratio / FCF-trend red-flag detector for fundamentals-wire (#cross-wire lane)
- mega-cap statement freshness in the massive refresh path
- discovery hand-off hardening: monster-discover hits above an RS bar should open coverage rows
- research-trace refresh cadence for active perspective layers
- macro-series intake (EIA first) — feasibility-gated, decision row not build row
Related
8 eventsNo direct external sources are attached to this read.