Rederivation benchmark: can the engine re-derive our follows' best calls?

Investigation

Rederivation benchmark: can the engine re-derive our follows' best calls?

Question: Could our own engine have derived the best calls our Twitter/Discord follows made, from our own data, at or before the date they published? 10 diversity-selected cases, graded LEAD / MATCH / PARTIAL / MISS. Verdict: assign-follow-up — the tape and fundamentals halves of the calls were consistently in our own data at or before publish; the misses cluster in four nameable gaps, each now a filed task.

Method

As-of tape replay: for each case, per-symbol change7d/30d/3m, acceleration (7d pace vs 30d weekly pace, the insight-wire formula), distance from trailing-252d high, and relative volume were computed at the trading day before the source's publish date, from research/market-engine/data/stocks/<SYM>/ohlc/12m.json (252 daily bars, close/volume/sma20/sma50). Basket relative strength = basket median 30d minus SPY 30d from the same files. Detector verdicts apply the live wire thresholds (accel > 3, quiet-turn week >= 5 on a negative month, fresh-high within 3% of the 252d high, strength >= 15%/30d). Caveat named honestly: this proves derivability from data we carry, not that the wires ran on those dates — the wires were built 2026-07-02/03.

Scorecard

# Case (source, date) Grade One-line receipt
1 Memorable-3 memory > Mag7 (@jiahanjimliu 06-28) LEAD (tape+fundamentals) / PARTIAL (2027 profit model) At 06-27: memory basket RS +13.2 vs SPY; MU +218.7%/3m vs NVDA -9.3%/30d; fundamentals-wire live-fired MU/STX/WDC revenue-accel
2 OUST industrial-lidar spine (@jiahanjimliu 06-03) PARTIAL — coverage gap At 06-02: OUST +74.0%/30d, at its 252d high, RS +68.6 — flaggable, but OUST was in no watchlist until robotics seeded 06-27; founder-positioning half is agent-read tier
3 Mag7 FCF collapse + rotation out (Apollo 06-28) PARTIAL — stale fundamentals + missing detector At 06-27: every Mag7 name negative 30d, basket RS -10.6 (rotation half = derivable); FCF-collapse half blocked by stale on-disk statements for mega-caps and no capex-ratio detector
4 Actuator/reducer humanoid chokepoint (MS 06-27) MISS — structural Harmonic Drive/Nabtesco are uncovered Japan listings; layer maps are research-trace tier, no trace had run on the actuator layer
5 Commercial electricity > residential 2027 (@ChurchillWw 05-19) MISS (expected) EIA series — no macro-data intake exists
6 Solar $54-82/MWh LCOE (@ChurchillWw 05-18) MISS (expected) IRENA/BNEF-class data — same macro gap
7 Memory +276%/3m, healthcare held (rotation-map 06-26, tweet later) LEAD At 06-25: MU and SNDK both accelerating + at fresh highs (accel +8.0/+8.3), memory RS +30.7; healthcare RS +9.5 with JNJ/XBI/IBB accelerating + fresh-high — both rotations mechanically flagged pre-tweet
8 HLIT hardware→software re-rate + breakout (@CKCapitalxx 05-23) LEAD (tape) / PARTIAL (story) At 05-22: HLIT +45.7%/30d, accel +10.5, at its high, 6.12x average volume — accelerating + fresh-high + volume-spike all fire; the segment-mix story needs a filing read (filing-diff exists now)
9 8-layer HBM bottleneck map (blind test 05-25) PARTIAL At 05-24: KLIC fresh-high +23.3%/30d, BESIY fresh-high at 3.55x volume — 2 of 8 layer names tape-flagged; the layer MAP itself is physics/trace tier, not tape-derivable
10 Physical-AI optical supercycle (@crux_capital_ 05-27) PARTIAL At 05-26: CIEN accelerating + fresh-high +75.8%/3m, GLW accelerating, basket RS +8.4 — convergence visible; the 11-layer book structure is trace tier

Totals: 3 LEAD-class, 4 PARTIAL, 3 MISS (2 by construction).

What the pattern says

  1. The tape half of nearly every follow call was in our data first. Cases 1, 7, 8: our detectors (as now built) would have fired before the tweet — including the single most "found" call in the set (HLIT: a 6x-volume breakout at a fresh high with +10.5 acceleration is a textbook wire hit). The follows' edge on these was surfacing, not information — consistent with the standing front-run finding.
  2. The story half is agent-read tier, and that's fine. Segment-mix re-rates (HLIT), founder positioning (OUST), layer books (crux) need a narrative read of filings/transcripts — this is exactly what the new deep-read queue routes: the wire flags the symbol, the agent does the read.
  3. The real gaps are nameable and small:
    • Coverage/discovery (case 2): OUST screamed for 3 weeks before we covered it. Names outside watchlists are invisible to the wires; monster-discover exists but its cadence and hand-off didn't catch this.
    • Mega-cap fundamentals freshness + one missing detector (case 3): hyperscaler statements were stale on disk, and no detector watches FCF-collapse / capex-as-%-of-OCF — the Apollo claim was free public data.
    • Structural layer maps (cases 4, 9, 10): research-trace machinery exists but runs ad-hoc; the misses were un-run traces, not missing tooling.
    • Macro series (cases 5, 6): EIA/IRENA-class data has no intake path; sized as its own feasibility-gated decision.

Follow-ups filed (see ops/tasks/TASKS-ENGINE.md + research/tasks/)

  • capex-ratio / FCF-trend red-flag detector for fundamentals-wire (#cross-wire lane)
  • mega-cap statement freshness in the massive refresh path
  • discovery hand-off hardening: monster-discover hits above an RS bar should open coverage rows
  • research-trace refresh cadence for active perspective layers
  • macro-series intake (EIA first) — feasibility-gated, decision row not build row
8 events

No direct external sources are attached to this read.