Day 3 of 78 · Run #2 · 2026-08-22
First real cell-eval run — mean-shift floor vs VCC 2025 validation (H1 hESC, 50 targets)
Audit & metrics DE gene recall 0% ceiling · Pearson Δ undefined
No audit flags.
Metrics vs ceiling (all scores)
Differential expression volcano plot (log2FC vs -log10 p)
Pre-registered hypotheses (2)
- Zero-shift floor: DESigGenesRecall lands at or near zero — a constant prediction has no DE genes to recall
- Mean absolute error will beat replicate-naive intuition but pearson_delta will sit far below the computed ceiling
Evidence Literature & entity enrichment
Literature
0 genesTavily · auxiliary, not scored
Literature pending — run tools/enrich_literature.py.
Field context
VCC / perturbation researchField research pending — run tools/enrich_newsroom.py.
Biomedical NER
0 entitiesPioneer GLiNER2 · fine-tuned on Tavily literature when available · regex fallback offline
NER pending — run tools/pioneer_ner.py --train once, then tools/pioneer_ner.py --run experiments/<run-id>.
Narrative digest run digest · traces to facts.json
> Fallback digest rendered deterministically from `facts.json` (no LLM call).
Full digest
Headline
First real cell-eval run — mean-shift floor vs VCC 2025 validation (H1 hESC, 50 targets)
Metrics
| metric | value | ceiling | |---|---|---| | de_sig_genes_recall | 0.0 | 0.49358 | | pearson_delta | None | 0.667108 |
Provenance
- commit: `86a5b5b0f23c294216ecc0bf2513ed7c10af2eac` - seed: 0 - code hash: `k002-mean-shift-vcc2025-v0` - hypotheses pre-registered: ['Zero-shift floor: DESigGenesRecall lands at or near zero — a constant prediction has no DE genes to recall', 'Mean absolute error will beat replicate-naive intuition but pearson_delta will sit far below the computed ceiling']
Trust & provenance Self-tests + reproduce command
6 pipeline steps · sourced from committed artifacts only
Evaluated 28 cell-eval metrics vs 2 ceiling bounds · data_status=final.
Ran 5 deterministic rules · 0 flags raised (none fired).
No audit-flagged genes to search.
No literature files to enrich — run enrich_literature.py first.
Deterministic fallback digest (no API key or LLM call failed).
Passed: planted-signal 13/13 caught · holo-agent 4/4 passed · narrative-check 4/4 checks passed.