Why
The 2026-06-28 audit's field/opportunity research (Code/Audits/Inv-Man-Intake/2026-06-28-05-field-and-opportunities.md)
found the contracts and the explainable-scoring stance are at/ahead of the institutional manager-DD field, but the
highest-leverage gaps are adjacent problems the existing machinery almost solves. This epic tracks them so
dimensions 5–7 (approach-vs-field / missed-opportunities / tools) are not lost; each child is a separately-scoped
2-week issue. This is a roadmap epic, not a single mergeable change.
Scope
Track and sequence the four highest-leverage opportunities. Each becomes its own AGENT_ISSUE_FORMAT child issue
with its own test/verification gate before implementation.
Non-Goals
- This epic itself ships no code; do NOT implement everything in one PR.
- Do NOT claim extraction accuracy/throughput on fixtures (the README disclaimer is correct) — any accuracy claim
must be backed by a real-document eval set.
- Scaffold-only completion does NOT count: closing this epic requires the child issues to be filed (and the
peer-group child to land with a test), not a single placeholder.
Tasks (each → its own child issue)
Acceptance Criteria
Implementation Notes
- Source analysis with citations:
2026-06-28-05-field-and-opportunities.md (commercial peers: Canoe, Accelex,
Backstop, Dynamo; tools: Docling, instructor, DeepEval).
- Sequence by leverage: peer-group → HITL calibration → lineage packet → real extraction.
Why
The 2026-06-28 audit's field/opportunity research (
Code/Audits/Inv-Man-Intake/2026-06-28-05-field-and-opportunities.md)found the contracts and the explainable-scoring stance are at/ahead of the institutional manager-DD field, but the
highest-leverage gaps are adjacent problems the existing machinery almost solves. This epic tracks them so
dimensions 5–7 (approach-vs-field / missed-opportunities / tools) are not lost; each child is a separately-scoped
2-week issue. This is a roadmap epic, not a single mergeable change.
Scope
Track and sequence the four highest-leverage opportunities. Each becomes its own AGENT_ISSUE_FORMAT child issue
with its own test/verification gate before implementation.
Non-Goals
must be backed by a real-document eval set.
peer-group child to land with a test), not a single placeholder.
Tasks (each → its own child issue)
(
scoring/weights.py:LAUNCH_ASSET_CLASSES) and component metrics already exist; add a cohort store(DuckDB/SQLite) and emit a percentile rank vs. asset-class cohort alongside the absolute 0–1 score. Child issue
must include a deterministic test on a seeded cohort.
CorrectionRecord+ escalations are captured but never measured. Builda job that joins corrections to extractions → field-level precision/recall + confidence calibration, used to
tune the
0.85/0.75/0.60thresholds inconfig/extraction_thresholds.yamlwith evidence.one-shot per-decision lineage export (no new capture needed).
implementation +
instructorfor schema-valid LLM output + DeepEval (PR-gate) / LangSmith datasets (drift).Keep the stlite demo fixture-backed.
Acceptance Criteria
test/verification gate, and linked back to this epic.
(
tests/scoring/test_peer_group.py::test_percentile_rank_against_seeded_cohort) as its gate.Code/Audits/Inv-Man-Intake/2026-06-28-05-field-and-opportunities.mdas the source analysis.Implementation Notes
2026-06-28-05-field-and-opportunities.md(commercial peers: Canoe, Accelex,Backstop, Dynamo; tools: Docling, instructor, DeepEval).