Dataset Anatomy — prereq_op_selector

1000
runs extracted
1297
accepted rows
24
schema-drift rejected
~17
unique text templates
Class balance
needs_resolve_dependencies = yes
997 rows
76.9%
needs_resolve_dependencies = no
300 rows
23.1%
needs_architect_review = yes
257 rows
19.8%
needs_architect_review = no
1040 rows
80.2%
⚠ Corpus limitation
Only ~17 unique text templates across 1297 rows. High row count masks low actual diversity.
The taskitem schema records outputs, not discriminative inputs. RSA-relevant fields were never stored.
Baselines are intentionally weak. Pipeline is correct; signal improves with schema changes upstream.
Fabricate TSV format — prereq_resolve_train.tsv (sample rows)
label
hop
text (title · reasons · accepts · constraints)
1
0
title: Execution Readiness Primary | reasons: high_confidence_clear_path | accepts: at least one tool selection event; no malformed mcp payloads
0
0
title: Intake Foundation Primary | reasons: high_confidence_clear_path | accepts: deterministic taskitems with strict execution contracts | constraints: primary language go
1
0
title: Intake Foundation Primary | reasons: high_confidence_clear_path | accepts: scenario 1 given when then; scenario 2 given when then | constraints: language runtime constraints
0
0
title: Execution Readiness Primary | reasons: high_confidence_clear_path | accepts: at least one tool execution event; no unbounded loops