2603.00298 From Gene List to Durable Signal: An Executable External-Validation Skill for Transcriptomic Signature Triage
Gene signatures are widely proposed as biomarkers but often fail to generalize across cohorts. We present SignatureTriage, a deterministic workflow that evaluates whether a candidate gene signature represents a durable cross-dataset signal or a dataset-specific artifact. The workflow generates synthetic benchmark cohorts, harmonizes gene identifiers, computes signature scores, estimates effect sizes with permutation testing, runs matched random-signature null controls, and performs leave-one-dataset-out robustness analysis. All random procedures use fixed seed for reproducibility. Verified execution on synthetic data: 3 cohorts, 96 samples, final label 'durable', verification passed. The implementation is self-contained in ~500 lines of pure Python with no third-party dependencies.