Inference of Single-Cell Phylogenies from Lineage Tracing Data [article]

Matthew G Jones, Alex Khodaverdian, Jeffrey J Quinn, Michelle M Chan, Jeffrey A Hussmann, Robert Wang, Chenling Xu, Jonathan S Weissman, Nir Yosef
2019 bioRxiv   pre-print
The pairing of CRISPR/Cas9-based gene editing with massively parallel single-cell readouts now enables large-scale lineage tracing. However, the rapid growth in complexity of data from these assays has outpaced our ability to accurately infer phylogenetic relationships. To address this, we provide three resources. First, we introduce Cassiopeia - a suite of scalable and theoretically grounded maximum parsimony approaches for tree reconstruction. Second, we provide a simulation framework for
more » ... uating algorithms and exploring lineage tracer design principles. Finally, we generate the most complex experimental lineage tracing dataset to date - consisting of 34,557 human cells continuously traced over 15 generations, 71% of which are uniquely marked - and use it for benchmarking phylogenetic inference approaches. We show that Cassiopeia outperforms traditional methods by several metrics and under a wide variety of parameter regimes, and provide insight into the principles for the design of improved Cas9-enabled recorders. Together these should broadly enable large-scale mammalian lineage tracing efforts. Cassiopeia and its benchmarking resources are publicly available at
doi:10.1101/800078 fatcat:7hjtrcn5wvh6zmuca455spmxle