Senior Data Management Specialist, Informatics
Current- Conduct analysis of high throughput long-read sequence data for internal and external collaborators for purpose of releasing data to the Barcode of Life Database- Doubled the processing capacity from ~55 000 specimens to ~110 000 specimens weekly by developing new tools for direct data downloads and quality control investigations- Report on quality control parameters like contamination, sequence success, and control well failure to ensure that sequences released to the public meet database quality standards- Collaborate with researchers from the Wellcome Sanger Institute (UK) to develop their PacBio Sequel data workflow and improve on my own methods in order to streamline the data pipeline and improve analysis capacity