Introduction to dorado as a basecaller for Nanopore data.
Oxford Nanopore sequencing is a relatively unique sequencing technology that can be used to generate ultra long-read DNA or RNA data reads. Raw sequencing files are stored as POD5 files which are a recording of fluctuating electrical current signals generated as DNA or RNA strands pass through the microscopic pores of the machine. Translating this signal into standard sequencing formats (BAM or FASTQ) requires a basecalling algorithm. This workshop introduces Dorado, Oxford Nanopore’s basecalling software, which uses machine learning models to decode the sequencing signal and produce nucleotide base data. In this workshop, we will discuss running Dorado on Quest’s GPUs, interpreting the output, and some of the additional functionality it provides beyond basecalling.
This workshop requires a laptop with a terminal application installed and expects familiarity with working from the command line or attendance at the ‘Working from the Command Line’ workshop.
Audience
- Faculty/Staff
- Student
- Post Docs/Docs
- Graduate Students
Contact
Leticia Vega
Email
Interest
- Academic (general)
- Data Science & AI