Showing posts with label marine genomes. Show all posts
Showing posts with label marine genomes. Show all posts

Thursday, 31 October 2024

BioDiversity Genomics Conference (BG24)

The global BioDiversity Genomics Conference (BG24) was another great success this year. As well as running a session on Marine Vertebrate Genomics, it was particularly rewarding to see some many quality contributions from lab members and alumni.

Biodiversity Genomics in Australasia

Jessica Pearce (UWA): Reconstructing tiger shark history using genomics

Sharks and rays are a clade of high evolutionary, ecological, economic, and cultural significance, and yet they are one of the most threatened taxa groups in the marine environment. Despite this, there remains a lack of molecular resources for this class to assist with their conservation. The tiger shark (Galeocerdo cuvier) is a near threatened, keystone species distributed circumglobally that is under substantial pressure from human impacts, making it a high priority for management worldwide. We sequenced and characterised a reference quality assembly for the tiger shark, the first genome for this family, and used this to dive deeper into the evolution, adaption, and demographic history of this ancient species. We investigated how its effective population size (Ne), genome-wide heterozygosity and inbreeding has changed over time to infer how this species has responded to past global events, and hence potential responses to ongoing and future accumulating threats. This aims to assist in effective management of this high-profile species. 

Katarina Stuart (University of Auckland): Lifetime fitness is correlated more strongly with structural variant than SNP mutational load in a threatened bird species

Conservation genomics is becoming increasingly interested in whether structural variant (SV) information can help the management of threatened species. The functional consequences of SVs are more complex than for single nucleotide polymorphisms (SNPs) and thus may be more likely to contribute to load. While the impacts of SV-specific genetic load may be less consequential for large populations, the interplay between weakened selection and stochastic processes mean that smaller populations, like those of the threatened Aotearoa hihi/New Zealand stitchbird (Notiomystis cincta), may harbour a high SV load. Hihi were once confined to a single remnant population, but have been reestablished into six sanctuaries and reserves, often via secondary bottlenecks, resulting in low genetic diversity, low adaptive potential and inbreeding depression. In this study, we use whole genome resequencing of 30 individuals from the Tiritiri Matangi population to identify the nature and distribution of both SNPs and SVs within this small avian population. We find that SNP and SV individual mutation load is only moderately correlated, likely because SVs arise in regions of high recombination and reduced evolutionary conservation. Finally, we leverage a long-term monitoring dataset of pedigree and fitness data to assess the impact of SNP and SV mutation load on individual fitness, and demonstrate that SV load correlates more strongly than SNP load with lifetime fitness. The results of this study indicate that only examining SNPs neglects important aspects of intraspecific variation, and that studying SVs has direct implications for linking genetic diversity and genetic health to inform management decisions.

Richard Edwards (UWA): Improving Hifiasm assemblies with 20 kb ONT reads

The quality and quantity of genome assembly has improved dramatically over recent years. Many large-scale genome projects assemble HiFi and HiC reads using Hifiasm to produce contiguous phased assemblies, scaffolded to chromosome-level. Nevertheless, HiFi reads are typically under 25 kb and can still struggle to assemble long, low-diversity repeat regions. Obtaining ‘ultra-long’ (100 kb or longer) ONT reads to solve this problem remains a significant challenge due to technical constraints and DNA sample requirements. Here, we explore the utility of using standard ONT long reads (20 kb or more) as ‘ultra-long’ input to improve phased Hifiasm assemblies for 22 species of bony fish (Genome Size, 627 Mb - 1.54 Gb). We also explore whether the new ‘telo-m’ mode in Hifiasm v0.9.0 improves telomere prediction in these species. Incorporating 20+ kb ONT reads (7.8X - 93.5X) significantly increased assembly contiguity. BUSCO completeness was not significantly altered, although there was some re-partitioning of BUSCO genes between phased haplotypes for some species. Improvement did not strongly correlate with read depth (neither HiFi nor ONT), suggesting that the underlying read length distributions and/or specific genome features are more important for determining the outcome. Hifiasm ‘telo-m’ mode significantly increased telomere recovery, assembling over six times the number of gapless telomere-to-telomere chromosomes when combined with incorporation of ONT reads. Verification of how these results translate to the quality and/or ease of curation of final HiC-scaffolded chromosome-level assemblies is ongoing, with a goal to determine whether the additional sample preparation and sequencing in the lab is cost-effective.

Emma de Jong (UWA): High-Quality Genomes for Australian Lutjanidae Species

Lutjanidae (snappers) are highly valued in commercial and recreational fisheries worldwide and some species serve as fisheries indicator species particularly for bioregions in Western Australia. Comprehensive genomic mapping of immune gene families of Lutjanidae species are lacking, but this information can inform understanding disease vulnerability, the impact of environmental stress, improving aquaculture efforts and to provide insights into the health of wild populations. Despite their importance, only 3 out of 113 Lutjanid species currently have available reference genomes, two of which are highly fragmented (>11,000 and >200,000 contigs), impacting studies on gene families relevant to aquaculture. In this study, we present high-quality chromosome-level reference genomes for 14 Australian lutjanid species across seven genera, generated using PacBio HiFi and Dovetail HiC data. We present initial comparative genomic analyses, including immune gene content and chromosomal synteny analyses across species. These analyses provide insights into the genomic architecture and evolutionary relationships within Lutjanidae. Ongoing work aims to comprehensively map and compare the immune gene family repertoire across genera in Lutjanidae, as well as lethrinid species as an outgroup, to determine genus-specific changes in genes (e.g., loss, selection, duplication) important for pathogen detection, antigen presentation, inflammation, and immune memory. These genome assemblies will serve as a foundational resource to the wider scientific community interested in these species.

Research of ECRs who work in biodiversity genomics

Lara Parata (UWA): Genome Evolution in Marine Ray-Finned Fishes

Approximately half of extant vertebrate species are fishes, with more than 30,000 species classified as ray-finned fishes (Actinopterygii). Actinopterygii represent diverse phenotypes, feeding strategies, life history traits and occupy distinct ecological niches, making them an ideal taxa for studying molecular drivers of diversity and adaptation. Despite their diversity, ecological, and economical importance, only 145 Illumina genome assemblies are available for marine Actinopterygii species. In this study we present 250 new marine Actinopterygii genome assemblies generated using Illumina whole genome sequencing and initial results from a large-scale study of these 395 genomes. Using reference-based annotation tools we determine which fish families have unique patterns of gene family frequency / structure (e.g., losses, expansions, contractions), and correlate these with predicted functional signatures to infer biological and ecological adaptations. We identify fish families with distinct rates of change in the gene families present within their genomes (e.g., more losses / expansions or diversity) and associate these patterns with increased rates of diversification or speciation to further elucidate the genomic attributes contributing to ecological success. The results of this work contribute to the growing understanding of fish genome evolution and provide new insights into the evolutionary history and ecological success of marine Actinopterygii.

Thursday, 29 June 2023

Three Pawsey Internship projects available for the Ocean Genomes Project

We have three Pawsey student internships available this summer with the Ocean Genomes Laboratory in the Minderoo OceanOmics Centre at UWA. Closing date: 07 August, 2023 at 17:00 AWST (Perth time). This is a 10-week, paid program open to exceptional undergrad (2nd/3rd year), Honours, Master’s and PhD students. Apply at the CSIRO Application page. Please get in touch if you want to know more and/or are interested in a student research project in the lab.

Optimising workflows for whole genome assembly for marine vertebrates (Project #04)

The biodiversity of marine vertebrates is critical for the health of our ocean’s ecosystem, but is under immediate threat from climate change, pollution, overfishing and habitat destruction. To advance our understanding of how best to protect and sustain our ocean life, global efforts are underway (such as the Vertebrate Genome Project; VGP) to establish a complete library of high-quality reference genomes for all ~22,000 marine vertebrates.

Reference genomes are pivotal not only for answering fundamental questions in marine biology and evolution, but also for guiding the conservation of species most at risk within our changing oceans, and for accurately monitoring biodiversity.

This project utilizes data generated in-house, either by Illumina short-read or PacBio high-fidelity long-read sequencing of Australian marine vertebrate species. The primary objective is to optimize analysis workflows on Pawsey, encompassing the entire life cycle of the data from its raw format to the ultimate outcome of a high-quality assembled genome. We have data across a diverse range of species covering small to large genome sizes.

A containerised Pawsey workflow for Diploidocus (Project #10)

Bioinformatics in general, and genomics specifically, is replete with complex workflows that do not translate easily to HPC. Frequently, genomics pipelines will incorporate many different tools and/or in-built functions with very different computational requirements in terms of multithreading, memory requirements and IO pressures. The Diploidocus genome curation pipeline exemplifies this problem with some lengthy single-processor steps building on data produced by highly parallelised tools, such as minimap2. As well as adapting a specific mission-critical tool, this project will help identify and establish some general principles for optimising genomics code/workflows for running on Setonix.

Diploidocus is a published genome curation and clean-up tool that utilises several different underlying bioinformatics tools and in-built algorithms. Different steps (and tools) in the pipeline have markedly different CPU, IO and memory requirements, including some lengthy non-parallelised portions. This makes it hard to run efficiently on HPC without wasting resource allocation and/or failing to take advantage of parallelisation when available.

The expected outcome of this project is a Nextflow workflow for the deployment of the Diploidocus pipeline on HPC. This will (a) increase in-house efficiency of HPC usage, and (b) make Diploidocus more attractive as a tool to other research groups.

A containerised Pawsey workflow high throughput phylogenomics (Project #14)

This project aims to produce a robust and efficient phylogenomics workflow for whole genome sequencing data.

One important application of genome assemblies is to test and improve the taxonomic classification of species using large-scale genome-wide phylogenetics, known as phylogenomics. There is a previously developed Snakemake workflow for the rapid generation of phylogenomic trees from low- to mid-coverage whole genome shotgun sequencing data. This pipeline (1) creates multiple rapid draft assemblies; (2) identifies an optimal set of orthologous genes per species using BUSCO and BUSCOMP; (3) generates a multiple sequence alignment per gene; (4) generates a phylogenetic tree per gene; and (5) generates a consensus tree from all the individual gene trees.

There is now a requirement to (1) update the pipeline to be optimised for the high-coverage draft and reference genomes created by the Ocean Genomes Project, and (2) convert this pipeline from PBS/Snakemake to SLURM/Nextflow in-line with other genomics workflows being developed at the Minderoo OceanOmics Centre at UWA.

This project will adapt the wgs2tree workflow to optionally start from a set of existing genome assemblies and BUSCO orthologue annotations and implement a Nextflow/SLURM workflow optimised to run efficiently on Pawsey.

Friday, 21 October 2022

The Ocean Genomes Lab is hiring - Bioinformatics and Sequencing technicians wanted!

Adding to the recently advertised Sequencing technician posts (closing 27 October), we are now pleased to advertise two bioinformatics research assistant positions to support our creation of marine vertebrate reference genome library. If you have experience with genome assembly or bioinformatics workflows, and are passionate about saving marine biodiversity, come and join us!

Two positions are available at Level 5 or 6, depending on your experience. Both roles will be providing bioinformatics support for our marine vertebrate reference genome project. You’ll get to play with data from the latest sequencing toys, including Illumina NovaSeq 6000, NextSeq 2000 and iSeq 100, the PacBio Sequel IIe, and ONT (probably PromethION and MinION).

Job roles will include developing and applying genome assembly workflows, data curation and QC, data sharing, and development/benchmarking of comparative genomics and genome assembly curation tools. If you have experience or passion for integrating bioinformatics workflows with Laboratory Information Management Systems and/or Electronic Laboratory Notebooks, we’d also love to hear to from you. SQL database skills would not go amiss too.

We’re a new team with lots to do, so there is plenty of scope to make the position your own and play to your strengths.

The closing date for applications is 11:55 PM AWST on Thursday 10 November 2022.

To learn more about these opportunities, please click here or contact Rich Edwards at rich.edwards@uwa.edu.au.

ABOUT THE TEAM

The Minderoo OceanOmics Centre at UWA combines a joint Ocean Genomes Laboratory, an OceanOmics Laboratory, and Computational Biology Services.

Equipped with the latest high-throughput sequencing technology and in collaboration with global partners, the Ocean Genomes Laboratory will generate a comprehensive library of high quality marine vertebrate reference genome assemblies. All such reference genome data will be subject to rigorous QA/QC and all assemblies will be released publicly with open access.

The Ocean Genomes Laboratory will undertake research and development under the direction of Minderoo’s ambitious OceanOmics Program which has the goal of revolutionising ocean conservation through novel marine sampling and genomics approaches and scaling these to significantly advance our knowledge of marine life. The Ocean Genomes Laboratory and Computational Biology Services will include state of the art infrastructure including sample and eDNA preparation areas, flow cytometry, single cell sequencing equipment and the latest bioinformatics and computational biology tools.

Thursday, 6 October 2022

The Ocean Genomes Laboratory is hiring!

The Minderoo OceanOmics Centre at UWA Ocean Genomes Laboratory is now hiring our technical team to support high throughput DNA sequencing and genome assembly. We currently have three "wet" lab positions going: a Sequencing Specialist Scientific Officer, and two Sequencing Technician positions. Both roles will be providing technical support in the lab, particularly with respect to all aspects of DNA sequencing (sample extraction, library preparation and setting up sequencing runs). You'll get to play with the latest sequencing toys, including Illumina NovaSeq 6000, NextSeq 2000 and iSeq 100, the PacBio Sequel IIe, and ONT (probably PromethION and MinION).

The closing date for applications is 11:55 PM AWST on Thursday 27 October 2022.

To learn more about these opportunities, please click on the links above or contact Rich Edwards at rich.edwards@uwa.edu.au. We will also be advertising some bioinformatics positions soon.

About the team

The Minderoo OceanOmics Centre at UWA combines a joint Ocean Genomes Laboratory, an OceanOmics Laboratory, and Computational Biology Services.

Equipped with the latest high-throughput sequencing technology and in collaboration with global partners, the Ocean Genomes Laboratory will generate a comprehensive library of high quality marine vertebrate reference genome assemblies. All such reference genome data will be subject to rigorous QA/QC and all assemblies will be released publicly with open access.

The Ocean Genomes Laboratory will undertake research and development under the direction of Minderoo’s ambitious OceanOmics Program which has the goal of revolutionising ocean conservation through novel marine sampling and genomics approaches and scaling these to significantly advance our knowledge of marine life. The Ocean Genomes Laboratory and Computational Biology Services will include state of the art infrastructure including sample and eDNA preparation areas, flow cytometry, single cell sequencing equipment and the latest bioinformatics and computational biology tools.

Monday, 8 August 2022

Senior Postdoc wanted for UWA Ocean Genomes Lab! (Closing soon)

The new Ocean Genomes Laboratory (part of the Minderoo OceanOmics Centre at the UWA Oceans Institute) is hiring a Level B postdoc in marine genomics. (Three-year fixed term full time role, or flexible working equivalent.)

This is a rare opportunity to work as part of a collaborative team in a high-profile state of the art genomics research facility dedicated to studying marine vertebrates. You should have a PhD in bioinformatics, computational biology, molecular genetics or genomics, plus an interest in marine vertebrates and postdoctoral experience in high throughput DNA sequencing and whole genome assembly. The lab is new and there is plenty of scope to shape its direction beyond the core mission creating a marine vertebrate reference genome library as part of the Vertebrate Genome Project. You will also have an important role in helping to supervise the lab staff and research team.

Closing date: 11:55pm AWST, Friday 12 August 2022

Please see the UWA job advert for more details.

About the team

The Minderoo OceanOmics Centre at UWA combines a joint Ocean Genomes Laboratory, an OceanOmics Laboratory, and a Computational Biology Program.

Equipped with the latest high-throughput sequencing technology, and in collaboration with global partners, the Ocean Genomes Laboratory will generate a comprehensive library of high-quality marine vertebrate reference genome assemblies. All reference genome data will be subject to rigorous QA/QC and all assemblies will be released publicly through open access.

The OceanOmics Centre will be located in the Bayliss Building on the UWA Crawley Campus, OceanOmics staff sharing the building with research and teaching staff primarily from the UWA School of Molecular Sciences and interacting with staff in the UWA Oceans Institute in the nearby IOMRC building.

About the opportunity

As a Research Fellow you will join a research group committed to applying modern molecular biological methods to marine research.

Using modern genomic approaches, you will undertake research on marine vertebrates, focussed on the production, QC and assembly of high-quality reference genome data. You will participate in the entire workflow from sample collection and processing, generating genomic sequence data in the laboratory using multiple modern genome sequencing technologies, with a focus on data processing, assembly, curation, analysis and dissemination.

In this unique role you will also be supported to develop your leadership skills. Working closely with the Centre’s UWA Principal Research Fellow, junior postdoctoral academics, the Centre’s Laboratory Manager, and diverse researchers from Minderoo Foundation you will contribute to decision making, oversee the work of technicians and PhD students and provide leadership in modern high-quality genome assembly production and publication.

Friday, 22 July 2022

The Edwards Lab is moving to the UWA Oceans Institute!

More details will follow but, in August, I will be starting a new position at the University of Western Australia Oceans Institute to head up the new Ocean Genomes Laboratory as part of the Minderoo OceanOmics Centre. This exciting project will collaborate closely with the Minderoo Foundation, the Vertebrate Genome Project, and scientists across Australia to create marine vertebrate reference genomes.

The goal of the Ocean Genomes Lab is "building and openly publishing the reference libraries for marine vertebrates ... to accurately detect, monitor and determine the health of these species". The lab is still being setup and we're hiring. Currently available is a Level B postdoc positions: http://bit.ly/OceanOmics. If building genomes is your thing, and you want to help fight the biodiversity crisis in our oceans, come and work with me! (Or pass it on if you know someone who does!) Research Assistant positions will follow.

Look out for a bunch of updates over the next few weeks, both as I update some of the outstanding presentations and posters from this year, and as the website rebrands. In the meantime, please get in touch if any of this sounds interesting!