<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/38561?offset=530</link>
	<atom:link href="https://bioinformaticsonline.com/related/38561?offset=530" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/40302/simug-a-general-purpose-genome-simulator</guid>
	<pubDate>Thu, 28 Nov 2019 04:33:18 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/40302/simug-a-general-purpose-genome-simulator</link>
	<title><![CDATA[simuG: a general-purpose genome simulator]]></title>
	<description><![CDATA[<p><span>Simulated genomes with pre-defined and random genomic variants can be very useful for benchmarking genomic and bioinformatics analyses. Here we introduce simuG, a lightweight tool for simulating the full-spectrum of genomic variants (single nucleotide polymorphisms, Insertions/Deletions, copy number variants, inversions and translocations) for any organisms (including human). The simplicity and versatility of simuG make it a unique general-purpose genome simulator for a wide-range of simulation-based applications.</span></p><p>Address of the bookmark: <a href="https://github.com/yjx1217/simuG" rel="nofollow">https://github.com/yjx1217/simuG</a></p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/researchlabs/view/40881/liu-lab</guid>
  <pubDate>Tue, 04 Feb 2020 06:27:02 -0600</pubDate>
  <link></link>
  <title><![CDATA[Liu Lab]]></title>
  <description><![CDATA[
<p>Shirley is a computational biologist with expertise in cancer epigenetics. Her research focuses on algorithm development and integrative mining from big data generated on microarrays, massively parallel sequencing, and other high throughput techniques to model the specificity and function of transcription factors, chromatin regulators and lncRNAs in tumor development, progression, drug response and resistance.</p>

<p>https://liulab-dfci.github.io/software/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/41493/coronavirus-resources</guid>
	<pubDate>Wed, 25 Mar 2020 17:11:33 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/41493/coronavirus-resources</link>
	<title><![CDATA[Coronavirus Resources !]]></title>
	<description><![CDATA[<p><span>2019nCoVR features comprehensive integration of genomic and proteomic sequences as well as their metadata information from the GISAID, NCBI, NMDC and CNCB/NGDC. It also incorporates a wide range of relevant information including scientific literatures, news, and popular articles for science dissemination, and provides visualization functionalities for genome variation analysis results based on all collected 2019-nCoV strains.</span></p>
<p><span>Annotation</span></p>
<p><span><a href="https://bigd.big.ac.cn/ncov/variation/annotation">https://bigd.big.ac.cn/ncov/variation/annotation</a></span></p>
<p><span>Genome wharehouse&nbsp;</span></p>
<p><span><a href="https://bigd.big.ac.cn/gwh/browse/index">https://bigd.big.ac.cn/gwh/browse/index</a></span></p>
<p>Released Genome</p>
<p><a href="https://bigd.big.ac.cn/ncov/release_genome">https://bigd.big.ac.cn/ncov/release_genome</a></p>
<p>Download data&nbsp;</p>
<p><a href="ftp://download.big.ac.cn/Genome/Viruses/Coronaviridae/">ftp://download.big.ac.cn/Genome/Viruses/Coronaviridae/</a></p>
<p>Raw data</p>
<p><a href="https://bigd.big.ac.cn/gsa/browse/run/?tag=Coronaviridae">https://bigd.big.ac.cn/gsa/browse/run/?tag=Coronaviridae</a></p><p>Address of the bookmark: <a href="https://bigd.big.ac.cn/ncov/about" rel="nofollow">https://bigd.big.ac.cn/ncov/about</a></p>]]></description>
	<dc:creator>Neel</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/researchlabs/view/42900/svardal-lab</guid>
  <pubDate>Sat, 20 Feb 2021 10:01:19 -0600</pubDate>
  <link></link>
  <title><![CDATA[Svardal lab]]></title>
  <description><![CDATA[
<p>In the Svardal lab they are interested how the astonishing natural diversity we see on earth came into being, by which forces it formed and how it is changing today. Hence, they are trying to understand the process of evolution, with mathematical models and through the analysis of genome sequencing data.</p>

<p>Genomes, and in particular differences between them, are a crucial source of information to understand evolution and biology in general. They provide a record of the evolutionary past of populations, their relatedness patterns, their demography, and their adaptations.</p>

<p>More at https://www.uantwerpen.be/en/staff/hannes-svardal/svardal-lab/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/43057/hapsolo-an-optimization-approach-for-removing-secondary-haplotigs-during-diploid-genome-assembly-and-scaffolding</guid>
	<pubDate>Sat, 08 May 2021 21:25:00 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/43057/hapsolo-an-optimization-approach-for-removing-secondary-haplotigs-during-diploid-genome-assembly-and-scaffolding</link>
	<title><![CDATA[HapSolo: An optimization approach for removing secondary haplotigs during diploid genome assembly and scaffolding]]></title>
	<description><![CDATA[<p><span>HapSolo, that identifies secondary contigs and defines a primary assembly based on multiple pairwise contig alignment metrics. HapSolo evaluates candidate primary assemblies using BUSCO scores and then distinguishes among candidate assemblies using a cost function. The cost function can be defined by the user but by default considers the number of missing, duplicated and single BUSCO genes within the assembly. HapSolo performs hill climbing to minimize cost over thousands of candidate assemblies.&nbsp;</span></p><p>Address of the bookmark: <a href="https://github.com/esolares/HapSolo" rel="nofollow">https://github.com/esolares/HapSolo</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/researchlabs/view/43293/josefa-gonzalez-lab</guid>
  <pubDate>Thu, 19 Aug 2021 08:52:56 -0500</pubDate>
  <link></link>
  <title><![CDATA[Josefa González Lab]]></title>
  <description><![CDATA[
<p>Lab focus on understanding how organisms adapt to their environments. They combine omics approaches with detailed molecular and phenotypic analyses to get a comprehensive picture of adaptation. Our aim at being internationally recognized as a leading lab in the field of environmental adaptation.<br />Lab share our passion for science with the general public by leading outreach projects aimed at increasing science awareness.</p>

<p>More at https://www.biologiaevolutiva.org/gonzalez_lab/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/43634/illumina-based-assembly-pipeline-steps</guid>
	<pubDate>Fri, 10 Dec 2021 06:22:54 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/43634/illumina-based-assembly-pipeline-steps</link>
	<title><![CDATA[Illumina based assembly pipeline steps !]]></title>
	<description><![CDATA[<h3 id="illumina">Illumina<a href="https://nf-co.re/viralrecon#illumina"><span></span></a></h3><ol>
<li>Merge re-sequenced FastQ files (<a href="http://www.linfo.org/cat.html"><code>cat</code></a>)</li>
<li>Read QC (<a href="https://www.bioinformatics.babraham.ac.uk/projects/fastqc/"><code>FastQC</code></a>)</li>
<li>Adapter trimming (<a href="https://github.com/OpenGene/fastp"><code>fastp</code></a>)</li>
<li>Removal of host reads (<a href="http://ccb.jhu.edu/software/kraken2/"><code>Kraken 2</code></a>; <em>optional</em>)</li>
<li>Variant calling<ol>
<li>Read alignment (<a href="http://bowtie-bio.sourceforge.net/bowtie2/index.shtml"><code>Bowtie 2</code></a>)</li>
<li>Sort and index alignments (<a href="https://sourceforge.net/projects/samtools/files/samtools/"><code>SAMtools</code></a>)</li>
<li>Primer sequence removal (<a href="https://github.com/andersen-lab/ivar"><code>iVar</code></a>; <em>amplicon data only</em>)</li>
<li>Duplicate read marking (<a href="https://broadinstitute.github.io/picard/"><code>picard</code></a>; <em>optional</em>)</li>
<li>Alignment-level QC (<a href="https://broadinstitute.github.io/picard/"><code>picard</code></a>, <a href="https://sourceforge.net/projects/samtools/files/samtools/"><code>SAMtools</code></a>)</li>
<li>Genome-wide and amplicon coverage QC plots (<a href="https://github.com/brentp/mosdepth/"><code>mosdepth</code></a>)</li>
<li>Choice of multiple variant calling and consensus sequence generation routes (<a href="https://github.com/andersen-lab/ivar"><code>iVar variants and consensus</code></a>; <em>default for amplicon data</em> <em>||</em> <a href="http://samtools.github.io/bcftools/bcftools.html"><code>BCFTools</code></a>, <a href="https://github.com/arq5x/bedtools2/"><code>BEDTools</code></a>; <em>default for metagenomics data</em>)
<ul>
<li>Variant annotation (<a href="http://snpeff.sourceforge.net/SnpEff.html"><code>SnpEff</code></a>, <a href="http://snpeff.sourceforge.net/SnpSift.html"><code>SnpSift</code></a>)</li>
<li>Consensus assessment report (<a href="http://quast.sourceforge.net/quast"><code>QUAST</code></a>)</li>
<li>Lineage analysis (<a href="https://github.com/cov-lineages/pangolin"><code>Pangolin</code></a>)</li>
<li>Clade assignment, mutation calling and sequence quality checks (<a href="https://github.com/nextstrain/nextclade"><code>Nextclade</code></a>)</li>
<li>Individual variant screenshots with annotation tracks (<a href="https://asciigenome.readthedocs.io/en/latest/"><code>ASCIIGenome</code></a>)</li>
</ul>
</li>
<li>Intersect variants across callers (<a href="http://samtools.github.io/bcftools/bcftools.html"><code>BCFTools</code></a>)</li>
</ol></li>
<li><em>De novo</em> assembly<ol>
<li>Primer trimming (<a href="https://cutadapt.readthedocs.io/en/stable/guide.html"><code>Cutadapt</code></a>; <em>amplicon data only</em>)</li>
<li>Choice of multiple assembly tools (<a href="http://cab.spbu.ru/software/spades/"><code>SPAdes</code></a> <em>||</em> <a href="https://github.com/rrwick/Unicycler"><code>Unicycler</code></a> <em>||</em> <a href="https://github.com/GATB/minia"><code>minia</code></a>)
<ul>
<li>Blast to reference genome (<a href="https://blast.ncbi.nlm.nih.gov/Blast.cgi?PAGE_TYPE=BlastSearch"><code>blastn</code></a>)</li>
<li>Contiguate assembly (<a href="https://www.sanger.ac.uk/science/tools/pagit"><code>ABACAS</code></a>)</li>
<li>Assembly report (<a href="https://github.com/BU-ISCIII/plasmidID"><code>PlasmidID</code></a>)</li>
<li>Assembly assessment report (<a href="http://quast.sourceforge.net/quast"><code>QUAST</code></a>)</li>
</ul>
</li>
</ol></li>
<li>Present QC and visualisation for raw read, alignment, assembly and variant calling results (<a href="http://multiqc.info/"><code>MultiQC</code></a>)</li>
</ol>]]></description>
	<dc:creator>Surabhi Chaudhary</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/43846/the-complete-sequence-of-a-human-genome</guid>
	<pubDate>Thu, 31 Mar 2022 23:58:18 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/43846/the-complete-sequence-of-a-human-genome</link>
	<title><![CDATA[The complete sequence of a human genome]]></title>
	<description><![CDATA[<p><span>The completed regions include all centromeric satellite arrays, recent segmental duplications, and the short arms of all five acrocentric chromosomes, unlocking these complex regions of the genome to variational and functional studies.</span></p><p>Address of the bookmark: <a href="https://www.science.org/doi/10.1126/science.abj6987" rel="nofollow">https://www.science.org/doi/10.1126/science.abj6987</a></p>]]></description>
	<dc:creator>Neel</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/pages/view/44352/bioinformatics-tools-for-genome-assembly</guid>
	<pubDate>Mon, 24 Jul 2023 07:04:26 -0500</pubDate>
	<link>https://bioinformaticsonline.com/pages/view/44352/bioinformatics-tools-for-genome-assembly</link>
	<title><![CDATA[Bioinformatics tools for genome assembly !]]></title>
	<description><![CDATA[<p>There are numerous genome assembly tools available, each with its strengths and weaknesses. Here is a list of some widely used genome assembly tools as of my last update in September 2021:</p><ol>
<li>
<p><span>SPAdes:</span> An assembler specifically designed for single-cell and multi-cell bacterial genomes, as well as small eukaryotic genomes.</p>
</li>
<li>
<p><span>ABySS:</span> A parallelized assembler for large genomes that uses de Bruijn graphs.</p>
</li>
<li>
<p><span>Velvet:</span> Another de Bruijn graph-based assembler optimized for short-read sequencing data.</p>
</li>
<li>
<p><span>SOAPdenovo:</span> A de Bruijn graph-based assembler designed for short reads, widely used for assembling large and complex genomes.</p>
</li>
<li>
<p><span>MaSuRCA:</span> A hybrid assembler that combines data from multiple sequencing technologies, such as Illumina and PacBio.</p>
</li>
<li>
<p><span>Canu:</span> A long-read assembler optimized for PacBio and Oxford Nanopore sequencing data.</p>
</li>
<li>
<p><span>Flye:</span> A long-read assembler suitable for bacterial and small eukaryotic genomes.</p>
</li>
<li>
<p><span>SMARTdenovo:</span> An assembler designed for long reads, particularly suited for PacBio data.</p>
</li>
<li>
<p><span>SPAdes Long Read (SPAdesLR):</span> An extension of SPAdes for long-read data, such as those from PacBio or Nanopore.</p>
</li>
<li>
<p><span>Minia:</span> An assembler optimized for low memory consumption, suitable for small and medium-sized genomes.</p>
</li>
<li>
<p><span>Unicycler:</span> A hybrid assembler that combines short and long reads for circular bacterial genome assembly.</p>
</li>
<li>
<p><span>wtdbg2:</span> A de Bruijn graph assembler for long reads, efficient for very large genomes.</p>
</li>
<li>
<p><span>Shasta:</span> A long-read assembler that uses the Overlap-Layout-Consensus approach, suitable for PacBio and Nanopore data.</p>
</li>
<li>
<p><span>Sparc:</span> An assembler designed to handle noisy long reads from Nanopore sequencing.</p>
</li>
<li>
<p><span>CANA:</span> An assembler for metagenomic data, particularly for complex and diverse microbial communities.</p>
</li>
<li>
<p><span>Ra</span> Assembler: A metagenome assembler for long reads, designed for highly complex metagenomic samples.</p>
</li>
</ol><p>Please note that the field of bioinformatics is constantly evolving, and new assembly tools may have emerged since my last update. Additionally, the performance of these tools can vary depending on the characteristics of the sequencing data and the genome being assembled. When selecting an assembly tool, consider the specific requirements of your project, the available data types, and the computational resources at your disposal. Always refer to the respective tool's documentation and publications for the most up-to-date information and recommendations.</p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/45278/the-day-we-began-reading-the-dna-of-the-world</guid>
	<pubDate>Sat, 05 Sep 2026 01:43:21 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/45278/the-day-we-began-reading-the-dna-of-the-world</link>
	<title><![CDATA[The Day We Began Reading the DNA of the World]]></title>
	<description><![CDATA[<div>Picture waking up and knowing that in one lab, a scientist is reading the DNA of a whale. In another, a team is sequencing a rare plant. Meanwhile, researchers in India are decoding the genomes of different human populations.</div><div>&nbsp;</div><div>Many organisms. Many countries. All share one remarkable goal: to understand the genetic story of life.</div><div>DNA is nature's instruction manual, written with just four letters: A, T, C, and G. For years, scientists could only read small sections of this huge book. Now, new sequencing technology lets us read entire genomes like never before.</div><div>&nbsp;</div><div>One of the most ambitious projects is the Earth BioGenome Project (EBP). Its goal is to create reference genomes for about 1.5 million known eukaryotic species. This project links genome research across continents, species, and scientific fields.</div><div>&nbsp;</div><div>But this story goes beyond just animals and plants.</div><div>&nbsp;</div><div><strong>A New Chapter in Human Genomics</strong></div><div>&nbsp;</div><div>Large human genome projects are changing how we understand our own genetic diversity.</div><div>India's GenomeIndia project has sequenced thousands of people from different populations. This work is uncovering genetic variation that global databases have often missed.</div><div>&nbsp;</div><div>Similar population-scale effSimilar large-scale projects are happening worldwide. The All of Us Research Program in the United States, Europe's 1+ Million Genomes initiative, South Korea's national genomic data program, and projects in Saudi Arabia, Australia, and Singapore are all building huge genomic resources.collect DNA.</div><div>&nbsp;</div><div>The real goal is to link genomes with health, disease, and other biological information. This creates a base for more accurate research and, in time, more personalized medicine.</div><div>&nbsp;</div><div><strong>A Race Against Time</strong></div><div>&nbsp;</div><div>There is also an important twist.</div><div>&nbsp;</div><div>We are sequencing Earth's biodiversity even as many species face growing environmental threats.</div><div>A genome alone cannot save a species, but it does keep important information about its biology, evolution, and genetic diversity. This knowledge helps scientists understand risks, guide conservation, and find traits that could help agriculture or medicine.</div><div>&nbsp;</div><div>Projects like the Darwin Tree of Life, which studies species in Britain and Ireland, and the African BioGenome Project, which is growing genomics work across Africa, are key parts of this worldwide effort.</div><div>&nbsp;</div><div>Projects to Watch</div><ol>
<li>Earth BioGenome Project (EBP) - A global effort to sequence and annotate roughly 1.5 million known eukaryotic species and create a comprehensive reference library of Earth's biodiversity.</li>
<li>Vertebrate Genomes Project (VGP) - Focuses on producing high-quality reference genomes for vertebrate species, providing a major foundation for the wider Tree of Life.</li>
<li>Darwin Tree of Life - A UK and Ireland initiative aiming to sequence approximately 70,000 species, creating a detailed genomic map of regional biodiversity.</li>
<li>African BioGenome Project (AfricaBP) - A pan-African initiative focused on sequencing Africa's biodiversity while strengthening genomics and bioinformatics capacity across the continent.</li>
<li>European Reference Genome Atlas (ERGA) - A European effort to generate high-quality reference genomes for Europe's biodiversity and connect national sequencing efforts.</li>
<li>Canada BioGenome Project - Building reference genomes for Canadian biodiversity, supporting conservation, research and the study of ecosystems.</li>
<li>California Conservation Genomics Project - Uses genomics to understand and protect California's biodiversity and help inform conservation decisions.</li>
<li>Bat1K - An ambitious global effort to sequence the genomes of all known bat species, helping scientists study their evolution, longevity, immunity and unique biology.</li>
<li>Bird 10,000 (B10K) - A global project aiming to create genome resources covering the world's bird diversity and reconstruct avian evolutionary history.</li>
<li>Global Invertebrate Genomics Alliance (GIGA) - Builds genomic resources for the enormous and genetically diverse world of invertebrates.</li>
<li>GenomeIndia - India's national human-genome initiative, designed to capture the country's extraordinary population diversity and create a reference resource for Indian genomics.</li>
<li>All of Us Research Program - A US national research program combining genomic information with health and lifestyle data at population scale.</li>
<li>1+ Million Genomes / Genome of Europe - A European effort to enable secure, cross-border use of genomic and health data and develop genome-scale resources involving more than one million people.</li>
<li>Korea National Integrated Bio Big Data Project - South Korea's large-scale national effort combining genomic and health information from a very large population cohort.</li>
<li>Saudi Genome Program - A national genomics initiative focused on genetic diseases, population genomics and precision medicine in Saudi Arabia.</li>
<li>Australian Genomics / Genomics Health Futures Mission - Australia's growing investment in genomic medicine, rare disease research and population-scale health genomics.</li>
</ol><div><strong>The Library of Life</strong></div><div>&nbsp;</div><div>The fuThe future of genomics is not just about sequencing one genome at a time. Now, it is about creating a worldwide network of genomes&mdash;human and non-human, individual and species&mdash;linked by powerful databases and advanced computational tools.ine a library without shelves.</div><div>&nbsp;</div><div>In this library, the books are genomes. The pages are made of DNA. Scientists and machines are the readers. And this library is still being written. Each genome sequenced today adds a new page to the story of life and helps us better understand our place in it.</div>]]></description>
	<dc:creator>Jitendra Narayan</dc:creator>
</item>

</channel>
</rss>