<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/4197?offset=980</link>
	<atom:link href="https://bioinformaticsonline.com/related/4197?offset=980" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/43725/comparative-genomics-workshops</guid>
	<pubDate>Tue, 25 Jan 2022 20:39:58 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/43725/comparative-genomics-workshops</link>
	<title><![CDATA[Comparative Genomics Workshops !]]></title>
	<description><![CDATA[<p><span>This meeting's objective was to obtain a big picture look at the current state of the field of comparative&nbsp;genomics with a focus on commonalities across genomic investigations into humans, model organisms&nbsp;(both traditional and non-traditional), agricultural species, wildlife species and microbes.</span></p>
<p>https://www.genome.gov/event-calendar/perspectives-in-comparative-genomics-and-evolution</p><p>Address of the bookmark: <a href="https://www.genome.gov/event-calendar/perspectives-in-comparative-genomics-and-evolution" rel="nofollow">https://www.genome.gov/event-calendar/perspectives-in-comparative-genomics-and-evolution</a></p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/researchlabs/view/44639/the-sheppard-lab</guid>
  <pubDate>Fri, 09 Aug 2024 02:48:34 -0500</pubDate>
  <link></link>
  <title><![CDATA[The Sheppard Lab]]></title>
  <description><![CDATA[
<p>Ineos Oxford Institute of Antimicrobial Research – Department of Biology – University of Oxford</p>

<p>Our research centres on the use of genetics/genomics and phenotypic studies to address complex questions in the ecology, epidemiology and evolution of microbes. Our most recent interest focuses upon comparative genome analysis to describe the core and flexible genome of pathogenic bacteria (Campylobacter, Acinetobacter, Escherichia coli, Helicobacter, Staphylococcus and Streptococcus suis) and how this is related to population genetic structuring, the maintenance of species, and the evolution of host/niche adaptation and virulence.</p>

<p>More at https://sheppardlab.com/research/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/14024/grapher</guid>
	<pubDate>Thu, 14 Aug 2014 14:02:17 -0500</pubDate>
	<link>https://bioinformaticsonline.com/news/view/14024/grapher</link>
	<title><![CDATA[GrapheR !!!]]></title>
	<description><![CDATA[<p>What a wonderful gem <em>GrapheR</em> is.... Oh yes it is. <em>GrapheR</em> is a GUI for base graphics in R by http://www.maximeherve.com/. The package provides a graphical user interface for creating base charts in R. It is ideal for beginners in R, as the user interface is very clear and the code is written along side into a text file, allowing users to recreate the charts directly in the console. <br /><br />Adding and changing legends? Messing around with the plotting window settings? It is much easier/quicker with this GUI than reading the help file and trying to understand the various parameters.<br />Here is a little example using the iris data set.<br /><br />library(GrapheR)<br />data(iris)<br />run.GrapheR()<br /><br />This will bring up a window that helps me to create the chart and tweak the various parameters.</p><p><img src="http://4.bp.blogspot.com/-NbnCM1dPh3E/U9aW9YxJ9oI/AAAAAAAABgo/gEPzPhOpf2Y/s1600/GrapheR.png" alt="image" width="878" height="868" style="border: 0px; border: 0px;"><br /><br />Finally, I find the underlying R code in a file created by <em>GrapheR</em>. For more details read also the <a href="http://cran.r-project.org/web/packages/GrapheR/index.html" target="_blank">package vignette</a>, which is available in <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_en.pdf" target="_blank">English</a>, <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_fr.pdf" target="_blank">French</a> and <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_de.pdf" target="_blank">German</a>!</p>]]></description>
	<dc:creator>John Parker</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/44758/the-ifs-and-buts-of-ngs-quality-control-and-trimming</guid>
	<pubDate>Thu, 02 Jan 2025 20:11:07 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/44758/the-ifs-and-buts-of-ngs-quality-control-and-trimming</link>
	<title><![CDATA[The &quot;Ifs&quot; and &quot;Buts&quot; of NGS Quality Control and Trimming]]></title>
	<description><![CDATA[<p>Next-Generation Sequencing (NGS) has revolutionized biological research, providing vast amounts of data for a wide range of applications. However, the reliability of NGS analyses heavily depends on the quality of raw sequencing data. Quality control (QC) and trimming are critical preprocessing steps that can make or break your downstream analyses. In this blog, we explore the "ifs" (why you should perform QC and trimming) and the "buts" (challenges or considerations) of this vital step in NGS workflows.</p><h3><strong>The "Ifs" of NGS QC and Trimming</strong></h3><ol>
<li>
<p><strong>Ensures Data Integrity</strong><br />If you want to minimize errors in downstream analyses, QC and trimming remove low-quality reads and bases, ensuring high-confidence data. This step is essential for reliable variant calling, assembly, and other applications.</p>
</li>
<li>
<p><strong>Removes Contaminants</strong><br />If adapter sequences or contaminants are present in the raw reads, trimming can eliminate them. This prevents issues like misalignment or incorrect biological interpretations, ensuring cleaner data for analysis.</p>
</li>
<li>
<p><strong>Improves Mapping and Assembly</strong><br />If your goal is better alignment to a reference genome or improved de novo assembly, trimming low-quality bases and adapters is critical. High-quality reads map more efficiently and generate more accurate assemblies.</p>
</li>
<li>
<p><strong>Reduces Computational Load</strong><br />If you want to save computational resources, trimming reduces the dataset size, which speeds up processing and analysis. Clean datasets mean less computational time spent on processing low-quality data.</p>
</li>
<li>
<p><strong>Prepares for Standardized Analyses</strong><br />If your project involves multiple datasets, QC and trimming ensure uniformity across them. This standardization makes comparisons valid and reproducible, particularly in large collaborative studies.</p>
</li>
</ol><h3><strong>The "Buts" of NGS QC and Trimming</strong></h3><ol>
<li>
<p><strong>Risk of Over-Trimming</strong><br />But excessive trimming can lead to the loss of informative sequences, reducing read depth and potentially discarding biologically relevant data. This is especially critical in studies with limited sequencing depth.</p>
</li>
<li>
<p><strong>Bias Introduction</strong><br />But trimming algorithms might introduce biases, especially if they inadvertently remove sequences with specific biological patterns. This can skew results and compromise biological insights.</p>
</li>
<li>
<p><strong>Loss of Context in Paired-End Reads</strong><br />But trimming one read in a pair more than the other can lead to loss of pairing information. This complicates downstream analyses that rely on paired-end data, such as structural variant detection.</p>
</li>
<li>
<p><strong>Time and Resource Intensive</strong><br />But running QC and trimming for large datasets can be computationally expensive and time-consuming. As sequencing depth increases, preprocessing becomes a bottleneck in the analysis pipeline.</p>
</li>
<li>
<p><strong>Variable Standards</strong><br />But the criteria for trimming (e.g., quality threshold, minimum read length) can vary between tools and datasets. This variability may affect reproducibility and comparability of results across studies.</p>
</li>
</ol><h3><strong>Balancing the "Ifs" and "Buts"</strong></h3><p>To maximize the benefits of QC and trimming while mitigating the challenges, consider the following best practices:</p><ul>
<li>
<p><strong>Use QC Tools Wisely:</strong> Start with tools like <strong>FastQC</strong> to identify quality issues in your raw data. Visualizing quality metrics helps tailor your trimming parameters.</p>
</li>
<li>
<p><strong>Choose Reliable Trimming Tools:</strong> Tools like <strong>Trimmomatic</strong>, <strong>Cutadapt</strong>, and <strong>BBduk</strong> offer adaptive and customizable trimming options. Select one that aligns with your dataset and project goals.</p>
</li>
<li>
<p><strong>Set Reasonable Parameters:</strong> Avoid over-trimming by setting quality thresholds and minimum read lengths that balance data retention and quality improvement.</p>
</li>
<li>
<p><strong>Test Downstream Effects:</strong> Validate the impact of QC and trimming on downstream analyses, such as alignment efficiency, variant calling accuracy, or assembly quality.</p>
</li>
<li>
<p><strong>Document Your Workflow:</strong> Maintain detailed records of the parameters and tools used for QC and trimming. This ensures reproducibility and enables better troubleshooting.</p>
</li>
</ul><h3><strong>Conclusion</strong></h3><p>NGS quality control and trimming are essential steps to ensure reliable and accurate data for analysis. While the "ifs" highlight the clear benefits of these steps, the "buts" remind us of the potential pitfalls. By adopting best practices and carefully balancing these considerations, you can optimize your preprocessing workflow and unlock the full potential of your sequencing data.</p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/14186/pybedtools</guid>
	<pubDate>Wed, 20 Aug 2014 01:03:41 -0500</pubDate>
	<link>https://bioinformaticsonline.com/news/view/14186/pybedtools</link>
	<title><![CDATA[pybedtools]]></title>
	<description><![CDATA[<p>pybedtools is a Python wrapper for Aaron Quinlan's BEDtools programs (https://github.com/arq5x/bedtools), which are widely used for genomic interval manipulation or "genome algebra". pybedtools extends BEDTools by offering feature-level manipulations from with Python. See full online documentation, including installation instructions, at http://pythonhosted.org/pybedtools/.</p><p>More at http://pythonhosted.org/pybedtools/</p><p>A powerful toolset for genome arithmetic.http://code.google.com/p/bedtools/</p>]]></description>
	<dc:creator>Shruti Paniwala</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/45310/when-genes-leave-clues-reading-the-evolutionary-footprints-of-life</guid>
	<pubDate>Fri, 11 Sep 2026 13:47:07 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/45310/when-genes-leave-clues-reading-the-evolutionary-footprints-of-life</link>
	<title><![CDATA[When Genes Leave Clues: Reading the Evolutionary Footprints of Life]]></title>
	<description><![CDATA[<p>Imagine walking through a forest and finding two sets of footprints that always appear, disappear, and reappear together. You might wonder whether the animals are connected. In genomics, genes leave similar footprints across evolution. When two genes repeatedly occur together&mdash;or disappear together&mdash;across different species, their shared evolutionary history can provide clues about their function.</p><p>This idea, known as phylogenetic profiling, is the foundation of Profylo (https://github.com/MartinSchoenstein/Profylo), an open-source Python toolkit developed for comparing evolutionary profiles and discovering co-evolving genes. Profylo brings together seven profile-comparison methods and four approaches for identifying co-evolving gene clusters, making it easier to explore hidden functional relationships within genomes.</p><p>The study demonstrates that even a simple matrix of gene presence and absence can reveal meaningful biological patterns. By following these evolutionary footprints, researchers can generate new hypotheses about unknown genes, pathways, and molecular interactions. In a world where genomes are being sequenced faster than ever, tools like Profylo offer an exciting way to let evolution itself become a guide to understanding gene function.</p><p>More at&nbsp;https://link.springer.com/article/10.1007/s00239-025-10280-6&nbsp;</p>]]></description>
	<dc:creator>LEGE</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/34485/phyloxml-xml-for-evolutionary-biology-and-comparative-genomics</guid>
	<pubDate>Wed, 29 Nov 2017 10:04:48 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/34485/phyloxml-xml-for-evolutionary-biology-and-comparative-genomics</link>
	<title><![CDATA[phyloXML:  XML for evolutionary biology and comparative genomics]]></title>
	<description><![CDATA[<p><a href="http://www.biomedcentral.com/1471-2105/10/356/">phyloXML</a><span>&nbsp;(</span><a href="http://www.phyloxml.org/examples_syntax/phyloxml_syntax_example_1.html">example</a><span>) is an&nbsp;</span><a href="http://en.wikipedia.org/wiki/XML">XML</a><span>&nbsp;language designed to describe phylogenetic trees (or networks) and associated data. PhyloXML provides elements for commonly used features, such as taxonomic information, gene names and identifiers, branch lengths, support values, and gene duplication and speciation events. Using these standardized elements allows interoperability between various applications and databases. Furthermore, both due to extensible nature of XML itself and the provision of &lt;property&gt; elements by phyloXML, extensibility as well as domain specific applications are ensured. The structure of phyloXML is described by&nbsp;</span><a href="http://en.wikipedia.org/wiki/XML_Schema_%28W3C%29">XML Schema Definition (XSD)</a><span>&nbsp;language.</span></p><p>Address of the bookmark: <a href="http://www.phyloxml.org/" rel="nofollow">http://www.phyloxml.org/</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/36267/pspairwise-sequentially-markovian-coalescent-psmc-model</guid>
	<pubDate>Thu, 19 Apr 2018 05:29:23 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/36267/pspairwise-sequentially-markovian-coalescent-psmc-model</link>
	<title><![CDATA[PSPairwise Sequentially Markovian Coalescent (PSMC) model]]></title>
	<description><![CDATA[<p><span>Implementation of the Pairwise Sequentially Markovian Coalescent (PSMC) model</span></p><p>Address of the bookmark: <a href="https://github.com/lh3/psmc" rel="nofollow">https://github.com/lh3/psmc</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/40834/nucleus-python-and-c-code-for-reading-and-writing-genomics-data</guid>
	<pubDate>Sun, 02 Feb 2020 08:14:19 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/40834/nucleus-python-and-c-code-for-reading-and-writing-genomics-data</link>
	<title><![CDATA[Nucleus: Python and C++ code for reading and writing genomics data.]]></title>
	<description><![CDATA[<p>Nucleus is a library of Python and C++ code designed to make it easy to read, write and analyze data in common genomics file formats like SAM and VCF. In addition, Nucleus enables painless integration with the TensorFlow machine learning framework, as anywhere a genomics file is consumed or produced, a TensorFlow tfrecords file may be used instead.</p><p>Address of the bookmark: <a href="https://github.com/google/nucleus" rel="nofollow">https://github.com/google/nucleus</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/42143/sibelia-a-comparative-genomics-tool</guid>
	<pubDate>Sat, 22 Aug 2020 02:49:00 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/42143/sibelia-a-comparative-genomics-tool</link>
	<title><![CDATA[Sibelia: A comparative genomics tool]]></title>
	<description><![CDATA[<p><strong>Sibelia</strong>: A comparative genomics tool: It assists biologists in analysing the genomic variations that correlate with pathogens, or the genomic changes that help microorganisms adapt in different environments. Sibelia will also be helpful for the evolutionary and genome rearrangement studies for multiple strains of microorganisms.&nbsp;</p>
<p><strong>Sibelia</strong>&nbsp;is useful in finding: (1) shared regions, (2) regions that present in one group of genomes but not in others, (3) rearrangements that transform one genome to other genomes.</p>
<p>More at&nbsp;<a href="http://bioinf.spbau.ru/sibelia">http://bioinf.spbau.ru/sibelia</a></p>
<p>Sibelia docs&nbsp;<a href="http://gensoft.pasteur.fr/docs/Sibelia/3.0.7/SIBELIA.md">http://gensoft.pasteur.fr/docs/Sibelia/3.0.7/SIBELIA.md</a></p><p>Address of the bookmark: <a href="https://github.com/bioinf/Sibelia" rel="nofollow">https://github.com/bioinf/Sibelia</a></p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>

</channel>
</rss>