<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/9586?offset=530</link>
	<atom:link href="https://bioinformaticsonline.com/related/9586?offset=530" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/13025/the-5-reasons-to-mistakes-at-bioinformatics-work</guid>
	<pubDate>Thu, 24 Jul 2014 02:51:41 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/13025/the-5-reasons-to-mistakes-at-bioinformatics-work</link>
	<title><![CDATA[The 5 reasons to mistakes at bioinformatics work !!!]]></title>
	<description><![CDATA[<p>When you're just starting out with biological programming, it's easy to run into complex problems that make you wonder how anyone has ever managed to write a program. There are some problems that trip up nearly every bioinformatician--everything from getting started understanding the biological problems to dealing with program design. Some random mistakes are so prominent that even experienced biological programmers do it. The 8 years in bioinformatics and my few random observations, most of them are snarky. These reasons will always take longer than expected and compel you to postpone your project deadline.</p><p><strong>1.Stupid for biologist:</strong> Biology is so complex that it will make bioinformatician feel stupid. There are no any universal fixed rules; it can surprise you any time. So be nice to biologists who ask questions and resolve your biological puzzles. Sometime you will have no idea what the hell you were doing either.<br /><br /><strong>2.Puzzling why:</strong> Do not hesitate to ask question. Especially. at the beginning of project you will have to ask a lot of questions. Instead of puzzling it out at end check out and clear your doubt even for a single error. It may can leads to wrong conclusion.<br /><br /><strong>3.Running marathon:</strong> The most of the biological software&rsquo;s documentation is always incomplete. In other word they are no more than 95 percent complete. Sometime a single problem can halt your entire project for months. Compilation and running the pipelines in tedious because almost all are interdependent and need proper configuration. I face the same kind of problem with Evolver :( &hellip; <br /><br /><strong>4.Folders missing:</strong> The pipelines generate lots of data, and we keep them in several folders for future use. But sometime we delete them by mistake and move to recovery&hellip;<br /><br /><strong>5.Digging deeper:</strong> Digging deeper is fruitful, but some time it can be catastrophic. You may get frustrated or direction less. So keep a biologist with you for rescue &hellip;. Sometime an expert computer programmer to handle your server. Remember, the server will always go down when you need it the most.<br /><br />The most common frustrating&nbsp; common line: Why do we do this again?</p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/44713/understanding-rna-seq-normalization-methods-tpm-vs-fpkm-vs-cpm</guid>
	<pubDate>Wed, 11 Dec 2024 00:59:15 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/44713/understanding-rna-seq-normalization-methods-tpm-vs-fpkm-vs-cpm</link>
	<title><![CDATA[Understanding RNA-Seq Normalization Methods: TPM vs. FPKM vs. CPM]]></title>
	<description><![CDATA[<p>RNA sequencing (RNA-Seq) is a powerful technology used to study transcriptomes, providing insights into gene expression levels. However, raw RNA-Seq data requires normalization to account for sequencing depth and gene length, enabling accurate comparisons between genes and samples. Among the most widely used normalization methods are TPM (Transcripts Per Million), FPKM (Fragments Per Kilobase Million), and CPM (Counts Per Million). Each method has its unique principles and applications, which we&rsquo;ll explore in this blog.</p><h2>Why Normalize RNA-Seq Data?</h2><p>Normalization is a crucial step in RNA-Seq analysis for the following reasons:</p><ul>
<li>
<p><strong>Sequencing depth:</strong> Different RNA-Seq experiments produce varying numbers of reads, making direct comparisons between samples misleading.</p>
</li>
<li>
<p><strong>Gene length:</strong> Longer genes inherently generate more reads, irrespective of their actual expression level.</p>
</li>
<li>
<p><strong>Bias reduction:</strong> Normalization mitigates technical biases, enabling meaningful biological interpretation.</p>
</li>
</ul><h2>TPM (Transcripts Per Million)</h2><p>TPM measures the proportion of reads mapped to a transcript, normalized by transcript length and sequencing depth. It is calculated as:</p><h3>Key Features:</h3><ol>
<li>
<p><strong>Proportionality:</strong> TPM values sum to 1,000,000 across all transcripts in a sample, making it easier to compare between samples.</p>
</li>
<li>
<p><strong>Intuitive interpretation:</strong> TPM values directly represent the abundance of transcripts in a sample.</p>
</li>
<li>
<p><strong>Preferred for comparisons:</strong> TPM facilitates between-sample comparisons better than FPKM.</p>
</li>
</ol><h2>FPKM (Fragments Per Kilobase Million)</h2><p>FPKM normalizes read counts by transcript length and sequencing depth, but without enforcing proportionality like TPM. It is defined as:</p><h3>Key Features:</h3><ol>
<li>
<p><strong>Historical significance:</strong> FPKM was one of the first normalization methods used for RNA-Seq.</p>
</li>
<li>
<p><strong>Single-end vs. paired-end:</strong> In paired-end sequencing, FPKM becomes RPKM (Reads Per Kilobase Million).</p>
</li>
<li>
<p><strong>Limited utility:</strong> FPKM values are not as robust as TPM for cross-sample comparisons due to lack of proportionality.</p>
</li>
</ol><h2>CPM (Counts Per Million)</h2><p>CPM normalizes raw read counts by sequencing depth, without considering gene length. It is expressed as:</p><h3>Key Features:</h3><ol>
<li>
<p><strong>Simplicity:</strong> CPM is straightforward and computationally less intensive.</p>
</li>
<li>
<p><strong>Application:</strong> Suitable for non-length-dependent analyses, such as comparing total expression levels or differential expression analysis.</p>
</li>
<li>
<p><strong>Gene length agnostic:</strong> CPM does not correct for gene length, making it less ideal for measuring expression levels.</p>
</li>
</ol><h2>When to Use Each Method</h2><ul>
<li>
<p><strong>TPM:</strong> Best for comparing expression levels between samples, especially when transcript length and sequencing depth vary.</p>
</li>
<li>
<p><strong>FPKM:</strong> Useful for historical consistency but generally replaced by TPM.</p>
</li>
<li>
<p><strong>CPM:</strong> Ideal for differential expression analysis when gene length normalization is unnecessary.</p>
</li>
</ul><h2>Conclusion</h2><p>Choosing the right normalization method depends on the specific objectives of your RNA-Seq analysis. TPM&rsquo;s proportionality and robustness make it the preferred choice for most applications, while CPM serves well for differential expression studies. Although FPKM paved the way for RNA-Seq normalization, it has largely been supplanted by TPM in modern workflows. Understanding these methods and their nuances ensures accurate and meaningful interpretations of RNA-Seq data.</p><h3>References:</h3><ol>
<li>
<p>Li, B., &amp; Dewey, C. N. (2011). RSEM: accurate transcript quantification from RNA-Seq data with or without a reference genome. <em>BMC Bioinformatics.</em></p>
</li>
<li>
<p>Trapnell, C., et al. (2010). Transcript assembly and quantification by RNA-Seq reveals unannotated transcripts and isoform switching during cell differentiation. <em>Nature Biotechnology.</em></p>
</li>
<li>
<p>Law, C. W., et al. (2014). voom: precision weights unlock linear model analysis tools for RNA-seq read counts. <em>Genome Biology.</em></p>
</li>
</ol>]]></description>
	<dc:creator>Neel</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/researchlabs/view/45309/ley-lab</guid>
  <pubDate>Fri, 11 Sep 2026 13:21:10 -0500</pubDate>
  <link></link>
  <title><![CDATA[Ley Lab !]]></title>
  <description><![CDATA[
<p>Research Program aims to elucidate how gut microbial species relate to genotypic differences between human individuals, to identify bacterial and archaeal species that share an evolutionary history with humans, and to determine the molecular basis of long-term host-microbial relationships. We approach these aims through a combination of population-level observations and laboratory-based molecular-level investigations. We link inter-microbial and microbial-host interactions at the molecular scale to patterns at population and evolutionary scales. </p>

<p>More at https://leylab.com/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/13523/megadock-40</guid>
	<pubDate>Thu, 07 Aug 2014 18:08:54 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/13523/megadock-40</link>
	<title><![CDATA[MEGADOCK 4.0]]></title>
	<description><![CDATA[<p>An ultra&ndash;high-performance protein&ndash;protein docking software for heterogeneous supercomputers</p>
<p id="p-4"><strong>Summary:</strong> The application of protein&ndash;protein docking in large-scale interactome analysis is a major challenge in structural bioinformatics and requires huge computing resources. In this work, we present MEGADOCK 4.0, an FFT-based docking software that makes extensive use of recent heterogeneous supercomputers and shows powerful, scalable performance of over 97% strong scaling.</p>
<p id="p-5"><strong>Availability and Implementation:</strong> MEGADOCK 4.0 is written in C++ with OpenMPI and NVIDIA CUDA 5.0 (or later) and is freely available to all academic and non-profit users at: <a href="http://www.bi.cs.titech.ac.jp/megadock">http://www.bi.cs.titech.ac.jp/megadock</a>.</p>
<p id="p-6"><strong>Contact:</strong> <a href="mailto:akiyama@cs.titech.ac.jp">akiyama@cs.titech.ac.jp</a></p><p>Address of the bookmark: <a href="http://bioinformatics.oxfordjournals.org/content/early/2014/08/06/bioinformatics.btu532.short" rel="nofollow">http://bioinformatics.oxfordjournals.org/content/early/2014/08/06/bioinformatics.btu532.short</a></p>]]></description>
	<dc:creator>Suleman Khan</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/14024/grapher</guid>
	<pubDate>Thu, 14 Aug 2014 14:02:17 -0500</pubDate>
	<link>https://bioinformaticsonline.com/news/view/14024/grapher</link>
	<title><![CDATA[GrapheR !!!]]></title>
	<description><![CDATA[<p>What a wonderful gem <em>GrapheR</em> is.... Oh yes it is. <em>GrapheR</em> is a GUI for base graphics in R by http://www.maximeherve.com/. The package provides a graphical user interface for creating base charts in R. It is ideal for beginners in R, as the user interface is very clear and the code is written along side into a text file, allowing users to recreate the charts directly in the console. <br /><br />Adding and changing legends? Messing around with the plotting window settings? It is much easier/quicker with this GUI than reading the help file and trying to understand the various parameters.<br />Here is a little example using the iris data set.<br /><br />library(GrapheR)<br />data(iris)<br />run.GrapheR()<br /><br />This will bring up a window that helps me to create the chart and tweak the various parameters.</p><p><img src="http://4.bp.blogspot.com/-NbnCM1dPh3E/U9aW9YxJ9oI/AAAAAAAABgo/gEPzPhOpf2Y/s1600/GrapheR.png" alt="image" width="878" height="868" style="border: 0px; border: 0px;"><br /><br />Finally, I find the underlying R code in a file created by <em>GrapheR</em>. For more details read also the <a href="http://cran.r-project.org/web/packages/GrapheR/index.html" target="_blank">package vignette</a>, which is available in <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_en.pdf" target="_blank">English</a>, <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_fr.pdf" target="_blank">French</a> and <a href="http://cran.r-project.org/web/packages/GrapheR/vignettes/manual_de.pdf" target="_blank">German</a>!</p>]]></description>
	<dc:creator>John Parker</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/26290/webinar-on-streamlining-large-scale-analysis-using-the-strand-ngs-pipeline-manager-on-24-feb-2016</guid>
	<pubDate>Fri, 05 Feb 2016 06:43:28 -0600</pubDate>
	<link>https://bioinformaticsonline.com/news/view/26290/webinar-on-streamlining-large-scale-analysis-using-the-strand-ngs-pipeline-manager-on-24-feb-2016</link>
	<title><![CDATA[Webinar on Streamlining large scale analysis using the Strand NGS Pipeline Manager on 24 Feb 2016]]></title>
	<description><![CDATA[<p><a href="http://www.strand-ngs.com/webinar_registration" title="webinar"><strong>Live Webinar on Streamlining large scale NGS data analysis using the Strand NGS Pipeline Manager on 24 Feb 2016</strong></a></p><p><strong>Abstract:</strong> Strand NGS includes comprehensive workflows for DNA-Seq, RNA-Seq, Small RNA-Seq, ChIP-Seq, MeDIP-Seq, and Methyl-Seq analysis. Each workflow includes a quality assessment and filter section, followed by a workflow-specific analysis section. The pipeline functionality in Strand NGS allows users to execute a sequence of analysis steps with specific parameters - all without any manual intervention. This simplifies the analysis in large scale sequencing projects where every sample needs to be processed identically.</p><p>In this webinar we will discuss the pre-packaged pipelines present in Strand NGS. The packaged pipelines have well-chosen default parameters and are suitable for users analyzing data for the first time in the tool. We will also show how advanced users can customize pipelines and share them with other Strand NGS users. Finally, we will show a brief glimpse of an elaborate pipeline that aligns reads, filters poor-quality matches, computes coverage metrics, identifies variants, checks for sample cross-contamination, and emails quality reports - all from within Strand NGS.</p><p><strong>Speaker:</strong> Dr. Vamsi Veeramachaneni, Vice President - Bioinformatics, Strand Life Sciences</p><p><strong>Details:</strong> Session 1: 2:30 PM IST, Session 2 : 10:30 PM IST<br /><strong>Register here:</strong> http://www.strand-ngs.com/webinar_registration</p><h3>&nbsp;</h3>]]></description>
	<dc:creator>Yeshodari</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/14186/pybedtools</guid>
	<pubDate>Wed, 20 Aug 2014 01:03:41 -0500</pubDate>
	<link>https://bioinformaticsonline.com/news/view/14186/pybedtools</link>
	<title><![CDATA[pybedtools]]></title>
	<description><![CDATA[<p>pybedtools is a Python wrapper for Aaron Quinlan's BEDtools programs (https://github.com/arq5x/bedtools), which are widely used for genomic interval manipulation or "genome algebra". pybedtools extends BEDTools by offering feature-level manipulations from with Python. See full online documentation, including installation instructions, at http://pythonhosted.org/pybedtools/.</p><p>More at http://pythonhosted.org/pybedtools/</p><p>A powerful toolset for genome arithmetic.http://code.google.com/p/bedtools/</p>]]></description>
	<dc:creator>Shruti Paniwala</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/34912/list-of-cancer-genomics-research-web-resources</guid>
	<pubDate>Wed, 27 Dec 2017 20:33:09 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/34912/list-of-cancer-genomics-research-web-resources</link>
	<title><![CDATA[List of cancer genomics research web resources !]]></title>
	<description><![CDATA[<p>Major web resources for cancer genomics research</p><p>CGHub <br />https://cghub.ucsc.edu/ <br />Comprehensive data repository; huge data size</p><p>EGA <br />https://www.ebi.ac.uk/ega/ <br />Comprehensive data repository; huge data size</p><p>COSMIC <br />http://cancer.sanger.ac.uk <br />Largest somatic mutation database; genome sequencing paper curation</p><p>CPRG <br />http://www.broadinstitute.org/software/cprg <br />Interface for cancer program resources</p><p>GDAC <br />http://gdac.broadinstitute.org/ <br />Data analysis; automatic pipelines; user-friendly reports</p><p>SNP500Cancer <br />http://snp500cancer.nci.nih.gov <br />Sequence and genotype verification of SNPs</p><p>canEvolve <br />www.canevolve.org/ <br />Comprehensive analysis of tumor profile; Data from 90 studies involving more than 10,000 patients</p><p>MethyCancer <br />http://methycancer.psych.ac.cn <br />Relationship among DNA methylation, gene expression and cancer</p><p>SomamiR <br />http://compbio.uthsc.edu/SomamiR/ <br />Correlation between somatic mutation and microRNA; genome-wide displaying</p><p>cBioPortal <br />http://www.cbioportal.org/public-portal/ <br />Graphical summaries; gene alteration; processed data; visualization</p><p>UCSC Cancer Genomics Browser <br />https://genome-cancer.soe.ucsc.edu/ <br />Clinical information; gene expression; copy number variation; visualization</p><p>CGWB <br />https://cgwb.nci.nih.gov/ <br />Visualization; gene mutation and variation; automated analysis pipeline</p><p>GDSC <br />http://www.cancerrxgene.org <br />Drug sensitivity information; drug response information</p><p>canSAR <br />https://cansar.icr.ac.uk/ <br />Multidisciplinary information; drug discovery</p><p>NONCODE <br />http://www.noncode.org/ ncRNAs; <br />lncRNAs; up-to-date and comprehensive resource</p>]]></description>
	<dc:creator>biogeek</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/38577/genoviz-visualization-software-for-genomics</guid>
	<pubDate>Wed, 02 Jan 2019 04:07:57 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/38577/genoviz-visualization-software-for-genomics</link>
	<title><![CDATA[GenoViz: Visualization software for genomics]]></title>
	<description><![CDATA[<p><span>GenoViz provides software applications and re-usable components for data visualization and data sharing in genomics. Our flagship product is Integrated Genome Browser (IGB).</span><br><br><span>For more information about IGB, visit&nbsp;</span><a href="http://bioviz.org/" target="_blank">http://bioviz.org<span></span></a><span>.</span><br><br><span>Source code for the project was hosted here for many years. In 2014, we moved to a new git repository at&nbsp;</span><a href="http://www.bitbucket.org/lorainelab/integrated-genome-browser" target="_blank">http://www.bitbucket.org/lorainelab/integrated-genome-browser<span></span></a><span>. We are still using SourceForge to distribute new releases of IGB as compiled code (igb.zip) you can use to run IGB on your computer.&nbsp;</span><br><br><span>If you have questions, feel free to get in touch. Contact project head Ann Loraine (</span><a href="mailto:aloraine@uncc.edu" target="_blank">aloraine@uncc.edu<span></span></a><span>) or lead developer David Norris (</span><a href="mailto:dcnorris@uncc.edu" target="_blank">dcnorris@uncc.edu<span></span></a><span>&gt;).</span></p><p>Address of the bookmark: <a href="https://sourceforge.net/projects/genoviz/" rel="nofollow">https://sourceforge.net/projects/genoviz/</a></p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/14904/bioinformatics-jrfsrf-position-at-iari</guid>
  <pubDate>Thu, 04 Sep 2014 04:14:01 -0500</pubDate>
  <link></link>
  <title><![CDATA[Bioinformatics JRF/SRF position at IARI]]></title>
  <description><![CDATA[
<p>DIVISION OF NEMATOLOGY<br />INDIAN AGRICULTURAL RESEARCH INSTITUTE<br />NEW DELHI 110012<br />Applications are invited for the posts of one Junior<br />Research Fellow and one RA in the DBT funded project entitled “ Plant parasitic nematode genome informatics - insilico resource development”. The project is for a period of three years. </p>

<p>Essential qualifications for JRF<br />: M. Sc. in Bioinformatics with experience in Proteomics, genomics and structural biology. Knowledge of programming language, pearl and database – HTML, CSS,php and Java script.<br />Essential qualifications for Research Associate:<br />MSc/MTech in Bioinformatics with three years experience or Ph.D in Bioinformatics with experience in proteomics, genomics and structural biology. Knowledge of programming language, perl and database<br />– HTML, CSS, Java script. NGS sequence assembly and analysis and algorithm designing.<br />Age limit : 35 years maximum (5 year relaxation for SC/ST and women candidates)<br />Emoluments:<br />JRF: 16,000 + 30% HRA<br />.<br />Res Assoc: Rs22,000 + 30% HRA<br />The post is purely temporary in nature and is co-terminus with the project. The appointment would be initially for one year and may be extended further upon satisfactory performance.<br />Interested candidates<br />should send the duly filled application forms (format in the following page ) so as to reach on or before 20.9.2014 along with all the relevant documents.</p>

<p>More at http://www.iari.res.in/files/JRF_RA-03092014-20140903-135319.pdf</p>
]]></description>
</item>

</channel>
</rss>