<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/30140?offset=500</link>
	<atom:link href="https://bioinformaticsonline.com/related/30140?offset=500" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/36723/hapsembler-an-assembler-for-highly-polymorphic-genomes</guid>
	<pubDate>Tue, 22 May 2018 04:09:53 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/36723/hapsembler-an-assembler-for-highly-polymorphic-genomes</link>
	<title><![CDATA[Hapsembler: An Assembler for Highly Polymorphic Genomes]]></title>
	<description><![CDATA[Hapsembler is a haplotype-specific genome assembly toolkit that is designed for genomes that are rich in SNPs and other types of polymorphism. Hapsembler can be used to assemble reads from a variety of platforms including Illumina and Roche/454. 

http://compbio.cs.toronto.edu/hapsembler/<p>Address of the bookmark: <a href="http://compbio.cs.toronto.edu/hapsembler/" rel="nofollow">http://compbio.cs.toronto.edu/hapsembler/</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/44413/bioinformatics-opening-at-nibmg-india</guid>
  <pubDate>Sun, 03 Dec 2023 00:16:59 -0600</pubDate>
  <link></link>
  <title><![CDATA[Bioinformatics Opening at NIBMG, India]]></title>
  <description><![CDATA[
<p>NIBMG is looking for motivated and bright individuals interested to explore career<br />opportunities for the position of Research Associate (Project Linked Person) for extramural<br />project funded by ICMR as per details given below.<br />Project Name: Fast detection of driver mutations and genes from cancer genomics data using<br />an integrative machine learning-based approach.</p>

<p>More at https://www.nibmg.ac.in/uploads/3c5d4da3fb31bef490a218805408c858.pdf</p>
]]></description>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/44726/postdoc-at-ubasel-comparative-single-cell-genomics</guid>
  <pubDate>Fri, 13 Dec 2024 12:46:19 -0600</pubDate>
  <link></link>
  <title><![CDATA[Postdoc at UBasel Comparative Single Cell Genomics]]></title>
  <description><![CDATA[
<p>A fully funded 4-year Postdoc position is available in the lab of Patrick<br />Tschopp at the University of Basel, Switzerland, study the molecular and<br />tissue-scale dynamics during the embryonic formation of the vertebrate<br />skeleton and compare it across different vertebrate species with distinct<br />habitats.</p>

<p>We are looking for a highly motivated candidate with a PhD degree in<br />Bioinformatics or a related field. Candidates are expected to have a<br />strong background in evolutionary biology and/or comparative functional<br />genomics. Additional experiences in single cell functional genomics<br />analyses, statistics and computational data analyses are a plus, as is<br />an interest in comparative developmental (EvoDevo) questions.</p>

<p>We offer a dynamic and interactive research environment with state-of-the<br />art research facilities, good research funding and internationally<br />competitive salaries.</p>

<p>The Tschopp lab (www.evolution.unibas.ch/tschopp/research/)<br />studies the gene regulatory mechanisms of cell type<br />specification and evolution in vertebrates. See also our<br />preprints at https://doi.org/10.1101/2024.03.26.586769 and<br />https://doi.org/10.1101/2024.11.28.625862 Applications should include<br />a motivation letter, a CV, a list of publications, a statement about<br />research interests, as well as the names and contact details of at<br />least two referees. Applications (in the form of a single .pdf file)<br />should be sent to Patrick Tschopp (patrick.tschopp@unibas.ch); review<br />of applications will begin on January 1st 2025, and will continue until<br />the position is filled.</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/38804/grabb-selective-assembly-of-genomic-regions-a-new-niche-for-genomic-research</guid>
	<pubDate>Sat, 26 Jan 2019 18:58:16 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/38804/grabb-selective-assembly-of-genomic-regions-a-new-niche-for-genomic-research</link>
	<title><![CDATA[GRAbB: Selective Assembly of Genomic Regions, a New Niche for Genomic Research]]></title>
	<description><![CDATA[<p><span>GRAbB is shown to be more efficient than MITObim in terms of speed, memory and disk usage. The other functionalities (handling multiple targets simultaneously and extracting homologous regions) of the new program are not matched by other programs. The program is available with explanatory documentation at&nbsp;</span><a href="https://github.com/b-brankovics/grabb">https://github.com/b-brankovics/grabb</a><span>. GRAbB has been tested on Ubuntu (12.04 and 14.04), Fedora (23), CentOS (7.1.1503) and Mac OS X (10.7). Furthermore, GRAbB is available as a docker repository: brankovics/grabb (</span><a href="https://hub.docker.com/r/brankovics/grabb/">https://hub.docker.com/r/brankovics/grabb/</a><span>).</span></p><p>Address of the bookmark: <a href="https://github.com/b-brankovics/grabb" rel="nofollow">https://github.com/b-brankovics/grabb</a></p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/44677/exploring-bioinformatics-job-websites-your-gateway-to-a-thriving-career</guid>
	<pubDate>Sat, 19 Oct 2024 13:43:06 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/44677/exploring-bioinformatics-job-websites-your-gateway-to-a-thriving-career</link>
	<title><![CDATA[Exploring Bioinformatics Job Websites: Your Gateway to a Thriving Career]]></title>
	<description><![CDATA[<p>Bioinformatics is a rapidly growing field at the intersection of biology, computer science, and data analytics, with applications in healthcare, genomics, drug discovery, and more. As demand increases for skilled professionals who can manage, analyze, and interpret biological data, finding the right job opportunities can be challenging. Fortunately, numerous online platforms cater specifically to bioinformatics professionals, from academia to industry positions.</p><p>Here&rsquo;s a curated list of the top websites offering bioinformatics job opportunities and postdoctoral fellowships worldwide.</p><h3>1. <strong>General Bioinformatics Job Portals</strong></h3><p>These platforms are ideal for bioinformaticians seeking jobs in diverse sectors:</p><ul>
<li>
<p><strong><a href="https://www.nature.com/naturecareers/" target="_new">Nature Careers</a>:</strong> A trusted resource for job seekers in the sciences, Nature Careers offers bioinformatics roles globally. Their specialized search function allows you to filter jobs by keyword, location, and more.</p>
<ul>
<li><a href="https://www.nature.com/naturecareers/searchjobs/?Keywords=bioinformatics" target="_new">Explore Bioinformatics Jobs on Nature Careers</a></li>
</ul>
</li>
<li>
<p><strong><a href="https://jobs.sciencecareers.org/searchjobs/?Keywords=bioinformatics" target="_new">Science Careers</a>:</strong> A job board from the AAAS, this site focuses on STEM jobs, including numerous bioinformatics opportunities in academia and industry.</p>
</li>
<li>
<p><strong><a href="https://euraxess.ec.europa.eu/" target="_new">Euraxess</a>:</strong> Euraxess is the go-to platform for researchers looking for jobs, fellowships, and funding across Europe and beyond. It lists both bioinformatics roles and research grants.</p>
<ul>
<li><a href="https://euraxess.ec.europa.eu/search?keys=bioinformatics" target="_new">Search Bioinformatics Jobs on Euraxess</a></li>
</ul>
</li>
<li>
<p><strong><a href="https://www.researchgate.net/jobs/search/bioinformatics" target="_new">ResearchGate Jobs</a>:</strong> ResearchGate is widely known as a platform for researchers to share publications, but it also has a robust job board featuring bioinformatics positions globally.</p>
</li>
<li>
<p><strong><a href="https://www.findapostdoc.com/?Keywords=bioinformatics" target="_new">FindAPostDoc</a>:</strong> This site is dedicated to helping postdoctoral researchers find positions, with bioinformatics being a popular category.</p>
</li>
<li>
<p><strong><a href="https://academicpositions.com/find-jobs?search=bioinformatics" target="_new">Academic Positions</a>:</strong> Targeting academic roles worldwide, Academic Positions lists bioinformatics jobs at universities and research institutions.</p>
</li>
<li>
<p><strong><a href="https://www.postdocjobs.com/job/search/index?keyword=bioinformatics&amp;location=" target="_new">PostdocJobs.com</a>:</strong> Specializing in postdoctoral roles, this platform is a great resource for early-career researchers looking for bioinformatics-related positions.</p>
</li>
<li>
<p><strong><a href="https://scholarship-positions.com/?s=bioinformatics" target="_new">Scholarship Positions</a>:</strong> In addition to jobs, Scholarship Positions provides information on scholarships, fellowships, and grants related to bioinformatics.</p>
</li>
</ul><h3>2. <strong>Fellowship and Training Opportunities in Bioinformatics</strong></h3><p>For those seeking fellowships or specialized training, these organizations offer postdoctoral programs, grants, and research opportunities:</p><ul>
<li>
<p><strong><a href="https://www.training.nih.gov/research-training/pd/" target="_new">NIH Office of Intramural Training and Education</a>:</strong> The National Institutes of Health offer extensive research training programs for postdocs, including those in bioinformatics.</p>
</li>
<li>
<p><strong><a href="https://new.nsf.gov/funding/opportunities/rui-roa-pui-facilitating-research-predominantly-undergraduate" target="_new">NSF Research Opportunity Awards</a>:</strong> The National Science Foundation funds bioinformatics research at predominantly undergraduate institutions, providing fellowships and grants.</p>
</li>
<li>
<p><strong>Top U.S. Universities:</strong> Many prestigious U.S. institutions, including <a href="https://postdoc.hms.harvard.edu/fellowships" target="_new">Harvard</a>, <a href="https://postdoc.berkeley.edu/" target="_new">Berkeley</a>, <a href="https://postdocs.yale.edu/" target="_new">Yale</a>, <a href="https://postdocs.mit.edu/" target="_new">MIT</a>, <a href="https://postdoc.jhu.edu/" target="_new">Johns Hopkins</a>, <a href="https://postdocs.ucsd.edu/" target="_new">UCSD</a>, and <a href="https://postdocs.cornell.edu/" target="_new">Cornell</a>, offer postdoctoral opportunities in bioinformatics.</p>
</li>
</ul><h3>3. <strong>Country-Specific Job and Fellowship Resources</strong></h3><p>If you're targeting a specific region, these platforms offer bioinformatics opportunities tailored to their respective countries:</p><h4><strong>Canada</strong></h4><ul>
<li><strong><a href="https://capsacpp.ca/" target="_new">CAPS/ACPP</a>:</strong> The Canadian Association of Postdoctoral Scholars provides a job board, including bioinformatics roles in academia.</li>
<li><strong><a href="https://banting.fellowships-bourses.gc.ca/" target="_new">Banting Postdoctoral Fellowships</a>:</strong> A prestigious fellowship program for postdocs in bioinformatics and related fields.</li>
<li><strong><a href="https://www.mitacs.ca/our-programs/elevate-business/" target="_new">Mitacs Elevate</a>:</strong> A Canadian initiative offering fellowships to connect postdoctoral researchers with industry partners.</li>
</ul><h4><strong>United Kingdom</strong></h4><ul>
<li><strong><a href="https://www.ukri.org/" target="_new">UKRI</a>:</strong> The UK Research and Innovation body funds bioinformatics research and offers various grants.</li>
<li><strong><a href="https://royalsociety.org/grants/" target="_new">The Royal Society</a>:</strong> Provides funding schemes for researchers in bioinformatics.</li>
<li><strong><a href="https://marie-sklodowska-curie-actions.ec.europa.eu/" target="_new">Marie Skłodowska-Curie Actions</a>:</strong> The MSCA funds fellowships and doctoral programs across Europe, including bioinformatics-related projects.</li>
<li><strong><a href="https://wellcome.org/grant-funding/schemes" target="_new">Wellcome Trust</a>:</strong> Offers research funding and career development opportunities in health-related fields, including bioinformatics.</li>
</ul><h4><strong>Europe</strong></h4><ul>
<li><strong><a href="https://www.embo.org/funding/fellowships-grants-and-career-support/" target="_new">EMBO Fellowships</a>:</strong> The European Molecular Biology Organization supports bioinformaticians through fellowships and career grants.</li>
<li><strong><a href="https://www.mpg.de/career-programs" target="_new">Max Planck Society</a>:</strong> A leading research organization offering bioinformatics positions and fellowships across Europe.</li>
<li><strong><a href="https://www.helmholtz.de/en/" target="_new">Helmholtz Association</a>:</strong> A major research organization in Germany offering bioinformatics roles in various disciplines.</li>
<li><strong><a href="https://www.leibniz-gemeinschaft.de/en/careers/careers-in-research" target="_new">Leibniz Association</a>:</strong> Offers research opportunities, including bioinformatics, across its numerous institutes.</li>
</ul><h4><strong>Australia and New Zealand</strong></h4><ul>
<li><strong><a href="https://www.arc.gov.au/funding-research/funding-schemes" target="_new">Australian Research Council</a>:</strong> Offers funding and research schemes, including in bioinformatics.</li>
<li><strong>Top Universities:</strong> Universities like <a href="https://www.sydney.edu.au/research.html" target="_new">Sydney</a>, <a href="https://research.unimelb.edu.au/" target="_new">Melbourne</a>, and <a href="https://research.uq.edu.au/" target="_new">Queensland</a> have research programs in bioinformatics.</li>
</ul><h4><strong>Asia</strong></h4><ul>
<li><strong><a href="https://www.jsps.go.jp/english/e-fellow/index.html" target="_new">Japan Society for the Promotion of Science (JSPS)</a>:</strong> Offers fellowships for international researchers in bioinformatics.</li>
<li><strong>Top Institutions:</strong> Universities like <a href="https://www.nus.edu.sg/careers/" target="_new">NUS</a>, <a href="https://english.cas.cn/" target="_new">CAS</a>, and <a href="https://iisc.ac.in/" target="_new">IISc</a> are leading hubs for bioinformatics research.</li>
</ul><h4><strong>Middle East</strong></h4><ul>
<li><strong><a href="https://qrdi.org.qa/en-us/" target="_new">Qatar Research, Development, and Innovation (QRDI)</a>:</strong> Offers research opportunities in bioinformatics.</li>
<li><strong><a href="https://www.kaust.edu.sa/en/" target="_new">KAUST</a>:</strong> A leading university in Saudi Arabia offering bioinformatics research positions.</li>
</ul><h4><strong>Africa</strong></h4><ul>
<li><strong><a href="https://aasciences.africa/" target="_new">African Academy of Sciences</a>:</strong> Provides career opportunities and research funding in bioinformatics across Africa.</li>
</ul><h3>Conclusion</h3><p>The field of bioinformatics is full of exciting opportunities for those with the right skills. Whether you are looking for a postdoc position, research funding, or a long-term job in industry, these platforms are an excellent starting point. Explore, apply, and take the next step in your bioinformatics career!</p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/40940/consed-a-finishing-package-bam-file-viewer-assembly-editor-autofinish-autoreport-autoedit-and-align-reads-to-reference-sequence</guid>
	<pubDate>Fri, 07 Feb 2020 07:16:22 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/40940/consed-a-finishing-package-bam-file-viewer-assembly-editor-autofinish-autoreport-autoedit-and-align-reads-to-reference-sequence</link>
	<title><![CDATA[Consed--A Finishing Package (BAM File Viewer, Assembly Editor, Autofinish, Autoreport, Autoedit, and Align Reads To Reference Sequence)]]></title>
	<description><![CDATA[<ul>
<li>Supports Illumina, 454, other Next-Gen and Sanger Reads and allows mixtures of these read types</li>
<li>Consed includes BamScape which can view bam files with unlimited numbers of reads. BamScape can bring up consed to edit reads and the reference sequence in targeted regions.</li>
<li>Consed is compatible with Newbler, Cross_match, Phrap, MIRA, Velvet and PCAP output.</li>
<li>Quickly takes the user to each variant site for viewing (also available as an automated report)</li>
<li>Overview of assembly can help detect and fix misassemblies</li>
<li>Editing time reduced by the program's ability to pin-point problem areas</li>
<li>Editing is guided by error probabilities</li>
</ul><p>Address of the bookmark: <a href="http://www.phrap.org/consed/consed.html" rel="nofollow">http://www.phrap.org/consed/consed.html</a></p>]]></description>
	<dc:creator>Neel</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/44703/the-role-of-lncrna-in-bioinformatics-unlocking-the-secrets-of-the-genome</guid>
	<pubDate>Sat, 07 Dec 2024 02:09:47 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/44703/the-role-of-lncrna-in-bioinformatics-unlocking-the-secrets-of-the-genome</link>
	<title><![CDATA[The Role of lncRNA in Bioinformatics: Unlocking the Secrets of the Genome]]></title>
	<description><![CDATA[<p>In the intricate dance of molecular biology, long non-coding RNAs (lncRNAs) have emerged as key players, capturing the interest of researchers worldwide. These RNA molecules, once dismissed as "junk," have proven to be vital in the regulation of gene expression, cellular processes, and the progression of diseases. The intersection of lncRNA studies and bioinformatics is transforming our understanding of these enigmatic molecules, offering profound insights into their structure, function, and therapeutic potential.</p><h3>What Are lncRNAs?</h3><p>lncRNAs are RNA transcripts longer than 200 nucleotides that do not code for proteins. Despite their non-coding nature, they play diverse roles in gene regulation, including chromatin remodeling, transcriptional control, and post-transcriptional processing. Unlike messenger RNAs (mRNAs), lncRNAs often function as scaffolds, decoys, or guides in cellular machinery, influencing biological processes such as cell differentiation, immune response, and even cancer metastasis.</p><h3>Challenges in lncRNA Research</h3><p>Identifying and understanding lncRNAs pose unique challenges:</p><ol>
<li><strong>High Sequence Variability</strong>: Unlike protein-coding genes, lncRNAs exhibit low sequence conservation across species, making functional predictions difficult.</li>
<li><strong>Low Expression Levels</strong>: lncRNAs are often expressed at low levels, complicating their detection in transcriptomic data.</li>
<li><strong>Diverse Functions</strong>: The multifunctional nature of lncRNAs requires advanced computational tools to decipher their roles in complex networks.</li>
</ol><h3>Bioinformatics: A Crucial Ally in lncRNA Research</h3><p>Bioinformatics bridges the gap between raw biological data and meaningful insights, making it indispensable in lncRNA research. Here&rsquo;s how:</p><h4>1. <strong>Identification and Annotation</strong></h4><p>High-throughput sequencing technologies like RNA-seq generate vast amounts of data. Bioinformatics tools such as <em>StringTie</em>, <em>Cufflinks</em>, and <em>HISAT2</em> help assemble and annotate lncRNAs from this data. Additionally, databases like NONCODE, LNCipedia, and Ensembl provide curated repositories of lncRNA sequences and annotations.</p><h4>2. <strong>Functional Prediction</strong></h4><p>Bioinformatics algorithms predict the potential functions of lncRNAs by analyzing their interactions with DNA, RNA, and proteins. Tools like LncRNA2Function and RIblast utilize sequence motifs and secondary structure predictions to hypothesize about the roles of specific lncRNAs.</p><h4>3. <strong>Network Construction</strong></h4><p>lncRNAs often act as regulatory hubs. Bioinformatics platforms such as Cytoscape enable the visualization of lncRNA-mediated networks, elucidating their roles in pathways like cell cycle regulation and apoptosis.</p><h4>4. <strong>Epigenetic Studies</strong></h4><p>lncRNAs are known to interact with chromatin-modifying complexes, influencing gene expression epigenetically. Tools like ChIP-seq and ATAC-seq, combined with computational pipelines, identify these interactions and map them to the genome.</p><h4>5. <strong>Clinical Applications</strong></h4><p>Bioinformatics aids in the discovery of lncRNA biomarkers for diseases like cancer and neurodegenerative disorders. Machine learning models analyze differential expression profiles, helping prioritize lncRNAs with therapeutic potential.</p><h3>Case Study: lncRNAs in Cancer Research</h3><p>lncRNAs such as HOTAIR and MALAT1 have been implicated in cancer progression. Bioinformatics analyses have revealed their roles in promoting metastasis and altering the tumor microenvironment. For example, transcriptome analysis in cancer patients identifies lncRNA expression signatures, enabling precision medicine approaches.</p><h3>Future Directions</h3><p>The fusion of bioinformatics with experimental biology is unlocking the secrets of lncRNAs. Advances in artificial intelligence, single-cell sequencing, and structural modeling promise to overcome current limitations. Here are some promising directions:</p><ul>
<li><strong>Integrative Analysis</strong>: Combining multi-omics data to understand the interplay of lncRNAs with other biomolecules.</li>
<li><strong>CRISPR Screens</strong>: Leveraging bioinformatics to design CRISPR-based functional screens for lncRNAs.</li>
<li><strong>Therapeutic Development</strong>: Using bioinformatics to design lncRNA-based therapeutics, including antisense oligonucleotides and RNA interference tools.</li>
</ul><h3>Conclusion</h3><p>lncRNAs are the hidden gems of the genome, and bioinformatics is the key to unearthing their full potential. As research progresses, lncRNAs could pave the way for novel diagnostics, targeted therapies, and personalized medicine, revolutionizing our approach to complex diseases.</p><p>The journey into the world of lncRNAs is only beginning, and bioinformatics will continue to play a pivotal role in decoding these molecular mysteries. Whether you&rsquo;re a researcher, clinician, or bioinformatics enthusiast, the study of lncRNAs offers a fascinating frontier of discovery.</p>]]></description>
	<dc:creator>LEGE</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/42633/protocol-for-de-novo-genome-assembly-using-illumina-reads</guid>
	<pubDate>Sat, 16 Jan 2021 21:42:11 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/42633/protocol-for-de-novo-genome-assembly-using-illumina-reads</link>
	<title><![CDATA[Protocol for De novo Genome Assembly using Illumina Reads]]></title>
	<description><![CDATA[<p>In this protocol, we address and describe the de novo assembly method for small to medium-sized genomes.</p><p><strong>What is de novo genome assembly?<br /></strong>The method of taking a large number of short DNA sequences and placing them back together to create a reflection of the original chromosomes from which the DNA originated relates to genome assembly. No previous knowledge of the source DNA sequence length, structure or composition is inferred by De novo genome assemblies. The DNA of the target organism is split up into millions of tiny parts and read on a sequencing computer in a genome sequencing experiment. Depending on the sequencing system used, these "reads" range from 20 to 1000 nucleotide base pairs (bp) in length. Usually, length reads of 36 - 150 bp are produced for Illumina style short read sequencing. These reads can be either &ldquo;single ended&rdquo; as described above or &ldquo;paired end.&rdquo;</p><p><strong>Why genome assembly?</strong><br />In basic research into why and how they live, as well as in applied topics, identifying the DNA sequence of an organism is useful. Awareness of a DNA sequence may be useful in virtually any biological research because of the relevance of DNA to living things. For example, it may be used in medicine to classify, diagnose and eventually improve genetic disorder therapies. Similarly, pathogens study can lead to treatments for infectious diseases.</p><p><strong>Raw NGS data</strong><br />Reads can be saved as a Fasta file as text or in a FastQ file with their attributes.&nbsp;FastQ is the most common read file format since this is what the Illumina sequencing pipeline creates. This will henceforth be the subject of our conversation.</p><p><strong>In a nutshell the protocol:</strong> <br />Get the sequence file(s) read from the sequencing machine (s). <br />Look at the readings - have an idea of what you have and what the standard is like. <br />If required, raw data cleanup/quality trimming. <br />Choose an adequate parameter set for assembly. <br />Assemble the data into scaffolds/contigs. <br />Examine the assembly performance and determine the efficiency of the assembly.</p><p><strong>Read Quality Control:</strong><br />Check the qualiy with fastQC.<br />Script<br />https://bioinformaticsonline.com/snippets/view/42540/install-fastqc-using-conda</p><p>Quality trimming/cleanup of read files.<br />This function trims adapters, barcodes and other contaminants from the reads.<br />Script<br />https://bioinformaticsonline.com/snippets/view/42542/trimmomatic-command</p><p><strong>Genome Assembly:</strong><br />The object of this portion of the protocol is to explain the method of assembling the reads trimmed by quality into draft contigs.</p><blockquote><p>spades.py -1 illumina_R1.fastq.gz -2 illumina_R2.fastq.gz --careful --cov-cutoff auto -o result_of_spades_assembly_all_illumina</p></blockquote><p>A significant range of short-read assemblers are available. Everyone with strengths and disadvantages of their own. <br /><em>Some of the assemblers available include:</em><br />Velvet<br />SOAP-denovo<br />MIRA<br />ALLPATHS</p><p>Next step is to assess the suitability and what to do with a draft package of contiguous details for the remainder of the study now.&nbsp;Few stuff you can note about the contigs you just created:&nbsp;They're the draft Contigs. Any mis-assemblies can occur.</p><p><strong>Mis-assembly checking and assembly metric tools:</strong><br />QUAST - Quality assessment tool for genome assembly http://bioinf.spbau.ru/quast<br />Mauve assembly metrics - http://code.google.com/p/ngopt/wiki/How_To_Score_Genome_Assemblies_with_Mauve<br />InGAP-SV - https://sites.google.com/site/nextgengenomics/ingap and http://ingap.sourceforge.net/<br />inGAP is also useful for finding structural variants between genomes from read mappings.</p><p><strong>Genome finishing tools:</strong><br />Semi-automated gap fillers:<br />Gap filler - http://www.baseclear.com/landingpages/basetools-a-wide-range-of-bioinformatics-solutions/gapfiller/</p><p>IMAGE (V2) - http://sourceforge.net/apps/mediawiki/image2/index.php?title=Main_Page</p><p><strong>Genome visualisers and editors:</strong><br />Artemis - http://www.sanger.ac.uk/resources/software/artemis/<br />IGV - http://www.broadinstitute.org/igv/</p><p><strong>Automated and semi automated annotation tools:</strong><br />Prokka - https://github.com/tseemann/prokka<br />RAST - http://www.nmpdr.org/FIG/wiki/view.cgi/FIG/RapidAnnotationServer<br />JCVI Annotation Service - http://www.jcvi.org/cms/research/projects/annotation-service/</p><p><strong>Frequent command use for the analysis are at:</strong></p><p>https://bioinformaticsonline.com/blog/view/38765/list-of-tools-frequently-used-while-genome-assembly<br />https://bioinformaticsonline.com/pages/view/42275/frequent-parameters-for-bioinformatics-tools</p>]]></description>
	<dc:creator>BioStar</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/44720/a-beginners-guide-to-using-kraken-for-taxonomic-classification</guid>
	<pubDate>Fri, 13 Dec 2024 11:29:03 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/44720/a-beginners-guide-to-using-kraken-for-taxonomic-classification</link>
	<title><![CDATA[A Beginner&#039;s Guide to Using Kraken for Taxonomic Classification]]></title>
	<description><![CDATA[<div>Kraken is a popular bioinformatics tool designed for fast and accurate taxonomic classification of metagenomic sequences. Its efficiency and precision make it a go-to resource for analyzing microbial communities, including bacteria, viruses, archaea, and fungi. Whether you're new to bioinformatics or experienced in the field, Kraken is an indispensable tool for taxonomic analysis.</div><div><div><div><div dir="auto"><div><div><p>In this blog, we&rsquo;ll walk through the basics of Kraken, from installation to running an analysis, and highlight its key features and applications.</p><h4><strong>What is Kraken?</strong></h4><p>Kraken is a sequence classification tool that assigns taxonomic labels to DNA sequences using exact k-mer matching. It uses a reference database of genomes, dividing sequences into k-mers and identifying matches in a computationally efficient way.</p><h4><strong>Key Features of Kraken</strong></h4><ul>
<li><strong>Speed</strong>: Kraken processes data much faster than alignment-based methods.</li>
<li><strong>Accuracy</strong>: It uses a precise k-mer matching algorithm for high-resolution taxonomic assignments.</li>
<li><strong>Scalability</strong>: It can handle large metagenomic datasets.</li>
<li><strong>Custom Databases</strong>: You can build and use custom databases tailored to your research needs.</li>
</ul><h4><strong>Installing Kraken</strong></h4><ol>
<li>
<p><strong>System Requirements</strong></p>
<ul>
<li>A Unix-based operating system (Linux/macOS).</li>
<li>Sufficient computational resources for database building (RAM and disk space).</li>
</ul>
</li>
<li>
<p><strong>Installation Steps</strong></p>
<ul>
<li>Clone the Kraken repository from GitHub:
<div>
<div>&nbsp;</div>
<div dir="ltr"><code>git <span style="font-size: 12.8px; font-weight: normal;">clone</span> https://github.com/DerrickWood/kraken.git <span style="font-size: 12.8px; font-weight: normal;">cd</span> kraken </code></div>
</div>
</li>
<li>Compile the Kraken binaries:
<div>
<div>&nbsp;</div>
<div dir="ltr"><code>make </code></div>
</div>
</li>
<li>Add Kraken to your PATH for easy access:
<div>
<div>&nbsp;</div>
<div dir="ltr"><code><span style="font-size: 12.8px; font-weight: normal;">export</span> PATH=<span style="font-size: 12.8px; font-weight: normal;">$PATH</span>:/path/to/kraken </code></div>
</div>
</li>
</ul>
</li>
</ol><h4><strong>Preparing a Database</strong></h4><p>Kraken requires a database of reference genomes. You can use a pre-built database or create a custom one.</p><ol>
<li>
<p><strong>Downloading a Pre-built Database</strong><br />Kraken offers pre-built databases, such as the <em>MiniKraken</em> database, which is lightweight and suitable for smaller datasets. Download it using:</p>
<div>
<div dir="ltr"><code>kraken-build --download-library minikraken </code></div>
</div>
</li>
<li>
<p><strong>Building a Custom Database</strong><br />To include specific genomes, download FASTA files and build the database:</p>
<div>
<div dir="ltr"><code>kraken-build --download-library bacteria --threads 4 --db my_database kraken-build --build --db my_database </code></div>
</div>
<p>This process may take considerable time and resources, depending on the size of the database.</p>
</li>
</ol><h4><strong>Running Kraken</strong></h4><p>Once the database is ready, you can classify sequences.</p><ol>
<li>
<p><strong>Basic Usage</strong><br />Use the following command to classify sequences:</p>
<div>
<div dir="ltr"><code>kraken --db my_database --threads 4 --fastq-input input_sequences.fastq --output kraken_output.txt </code></div>
</div>
<p>Key options:</p>
<ul>
<li><code>--db</code>: Specifies the database.</li>
<li><code>--threads</code>: Number of threads for parallel processing.</li>
<li><code>--fastq-input</code>: Indicates input file format (FASTQ/FASTA).</li>
</ul>
</li>
<li>
<p><strong>Interpreting Results</strong><br />Kraken generates an output file with columns for sequence IDs, taxonomic classifications, and the confidence score.</p>
</li>
</ol><h4><strong>Visualizing Kraken Results</strong></h4><p>Kraken results can be visualized using tools like <strong>Krona</strong> or converted to human-readable reports using <code>kraken-report</code>.</p><ol>
<li>
<p><strong>Generate a Report</strong></p>
<div>
<div dir="ltr"><code>kraken-report --db my_database kraken_output.txt &gt; kraken_report.txt </code></div>
</div>
</li>
<li>
<p><strong>Krona Visualization</strong><br />Install Krona and convert Kraken output for visualization:</p>
<div>
<div dir="ltr"><code>cut -f2,3 kraken_output.txt | ktImportTaxonomy -o krona_output.html </code></div>
</div>
<p>Open the HTML file in your browser to interactively explore the taxonomic classifications.</p>
</li>
</ol><h4><strong>Advanced Usage</strong></h4><ol>
<li>
<p><strong>Confidence Thresholds</strong><br />Adjust the confidence threshold for classification using the <code>--confidence</code> option. Higher values reduce false positives but may miss some true positives:</p>
<div>
<div dir="ltr"><code>kraken --db my_database --confidence 0.1 --fastq-input input.fastq </code></div>
</div>
</li>
<li>
<p><strong>Paired-End Reads</strong><br />For paired-end sequencing data, use:</p>
<div>
<div dir="ltr"><code>kraken --db my_database --paired reads_1.fastq reads_2.fastq </code></div>
</div>
</li>
<li>
<p><strong>Customizing K-mers</strong><br />Kraken allows you to set custom k-mer lengths during database building for specific applications.</p>
</li>
</ol><h4><strong>Applications of Kraken</strong></h4><ul>
<li><strong>Microbial Ecology</strong>: Characterizing microbial communities in soil, water, and the human microbiome.</li>
<li><strong>Pathogen Detection</strong>: Identifying pathogens in clinical samples.</li>
<li><strong>Fungal Research</strong>: Analyzing fungal diversity in metagenomic datasets.</li>
<li><strong>Environmental Monitoring</strong>: Tracking microbial populations in diverse habitats.</li>
</ul><h4><strong>Conclusion</strong></h4><p>Kraken is a versatile and efficient tool for taxonomic classification in metagenomics. Its speed, accuracy, and flexibility make it a favorite among bioinformaticians. By following this guide, you can set up and use Kraken to unlock insights into microbial and fungal communities, paving the way for discoveries in ecology, medicine, and biotechnology.</p></div></div></div></div></div></div>]]></description>
	<dc:creator>Neel</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/43652/peregrine-shimmer-genome-assembly-toolkit</guid>
	<pubDate>Thu, 16 Dec 2021 02:50:19 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/43652/peregrine-shimmer-genome-assembly-toolkit</link>
	<title><![CDATA[Peregrine &amp; SHIMMER Genome Assembly Toolkit]]></title>
	<description><![CDATA[<p><span>Peregrine is a fast genome assembler for accurate long reads (length &gt; 10kb, accuracy &gt; 99%). It can assemble a human genome from 30x reads within 20 cpu hours from reads to polished consensus. It uses Sparse HIereachical MimiMizER (SHIMMER) for fast read-to-read overlaping without quadratic comparisions used in other OLC assemblers.</span></p><p>Address of the bookmark: <a href="https://github.com/cschin/Peregrine" rel="nofollow">https://github.com/cschin/Peregrine</a></p>]]></description>
	<dc:creator>Abhi</dc:creator>
</item>

</channel>
</rss>