<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/32709?offset=330</link>
	<atom:link href="https://bioinformaticsonline.com/related/32709?offset=330" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/26629/computer-simulation-of-genetic-mechanism</guid>
	<pubDate>Sun, 13 Mar 2016 09:29:56 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/26629/computer-simulation-of-genetic-mechanism</link>
	<title><![CDATA[Computer simulation of genetic mechanism !!]]></title>
	<description><![CDATA[<p>Computer simulation is the discipline of designing a model of an actual or theoretical physical/biological system, executing the model on a digital computer, and analyzing the execution output. Simulation embodies the principle of ``learning by doing'' --- to learn about the system we must first build a model of some sort and then operate the model. The use of simulation is an activity that is as natural as a child who role plays. Children understand the world around them by simulating (with toys and figurines) most of their interactions with other people, animals and objects. As adults, we lose some of this childlike behavior but recapture it later on through computer simulation. To understand reality and all of its complexity, we must build artificial objects and dynamically act out roles with them. Computer simulation is the electronic equivalent of this type of role playing and it serves to drive synthetic environments and virtual worlds. Within the overall task of simulation, there are three primary sub-fields: model design, model execution and model analysis<br /><br />Simulation models have become important tools in Bioinformatics studies. There are many reasons for this, but we emphasize three of the more important:</p><p>(1) they enable exploration of hypotheses, and as such, have become invaluable means to guide research;</p><p>(2) they are unique approaches to integrate (in the literal term of the word) biological knowledge, in the form of experimental results; and</p><p>(3) they enable connecting biology with other fields of study ranging from physiology to genomics;</p><p>This blog, and this software list, is intended to guide the potential user of simulation models.<br />It is not, in any way, meant to be comprehensive on the very diverse simulation tools that already exist, but focuses on mechanistic, dynamic models. Similarly, it is not meant to provide any coverage of the breadth of applications; however, for interested readers, we provide references to use as a possible starting point.<br /><br />Simulation models are meant to answer questions which scientists have in a dynamic, quantitative, and often, a pictorial way. Much of the bioinformatics research and its applications, in particular, involve a large number of components, actors, and factors. Assembling these in a coherent framework may seem a daunting task, especially for beginners, and can lead to confusion, even for experienced scientists, especially if the objectives of such an exercise are not well defined. Followings are the list of tools bioinformatician may use to analyze and provide answers to complex biological mechanisms and related problems.</p><p style="margin-bottom: 0in;">&nbsp;</p><table width="718" cellspacing="0" cellpadding="2"><colgroup><col width="134"> <col width="501"> </colgroup>
<tbody>
<tr><th style="border: none; padding: 0in;">
<p>Software Resource</p>
</th><th style="border: none; padding: 0in;">
<p>Brief Description and Homepage</p>
</th></tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/aladyn/">Aladyn </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Tools to investigate how demographic parameters, populations genetics and abiotic conditions affect the rate of adaptation <br /><a href="http://www.katja-schiffers.eu/research.html">http://www.katja-schiffers.eu/research.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/alf/">ALF </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A Simulation Framework for Genome Evolution <br /><a href="http://www.cbrg.ethz.ch/alf">http://www.cbrg.ethz.ch/alf</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/art/">ART </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>ART is a set of simulation tools to generate synthetic next-generation sequencing data by mimicking real sequencing process with empirical error models or quality profiles. <br /><a href="http://www.niehs.nih.gov/research/resources/software/biostatistics/art/">http://www.niehs.nih.gov/research/resources/software/biostatistics/art/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/bamsurgeon/">BAMSurgeon </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Methods for realistic simulation of mutations in real data. <br /><a href="https://github.com/adamewing/bamsurgeon">https://github.com/adamewing/bamsurgeon</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/bayesian-serial-simcoal/">Bayesian Serial SimCoal </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Bayesian Serial SimCoal, (BayeSSC) is a modification of SIMCOAL 1.0, a program written by Laurent Excoffier, John Novembre, and Stefan Schneider. <br /><a href="http://www.stanford.edu/group/hadlylab/ssc/index.html">http://www.stanford.edu/group/hadlylab/ssc/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/baysics/">BaySICS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An integral platform with a graphical interface for statistical inference based on approximate Bayesian computation. <br /><a href="https://sites.google.com/site/baysicsabc/">https://sites.google.com/site/baysicsabc/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/beers/">BEERS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>BEERS was designed to benchmark RNA-Seq alignment algorithms and also algorithms that aim to reconstruct different isoforms and alternate splicing from RNA-Seq data <br /><a href="http://cbil.upenn.edu/BEERS/">http://cbil.upenn.edu/beers/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/bottleneck/">BOTTLENECK </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Bottleneck is a program for detecting recent effective population size reductions from allele data frequencies <br /><a href="http://www.ensam.inra.fr/URLB/bottleneck/bottleneck.html">http://www.ensam.inra.fr/urlb/bottleneck/bottleneck.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/bottlesim/">BottleSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>BottleSim is a computer simulation program for simulating the process of population bottlenecks <br /><a href="http://chkuo.name/software/BottleSim.html">http://chkuo.name/software/bottlesim.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/cass/">CASS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Protein Sequence Simulation <br /><a href="https://liberles.cst.temple.edu/Software/CASS/index.html">https://liberles.cst.temple.edu/software/cass/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/cdpop/">CDPOP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>CDPOP is a landscape genetics tool for simulating the emergence of spatial genetic structure in populations resulting from specified landscape processes governing organism movement behavior. <br /><a href="http://cel.dbs.umt.edu/CDPOP">http://cel.dbs.umt.edu/cdpop</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/classical-genetics-simulator/">Classical Genetics Simulator </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Web-based simulation software <br /><a href="http://www.cgslab.com/">http://www.cgslab.com/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/coasim/">CoaSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>CoaSim is a tool for simulating the coalescent process with recombination and geneconversion under various demographic models. <br /><a href="http://users-birc.au.dk/mailund/CoaSim/index.html">http://users-birc.au.dk/mailund/coasim/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/cosi/">cosi </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>The cosi package is written in C and is available as a tar file. <br /><a href="http://www.broadinstitute.org/%7Esfs/cosi/">http://www.broadinstitute.org/~sfs/cosi/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/cs-pseq-gen/">CS-PSeq-Gen </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A program to simulate the evolution of protein sequences under the constraints of the information of a particular reconstructed phylogeny <br /><a href="http://bioserv.rpbs.univ-paris-diderot.fr/software/CS-PSeq-Gen/">http://bioserv.rpbs.univ-paris-diderot.fr/software/cs-pseq-gen/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/dawg/">DAWG </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An application designed to simulate the evolution of recombinant DNA sequences in continuous time <br /><a href="http://scit.us/projects/dawg">http://scit.us/projects/dawg</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/easypop/">Easypop </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>EASYPOP is an individual based model intended to simulate datasets under a very broad range of conditions <br /><a href="http://www.unil.ch/dee/en/home/menuinst/softwares--dataset/softwares/easypop.html">http://www.unil.ch/dee/en/home/menuinst/softwares--dataset/softwares/easypop.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/egglib/">EggLib </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>EggLib is a C++/Python library and program package for evolutionary genetics and genomics. <br /><a href="http://egglib.sourceforge.net/">http://egglib.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/episim/">EpiSIM </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>EpiSIM: simulation of multiple epistasis, linkage disequilibrium patterns and haplotype blocks for genome-wide interaction analysis <br /><a href="https://sourceforge.net/projects/episimsimulator/files/">https://sourceforge.net/projects/episimsimulator/files/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/evolsimulator/">EvolSimulator </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A simulation test bed for hypotheses of genome evolution <br /><a href="http://acb.qfab.org/acb/evolsim/">http://acb.qfab.org/acb/evolsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/evolveagene/">EvolveAGene </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A realistic coding sequence simulation program that separates mutation from selection and allows the user to set selection conditions <br /><a href="http://bellinghamresearchinstitute.com/software/index.html">http://bellinghamresearchinstitute.com/software/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/fastsimcoal/">fastsimcoal </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A continuous-&shy;‐time coalescent simulator of genomic diversity under arbitrarily complex evolutionary scenarios <br /><a href="http://cmpg.unibe.ch/software/fastsimcoal/">http://cmpg.unibe.ch/software/fastsimcoal/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/fastslink/">FastSLINK </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulation of Marker and Phenotype Data in Pedigrees <br /><a href="https://watson.hgen.pitt.edu/">https://watson.hgen.pitt.edu/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/ffpopsim/">FFPopSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>C++/Python library for population genetics. <br /><a href="http://webdav.tuebingen.mpg.de/ffpopsim/">http://webdav.tuebingen.mpg.de/ffpopsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/flux-simulator/">FLUX SIMULATOR </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>The Flux Simulator aims at providing a deterministic in silico reproduction of the experimental pipelines for RNA-Seq, employing a minimal set of parameters. <br /><a href="http://sammeth.net/confluence/display/SIM/Home">http://sammeth.net/confluence/display/sim/home</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/forqs/">forqs </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Forward-in-time simulation of Recombination, Quantitative Traits, and Selection <br /><a href="https://bitbucket.org/dkessner/forqs">https://bitbucket.org/dkessner/forqs</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/forsim/">ForSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>ForSim: A Forward Evolutionary Computer Simulation <br /><a href="http://anth.la.psu.edu/research/weiss-lab/research/research">http://anth.la.psu.edu/research/weiss-lab/research/research</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/forwsim/">ForwSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>The program given below is based on the algorithm described in Padhukasahasram et al. 2008 to simulate genetic drift in a standard Wright-Fisher process. <br /><a href="http://badri-populationgeneticsimulators.blogspot.com/">http://badri-populationgeneticsimulators.blogspot.com/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/fpg/">FPG </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Forward Population Genetic simulation <br /><a href="https://bio.cst.temple.edu/%7Ehey/software/software.htm#FPG">https://bio.cst.temple.edu/~hey/software/software.htm#fpg</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/fregene/">FREGENE </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>FREGENE is a C++ program that simulates sequence-like data over large genomic regions in large diploid populations. <br /><a href="http://www.ebi.ac.uk/projects/BARGEN">http://www.ebi.ac.uk/projects/bargen</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/frequency-based-insilico-genome-generator-figg/">FIGG </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>FIGG is a genome simulation tool that uses known or theorized variation frequency, per a given fragment size and grouped by GC content across a genome to model new genomes in FASTA format while tracking applied mutations for use in analysis <br /><a href="http://insilicogenome.sourceforge.net/">http://insilicogenome.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/fwdpp/">fwdpp </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A C++ template library for implementing efficient forward simulations. <br /><a href="http://molpopgen.github.io/fwdpp/">http://molpopgen.github.io/fwdpp/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gametes/">GAMETES </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Genetic Architecture Model Emulator for Testing and Evaluating Software: Simulates complex SNP models with pure, strict epistatic interactions with n-loci. <br /><a href="http://sourceforge.net/projects/gametes/?source=navbar">http://sourceforge.net/projects/gametes/?source=navbar</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gasp/">GASP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Genometric Analysis Simulation Program. A software tool for testing and investigating methods in statistical genetics by generating samples of family data based on user specified models. <br /><a href="http://research.nhgri.nih.gov/gasp/">http://research.nhgri.nih.gov/gasp/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gcta/">GCTA </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Genome-wide Complex Trait Analysis <br /><a href="http://www.complextraitgenomics.com/software/gcta/download.html">http://www.complextraitgenomics.com/software/gcta/download.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gemsim/">GemSIM </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Next generation sequencing read simulator <br /><a href="http://sourceforge.net/projects/gemsim/">http://sourceforge.net/projects/gemsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/geneartisan/">GeneArtisan </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulation of Markers in Case-Control Study Designs <br /><a href="http://www.rannala.org/?page_id=241">http://www.rannala.org/?page_id=241</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/genome/">GENOME </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A rapid coalescent-based whole genome simulator <br /><a href="http://www.sph.umich.edu/csg/liang/genome/">http://www.sph.umich.edu/csg/liang/genome/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/genomepop2/">GenomePop2 </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>GenomePop2 is a specialization of the program GenomePop just to manage SNPs under more flexible and useful settings. If you need models with more than 2 alleles please use the GenomePop program version. <br /><a href="https://ritchielab.psu.edu/research/research-areas/statistical-genetics-and-gen-epi/methods/genomesimla">https://ritchielab.psu.edu/research/research-areas/statistical-genetics-and-gen-epi/methods/genomesimla</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/genomesimla/">GenomeSimla </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>GenomeSIMLA is currently under development- however, we have a beta release that we are asking to be tested <br /><a href="http://chgr.mc.vanderbilt.edu/genomeSIMLA/">http://chgr.mc.vanderbilt.edu/genomesimla/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gens2/">GENS2 </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulates interactions among two genetic and one environmental factor and also allows for epistatic interactions. <br /><a href="https://sourceforge.net/projects/gensim/">https://sourceforge.net/projects/gensim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/gwasimulator/">GWAsimulator </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A rapid whole genome simulation program <br /><a href="http://biostat.mc.vanderbilt.edu/wiki/Main/GWAsimulator">http://biostat.mc.vanderbilt.edu/wiki/main/gwasimulator</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/hap-sample/">HAP-SAMPLE </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An association simulator for candidate regions or genome scans <br /><a href="http://www.hapsample.org/">http://www.hapsample.org/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/hapgen/">HAPGEN </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A simulator for the simulation of case control datasets at SNP markers <br /><a href="https://mathgen.stats.ox.ac.uk/genetics_software/hapgen/hapgen2.html">https://mathgen.stats.ox.ac.uk/genetics_software/hapgen/hapgen2.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/hapsim/">HapSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A simulation tool for generating haplotype data with pre-specified allele frequencies and LD coefficients <br /><a href="http://cran.r-project.org/web/packages/hapsim/index.html">http://cran.r-project.org/web/packages/hapsim/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/hapsimu/">HAPSIMU </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A program that simulates heterogeneous populations with various known and controllable structures under the continuous migration model or the discrete model <br /><a href="http://l.web.umkc.edu/liujian/">http://l.web.umkc.edu/liujian/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/ibdsim/">IBDsim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>IBDSim is a computer package for the simulation of genotypic data under general isolation by distance models. <br /><a href="http://raphael.leblois.free.fr/">http://raphael.leblois.free.fr/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/indel-seq-gen/">indel-Seq-Gen </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A biological sequence simulation program that simulates highly divergent DNA sequences and protein superfamilies <br /><a href="http://bioinfolab.unl.edu/%7Ecstrope/iSG/">http://bioinfolab.unl.edu/~cstrope/isg/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/indelible/">Indelible </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A powerful and flexible simulator of biological evolution <br /><a href="http://abacus.gene.ucl.ac.uk/software/indelible/">http://abacus.gene.ucl.ac.uk/software/indelible/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/invertfregene/">invertFREGENE </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>InvertFREGENE is a forward-in-time simulator of inversions in population genetic data <br /><a href="http://www.ebi.ac.uk/projects/BARGEN/">http://www.ebi.ac.uk/projects/bargen/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/kernalpop/">kernalPop </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A spatially explicit population genetic simulation engine <br /><a href="http://cran.r-project.org/src/contrib/Archive/kernelPop/">http://cran.r-project.org/src/contrib/archive/kernelpop/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/macs/">MaCS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Markovian Coalescent Simulator <br /><a href="http://www-hsc.usc.edu/%7Egarykche/">http://www-hsc.usc.edu/~garykche/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/marlin/">Marlin </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Marlin provides a user-friendly interface for performing forward-in-time population genetic simulations. <br /><a href="http://www.patrickmeirmans.com/software/Marlin.html">http://www.patrickmeirmans.com/software/marlin.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/mason/">Mason </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A package for the simulation of nucleotide data. <br /><a href="http://www.seqan.de/projects/mason/">http://www.seqan.de/projects/mason/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/mbs/">mbs </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>modifying Hudson's ms software to generate samples of DNA sequences with a biallelic site under selection <br /><a href="http://www.sendou.soken.ac.jp/esb/innan/InnanLab/software.html">http://www.sendou.soken.ac.jp/esb/innan/innanlab/software.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/mendels-accountant/">Mendel's Accountant </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Mendel's Accountant (MENDEL) is an advanced numerical simulation program for modeling genetic change over time and was developed collaboratively by Sanford, Baumgardner, Brewer, Gibson and ReMine <br /><a href="http://mendelsaccount.sourceforge.net/">http://mendelsaccount.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/metapopgen/">MetaPopGen </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulates genetics in large size metapopulations <br /><a href="https://sites.google.com/site/marcoandrello/metapopgen">https://sites.google.com/site/marcoandrello/metapopgen</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/metasim/">MetaSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A tool to generate collections of synthetic reads that reflect the diverse taxonomical composition of typical metagenome data sets <br /><a href="http://ab.inf.uni-tuebingen.de/software/metasim/">http://ab.inf.uni-tuebingen.de/software/metasim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/mlcoalsim/">mlcoalsim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Multilocus Coalescent Simulations <br /><a href="http://code.google.com/p/mlcoalsim-v1/">http://code.google.com/p/mlcoalsim-v1/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/ms/">ms </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>The purpose of this program is to allow one to investigate the statistical properties of such samples, to evaluate estimators or statistical tests, and generally to aid in the interpretation of polymorphism data sets. <br /><a href="http://home.uchicago.edu/%7Erhudson1/source/mksamples.html">http://home.uchicago.edu/~rhudson1/source/mksamples.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/mshot/">msHOT </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>The purpose of this program is to allow one to investigate the statistical properties of such samples, to evaluate estimators or statistical tests, and generally to aid in the interpretation of polymorphism data sets. <br /><a href="http://home.uchicago.edu/%7Erhudson1/">http://home.uchicago.edu/~rhudson1/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/msms/">msms </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A coalescent Simlation tool with selection. <br /><a href="http://www.mabs.at/ewing/msms/index.shtml">http://www.mabs.at/ewing/msms/index.shtml</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/myssp/">MySSP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A program for the simulation of DNA sequence evolution across a phylogenetic tree <br /><a href="http://www.rosenberglab.net/software.html">http://www.rosenberglab.net/software.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/nemo/">Nemo </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A forward-time, individual-based, genetically explicit, and stochastic simulation program designed to study the evolution of genetic markers, life history traits, and phenotypic traits in a flexible (meta-)population framework. <br /><a href="http://nemo2.sourceforge.net/">http://nemo2.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/netrecodon/">NetRecodon </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Coalescent simulation of coding DNA sequences with recombination (inter and intracodon), migration and demography <br /><a href="http://code.google.com/p/netrecodon/">http://code.google.com/p/netrecodon/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/oncosimulr/">OncoSimulR </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>BioConductor package for Forward Genetic Simulation of Cancer Progresion with Epistasis <br /><a href="https://github.com/rdiaz02/OncoSimul">https://github.com/rdiaz02/oncosimul</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/pedagog/">PEDAGOG </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Software for simulating eco-evolutionary population dynamics <br /><a href="https://bcrc.bio.umass.edu/pedigreesoftware/node/5">https://bcrc.bio.umass.edu/pedigreesoftware/node/5</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/phenosim/">phenosim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A tool to add phenotypes to simulated genotypes <br /><a href="http://evoplant.uni-hohenheim.de/doku.php?id=software:software">http://evoplant.uni-hohenheim.de/doku.php?id=software:software</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/phylosim/">PhyloSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An R package for the Monte Carlo simulation of sequence evolution <br /><a href="http://www.ebi.ac.uk/goldman-srv/phylosim/">http://www.ebi.ac.uk/goldman-srv/phylosim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/pirs/">pIRS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Profile-based Illumina pair-end reads simulator <br /><a href="https://code.google.com/p/pirs/">https://code.google.com/p/pirs/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/proteinevolver/">ProteinEvolver </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulation of protein evolution along phylogenies under structure-based substitution models <br /><a href="http://code.google.com/p/proteinevolver/">http://code.google.com/p/proteinevolver/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/qmsim/">QMSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>QTL and Marker Simulator <br /><a href="http://www.aps.uoguelph.ca/%7Emsargol/qmsim/">http://www.aps.uoguelph.ca/~msargol/qmsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/quantinemo/">quantiNEMO </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An individual-based program for the analysis of quantitative traits with explicit genetic architecture potentially under selection in a structured population <br /><a href="http://www2.unil.ch/popgen/softwares/quantinemo/">http://www2.unil.ch/popgen/softwares/quantinemo/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/recoal/">RECOAL </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulates new haplotype data from a reference population of haplotypes. <br /><a href="ftp://popgen.usc.edu/">ftp://popgen.usc.edu/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/recodon/">Recodon </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Coalescent simulation of coding DNA sequences with recombination, migration and demography <br /><a href="http://code.google.com/p/recodon/">http://code.google.com/p/recodon/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/rlsim/">rlsim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A package for simulating RNA-seq library preparation with parameter estimation <br /><a href="http://bit.ly/rlsim-git">http://bit.ly/rlsim-git</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/rmetasim/">Rmetasim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Rmetasim is a front-end for the metasim engine that is implemented as a package that runs in the statistical computing environment R <br /><a href="http://cran.r-project.org/web/packages/rmetasim/index.html">http://cran.r-project.org/web/packages/rmetasim/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/rna-seq-simulator/">RNA Seq Simulator </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>RSS takes SAM alignment files from RNA-Seq data and simulates over dispersed, multiple replica, differential, non-stranded RNA-Seq datasets. <br /><a href="http://useq.sourceforge.net/cmdLnMenus.html#RNASeqSimulator">http://useq.sourceforge.net/cmdlnmenus.html#rnaseqsimulator</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/rose/">Rose </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Random model of sequence evolution <br /><a href="http://bibiserv.techfak.uni-bielefeld.de/rose/">http://bibiserv.techfak.uni-bielefeld.de/rose/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/scrm/">scrm </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A coalescent simulator optimized for long sequences and large samples. <br /><a href="https://scrm.github.io/">https://scrm.github.io/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/selsim/">SelSim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SelSim is a program for Monte Carlo simulation of DNA polymorphism data for a recom- bining region within which a single bi-allelic site has experienced natural selection <br /><a href="http://www.well.ox.ac.uk/%7Espencer/SelSim/">http://www.well.ox.ac.uk/~spencer/selsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/seq-gen/">Seq-Gen </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An application for the Monte Carlo simulation of molecular sequence evolution along phylogenetic trees. <br /><a href="http://tree.bio.ed.ac.uk/software/seqgen/">http://tree.bio.ed.ac.uk/software/seqgen/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/seqpower/">SEQPower </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Statistical power analysis for sequence-based association studies <br /><a href="http://bioinformatics.org/spower/">http://bioinformatics.org/spower/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/seqsimla/">SeqSIMLA </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SeqSIMLA can simulate sequence data with user-specified disease and quantitative trait models. Family or unrelated case-control data can be simulated. <br /><a href="http://seqsimla.sourceforge.net/">http://seqsimla.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/serial-netevolve/">Serial NetEvolve </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A flexible utility for generating serially-sampled sequences along a tree or recombinant network <br /><a href="http://biorg.cis.fiu.edu/SNE/">http://biorg.cis.fiu.edu/sne/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/sfs_code/">SFS_CODE </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SFS_CODE can perform forward population genetic simulations under a general Wright-Fisher model with arbitrary migration, demographic, selective, and mutational effects. <br /><a href="http://sfscode.sourceforge.net/SFS_CODE/index/index.html">http://sfscode.sourceforge.net/sfs_code/index/index.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/sibsim/">SIBSIM </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Quantitative phenotype simulation in extended pedigrees <br /><a href="http://sourceforge.net/projects/sibsim/">http://sourceforge.net/projects/sibsim/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simadapt/">SimAdapt </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A spatially explicit, individual-based, forward-time, landscape-genetic simulation model combined with a landscape cellular automaton. <br /><a href="https://www.openabm.org/model/3137">https://www.openabm.org/model/3137</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simcoal2/">SIMCOAL2 </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A coalescent program for the simulation of complex recombination patterns over large genomic regions under various demographic models <br /><a href="http://cmpg.unibe.ch/software/simcoal2/">http://cmpg.unibe.ch/software/simcoal2/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simcopy/">SimCopy </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>An R package simulating the evolution of copy number profiles along a tree. <br /><a href="http://bit.ly/simcopy">http://bit.ly/simcopy</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simla/">SIMLA </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SIMLA is a SIMuLAtion program that generates data sets of families for use in Linkage and Association studies. <br /><a href="http://dmpi.duke.edu/simla-simulation-software-version-32">http://dmpi.duke.edu/simla-simulation-software-version-32</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simped/">SimPed </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A Simulation Program to Generate Haplotype and Genotype Data for Pedigree Structures <br /><a href="http://bioinformatics.org/simped/">http://bioinformatics.org/simped/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simprot/">Simprot </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A program to simulate protein evolution by substitution, insertion and deletion <br /><a href="http://www.uhnresearch.ca/labs/tillier/software.htm#3">http://www.uhnresearch.ca/labs/tillier/software.htm#3</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simrare/">SimRare </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Rare variant simulation and analysis tool <br /><a href="http://code.google.com/p/simrare/">http://code.google.com/p/simrare/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simugwas/">simuGWAS </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A forward-time simulator that simulates realistic samples for genome-wide association studies. <br /><a href="http://simupop.sourceforge.net/Cookbook/SimuGWAS">http://simupop.sourceforge.net/cookbook/simugwas</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/simupop/">simuPOP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>simuPOP is a general-purpose individual-based forward-time population genetics simulation environment. <br /><a href="http://simupop.sourceforge.net/">http://simupop.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/sissi/">SISSI </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A software tool to generate data of related sequences along a given phylogeny, taking into account user defined system of neighbourhoods and instantaneous rate matrices. <br /><a href="http://www.cibiv.at/software/sissi/">http://www.cibiv.at/software/sissi/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/smartpop/">SMARTPOP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulating Mating Alliance as a Reproductive Tactic for Populations <br /><a href="http://smartpop.sourceforge.net/">http://smartpop.sourceforge.net/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/snpsim/">SNPsim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Coalescent simulation of hotspot recombination <br /><a href="http://code.google.com/p/phylosoftware/">http://code.google.com/p/phylosoftware/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/spip/">SPIP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SPIP simulates the transmission of genes from parents to offspring in a population having demographic structure defined by the user <br /><a href="http://swfsc.noaa.gov/textblock.aspx?Division=FED&amp;id=3434">http://swfsc.noaa.gov/textblock.aspx?division=fed&amp;id=3434</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/splatche/">Splatche </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Spatial and Temporal Coalescences in Heterogeneous Environment <br /><a href="http://www.splatche.com/">http://www.splatche.com/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/srv/">srv </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Simulator of Rare Varaints (srv) is a simulator for the simulation of the introduction and evolution of (rare) genetic variants. <br /><a href="http://simupop.sourceforge.net/Cookbook/SimuRareVariants">http://simupop.sourceforge.net/cookbook/simurarevariants</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/sup/">SUP </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>SLINK/FastSLINK utility program <br /><a href="http://mlemire.freeshell.org/software.html">http://mlemire.freeshell.org/software.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/treesimj/">TreesimJ </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A flexible, forward-time population genetic simulator <br /><a href="http://code.google.com/p/treesimj/">http://code.google.com/p/treesimj/</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/variant-simulation-tools/">Variant Simulation Tools </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>A simulation tool for post-GWAS genetic epidemiological studies using whole-genome or whole-exome next-gen sequencing data, with an emphasis on user-friendliness and reproducibility. <br /><a href="http://varianttools.sourceforge.net/Simulation/HomePage">http://varianttools.sourceforge.net/simulation/homepage</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/vortex/">Vortex </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>VORTEX is an individual-based simulation model for population viability analysis (PVA). <br /><a href="http://www.vortex9.org/vortex.html">http://www.vortex9.org/vortex.html</a></p>
</td>
</tr>
<tr><th style="border: none; padding: 0in;">
<p><a href="https://popmodels.cancercontrol.cancer.gov/gsr/packages/wessim/">Wessim </a></p>
</th>
<td style="border: none; padding: 0in;">
<p>Whole Exome Sequencing SIMulator <br /><a href="http://sak042.github.io/Wessim/">http://sak042.github.io/wessim/</a></p>
</td>
</tr>
</tbody>
</table><p style="margin-bottom: 0in;">&nbsp;</p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/27331/andi</guid>
	<pubDate>Fri, 13 May 2016 05:16:35 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/27331/andi</link>
	<title><![CDATA[Andi]]></title>
	<description><![CDATA[<p>This is the <code>andi</code> program for estimating the evolutionary distance between closely related genomes. These distances can be used to rapidly infer phylogenies for big sets of genomes. Because <code>andi</code> does not compute full alignments, it is so efficient that it scales even up to thousands of bacterial genomes.</p>
<p>This readme covers all necessary instructions for the impatient to get <code>andi</code> up and running. For extensive instructions please consult the <a href="https://github.com/EvolBioInf/andi/blob/master/andi-manual.pdf">manual</a>.</p>
<p>More at https://github.com/evolbioinf/andi/</p><p>Address of the bookmark: <a href="http://bioinformatics.oxfordjournals.org/content/early/2015/01/13/bioinformatics.btu815.full" rel="nofollow">http://bioinformatics.oxfordjournals.org/content/early/2015/01/13/bioinformatics.btu815.full</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/pages/view/27799/bbmapbbtools-package-multipurpose-tool-designed-for-converting-reads-or-other-nucleotide-data-between-different-formats</guid>
	<pubDate>Mon, 13 Jun 2016 05:47:21 -0500</pubDate>
	<link>https://bioinformaticsonline.com/pages/view/27799/bbmapbbtools-package-multipurpose-tool-designed-for-converting-reads-or-other-nucleotide-data-between-different-formats</link>
	<title><![CDATA[BBMap/BBTools package: Multipurpose tool designed for converting reads or other nucleotide data between different formats.]]></title>
	<description><![CDATA[<div id="post_message_148585"><a href="https://sourceforge.net/projects/bbmap/" target="_blank">Reformat</a>is a member of the <a href="https://sourceforge.net/projects/bbmap/" target="_blank">BBMap/BBTools package</a>. It is a multipurpose tool designed for converting reads or other nucleotide data between different formats. It supports, and can inter-convert:<br /> <br /> fastq<br /> fasta<br /> fasta+qual<br /> sam<br /> scarf (an old Illumina format)<br /> bam (if samtools is installed)<br /> gzip<br /> zip<br /> ascii-33 (sanger)<br /> ascii-64 (old Illumina)<br /> paired files<br /> interleaved files<br /> <br /> It is multithreaded and can process data at over 500 megabytes per second, and can accept streams from standard in and write to standard out, allowing it to be easily dropped into the middle of a pipeline for format conversion. Reformat autodetects formats based on file extensions and content, making it very easy to use; and the autodetection can be overridden, allowing flexibility for people who don't like to follow naming conventions, or out-of-spec fastq files with qualities values like -17 or 120.<br /> <br /> The program has been gradually expanded, and can now perform various other functions. None of these will break pairing, if the input is paired.<br /> <br /> Quality trimming (either or both ends)<br /> Quality filtering<br /> Fixed-length trimming<br /> Generation of histograms (base composition, quality, etc)<br /> Subsampling (to a fraction of input reads, or an exact number of reads or bases)<br /> Changing fasta line-wrapping length<br /> Reverse-complementing (all reads or only read 2)<br /> Adding /1 and /2 suffix to read names<br /> GC-content filtering<br /> Length-filtering<br /> Testing for corrupted interleaved files<br /> <br /> Reformat is compatible with any platform that supports Java 1.7 or higher. It also has a bash shellscript for simpler invocation. Typical usage examples:<br /> <br /> Reformat fastq into fasta:<br /> <strong>reformat.sh in=x.fq out=y.fa</strong><br /> <br /> Interleave paired reads:<br /> <strong>reformat.sh in1=x1.fq in2=x2.fq out=y.fq</strong><br /> <br /> Note - you can actually use a shortcut if paired read files have the same name with a 1 and a 2. This is equivalent to the above command:<br /> <strong>reformat.sh in=x#.fq out=y.fq</strong><br /> <br /> De-interleave reads:<br /> <strong>reformat.sh in=x.fq out1=y1.fq out2=y2.fq</strong><br /> <br /> Verify that interleaving appears correct, assuming Illumina namimg conventions:<br /> <strong>reformat.sh in=x.fq vint</strong><br /> <br /> Convert ASCII-33 to ASCII-64:<br /> <strong>reformat.sh in=x.fq out=y.fq qin=33 qout=64</strong><br /> <br /> Quality-trim paired reads to Q10 on the left and right ends and discard reads shorter than 50bp after trimming:<br /> <strong>reformat.sh in1=x1.fq in2=x2.fq out1=y1.fq out2=y2.fq outsingle=singletons.fq qtrim=rl trimq=10 minlength=50</strong><br /> <br /> Subsample 10% of the first 20000 pairs in an interleaved file:<br /> <strong>reformat.sh in=x.fq out=y.fq reads=20000 samplerate=0.1 int=t</strong><br /> (in this case "int=t" overrides interleaving autodetection, to ensure reads are treated as pairs)<br /> <br /> Pipe in a gzipped sam file and pipe out fasta:<br /> <strong>reformat.sh in=stdin.sam.gz out=stdout.fa</strong><br /> <br /> Reverse-complement reads:<br /> <strong>reformat.sh in=x.fq out=y.fq rcomp</strong><br /> <br /> For reformatting a file with very long sequences, Reformat will need more memory; just add the additional flag "-Xmx2g". For example, to change the line-wrapping length on the human genome (which has individual sequences over 200Mbp long) to 70 characters:<br /> <strong>reformat.sh -Xmx2g in=HG19.fa.gz out=HG19_wrapped.fa.gz fastawrap=70</strong><br /> <br /> For additional functions, please run the shellscript with no arguments, or just read it with a text editor. If you have any questions, please post them in this thread.<br /> <br /> For people using a non-bash terminal, you may need to type "bash reformat.sh" instead of just "reformat.sh".<br /> For users of Windows or other platforms that do not support bash shellscripts, replace "reformat.sh" with "java -ea -Xmx200m /path/to/bbmap/current/ jgi.ReformatReads"<br /> for example,<br /> <strong>java -ea -Xmx200m C:\bbmap\current\ jgi.ReformatReads in=x.fq out=y.fa</strong><br /> <br /> Reformat can be downloaded with BBTools here:<br /> <a href="https://sourceforge.net/projects/bbmap/" target="_blank">https://sourceforge.net/projects/bbmap/</a></div>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/27839/lorma-a-tool-for-correcting-sequencing-errors-in-long-reads-such-those-produced-by-pacific-biosciences-sequencing-machines</guid>
	<pubDate>Wed, 15 Jun 2016 17:18:36 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/27839/lorma-a-tool-for-correcting-sequencing-errors-in-long-reads-such-those-produced-by-pacific-biosciences-sequencing-machines</link>
	<title><![CDATA[LoRMA: a tool for correcting sequencing errors in long reads such those produced by Pacific Biosciences sequencing machines]]></title>
	<description><![CDATA[<p>LoRMA is a tool for correcting sequencing errors in long reads such those produced by Pacific Biosciences sequencing machines.</p>
<p>Publication:</p>
<ul>
<li>L. Salmela, R. Walve, E. Rivals, and E. Ukkonen: Accurate selfcorrection of errors in long reads using de Bruijn graphs. Accepted to RECOMB-Seq 2016.</li>
</ul>
<p>Download:</p>
<ul>
<li><a href="https://www.cs.helsinki.fi/u/lmsalmel/LoRMA/LoRMA-0.3.tar.gz">LoRMA 0.3 source files</a></li>
<li><a href="https://www.cs.helsinki.fi/u/lmsalmel/LoRMA/README.txt">README</a></li>
</ul><p>Address of the bookmark: <a href="https://www.cs.helsinki.fi/u/lmsalmel/LoRMA/" rel="nofollow">https://www.cs.helsinki.fi/u/lmsalmel/LoRMA/</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/28168/sam-flags</guid>
	<pubDate>Wed, 29 Jun 2016 15:38:15 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/28168/sam-flags</link>
	<title><![CDATA[SAM flags]]></title>
	<description><![CDATA[<p>Decoding SAM flags</p>
<p>This utility makes it easy to identify what are the properties of a read based on its SAM flag value, or conversely, to find what the SAM Flag value would be for a given combination of properties.</p>
<p>To decode a given SAM flag value, just enter the number in the field below. The encoded properties will be listed under Summary below, to the right.</p><p>Address of the bookmark: <a href="https://broadinstitute.github.io/picard/explain-flags.html" rel="nofollow">https://broadinstitute.github.io/picard/explain-flags.html</a></p>]]></description>
	<dc:creator>Poonam Mahapatra</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/28121/kaiju</guid>
	<pubDate>Mon, 27 Jun 2016 11:23:04 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/28121/kaiju</link>
	<title><![CDATA[Kaiju]]></title>
	<description><![CDATA[<p>Kaiju is a program for the taxonomic classification of metagenomic high-throughput sequencing reads. Each read is directly assigned to a taxon within the NCBI taxonomy by comparing it to a reference database containing microbial and viral protein sequences.</p>
<p>By default, Kaiju uses either the available complete genomes from NCBI RefSeq or the microbial subset of the non-redundant protein database <em>nr</em> used by NCBI BLAST, optionally also including fungi and microbial eukaryotes.</p>
<p>Kaiju translates reads into amino acid sequences, which are then searched in the database using a modified backward search on a memory-efficient implementation of the Burrows-Wheeler transform, which finds maximum exact matches (MEMs), optionally allowing mismatches in the protein alignment. The search can process up to millions of reads per minute using, for example, only 10 GB RAM with a protein database comprising 4821 microbial genomes. Kaiju can also be used for querying any other protein database without taxonomic classification, using either protein or nucleotide queries.</p>
<p>Kaiju is described in <a href="http://www.nature.com/ncomms/2016/160413/ncomms11257/full/ncomms11257.html">Menzel, P. et al. (2016) Fast and sensitive taxonomic classification for metagenomics with Kaiju. <em>Nat. Commun.</em> 7:11257</a> (open access).</p><p>Address of the bookmark: <a href="http://kaiju.binf.ku.dk/" rel="nofollow">http://kaiju.binf.ku.dk/</a></p>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/28415/scarpa</guid>
	<pubDate>Wed, 13 Jul 2016 07:59:25 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/28415/scarpa</link>
	<title><![CDATA[Scarpa]]></title>
	<description><![CDATA[<p><strong>Scarpa</strong>&nbsp;is a stand-alone scaffolding tool for NGS data. It can be used together with virtually any genome assembler and any NGS read mapper that supports SAM format. Other features include support for multiple libraries and an option to estimate insert size distributions from data. Scarpa is available free of charge for academic and commercial use under the GNU General Public License (GPL).</p>
<p>See the&nbsp;<a href="http://compbio.cs.toronto.edu/hapsembler/hapsembler-2.21_manual.pdf">user manual</a>&nbsp;or the&nbsp;<a href="http://compbio.cs.toronto.edu/hapsembler/scarpa_paper.pdf">paper</a>&nbsp;for more information about Scarpa. Click&nbsp;<a href="http://compbio.cs.toronto.edu/hapsembler/ScarpaSupplementary.pdf">here</a>&nbsp;for the supplementary material.</p><p>Address of the bookmark: <a href="http://compbio.cs.toronto.edu/hapsembler/scarpa.html" rel="nofollow">http://compbio.cs.toronto.edu/hapsembler/scarpa.html</a></p>]]></description>
	<dc:creator>Poonam Mahapatra</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/pages/view/34418/spades-hybrid-genome-assembly</guid>
	<pubDate>Mon, 27 Nov 2017 08:05:40 -0600</pubDate>
	<link>https://bioinformaticsonline.com/pages/view/34418/spades-hybrid-genome-assembly</link>
	<title><![CDATA[SPAdes hybrid genome assembly]]></title>
	<description><![CDATA[<p>When you have both Illumina and Nanopore data, then SPAdes remains a good option for hybrid assembly - SPAdes was used to produce the&nbsp;<a href="https://gigascience.biomedcentral.com/articles/10.1186/s13742-015-0101-6">B fragilis assembly</a>&nbsp;by Mick Watson&rsquo;s group.</p><p>Again, running spades.py will show you the options:</p><div><pre><code>spades.py
</code></pre></div><p>This produces:</p><div><pre><code>SPAdes genome assembler v3.10.1

Usage: /usr/local/SPAdes-3.10.1-Linux/bin/spades.py [options] -o &lt;output_dir&gt;

Basic options:
-o      &lt;output_dir&gt;    directory to store all the resulting files (required)
--sc                    this flag is required for MDA (single-cell) data
--meta                  this flag is required for metagenomic sample data
--rna                   this flag is required for RNA-Seq data
--plasmid               runs plasmidSPAdes pipeline for plasmid detection
--iontorrent            this flag is required for IonTorrent data
--test                  runs SPAdes on toy dataset
-h/--help               prints this usage message
-v/--version            prints version

Input data:
--12    &lt;filename&gt;      file with interlaced forward and reverse paired-end reads
-1      &lt;filename&gt;      file with forward paired-end reads
-2      &lt;filename&gt;      file with reverse paired-end reads
-s      &lt;filename&gt;      file with unpaired reads
--pe&lt;#&gt;-12      &lt;filename&gt;      file with interlaced reads for paired-end library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--pe&lt;#&gt;-1       &lt;filename&gt;      file with forward reads for paired-end library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--pe&lt;#&gt;-2       &lt;filename&gt;      file with reverse reads for paired-end library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--pe&lt;#&gt;-s       &lt;filename&gt;      file with unpaired reads for paired-end library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--pe&lt;#&gt;-&lt;or&gt;    orientation of reads for paired-end library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9; &lt;or&gt; = fr, rf, ff)
--s&lt;#&gt;          &lt;filename&gt;      file with unpaired reads for single reads library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--mp&lt;#&gt;-12      &lt;filename&gt;      file with interlaced reads for mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--mp&lt;#&gt;-1       &lt;filename&gt;      file with forward reads for mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--mp&lt;#&gt;-2       &lt;filename&gt;      file with reverse reads for mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--mp&lt;#&gt;-s       &lt;filename&gt;      file with unpaired reads for mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--mp&lt;#&gt;-&lt;or&gt;    orientation of reads for mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9; &lt;or&gt; = fr, rf, ff)
--hqmp&lt;#&gt;-12    &lt;filename&gt;      file with interlaced reads for high-quality mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--hqmp&lt;#&gt;-1     &lt;filename&gt;      file with forward reads for high-quality mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--hqmp&lt;#&gt;-2     &lt;filename&gt;      file with reverse reads for high-quality mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--hqmp&lt;#&gt;-s     &lt;filename&gt;      file with unpaired reads for high-quality mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--hqmp&lt;#&gt;-&lt;or&gt;  orientation of reads for high-quality mate-pair library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9; &lt;or&gt; = fr, rf, ff)
--nxmate&lt;#&gt;-1   &lt;filename&gt;      file with forward reads for Lucigen NxMate library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--nxmate&lt;#&gt;-2   &lt;filename&gt;      file with reverse reads for Lucigen NxMate library number &lt;#&gt; (&lt;#&gt; = 1,2,..,9)
--sanger        &lt;filename&gt;      file with Sanger reads
--pacbio        &lt;filename&gt;      file with PacBio reads
--nanopore      &lt;filename&gt;      file with Nanopore reads
--tslr  &lt;filename&gt;      file with TSLR-contigs
--trusted-contigs       &lt;filename&gt;      file with trusted contigs
--untrusted-contigs     &lt;filename&gt;      file with untrusted contigs

Pipeline options:
--only-error-correction runs only read error correction (without assembling)
--only-assembler        runs only assembling (without read error correction)
--careful               tries to reduce number of mismatches and short indels
--continue              continue run from the last available check-point
--restart-from  &lt;cp&gt;    restart run with updated options and from the specified check-point ('ec', 'as', 'k&lt;int&gt;', 'mc')
--disable-gzip-output   forces error correction not to compress the corrected reads
--disable-rr            disables repeat resolution stage of assembling

Advanced options:
--dataset       &lt;filename&gt;      file with dataset description in YAML format
-t/--threads    &lt;int&gt;           number of threads
                                [default: 16]
-m/--memory     &lt;int&gt;           RAM limit for SPAdes in Gb (terminates if exceeded)
                                [default: 250]
--tmp-dir       &lt;dirname&gt;       directory for temporary files
                                [default: &lt;output_dir&gt;/tmp]
-k              &lt;int,int,...&gt;   comma-separated list of k-mer sizes (must be odd and
                                less than 128) [default: 'auto']
--cov-cutoff    &lt;float&gt;         coverage cutoff value (a positive float number, or 'auto', or 'off') [default: 'off']
--phred-offset  &lt;33 or 64&gt;      PHRED quality offset in the input reads (33 or 64)
                                [default: auto-detect]
</code></pre></div><p>As you can see this is also a &ldquo;pipeline&rdquo; of tools that can be switched on or off. SPAdes takes quite a long time, so for the purposes of this practical, something like this may suffice:</p><div><pre><code>spades.py -t 4 <span>\</span>
          -m 32 <span>\</span>
          -k 31,51,71 <span>\</span>
          --only-assembler <span>\</span>
          -1 miseq.1.fastq -2 miseq.2.fastq <span>\</span>
          --nanopore minion.fastq <span>\</span>
          -o hybrid_assembly
</code></pre></div><p>In turn, these parameters mean</p><ul>
<li>use 4 threads</li>
<li>max memory is 32Gb</li>
<li>use 3 kmer values to build the de bruijn graph(s) - 31, 51 and 71</li>
<li>only run the assembler, not the correction algorithm (for speed)</li>
<li>read 1 and read 2 of the MiSeq data</li>
<li>the nanopore data</li>
<li>put the output in folder &ldquo;hybrid_assembly&rdquo;</li>
</ul>]]></description>
	<dc:creator>Jit</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/34707/string-graph-based-genome-assembly-software-and-tools</guid>
	<pubDate>Tue, 19 Dec 2017 17:17:38 -0600</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/34707/string-graph-based-genome-assembly-software-and-tools</link>
	<title><![CDATA[String graph based genome assembly software and tools !]]></title>
	<description><![CDATA[<p>In&nbsp;<a href="https://en.wikipedia.org/wiki/Graph_theory" title="Graph theory">graph theory</a>, a&nbsp;<strong>string graph</strong>&nbsp;is an&nbsp;<a href="https://en.wikipedia.org/wiki/Intersection_graph" title="Intersection graph">intersection graph</a>&nbsp;of&nbsp;<a href="https://en.wikipedia.org/wiki/Curve" title="Curve">curves</a>&nbsp;in the plane; each curve is called a "string".&nbsp; String graphs were first proposed by E. W. Myers in a&nbsp;<a href="http://bioinformatics.oxfordjournals.org/content/21/suppl_2/ii79.full.pdf+html">2005 publication</a>.&nbsp;In&nbsp;recent&nbsp;<a href="http://genome.cshlp.org/content/early/2012/01/22/gr.126953.111">Genome Research paper</a>&nbsp;describing an innovative approach for assembling large genomes from NGS data caught our attention for several reasons. i) it give different "string graph" prospective of long lasting genome assembly problem ii) the&nbsp;paper is coauthored by Jared Simpson, the developer of&nbsp;<a href="http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2694472/">ABySS assembler</a>&nbsp;and Richard Durbin. iii)&nbsp;Simpson-Durbin algorithm is that it does not rely on de Bruijn graphs, and instead employs a different graph construction approach called &lsquo;string graph&rsquo;.</p><p>Following are the genome assembly tools based on string graph:</p><p>1.SGA (String Graph Assembler)&nbsp;https://github.com/jts/sga</p><p>Assembles large genomes from high coverage short read data. SGA is designed as a modular set of programs, which are used to form an assembly pipeline. SGA implements a set of assembly algorithms based on the FM-index. As the FM-index is a compressed data structure, the algorithms are very memory efficient. The SGA assembly has three distinct phases. The first phase corrects base calling errors in the reads. The second phase assembles contigs from the corrected reads. The third phase uses paired end and/or mate pair data to build scaffolds from the contigs. The output of this software is a PDF report that allows the properties of the genome and data quality to be visually explored. By providing more information to the user at the start of an assembly project, this software will help increase awareness of the factors that make a given assembly easy or difficult, assist in the selection of software and parameters and help to troubleshoot an assembly if it runs into problems.</p><p>2.&nbsp;SAGE: String-overlap Assembly of GEnomes&nbsp;https://github.com/lucian-ilie/SAGE2</p><p>SAGE, for de novo genome assembly. As opposed to most assemblers, which are de Bruijn graph based, SAGE uses the string-overlap graph. SAGE builds upon great existing work on string-overlap graph and maximum likelihood assembly, bringing an important number of new ideas, such as the efficient computation of the transitive reduction of the string overlap graph, the use of (generalized) edge multiplicity statistics for more accurate estimation of read copy counts, and the improved use of mate pairs and min-cost flow for supporting edge merging. The assemblies produced by SAGE for several short and medium-size genomes compared favourably with those of existing leading assemblers.</p><p>3. FSG: Fast String Graph</p><p>The new integrated assembler has been assessed on a standard benchmark, showing that fast string graph (FSG) is significantly faster than SGA while maintaining a moderate use of main memory, and showing practical advantages in running FSG on multiple threads. Moreover, we have studied the effect of coverage rates on the running times.</p><p>4.&nbsp;&nbsp;BASE&nbsp;https://github.com/dhlbh/BASE</p><p>It enhances the classic seed-extension approach by indexing the reads efficiently to generate adaptive seeds that have high probability to appear uniquely in the genome. Such seeds form the basis for BASE to build extension trees and then to use reverse validation to remove the branches based on read coverage and paired-end information, resulting in high-quality consensus sequences of reads sharing the seeds. Such consensus sequences are then extended to contigs.&nbsp;BASE is a practically efficient tool for constructing contig, with significant improvement in quality for long NGS reads. It is relatively easy to extend BASE to include scaffolding.</p><p>5.&nbsp;Fermi&nbsp;https://github.com/lh3/fermi/</p><p>Fermi is a de novo assembler with a particular focus on assembling Illumina&nbsp;short sequence reads from a mammal-sized genome. In addition to the role of a&nbsp;typical assembler, fermi also aims to preserve heterozygotes which are often&nbsp;collapsed by other assemblers. Its ultimate goal is to find a minimal set of&nbsp;unitigs to represent all the information in raw reads.</p><p>If you want to learn about String Graph assembler, please read the following papers -</p><p>i)&nbsp;<a href="http://bioinformatics.oxfordjournals.org/content/21/suppl_2/ii79.full.pdf+html">The Fragment Assembly String Graph - E. W. Myers</a></p><p>This paper describes the String Graph concept.</p><p>ii)&nbsp;<a href="http://bioinformatics.oxfordjournals.org/content/26/12/i367.full#ref-20">Efficient construction of an assembly string graph using the FM-index - Jared T. Simpson and Richard Durbin</a></p><p>This earlier paper from Simpson and Durbin</p><p>iii)&nbsp;<a href="http://genome.cshlp.org/content/early/2012/01/22/gr.126953.111">Efficient de novo assembly of large genomes using compressed data structures - Jared T. Simpson and Richard Durbin</a></p><p>&nbsp;</p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/36257/aligngraph-algorithm-for-secondary-de-novo-genome-assembly-guided-by-closely-related-references</guid>
	<pubDate>Tue, 17 Apr 2018 16:21:20 -0500</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/36257/aligngraph-algorithm-for-secondary-de-novo-genome-assembly-guided-by-closely-related-references</link>
	<title><![CDATA[AlignGraph: algorithm for secondary de novo genome assembly guided by closely related references]]></title>
	<description><![CDATA[<p>AlignGraph is a software that extends and joins contigs or scaffolds by reassembling them with help provided by a reference genome of a closely related organism.</p>
<p>Using AlignGraph</p>
<pre><code>AlignGraph --read1 reads_1.fa --read2 reads_2.fa --contig contigs.fa --genome genome.fa --distanceLow distanceLow --distanceHigh distancehigh --extendedContig extendedContigs.fa --remainingContig remainingContigs.fa [--kMer k --insertVariation insertVariation --coverage coverage --part p --fastMap --ratioCheck --iterativeMap --misassemblyRemoval --resume]</code></pre>
<h3>&nbsp;</h3><p>Address of the bookmark: <a href="https://github.com/baoe/AlignGraph" rel="nofollow">https://github.com/baoe/AlignGraph</a></p>]]></description>
	<dc:creator>Manisha Mishra</dc:creator>
</item>

</channel>
</rss>