<?xml version='1.0'?><rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:georss="http://www.georss.org/georss" xmlns:atom="http://www.w3.org/2005/Atom" >
<channel>
	<title><![CDATA[BOL: Related items]]></title>
	<link>https://bioinformaticsonline.com/related/29270?offset=1390</link>
	<atom:link href="https://bioinformaticsonline.com/related/29270?offset=1390" rel="self" type="application/rss+xml" />
	<description><![CDATA[]]></description>
	
	
<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/21021/ra-bioinformatics-at-iiser</guid>
  <pubDate>Fri, 06 Feb 2015 04:05:49 -0600</pubDate>
  <link></link>
  <title><![CDATA[RA Bioinformatics at IISER]]></title>
  <description><![CDATA[
<p>Advertisement: Research Position in Computational Biology in the group of Shree P. Pandey Positions available in the area of NGS data analysis, bioinformatics, plant genomics Project Description: Projects involves high throughput analysis of data mostly generated by massively parallel sequencing (RNA-Seq and small-RNA-Seq), microarrays and related platforms.</p>

<p>We are looking for highly motivated and bright individuals interested in high-throughput cutting-edge data analyses methods in genomics (computational positions). Available positions:</p>

<p>Applications are invited from suitable candidates in both, the Max Planck India Partner Program and the CRP Wheat Program for openings at the levels:</p>

<p>Minimum qualification Salary scale (per month)</p>

<p>Project assistant Bachelor’s / Master’s Rs. 10000 / Rs. 14000</p>

<p>Project fellow (junior data analyst) Masters + research experience Rs. 16000</p>

<p>Research fellow (senior data analyst) Masters + adequate research experience/desirable skill sets Rs. 22000</p>

<p>Research Associated PhD (&lt;1yr) / &gt; 1 yr experience Rs. 28000 / Rs. 32000</p>

<p>Condition to satisfactory performance, availability of funds and requirements of the project, the positions could be available upto a period of ~2 years (or funding of the project).</p>

<p>Essential qualification: MSc/BTech/MTech/PhD (or other suitable qualification) in discipline related to bioinformatics, computational biology, computer application (or equivalent)/ ‘Advance PostGraduate Diploma in Bioinformatics’. Proficiency in one of the programming languages or statistics (proficient in R for example) is compulsory.</p>

<p>Desirable qualification:</p>

<p>1. Programming experiences in at least one low level language such as C/C++ and one scripting language such as Perl/Python/PHP and knowledge of SQL/MySQL.</p>

<p>2. Substantial experience in the linux or other unix environments.</p>

<p>Application process: Applications should contain CV along with brief description (maximum 1 page) of research conducted (highlighting skills and experience) till now. Applications should be sent by email to Shree P. Pandey, Department of Biological Sciences, IISER-Kolkata, Mohanpur Campus, West Bengal within 2 weeks (Feb 19th 2015). E-mail: sppiiserkol@gmail.com, sppandey@iiserkol.ac.in Brief description of the group: We are an interdisciplinary group focusing on small-RNA (miRNA, siRNA) mediated regulation of signaling and defense. Project equally involve bioinformatics and systems biology (specially microarrays and next-generation sequencing (NGS) data analysis and its use), along with plant molecular biology, genetic engineering, field biology, and analytical plant chemistry for understanding response of plants to biotic stresses. For more details visit: http://www.iiserkol.ac.in/~sppandey/</p>

<p>Advertisement:</p>

<p>www.iiserkol.ac.in/images/iiserk/advertisements/advertisement_7_spp_feb_2015.pdf</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/45314/nf-coremag-v5-advancing-genome-resolved-metagenomics</guid>
	<pubDate>Sat, 12 Sep 2026 20:27:51 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/45314/nf-coremag-v5-advancing-genome-resolved-metagenomics</link>
	<title><![CDATA[nf-core/mag v5: Advancing Genome-Resolved Metagenomics]]></title>
	<description><![CDATA[<p>Metagenomics is rapidly moving beyond short-read sequencing. With the increasing adoption of long-read technologies, researchers can generate more contiguous assemblies&mdash;but converting these reads into reliable metagenome-assembled genomes (MAGs) remains computationally challenging.</p><p>nf-core/mag v5 (https://github.com/nf-core/mag) addresses this challenge by extending its reproducible Nextflow-based workflow for modern genome-resolved metagenomics.</p><p>A key addition is support for long-read-only metagenomic assembly and bin refinement, enabling long-read data to be processed through an integrated workflow. The release also introduces five additional binning tools, providing complementary strategies for recovering genomes from complex microbial communities.</p><p>The workflow has also expanded beyond conventional bacterial and archaeal MAGs, with improved classification of viruses and eukaryotes, together with enhanced genome-quality assessment.</p><p>Conceptually, the workflow brings together:</p><p>Reads &rarr; QC &rarr; Assembly &rarr; Binning &rarr; Bin refinement &rarr; Taxonomic classification &rarr; MAG quality assessment</p><p>The real strength of nf-core/mag is not any single algorithm, but the integration and reproducibility of multiple tools within a standardized workflow. This is particularly important for large-scale metagenomic studies where software versions, parameters, databases and computational environments can strongly influence results.</p><p>After seven years of development involving multiple curator teams and the wider nf-core community, v5 demonstrates how community-driven workflow development can keep metagenomic analysis aligned with rapidly evolving sequencing technologies.</p><p>The future of genome-resolved metagenomics is not simply longer reads&mdash;it is better integration of assembly, binning, refinement and quality control within reproducible computational workflows.</p><p>Read the full paper in Bioinformatics https://academic.oup.com/bioinformatics/article/42/9/btag628/8770536?</p>]]></description>
	<dc:creator>LEGE</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/21065/ra-bioinformatics-at-north-eastern-hill-university</guid>
  <pubDate>Sat, 07 Feb 2015 06:06:05 -0600</pubDate>
  <link></link>
  <title><![CDATA[RA Bioinformatics at North Eastern Hill University]]></title>
  <description><![CDATA[
<p>Bioinformatics Infrastructure Facility, Department of RDAP, NEHU vacancy of Research Associate</p>

<p>Name of the Post: Research Associate<br />No. of the Post: 01 One<br />Age Limit: Max. 35 years<br />Salary: Rs. 22000/- per month plus HRA</p>

<p>Required Job Profile:<br />Candidate must possess M.Sc. in bioinformatics or biotechnology from recognized university or institute.<br />Desired Job Profile;<br />Candidate having Ph.D. or pursuing Ph.D. in the related subject or equivalent published work in reputed peer reviewed journals or advance PG dipoma in bioinformatics course.</p>

<p>How to apply:<br />Eligible and interested candidates should need to send the bio-data and bring all related documents in original and set of attested copies of the same in the time of interview.</p>

<p>Last date: 16.02.2015<br />Refer to http://www.nehu.ac.in/Advertisements/BIFTuraAdvt_221214.pdf</p>

<p>Summary <br />Employer Address:	Dr.B.K. Mishra Coordinator BIF, RDAP Department, North Eastern Hill University, Tura Campus, Tura, Meghalaya<br />Email:	drbkm1972@yahoo.co.in;birendramishra14@gmail.com<br />URL:	http://www.nehu.ac.in/Advertisements/BIFTuraAdvt_221214.pdf<br />Phone:	03651-223107<br />Required Skills:	not mentioned / required for this post<br />Required Experience:	not mentioned / required for this job post<br />Required Education:	M.Sc. in bioinformatics or biotechnology from recognized university or institute.<br />Job Location:	Tura, Meghalaya, India</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/4183/320000-viruses-in-mammals-yet-to-sequenced-in-future</guid>
	<pubDate>Tue, 03 Sep 2013 08:35:30 -0500</pubDate>
	<link>https://bioinformaticsonline.com/news/view/4183/320000-viruses-in-mammals-yet-to-sequenced-in-future</link>
	<title><![CDATA[320000 viruses in mammals yet to sequenced in future!!!]]></title>
	<description><![CDATA[<p>With current biological technique improvements, finally it is now possible to look at millions of unknown viruses at genomic level and understand the mechanism. According to available data, close to 70 per cent of emerging viral diseases such as HIV/AIDS, West Nile, Ebola, SARS, and influenza, are zoonoses - infections of animals that cross into humans.</p><p>To address the challenges of describing and estimating virodiversity, a team of investigators from Center for Infection and Immunity (CII) and EcoHealth Alliance began in jungles of Bangladesh - home to the flying fox.</p><p>Reference:</p><p><a href="http://economictimes.indiatimes.com/news/news-by-industry/et-cetera/mammals-harbour-at-least-320000-new-viruses/articleshow/22253268.cms">http://economictimes.indiatimes.com/news/news-by-industry/et-cetera/mammals-harbour-at-least-320000-new-viruses/articleshow/22253268.cms</a></p><p><a href="http://www.bbc.co.uk/news/science-environment-23932400">http://www.bbc.co.uk/news/science-environment-23932400</a></p>]]></description>
	<dc:creator>Rahul Agarwal</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/news/view/21241/pacman</guid>
	<pubDate>Mon, 16 Feb 2015 12:15:17 -0600</pubDate>
	<link>https://bioinformaticsonline.com/news/view/21241/pacman</link>
	<title><![CDATA[Pacman]]></title>
	<description><![CDATA[<p><span>The pacman package is an R package management tool that combines the functionality of base library related functions into intuitively named functions. This package is ideally added to .Rprofile to increase workflow by reducing time recalling obscurely named functions, reducing code and integrating functionality of base functions to simultaneously perform multiple actions.<br /><br />Function names in the pacman package follow the format of p_xxx where &lsquo;xxx&rsquo; is the task the function performs. For instance the p_load function allows the user to load one or more packages as a more generic substitute for the library or require functions and if the package isn&rsquo;t available locally it will install it for you.<br /><br /></span></p><p><strong>Installation</strong></p><p><span>To download the development version of pacman:</span></p><p><span>Download the </span><a href="https://github.com/trinker/pacman/zipball/master">zip ball</a><span> or </span><a href="https://github.com/trinker/pacman/tarball/master">tar ball</a><span>, decompress and run </span><code>R CMD INSTALL</code><span> on it, or use th</span><span>e </span><strong>devtools</strong><span> package to install the development version:</span></p><pre title="">## Make sure your current packages are up to date
update.packages()
## devtools is required
devtools::install_github("trinker/pacman")
</pre><p>Note: Windows users need <a href="http://www.murdoch-sutherland.com/Rtools/">Rtools</a> and <a href="http://CRAN.R-project.org/package=devtools">devtools</a> to install this way.</p><p>More at https://github.com/trinker/pacman</p><p>&nbsp;</p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/blog/view/4574/tools-to-detect-synteny-blocks-regions-among-multiple-genomes</guid>
	<pubDate>Mon, 16 Sep 2013 17:12:02 -0500</pubDate>
	<link>https://bioinformaticsonline.com/blog/view/4574/tools-to-detect-synteny-blocks-regions-among-multiple-genomes</link>
	<title><![CDATA[Tools to detect synteny blocks regions among multiple genomes]]></title>
	<description><![CDATA[<p>The synteny block (which etymologically means &ldquo;on the same ribbon&rdquo;) is a collection of contiguous genes located on the same chromosome. These block regions have mostly been preserved by genome rearrangements, and so synteny blocks from two related species (e.g., humans and mice) will be roughly similar but flipped around on the respective genomes. Ovcharenko et. al. define it as &lsquo;any conserved sequence blocks, regardless of whether it encompasses multiple genes, an area containing single genes, or areas devoid of known genes to be considers as synteny block as long as there is conservation at the sequence level. Today, however, biologists usually refer to synteny as the conservation of blocks of order within two sets of chromosomes that are being compared with each other. This concept can also be referred to as shared synteny. The NHBLI/NCBI Glossary define synteny as &ldquo;Two genes which occur on the same chromosome are syntenic; however, syntenic genes may or may not be "linked."</p><p>Now a day, geneticists have developed a language of their own. They are pouring lots of money and energy to read the entire genomic text and understand the gods own code ATGC. It is somewhat fascinating, not only for geneticist but also for non-biologist to know that there are several conserved blocks in genome which remain conserved over hundreds of millions of years. There have been several researches on conserved blocks and non-conserved regions to understand the mechanism and importance of all these regions (http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2675965/). The finding indicates conservation and rearrangements of certain evolutionary important genes play an important role in evolution/adaptive changes (http://www.nature.com/nature/journal/v491/n7424/abs/nature11622.html https://academic.oup.com/gbe/article/8/8/2442/2198198/Novel-Insights-into-Chromosome-Evolution-in-Birds , http://science.sciencemag.org/content/346/6215/1311).</p><p>But the puzzle remains open, how to correctly define the synteny (presence of two or more genes on the same chromosome) and conserved synteny (presence of two or more genes on chromosome of each of the two species) on several genomes.</p><p><img src="http://bioinformaticsonline.com/mod/photo/syntenyImg.jpg" alt="image" width="720" height="179" style="border: 0px; border: 0px;"></p><p>Figure: Image generated with Evolution Highway (EH) tool http://eh-demo.ncsa.illinois.edu/&nbsp;</p><p>Keeping the new approach to define conserved synteny in mind there have been various algorithms developed to identify the conserved homologous synteny blocks (HSB) amongst species. Some of them which were commonly used for synteny detections are:</p><p>SyntenyTracker ( http://www-app.igb.uiuc.edu/labs/lewin/donthu/Synteny_assign/html/),</p><p>SyntenyTracker was shown to be an efficient and accurate automated tool for defining HSBs using datasets that may contain minor errors resulting from limitations in map construction methodologies.</p><p>CoGe (http://genomevolution.org/CoGe/SynFind.pl )</p><p>Satsuma (http://evomics.org/learning/genomics/satsuma/)</p><p>Cinteny (http://cinteny.cchmc.org/) ,</p><p>Cinteny server can be used for finding regions syntenic across multiple genomes and measuring the extent of genome rearrangement using reversal distance as a measure.</p><p>OrthoCluster (http://krono.act.uji.es/noticias/orthocluster-a-new-tool-for-mining-syntenic-blocks)</p><p>A new tool for mining syntenic blocks in comparative genomics</p><p>SynMap (http://genomevolution.org/wiki/index.php/SynMap),</p><p>SyMAP (http://www.symapdb.org/)</p><p>SyMAP (Synteny Mapping and Analysis Program) v4.0 is an automated system for identifying and displaying genome synteny alignments. The genomes may be represented by sequenced chromosomes (pseudomolecules), by draft sequence contigs, or by FPC physical maps (with BAC-end or marker sequence).</p><p>http://genomevolution.org/CoGe/SynMap.pl</p><p>RegionMiner (http://www.genomatix.de/online_help/help_regionminer/orthologous.html)</p><p>SyntenyMiner is being developed as an application to visualize and interrogate comparisons among multiple complete genome sequences. http://syntenyminer.sourceforge.net/</p><p>AutoGRAPH ( http://autograph.genouest.org/),</p><p>AutoGRAPH is an integrated web server for multi-species comparative genomic analysis. It is designed for constructing and visualizing synteny maps between two or three species, determination and display of macrosynteny and microsynteny relationships among species, and for highlighting evolutionary breakpoints.</p><p>SynChro(http://www.lgm.upmc.fr/CHROnicle/SynChro.html)</p><p>SynChro is a tool designed to define conserved synteny blocks. It reconstructs synteny blocks between pairwise comparison of multiple genomes. The reconstructed synteny blocks may overlap each other, be included in one another or duplicated due to micro-rearrangements.</p><p>SyntenyView ( http://www.cbs.dtu.dk/dtucourse/cookbooks/nikob/exercises/gf1_output_5.html),</p><p>Ensembl 'SyntenyView' shows conservation of large-scale gene order between species pairs. A brief summary of the calculation method appears at the bottom of this help page.&nbsp; The left of a 'SyntenyView' page displays a diagram of chromosomes with blocks of conserved synteny. The right of a page shows homology matches between individual genes within syntenic blocks.</p><p>SynBrowse ( http://www.synbrowse.org/),</p><p>SynBrowse (Synteny Browser) is a generic sequence comparison tool for visualizing genome alignments both within and between species. It is intended to help scientists study and analyze synteny, homologous genes and other conserved elements between sequences. This software is useful in studying genome duplication and evolution. It can also aid in identifying uncharacterized genes, putative regulatory elements and novel structural features of study species by comparing to a well annotated reference sequence, thus enabling genome curators to refine and edit annotations of species that have incomplete genome annotations.</p><p>Sibelia (http://arxiv.org/abs/1307.7941).</p><p>A comparative genomic tool: It assists biologists in analysing the genomic variations that correlate with pathogens, or the genomic changes that help microorganisms adapt in different environments. Sibelia will also be helpful for the evolutionary and genome rearrangement studies for multiple strains of microorganisms.</p><p>GSV (http://cas-bioinfo.cas.unt.edu/gsv/homepage.php)</p><p>Genome Synteny Viewer allows users to upload files which contain synteny regions between two or more genomes and interactively visualize the synteny between them. GSV also allows users to upload annotation files to visualize annotated regions in addition to synteny regions.</p><p>MicroSyn (http://www.lgm.upmc.fr/CHROnicle/SynChro.html)</p><p>MicroSyn software as a means of detecting microsynteny in adjacent genomic regions surrounding genes in gene families. MicroSyn searches for conserved, flanking colinear homologous gene pairs between two genomic fragments to determine the relationship between two members in a gene family.</p><p>SynOrth (http://synorth.genereg.net/)</p><p>Synorth [s n &ocirc;rth], named in combination of "synteny" and "ortholog", is designed for the study of evolutionary changes of genomic regulatory blocks (GRBs) in vertebrate genomes, and especially the changes following the whole-genome duplication in teleost fish, by tracing the ortholog genes gain and loss in ancient synteny blocks.</p><p>SyDiG (http://www.ncbi.nlm.nih.gov/pubmed/21441096)</p><p>Uncovering Synteny in Distant Genomes.</p><p>MapSynteny&nbsp; (http://www.automatizacionysistemas.com/download.html)</p><p>MapSynteny is a macro in MS Excel&reg; able to create images to show the relationship between genetic maps and large sequences (scaffolds, chromosomes, BACs, etc.). Based on tab &ndash; delimited BLAST results and some formulas, a suitable image of syntenic relationships or physical mapping can be obtained. http://www.automatizacionysistemas.com/Poster_MapSynteny.pdf</p><p>One of the best synteny tutorial for beginer @&nbsp;http://www.nature.com/scitable/topicpage/synteny-inferring-ancestral-genomes-44022</p><p>Reference:</p><p><a href="http://www.nature.com/scitable/topicpage/synteny-inferring-ancestral-genomes-44022">http://www.nature.com/scitable/topicpage/synteny-inferring-ancestral-genomes-44022</a></p><p><a href="http://www.nature.com/nature/journal/v491/n7424/full/nature11622.html">http://www.nature.com/nature/journal/v491/n7424/full/nature11622.html</a></p><p><a href="http://en.wikipedia.org/wiki/Synteny">http://en.wikipedia.org/wiki/Synteny</a></p><p><a href="http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2675965/">http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2675965/</a></p>]]></description>
	<dc:creator>Jitendra Narayan</dc:creator>
</item>

<item>
  <guid isPermaLink='true'>https://bioinformaticsonline.com/opportunity/view/21435/ra-walk-in-interview-nbfgr-lucknow</guid>
  <pubDate>Tue, 24 Feb 2015 08:23:48 -0600</pubDate>
  <link></link>
  <title><![CDATA[RA WALK-IN-INTERVIEW @ NBFGR, Lucknow]]></title>
  <description><![CDATA[
<p>F.No. 1(122)/2015-Admn. (CABin Project)<br />Research Associate/Young Professional/SRF Zoology job vacancies in National Bureau of Fish Genetic Resources (NBFGR)<br />Post Name: Research Associate (Computer Science/ Applications)                <br />Qualification: Ph.D. In Computer Science/Computer Applications or equivalent. Or Post-Graduation in Computer Science/ Computer Applications with 1st Division or 60% marks or equivalent overall grade point average with at least two years of research experience. Desirable: 1. Expertise and experience of working/ handling High Performance Computing (H PC) and genomic resource data. 2. Expertise on database management, data mining technologies/ softwares/tools. 3. Published Research papers	<br />No.of Post: 1<br />Pay Scale: Consolidated Rs.24,000/- p.m. + HRA (as admissible) for Ph.D. holders and consolidated `23,000/- + HRA (as admissible) for Master degree holder.	<br />Age:40 years</p>

<p>Young Professional II (Computer Science/Applications)	<br />Master degree in Computer Science/Computer Applications/B.Tech (Computer Science) or equivalent. <br />Desirable: 1. Knowledge of Statistical and Computational Genomics/ Proteomics/ Bioinformatics/Data mining tools. 2. Experience in handling HPC, programming languages and database management packages.	<br />A consolidated salary of Rs.25,000/- per month.	<br />21 to 45 year</p>

<p>Young Professional II (Biotechnology/ Bioinformatics)	<br />Master degree in Bioinformatics/ Biotechnology/ B. Tech(Biotech) or equivalent. Desirable: 1. Knowledge of Computational Genomics/Proteomics/Bioinformatics. 2. Expertise in NGS data analysis and knowledge of allied software and tools.	<br />A consolidated salary of Rs.25,000/- per month.	</p>

<p>Senior Research Fellow	<br />1. Bachelors degree with Zoology, Fisheries and 2. Master's degree in Fishery science/ Zoology with Fisheries/ Biotechnology/ Life Sciences with specialization in Fisheries/ Molecular Biology. 3. 1 st Division or 60% marks or equivalent overall grade point average. <br />Desirable: Work experience in Fisheries, molecular research techniques, bioinformatics and Computer skills. NET qualified <br />Note: The project involves extensive exploration tours and sampling from water bodies all over India	<br />Rs.16,000/- p.m. for 1st &amp; 2nd year and `18,000/- p.m. for 3rd and subsequent years +HRA (as per rules)	35 years for male and 40 years for female candidate</p>

<p>How to apply</p>

<p>A walk-in-interview will be held on 04th March, 2015 at 10:00 hrs at National Bureau of Fish Genetic Resources, Lucknow. Eligible and desirous candidates fulfilling all the requirements may appear for the interview with duly filled in application giving full details of academic records and experience(s) along with attested photocopy as well as original copy of the relevant documents and a passport size photograph on the attached proforma.</p>

<p>http://www.nbfgr.res.in/</p>
]]></description>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/file/view/8650/bioinformatician-duties-and-jobs</guid>
	<pubDate>Wed, 05 Mar 2014 14:32:26 -0600</pubDate>
	<link>https://bioinformaticsonline.com/file/view/8650/bioinformatician-duties-and-jobs</link>
	<title><![CDATA[Bioinformatician duties and jobs !!!]]></title>
	<description><![CDATA[<p><span><em>Needle</em> in a haystack</span> ... ohh yes this is what bioinformatician do. We handle and analyse, Terabytes and Petabytes of genomic data on daily basis.</p>]]></description>
	<dc:creator>Jitendra Narayan</dc:creator>
	<enclosure url="https://bioinformaticsonline.com/file/download/8650" length="37079" type="image/gif" />
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/pages/view/21443/a-guide-for-complete-r-beginners-getting-data-into-r</guid>
	<pubDate>Tue, 24 Feb 2015 20:15:08 -0600</pubDate>
	<link>https://bioinformaticsonline.com/pages/view/21443/a-guide-for-complete-r-beginners-getting-data-into-r</link>
	<title><![CDATA[A guide for complete R beginners :- Getting data into R]]></title>
	<description><![CDATA[<p>For a beginner this can be is the hardest part, it is also the most important to get right.</p><p>It is possible to create a vector by typing data directly into R using the combine function &lsquo;c&rsquo;</p><blockquote><p><strong>x </strong></p></blockquote><p>same as</p><blockquote><p><strong>x </strong></p></blockquote><p>creates the vector x with the numbers between 1 and 5.</p><p>You can see what is in an object at any time by typing its name;</p><blockquote><p><strong>x</strong></p></blockquote><p>will produce the output<strong> &lsquo;[1] 1 2 3 4 5&prime;</strong></p><p>Note that names need to be quoted</p><blockquote><p><strong>daysofweek </strong><strong>&larr; c(&lsquo;Monday&rsquo;, &lsquo;Tuesday&rsquo;, &lsquo;Wednesday&rsquo;, &lsquo;Thursday&rsquo;, &lsquo;Friday&rsquo;);</strong></p></blockquote><p>Usually however you want to input from a file. We have touched on the &lsquo;read.table&rsquo; function already.</p><blockquote><p><strong>mydata </strong></p></blockquote><p>Now <strong>mydata</strong> is a data frame with multiple vectors</p><p>each vector can be identified by the default syntax</p><p>#if any of these are typed it will print to screen</p><blockquote><p><strong>mydata$V1 mydata$V2 mydata$V3 </strong></p></blockquote><p>By default the function assumes certain things from the file</p><ul>
<li>The file is a plain text file (there are function to read excel files: <em>not covered here</em>)</li>
<li>columns are separated by any number of tabs or spaces</li>
<li>there is the same number of data points in each column</li>
<li>there is no header row (labels for the columns)</li>
<li>there is no column with names for the rows** [I&rsquo;ll explain].</li>
</ul><p><span style="text-decoration: underline;">If any of these are false, we need to tell that to the function</span></p><p>If it has a header column</p><blockquote><p><strong>mydata <em>header=T also works</em></strong></p></blockquote><p>Note that there is a comma between different parts of the functions arguments</p><p>If there is one less column in the header row, then R assumes that the 1<sup>st</sup> column of data after the header are the row names</p><p>Now the vectors (columns) are identified by their name</p><p>#if any of these are typed it will print to screen</p><blockquote><p><strong>mydata$A mydata$B mydata$C </strong></p></blockquote><p># Summary about the whole data frame</p><blockquote><p><strong>summary(mydata)</strong></p></blockquote><p># Summary information of column A</p><blockquote><p><strong>summary(mydata$A) </strong></p></blockquote><p>We can shortcut having to type the data frame each time by attaching it</p><blockquote><p><strong>attach(mydata)</strong></p></blockquote><p># summary of column B as &lsquo;mydata&rsquo; is attached</p><blockquote><p><strong>summary(B)</strong></p></blockquote><p><span style="text-decoration: underline;">Two other important options for </span><em><span style="text-decoration: underline;">read.table</span></em></p><p>If is is separated only by tabs and has a header</p><blockquote><p><strong>mydata </strong></p></blockquote><p>Really useful if you have spaces in the contents of some columns, so R does not mess up reading the columns . However if the columns or of an uneven length it will tell you.</p><p>If you know that the file has uneven columns</p><blockquote><p><strong>mydata </strong></p></blockquote><p>This causes R to fill empty spaces in a columns with &lsquo;NA&rsquo; .</p><p>The last two examples will still work with our file and give the same result as with only headers=T</p><p><span style="text-decoration: underline;">Graphs</span></p><p>to get an idea of what R is capable of type</p><blockquote><p><strong>demo(graphics)</strong></p></blockquote><p>steps through the examples, and the code is printed to the screen</p><p>We will work with simpler examples that have immediate use to biologists.</p><p>Remember to get more information about the options to a function type &lsquo;?function&rsquo;</p><p><span style="text-decoration: underline;">Histogram of A</span><span style="text-decoration: underline;"></span></p><blockquote><p><strong>hist(mydata$A)</strong></p></blockquote><p>If there was more data we could increase the number of vertical columns with the option, breaks=50 (or another relevant number).</p><blockquote><p><strong>boxplot(mydata)</strong></p></blockquote><p>We can get rid of the need to type the data frame each time by using the <strong>attach</strong> function</p><p># if not already done so</p><blockquote><p><strong>attach(mydata) </strong></p><p><strong>boxplot(mydata$A, mydata$B, name=c(&ldquo;Value A&rdquo;, &ldquo;Value B&rdquo;) , ylab=&ldquo;Count of Something&rdquo;)</strong></p></blockquote><p>same as</p><blockquote><p><strong>boxplot(A, B, name=c(&ldquo;Value A&rdquo;, &ldquo;Value B&rdquo;) , ylab=&ldquo;Count of Something&rdquo;)</strong></p></blockquote><p><span style="text-decoration: underline;">Scatter plot</span></p><p># if not already done so</p><blockquote><p><strong>attach(mydata) </strong></p><p><strong>plot(A,B) # or plot(mydata$A, mydata$B)</strong></p></blockquote><p><strong><span style="text-decoration: underline;">SAVING an image</span></strong></p><p>Windows users (Rgui) RIGHT click on image and select which you want.</p><p><span style="text-decoration: underline;">These instructions work for everyone.</span></p><p>You need to create a new device of the type of file you need, then send the data to that device</p><p>to save as a png file (easy to load into the likes of powerpoint, also great for web applications.</p><blockquote><p><strong>png(&lsquo;filename&rsquo;) </strong></p><p><strong>boxplot(A, B, name=c(&ldquo;Value A&rdquo;, &ldquo;Value B&rdquo;) , ylab=&ldquo;Count of Something&rdquo;)</strong></p></blockquote><p>or to save as a pdf</p><blockquote><p><strong>pdf(&lsquo;filename&rsquo;) </strong></p><p><strong>boxplot(A, B, name=c(&ldquo;Value A&rdquo;, &ldquo;Value B&rdquo;) , ylab=&ldquo;Count of Something&rdquo;)</strong></p></blockquote><p><span style="text-decoration: underline;">Note</span></p><ul>
<li>Nothing will appear on screen, the output is going to the file</li>
<li>Also it may not be saved immediately but will once the device (or R) is turned quit.</li>
</ul><p>To quit R type</p><p><strong>q() # </strong>If you save your session, next time you start R, you will have your data preloaded.</p><p>Or if you want to remain in R</p><blockquote><pre><strong>dev.off() #</strong>turns of the png (or pdf etc) device, thus forces the data to save</pre></blockquote>]]></description>
	<dc:creator>Archana Malhotra</dc:creator>
</item>
<item>
	<guid isPermaLink="true">https://bioinformaticsonline.com/bookmarks/view/26179/alignment-of-closely-related-whole-genomesscaffolds</guid>
	<pubDate>Fri, 29 Jan 2016 10:37:27 -0600</pubDate>
	<link>https://bioinformaticsonline.com/bookmarks/view/26179/alignment-of-closely-related-whole-genomesscaffolds</link>
	<title><![CDATA[Alignment of closely related whole genomes/scaffolds]]></title>
	<description><![CDATA[<p>With the relative ease and low cost of current generation sequencing technologies has led to a dramatic increase in the number of sequenced genomes for species across the tree of life. This increasing volume of data requires tools that can quickly compare multiple whole-genome sequences, millions of base pairs in length, to aid in the study of populations, pan-genomes, and genome evolution.This bookmaks have been created to report new tools for whole genome alignments.</p>
<p>Please report new whole genome alignment tools under comment sections.</p><p>Address of the bookmark: <a href="http://www.cs.utoronto.ca/~brudno/721.full.pdf" rel="nofollow">http://www.cs.utoronto.ca/~brudno/721.full.pdf</a></p>]]></description>
	<dc:creator>Rahul Nayak</dc:creator>
</item>

</channel>
</rss>