Finding contig repeat counts by mapping contigs to the reference genome

misagh

Junior Member

Join Date: Aug 2011

Posts: 2
- Share
- Tweet
#1

Finding contig repeat counts by mapping contigs to the reference genome

04-26-2013, 08:12 PM

Hi guys, I have a set of contigs of genome G (using de novo assembly by Velvet) and I also have the complete sequence of the reference genome G. I want to know the repeat count (an integer number) of each contig in the reality by mapping them to the reference genome and finding and counting exact matches.

Which tools are easier to use? At the moment I'm just interested to have a 2 column result, one column showing the contig names and the other showing an integer number which is the repeat count of that contig in the reference genome. Everything else is just a bonus. Would you please let me know which tool is better or how I can easily produce this result based on MUMmer or BLAST output?

Thanks.
Tags: alignment, contigs, mapping, reference, repeat count
bambus

Member

Join Date: Nov 2013

Posts: 20
- Share
- Tweet
#2

01-09-2014, 06:41 AM

Hi,

I hope you would have got the solution for your query,if so can you please share it here as I too got stuck with the same problem.But,the only difference is that I am working with meta-transcriptomics data for which no reference genomes are available.So after assembly I mapped the original read file against the contig file and obtained an output in sam format.

Assemblers used:- SoapDenovo-Trans,Trinity,Metavelvet(didn't worked well with my data)
Mapping tools used :- Bowtie2,segemehl

It will be very helpful if you can post the procedure or steps you have gone through to get the desired information.

Thank you in advance.
Comment

Previous template Next

Essential Discoveries and Tools in Epitranscriptomics

by seqadmin

The field of epigenetics has traditionally concentrated more on DNA and how changes like methylation and phosphorylation of histones impact gene expression and regulation. However, our increased understanding of RNA modifications and their importance in cellular processes has led to a rise in epitranscriptomics research. “Epitranscriptomics brings together the concepts of epigenetics and gene expression,” explained Adrien Leger, PhD, Principal Research Scientist...
- Channel: Articles
04-22-2024, 07:01 AM
Current Approaches to Protein Sequencing

by seqadmin

Proteins are often described as the workhorses of the cell, and identifying their sequences is key to understanding their role in biological processes and disease. Currently, the most common technique used to determine protein sequences is mass spectrometry. While still a valuable tool, mass spectrometry faces several limitations and requires a highly experienced scientist familiar with the equipment to operate it. Additionally, other proteomic methods, like affinity assays, are constrained...
- Channel: Articles
04-04-2024, 04:25 PM

Topics	Statistics	Last Post
Expanding the Horizons of Cellular Research with the Single Cell Atlas by seqadmin Started by seqadmin, Yesterday, 11:49 AM	0 responses 15 views 0 likes	Last Post by seqadmin Yesterday, 11:49 AM
Genetic Variants and Diabetes Risk in Childhood Cancer Survivors by seqadmin Started by seqadmin, 04-24-2024, 08:47 AM	0 responses 16 views 0 likes	Last Post by seqadmin 04-24-2024, 08:47 AM
Cancer Metastasis: A Deep Dive into Cellular Plasticity by seqadmin Started by seqadmin, 04-11-2024, 12:08 PM	0 responses 62 views 0 likes	Last Post by seqadmin 04-11-2024, 12:08 PM
Proteogenomic Profiles Offer New Clues in Prostate Cancer by seqadmin Started by seqadmin, 04-10-2024, 10:19 PM	0 responses 60 views 0 likes	Last Post by seqadmin 04-10-2024, 10:19 PM

Seqanswers Leaderboard Ad

Announcement

Finding contig repeat counts by mapping contigs to the reference genome

Comment

Latest Articles

ad_right_rmr

News