Searching the Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

MEGAN-LR: new algorithms allow accurate binning and easy interactive exploration of metagenomic long reads and contigs.

Daniel H Huson | Benjamin Albrecht | Caner Bağcı | Irina Bessarab | Anna Górska | Dino Jolic | Rohan B H Williams
Biology direct | 2018

There are numerous computational tools for taxonomic or functional analysis of microbiome samples, optimized to run on hundreds of millions of short, high quality sequencing reads. Programs such as MEGAN allow the user to interactively navigate these large datasets. Long read sequencing technologies continue to improve and produce increasing numbers of longer reads (of varying lengths in the range of 10k-1M bps, say), but of low quality. There is an increasing interest in using long reads in microbiome sequencing, and there is a need to adapt short read tools to long read datasets.

Pubmed ID: 29678199

Research resources used in this publication

None found

Antibodies used in this publication

None found

Associated grants

  • Agency: National Science Foundation, International
    Id: NSF PHY-1748958

Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.

This is a list of tools and resources that we have found mentioned in this publication.


BLASTX (tool)

RRID:SCR_001653

Web application to search protein databases using a translated nucleotide query. Translated BLAST services are useful when trying to find homologous proteins to a nucleotide coding region. Blastx compares translational products of the nucleotide query sequence to a protein database. Because blastx translates the query sequence in all six reading frames and provides combined significance statistics for hits to different frames, it is particularly useful when the reading frame of the query sequence is unknown or it contains errors that may lead to frame shifts or other coding errors. Thus blastx is often the first analysis performed with a newly determined nucleotide sequence and is used extensively in analyzing EST sequences. This search is more sensitive than nucleotide blast since the comparison is performed at the protein level.

View all literature mentions

eggNOG (tool)

RRID:SCR_002456

A database of orthologous groups of genes. The orthologous groups are annotated with functional description lines (derived by identifying a common denominator for the genes based on their various annotations), with functional categories (i.e derived from the original COG/KOG categories). eggNOG's database currently counts 1.7 million orthologous groups in 3686 species, covering over 7.7 million proteins (built from 9.6 million proteins). (Jan 30, 2014)

View all literature mentions

LAST (tool)

RRID:SCR_006119

THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software tool for aligning sequences, similar to BLAST 2 sequences that colour-codes the alignments by reliability. Another useful feature of LAST is that it can compare huge (vertebrate-genome-sized) datasets. Unfortunately, this only applies to the downloadable version of LAST, not the web service. The web service can just about handle bacterial genomes, but it will take a few minutes and the output will be large. LAST can: * Handle big sequence data, e.g: ** Compare two vertebrate genomes ** Align billions of DNA reads to a genome * Indicate the reliability of each aligned column. * Use sequence quality data properly. * Compare DNA to proteins, with frameshifts. * Compare PSSMs to sequences * Calculate the likelihood of chance similarities between random sequences. LAST cannot (yet): * Do spliced alignment.

View all literature mentions

InterPro (tool)

RRID:SCR_006695

Service providing functional analysis of proteins by classifying them into families and predicting domains and important sites. They combine protein signatures from a number of member databases into a single searchable resource, capitalizing on their individual strengths to produce a powerful integrated database and diagnostic tool. This integrated database of predictive protein signatures is used for the classification and automatic annotation of proteins and genomes. InterPro classifies sequences at superfamily, family and subfamily levels, predicting the occurrence of functional domains, repeats and important sites. InterPro adds in-depth annotation, including GO terms, to the protein signatures. You can access the data programmatically, via Web Services. The member databases use a number of approaches: # ProDom: provider of sequence-clusters built from UniProtKB using PSI-BLAST. # PROSITE patterns: provider of simple regular expressions. # PROSITE and HAMAP profiles: provide sequence matrices. # PRINTS provider of fingerprints, which are groups of aligned, un-weighted Position Specific Sequence Matrices (PSSMs). # PANTHER, PIRSF, Pfam, SMART, TIGRFAMs, Gene3D and SUPERFAMILY: are providers of hidden Markov models (HMMs). Your contributions are welcome. You are encouraged to use the ''''Add your annotation'''' button on InterPro entry pages to suggest updated or improved annotation for individual InterPro entries.

View all literature mentions

DIAMOND (tool)

RRID:SCR_009457

Software to: view dicom files and assemble them into 3D volumes. View and convert between Analyze, Nifti, and Interfile. Classify and organize dicoms and 3D volumes using metadata. Search and report on a collection of scans.

View all literature mentions

MEGAN (tool)

RRID:SCR_011942

Software for analyzing metagenomes.

View all literature mentions

New England Biolabs (tool)

RRID:SCR_013517

An Antibody supplier

View all literature mentions