Searching the Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

A microbial gene catalog of anaerobic digestion from full-scale biogas plants.

Shichun Ma | Fan Jiang | Yan Huang | Yan Zhang | Sen Wang | Hui Fan | Bo Liu | Qiang Li | Lijuan Yin | Hengchao Wang | Hangwei Liu | Yuwei Ren | Shuqu Li | Lei Cheng | Wei Fan | Yu Deng
GigaScience | 2021

Biogas production with anaerobic digestion (AD) is one of the most promising solutions for both renewable energy production and resolving the environmental problem caused by the worldwide increase in organic waste. However, the complex structure of the microbiome in AD is poorly understood.

Pubmed ID: 33506264

Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.

This is a list of tools and resources that we have found mentioned in this publication.


MetaBAT (software resource)

RRID:SCR_019134

Software statistical framework for reconstructing genomes from metagenome data. Open source software tool for accurately reconstructing single genomes from complex microbial communities.

View all literature mentions

Prodigal (software resource)

RRID:SCR_011936

Software tool for protein coding gene prediction for prokaryotic genomes.

View all literature mentions

SAMTOOLS (software resource)

RRID:SCR_002105

Original SAMTOOLS package has been split into three separate repositories including Samtools, BCFtools and HTSlib. Samtools for manipulating next generation sequencing data used for reading, writing, editing, indexing,viewing nucleotide alignments in SAM,BAM,CRAM format. BCFtools used for reading, writing BCF2,VCF, gVCF files and calling, filtering, summarising SNP and short indel sequence variants. HTSlib used for reading, writing high throughput sequencing data.

View all literature mentions

CARMA (software resource)

RRID:SCR_004999

A software pipeline for characterizing the taxonomic composition and genetic diversity of short-read metagenomes. The software was originally designed for the analysis of environmental metagenomes obtained by the ultra-fast 454 pyrosequencing system.

View all literature mentions

Hmmer (software resource)

RRID:SCR_005305

Tool for searching sequence databases for homologs of protein sequences, and for making protein sequence alignments. It implements methods using probabilistic models called profile hidden Markov models (profile HMMs). Compared to BLAST, FASTA, and other sequence alignment and database search tools based on older scoring methodology, HMMER aims to be significantly more accurate and more able to detect remote homologs because of the strength of its underlying mathematical models. In the past, this strength came at significant computational expense, but in the new HMMER3 project, HMMER is now essentially as fast as BLAST.

View all literature mentions

CD-HIT (data processing software)

RRID:SCR_007105

THIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute).

View all literature mentions

BLAT (software resource)

RRID:SCR_011919

Software designed to quickly find sequences of 95% and greater similarity of length 25 bases or more.

View all literature mentions

CheckM (software resource)

RRID:SCR_016646

Software tool to assess the quality of microbial genomes recovered from isolates, single cells, and metagenomes by using a broader set of marker genes specific to the position of a genome within a reference genome tree and information about the collocation of these genes.

View all literature mentions

GTDB-Tk (software resource)

RRID:SCR_019136

Open source software tool for assigning objective taxonomic classifications to bacterial and archaeal genomes based on Genome Database Taxonomy. Designed to work with recent advances that allow metagenome assembled genomes to be obtained directly from environmental samples. Can also be applied to isolate and single cell genomes.

View all literature mentions

Mash (software application)

RRID:SCR_019135

Software tool for genome and metagenome distance estimation using MinHash. Reduces large sequences and sequence sets to small, representative sketches, from which global mutation distances can be rapidly estimated.

View all literature mentions

MEGAHIT (software resource)

RRID:SCR_018551

Software tool as Next Generation Sequencing assembler. Optimized for metagenomes, but also works well on generic single genome assembly (small or mammalian size) and single cell assembly. Can assemble genome sequences from metagenomic datasets of hundreds of Giga base-pairs in time and memory efficient manner on single server.

View all literature mentions

BWA (software resource)

RRID:SCR_010910

Software for aligning sequencing reads against large reference genome. Consists of three algorithms: BWA-backtrack, BWA-SW and BWA-MEM. First for sequence reads up to 100bp, and other two for longer sequences ranged from 70bp to 1Mbp.

View all literature mentions

BBmap (software resource)

RRID:SCR_016965

Software tool as a short read aligner for DNA and RNA seq data. Used for large genomes with millions of scaffolds. Can align reads from Illumina, PacBio, 454, Sanger, Ion Torrent, Nanopore. Fast and accurate, particularly with highly mutated genomes or reads with long indels, even whole gene deletions over 100kbp long. It has no upper limit to genome size or number of contigs. Written in Java, can run on any platform.

View all literature mentions

DIAMOND (software resource)

RRID:SCR_016071

Software that performs sequence alignment for protein and translated DNA searches and functions. Used for high performance analysis of big sequence data, protein-protein search, and DNA-protein search.

View all literature mentions