Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.
Biogas production with anaerobic digestion (AD) is one of the most promising solutions for both renewable energy production and resolving the environmental problem caused by the worldwide increase in organic waste. However, the complex structure of the microbiome in AD is poorly understood.
Pubmed ID: 33506264
Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.
Software statistical framework for reconstructing genomes from metagenome data. Open source software tool for accurately reconstructing single genomes from complex microbial communities.
View all literature mentionsSoftware tool for protein coding gene prediction for prokaryotic genomes.
View all literature mentionsOriginal SAMTOOLS package has been split into three separate repositories including Samtools, BCFtools and HTSlib. Samtools for manipulating next generation sequencing data used for reading, writing, editing, indexing,viewing nucleotide alignments in SAM,BAM,CRAM format. BCFtools used for reading, writing BCF2,VCF, gVCF files and calling, filtering, summarising SNP and short indel sequence variants. HTSlib used for reading, writing high throughput sequencing data.
View all literature mentionsA software pipeline for characterizing the taxonomic composition and genetic diversity of short-read metagenomes. The software was originally designed for the analysis of environmental metagenomes obtained by the ultra-fast 454 pyrosequencing system.
View all literature mentionsTool for searching sequence databases for homologs of protein sequences, and for making protein sequence alignments. It implements methods using probabilistic models called profile hidden Markov models (profile HMMs). Compared to BLAST, FASTA, and other sequence alignment and database search tools based on older scoring methodology, HMMER aims to be significantly more accurate and more able to detect remote homologs because of the strength of its underlying mathematical models. In the past, this strength came at significant computational expense, but in the new HMMER3 project, HMMER is now essentially as fast as BLAST.
View all literature mentionsTHIS RESOURCE IS NO LONGER IN SERVICE. Documented on February 28,2023. Software program for clustering biological sequences with many applications in various fields such as making non-redundant databases, finding duplicates, identifying protein families, filtering sequence errors and improving sequence assembly etc. It is very fast and can handle extremely large databases. CD-HIT helps to significantly reduce the computational and manual efforts in many sequence analysis tasks and aids in understanding the data structure and correct the bias within a dataset. The CD-HIT package has CD-HIT, CD-HIT-2D, CD-HIT-EST, CD-HIT-EST-2D, CD-HIT-454, CD-HIT-PARA, PSI-CD-HIT, CD-HIT-OTU and over a dozen scripts. * CD-HIT (CD-HIT-EST) clusters similar proteins (DNAs) into clusters that meet a user-defined similarity threshold. * CD-HIT-2D (CD-HIT-EST-2D) compares 2 datasets and identifies the sequences in db2 that are similar to db1 above a threshold. * CD-HIT-454 identifies natural and artificial duplicates from pyrosequencing reads. * CD-HIT-OTU cluster rRNA tags into OTUs The usage of other programs and scripts can be found in CD-HIT user''s guide. CD-HIT was originally developed by Dr. Weizhong Li at Dr. Adam Godzik''s Lab at the Burnham Institute (now Sanford-Burnham Medical Research Institute).
View all literature mentionsSoftware designed to quickly find sequences of 95% and greater similarity of length 25 bases or more.
View all literature mentionsSoftware tool to assess the quality of microbial genomes recovered from isolates, single cells, and metagenomes by using a broader set of marker genes specific to the position of a genome within a reference genome tree and information about the collocation of these genes.
View all literature mentionsOpen source software tool for assigning objective taxonomic classifications to bacterial and archaeal genomes based on Genome Database Taxonomy. Designed to work with recent advances that allow metagenome assembled genomes to be obtained directly from environmental samples. Can also be applied to isolate and single cell genomes.
View all literature mentionsSoftware tool for genome and metagenome distance estimation using MinHash. Reduces large sequences and sequence sets to small, representative sketches, from which global mutation distances can be rapidly estimated.
View all literature mentionsSoftware tool as Next Generation Sequencing assembler. Optimized for metagenomes, but also works well on generic single genome assembly (small or mammalian size) and single cell assembly. Can assemble genome sequences from metagenomic datasets of hundreds of Giga base-pairs in time and memory efficient manner on single server.
View all literature mentionsSoftware for aligning sequencing reads against large reference genome. Consists of three algorithms: BWA-backtrack, BWA-SW and BWA-MEM. First for sequence reads up to 100bp, and other two for longer sequences ranged from 70bp to 1Mbp.
View all literature mentionsSoftware tool as a short read aligner for DNA and RNA seq data. Used for large genomes with millions of scaffolds. Can align reads from Illumina, PacBio, 454, Sanger, Ion Torrent, Nanopore. Fast and accurate, particularly with highly mutated genomes or reads with long indels, even whole gene deletions over 100kbp long. It has no upper limit to genome size or number of contigs. Written in Java, can run on any platform.
View all literature mentionsSoftware that performs sequence alignment for protein and translated DNA searches and functions. Used for high performance analysis of big sequence data, protein-protein search, and DNA-protein search.
View all literature mentions