[LCRC Accounts] Yearly Allocation Request from MG-RAST
Hello, A yearly allocation for the LCRC cluster has been requested with the following updated information: Submitter/PI: Jared Wilkening Project Name: MG-RAST Division: IGSB Project title: Assembly of metagenomes using MG-RAST Associated funding: DOE, NIH, Sloan foundation Other Systems: local cluster (shared ~200 cores), Magellan, PADS Science: Metagenomics is the study of genomic DNA isolated from environmental samples. Due to some recent advances in sequencing technology and the readily available Metagenomics RAST server (MG-RAST) thousands of groups have started to use metagenomic sequencing strategies as part of the research. Metagenomic sequencing can be used to reveal insights into the microbial community structure and function present in a particular sample. With thousands of data sets available and numbers growing fast, comparison between different environments or samples from similar environments is one key feature of metagenomic studies. Metagenomics is applied to environments with relevance to Climate Change and environmental remediation, helping researchers better understand the roles played by the microbial players in those ecosystems. Many researchers have started using metagenomics and MG-RAST as the leading metagenomic analysis and comparison platform to shed light on the contribution of microbes to their research. MG-RAST has changed dramatically over the course of the last 12 month, becoming more efficient and allowing more users to benefit from the analysis offered. The number of data submitting users has reached 2000 and the number of data sets now exceeds 9000. Project description: MG-RAST is the premiere metagenomic portal in use today, having analyzed over 35000 data sets (3TBp of data), representing the analysis of millions of dollars of sequencing. MG-RAST is a major component in the lab's BESI initiative goals. This allocation would focus our the next area of scaling interest: genomic / metagenomic assemble. Our previous allocation of Fusion provided us with invaluable data on scaling sequence similarity search analysis that we have since off-loaded to Magellan and other cloud based resources. While many of the challenges in scaling genomic / metagenomic assemble are similar to those of sequence similarity search, Fusion is an ideal platform for this work because of fast interconnect, high speed disk and the available high memory machines. Over the course of the year we anticipate approximately 1,000 - 2,000 jobs of 2-8 nodes in 60-300 minute time blocks. Project URL: http://metagenomics.anl.gov Current FY Hours Used: undetermined amount New FY Requested allocation: 475000 Q1: 100000 Q2: 125000 Q3: 125000 Q4: 125000 Justification: Thank You, The LCRC Accounts System
participants (1)
-
accounts@lcrc.anl.gov