This $565,783 National Science Foundation project grant under the Computer and Information Science and Engineering program will support the development of new data structures and software for computational biology and big data storage systems at the University of Maryland, College Park. The grantee will create compact summary data structures that can represent huge datasets and fit within computer memory, enabling applications to run much more quickly and scale to larger data sets. The data structures will allow computational biology and big data applications to maintain summaries of genetic and other biological information for thousands to millions of individuals to detect genetic variations correlated with disease or other traits. All software and documentation will be released as open source. $100,000 of the funding will be provided to Stony Brook University as a subaward to support related work. The new data structures and software have the potential to accelerate genomic and biomedical research through more efficient analysis of massive sequencing datasets.