{"response":{"award":[{"abstractText":"Computation is critical to our nation's progress in science and engineering. Whether through simulation of phenomena where experiments are costly or impossible, large scale data analysis to sift the enormous quantities of digital data scientific instruments can produce, or machine learning to find patterns and suggest hypothesis from this vast array of data, computation is the universal tool upon which nearly every field of science and engineering relies upon to hasten their advance. This project will deploy a powerful new system, called \"Frontier\", that builds upon a design philosophy and operations approach proven by the success of the Texas Advanced Computing Center (TACC) in delivering leading instruments for computational science. Frontier provides a system of unprecedented scale in the NSF cyberinfrastructure that will yield productive science on day one, while also preparing the research community for the shift to much more capable systems in the future.  Frontier is a hybrid system of conventional Central Processing Units (CPU) and Graphics Processing Units (GPU), with performance capabilities that significantly exceeds prior leadership-class computing investments made by NSF.  Importantly, the design of Frontier will support the seamless transition of current NSF leadership-class computing applications to the new system, as well as enable new large-scale data-intensive and machine learning workloads that are expected in the future.  Following deployment, the project will operate the system in partnership with ten academic partners.  In addition, the project will begin planning activities in collaboration with leading computational scientists and technologists from around the country, and will leverage strategic public-private partnerships to design a leadership-class computing facility with at least ten times more performance capabilities for Science and Engineering research, ensuring the economic competitiveness and prosperity for our nation at large.\r\n\r\nTACC, in partnerships with Dell EMC and Intel, will deploy Frontier, a hybrid system offering 39 PF (double precision) of Intel Xeon processors, complemented by 11 PF (single precision) of GPU cards for machine learning applications. In addition to 3x the per node memory of NSF's prior leadership-class computing system primary compute nodes, Frontier will have 2x the storage bandwidth in a storage hierarchy that includes 55PB of usable disk-based storage and 3PB of 'all flash' storage, to enable next generation data-intensive applications and support for the data science community.  Frontier will be deployed in TACC's state-of-the-art datacenter which is configured to supply 30% of the system's power needs from renewable energy.  Frontier will include support for science and engineering in virtually all disciplines through its software environment support for application containers, as well as through its partnership with ten academic institutions providing deep computational science expertise in support of users on the system. The project planning effort for a Phase 2 system with at least 10x performance improvement will incorporate a community-driven process that will include leading computational scientists and technologists from around the country and leverage strategic public-private partnerships.  This process will ensure the design of a future NSF leadership-class computing facility that incorporates the most productive near-term technologies, and anticipates the most likely future technological capabilities for all of science and engineering requiring leadership-class computational and data-analytics capabilities.  Furthermore, the project is expected to develop new expertise and techniques for leadership-class computing and data-driven applications that will benefit future users worldwide through publications, training, and consulting.  The project will leverage the team's unique approach to education, outreach, and training activities to encourage, educate, and develop the next generation of leadership-class computational science researchers. The team includes leaders in campus bridging, minority-serving institute (MSI) outreach, and data technologies who will oversee efforts to use Frontier to increase the diversity of groups using leadership-class computing for traditional and data-driven applications.\r\n\r\nThis award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.","activeAwd":"false","agency":"NSF","awardAgencyCode":"4900","awardee":"UNIVERSITY OF TEXAS AT AUSTIN","awardeeAddress":"110 INNER CAMPUS DR","awardeeCity":"AUSTIN","awardeeCountryCode":"US","awardeeDistrict":"25","awardeeDistrictCode":"TX25","awardeeName":"University of Texas at Austin","awardeePhone":"5124716424","awardeeStateCode":"TX","awardeeZipCode":"787121139","cfdaNumber":"47.070","coPDPI":["Dhabaleswar K Panda panda.2@osu.edu","Omar N Ghattas omar@ices.utexas.edu","Tommy K Minyard minyard@tacc.utexas.edu","John E West john@tacc.utexas.edu"],"date":"08/28/2018","dirAbbr":"CSE","divAbbr":"OAC","estimatedTotalAmt":"60000000","expDate":"02/28/2025","fundAgencyCode":"4900","fundProgramName":"CYBERINFRASTRUCTURE, Leadership-Class Computing","fundsObligated":["FY 2018 = $60,000,000.00","FY 2019 = $2,999,135.00","FY 2020 = $4,000,000.00","FY 2023 = $11,999,999.00"],"fundsObligatedAmt":"78999136","histAwd":"false","id":"1818253","initAmendmentDate":"08/28/2018","jrnl":[{"artTitl":"Spark Meets MPI: Towards High-Performance Communication Framework for Spark using MPI","auth":"Al Attar, K. and Shafi, A. and Abduljabbar, M. and H. Subramoni, H. and Panda, DK","jrnlTitl":"2022 IEEE International Conference on Cluster Computing","jrnlYr":"2022","parPblcId":"10355057"},{"artTitl":"Towards Java-based HPC using the MVAPICH2 Library: Early Experiences","auth":"Al-Attar, Kinan and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.","dgtlObjId":"https://doi.org/10.1109/IPDPSW55747.2022.00091","jrnlTitl":"2022 IEEE International Parallel and Distributed Processing Symposium Workshops (","jrnlYr":"2022","parPblcId":"10355059"},{"artTitl":"OMB-Py: Python Micro-Benchmarks for Evaluating Performance of MPI Libraries on HPC Systems","auth":"Alnaasan, Nawras and Jain, Arpan and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K","dgtlObjId":"https://doi.org/10.1109/IPDPSW55747.2022.00143","jrnlTitl":"23rd Parallel and Distributed Scientific and Engineering Computing Workshop (PDSEC) at IPDPS22","jrnlYr":"2022","parPblcId":"10355045"},{"artTitl":"Highly Efficient Alltoall and Alltoallv Communication Algorithms for GPU Systems","auth":"Chen, Chen-Chun and Khorassani, Kawthar Shafie and Anthony, Quentin G. and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.","dgtlObjId":"https://doi.org/10.1109/IPDPSW55747.2022.00014","jrnlTitl":"Heterogeneity in Computing Workshop","jrnlYr":"2022","parPblcId":"10355071"},{"artTitl":"Hy-Fi: Hybrid Five-Dimensional Parallel DNN Training on High-Performance GPU Clusters","auth":"Jain, A and Shafi, A. and Anthony, Q. and Kousha, P. and Subramoni, H. and Panda, DK.","dgtlObjId":"https://doi.org/10.1007/978-3-031-07312-0_6","jrnlTitl":"Proceedings International Conference on High Performance Computing","jrnlYr":"2022","parPblcId":"10355061"},{"artTitl":"Hey CAI - Conversational AI Enabled User Interface for HPC Tools","auth":"Kousha, P. and Jain, A. and Kolli, A. and Prasanna, S. and Miriyala, S. and Subramoni, H. and Shafi, A. and Panda, DK.","dgtlObjId":"https://doi.org/10.1007/978-3-031-07312-0_5","jrnlTitl":"Proceedings International Conference on High Performance Computing","jrnlYr":"2022","parPblcId":"10355064"},{"artTitl":"Towards Architecture-aware Hierarchical Communication Trees on Modern HPC Systems","auth":"Ramesh, Bharath and Hashmi, Jahanzeb Maqbool and Xu, Shulei and Shafi, Aamir and Ghazimirsaeed, Mahdieh and Bayatpour, Mohammadreza and Subramoni, Hari and Panda, Dhabaleswar K.","dgtlObjId":"https://doi.org/10.1109/HIPC53243.2021.00041","jrnlTitl":"International Conference on High Performance Computing, Data, and Analytics","jrnlYr":"2021","parPblcId":"10355065"},{"artTitl":"Network Assisted Non-Contiguous Transfers for GPU-Aware MPI Libraries","auth":"Suresh, K. and Khorassani, K. and Chen, C. and Ramesh, B. and Abduljabbar, M. and Shafi, A. and Panda, DK.","jrnlTitl":"Hot Interconnects","jrnlYr":"2022","parPblcId":"10355066"},{"artTitl":"Layout-aware Hardware-assisted Designs for Derived Data Types in MPI","auth":"Suresh, Kaushik Kandadi and Ramesh, Bharath and Chen, Chen Chun and Ghazimirsaeed, Seyedeh Mahdieh and Bayatpour, Mohammadreza and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.","dgtlObjId":"https://doi.org/10.1109/HiPC53243.2021.00044","jrnlTitl":"2021 IEEE 28th International Conference on High Performance Computing, Data, and Analytics (HiPC)","jrnlYr":"2021","parPblcId":"10334470"},{"artTitl":"Designing Hierarchical Multi-HCA Aware Allgather in MPI","auth":"Tran, A. and Michalowicz, B. and Ramesh, B. and Subramoni, H. and Shafi, A. and Panda, DK.","dgtlObjId":"https://doi.org/10.1145/3547276.3548524","jrnlTitl":"International Workshop on Parallel Programming Models and Systems Software for High-End Computing","jrnlYr":"2022","parPblcId":"10355067"},{"artTitl":"Arm meets Cloud: A Case Study of MPI Library Performance on AWS Arm-based HPC Cloud with Elastic Fabric Adapter","auth":"Xu, Shulei and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.","dgtlObjId":"https://doi.org/10.1109/IPDPSW55747.2022.00083","jrnlTitl":"IEEE International Parallel and Distributed Processing Symposium Workshops","jrnlYr":"2022","parPblcId":"10355068"},{"artTitl":"Accelerating MPI All-to-All Communication with Online Compression on Modern GPU Clusters","auth":"Zhou, Q. and Kousha, P. and Anthony, Q. and Khorassani, K. and Shafi, A. and Subramoni, H. and Panda, DK.","dgtlObjId":"https://doi.org/10.1007/978-3-031-07312-0_1","jrnlTitl":"ISC HIGH PERFORMANCE","jrnlYr":"2022","parPblcId":"10355069"}],"latestAmendmentDate":"02/29/2024","managingPec":"778100","orgCodeDir":"05000000","orgCodeDiv":"05090000","orgLongName":"Directorate for Computer and Information Science and Engineering","orgLongName2":"Office of Advanced Cyberinfrastructure (OAC)","orgUrl":"https://www.nsf.gov/div/index.jsp?div=OAC","parentUeiNumber":"X5NKD2NFF2V3","pdPIName":"Daniel Stanzione","perfAddress":"3925 West Braker Lane, Suite 156","perfCity":"Austin","perfCountryCode":"US","perfDistrict":"37","perfDistrictCode":"TX37","perfLocation":"University of Texas at Austin","perfStateCode":"TX","perfZipCode":"787595316","pi":["Daniel Stanzione dan@tacc.utexas.edu"],"piEmail":"dan@tacc.utexas.edu","piFirstName":"Daniel","piId":"269709413","piLastName":"Stanzione","poEmail":"edwalker@nsf.gov","poName":"Edward Walker","poPhone":"7032924863","primaryProgram":["01002324DB NSF RESEARCH & RELATED ACTIVIT","01001819DB NSF RESEARCH & RELATED ACTIVIT","01001920DB NSF RESEARCH & RELATED ACTIVIT","01002021DB NSF RESEARCH & RELATED ACTIVIT"],"progEleCode":"723100, 778100","program":"NSCI: National Strategic Computing Initi, COVID-19 Impacts on Existing Activities, PETASCALE - TRACK 1","progRefCode":"026Z, 097Z, 7781","projectOutComesReport":"<div class=\"porColContainerWBG\">\n<div class=\"porContentCol\"><p>This award led to the 2019 deployment of the Frontera supercomputer at the Texas Advanced Computing Center at the University of Texas at Austin.&nbsp; Built with components from Dell, Intel, and Mellanox, Frontera debuted as the #5 fastest supercomputer in the world, and for the last six years has been the fastest supercomputer at any University in the United States.&nbsp; &nbsp;</p>\r\n<p>With operations still ongoing, through six years of life Frontera has delivered more than 7.2M simulations to more than 1,600 unclassified research projects, and supported nearly a billion dollars worth of open science research supported by the National Science Foundation.&nbsp; &nbsp;Frontera has delivered more than 370M node hours (21 Billion CPU core hours) to researcher in fields such as Materials Science, Electronics, Astronomy, Weather and Climate, Chemistry, Physics, Engineering, and many more.&nbsp; &nbsp;Frontera was instrumental in research during the COVID pandemic running drug discovery pipelines and pandemic modeling, was involved in the imaging of the Black Hole at the center of our galaxy, processing some of the first datasets from the James Webb Space Telescope, countless hurricane impact forecasts, and many other projects.&nbsp; &nbsp;Frontera was used from everyone from Nobel Prize winners to high school students learning AI and coding skills.&nbsp; &nbsp;</p>\r\n<p>Frontera initially consisted of 8,008 56-core Intel Xeon \"Cascade Lake\" nodes from Dell, 360 NVIDIA RTX 5000 GPUs, and roughly 100 V100 GPUs in IBM Power9 nodes, along with a 200Gbps fabric from Mellanox, and more than 50PB of fast storage from DataDirect Networks.&nbsp; &nbsp;Frontera was upgraded several times, with the addition of more than 300 additional compute nodes during the pandemic, replacement of the V100s with NVIDIA A100 GPUs in Dell servers, and most recently the project was used to stand up the bridge system to the next Leadership System, Horizon, coming in 2026.&nbsp;&nbsp;</p>\r\n<p>Throughout its production life, Frontera maintained uptime in excess of 99%, with more than 95% of compute nodes in use at all times.&nbsp; &nbsp;It has been a remarkably productive system that has helped make thousands of advances in engineering and science.&nbsp;</p><br>\n<p>\n Last Modified: 07/01/2025<br>\nModified by: Daniel&nbsp;Stanzione</p></div>\n<div class=\"porSideCol\"\n><div class=\"each-gallery\">\n<div class=\"galContent\" id=\"gallery0\">\n<div class=\"photoCount\" id=\"photoCount0\">\n\t\t\t\t\t\t\t\t\tImage\n\t\t\t\t\t\t\t\t</div>\n<div class=\"galControls onePhoto\" id=\"controls0\"></div>\n<div class=\"galSlideshow\" id=\"slideshow0\"></div>\n<div class=\"galEmbox\" id=\"embox\">\n<div class=\"image-title\"></div>\n</div>\n</div>\n<div class=\"galNavigation onePhoto\" id=\"navigation0\">\n<ul class=\"thumbs\" id=\"thumbs0\">\n<li>\n<a href=\"/por/images/Reports/POR/2025/1818253/1818253_10576865_1751382874010_Frontera_Img--rgov-214x142.png\" original=\"/por/images/Reports/POR/2025/1818253/1818253_10576865_1751382874010_Frontera_Img--rgov-800width.png\" title=\"The Frontera Supercomputer\"><img src=\"/por/images/Reports/POR/2025/1818253/1818253_10576865_1751382874010_Frontera_Img--rgov-66x44.png\" alt=\"The Frontera Supercomputer\"></a>\n<div class=\"imageCaptionContainer\">\n<div class=\"imageCaption\">Two of the rows of the Frontera supercomputer in Austin, Texas.</div>\n<div class=\"imageCredit\">UT-Austin</div>\n<div class=\"imagePermisssions\">Public Domain</div>\n<div class=\"imageSubmitted\">Daniel&nbsp;Stanzione\n<div class=\"imageTitle\">The Frontera Supercomputer</div>\n</div>\n</li></ul>\n</div>\n</div></div>\n</div>\n","publicAccessMandate":"1","publicationResearch":["2022 IEEE International Conference on Cluster Computing~2022~Al Attar, K. and Shafi, A. and Abduljabbar, M. and H. Subramoni, H. and Panda, DK~Spark Meets MPI: Towards High-Performance Communication Framework for Spark using MPI~10355057~10355057~OSTI~2022-09-09 21:03:24.406","2022 IEEE International Parallel and Distributed Processing Symposium Workshops (~2022~Al-Attar, Kinan and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.~https://doi.org/10.1109/IPDPSW55747.2022.00091~Towards Java-based HPC using the MVAPICH2 Library: Early Experiences~510 to 519~10355059~10355059~OSTI~2022-09-09 21:03:20.843","23rd Parallel and Distributed Scientific and Engineering Computing Workshop (PDSEC) at IPDPS22~2022~Alnaasan, Nawras and Jain, Arpan and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K~https://doi.org/10.1109/IPDPSW55747.2022.00143~OMB-Py: Python Micro-Benchmarks for Evaluating Performance of MPI Libraries on HPC Systems~870 to 879~10355045~10355045~OSTI~2022-09-09 21:03:20.696","Heterogeneity in Computing Workshop~2022~Chen, Chen-Chun and Khorassani, Kawthar Shafie and Anthony, Quentin G. and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.~https://doi.org/10.1109/IPDPSW55747.2022.00014~Highly Efficient Alltoall and Alltoallv Communication Algorithms for GPU Systems~24 to 33~10355071~10355071~OSTI~2022-09-09 21:03:20.996","Proceedings International Conference on High Performance Computing~2022~Jain, A and Shafi, A. and Anthony, Q. and Kousha, P. and Subramoni, H. and Panda, DK.~https://doi.org/10.1007/978-3-031-07312-0_6~Hy-Fi: Hybrid Five-Dimensional Parallel DNN Training on High-Performance GPU Clusters~10355061~10355061~OSTI~2022-09-09 21:03:19.87","Proceedings International Conference on High Performance Computing~2022~Kousha, P. and Jain, A. and Kolli, A. and Prasanna, S. and Miriyala, S. and Subramoni, H. and Shafi, A. and Panda, DK.~https://doi.org/10.1007/978-3-031-07312-0_5~Hey CAI - Conversational AI Enabled User Interface for HPC Tools~10355064~10355064~OSTI~2022-09-09 21:03:19.72","International Conference on High Performance Computing, Data, and Analytics~2021~Ramesh, Bharath and Hashmi, Jahanzeb Maqbool and Xu, Shulei and Shafi, Aamir and Ghazimirsaeed, Mahdieh and Bayatpour, Mohammadreza and Subramoni, Hari and Panda, Dhabaleswar K.~https://doi.org/10.1109/HIPC53243.2021.00041~Towards Architecture-aware Hierarchical Communication Trees on Modern HPC Systems~272 to 281~10355065~10355065~OSTI~2022-09-09 21:03:25.106","Hot Interconnects~2022~Suresh, K. and Khorassani, K. and Chen, C. and Ramesh, B. and Abduljabbar, M. and Shafi, A. and Panda, DK.~Network Assisted Non-Contiguous Transfers for GPU-Aware MPI Libraries~10355066~10355066~OSTI~2022-09-09 21:03:23.983","2021 IEEE 28th International Conference on High Performance Computing, Data, and Analytics (HiPC)~2021~Suresh, Kaushik Kandadi and Ramesh, Bharath and Chen, Chen Chun and Ghazimirsaeed, Seyedeh Mahdieh and Bayatpour, Mohammadreza and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.~https://doi.org/10.1109/HiPC53243.2021.00044~Layout-aware Hardware-assisted Designs for Derived Data Types in MPI~302 to 311~10334470~10334470~OSTI~2022-09-09 21:03:24.96","International Workshop on Parallel Programming Models and Systems Software for High-End Computing~2022~Tran, A. and Michalowicz, B. and Ramesh, B. and Subramoni, H. and Shafi, A. and Panda, DK.~https://doi.org/10.1145/3547276.3548524~Designing Hierarchical Multi-HCA Aware Allgather in MPI~10355067~10355067~OSTI~2022-09-09 21:03:24.116","IEEE International Parallel and Distributed Processing Symposium Workshops~2022~Xu, Shulei and Shafi, Aamir and Subramoni, Hari and Panda, Dhabaleswar K.~https://doi.org/10.1109/IPDPSW55747.2022.00083~Arm meets Cloud: A Case Study of MPI Library Performance on AWS Arm-based HPC Cloud with Elastic Fabric Adapter~449 to 456~10355068~10355068~OSTI~2022-09-09 21:03:21.146","ISC HIGH PERFORMANCE~2022~Zhou, Q. and Kousha, P. and Anthony, Q. and Khorassani, K. and Shafi, A. and Subramoni, H. and Panda, DK.~https://doi.org/10.1007/978-3-031-07312-0_1~Accelerating MPI All-to-All Communication with Online Compression on Modern GPU Clusters~3-25~10355069~10355069~OSTI~2022-09-09 21:03:24.276"],"startDate":"09/01/2018","title":"Computation for the Endless Frontier","transType":"Cooperative Agreement","ueiNumber":"V6AFQPN18437"}],"metadata":{"offset":0,"rpp":25,"totalCount":1}}}