AWS Public Sector Blog

Erin Chu

Author: Erin Chu

Erin Chu is the life sciences lead on the Amazon Web Services (AWS) open data team. Trained to bridge the gap between the clinic and the lab, Erin is a veterinarian and a molecular geneticist, and spent the last four years in the companion animal genomics space. She is dedicated to helping speed time to science through interdisciplinary collaboration, communication, and learning.

36 new or updated datasets on the Registry of Open Data: AI analysis-ready datasets and more

36 new or updated datasets on the Registry of Open Data: AI analysis-ready datasets and more

This quarter, AWS released 36 new or updated datasets. As July 16 is Artificial Intelligence (AI) Appreciation Day, the AWS Open Data team is highlighting three unique datasets that are analysis-ready for AI. What will you build with these datasets?

33 new or updated datasets on the Registry of Open Data for Earth Day and more

The AWS Open Data Sponsorship Program makes high-value, cloud-optimized datasets publicly available on AWS. Through this program, customers are making over 100PB of high-value, cloud-optimized data available for public use. As April 22 is Earth Day, the AWS Open Data team wanted to highlight some new datasets from our geospatial and environmental communities of practice, as well as the other new or updated datasets available now on the Registry of Open Data on AWS and also discoverable on AWS Data Exchange.

Scientist looks at an image of a brain scan on a computer.

34 new or updated datasets on the Registry of Open Data: New data for land use, Alzheimer’s Disease, and more

The AWS Open Data Sponsorship Program makes high-value, cloud-optimized datasets publicly available on AWS. This quarter, AWS released 34 new or updated datasets from Impact Observatory, The Allen Institute for Brain Science, Common Screens, and others, which are available now on the Registry of Open Data in the following categories.

NYU Langone Center increases MRI accessibility through cooperative data sharing and research

About 40 million MRI scans are performed in the United States every year. MRIs are a valuable part of diagnostic plans, but as they exist today, they may not always be a part of a patient’s care plan. A research team at the New York University (NYU) Langone Center set out to make MRIs more accessible for more patients by using artificial intelligence (AI), machine learning (ML), and the power of cooperative open data sharing.

coronavirus

Taking COVID in STRIDES: The National Center for Biotechnology Information makes coronavirus genomic data available on AWS

AWS and the National Institutes of Health’s (NIH) National Center for Biotechnology Information (NCBI) announced the creation of the Coronavirus Genome Sequence Dataset to support COVID-19 research. The dataset is hosted by the AWS Open Data Sponsorship Program and accessible on the Registry of Open Data on AWS, providing researchers quick and easy access to coronavirus sequence data at no cost for use in their COVID-19 research.

genomic makeup data

Stanford researchers accelerate autism research by sharing genomic data in the cloud

In 2014, the Wall Lab at Stanford University sought to answer one of the most pressing questions in neuroscience: What genes influence autism spectrum disorder (ASD)? According to the Centers for Disease Control (CDC), this neurodevelopmental disorder affects roughly one in 54 children in America and is on the rise—nearly tripling since 1992. In the lab’s study of ASD genetics, they chose the cloud—and a unique experimental approach—to speed the time to science.