Scientific Data Management on AWS with Open Source Quilt Data Packages
Industries Blog
This article discusses how to manage scientific data on AWS using open source Quilt Data Packages. It covers best practices for setting up Amazon S3 for scientific data management, packaging data to improve visibility and trust, moving on-premises data to AWS while capturing metadata, and leveraging event-driven automation to connect applications.
Specifically, the article covers:
- Setting up Amazon S3 for scientific data management with features like versioning, object lock, checksums, lifecycle configuration, encryption, and CloudTrail for auditing.
- Using Quilt Data Packages to logically organize data, capture metadata, and ensure data integrity through top-hashes and revisions.
- Moving on-premises data to AWS using services like DataSync, Storage Gateway, Lambda, and IoT Greengrass while capturing metadata.
- Leveraging the Quilt Data Platform for data discovery, cataloging, search, and integration with other AWS services.
- Connecting systems in a laboratory data mesh using event-driven automation and Quilt's open APIs.
- A case study of how Inari Agriculture uses Quilt and AWS to accelerate product development and commercialization.
- Conclusion and resources for getting started with Quilt Data Packages on AWS.
The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.
Related articles
2024
2024
2026
2024
The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.