filplus-bookkeeping / FF-Social-Impact

Bookkeeping repo for Allocator #1096
1 stars 0 forks source link

[DataCap Application] NASA - GEDI Mission Collection #6

Open SethDocherty opened 5 months ago

SethDocherty commented 5 months ago

Data Owner Name

NASA

Data Owner Country/Region

United States

Data Owner Industry

Environment

Website

https://gedi.umd.edu/data/products/

Social Media Handle

https://twitter.com/GEDI_Knights

Social Media Type

Twitter

What is your role related to the dataset

Data Preparer

Total amount of DataCap being requested

1.2PiB

Expected size of single dataset (one copy)

300TiB

Number of replicas to store

4

Weekly allocation of DataCap requested

15TiB

On-chain address for first allocation

f27ebdfco3lzldec53embc36tnysdmreigpnii7ti f154wr5fvxypdjbntk5xxpoifh3hfpjf2xbzfpmpa

Data Type of Application

Public, Open Dataset (Research/Non-Profit)

Custom multisig

Identifier

No response

Share a brief history of your project and organization

The EASIER Data initiative kicked off during the Summer 2022 and is a two-year project in collaboration with the Filecoin Foundation for the Decentralized Web to build pipelines for storing and extracting geospatial data on Filecoin and IPFS. We have onboarded a near 200+ TB (per replication) of Landsat 9 satellite data for a single year. Our pipeline development will utilize this content to prototype and demonstrate the capabilities of decentralized access of existing tools our team has built and tutorials that have been developed by the GEDI Science team.

Is this project associated with other projects/ecosystem stakeholders?

Yes

If answered yes, what are the other projects/ecosystem stakeholders

University of Maryland
Filecoin Foundation for the Decentralized Web

Describe the data being stored onto Filecoin

Deployed to the ISS, the Global Ecosystem Dynamics Investigation (GEDI) mission is a full-waveform lidar instrument that makes detailed measurements of the 3D structure of the Earth’s surface. Lidar is an active remote sensing technology (the laser version of radar) which uses pulses of laser light to measure 3D structure. The light is reflected by the ground, vegetation and any clouds and is then collected by GEDI’s telescope. The sole GEDI observable is the waveform from which all other data products are derived. Signal processing is used to identify the ground within the waveform. The distribution of laser energy above the ground can be used to determine the height and density of objects within the footprint. The view geometry and active use of light by lidar allows the ground to be identified through small gaps in the tree canopy, enabling unsaturated measurements of much denser forests than is possible with either passive optical (such as spaceborne cameras) or short wavelength radar systems. In addition, uniquely amongst satellite remote sensing, the height and vertical distribution are direct measurements than can be compared to field observations.

GEDI science data products include footprint and gridded data sets that describe the 3D features of the Earth in HDF5 and Geotiff format.

Additional details on the science objectives of GEDI can be found [here](https://gedi.umd.edu/science/objectives-overview/)

Where was the data currently stored in this dataset sourced from

AWS Cloud

If you answered "Other" in the previous question, enter the details here

NASA's Earth Observing System Data and Information System (EOSDIS) are stored and maintained across 12 Distributed Active Achieve Centers (DAACs).  Each of the DAACs have unique expertise serving science discipline-oriented user communities, therefore data retrieval can be served directly from an AWS S3 bucket (us-west-2 region) or HTTP endpoint maintained by the DAACs infrastructure.

If you are a data preparer. What is your location (Country/Region)

United States

If you are a data preparer, how will the data be prepared? Please include tooling used and technical details?

We'll be preparing data using Singularity-V2 by using the inline preparation capability and content distribution will be accessible to storage providers through our hosted EC2 instance.

If you are not preparing the data, who will prepare the data? (Provide name and business)

No response

Has this dataset been stored on the Filecoin network before? If so, please explain and make the case why you would like to store this dataset again to the network. Provide details on preparation and/or SP distribution.

No response

Please share a sample of the data

The following is from Earthdata Search, containing a list of publicly available collections from the GEDI mission.

https://search.earthdata.nasa.gov/search?fdc=Land%2BProcess%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528LPDAAC%2529!Oak%2BRidge%2BNational%2BLaboratory%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528ORNL%2BDAAC%2529&fpj=GEDI

Confirm that this is a public dataset that can be retrieved by anyone on the Network

If you chose not to confirm, what was the reason

The data are supported by open missions and are meant to be public.

What is the expected retrieval frequency for this data

Weekly

For how long do you plan to keep this dataset stored on Filecoin

Permanently

In which geographies do you plan on making storage deals

Asia other than Greater China, North America, Europe, Australia (continent)

How will you be distributing your data to storage providers

HTTP or FTP server

How did you find your storage providers

Slack, Partners

If you answered "Others" in the previous question, what is the tool or platform you used

No response

Please list the provider IDs and location of the storage providers you will be working with.

f02639429, United States

How do you plan to make deals to your storage providers

Singularity

If you answered "Others/custom tool" in the previous question, enter the details here

No response

Can you confirm that you will follow the Fil+ guideline

Yes

datacap-bot[bot] commented 5 months ago

Application is waiting for allocator review

datacap-bot[bot] commented 5 months ago

Datacap Request Trigger

Total DataCap requested

1.2PiB

Expected weekly DataCap usage rate

15TiB

DataCap Amount - First Tranche

15TiB

Client address

f27ebdfco3lzldec53embc36tnysdmreigpnii7ti

datacap-bot[bot] commented 5 months ago

DataCap Allocation requested

Multisig Notary address

Client address

f27ebdfco3lzldec53embc36tnysdmreigpnii7ti

DataCap allocation requested

15TiB

Id

6af65d13-e6ea-481d-ab94-01ce79486f86

datacap-bot[bot] commented 5 months ago

Application is ready to sign

datacap-bot[bot] commented 5 months ago

Request Approved

Your Datacap Allocation Request has been approved by the Notary

Message sent to Filecoin Network

bafy2bzacedrly3vwp6mzr7q3c5cnrpx2vxh7znz3mqsvjb3hlpqjvwnteetxm

Address

f27ebdfco3lzldec53embc36tnysdmreigpnii7ti

Datacap Allocated

15TiB

Signer Address

f13kcetuvgmoebigp2efnpbiynmcpmenrspygwcpi

Id

6af65d13-e6ea-481d-ab94-01ce79486f86

You can check the status of the message here: https://filfox.info/en/message/bafy2bzacedrly3vwp6mzr7q3c5cnrpx2vxh7znz3mqsvjb3hlpqjvwnteetxm

datacap-bot[bot] commented 5 months ago

Application is Granted

datacap-bot[bot] commented 5 months ago

Issue has been modified. Changes below:

(OLD vs NEW)

Confirm that this is a public dataset that can be retrieved by anyone on the network: [x] I confirm vs [X] I confirm State: ChangesRequested vs Granted

datacap-bot[bot] commented 5 months ago

Issue has been modified. Changes below:

(OLD vs NEW)

datacap-bot[bot] commented 5 months ago

Issue information change request has been approved.

datacap-bot[bot] commented 5 months ago

Application not found. If you have modified the wallet address, please create a new application.

datacap-bot[bot] commented 5 months ago

Issue information change request has been approved.