filplus-bookkeeping / FF-Social-Impact

Bookkeeping repo for Allocator #1096
1 stars 0 forks source link

[DataCap Application] EASIER Data Initiative - GEDI Mission Collection #10

Open SethDocherty opened 3 weeks ago

SethDocherty commented 3 weeks ago

Data Owner Name

NASA

Data Owner Country/Region

United States

Data Owner Industry

Environment

Website

https://gedi.umd.edu/data/products/

Social Media Handle

https://twitter.com/GEDI_Knights

Social Media Type

Twitter

What is your role related to the dataset

Data Preparer

Total amount of DataCap being requested

1.2PiB

Expected size of single dataset (one copy)

300TiB

Number of replicas to store

4

Weekly allocation of DataCap requested

15TiB

On-chain address for first allocation

f154wr5fvxypdjbntk5xxpoifh3hfpjf2xbzfpmpa

Data Type of Application

Public, Open Dataset (Research/Non-Profit)

Custom multisig

Identifier

No response

Share a brief history of your project and organization

The EASIER Data initiative kicked off during the Summer 2022 and is a two-year project in collaboration with the Filecoin Foundation for the Decentralized Web to build pipelines for storing and extracting geospatial data on Filecoin and IPFS. We have onboarded a near 200+ TB (per replication) of Landsat 9 satellite data for a single year. Our pipeline development will utilize this content to prototype and demonstrate the capabilities of decentralized access of existing tools our team has built and tutorials that have been developed by the GEDI Science team.

Is this project associated with other projects/ecosystem stakeholders?

Yes

If answered yes, what are the other projects/ecosystem stakeholders

University of Maryland
Filecoin Foundation for the Decentralized Web

Describe the data being stored onto Filecoin

Deployed to the ISS, the Global Ecosystem Dynamics Investigation (GEDI) mission is a full-waveform lidar instrument that makes detailed measurements of the 3D structure of the Earth’s surface. Lidar is an active remote sensing technology (the laser version of radar) which uses pulses of laser light to measure 3D structure. The light is reflected by the ground, vegetation and any clouds and is then collected by GEDI’s telescope. The sole GEDI observable is the waveform from which all other data products are derived. Signal processing is used to identify the ground within the waveform. The distribution of laser energy above the ground can be used to determine the height and density of objects within the footprint. The view geometry and active use of light by lidar allows the ground to be identified through small gaps in the tree canopy, enabling unsaturated measurements of much denser forests than is possible with either passive optical (such as spaceborne cameras) or short wavelength radar systems. In addition, uniquely amongst satellite remote sensing, the height and vertical distribution are direct measurements than can be compared to field observations.

GEDI science data products include footprint and gridded data sets that describe the 3D features of the Earth in HDF5 and Geotiff format.

Additional details on the science objectives of GEDI can be found [here](https://gedi.umd.edu/science/objectives-overview/)

Where was the data currently stored in this dataset sourced from

AWS Cloud

If you answered "Other" in the previous question, enter the details here

NASA's Earth Observing System Data and Information System (EOSDIS) are stored and maintained across 12 Distributed Active Achieve Centers (DAACs).  Each of the DAACs have unique expertise serving science discipline-oriented user communities, therefore data retrieval can be served directly from an AWS S3 bucket (us-west-2 region) or HTTP endpoint maintained by the DAACs infrastructure.

If you are a data preparer. What is your location (Country/Region)

United States

If you are a data preparer, how will the data be prepared? Please include tooling used and technical details?

The following is from Earthdata Search, containing a list of publicly available collections from the GEDI mission.

https://search.earthdata.nasa.gov/search?fdc=Land%2BProcess%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528LPDAAC%2529!Oak%2BRidge%2BNational%2BLaboratory%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528ORNL%2BDAAC%2529&fpj=GEDI

If you are not preparing the data, who will prepare the data? (Provide name and business)

No response

Has this dataset been stored on the Filecoin network before? If so, please explain and make the case why you would like to store this dataset again to the network. Provide details on preparation and/or SP distribution.

No response

Please share a sample of the data

The following is from Earthdata Search, containing a list of publicly available collections from the GEDI mission.

https://search.earthdata.nasa.gov/search?fdc=Land%2BProcess%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528LPDAAC%2529!Oak%2BRidge%2BNational%2BLaboratory%2BDistributed%2BActive%2BArchive%2BCenter%2B%2528ORNL%2BDAAC%2529&fpj=GEDI

Confirm that this is a public dataset that can be retrieved by anyone on the Network

If you chose not to confirm, what was the reason

The data are supported by open missions and are meant to be public.

What is the expected retrieval frequency for this data

Weekly

For how long do you plan to keep this dataset stored on Filecoin

Permanently

In which geographies do you plan on making storage deals

Asia other than Greater China, North America, Europe, Australia (continent)

How will you be distributing your data to storage providers

HTTP or FTP server

How did you find your storage providers

Slack, Partners

If you answered "Others" in the previous question, what is the tool or platform you used

No response

Please list the provider IDs and location of the storage providers you will be working with.

f02639429, United States

How do you plan to make deals to your storage providers

Singularity

If you answered "Others/custom tool" in the previous question, enter the details here

No response

Can you confirm that you will follow the Fil+ guideline

Yes

datacap-bot[bot] commented 3 weeks ago

Application is waiting for allocator review

datacap-bot[bot] commented 3 weeks ago

Datacap Request Trigger

Total DataCap requested

1.2PiB

Expected weekly DataCap usage rate

15TiB

DataCap Amount - First Tranche

32GiB

Client address

f154wr5fvxypdjbntk5xxpoifh3hfpjf2xbzfpmpa

datacap-bot[bot] commented 3 weeks ago

DataCap Allocation requested

Multisig Notary address

Client address

f154wr5fvxypdjbntk5xxpoifh3hfpjf2xbzfpmpa

DataCap allocation requested

32GiB

Id

69806a74-958f-4924-a32b-01d8d9456c2f

datacap-bot[bot] commented 3 weeks ago

Application is ready to sign

datacap-bot[bot] commented 3 weeks ago

Request Approved

Your Datacap Allocation Request has been approved by the Notary

Message sent to Filecoin Network

bafy2bzacea2njcxoivntiyvyjy4qvghxdmlp4drt55gdrgpuz6zrppn44eroy

Address

f154wr5fvxypdjbntk5xxpoifh3hfpjf2xbzfpmpa

Datacap Allocated

32GiB

Signer Address

f13kcetuvgmoebigp2efnpbiynmcpmenrspygwcpi

Id

69806a74-958f-4924-a32b-01d8d9456c2f

You can check the status of the message here: https://filfox.info/en/message/bafy2bzacea2njcxoivntiyvyjy4qvghxdmlp4drt55gdrgpuz6zrppn44eroy

datacap-bot[bot] commented 3 weeks ago

Application is Granted

datacap-bot[bot] commented 3 weeks ago

Application is in Refill