NASA-IMPACT / NASA-IMPACT/veda-odd

ODD PI 26.2 Objective 6: 🛰️ Propose unified strategy for virtualization of orbital swath (“L2”) data

Open
#309 0 comments 0 reactions 3 assignees View on GitHub

@hrodmn is already working on this.

Since Jan 20, 2026.

initiative: Virtualization pi-26.2-objective repo:datafusion-contrib/arrow-zarr repo:developmentseed/obspec-utils repo:ds/async-tiff repo:ds/obspec repo:ds/obstore repo:ds/zarr-datafusion-search repo:geoarrow/geoarrow-rs repo:vz/virtual-tiff repo:zarr/virtualizarr team: Geospatial Data
Dominant language
No language data
Stars
5
Forks
0
Avg merge
4d 19h
Merged PRs (30d)
3

Description

Why us?

We have developed many of the underlying technologies for virtualization, led numerous workshops and presentations on the concept, and are well-positioned to collaborate across NASA to propose a unified strategy.

Why now?

Cloud-native access to L2 data is particularly difficult. These challenges will be compounded by the influx of data from the NISAR project. We aim to develop a strategy to solve these issues.

Milestone
  • Sprint 1: Identify use cases, design benchmarking plan
  • Sprint 2 + 3: Execute benchmarking
  • Sprint 4: Completion of unified strategy document
Acceptance Criteria
  • Produce a diagram or ADR of what problems/use cases this is trying to solve
  • Benchmarking of at least 1 approach
  • Stretch: Production of a unified strategy document for virtualization and metadata querying of L2 data
Tasks
  • Groundwork phase (sprint 1)
    • Identify use-cases for new L2 solutions - e.g., WorldView, MAAP?, ISMIP6/7, Project focused virtualization
    • Produce a design document for benchmarking scalability across multiple metadata management schemes, including zarr-datafusion-search, stac-geoparquet, and icechunk (e.g., time to write, time to query, time to query with different filters)
      • Proposed: Generate a demonstration Icechunk store containing a metadata schema with 5 or more "columns" (with one of them being geometry) and 1 million entries for HLS data stored in a bucket in us-west-2.
  • Evaluation phase (sprint 2-3)
    • Implement and run benchmarks for comparison to other data and metadata management options
  • Reporting phase (sprint 4)
    • Collaborate on a unified strategy document for metadata management for L2 data
  • Ongoing
    • Prepare Zarr V3 standard to the ESCO office to enable virtualization of L2 products
  • Stretch:
    • Investigate visualization approaches for virtualized L2 data
    • Incrementally add user required functionality such as spatial search and schema projections.
    • Prepare Icechunk standard for the ESCO office to enable virtualization of L2 products

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.