Core Components
VEDA is made up of three primary components that work together to deliver an integrated data exploration and experience.
VEDA Data Catalog and Data Services
The interoperable data catalog and standards-based data services form the backbone of VEDA’s infrastructure. These services enable access and discovery of distributed data sources through a centralized interface, including those from NASA’s Earth Observing System Data and Information System (EOSDIS), which supports scalable access to large volumes of science data.
The data catalog and services provide numerous key capabilities:
- Scalable data access for modeling and product development
- Open-source best practices to enable accessing, visualizing and analyzing multiple data formats on the fly, all in the cloud
- Support for both existing Analysis-Ready Cloud Optimized (ARCO) datasets and external sources via integrations like ESRI services and NASA’s Global Imagery Browse Service (GIBS)
- Integration with NASA’s authoritative data archive, as well as externally hosted and archived data within the commercial cloud
- A cloud-first approach for storing and visualizing Earth observation data, especially data not yet formalized in long-term repositories
This flexible and modular cataloging approach makes VEDA accessible to other organizations interested in adopting open-source, cloud-based science infrastructure.
VEDA Dashboard
The VEDA Dashboard is a public-facing, browser-based visualization and data-driven storytelling platform. It’s designed for science teams and organizations that want to share their data and stories broadly but lack the time, budget, or capacity to build visualization tools from scratch. The VEDA Dashboard also supports the use of open-source solutions and specialized visualization tools providing content providers with more control in communicating information to the public.
The dashboard is modular and replicable, allowing organizations to adopt only the tools they need. All components are purpose-built for science-use cases and work seamlessly with ARCO datasets stored in the cloud. Capabilities of the dashboard include visualization tools, data-driven storytelling, data summaries, and science examples, allowing for tailored site structure and content.
VEDA Hub
The VEDA Hub is a dedicated JupyterHub environment that supports a users transition from on-premise computing to the cloud to perform data analysis at scale. With the VEDA Hub, scientists can more efficiently progress through the science data lifecycle through:
- Access to datasets at scale for compute intensive analysis without managing data and compute resources manually
- Customizable cloud compute resources that streamline reproducibility of scientific analysis
- Collaborative shared workspaces and data
- Support for multiple programming languages, paradigms, and tooling
- Training materials for and hosting stakeholder workshops
By reducing the difficulty of managing infrastructure, the VEDA Hub makes it easier for teams to focus on advancing their science.