Overview of ARCO ERA5 data types
mainThe ARCO ERA5 project provides three distinct types of data stored in the gcp-public-data-arco-era5 bucket (located in the us-central1 region) to serve different research and machine learning needs:
- Analysis Ready (
$BUCKET/ar/): A unified, ML-ready version of both surface and atmospheric data in Zarr format. This version includes standard pre-processing and chunk optimization for common research workflows. - Cloud Optimized (
$BUCKET/co/): A direct port of the gaussian-gridded ERA5 data to Zarr format, providing a cloud-optimized alternative to the original GRIB files without additional pre-processing. - Raw Data (
$BUCKET/raw/): The original raw GRIB and NetCDF data files.
Data is updated on a monthly cadence (around the 9th of each month) with a 3-month delay for stable ERA5. Preliminary ERA5T data is available with approximately a 1-week delay.