# Public Virtual Zarr of the NASA (NEX-GDDP-CMIP6) Dataset

**URL:** <https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802>\
**Category:** Data\
**Tags:** zarr, virtual-zarr\
**Created:** [July 28, 2026, 10:12pm UTC](https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802 "2026-07-28T22:12:48Z")\
**Posts on this page:** 4\
**Page:** 1

<div class="post-metadata">

**Author:** ![norlandrhagen](https://yyz2.discourse-cdn.com/flex030/user_avatar/discourse.pangeo.io/norlandrhagen/32/1966_2.png) [@norlandrhagen](https://discourse.pangeo.io/u/norlandrhagen)\
**Post date:** [July 28, 2026, 10:12pm UTC](https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802/1 "2026-07-28T22:12:48Z")

</div>

In a recent project, I needed some downscaled CMIP6 data. I created a [Virtual Zarr store](https://github.com/virtual-zarr/nex-gddp-cmip6/) of the NEX-GDDP-CMIP6 dataset. I wanted to share it if anyone has use of this dataset and doesn’t want to go through the work of re-virtualizing it.

With [VirtualiZarr](https://github.com/zarr-developers/VirtualiZarr), we can take 38TB of NetCDF files and make a virtual Icechunk datatree that is under a GB (740MB).

---

<div class="post-metadata">

**Author:** ![Nicola\_Masotti](https://yyz2.discourse-cdn.com/flex030/user_avatar/discourse.pangeo.io/nicola_masotti/32/3792_2.png) [@Nicola\_Masotti](https://discourse.pangeo.io/u/Nicola_Masotti)\
**Post date:** [July 29, 2026, 7:34pm UTC](https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802/2 "2026-07-29T19:34:39Z")

</div>

Hi, not sure if this is of any interest to you, but we also zarrified a bunch of CMIP6 datasets. As opposed to using [VirtualiZarr](https://github.com/zarr-developers/VirtualiZarr) , these are native Zarrs (we did it the hard way). You can find them at [CMIP6 - Earth Data Hub](https://earthdatahub.destine.eu/collections/cmip6)

We are considering to do more.

---

<div class="post-metadata">

**Author:** ![norlandrhagen](https://yyz2.discourse-cdn.com/flex030/user_avatar/discourse.pangeo.io/norlandrhagen/32/1966_2.png) [@norlandrhagen](https://discourse.pangeo.io/u/norlandrhagen)\
**Post date:** [July 31, 2026, 4:02pm UTC](https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802/3 "2026-07-31T16:02:01Z")

</div>

Hey @Nicola_Masotti, super cool to see people making CMIP easier to use!

Any reason why you preferred native Zarr?

---

<div class="post-metadata">

**Author:** ![Nicola\_Masotti](https://yyz2.discourse-cdn.com/flex030/user_avatar/discourse.pangeo.io/nicola_masotti/32/3792_2.png) [@Nicola\_Masotti](https://discourse.pangeo.io/u/Nicola_Masotti)\
**Post date:** [August 24, 2026, 8:52am UTC](https://discourse.pangeo.io/t/public-virtual-zarr-of-the-nasa-nex-gddp-cmip6-dataset/5802/4 "2026-08-24T08:52:30Z")

</div>

In our case we wanted to fine tune the read performance of the result, thus wee needed to fully control the chunking. Considering the relatively low complexity of the CMIP6 daily data we converted ou to now, I do agree that also using VirtualiZarr could have been an option 😉
