# Uncategorized

**URL:** https://discourse.pangeo.io/c/uncategorized/1.md

[Latest](https://discourse.pangeo.io/latest.md) · [Categories](https://discourse.pangeo.io/categories.md) · [Tags](https://discourse.pangeo.io/tags.md)

---

## [CNG Forum 2026 - Saturday is the final day for the room block](https://discourse.pangeo.io/t/cng-forum-2026-saturday-is-the-final-day-for-the-room-block/5810)

<div class="topic-metadata">

**Author:** [@maxrjones](https://discourse.pangeo.io/u/maxrjones)\
**Replies:** 2\
**Last updated:** [September 23, 2026, 3:40pm UTC](https://discourse.pangeo.io/t/cng-forum-2026-saturday-is-the-final-day-for-the-room-block/5810 "2026-09-23T15:40:59Z")

</div>

Hey Pangeo Community, Several of my Development Seed colleagues will be at the 2026 CNG Forum in Snowbird, Utah. I’m sad to not be attending, but I’d encourage others who are interested in sharing knowledge and advancin…

---

## [Dask in Rust -- Frisky](https://discourse.pangeo.io/t/dask-in-rust-frisky/5755)

<div class="topic-metadata">

**Author:** [@mrocklin](https://discourse.pangeo.io/u/mrocklin)\
**Replies:** 6\
**Last updated:** [July 24, 2026, 7:12pm UTC](https://discourse.pangeo.io/t/dask-in-rust-frisky/5755 "2026-07-24T19:12:10Z")

</div>

I rebuilt the Dask scheduler and distributed compute system in Rust. It’s called Frisky. It’s fast, but also very early and could use some adventurous users to hammer on it. Here is a general overview: Why - Frisky H…

---

## [Best way of copying and rechunking a remote zarr store](https://discourse.pangeo.io/t/best-way-of-copying-and-rechunking-a-remote-zarr-store/5789)

<div class="topic-metadata">

**Author:** [@guidocioni](https://discourse.pangeo.io/u/guidocioni)\
**Replies:** 2\
**Last updated:** [July 23, 2026, 1:10pm UTC](https://discourse.pangeo.io/t/best-way-of-copying-and-rechunking-a-remote-zarr-store/5789 "2026-07-23T13:10:01Z")

</div>

I want to create a local copy of some of the variables contained in one of the many ERA5 zarr replicas available (here I chose to use the google cloud storage one for convenience). Why? Because we need quick access for a…

---

## [AquaScope: one Python schema over 18 water-data collectors (12 agencies), plus hydrology analysis. Looking for feedback on the xarray/interop story](https://discourse.pangeo.io/t/aquascope-one-python-schema-over-18-water-data-collectors-12-agencies-plus-hydrology-analysis-looking-for-feedback-on-the-xarray-interop-story/5753)

<div class="topic-metadata">

**Author:** [@Rekin226](https://discourse.pangeo.io/u/Rekin226)\
**Replies:** 0\
**Last updated:** [June 16, 2026, 2:47am UTC](https://discourse.pangeo.io/t/aquascope-one-python-schema-over-18-water-data-collectors-12-agencies-plus-hydrology-analysis-looking-for-feedback-on-the-xarray-interop-story/5753 "2026-06-16T02:47:11Z")

</div>

Hi all, I’m a postdoc working on water resources, and I’ve been building AquaScope, an MIT-licensed Python toolkit that does two things: it pulls from 18 collectors across 12 water-data agencies (USGS, FAO AQUASTAT, FAO …

---

## [GOES-16 as a single virtual Zarr store](https://discourse.pangeo.io/t/goes-16-as-a-single-virtual-zarr-store/5721)

<div class="topic-metadata">

**Author:** [@TomNicholas](https://discourse.pangeo.io/u/TomNicholas)\
**Replies:** 2\
**Last updated:** [June 4, 2026, 2:17am UTC](https://discourse.pangeo.io/t/goes-16-as-a-single-virtual-zarr-store/5721 "2026-06-04T02:17:54Z")

</div>

I stress-tested VirtualiZarr and Icechunk by ingesting all of the GOES-16 Cloud and Moisture Imagery Product archive into a single store. As far as I know this is the biggest virtual Zarr store ever (by number of chunks…

---

## [Dawnloading ICON D2 data june 2021](https://discourse.pangeo.io/t/dawnloading-icon-d2-data-june-2021/5536)

<div class="topic-metadata">

**Author:** [@Nishant\_Pawar](https://discourse.pangeo.io/u/Nishant_Pawar)\
**Replies:** 1\
**Last updated:** [May 15, 2026, 7:07pm UTC](https://discourse.pangeo.io/t/dawnloading-icon-d2-data-june-2021/5536 "2026-05-15T19:07:07Z")

</div>

im looking for icon d2 data for the month june 2021, if anyone have leads share your thoughts.

---

## [Building Earth Data Workflows with Agents](https://discourse.pangeo.io/t/building-earth-data-workflows-with-agents/5586)

<div class="topic-metadata">

**Author:** [@rsignell](https://discourse.pangeo.io/u/rsignell)\
**Replies:** 5\
**Last updated:** [May 12, 2026, 2:53pm UTC](https://discourse.pangeo.io/t/building-earth-data-workflows-with-agents/5586 "2026-05-12T14:53:02Z")

</div>

How best to use Agents to help build Earth Data workflows? I’ve been just feeding the agents a bunch of existing workflows (e.g. Notebooks) and letting the agents use them as source material to construct new workflows. …

---

## [Optimizing data load/compute when using STAC API](https://discourse.pangeo.io/t/optimizing-data-load-compute-when-using-stac-api/5563)

<div class="topic-metadata">

**Author:** [@Valentina\_Premier](https://discourse.pangeo.io/u/Valentina_Premier)\
**Replies:** 1\
**Last updated:** [April 20, 2026, 4:00pm UTC](https://discourse.pangeo.io/t/optimizing-data-load-compute-when-using-stac-api/5563 "2026-04-20T16:00:47Z")

</div>

Hi all, I am wondering if I can use dask to speed up the computation/data load over a relatively large area (Euregio region). I have a single time step and I want to upload Sentinel-2 data though the CDSE STAC API. I a…

---

## [Efficient extraction of concurrent variables using rolling-window argmax indices](https://discourse.pangeo.io/t/efficient-extraction-of-concurrent-variables-using-rolling-window-argmax-indices/5541)

<div class="topic-metadata">

**Author:** [@chaithra](https://discourse.pangeo.io/u/chaithra)\
**Replies:** 8\
**Last updated:** [March 28, 2026, 10:05am UTC](https://discourse.pangeo.io/t/efficient-extraction-of-concurrent-variables-using-rolling-window-argmax-indices/5541 "2026-03-28T10:05:18Z")

</div>

Hi, I am working with CMIP6 daily piControl data. My objective is to identify the 5-day period of maximum precipitation (Rx5day) for each year, and then extract the concurrent daily data for other variables (temperature…

---

## [Chunk-level GPU parallelism in using cutile-python?](https://discourse.pangeo.io/t/chunk-level-gpu-parallelism-in-using-cutile-python/5469)

<div class="topic-metadata">

**Author:** [@TomNicholas](https://discourse.pangeo.io/u/TomNicholas)\
**Replies:** 5\
**Last updated:** [March 23, 2026, 11:10pm UTC](https://discourse.pangeo.io/t/chunk-level-gpu-parallelism-in-using-cutile-python/5469 "2026-03-23T23:10:55Z")

</div>

NVIDIA recently released CUDA-tiles, which basically allows you to specify how you want to do chunk-level parallelism in python code and have the GPU just do it. See cutile-python. I’m not really a GPU person, but could…

---

## [RasterioIOError when using CDSE STAC API](https://discourse.pangeo.io/t/rasterioioerror-when-using-cdse-stac-api/5540)

<div class="topic-metadata">

**Author:** [@Valentina\_Premier](https://discourse.pangeo.io/u/Valentina_Premier)\
**Replies:** 2\
**Last updated:** [March 23, 2026, 8:51am UTC](https://discourse.pangeo.io/t/rasterioioerror-when-using-cdse-stac-api/5540 "2026-03-23T08:51:08Z")

</div>

Hi all, I need someone’s help to understand what is going wrong in my workflow! I am trying to read Sentinel-2 data through the CDSE STAC API and stackstac pyhton library and applying a workflow for snow classification…

---

## [\`scipy\` is considering to mark the netcdf3-reader in \`scipy.io\` as "legacy"](https://discourse.pangeo.io/t/scipy-is-considering-to-mark-the-netcdf3-reader-in-scipy-io-as-legacy/5542)

<div class="topic-metadata">

**Author:** [@keewis](https://discourse.pangeo.io/u/keewis)\
**Replies:** 0\
**Last updated:** [March 22, 2026, 1:13am UTC](https://discourse.pangeo.io/t/scipy-is-considering-to-mark-the-netcdf3-reader-in-scipy-io-as-legacy/5542 "2026-03-22T01:13:41Z")

</div>

scipy is considering to mark the netcdf reader in scipy.io as “legacy” (which means “not deprecated, but minimally maintained”). See RFC: declaring netcdf IO functions in \`scipy.io\` as legacy · Issue #24885 · scipy/scipy…

---

## [Curated GIS and Remote Sensing resource list](https://discourse.pangeo.io/t/curated-gis-and-remote-sensing-resource-list/5476)

<div class="topic-metadata">

**Author:** [@dhersh3094](https://discourse.pangeo.io/u/dhersh3094)\
**Replies:** 2\
**Last updated:** [January 7, 2026, 9:33am UTC](https://discourse.pangeo.io/t/curated-gis-and-remote-sensing-resource-list/5476 "2026-01-07T09:33:26Z")

</div>

Hi everyone, I maintain a database of GIS and remote sensing resources at Geospatial Catalog. It’s a collection of 700+ links organized into categories and with tags. I hope this may be helpful for you as a place to se…

---

## [Installing ESMF](https://discourse.pangeo.io/t/installing-esmf/5210)

<div class="topic-metadata">

**Author:** [@ulfata](https://discourse.pangeo.io/u/ulfata)\
**Replies:** 2\
**Last updated:** [December 30, 2025, 9:04pm UTC](https://discourse.pangeo.io/t/installing-esmf/5210 "2025-12-30T21:04:35Z")

</div>

Hi Pangeo/xESMF community, I’m trying to use xESMF with the ESMF/ESMPy backend on Google Colab for regridding, but I keep running into problems with the ESMF installation. Specifically, I get errors like: ModuleNotFoun…

---

## [Fall showcase close-out: Lightning talks! (December 10, 2025 at 12 PM ET)](https://discourse.pangeo.io/t/fall-showcase-close-out-lightning-talks-december-10-2025-at-12-pm-et/5467)

<div class="topic-metadata">

**Author:** [@maxrjones](https://discourse.pangeo.io/u/maxrjones)\
**Replies:** 8\
**Last updated:** [December 23, 2025, 4:33pm UTC](https://discourse.pangeo.io/t/fall-showcase-close-out-lightning-talks-december-10-2025-at-12-pm-et/5467 "2025-12-23T16:33:17Z")

</div>

Title: “Fall showcase close-out: Lightning talks!!!” When: Wednesday, December 10, 2025 at 12 PM EST (2025-11-12T17:00:00…

---

## [Pangeo and friends at AGU25](https://discourse.pangeo.io/t/pangeo-and-friends-at-agu25/5472)

<div class="topic-metadata">

**Author:** [@maxrjones](https://discourse.pangeo.io/u/maxrjones)\
**Replies:** 0\
**Last updated:** [December 15, 2025, 3:09am UTC](https://discourse.pangeo.io/t/pangeo-and-friends-at-agu25/5472 "2025-12-15T03:09:43Z")

</div>

Hey folks, I started a spreadsheet to share any events of interest (sessions, talks, posters, social events, etc) during the AGU25 conference this week. Please add content! Hope to see many of you during the week!

---

## [New Cloud Tensor I/O Benchmarks - Zarr is fast now!](https://discourse.pangeo.io/t/new-cloud-tensor-i-o-benchmarks-zarr-is-fast-now/5459)

<div class="topic-metadata">

**Author:** [@rabernat](https://discourse.pangeo.io/u/rabernat)\
**Replies:** 4\
**Last updated:** [December 1, 2025, 3:03pm UTC](https://discourse.pangeo.io/t/new-cloud-tensor-i-o-benchmarks-zarr-is-fast-now/5459 "2025-12-01T15:03:35Z")

</div>

Given all of the recent development on Zarr, we decided to revisit the question of Zarr performance in the cloud. We recently published a blog post with our results. Here’s the key figure Bottom line: Zarr Python …

---

## [Tips for platform for teaching climate data analysis with python?](https://discourse.pangeo.io/t/tips-for-platform-for-teaching-climate-data-analysis-with-python/5005)

<div class="topic-metadata">

**Author:** [@sarambl](https://discourse.pangeo.io/u/sarambl)\
**Replies:** 10\
**Last updated:** [November 24, 2025, 3:17pm UTC](https://discourse.pangeo.io/t/tips-for-platform-for-teaching-climate-data-analysis-with-python/5005 "2025-11-24T15:17:55Z")

</div>

Hey all :slight\_smile: I’m a researcher at Stockholm University and we want to teach our students climate data analysis with python (xarray, intake, dask etc.) but we always spend so much time on installation. Does anyo…

---

## [Guidance on Correctly Serving Zarr Datasets Through pygeoapi](https://discourse.pangeo.io/t/guidance-on-correctly-serving-zarr-datasets-through-pygeoapi/5448)

<div class="topic-metadata">

**Author:** [@Kaboom\_Official](https://discourse.pangeo.io/u/Kaboom_Official)\
**Replies:** 0\
**Last updated:** [November 18, 2025, 7:47pm UTC](https://discourse.pangeo.io/t/guidance-on-correctly-serving-zarr-datasets-through-pygeoapi/5448 "2025-11-18T19:47:57Z")

</div>

Hello everyone. I’m trying to serve my data using pygeoapi, but I’m not able to get it working properly. I keep running into errors even after following the documentation closely. I tried loading the example Zarr file u…

---

## [Writing larger-than-memory COG from many NetCDF files](https://discourse.pangeo.io/t/writing-larger-than-memory-cog-from-many-netcdf-files/5429)

<div class="topic-metadata">

**Author:** [@ZZMitch](https://discourse.pangeo.io/u/ZZMitch)\
**Replies:** 3\
**Last updated:** [October 27, 2025, 6:25pm UTC](https://discourse.pangeo.io/t/writing-larger-than-memory-cog-from-many-netcdf-files/5429 "2025-10-27T18:25:24Z")

</div>

Hello, Over the last few days I have been struggling with this issue and was hoping to get some tips from the Pangeo community… I have 3390 NetCDF files that together represent environmental data at 30 m spatial resolu…

---

## [Xarray.dataset.grouby\_bins without squishing other dimensions](https://discourse.pangeo.io/t/xarray-dataset-grouby-bins-without-squishing-other-dimensions/1692)

<div class="topic-metadata">

**Author:** [@nwilliams6](https://discourse.pangeo.io/u/nwilliams6)\
**Replies:** 5\
**Last updated:** [October 22, 2025, 8:21am UTC](https://discourse.pangeo.io/t/xarray-dataset-grouby-bins-without-squishing-other-dimensions/1692 "2025-10-22T08:21:16Z")

</div>

I’m trying to use xarray.dataset.groupby\_bins to get a 1x1 monthly ocean dataset (dimensions of lat, lon time) averaged into potential density bins. Because this is a Southern Ocean dataset, what I want to end up with is…

---

## [Help with concurrent.futures](https://discourse.pangeo.io/t/help-with-concurrent-futures/5415)

<div class="topic-metadata">

**Author:** [@Michael\_Sumner](https://discourse.pangeo.io/u/Michael_Sumner)\
**Replies:** 3\
**Last updated:** [September 29, 2025, 3:39pm UTC](https://discourse.pangeo.io/t/help-with-concurrent-futures/5415 "2025-09-29T15:39:40Z")

</div>

Can anyone provide a simple example (using map() or submit() as here with ThreadPoolExecutor that actually processes data from objects and delivers a speedup when max\_workers is set \> 1? I’m struggling to find a real ex…

---

## [Fixing GHRSST - seeking opinions](https://discourse.pangeo.io/t/fixing-ghrsst-seeking-opinions/3833)

<div class="topic-metadata">

**Author:** [@Michael\_Sumner](https://discourse.pangeo.io/u/Michael_Sumner)\
**Replies:** 7\
**Last updated:** [September 26, 2025, 11:40am UTC](https://discourse.pangeo.io/t/fixing-ghrsst-seeking-opinions/3833 "2025-09-26T11:40:51Z")

</div>

The GHRSST NetCDF product (Multi-scale\_Ultra-high\_Resolution\_MUR-SST) has a number of problems: units are Kelvin, this destroys lazy load strategies with the need to recalculate having to be pushed all the way to the u…

---

## [Interested to learn hybrid modeling](https://discourse.pangeo.io/t/interested-to-learn-hybrid-modeling/5398)

<div class="topic-metadata">

**Author:** [@nrchow](https://discourse.pangeo.io/u/nrchow)\
**Replies:** 5\
**Last updated:** [September 16, 2025, 7:01am UTC](https://discourse.pangeo.io/t/interested-to-learn-hybrid-modeling/5398 "2025-09-16T07:01:04Z")

</div>

I am interested to learn about hybrid modeling (machine learning incorporated to conventional physics-based models) but unsure how to start. I have some background in numerical modeling, as in I have experience in prepar…

---

## [.to\_zarr(..., compute=False) gives zeros not NaNs](https://discourse.pangeo.io/t/to-zarr-compute-false-gives-zeros-not-nans/5388)

<div class="topic-metadata">

**Author:** [@Len](https://discourse.pangeo.io/u/Len)\
**Replies:** 5\
**Last updated:** [September 9, 2025, 1:12pm UTC](https://discourse.pangeo.io/t/to-zarr-compute-false-gives-zeros-not-nans/5388 "2025-09-09T13:12:36Z")

</div>

I was hoping that \_FillValue would be used if no chunks have been written yet. For example, import dask.array as da import numpy as np import xarray as xr ex = xr.Dataset({"data": xr.DataArray(da.full((10, 10), fill\_va…

---

## [Lat/Lon grid mismatch in variables from the same CMIP6 model](https://discourse.pangeo.io/t/lat-lon-grid-mismatch-in-variables-from-the-same-cmip6-model/4448)

<div class="topic-metadata">

**Author:** [@chaithra](https://discourse.pangeo.io/u/chaithra)\
**Replies:** 10\
**Last updated:** [August 28, 2025, 7:39pm UTC](https://discourse.pangeo.io/t/lat-lon-grid-mismatch-in-variables-from-the-same-cmip6-model/4448 "2025-08-28T19:39:34Z")

</div>

Hi, I’m working on some calculations involving precipitation and omega using CMIP6 models. However, I’ve encountered an issue with some models, such as ‘ACCESS-ESM1-5’, where the latitude and longitude grids for precipi…

---

## [\`.to\_zarr(..., compute=False)\` optimizations](https://discourse.pangeo.io/t/to-zarr-compute-false-optimizations/5353)

<div class="topic-metadata">

**Author:** [@Len](https://discourse.pangeo.io/u/Len)\
**Replies:** 4\
**Last updated:** [August 20, 2025, 3:52pm UTC](https://discourse.pangeo.io/t/to-zarr-compute-false-optimizations/5353 "2025-08-20T15:52:50Z")

</div>

I am looking for some tips to speed up .to\_zarr(…, compute=False) when the dataset is large and coordinates many. import dask.array as da import xarray as xr import numpy as np example = xr.Dataset( data\_vars={ …

---

## [How to efficiently merge 500 Zarr files with overlapping coordinates using Dask on S3?](https://discourse.pangeo.io/t/how-to-efficiently-merge-500-zarr-files-with-overlapping-coordinates-using-dask-on-s3/5301)

<div class="topic-metadata">

**Author:** [@BathyMapsAustralia](https://discourse.pangeo.io/u/BathyMapsAustralia)\
**Replies:** 8\
**Last updated:** [August 4, 2025, 12:20am UTC](https://discourse.pangeo.io/t/how-to-efficiently-merge-500-zarr-files-with-overlapping-coordinates-using-dask-on-s3/5301 "2025-08-04T00:20:48Z")

</div>

G’Day Pangeo Community, I’m hoping for some suggestions about how to efficiently combine approx 500 zarr files into a single DataArray, so I can then feed that DataArray into odc.geo.xr.write\_cog. I’ve tried a number of…

---

## [Parcels Design Decisions: Effectively operating with xarray for point based interpolation](https://discourse.pangeo.io/t/parcels-design-decisions-effectively-operating-with-xarray-for-point-based-interpolation/5312)

<div class="topic-metadata">

**Author:** [@NickHodgskin](https://discourse.pangeo.io/u/NickHodgskin)\
**Replies:** 0\
**Last updated:** [July 31, 2025, 9:07am UTC](https://discourse.pangeo.io/t/parcels-design-decisions-effectively-operating-with-xarray-for-point-based-interpolation/5312 "2025-07-31T09:07:13Z")

</div>

TLDR: We’d like to have insight from xarray advanced users/devs on the design decisions of our software. Problem: We want to do repeated interpolation of millions of points on top of xarray datasets. We can use vectorize…

---

## [Google Drive for Weekly Checkin Notes is full](https://discourse.pangeo.io/t/google-drive-for-weekly-checkin-notes-is-full/5303)

<div class="topic-metadata">

**Author:** [@TomAugspurger](https://discourse.pangeo.io/u/TomAugspurger)\
**Replies:** 2\
**Last updated:** [July 24, 2025, 7:55am UTC](https://discourse.pangeo.io/t/google-drive-for-weekly-checkin-notes-is-full/5303 "2025-07-24T07:55:50Z")

</div>

Currently, if you visit our check-in notes you’ll see that it can’t be edited because we’ve filled our drive Does anyone know what Google Drive that doc is in? (maybe another one for https://github.com/pangeo-data/bu…

[Next page](https://discourse.pangeo.io/c/uncategorized/1.md?page=1)
