# Best practices to go from 1000s of netcdf files to analyses on a HPC cluster?

**URL:** https://discourse.pangeo.io/t/best-practices-to-go-from-1000s-of-netcdf-files-to-analyses-on-a-hpc-cluster/588
**Category:** HPC
**Created:** [May 1, 2020, 9:00pm UTC](https://discourse.pangeo.io/t/best-practices-to-go-from-1000s-of-netcdf-files-to-analyses-on-a-hpc-cluster/588 "2020-05-01T21:00:04Z")
**Posts on this page:** 1
**Showing post:** 32

<div class="post-metadata">

### Author: ![rabernat](https://yyz2.discourse-cdn.com/flex030/user_avatar/discourse.pangeo.io/rabernat/32/22_2.png) [@rabernat](https://discourse.pangeo.io/u/rabernat)
#### Post date: [May 20, 2020, 3:15pm UTC](https://discourse.pangeo.io/t/best-practices-to-go-from-1000s-of-netcdf-files-to-analyses-on-a-hpc-cluster/588/32 "2020-05-20T15:15:24Z")

</div>

> [@nbren12](#):
>
> What if the GCD of the chunk sizes is 1 though?

Actually, I realized it can be a lot simpler. I think you can just take the min of `read_chunks` and `target_chunks`. Testing this now.

---

_[View the full topic](https://discourse.pangeo.io/t/best-practices-to-go-from-1000s-of-netcdf-files-to-analyses-on-a-hpc-cluster/588)._
