Skip to content

Support duplicate dimensions in .chunk - #9099

Merged
dcherian merged 7 commits into
pydata:mainfrom
mraspaud:fix-duplicate-dimensions
Jun 17, 2024
Merged

Support duplicate dimensions in .chunk#9099
dcherian merged 7 commits into
pydata:mainfrom
mraspaud:fix-duplicate-dimensions

Conversation

@mraspaud

@mraspaudmraspaud commented Jun 12, 2024

Copy link
Copy Markdown
Contributor

This PR allows duplicate dimension when chunking an array, when the chunk sizes is provided as a dict.
A typical example of the usefulness of this PR is trying to open a netcdf file (with chunking) containing a covariance matrix.

@mraspaud

Copy link
Copy Markdown
ContributorAuthor

Any feedback welcome on how I can improve this PR!

Comment threadxarray/tests/test_dask.py Outdated
Comment threadxarray/namedarray/core.py Outdated
* main:
new whats-new section (pydata#9115)
release v2024.06.0 (pydata#9113)
release notes for 2024.06.0 (pydata#9092)
[skip-ci] Try fixing hypothesis CI trigger (pydata#9112)
Undo custom padding-top. (pydata#9107)
add remaining core-dev citations [skip-ci][skip-rtd] (pydata#9110)
Add user survey announcement to docs (pydata#9101)
skip the `pandas` datetime roundtrip test with `pandas=3.0` (pydata#9104)
Adds Matt Savoie to CITATION.cff (pydata#9103)
[skip-ci] Fix skip-ci for hypothesis (pydata#9102)
open_datatree performance improvement on NetCDF, H5, and Zarr files (pydata#9014)

@dcheriandcherian left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Thanks @mraspaud

@dcheriandcherian added the plan to merge Final call for comments label Jun 13, 2024
@dcheriandcherian changed the title Allow duplicate dimensions in chunkingSupport duplicate dimensions in .chunkJun 13, 2024
Comment threadxarray/tests/test_dask.py
@dcherian
dcherian merged commit be8e17e into pydata:mainJun 17, 2024
@mraspaud
mraspaud deleted the fix-duplicate-dimensions branch June 18, 2024 12:20
dcherian added a commit to dcherian/xarray that referenced this pull request Jun 21, 2024
* main:
Split out distributed writes in zarr docs (pydata#9132)
Update zendoo badge link (pydata#9133)
Support duplicate dimensions in `.chunk` (pydata#9099)
Bump the actions group with 2 updates (pydata#9130)
adjust repr tests to account for different platforms (pydata#9127) (pydata#9128)
dcherian added a commit that referenced this pull request Jul 24, 2024
* main: (48 commits)
Add test for #9155 (#9161)
Remove mypy exclusions for a couple more libraries (#9160)
Include numbagg in type checks (#9159)
Improve zarr chunks docs (#9140)
groupby: remove some internal use of IndexVariable (#9123)
Improve `to_zarr` docs (#9139)
Split out distributed writes in zarr docs (#9132)
Update zendoo badge link (#9133)
Support duplicate dimensions in `.chunk` (#9099)
Bump the actions group with 2 updates (#9130)
adjust repr tests to account for different platforms (#9127) (#9128)
Grouper refactor (#9122)
Update docstring in api.py for open_mfdataset(), clarifying "chunks" argument (#9121)
Add test for rechunking to a size string (#9117)
Move Sphinx directives out of `See also` (#8466)
new whats-new section (#9115)
release v2024.06.0 (#9113)
release notes for 2024.06.0 (#9092)
[skip-ci] Try fixing hypothesis CI trigger (#9112)
Undo custom padding-top. (#9107)
...
wavebyrd pushed a commit to wavebyrd/xarray that referenced this pull request Mar 13, 2026
* main: (48 commits)
Add test for pydata#9155 (pydata#9161)
Remove mypy exclusions for a couple more libraries (pydata#9160)
Include numbagg in type checks (pydata#9159)
Improve zarr chunks docs (pydata#9140)
groupby: remove some internal use of IndexVariable (pydata#9123)
Improve `to_zarr` docs (pydata#9139)
Split out distributed writes in zarr docs (pydata#9132)
Update zendoo badge link (pydata#9133)
Support duplicate dimensions in `.chunk` (pydata#9099)
Bump the actions group with 2 updates (pydata#9130)
adjust repr tests to account for different platforms (pydata#9127) (pydata#9128)
Grouper refactor (pydata#9122)
Update docstring in api.py for open_mfdataset(), clarifying "chunks" argument (pydata#9121)
Add test for rechunking to a size string (pydata#9117)
Move Sphinx directives out of `See also` (pydata#8466)
new whats-new section (pydata#9115)
release v2024.06.0 (pydata#9113)
release notes for 2024.06.0 (pydata#9092)
[skip-ci] Try fixing hypothesis CI trigger (pydata#9112)
Undo custom padding-top. (pydata#9107)
...
Sign up for freeto join this conversation on GitHub. Already have an account? Sign in to comment

Labels

plan to mergeFinal call for comments

Projects

None yet

Development

Successfully merging this pull request may close these issues.

Allow .chunk for datasets with duplicated dimension names, e.g. Sentinel-3 OLCI files

2 participants

@mraspaud@dcherian