Add `.cumulative` to `cumsum` & `cumprod` docstrings #9533

max-sixty · 2024-09-23T00:59:00Z

As discussed in a couple of issues, we should be directing folks towards the .cumulative. (the only missing piece is skip_na...

I thought this was a reasonable way to have the generation script work for these; ofc open to feedback.

I also added the namedarray file to the instructions for generating

for more information, see https://pre-commit.ci

dcherian · 2024-09-24T21:09:01Z

Note that the methods on the cumulative method are more performant and better supported

I'm surprised they're more performant than using numpy's cumprod, cumsum directly. Is this because cumulative redirects to numbagg?

max-sixty · 2024-09-24T22:00:31Z

Yes! :)

I don't have benchmarks at https://github.com/numbagg/numbagg, since cumulative().sum() just calls move_sum with the appropriate window, and numpy doesn't have a comparable function.

But to confirm — 10x faster one over 10 columns, 2x faster over 1 column:

[nav] In [36]: A = np.random.rand(60000, 10)

[ins] In [40]: %timeit np.cumsum(A, axis=0)
2.44 ms ± 82.8 µs per loop (mean ± std. dev. of 7 runs, 100 loops each)

[ins] In [47]: %timeit numbagg.move_sum(A, window=60000, min_count=0, axis=0)
371 µs ± 16.5 µs per loop (mean ± std. dev. of 7 runs, 1,000 loops each)

A = np.random.rand(60000, 1)

[nav] In [50]: %timeit np.cumsum(A, axis=0)
211 µs ± 5.34 µs per loop (mean ± std. dev. of 7 runs, 1,000 loops each)

[ins] In [49]: %timeit numbagg.move_sum(A, window=60000, min_count=0, axis=0)
106 µs ± 1.62 µs per loop (mean ± std. dev. of 7 runs, 10,000 loops each)

dcherian · 2024-09-24T22:02:40Z

we should clarify that then!

for the equivalent non-numbagg benchmark you could use xarray with use_numbagg=False?

max-sixty · 2024-09-24T22:08:06Z

we should clarify that then!

In the .cumulative docstring?

xarray/util/generate_aggregations.py

Co-authored-by: Deepak Cherian <dcherian@users.noreply.github.com>

max-sixty · 2024-10-03T01:16:19Z

Merged, also improved the docs & suggested commands in generate_aggregations.py

max-sixty and others added 7 commits September 22, 2024 16:17

Remove doctests so we can work with it

9c653fe

Add generated docstrings

4089471

Add generated doctest outputs

2064542

Update docs for running generations

5fb2794

add named array generations

Loading
Loading status checks…

80b4c6a

[pre-commit.ci] auto fixes from pre-commit.com hooks

Loading
Loading status checks…

fe1ade7

for more information, see https://pre-commit.ci

.

Loading
Loading status checks…

131e606

dcherian reviewed Sep 24, 2024

View reviewed changes

xarray/util/generate_aggregations.py Outdated Show resolved Hide resolved

max-sixty and others added 6 commits October 2, 2024 12:29

Update xarray/util/generate_aggregations.py

Loading
Loading status checks…

1f8fd91

Co-authored-by: Deepak Cherian <dcherian@users.noreply.github.com>

Loading
Loading status checks…

b89ecf3

Merge branch 'main' into see-also-cumulative

Loading
Loading status checks…

69b6b80

Loading
Loading status checks…

e300cac

1364822

Loading
Loading status checks…

53d3606

max-sixty merged commit e227c0b into pydata:main Oct 3, 2024
29 checks passed

max-sixty deleted the see-also-cumulative branch October 3, 2024 01:15

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

GitHub Sponsors

Add `.cumulative` to `cumsum` & `cumprod` docstrings #9533

Add `.cumulative` to `cumsum` & `cumprod` docstrings #9533

max-sixty commented Sep 23, 2024

dcherian commented Sep 24, 2024

max-sixty commented Sep 24, 2024 •

edited

Loading

dcherian commented Sep 24, 2024

max-sixty commented Sep 24, 2024

max-sixty commented Oct 3, 2024

Add .cumulative to cumsum & cumprod docstrings #9533

Add .cumulative to cumsum & cumprod docstrings #9533

Conversation

max-sixty commented Sep 23, 2024

dcherian commented Sep 24, 2024

max-sixty commented Sep 24, 2024 • edited Loading

dcherian commented Sep 24, 2024

max-sixty commented Sep 24, 2024

max-sixty commented Oct 3, 2024

Add `.cumulative` to `cumsum` & `cumprod` docstrings #9533

Add `.cumulative` to `cumsum` & `cumprod` docstrings #9533

max-sixty commented Sep 24, 2024 •

edited

Loading