I am trying to upsample within a grouped DataFrame but am unsure how to get it to only upsample within the groups. I have a DataFrame that looks like:
cat weekstart date
0.0 2016-07-04 00:00:00+00:00 2016-07-04 1
2016-07-06 1
2016-07-07 2
2016-08-15 00:00:00+00:00 2016-08-16 1
2016-08-19 1
2016-09-19 00:00:00+00:00 2016-09-20 1
2016-09-21 1
2016-12-19 00:00:00+00:00 2016-12-19 1
2016-12-21 1
1.0 2016-07-25 00:00:00+00:00 2016-07-26 2
2016-08-01 00:00:00+00:00 2016-08-03 1
2016-08-08 00:00:00+00:00 2016-08-12 1
If I do something like df.unstack().fillna(0).stack() leads to:
cat weekstart date
0.0 2016-07-04 00:00:00+00:00 2016-1-1 0
.
.
.
2016-07-04 1
2016-07-06 1
2016-07-07 2
because the minimum in the date column is 2016-1-1. What i'm after though is only sampling business days within each 'cat' and 'weekstart', like:
cat weekstart date
0.0 2016-07-04 00:00:00+00:00 2016-07-04 1
2016-07-05 0
2016-07-06 1
2016-07-07 2
2016-07-8 0
2016-08-15 00:00:00+00:00 2016-08-15 0
2016-08-16 1
2016-08-17 0
2016-08-18 0
2016-08-19 1
I've tried using:
level_values = df.index.get_level_values
df.groupby(
[level_values(i) for i in [0, 1]] + [pd.Grouper('B', level=-1)]
)
.sum()
but it isn't working as expected.