R: cumulative amount based on another column and maximum amount

Question

R: cumulative amount based on another column and maximum amount

This is similar to the previous post for the total amount that is reset based on the value in another column, except that I want to limit the amount so that it is also reset when it reaches the maximum value. For example, if the maximum value is 3:

> data.frame(x=rep(1,10), 
+ y=c(0,0,1,0,0,0,0,1,0,0), 
+ cum_sum_mod=c(1, 2, 1, 2, 3, 1, 2, 1, 2, 3))

   x y cum_sum_mod
1  1 0           1
2  1 0           2
3  1 1           1
4  1 0           2
5  1 0           3
6  1 0           1
7  1 0           2
8  1 1           1
9  1 0           2
10 1 0           3

cum_sum_mod sums column x until it reaches its maximum value (3), or the value in column y is 1. I want to avoid using a loop.

+4

r

Vlad 15 sept. '17 at 14:27

source share

2 answers

Base R

ave(df$x, cumsum(df$y == 1), FUN = function(x){
    temp = cumsum(x)
    replace(temp, temp > 3, rep(1:3, length.out = sum(temp > 3)))
})
# [1] 1 2 1 2 3 1 2 1 2 3

+3

db 15 sept. '17 at 14:50

source share

Wen · Accepted Answer · 2017-09-15T14:43:05+0000

Using dplyr

 library(dplyr)

 dat=data.frame(x=rep(1,10), 
             y=c(0,0,1,0,0,0,0,1,0,0))
 dat$B=cumsum(dat$y)
 dat%>%group_by(B)%>%mutate(cum_sum_mod=ifelse(cumsum(x)%%3==0,3,cumsum(x)%%3))

# A tibble: 10 x 4
# Groups:   B [3]
       x     y     B cum_sum_mod
   <dbl> <dbl> <dbl>       <dbl>
 1     1     0     0           1
 2     1     0     0           2
 3     1     1     1           1
 4     1     0     1           2
 5     1     0     1           3
 6     1     0     1           1
 7     1     0     1           2
 8     1     1     2           1
 9     1     0     2           2
10     1     0     2           3

R: cumulative amount based on another column and maximum amount

More articles: