Running count based on field in R

I have a data set of this format

Now I want to add a column saying count which counts the occurrence of the user. I want output in the below format.

User    Count
1       1
2       1 
3       1
2       2
3       2
1       2
1       3

I have few solutions but all those solutions are somewhat slow.

Running count variable in R

My data.frame has 100,000 rows now and soon it may go up to 1 million. I need a solution which is also fast.

标签： r running-total

3条回答

闹够了就滚

2楼-- · 2019-04-25 12:12

This is fairly easy with ave and seq.int:

> ave(User,User, FUN= seq.int)
[1] 1 1 1 2 2 2 3

This is a common strategy and is often used when the items are adjacent to each other. The second argument is the grouping variable and in this case the first argument is really kind of a dummy argument since the only thing that it contributes is a length, and it is not a requirement for ave to have adjacent rows for the values determined within groupings.

0人赞添加讨论(0) 举报

放荡不羁爱自由

3楼-- · 2019-04-25 12:18

You can use getanID from my "splitstackshape" package:

library(splitstackshape)
getanID(mydf, "User")
##    User .id
## 1:    1   1
## 2:    2   1
## 3:    3   1
## 4:    2   2
## 5:    3   2
## 6:    1   2
## 7:    1   3

This is essentially an approach with "data.table" that looks something like the following:

as.data.table(mydf)[, count := seq(.N), by = "User"][]

0人赞添加讨论(0) 举报

成全新的幸福

4楼-- · 2019-04-25 12:19

An option using dplyr

 library(dplyr)
 df1 %>%
      group_by(User) %>%
      mutate(Count=row_number())
 #    User Count
 #1    1     1
 #2    2     1
 #3    3     1
 #4    2     2
 #5    3     2
 #6    1     2
 #7    1     3

Using sqldf

library(sqldf)
sqldf('select a.*, 
           count(*) as Count
           from df1 a, df1 b
           where a.User = b.User and b.rowid <= a.rowid
           group by a.rowid')
#   User Count
#1    1     1
#2    2     1
#3    3     1
#4    2     2
#5    3     2
#6    1     2
#7    1     3

0人赞添加讨论(0) 举报

Running count based on field in R

采纳回答

编辑标签

举报内容

检举类型

检举原因

检举说明(必填)

打开微信“扫一扫”，打开网页后点击屏幕右上角分享按钮

付费偷看金额在0.1-10元之间