data.table. The speed and simplicity of this R package are astonishing.
Here is a simple example: I have a data frame showing incremental claims development by line of business and origin year. Now I would like add a column with the cumulative claims position for each line of business and each origin year along the development years.
It's one line with
data.table! Here it is:
It is even easy to read! Notice also that I don't have to copy the data. The operator
myData[order(dev), cvalue:=cumsum(value), by=list(origin, lob)]
':='works by reference and is one of the reasons why
data.tableis so fast.
And it is getting even better. Suppose you want to get the latest claims development position for each line of business and origin year. Again, it is only one line: