I have many data.frames, for example: df1 = data.frame(names=c(‘a’,’b’,’c’,’c’,’d’),data1=c(1,2,3,4,5)) df2 = data.frame(names=c(‘a’,’e’,’e’,’c’,’c’,’d’),data2=c(1,2,3,4,5,6)) df3 =

Question

0

Asked: May 31, 20262026-05-31T19:15:33+00:00 2026-05-31T19:15:33+00:00

I have many data.frames, for example: df1 = data.frame(names=c(‘a’,’b’,’c’,’c’,’d’),data1=c(1,2,3,4,5)) df2 = data.frame(names=c(‘a’,’e’,’e’,’c’,’c’,’d’),data2=c(1,2,3,4,5,6)) df3 =

0

I have many data.frames, for example:

df1 = data.frame(names=c('a','b','c','c','d'),data1=c(1,2,3,4,5))
df2 = data.frame(names=c('a','e','e','c','c','d'),data2=c(1,2,3,4,5,6))
df3 = data.frame(names=c('c','e'),data3=c(1,2))

and I need to merge these data.frames, without delete the name duplicates

> result
  names data1 data2 data3
1  'a'    1    1      NA
2  'b'    2    NA     NA
3  'c'    3    4      1
4  'c'    4    5      NA
5  'd'    5    6      NA
6  'e'    NA   2      2       
7  'e'    NA   3      NA

I cant find function like merge with option to handle with name duplicates. Thank you for your help.
To define my problem. The data comes from biological experiment where one sample have a different number of replicates. I need to merge all experiment, and I need to produce this table. I can’t generate unique identifier for replicates.

Report

Leave an answer
Cancel reply

You must login to add an answer.

Need An Account,

1 Answer

Editorial Team · Answer 1 · 2026-05-31T19:15:34+00:00

First define a function, run.seq, which provides sequence numbers for duplicates since it appears from the output that what is desired is that the ith duplicate of each name in each component of the merge be associated. Then create a list of the data frames and add a run.seq column to each component. Finally use Reduce to merge them all.

run.seq <- function(x) as.numeric(ave(paste(x), x, FUN = seq_along))

L <- list(df1, df2, df3)
L2 <- lapply(L, function(x) cbind(x, run.seq = run.seq(x$names)))

out <- Reduce(function(...) merge(..., all = TRUE), L2)[-2]

The last line gives:

> out
  names data1 data2 data3
1     a     1     1    NA
2     b     2    NA    NA
3     c     3     4     1
4     c     4     5    NA
5     d     5     6    NA
6     e    NA     2     2
7     e    NA     3    NA

EDIT: Revised run.seq so that input need not be sorted.

Sign Up

Sign In

Forgot Password

The Archive Base Latest Questions

I have many data.frames, for example: df1 = data.frame(names=c(‘a’,’b’,’c’,’c’,’d’),data1=c(1,2,3,4,5)) df2 = data.frame(names=c(‘a’,’e’,’e’,’c’,’c’,’d’),data2=c(1,2,3,4,5,6)) df3 =

Leave an answerCancel reply

1 Answer

Leave an answer
Cancel reply