I have 2 datasets (lists/columns of gene names) such as:
df1
Gene_id
SUMO2
CDC37
COPB2
BECN1
CAPNS1
and
df2
Gene_id
SUMO2
BECN1
CAPNS1
I want to make a new dataset that has 2 columns with the gene names matched up. 1 with all of df1 genes and the 2nd with all of df2 genes matched up in column 1. And NA's where column 2 does not have a match which would look like below. Preferably with dplyr in R or Python. Thanks
Gene_id Gene_id
SUMO2 SUMO2
CDC37 NA
COPB2 NA
BECN1 BECN1
CAPNS1 CAPNS1