繁体   English   中英

通过 cols 合并/连接两个 dataframe

[英]Merge/concat two dataframe by cols

我有两个数据框:

import pandas as pd
from numpy import nan
df1 = pd.DataFrame({'key':[1,2,3,4],
                    'only_at_df1':['a','b','c','d'],
                    'col2':['e','f','g','h'],})

df2 = pd.DataFrame({'key':[1,9],
                    'only_at_df2':[nan,'x'],
                    'col2':['e','z'],})

如何获得这个:

df3 = pd.DataFrame({'key':[1,2,3,4,9],
                    'only_at_df1':['a','b','c','d',nan],
                    'only_at_df2':[nan,nan,nan,nan,'x'],
                    'col2':['e','f','g','h','z'],})

任何帮助表示赞赏。

最好的方法可能是在临时将“key”设置为索引后使用combine_first

df1.set_index('key').combine_first(df2.set_index('key')).reset_index()

output:

   key col2 only_at_df1 only_at_df2
0    1    e           a         NaN
1    2    f           b         NaN
2    3    g           c         NaN
3    4    h           d         NaN
4    9    z         NaN           x

这似乎是mergehow="outer"的直接使用:

df1.merge(df2, how="outer")

Output:

   key only_at_df1 col2 only_at_df2
0    1           a    e         NaN
1    2           b    f         NaN
2    3           c    g         NaN
3    4           d    h         NaN
4    9         NaN    z           x

暂无
暂无

声明:本站的技术帖子网页,遵循CC BY-SA 4.0协议,如果您需要转载,请注明本站网址或者原文地址。任何问题请咨询:yoyou2525@163.com.

 
粤ICP备18138465号  © 2020-2024 STACKOOM.COM