如何更改DataFrame列的顺序？

我有以下DataFrame（df）：

import numpy as np
import pandas as pd

df = pd.DataFrame(np.random.rand(10, 5))

我通过分配添加更多列：

df['mean'] = df.mean(1)

如何将列的意思移到前面，即将其设置为第一列，而其他列的顺序保持不变？

当前回答

熊猫>=1.3（2022年编辑）：

df.insert(0, 'mean', df.pop('mean'))

怎么样（对于熊猫<1.3，原始答案）

df.insert(0, 'mean', df['mean'])

https://pandas.pydata.org/pandas-docs/stable/user_guide/dsintro.html#column-选择添加删除

2012-11-09 21:04:03

其他回答

这里有一个函数可以对任意数量的列执行此操作。

def mean_first(df):
    ncols = df.shape[1]        # Get the number of columns
    index = list(range(ncols)) # Create an index to reorder the columns
    index.insert(0,ncols)      # This puts the last column at the front
    return(df.assign(mean=df.mean(1)).iloc[:,index]) # new df with last column (mean) first

2018-01-29 18:57:18

只需键入要更改的列名，然后为新位置设置索引。

def change_column_order(df, col_name, index):
    cols = df.columns.tolist()
    cols.remove(col_name)
    cols.insert(index, col_name)
    return df[cols]

对于您的情况，这将是：

df = change_column_order(df, 'mean', 0)

2016-05-06 11:39:33

另一种选择是使用set_index（）方法，后跟reset_index（）。请注意，我们首先pop（）将要移动到数据帧前面的列，以便在重置索引时避免名称冲突：

df.set_index(df.pop('column_name'), inplace=True)
df.reset_index(inplace=True)

有关详细信息，请参阅How to change the order of dataframe columns in panda。

2021-08-15 22:41:00

只需按所需顺序分配列名：

In [39]: df
Out[39]: 
          0         1         2         3         4  mean
0  0.172742  0.915661  0.043387  0.712833  0.190717     1
1  0.128186  0.424771  0.590779  0.771080  0.617472     1
2  0.125709  0.085894  0.989798  0.829491  0.155563     1
3  0.742578  0.104061  0.299708  0.616751  0.951802     1
4  0.721118  0.528156  0.421360  0.105886  0.322311     1
5  0.900878  0.082047  0.224656  0.195162  0.736652     1
6  0.897832  0.558108  0.318016  0.586563  0.507564     1
7  0.027178  0.375183  0.930248  0.921786  0.337060     1
8  0.763028  0.182905  0.931756  0.110675  0.423398     1
9  0.848996  0.310562  0.140873  0.304561  0.417808     1

In [40]: df = df[['mean', 4,3,2,1]]

现在，“mean”列出现在前面：

In [41]: df
Out[41]: 
   mean         4         3         2         1
0     1  0.190717  0.712833  0.043387  0.915661
1     1  0.617472  0.771080  0.590779  0.424771
2     1  0.155563  0.829491  0.989798  0.085894
3     1  0.951802  0.616751  0.299708  0.104061
4     1  0.322311  0.105886  0.421360  0.528156
5     1  0.736652  0.195162  0.224656  0.082047
6     1  0.507564  0.586563  0.318016  0.558108
7     1  0.337060  0.921786  0.930248  0.375183
8     1  0.423398  0.110675  0.931756  0.182905
9     1  0.417808  0.304561  0.140873  0.310562

2015-04-28 14:19:49

这个问题以前已经回答过，但reindex_axis现在已被弃用，因此我建议使用：

df = df.reindex(sorted(df.columns), axis=1)

对于那些想要指定他们想要的顺序而不是仅仅对它们进行排序的人来说，下面列出了解决方案：

df = df.reindex(['the','order','you','want'], axis=1)

现在，如何对列名列表排序真的不是熊猫问题，而是Python列表操作问题。有很多方法可以做到这一点，我认为这个答案有一个非常简洁的方法。

2013-01-04 06:04:46

如何更改DataFrame列的顺序？

推荐文章

最新文章

标签