如何确定Pandas列是否包含特定值

I am trying to determine whether there is an entry in a Pandas column that has a particular value. I tried to do this with if x in df['id']. I thought this was working, except when I fed it a value that I knew was not in the column 43 in df['id'] it still returned True. When I subset to a data frame only containing entries matching the missing id df[df['id'] == 43] there are, obviously, no entries in it. How to I determine if a column in a Pandas data frame contains a particular value and why doesn't my current method work? (FYI, I have the same problem when I use the implementation in this answer to a similar question).

当前回答

你也可以使用pandas.Series.isin，尽管它比s.values中的'a'长一点:

In [2]: s = pd.Series(list('abc'))

In [3]: s
Out[3]: 
0    a
1    b
2    c
dtype: object

In [3]: s.isin(['a'])
Out[3]: 
0    True
1    False
2    False
dtype: bool

In [4]: s[s.isin(['a'])].empty
Out[4]: False

In [5]: s[s.isin(['z'])].empty
Out[5]: True

但是，如果您需要为一个DataFrame同时匹配多个值，这种方法可以更加灵活。

>>> df = DataFrame({'A': [1, 2, 3], 'B': [1, 4, 7]})
>>> df.isin({'A': [1, 3], 'B': [4, 7, 12]})
       A      B
0   True  False  # Note that B didn't match 1 here.
1  False   True
2   True   True

2016-11-04 18:41:03

其他回答

found = df[df['Column'].str.contains('Text_to_search')]
print(found.count())

find .count()将包含匹配数

如果它是0，那么意味着字符串没有在列中找到。

2019-02-01 07:03:07

假设你的数据框架是这样的:

现在你要检查文件名“80900026941984”是否存在于数据帧中。

你可以简单地写:

if sum(df["filename"].astype("str").str.contains("80900026941984")) > 0:
    print("found")

2020-01-21 11:19:21

我有一个CSV文件要读取:

df = pd.read_csv('50_states.csv')

在尝试之后:

if value in df.column:
    print(True)

即使值在列中，它也不会输出true;

我试着:

for values in df.column:
    if value == values:
        print(True)
        #Or do something
    else:
        print(False)

这工作。希望这能有所帮助!

2022-08-21 02:52:52

你也可以使用pandas.Series.isin，尽管它比s.values中的'a'长一点:

In [2]: s = pd.Series(list('abc'))

In [3]: s
Out[3]: 
0    a
1    b
2    c
dtype: object

In [3]: s.isin(['a'])
Out[3]: 
0    True
1    False
2    False
dtype: bool

In [4]: s[s.isin(['a'])].empty
Out[4]: False

In [5]: s[s.isin(['z'])].empty
Out[5]: True

但是，如果您需要为一个DataFrame同时匹配多个值，这种方法可以更加灵活。

>>> df = DataFrame({'A': [1, 2, 3], 'B': [1, 4, 7]})
>>> df.isin({'A': [1, 3], 'B': [4, 7, 12]})
       A      B
0   True  False  # Note that B didn't match 1 here.
1  False   True
2   True   True

2016-11-04 18:41:03

或者用级数。tolist或Series.any:

>>> s = pd.Series(list('abc'))
>>> s
0    a
1    b
2    c
dtype: object
>>> 'a' in s.tolist()
True
>>> (s=='a').any()
True

系列。tolist做了一个关于一个系列的列表，而另一个我只是从一个常规系列中获得一个布尔系列，然后检查是否有任何真布尔系列。

2018-12-05 09:43:26

如何确定Pandas列是否包含特定值

推荐文章

最新文章

标签