从字符串中删除所有特殊字符、标点符号和空格

我需要从字符串中删除所有特殊字符，标点符号和空格，以便我只有字母和数字。

当前回答

假设你想要使用正则表达式并且你想要/需要unicode - cognant 2。X代码是2to3-ready:

>>> import re
>>> rx = re.compile(u'[\W_]+', re.UNICODE)
>>> data = u''.join(unichr(i) for i in range(256))
>>> rx.sub(u'', data)
u'0123456789ABCDEFGHIJKLMNOPQRSTUVWXYZabcdefghijklmnopqrstuvwxyz\xaa\xb2 [snip] \xfe\xff'
>>>

2011-04-30 21:07:48

其他回答

这可以不使用regex完成:

>>> string = "Special $#! characters   spaces 888323"
>>> ''.join(e for e in string if e.isalnum())
'Specialcharactersspaces888323'

你可以使用str.isalnum:

S.isalnum() -> bool 如果S中的所有字符都是字母数字，则返回True 且S中至少有一个字符，否则为假。

如果坚持使用正则表达式，其他解决方案也可以。但是请注意，如果可以在不使用正则表达式的情况下完成，那么这是最好的方法。

2011-04-30 17:47:39

下面是一个正则表达式，用于匹配不是字母或数字的字符串:

[^A-Za-z0-9]+

下面是执行正则表达式替换的Python命令:

re.sub('[^A-Za-z0-9]+', '', mystring)

2011-04-30 17:46:17

import re
my_string = """Strings are amongst the most popular data types in Python. We can create the strings by enclosing characters in quotes. Python treats single quotes the

和双引号一样。”＂＂

# if we need to count the word python that ends with or without ',' or '.' at end

count = 0
for i in text:
    if i.endswith("."):
        text[count] = re.sub("^([a-z]+)(.)?$", r"\1", i)
    count += 1
print("The count of Python : ", text.count("python"))

2018-07-16 11:52:40

s = re.sub(r"[-()\"#/@;:<>{}`+=~|.!?,]", "", s)

2018-06-15 12:09:44

import re
abc = "askhnl#$%askdjalsdk"
ddd = abc.replace("#$%","")
print (ddd)

你会看到你的结果是

'Askhnlaskdjalsdk

2016-02-25 08:00:02

从字符串中删除所有特殊字符、标点符号和空格

推荐文章

最新文章

标签