【发布时间】:2019-05-26 21:55:56
【问题描述】:
鉴于下面的代码,来自this question 的接受答案:
import re
pathD = "M30,50.1c0,0,25,100,42,75s10.3-63.2,36.1-44.5s33.5,48.9,33.5,48.9l24.5-26.3"
print(re.findall(r'[A-Za-z]|-?\d+\.\d+|\d+',pathD))
['M', '30', '50.1', 'c', '0', '0', '25', '100', '42', '75', 's', '10.3', '-63.2', '36.1', '-44.5', 's', '33.5', '48.9', '33.5', '48.9', 'l', '24.5', '-26.3']
如果我在 pathD 变量中包含诸如“$”或“£”之类的符号,re 表达式将跳过它们,因为它以 [A-Za-z] 和数字为目标
[A-Za-z] # words
|
-?\d+\.\d+ # floating point numbers
|
\d+ # integers
我如何修改上面的正则表达式模式以同时保留非字母数字符号,根据下面的所需输出?
new_pathD = '$100.0thousand'
new_re_expression = ???
print(re.findall(new_re_expression, new_pathD))
['$', '100.0', 'thousand']
~~~
下面的相关 SO 帖子,尽管我无法完全找到如何在拆分练习中保留符号:
Split string into letters and numbers
split character data into numbers and letters
Python regular expression split string into numbers and text/symbols
Python - Splitting numbers and letters into sub-strings with regular expression
【问题讨论】: