我們如何在Python的正規表示式中找到每個匹配的精確位置？-Python教學-PHP中文網

我們如何在Python的正規表示式中找到每個匹配的精確位置？

王林

發布： 2023-08-31 12:14:13

轉載

715 人瀏覽過

我們如何在Python的正規表示式中找到每個匹配的精確位置？

介紹

re模組是我們在Python中使用的正規表示式。文字搜尋和更複雜的文字操作都使用正規表示式。像grep和sed這樣的工具，像vi和emacs這樣的文字編輯器，以及像Tcl、Perl和Python這樣的電腦語言都內建了正規表示式支援。

Python中的re模組提供了用於匹配正規表示式的函數。

定義我們要尋找或修改的文字的正規表示式稱為模式。文字字面量和元字元構成了這個字串。編譯函數用於建立模式。建議使用原始字串，因為正規表示式經常包含特殊字元。（r字元用於指示原始字串。）這些字元在組合成模式之前不會被解釋。

可以使用其中一個函數將模式應用於文字字串，模式在組裝完成後使用。可用的函式包括Match、Search、Find和Finditer。

使用的語法

在這裡使用的正規表示式函數是：我們使用正規表示式函數來尋找匹配項。

re.match(): Determines if the RE matches at the beginning of the string. If zero or more characters at the beginning of the string match the regular expression pattern, the match method returns a match object.

p.finditer(): Finds all substrings where the RE matches and returns them as an iterator. An iterator delivering match objects across all non-overlapping matches for the pattern in a string is the result of the finditer method.

re.compile(): Compile a regular expression pattern into a regular expression object, which can be used for matching using its match(), search(), and other methods described below. The expression’s behavior can be modified by specifying a flag&#39;s value. Values can be any of the following variables combined using bitwise OR (the | operator).

m.start(): m.start() returns the offset in the string at the match&#39;s start.

m.group(): You may use the multiple-assignment approach to assign each value to a different variable when mo.groups() returns a tuple of values, as in the areaCode, mainNumber = mo.groups() line below.

search: It is comparable to re.match() but does not require that we just look for matches at the beginning of the text. The search() function can locate a pattern in the string at any location, but it only returns the first instance of the pattern.

登入後複製

演算法

使用import re導入正規表示式模組。
使用re.compile()函數建立一個正規表示式物件。（記得使用原始字串。）
將要搜尋的字串傳遞給Regex物件的finditer()方法。這將會傳回一個Match物件。
呼叫Match物件的group()方法傳回實際符合的文字字串。
我們也可以使用span()方法在一個元組中取得起始和結束索引。

範例

 #importing re functions
import re
#compiling [A-Z0-9] and storing it in a variable p
p = re.compile("[A-Z0-9]")
#looping m times in p.finditer
for m in p.finditer(&#39;A5B6C7D8&#39;):
#printing the m.start and m.group
   print m.start(), m.group()

登入後複製

輸出

這將產生輸出−

登入後複製

程式碼解釋

使用import re導入正規表示式模組。使用re.compile()函數建立一個正規表示式物件（“[A-Z0-9]”）並將其賦值給變數p。使用迴圈遍歷m，並將要搜尋的字串傳遞給正規表示式物件的finditer()方法。這將會傳回一個Match物件。呼叫Match物件的m.group()和m.start()方法以傳回實際匹配文字的字串。

範例

# Python program to illustrate
# Matching regex objects
# with groups
import re
phoneNumRegex = re.compile(r&#39;(\d\d\d)-(\d\d\d-\d\d\d\d)&#39;)
mo = phoneNumRegex.search(&#39;My number is 415-555-4242.&#39;)
print(mo.groups())

登入後複製

輸出

這將產生輸出−

(&#39;415&#39;, &#39;555-4242&#39;)

登入後複製

程式碼解釋

使用import re導入正規表示式模組。使用re.compile()函數建立一個正規表示式物件(r'(\d\d\d)-(\d\d\d-\d\d\d\d)')，並將其賦值給變數phoneNumRegex。將要搜尋的字串傳遞給Regex物件的search()方法，並將其儲存在變數mo中。這將會傳回一個Match物件。呼叫Match物件的mo.groups()方法以傳回實際匹配的文字字串。