Python正則表達式匹配一段英文中包含關鍵字的句子

-Advertisement-

簡單又高大上的項目圖形識別、自然語言處理(語言識別、語音轉文字)、文字識別、區塊鏈 1.java實現一個基本的文字識別引入依賴  <dependency> <groupId>com.baidu.aip</groupId> <artifactId>java-sdk< ...

1.問題/需求

在含有多行文字的英文段落或一篇英文中查找匹配含有關鍵字的句子。

例如在以下字元串：

text = '''Today I registered my personal blog in the cnblogs and wrote my first essay. 
The main problem of this essay is to use python regular expression matching to filter out 
sentences containing keywords in the paper. To solve this problem, I made many attempts 
and finally found a regular expression matching method that could meet the requirements 
through testing. So I've documented the problem and the solution in this blog post and 
shared it for reference to others who are having the same problem. At the same time, 
this text is also used to test the feasibility of this matching approach. Some additional 
related thoughts and ideas will be added to this blog later.'''

中匹配含有’blog‘的句子。

2.解決方法

因為要找出所有含有關鍵字的句子，所以這裡採用re庫中findall()方法。同時，由於所搜索的字元串中含有換行符'\n'，因此向re.compilel()傳入re.DOTALL參數，以使'.'字元能夠匹配所有字元，包括換行符'\n'。這樣我們匹配創建Pattern對象為：

newre = re.compile('[A-Z][^.]*blog[^.]*[.]', re.DOTALL)
newre.findall(text)  # 進行匹配
# 結果為：
['Today I registered my personal blog in the cnblogs and wrote my first essay.',
"So I've documented the problem and the solution in this blog post and \nshared it for reference to others who are having the same problem.",
'Some additional \nrelated thoughts and ideas will be added to this blog later.']  # 這其中的'\n'就是換行符, 它在字元串中是不顯示的, 但是匹配結果中又顯示出來了

您的分享是我們最大的動力!

-Advertisement-

更多相關文章

隨機高併發查詢結果一致性設計實踐

物流合約中心是京東物流合同管理的唯一入口。為商家提供合同的創建，蓋章等能力，為不同業務條線提供合同的定製，歸檔，查詢等功能。由於各個業務條線眾多，為各個業務條線提供高可用查詢能力是物流合約中心重中之重。同時計費系統在每個物流單結算時，都需要查詢合約中心，確保商家簽署的合同內容來保證計費的準確性。 ...
風控核心子域——名單服務構建及挑戰

名單服務是風控架構中重要子域，對風險決策的性能、用戶體驗、成本管控、風險治理沉澱都有重要影響，本文將詳細介紹名單服務設計思路和實現。 ...
全球首個面向遙感任務設計的億級視覺Transformer大模型

深度學習在很大程度上影響了遙感影像分析領域的研究。然而，大多數現有的遙感深度模型都是用ImageNet預訓練權重初始化的，其中自然圖像不可避免地與航拍圖像相比存在較大的域差距，這可能會限制下游遙感場景任務上的微調性能。 ...
SpringBoot學習筆記 - 構建、簡化原理、快速啟動、配置文件與多環境配置、技術整合案例

【前置內容】Spring 學習筆記全系列傳送門： Spring學習筆記 - 第一章 - IoC（控制反轉）、IoC容器、Bean的實例化與生命周期、DI（依賴註入） Spring學習筆記 - 第二章 - 註解開發、配置管理第三方Bean、註解管理第三方Bean、Spring 整合 MyBatis 和 ...
讓Apache Beam在GCP Cloud Dataflow上跑起來

簡介在文章《Apache Beam入門及Java SDK開發初體驗》中大概講了Apapche Beam的簡單概念和本地運行，本文將講解如何把代碼運行在GCP Cloud Dataflow上。本地運行通過maven命令來創建項目： mvn archetype:generate \ -Darche ...
day16-聲明式事務-02

聲明式事務-02 3.事務的傳播機制事務的傳播機制說明：當有多個事務處理並存時，如何控制？比如用戶去購買兩次商品（使用不同的方法），每個方法都是一個事務，那麼如何控制呢？也就是說，某個方法本身是一個事務，然後該方法中又調用了其他一些方法，這些方法也是被@Transactional 修飾的，同 ...
QPython實例02-調用其他app實例

一、前言使用版本：QPython 3c 下載地址：百度搜索QPython 3C開源版即可下載或關註【產品經理不是經理】gzh，回覆【qpython 3c】即可獲取下載鏈接。二、代碼實例註意 # 執行以下方法前，請加上以下代碼 from androidhelper import Android ...
《RPC實戰與核心原理》學習筆記Day15

這篇文章主要關註流量回放和動態分組，主要包括流量回放的使用背景，RPC中流量回放的實現方式，動態分組要解決的問題以及如何實現動態分組。 ...