本文介绍了如何将段落分为句子的处理方法,对大家解决问题具有一定的参考价值,需要的朋友们下面随着小编来一起学习吧!
问题描述
我一直在尝试使用:
$string="The Dr. is here!!! I am glad I'm in the U.S.A. for the Dr. quality is great!!!!!!";
preg_match_all('~.*?[?.!]~s',$string,$sentences);
print_r($sentences);
但这不适用于美国博士等.
But it doesn't work on Dr., U.S.A., etc.
有人有更好的建议吗?
Does anyone have any better suggestions?
推荐答案
没有任何简单的解决方案.您需要在应用程序中进行一些自然语言处理(NLP)并识别每个句子.有一个叫 OpenNLP 的东西,它是一个基于Java的NLP解析器工具.或在Ruby中 Stanford NLP解析器.您可以为php找到类似的内容.
there is not any simple solution for that. you need do some natural language processing(NLP) in your application and recognize each sentence. there is something call OpenNLP, it's a JAVA-based NLP parser tool. Or Stanford NLP parser in Ruby. you can find something like that for php.
此处我找到了一组用于PHP中自然语言处理的类.
here I found a set of classes for natural language processing in PHP.
这篇关于如何将段落分为句子的文章就介绍到这了,希望我们推荐的答案对大家有所帮助,也希望大家多多支持!