Path: csiph.com!au2pb.net!usenet.blueworldhosting.com!feeder01.blueworldhosting.com!border2.nntp.dca1.giganews.com!nntp.giganews.com!news.iecc.com!.POSTED!nerds-end From: Seima Rao Newsgroups: comp.compilers Subject: Natural Language Parser Date: Tue, 29 Sep 2015 06:15:22 +0530 Organization: Compilers Central Lines: 44 Sender: news@iecc.com Approved: comp.compilers@iecc.com Message-ID: <15-09-025@comp.compilers> NNTP-Posting-Host: news.iecc.com Mime-Version: 1.0 Content-Type: text/plain; charset=UTF-8 X-Trace: miucha.iecc.com 1443541934 67153 2001:470:1f07:1126:0:676f:7373:6970 (29 Sep 2015 15:52:14 GMT) X-Complaints-To: abuse@iecc.com NNTP-Posting-Date: Tue, 29 Sep 2015 15:52:14 +0000 (UTC) Keywords: parse, question Posted-Date: 29 Sep 2015 11:52:14 EDT X-submission-address: compilers@iecc.com X-moderator-address: compilers-request@iecc.com X-FAQ-and-archives: http://compilers.iecc.com Xref: csiph.com comp.compilers:1620 Hi, I am looking for a C++ API based English Language Parser. The specific task I want to do on a *regular basis* is to parse English Language Documents and arrive at a Mathematical artefact. The Mathematical artefact will (have) built a Syntax Tree of the English text that is input to the NLP parser. This is my only requirement. I dont need any semanticizing artefacts. So, my requirement is limited to parsing english language documents and arriving at a tree( or any other mathematical structure that binds the English grammar to the input document aka syntax trees). My guess is that the specific C++ API based NL parser will be using some english dictionary inside the tool to do all the jobs that are advertised. Can readers of this forum direct me to a stable and active C++ API based NL parser ? I have zero experience in natural language parsing(compiling) and zero experience in using such tools. However, I intend to maintain internally the source code of the toolkit via whatever version control software that is used by the developers of the tool so that I am able to get regular updates and not break anything. Sincerely, Seima Rao. [You might start with this parser from Stanford: http://nlp.stanford.edu/software/lex-parser.shtml Or this one in python: http://spacy.io/ Parsing English or any natural language is very hard, and you'll never get more than an approximate result. Modern language translation systems don't even try and use machine learning on large corpora. -John]