Preply — Study more efficiently by working with a personal tutor. Get 50% off.Affiliate

Wikipedia

Natural Language Toolkit

The Natural Language Toolkit, or more commonly NLTK, is a suite of libraries and programs for symbolic and statistical natural language processing (NLP) for English written in the Python programming language. It supports classification, tokenization, stemming, tagging, parsing, and semantic reasoning functionalities. It was developed by Steven Bird and Edward Loper in the Department of Computer and Information Science at the University of Pennsylvania. NLTK includes graphical demonstrations and sample data. It is accompanied by a book that explains the underlying concepts behind the language processing tasks supported by the toolkit, plus a cookbook. NLTK is intended to support research and teaching in NLP or closely related areas, including empirical linguistics, cognitive science, artificial intelligence, information retrieval, and machine learning. NLTK has been used successfully as a teaching tool, as an individual study tool, and as a platform for prototyping and building research systems.

Library highlights Discourse representation Lexical analysis: Word and text tokenizer n-gram and collocations Part-of-speech tagger Tree model and Text chunker for capturing Named-entity recognition

See also

Gensim SpaCy

References

External links Official website

Tags

  • Data analysis software
  • Free linguistic software
  • Free science software
  • Free software programmed in Python
  • Natural language parsing
  • Natural language processing
  • Natural language processing stubs
  • Natural language processing toolkits
  • Programming language topic stubs
  • Python (programming language) libraries
  • Software using the Apache license
  • Statistical natural language processing