textacy

NLP, before and after spaCy

These details have not been verified by PyPI

Project links

Development Status
- 4 - Beta
Intended Audience
- Developers
- Science/Research
License
- OSI Approved :: Apache Software License
Natural Language
- English
Programming Language
Topic
- Text Processing :: Linguistic

Project description

textacy: NLP, before and after spaCy

textacy is a Python library for performing a variety of natural language processing (NLP) tasks, built on the high-performance spaCy library. With the fundamentals --- tokenization, part-of-speech tagging, dependency parsing, etc. --- delegated to another library, textacy focuses primarily on the tasks that come before and follow after.

features

Access and extend spaCy's core functionality for working with one or many documents through convenient methods and custom extensions
Load prepared datasets with both text content and metadata, from Congressional speeches to historical literature to Reddit comments
Clean, normalize, and explore raw text before processing it with spaCy
Extract structured information from processed documents, including n-grams, entities, acronyms, keyterms, and SVO triples
Compare strings and sequences using a variety of similarity metrics
Tokenize and vectorize documents then train, interpret, and visualize topic models
Compute text readability and lexical diversity statistics, including Flesch-Kincaid grade level, multilingual Flesch Reading Ease, and Type-Token Ratio

... and much more!

maintainer

Howdy, y'all. 👋

Burton DeWilde (burtdewilde@gmail.com)

Algorithm	Hash digest
SHA256	`6be02448c08fc7d7c4edf85289006e39a4a53ef747201ff24b675c652f40c686`
MD5	`54f049988924accaba14c18c268b0c34`
BLAKE2b-256	`04fe4a578d9f68e7aaf6b7be7d8df974ab3b1b21f2e64d492919adda3cd80b71`

Algorithm	Hash digest
SHA256	`0e150ce52c8366ccd26650ac310478bbe19604a16fd35a97659973f9d172573c`
MD5	`5e1b916d0c77659484bdefc00c72c8f1`
BLAKE2b-256	`8092a3593873fbd531f8430c4a2958611280dd33ace14ead14a6c43e61675e55`

textacy 0.13.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

textacy: NLP, before and after spaCy

features

links

maintainer

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes