skip to navigation
skip to content

Pattern 2.2

web mining module

Latest Version: 2.3

Pattern is a web mining module for Python 2.4+. It bundles tools for data retrieval (Google + Twitter + Wikipedia API, web spider, HTML DOM parser), text analysis (rule-based shallow parser, WordNet interface, syntactical + semantical n-gram search algorithm, tf-idf + cosine similarity + LSA metrics) and data visualization (graph networks).