Part of Speech Tagging
A Part of Speech tagger using the Average Perceptron.
Based on the tagger from here
This uses the following features:
- The Suffix (last 3 characters) of the current word (unnormalized).
- The Prefix (first character) of the current word (unnormalized).
- The current word.
- The previous Part of Speech tag and the current word.
- The Previous Part of Speech tag.
- The Part of Speech tag from the word before last.
- Both of the previous Part of Speech tags.
- The previous word.
- The previous word suffix.
- The word from 2 steps back.
- The next word.
- The next word suffix.
- The word after next.
- A Bias
Includes the following Pretrained models.
- POS Tagger, Trained on the CoNLL 2000 Chunking data
- Chunker, Trained on the CoNLL 2000 Chunking data
- Slot filler, Trained on ATIS data
Metadata
Release files for sequence-tagging 0.1.6
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| sequence_tagging-0.1.6.tar.gz | 2.9 MB | Details |
Release files / sequence_tagging-0.1.6.tar.gz
| Download URL | sequence_tagging-0.1.6.tar.gz |
|---|---|
| Size | 2.9 MB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
daf080cca294ae075d6f02e214019e1d242a7e9275be31bd9634dcde04675adf
|
|
BLAKE2b-256 checksum How to use checksums |
8f281eb9cd5d5ea6dcb8dbfcc58df3dc7b3333a3db16da2ee18717aaf876db6d
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |