Skip to main content

Thai Word Segmentation using TCC + Bidirectional RNNs

Project description


Thai Word Segmentation using TCC + Bidirectional RNNs

Credit code from A Beginner's Guide to Deep NLP with PyTorch - Dr. Prachya Boonkwan

Colab Notebook :

Train by BEST I Corpus Training set. (90% training , 10% test)

ep 6
loss: 0.017879242024514966
f1 : 98.47012481095481

F1 From BEST I Corpus Test set

F-measure: 96.94929
Recall: 122271.00000/125850.00000 = 97.15614

Precision: 122271.00000/126387.00000 = 96.74333

Number of incorrect : 3579.00000 words

Mr. Wannaphong Phatthiyaphaibun

Project details

Release history Release notifications | RSS feed

This version


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Built Distribution

nokcut-0.4-py3-none-any.whl (5.9 MB view hashes)

Uploaded py3

Supported by

AWS AWS Cloud computing Datadog Datadog Monitoring Facebook / Instagram Facebook / Instagram PSF Sponsor Fastly Fastly CDN Google Google Object Storage and Download Analytics Huawei Huawei PSF Sponsor Microsoft Microsoft PSF Sponsor NVIDIA NVIDIA PSF Sponsor Pingdom Pingdom Monitoring Salesforce Salesforce PSF Sponsor Sentry Sentry Error logging StatusPage StatusPage Status page