Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
The Annotated Corpus of Classical Tibetan (ACTib), Part I is a part-of-speech tagged version of the Buddhist Digital Resource Center's digitized Tibetan etext collection. It was created using a memory-based tagger trained on a separate POS-tagged corpus of Classical Tibetan. The corpus includes files that were not manually corrected and contains some annotations from corrupted source files.
The dataset is tagged but not manually corrected. Some annotations were performed on corrupted source files. License is listed as 'Open Access (green)'.