data source
Stanza
Open StanzaPer-language pipeline resource listing (tokenize/mwt/pos/lemma/depparse/ner/sentiment/constituency/coref) -- the cleanest pipeline-capability source.
tokenizationmorphological-analysis
Languages with a pipeline
89
Max processors
9
Raw entries scanned
204
Unmatched codes
3
not in the crosswalk
Languages in Stanza
89 of 89 shown
Community portal
Opportunities, events, and experts across the digital language inclusion ecosystem.