← All data sources← All Language Processing observatories on IDLI

data source

Per-language pipeline resource listing (tokenize/mwt/pos/lemma/depparse/ner/sentiment/constituency/coref) -- the cleanest pipeline-capability source.

tokenizationmorphological-analysis
Languages with a pipeline
89
Max processors
9
Raw entries scanned
204
Unmatched codes
3
not in the crosswalk

Languages in Stanza

1 of 89 shown

Detail
Ancient Hebrew
hbo
523detail →

Community portal

Opportunities, events, and experts across the digital language inclusion ecosystem.