NLP for Welsh

Measured on OPUS, spaCy, Stanza, Tatoeba, NLLB-200, FLORES-200, Belebele.

Observatories covering Welsh

7 observatories measure NLP. Cards linking straight to this selection have data for it.

101,882,966 sentence pairs

15 corpora

No data for this selection

85 language directories; presence of tokenizer_exceptions and a lemmatizer distinguishes a real pipeline from a stub.

5 pipeline processors

of 9 possible

1,803 sentences

in this language

Yes in nllb-200

FLORES-200

Open index →

Yes in flores-200

No data for this selection

Reading comprehension across 100+ language variants. Presence means evaluable, not merely trainable.