spaCy/spacy/lang/fi
Antti Ajanki e626a011cc Improvements to the Finnish language data (#4738)
* Enable lex_attrs on Finnish

* Copy the Danish tokenizer rules to Finnish

Specifically, don't break hyphenated compound words

* Contributor agreement

* A new file for Finnish tokenizer rules instead of including the Danish ones
2019-12-03 12:55:28 +01:00
..
__init__.py
examples.py
lex_attrs.py
punctuation.py
stop_words.py
tokenizer_exceptions.py