spaCy/spacy/tests/lang/fi
Antti Ajanki e626a011cc Improvements to the Finnish language data (#4738)
* Enable lex_attrs on Finnish

* Copy the Danish tokenizer rules to Finnish

Specifically, don't break hyphenated compound words

* Contributor agreement

* A new file for Finnish tokenizer rules instead of including the Danish ones
2019-12-03 12:55:28 +01:00
..
__init__.py Revert #4334 2019-09-29 17:32:12 +02:00
test_text.py Improvements to the Finnish language data (#4738) 2019-12-03 12:55:28 +01:00
test_tokenizer.py Improvements to the Finnish language data (#4738) 2019-12-03 12:55:28 +01:00