mirror of
https://github.com/explosion/spaCy.git
synced 2024-11-14 13:47:13 +03:00
9a478b6db8
* splitting up latin unicode interval * removing hyphen as infix for French * adding failing test for issue 1235 * test for issue #3002 which now works * partial fix for issue #2070 * keep the hyphen as infix for French (as it was) * restore french expressions with hyphen as infix (as it was) * added succeeding unit test for Issue #2656 * Fix issue #2822 with custom Italian exception * Fix issue #2926 by allowing numbers right before infix / * splitting up latin unicode interval * removing hyphen as infix for French * adding failing test for issue 1235 * test for issue #3002 which now works * partial fix for issue #2070 * keep the hyphen as infix for French (as it was) * restore french expressions with hyphen as infix (as it was) * added succeeding unit test for Issue #2656 * Fix issue #2822 with custom Italian exception * Fix issue #2926 by allowing numbers right before infix / * remove duplicate * remove xfail for Issue #2179 fixed by Matt * adjust documentation and remove reference to regex lib
10 lines
173 B
Python
10 lines
173 B
Python
# coding: utf8
|
|
from __future__ import unicode_literals
|
|
from ...symbols import ORTH, LEMMA
|
|
|
|
_exc = {
|
|
"po'": [{ORTH: "po'", LEMMA: 'poco'}]
|
|
}
|
|
|
|
TOKENIZER_EXCEPTIONS = _exc
|