spaCy/spacy/lang/lb/__init__.py

# coding: utf8
from __future__ import unicode_literals

from .tokenizer_exceptions import TOKENIZER_EXCEPTIONS
from .norm_exceptions import NORM_EXCEPTIONS
from .punctuation import TOKENIZER_INFIXES
from .lex_attrs import LEX_ATTRS
from .tag_map import TAG_MAP
from .stop_words import STOP_WORDS

from ..tokenizer_exceptions import BASE_EXCEPTIONS
from ..norm_exceptions import BASE_NORMS
from ...language import Language
from ...attrs import LANG, NORM
from ...util import update_exc, add_lookups


class LuxembourgishDefaults(Language.Defaults):
    lex_attr_getters = dict(Language.Defaults.lex_attr_getters)
    lex_attr_getters.update(LEX_ATTRS)
    lex_attr_getters[LANG] = lambda text: "lb"
    lex_attr_getters[NORM] = add_lookups(
        Language.Defaults.lex_attr_getters[NORM], NORM_EXCEPTIONS, BASE_NORMS
    )
    tokenizer_exceptions = update_exc(BASE_EXCEPTIONS, TOKENIZER_EXCEPTIONS)
    stop_words = STOP_WORDS
    tag_map = TAG_MAP
    infixes = TOKENIZER_INFIXES


class Luxembourgish(Language):
    lang = "lb"
    Defaults = LuxembourgishDefaults


__all__ = ["Luxembourgish"]
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`# coding: utf8`
			`from __future__ import unicode_literals`

Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`from .tokenizer_exceptions import TOKENIZER_EXCEPTIONS`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`from .norm_exceptions import NORM_EXCEPTIONS`
Fix basic language support for Luxembourgish (by adding punctuation.py) (#4648) * Update __init__.py * Create punctuation.py * Update tokenizer_exceptions.py * Create questoph.md * Update questoph.md * Update test_text.py * Update test_text.py * Update test_text.py * Update test_text.py 2019-11-15 18:16:47 +03:00			`from .punctuation import TOKENIZER_INFIXES`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`from .lex_attrs import LEX_ATTRS`
			`from .tag_map import TAG_MAP`
			`from .stop_words import STOP_WORDS`

			`from ..tokenizer_exceptions import BASE_EXCEPTIONS`
			`from ..norm_exceptions import BASE_NORMS`
			`from ...language import Language`
			`from ...attrs import LANG, NORM`
			`from ...util import update_exc, add_lookups`


			`class LuxembourgishDefaults(Language.Defaults):`
			`lex_attr_getters = dict(Language.Defaults.lex_attr_getters)`
			`lex_attr_getters.update(LEX_ATTRS)`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`lex_attr_getters[LANG] = lambda text: "lb"`
			`lex_attr_getters[NORM] = add_lookups(`
			`Language.Defaults.lex_attr_getters[NORM], NORM_EXCEPTIONS, BASE_NORMS`
			`)`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`tokenizer_exceptions = update_exc(BASE_EXCEPTIONS, TOKENIZER_EXCEPTIONS)`
			`stop_words = STOP_WORDS`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`tag_map = TAG_MAP`
Fix basic language support for Luxembourgish (by adding punctuation.py) (#4648) * Update __init__.py * Create punctuation.py * Update tokenizer_exceptions.py * Create questoph.md * Update questoph.md * Update test_text.py * Update test_text.py * Update test_text.py * Update test_text.py 2019-11-15 18:16:47 +03:00			`infixes = TOKENIZER_INFIXES`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00
			`class Luxembourgish(Language):`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`lang = "lb"`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`Defaults = LuxembourgishDefaults`


Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`__all__ = ["Luxembourgish"]`