spaCy/spacy/lang/lb/__init__.py

from .tokenizer_exceptions import TOKENIZER_EXCEPTIONS
from .punctuation import TOKENIZER_INFIXES
from .lex_attrs import LEX_ATTRS
from .stop_words import STOP_WORDS
from ...language import Language
from ...util import load_config_from_str


DEFAULT_CONFIG = """
[initialize]

[initialize.lookups]
@misc = "spacy.LookupsDataLoader.v1"
lang = ${nlp.lang}
tables = ["lexeme_norm"]
"""


class LuxembourgishDefaults(Language.Defaults):
    config = load_config_from_str(DEFAULT_CONFIG)
    tokenizer_exceptions = TOKENIZER_EXCEPTIONS
    infixes = TOKENIZER_INFIXES
    lex_attr_getters = LEX_ATTRS
    stop_words = STOP_WORDS


class Luxembourgish(Language):
    lang = "lb"
    Defaults = LuxembourgishDefaults


__all__ = ["Luxembourgish"]
Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`from .tokenizer_exceptions import TOKENIZER_EXCEPTIONS`
Fix basic language support for Luxembourgish (by adding punctuation.py) (#4648) * Update __init__.py * Create punctuation.py * Update tokenizer_exceptions.py * Create questoph.md * Update questoph.md * Update test_text.py * Update test_text.py * Update test_text.py * Update test_text.py 2019-11-15 18:16:47 +03:00			`from .punctuation import TOKENIZER_INFIXES`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`from .lex_attrs import LEX_ATTRS`
			`from .stop_words import STOP_WORDS`
			`from ...language import Language`
Add lexeme norm defaults 2020-09-30 11:20:14 +03:00			`from ...util import load_config_from_str`


			`DEFAULT_CONFIG = """`
			`[initialize]`

			`[initialize.lookups]`
			`@misc = "spacy.LookupsDataLoader.v1"`
			`lang = ${nlp.lang}`
			`tables = ["lexeme_norm"]`
			`"""`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00

			`class LuxembourgishDefaults(Language.Defaults):`
Add lexeme norm defaults 2020-09-30 11:20:14 +03:00			`config = load_config_from_str(DEFAULT_CONFIG)`
Tidy up and move noun_chunks, token_match, url_match 2020-07-22 23:18:46 +03:00			`tokenizer_exceptions = TOKENIZER_EXCEPTIONS`
Fix basic language support for Luxembourgish (by adding punctuation.py) (#4648) * Update __init__.py * Create punctuation.py * Update tokenizer_exceptions.py * Create questoph.md * Update questoph.md * Update test_text.py * Update test_text.py * Update test_text.py * Update test_text.py 2019-11-15 18:16:47 +03:00			`infixes = TOKENIZER_INFIXES`
Simplify language data and revert detailed configs 2020-07-24 15:50:26 +03:00			`lex_attr_getters = LEX_ATTRS`
			`stop_words = STOP_WORDS`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00
			`class Luxembourgish(Language):`
Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`lang = "lb"`
Initial commit: New language Luxembourgish (lb) (#4424) * new language: Luxembourgish (lb) * update * update * Update and rename .github/CONTRIBUTOR_AGREEMENT.md to .github/contributors/PeterGilles.md * Update and rename .github/contributors/PeterGilles.md to .github/CONTRIBUTOR_AGREEMENT.md * Update norm_exceptions.py * Delete README.md * moved test_lemma.py * deactivated 'lemma_lookup = LOOKUP' * update * Update conftest.py * update * tests updated * import unicode_literals * Update spacy/tests/lang/lb/test_text.py Co-Authored-By: Ines Montani <ines@ines.io> * Create PeterGilles.md 2019-10-14 13:27:50 +03:00			`Defaults = LuxembourgishDefaults`


Tidy up and auto-format 2019-10-18 12:27:38 +03:00			`__all__ = ["Luxembourgish"]`