spaCy/spacy/cli
Adriane Boyd f94168a41e
Backport bugfixes from v3.1.0 to v3.0 (#8739)
* Fix scoring normalization (#7629)

* fix scoring normalization

* score weights by total sum instead of per component

* cleanup

* more cleanup

* Use a context manager when reading model (fix #7036) (#8244)

* Fix other open calls without context managers (#8245)

* Don't add duplicate patterns all the time in EntityRuler (fix #8216) (#8246)

* Don't add duplicate patterns (fix #8216)

* Refactor EntityRuler init

This simplifies the EntityRuler init code. This is helpful as prep for
allowing the EntityRuler to reset itself.

* Make EntityRuler.clear reset matchers

Includes a new test for this.

* Tidy PhraseMatcher instantiation

Since the attr can be None safely now, the guard if is no longer
required here.

Also renamed the `_validate` attr. Maybe it's not needed?

* Fix NER test

* Add test to make sure patterns aren't increasing

* Move test to regression tests

* Exclude generated .cpp files from package (#8271)

* Fix non-deterministic deduplication in Greek lemmatizer (#8421)

* Fix setting empty entities in Example.from_dict (#8426)

* Filter W036 for entity ruler, etc. (#8424)

* Preserve paths.vectors/initialize.vectors setting in quickstart template

* Various fixes for spans in Docs.from_docs (#8487)

* Fix spans offsets if a doc ends in a single space and no space is
  inserted
* Also include spans key in merged doc for empty spans lists

* Fix duplicate spacy package CLI opts (#8551)

Use `-c` for `--code` and not additionally for `--create-meta`, in line
with the docs.

* Raise an error for textcat with <2 labels (#8584)

* Raise an error for textcat with <2 labels

Raise an error if initializing a `textcat` component without at least
two labels.

* Add similar note to docs

* Update positive_label description in API docs

* Add Macedonian models to website (#8637)

* Fix Azerbaijani init, extend lang init tests (#8656)

* Extend langs in initialize tests

* Fix az init

* Fix ru/uk lemmatizer mp with spawn (#8657)

Use an instance variable instead a class variable for the morphological
analzyer so that multiprocessing with spawn is possible.

* Use 0-vector for OOV lexemes (#8639)

* Set version to v3.0.7

Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com>
Co-authored-by: Paul O'Leary McCann <polm@dampfkraft.com>
2021-07-19 09:20:40 +02:00
..
project Support env vars and CLI overrides for project.yml 2021-02-10 13:45:27 +11:00
templates Backport bugfixes from v3.1.0 to v3.0 (#8739) 2021-07-19 09:20:40 +02:00
__init__.py assemble CLI command (#7783) 2021-04-19 18:39:11 +10:00
_util.py Add hint for --gpu-id to CLI device info (#7234) 2021-03-03 01:11:18 +11:00
assemble.py assemble CLI command (#7783) 2021-04-19 18:39:11 +10:00
convert.py Backport bugfixes from v3.1.0 to v3.0 (#8739) 2021-07-19 09:20:40 +02:00
debug_config.py Add --code to spacy debug CLI 2021-03-12 09:51:26 +01:00
debug_data.py Update debug data for textcat (#8066) 2021-05-17 13:27:04 +02:00
debug_model.py Fix 'debug model' for transformers + generalize (#7973) 2021-05-06 18:43:32 +10:00
download.py textcat scoring fix and multi_label docs (#6974) 2021-03-09 23:04:22 +11:00
evaluate.py Fix displacy output in evaluate CLI (#7122) 2021-02-19 23:01:20 +11:00
info.py Replace links to nightly docs [ci skip] 2021-01-30 20:09:38 +11:00
init_config.py Add --code option to init fill-config 2021-03-12 10:03:57 +01:00
init_pipeline.py WIP: Various small training changes (#6818) 2021-01-26 14:51:52 +11:00
package.py Backport bugfixes from v3.1.0 to v3.0 (#8739) 2021-07-19 09:20:40 +02:00
pretrain.py Check if the resume path points to a directory (#7919) 2021-04-28 09:17:15 +02:00
profile.py Replace links to nightly docs [ci skip] 2021-01-30 20:09:38 +11:00
train.py Replace links to nightly docs [ci skip] 2021-01-30 20:09:38 +11:00
validate.py Update validate CLI to fix compat and ignore warnings (#8423) 2021-07-14 23:28:08 +02:00