spaCy/spacy/tests/parser
Adriane Boyd 8b650f3a78 Modify setting missing and blocked entity tokens
In order to make it easier to construct `Doc` objects as training data,
modify how missing and blocked entity tokens are set to prioritize
setting `O` and missing entity tokens for training purposes over setting
blocked entity tokens.

* `Doc.ents` setter sets tokens outside entity spans to `O` regardless
of the current state of each token

* For `Doc.ents`, setting a span with a missing label sets the `ent_iob`
to missing instead of blocked

* `Doc.block_ents(spans)` marks spans as hard `O` for use with the
`EntityRecognizer`
2020-09-17 21:27:42 +02:00
..
__init__.py Revert #4334 2019-09-29 17:32:12 +02:00
test_add_label.py Tidy up and auto-format [ci skip] 2020-09-13 10:55:36 +02:00
test_arc_eager_oracle.py Renaming gold & annotation_setter (#6042) 2020-09-09 10:31:03 +02:00
test_ner.py Modify setting missing and blocked entity tokens 2020-09-17 21:27:42 +02:00
test_neural_parser.py Renaming gold & annotation_setter (#6042) 2020-09-09 10:31:03 +02:00
test_nn_beam.py Improve spacy.gold (no GoldParse, no json format!) (#5555) 2020-06-26 19:34:12 +02:00
test_nonproj.py Tidy up and auto-format 2020-08-05 16:00:59 +02:00
test_parse_navigate.py Refactor Docs.is_ flags (#6044) 2020-09-17 00:14:01 +02:00
test_parse.py Refactor Docs.is_ flags (#6044) 2020-09-17 00:14:01 +02:00
test_preset_sbd.py Renaming gold & annotation_setter (#6042) 2020-09-09 10:31:03 +02:00
test_space_attachment.py Refactor Docs.is_ flags (#6044) 2020-09-17 00:14:01 +02:00