Adriane Boyd
5e47a54d29
Include noun chunks method when pickling Vocab
2021-02-12 13:27:46 +01:00
Ines Montani
26bf642afd
Fix issue #7019 : Handle None scores in evaluate printer ( #7026 )
2021-02-11 16:45:23 +11:00
Ines Montani
6b9026a219
Merge pull request #7000 from explosion/feature/project-yml-overrides
...
Support env vars and CLI overrides for project.yml
2021-02-11 12:31:45 +11:00
Ines Montani
e21639fa56
Merge pull request #7025 from explosion/fix/6950
...
Fix issue #6950 : allow pickling Tok2Vec with listeners
2021-02-11 12:25:16 +11:00
Ines Montani
ad9ce3c8f6
Fix issue #6950 : allow pickling Tok2Vec with listeners
2021-02-11 11:37:39 +11:00
Peter Baumann
61b04a70d5
Run PhraseMatcher on Spans ( #6918 )
...
* Add regression test
* Run PhraseMatcher on Spans
* Add test for PhraseMatcher on Spans and Docs
* Add SCA
* Add test with 3 matches in Doc, 1 match in Span
* Update docs
* Use doc.length for find_matches in tokenizer
Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com>
2021-02-10 23:43:32 +11:00
Ines Montani
a96f850cb7
Merge pull request #6983 from KoichiYasuoka/patch-2
2021-02-10 15:28:52 +11:00
Ines Montani
21176c69b0
Update and add test
2021-02-10 14:12:00 +11:00
Ines Montani
c08b3f294c
Support env vars and CLI overrides for project.yml
2021-02-10 13:45:27 +11:00
Ines Montani
4bf7f63c22
Merge pull request #6996 from svlandeg/fix/accuracy [ci skip]
...
final 3.0 benchmark numbers
2021-02-10 12:10:16 +11:00
svlandeg
9a7f33c916
final 3.0 benchmark numbers
2021-02-09 21:28:33 +01:00
Koichi Yasuoka
8ed788660b
Several callable objects do not have __qualname__
2021-02-09 14:43:02 +09:00
Ines Montani
ca3f8386d7
Merge pull request #6975 from svlandeg/fix/link [ci skip]
...
fix link
2021-02-09 14:34:11 +11:00
Ines Montani
050f089a2e
Merge pull request #6976 from tarskiandhutch/patch-2 [ci skip]
...
Syntax error in Lemmatizer docs
2021-02-09 14:33:55 +11:00
tarskiandhutch
e897e7aaad
Line 70: syntax error
...
Original config definition treated dictionary key as a function argument.
2021-02-08 15:24:57 -05:00
svlandeg
bb7482bef8
fix link
2021-02-08 18:39:59 +01:00
Ines Montani
25a0c6c32f
Merge pull request #6965 from adrianeboyd/feature/reword-e923 [ci skip]
...
Rephrase error related to sample data initialization
2021-02-08 20:17:14 +11:00
Adriane Boyd
6108dabdc8
Rephrase error related to sample data initialization
...
Now that the initialize step is fully implemented, the source of E923 is
typically missing or improperly converted/formatted data rather than a
bug in spaCy, so rephrase the error and message and remove the prompt to
open an issue.
2021-02-08 09:21:36 +01:00
Sofie Van Landeghem
6ed423c16c
reduce memory load when reading all vectors from file ( #6945 )
...
* reduce memory load when reading all vectors from file
* one more small typo fix
2021-02-07 08:05:43 +08:00
Sofie Van Landeghem
a323ef90df
ensure the loss value is cast as float ( #6928 )
2021-02-07 07:51:56 +08:00
melonwater211
a7977b5143
The test spacy/tests/vocab_vectors/test_lexeme.py::test_vocab_lexeme_add_flag_auto_id
seems to fail occasionally when the test suite is run in a random order. ( #6956 )
...
```python
def test_vocab_lexeme_add_flag_auto_id(en_vocab):
is_len4 = en_vocab.add_flag(lambda string: len(string) == 4)
assert en_vocab["1999"].check_flag(is_len4) is True
assert en_vocab["1999"].check_flag(IS_DIGIT) is True
assert en_vocab["199"].check_flag(is_len4) is False
> assert en_vocab["199"].check_flag(IS_DIGIT) is True
E assert False is True
E + where False = <built-in method check_flag of spacy.lexeme.Lexeme object at 0x7fa155c36840>(3)
E + where <built-in method check_flag of spacy.lexeme.Lexeme object at 0x7fa155c36840> = <spacy.lexeme.Lexeme object at 0x7fa155c36840>.check_flag
spacy/tests/vocab_vectors/test_lexeme.py:49: AssertionError
```
> `pytest==6.1.1`
>
> `numpy==1.19.2`
>
> `Python version: 3.8.3`
To reproduce the error, run `pytest --random-order-bucket=global --random-order-seed=170158 -v spacy/tests`
If `test_vocab_lexeme_add_flag_auto_id` is run after `test_vocab_lexeme_add_flag_provided_id`, it fails.
It seems like `test_vocab_lexeme_add_flag_provided_id` uses the `IS_DIGIT` bit for testing purposes but does not reset the bit.
This solution seems to work but, if anyone has a better fix, please let me know and I will integrate it.
2021-02-07 07:51:34 +08:00
René Octavio Queiroz Dias
59271e887a
fix: TransformerListener with TextCatEnsemble ( #6951 )
...
* bug: Regression test
Issue #6946
* fix: Fix issue #6946
* chore: Remove regression test
2021-02-06 13:44:51 +01:00
Ines Montani
433835d9b0
Merge pull request #6889 from adrianeboyd/docs/source-install-dup [ci skip]
2021-02-05 13:35:16 +11:00
René Octavio Queiroz Dias
999ff03b19
fix: Fix textcat labels to expect a Optional[Iterable[str]] instead of Optional[Dict] ( #6911 )
...
* docs: Add agreement
* bug: Regression test
Issue #6908
* fix: Changed from Dict to Iterable[str]
Fix #6908
* Update test to use make_tempdir
* fix: Fix WindowsPath error
Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com>
2021-02-04 23:37:13 +01:00
Helio Machado
20a97cda38
Create 0x2b3bfa0.md ( #6916 )
2021-02-04 23:25:11 +01:00
Adriane Boyd
b903de3fcb
Pass on vocab arg in spacy.blank() ( #6924 )
2021-02-04 15:09:01 +01:00
Ines Montani
efdeb9b53f
Merge pull request #6909 from svlandeg/fix/docs [ci skip]
2021-02-03 23:59:15 +11:00
svlandeg
7cda5605a0
add type
2021-02-03 13:13:58 +01:00
svlandeg
94929c2b98
small doc fixes
2021-02-03 13:10:22 +01:00
Ines Montani
2cdfcd2d19
Update naming [ci skip]
2021-02-03 12:48:31 +11:00
Ines Montani
809f6282f2
Update README.md [ci skip]
2021-02-03 12:48:25 +11:00
Ines Montani
ae4fcabf20
Merge pull request #6896 from svlandeg/fix/capture
...
add capture arg
2021-02-03 12:35:06 +11:00
svlandeg
f852af2acf
add capture arg
2021-02-02 19:47:12 +01:00
Adriane Boyd
37a68a06ab
Update to recommend editable installs for source installs
2021-02-02 16:51:27 +01:00
Adriane Boyd
3a3e4daf60
Update install instructions
...
* Remove duplicate section about compiling from source
2021-02-02 14:44:15 +01:00
Matthew Honnibal
91a3cab1ca
Require spacy-transformers 1.0.1 for v3.0.1
2021-02-02 20:46:56 +11:00
Matthew Honnibal
b6a198481b
Set version to v3.0.0
2021-02-02 20:26:17 +11:00
Ines Montani
ff6a21cd18
Update GitHub link [ci skip]
2021-02-02 14:27:46 +11:00
Sofie Van Landeghem
f319d2765f
Add capture argument to project_run ( #6878 )
...
* add capture argument to project_run and run_commands
* git bump to 3.0.1
* Set version to 3.0.1.dev0
Co-authored-by: Matthew Honnibal <honnibal+gh@gmail.com>
2021-02-02 10:11:15 +08:00
Sofie Van Landeghem
f638306598
remove link_components flag again ( #6883 )
2021-02-02 10:08:40 +08:00
Ines Montani
e97d3f3c69
Merge pull request #6884 from pcyin/patch-1 [ci skip]
...
Fix a typo
2021-02-02 12:59:39 +11:00
Pengcheng YIN
6fdc33203a
Fix a typo
2021-02-01 17:26:28 -05:00
Ines Montani
a59f3fcf5d
Make wheel the default format and update docs [ci skip]
2021-02-01 23:18:43 +11:00
Ines Montani
b9573e9e22
Fix pip args
2021-02-01 23:15:00 +11:00
Ines Montani
b46073234a
Fix default clone branch and error handling [ci skip]
2021-02-01 22:29:04 +11:00
Sofie Van Landeghem
acabb284dd
Fix linking resumed components ( #6859 )
...
* link components across enabled, resumed and frozen
* revert renaming
* revert renaming, the sequel
2021-02-01 22:19:58 +11:00
Ines Montani
8a245076c4
Update spacy-transformers pin [ci skip]
2021-02-01 22:04:07 +11:00
Ines Montani
e17ea88e54
Fix config quickstart and download [ci skip]
2021-02-01 21:44:55 +11:00
Ines Montani
3b9ecd25d8
Merge pull request #6870 from adrianeboyd/bugfix/quickstart-lang-tokenizers [ci skip]
...
Remove nlp.tokenizer from quickstart template
2021-02-01 21:24:52 +11:00
Adriane Boyd
35a863cd27
Remove nlp.tokenizer from quickstart template
...
Remove `nlp.tokenizer` from quickstart template so that the default
language-specific tokenizer settings are filled instead.
2021-02-01 11:20:12 +01:00