Adriane Boyd
abd3f2b65a
Rename Polish lemmatizer method ( #5960 )
...
Rename Polish lemmatizer method to `pos_lookup` to distinguish it from
pure token-based lookup methods.
2020-08-25 00:22:27 +02:00
Ines Montani
e12b03358b
Support removing extra values in fill-config ( #5966 )
...
* Support removing extra values in fill-config
* Fix test
2020-08-24 22:53:47 +02:00
Matthew Honnibal
f232d8db96
Report p/r/f out of 100
2020-08-24 17:17:23 +02:00
Matthew Honnibal
311e1593e6
Fix makefile
2020-08-24 17:09:05 +02:00
Matthew Honnibal
5febbee0ff
Fix makefile
2020-08-24 16:59:09 +02:00
Matthew Honnibal
4fc1e57cf8
Fix makefile
2020-08-24 16:46:03 +02:00
Matthew Honnibal
f25cff1e38
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
2020-08-24 16:37:34 +02:00
Matthew Honnibal
cab286fbb2
Update makefile
2020-08-24 16:32:21 +02:00
Ines Montani
0e7f99da58
Fix handling of optional [pretraining] block ( #5954 )
...
* Fix handling of optional [pretraining] block
* Remote pretraining from default config
* Fix test
* Add schema option for empty pretrain block
2020-08-24 15:56:03 +02:00
Matthew Honnibal
463f1c8623
Avoid requiring smart-open directly
2020-08-24 14:49:17 +02:00
Matthew Honnibal
2ff29e603a
Update Makefile
2020-08-24 14:48:32 +02:00
Matthew Honnibal
6963260ff6
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
2020-08-24 14:42:10 +02:00
Matthew Honnibal
ecd86bae84
Update Makefile
2020-08-24 14:41:56 +02:00
Matthew Honnibal
944b1246f0
Add script to get package name
2020-08-24 14:41:49 +02:00
idoshr
b10c7bc56e
Hebrew like num ( #5952 )
...
* Update stop_words.py
Hebrew STOP WORDS
* Update stop_words.py
* contributor
* contributor
* add some common domain extentions
support human number 1K/1M....
* support human number 1K/1M....
* hebrew number tokenize
1K/1M implement in EN
* test human tokenize fix
* test
* heb like num
revert human number change
* heb like num
2020-08-24 14:30:05 +02:00
Ines Montani
967d69ec50
Fix website deployment [ci skip]
2020-08-24 14:28:24 +02:00
Ines Montani
26405710e0
Add icon credit [ci skip]
2020-08-24 10:28:15 +02:00
Matthew Honnibal
64df37643f
Update lockfile after project pull
2020-08-24 03:27:09 +02:00
Matthew Honnibal
588c28fe45
Fix project pull when deps missing
2020-08-24 01:23:36 +02:00
Matthew Honnibal
001546c19e
Set version to v3.0.0a10
2020-08-23 21:15:38 +02:00
Matthew Honnibal
160a855246
Format
2020-08-23 21:15:12 +02:00
Matthew Honnibal
89f5b8abb3
Fix project push
2020-08-23 21:14:44 +02:00
Matthew Honnibal
3828bc3ed0
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
2020-08-23 18:32:24 +02:00
Matthew Honnibal
e559867605
Allow spacy project to push and pull to/from remote storage ( #5949 )
...
* Add utils for working with remote storage
* WIP add remote_cache for project
* WIP add push and pull commands
* Use pathy in remote_cache
* Updarte util
* Update remote_cache
* Update util
* Update project assets
* Update pull script
* Update push script
* Fix type annotation in util
* Work on remote storage
* Remove site and env hash
* Fix imports
* Fix type annotation
* Require pathy
* Require pathy
* Fix import
* Add a util to handle project variable substitution
* Import push and pull commands
* Fix pull command
* Fix push command
* Fix tarfile in remote_storage
* Improve printing
* Fiddle with status messages
* Set version to v3.0.0a9
* Draft docs for spacy project remote storages
* Update docs [ci skip]
* Use Thinc config to simplify and unify template variables
* Auto-format
* Don't import Pathy globally for now
Causes slow and annoying Google Cloud warning
* Tidy up test
* Tidy up and update tests
* Update to latest Thinc
* Update docs
* variables -> vars
* Update docs [ci skip]
* Update docs [ci skip]
Co-authored-by: Ines Montani <ines@ines.io>
2020-08-23 18:32:09 +02:00
Matthew Honnibal
fe1cf7e124
Allow score_weights to list extra scores
2020-08-23 18:31:30 +02:00
Ines Montani
9bdc9e81f5
Fix error message [ci skip]
2020-08-23 12:14:02 +02:00
Ines Montani
f27aecac14
Update formatting [ci skip]
2020-08-23 11:57:56 +02:00
Ines Montani
98a9e063b6
Update docs [ci skip]
2020-08-22 17:15:05 +02:00
Matthew Honnibal
8dfc4cbfe7
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
2020-08-22 17:12:09 +02:00
Matthew Honnibal
048de64d4c
Suggest edits
2020-08-22 17:11:28 +02:00
Ines Montani
adcf790b96
Update docs[ci skip]
2020-08-22 17:04:16 +02:00
Ines Montani
37ebff6997
Update docs [ci skip]
2020-08-22 16:47:03 +02:00
Matthew Honnibal
8685229891
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
2020-08-22 16:06:59 +02:00
Matthew Honnibal
d97695d09d
Update embeddings-transformers.md
2020-08-22 15:41:35 +02:00
Ines Montani
c7c9b0451f
Update docs [ci skip]
2020-08-22 13:52:52 +02:00
Ines Montani
9740f1712b
Re-add font with lowercase name [ci skip]
2020-08-22 12:27:02 +02:00
Ines Montani
ff5ba14d06
Remove font [ci skip]
2020-08-22 12:26:50 +02:00
Ines Montani
71aeae89c5
Merge pull request #5948 from svlandeg/feature/docs-docs-docs [ci skip]
2020-08-22 12:18:47 +02:00
Ines Montani
27f81109d6
Update docs [ci skip]
2020-08-21 20:02:18 +02:00
Ines Montani
f102164a1f
Update docs [ci skip]
2020-08-21 19:34:06 +02:00
svlandeg
1b7cfa7347
Merge remote-tracking branch 'upstream/develop' into feature/docs-docs-docs
2020-08-21 18:36:18 +02:00
svlandeg
942adf0f4d
comma
2020-08-21 18:36:02 +02:00
svlandeg
262552010d
context manager with space (for consistency)
2020-08-21 18:34:02 +02:00
svlandeg
da48c6a2a2
several small updates
2020-08-21 18:25:26 +02:00
svlandeg
ad2332d4b7
alphabetize registries
2020-08-21 18:10:31 +02:00
svlandeg
dc98f69b57
alphabetize registries
2020-08-21 18:10:21 +02:00
svlandeg
c6659e37d8
small fixes
2020-08-21 18:02:20 +02:00
svlandeg
518a1f97f3
remove outdated TODO's
2020-08-21 17:55:15 +02:00
svlandeg
e92bd6e1c1
alphabetize training lists
2020-08-21 17:42:19 +02:00
Sofie Van Landeghem
56eabcb2f2
Adding num_like test for Czech ( #5946 )
...
* Create lex_attrs.py
Hello,
I am missing a CZECH language in SpaCy. So I would like to help to push it a little. This file is base on others lex_attrs.py files just with translation to Czech.
* Update __init__.py
Updated for use with new Czech Lex_attrs file
* Update stop_words.py
* Create test_text.py
* add like_num testing for czech
Co-authored-by: holubvl3 <47881982+holubvl3@users.noreply.github.com>
Co-authored-by: holubvl3 <vilemrousi@gmail.com>
Co-authored-by: Vladimír Holubec <vholubec@arcdata.cz>
2020-08-21 17:06:33 +02:00