Motoki Wu
|
7b5b49eef0
|
added contributor agreement
|
2017-11-17 17:27:20 -08:00 |
|
Motoki Wu
|
a52e195a0a
|
Fixes Issue #1207 where noun_chunks of Span gives an error.
Make sure to reference `self.doc` when getting the noun chunks.
Same fix as 9750a0128c
|
2017-11-17 17:16:20 -08:00 |
|
Motoki Wu
|
b818afaa0e
|
Added failing test for Issue #1207.
The noun chunk iterator should work for `Doc` but not for `Span`.
|
2017-11-17 17:04:27 -08:00 |
|
ines
|
c3051e95f7
|
Add note on attribute extension defaults (resolves #1587)
|
2017-11-17 19:14:29 +01:00 |
|
ines
|
954f8cc6d1
|
Update syntax theme (should move the modifications out to an extension sometime)
|
2017-11-17 19:13:53 +01:00 |
|
Ines Montani
|
f1a5c33294
|
Merge pull request #1604 from raphael0202/patch-1
Fix typo in documentation
|
2017-11-17 17:21:56 +00:00 |
|
Raphaël Bournhonesque
|
a0793fd4cc
|
Fix typo
|
2017-11-17 17:57:55 +01:00 |
|
Ines Montani
|
eee9cc41f4
|
Merge pull request #1602 from MartinoMensio/master (resolves #1599)
small typo on docs
|
2017-11-17 16:19:04 +00:00 |
|
Martino Mensio
|
239a0f391d
|
added contributor agreement
|
2017-11-17 16:30:09 +01:00 |
|
Martino Mensio
|
ce1aade41e
|
small typo on docs
|
2017-11-17 16:20:22 +01:00 |
|
Vadim Mazaev
|
a0739a06d4
|
Returned russian support from v1.10 branch
|
2017-11-17 17:06:15 +03:00 |
|
yuukos
|
7401152289
|
updated Russian tokenizer
moved the trying to import pymorph into __init__
|
2017-11-17 17:04:50 +03:00 |
|
yuukos
|
3aad66cf00
|
added russian language support
|
2017-11-17 17:04:22 +03:00 |
|
ines
|
7d5afadf5e
|
Update vectors_loc description
|
2017-11-17 14:57:11 +01:00 |
|
ines
|
4187cfe1ea
|
Merge branch 'master' of https://github.com/explosion/spaCy
|
2017-11-17 14:56:29 +01:00 |
|
ines
|
c57e05bec1
|
Make sure nr_dim is an int
In some languages (e.g. Dutch), the nr_dim is extracted as a byte string, causing an error down the line.
|
2017-11-17 14:56:27 +01:00 |
|
Ines Montani
|
1e3068ec33
|
Merge pull request #1594 from pavillet/patch-1
Update _spacy.jade
|
2017-11-16 23:42:59 +00:00 |
|
pavillet
|
ad2935f0c3
|
Update _spacy.jade
Doc example gives 'object is not subscriptable' error.
Correcting as an attribuet
|
2017-11-17 00:02:20 +01:00 |
|
ines
|
a3d4dd1a5d
|
Test adding of lots of pipeline components (see #1585)
Just to make sure that there's no error now or in the future with adding a large number of pipeline components.
|
2017-11-15 17:28:06 +01:00 |
|
Roman Domrachev
|
61d28d03e4
|
Try again to do selective remove cache
|
2017-11-15 19:11:12 +03:00 |
|
Roman Domrachev
|
b3311100c7
|
Merge branch 'master' of github.com:explosion/spaCy
|
2017-11-15 18:30:04 +03:00 |
|
Matthew Honnibal
|
b60d92aca8
|
Increment version
|
2017-11-15 16:14:46 +01:00 |
|
Roman Domrachev
|
505c6a2f2f
|
Completely cleanup tokenizer cache
Tokenizer cache can have be different keys than string
That modification can slow down tokenizer and need to be measured
|
2017-11-15 17:55:48 +03:00 |
|
Matthew Honnibal
|
cf0be62096
|
Increment version
|
2017-11-15 15:00:18 +01:00 |
|
Matthew Honnibal
|
716ccbb71e
|
Require thinc 6.10.1
|
2017-11-15 14:59:34 +01:00 |
|
ines
|
97a4f9362b
|
Merge branch 'master' of https://github.com/explosion/spaCy
|
2017-11-15 14:24:00 +01:00 |
|
ines
|
8e65247886
|
Fix lex.id if vectors is None
|
2017-11-15 14:23:58 +01:00 |
|
Matthew Honnibal
|
437ad1a852
|
Merge pull request #1570 from explosion/feature/fix-beam-leak
Fix memory leak in beam parser
|
2017-11-15 14:15:05 +01:00 |
|
Matthew Honnibal
|
2f169fdb0a
|
Set lex ID correctly for new tokens in Vocab
|
2017-11-15 13:58:03 +01:00 |
|
Matthew Honnibal
|
fe3c42a06b
|
Fix caching in tokenizer
|
2017-11-15 13:55:46 +01:00 |
|
Matthew Honnibal
|
8d692771f6
|
Improve profiling
|
2017-11-15 13:51:25 +01:00 |
|
Matthew Honnibal
|
b797dca977
|
Merge branch 'master' of https://github.com/explosion/spaCy
|
2017-11-15 13:11:43 +01:00 |
|
Ines Montani
|
9177c7d7aa
|
Merge pull request #1583 from yogendrasoni/master (resolves #1582)
Add rstrip after reading line from vec file #1582
|
2017-11-15 12:02:48 +00:00 |
|
ines
|
c9d72de0fb
|
Add dummy serialization methods for Japanese and missing lang getter (resolves #1557)
|
2017-11-15 12:44:02 +01:00 |
|
yogendrasoni
|
334ed433b2
|
rstrip line before rsplit
loading english fast text giving error because line contains new line at the end and rsplit is splitting it incorrectly
|
2017-11-15 13:55:08 +05:30 |
|
Matthew Honnibal
|
d274d3a3b9
|
Let beam forward use minibatches
|
2017-11-15 00:51:42 +01:00 |
|
Matthew Honnibal
|
855872f872
|
Remove state hashing
|
2017-11-14 23:36:46 +01:00 |
|
Roman Domrachev
|
3e21680814
|
Use safer method to get string without hit
|
2017-11-14 22:58:46 +03:00 |
|
Roman Domrachev
|
a33d5a068d
|
Try to hold origin data instead of restore it
|
2017-11-14 22:40:03 +03:00 |
|
ines
|
40c4e8fc09
|
Remove "optional" from dev_data arg and add more info (see #1578)
|
2017-11-14 20:26:05 +01:00 |
|
Roman Domrachev
|
91e2fa6561
|
Clean all caches
|
2017-11-14 21:15:04 +03:00 |
|
Roman Domrachev
|
4e378dc4a4
|
Remove all obsolete code and test only initial problem
|
2017-11-14 20:45:04 +03:00 |
|
Roman
|
47ce2347b0
|
Create test that fails when actual cleanup caused
|
2017-11-14 20:28:13 +03:00 |
|
Roman
|
caae77f72d
|
Update strings.pyx
|
2017-11-14 19:44:40 +03:00 |
|
Roman Domrachev
|
3d247d2bb8
|
Get back previous testcase
|
2017-11-14 18:01:37 +03:00 |
|
Roman Domrachev
|
870defa815
|
Swap keys in proper place
Remove unnecessary clear of the hits
|
2017-11-14 17:56:30 +03:00 |
|
Roman Domrachev
|
86ca434c93
|
Merge github.com:explosion/spaCy
|
2017-11-14 17:46:22 +03:00 |
|
Roman Domrachev
|
a2745b0e84
|
StringStore now actually cleaned
Do not lose docs in ref tracking
|
2017-11-14 17:45:50 +03:00 |
|
Matthew Honnibal
|
2512ea9eeb
|
Fix memory leak in beam parser
|
2017-11-14 02:11:40 +01:00 |
|
Ines Montani
|
48b6cfe59e
|
Merge pull request #1569 from KMLDS/patch-1
trivial typo in docs
|
2017-11-14 01:46:34 +01:00 |
|