ines
|
8a29308d0b
|
Remove unused imports
|
2017-06-04 22:39:29 +02:00 |
|
Ines Montani
|
112c5787eb
|
Merge pull request #1101 from oroszgy/hu_tokenizer_fix
More robust Hungarian tokenizer.
|
2017-06-04 22:37:51 +02:00 |
|
ines
|
96867a24ae
|
Fix typo
|
2017-06-04 22:36:40 +02:00 |
|
ines
|
f432bb4b48
|
Fix fixture scopes
|
2017-06-04 22:34:31 +02:00 |
|
Matthew Honnibal
|
6d0356e6cc
|
Whitespace
|
2017-06-04 14:55:24 -05:00 |
|
Matthew Honnibal
|
8a683a4494
|
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
|
2017-06-04 21:53:56 +02:00 |
|
Matthew Honnibal
|
92ae36f84e
|
Improve way noun chunks iterator is looked up
|
2017-06-04 21:53:39 +02:00 |
|
ines
|
9254a3dd78
|
Import and add Spanish syntax iterators
|
2017-06-04 21:42:15 +02:00 |
|
ines
|
7db1a0e83e
|
Make sure printed values are always strings
|
2017-06-04 21:27:20 +02:00 |
|
Matthew Honnibal
|
51e1541ddb
|
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
|
2017-06-04 14:26:29 -05:00 |
|
Matthew Honnibal
|
add9a33782
|
Return False for vocab.has_vector
|
2017-06-04 14:26:14 -05:00 |
|
Matthew Honnibal
|
675f448313
|
Fix vector linkage on Doc
|
2017-06-04 14:25:30 -05:00 |
|
Matthew Honnibal
|
f4662e9218
|
Fix vector linkage for token
|
2017-06-04 14:19:58 -05:00 |
|
ines
|
070e026ed9
|
Ensure path on read_json
|
2017-06-04 20:44:37 +02:00 |
|
ines
|
e1e73936b1
|
Raise correct error
|
2017-06-04 20:44:27 +02:00 |
|
ines
|
848e47669e
|
Fix typo
|
2017-06-04 20:44:15 +02:00 |
|
ines
|
c4614c02a2
|
Fix dev resources URL
|
2017-06-04 15:45:50 +02:00 |
|
ines
|
a66cf24ee8
|
xfail tokenizer serialization tests for now
Tests pass locally, but not on Travis – needs more investigation
|
2017-06-04 13:58:20 +02:00 |
|
ines
|
7b7d46b64e
|
Fix typo and success message
|
2017-06-04 13:45:50 +02:00 |
|
ines
|
90d117f378
|
Update version
|
2017-06-04 13:41:16 +02:00 |
|
Matthew Honnibal
|
7ca215bc26
|
Resolve lex_attr_getters conflict
|
2017-06-03 16:12:01 -05:00 |
|
Matthew Honnibal
|
21eef90dbc
|
Support specifying which GPU
|
2017-06-03 16:10:23 -05:00 |
|
Matthew Honnibal
|
d0e42f9275
|
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
|
2017-06-03 15:30:32 -05:00 |
|
Matthew Honnibal
|
8a17b99b1c
|
Use NORM attribute, not LOWER
|
2017-06-03 15:30:16 -05:00 |
|
ines
|
4c643d74c5
|
Add norm exceptions to other Language classes
|
2017-06-03 22:29:21 +02:00 |
|
ines
|
fa7e576c57
|
Change order of exception dicts
|
2017-06-03 21:52:06 +02:00 |
|
Matthew Honnibal
|
3f5c85d8de
|
Reorder setting of lex attrs, to avoid clobbering
|
2017-06-03 14:47:55 -05:00 |
|
Matthew Honnibal
|
aeb7520133
|
Make norm use lower-case
|
2017-06-03 14:47:38 -05:00 |
|
Matthew Honnibal
|
de3954843e
|
Populate norm exceptions with lower-case
|
2017-06-03 14:47:12 -05:00 |
|
Matthew Honnibal
|
f6955a459c
|
Fix prev commit
|
2017-06-03 14:38:37 -05:00 |
|
Matthew Honnibal
|
468ca6c760
|
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
|
2017-06-03 14:33:51 -05:00 |
|
Matthew Honnibal
|
c647a0d33e
|
Fix training counter for gold preprocessing
|
2017-06-03 14:33:39 -05:00 |
|
ines
|
e47eef5e03
|
Update German tokenizer exceptions and tests
|
2017-06-03 21:07:44 +02:00 |
|
ines
|
d77c2cc8bb
|
Add tests for English norm exceptions
|
2017-06-03 20:59:50 +02:00 |
|
ines
|
0d6fa8b241
|
Add German norm exceptions
|
2017-06-03 20:54:18 +02:00 |
|
ines
|
5bd311c77e
|
Fix update of norm exceptions
|
2017-06-03 20:54:09 +02:00 |
|
Matthew Honnibal
|
94e063ae2a
|
Merge branch 'develop' of https://github.com/explosion/spaCy into develop
|
2017-06-03 13:31:40 -05:00 |
|
Matthew Honnibal
|
fea1144e6d
|
Set max batch size in evaluate
|
2017-06-03 13:31:33 -05:00 |
|
Matthew Honnibal
|
805495af27
|
Fix off-by-one in number of tags
|
2017-06-03 13:29:23 -05:00 |
|
Matthew Honnibal
|
e62f46d39f
|
Clarify gold.pyx slightly
|
2017-06-03 13:28:52 -05:00 |
|
Matthew Honnibal
|
43353b5413
|
Improve train CLI script
|
2017-06-03 13:28:20 -05:00 |
|
ines
|
746653880c
|
Add English norm exceptions to lex_attrs
|
2017-06-03 20:27:28 +02:00 |
|
ines
|
095eeeb12f
|
Update English tokenizer exceptions and add norms
|
2017-06-03 20:27:16 +02:00 |
|
ines
|
e5d426406a
|
Add base norm exceptions
|
2017-06-03 20:27:05 +02:00 |
|
ines
|
4c2bbc3ccc
|
Add add_lookups util function
|
2017-06-03 19:44:47 +02:00 |
|
ines
|
05fe6758a7
|
Set lexeme attributes for tokenizer special cases
|
2017-06-03 19:44:39 +02:00 |
|
ines
|
3152ee5ca2
|
Update serialization tests for tokenizer
|
2017-06-03 17:05:28 +02:00 |
|
ines
|
7c919aeb09
|
Make sure serializers and deserializers are ordered
|
2017-06-03 17:05:09 +02:00 |
|
ines
|
1ebd0d3f27
|
Add assert_packed_msg_equal util function
|
2017-06-03 17:04:30 +02:00 |
|
ines
|
de974f7bef
|
Add serializer tests for tokenizer
|
2017-06-03 13:26:34 +02:00 |
|