Matthew Honnibal
|
e889cd698e
|
Increment version
|
2017-01-16 14:01:35 +01:00 |
|
Matthew Honnibal
|
e7f8e13cf3
|
Make Token hashable. Fixes #743
|
2017-01-16 13:27:57 +01:00 |
|
Matthew Honnibal
|
2c60d0cb1e
|
Test #743: Tokens unhashable.
|
2017-01-16 13:27:26 +01:00 |
|
Matthew Honnibal
|
48c712f1c1
|
Merge branch 'master' of ssh://github.com/explosion/spaCy
|
2017-01-16 13:18:06 +01:00 |
|
Matthew Honnibal
|
7ccf490c73
|
Increment version
|
2017-01-16 13:17:58 +01:00 |
|
Matthew Honnibal
|
d4e6d4c1c4
|
Use new thinc
|
2017-01-16 13:17:14 +01:00 |
|
Ines Montani
|
50878ef598
|
Exclude "were" and "Were" from tokenizer exceptions and add regression test (resolves #744)
|
2017-01-16 13:10:38 +01:00 |
|
Ines Montani
|
e053c7693b
|
Fix formatting
|
2017-01-16 13:09:52 +01:00 |
|
Ines Montani
|
116c675c3c
|
Merge pull request #742 from oroszgy/hu_tokenizer_fix
Improved Hungarian tokenizer
|
2017-01-14 23:52:44 +01:00 |
|
Gyorgy Orosz
|
92345b6a41
|
Further numeric test.
|
2017-01-14 22:44:19 +01:00 |
|
Gyorgy Orosz
|
b4df202bfa
|
Better error handling
|
2017-01-14 22:24:58 +01:00 |
|
Ines Montani
|
853130bcf8
|
Update installation instructions (see #727)
|
2017-01-14 22:12:42 +01:00 |
|
Gyorgy Orosz
|
b03a46792c
|
Better error handling
|
2017-01-14 22:09:29 +01:00 |
|
Gyorgy Orosz
|
a45f22913f
|
Added further abbreviations present in the Szeged corpus
|
2017-01-14 22:08:55 +01:00 |
|
Ines Montani
|
a3e3df3e33
|
Clean up fabfile
|
2017-01-14 21:30:38 +01:00 |
|
Ines Montani
|
332ce2d758
|
Update README.md
|
2017-01-14 21:12:11 +01:00 |
|
Ines Montani
|
c77698af25
|
Update CONTRIBUTING.md
|
2017-01-14 21:02:35 +01:00 |
|
Ines Montani
|
8dcb1c183d
|
Update CONTRIBUTING.md
|
2017-01-14 21:01:46 +01:00 |
|
Gyorgy Orosz
|
9505c6a72b
|
Passing all old tests.
|
2017-01-14 20:39:21 +01:00 |
|
Ines Montani
|
a538723047
|
Update README.rst
|
2017-01-14 19:55:38 +01:00 |
|
Ines Montani
|
696408e54f
|
Update README.rst
|
2017-01-14 19:02:05 +01:00 |
|
Ines Montani
|
89f85f0918
|
Update README.rst
|
2017-01-14 19:00:12 +01:00 |
|
Ines Montani
|
258e322e9c
|
Allow control of virtualenv in fabfile via environment variable
|
2017-01-14 17:49:32 +01:00 |
|
Gyorgy Orosz
|
63037e79af
|
Fixed hyphen handling in the Hungarian tokenizer.
|
2017-01-14 16:30:11 +01:00 |
|
Gyorgy Orosz
|
f77c0284d6
|
Maintaining compatibility with other spacy tokenizers.
|
2017-01-14 16:19:15 +01:00 |
|
Gyorgy Orosz
|
be7a7aeb1a
|
Reversed accidental changes.
|
2017-01-14 15:59:36 +01:00 |
|
Gyorgy Orosz
|
1be5da1ac6
|
Fixed Hungarian tokenizer for numbers
|
2017-01-14 15:51:59 +01:00 |
|
Ines Montani
|
a89e269a5a
|
Fix test formatting and consistency
|
2017-01-14 13:41:19 +01:00 |
|
Ines Montani
|
3424e3a7e5
|
Update README.md
|
2017-01-13 15:54:54 +01:00 |
|
Ines Montani
|
d12dabb8bc
|
Update CONTRIBUTING.md
|
2017-01-13 15:51:22 +01:00 |
|
Ines Montani
|
155a7b4bea
|
Update CONTRIBUTING.md
|
2017-01-13 15:25:05 +01:00 |
|
Ines Montani
|
b592b98a5a
|
Update contributing guidelines to add info on tests
|
2017-01-13 15:23:58 +01:00 |
|
Ines Montani
|
49186b34a1
|
Mark lemmatizer tests as models since they use installed data
|
2017-01-13 15:12:07 +01:00 |
|
Ines Montani
|
138deb80a1
|
Modernise vector tests, use add_vecs_to_vocab and don't depend on models
|
2017-01-13 15:12:07 +01:00 |
|
Ines Montani
|
96f0caa28a
|
Fix test name for consistency
|
2017-01-13 15:12:07 +01:00 |
|
Ines Montani
|
dc2bb1259f
|
Add util function to add vectors to vocab
|
2017-01-13 15:12:07 +01:00 |
|
Ines Montani
|
db9b25663d
|
Reformat add_docs_equal and add docstring
|
2017-01-13 15:12:07 +01:00 |
|
Ines Montani
|
62ce0a0073
|
Add README.md to tests to explain organisation and conventions
|
2017-01-13 15:11:18 +01:00 |
|
Ines Montani
|
01c2daf0c9
|
Update CONTRIBUTING.md
|
2017-01-13 11:27:18 +01:00 |
|
Ines Montani
|
38d60f6b90
|
Modernise serializer I/O tests and don't depend on models where possible
|
2017-01-13 02:24:56 +01:00 |
|
Ines Montani
|
4bb5b89ee4
|
Add text_file_b fixture using BytesIO
|
2017-01-13 02:23:50 +01:00 |
|
Ines Montani
|
49febd8c62
|
Modernise noun chunks tests and don't depend on models
|
2017-01-13 02:01:00 +01:00 |
|
Ines Montani
|
3ee97b5686
|
Rename test_parser to test_noun_chunks
|
2017-01-13 01:36:33 +01:00 |
|
Ines Montani
|
a308703f47
|
Remove old tests
|
2017-01-13 01:34:48 +01:00 |
|
Ines Montani
|
12eb8edf26
|
Move parser tests from unit to parser
|
2017-01-13 01:34:38 +01:00 |
|
Ines Montani
|
138c53ff2e
|
Merge tokenizer tests
|
2017-01-13 01:34:14 +01:00 |
|
Ines Montani
|
01f36ca3ff
|
Move attrs tests from unit to root and modernise
|
2017-01-13 01:33:50 +01:00 |
|
Ines Montani
|
3610d27967
|
Move alignment tests from munge to gold and modernise
|
2017-01-13 01:33:31 +01:00 |
|
Ines Montani
|
094ff7396a
|
Reformat and rename Pragmatic Segmenter tests and mark xfails
|
2017-01-13 01:30:20 +01:00 |
|
Ines Montani
|
affcf1b19d
|
Modernise lemmatizer tests
|
2017-01-12 23:41:17 +01:00 |
|