Commit Graph

5618 Commits

Author SHA1 Message Date
Henning Peters
4d375afb91 run tests for wheels 2016-02-24 19:59:08 +01:00
Henning Peters
f3df736e0a remove unidecode-related test 2016-02-24 18:22:22 +01:00
Matthew Honnibal
db87db87ea * Update tagger.pxd for CharacterTagger model 2016-02-24 18:20:47 +01:00
Matthew Honnibal
77f2b218f9 * Update conll_train script 2016-02-24 18:19:38 +01:00
Matthew Honnibal
fab538672e * Refactor CharacterTagger 2016-02-24 18:17:16 +01:00
Matthew Honnibal
1ba31f6229 Merge pull request #275 from henningpeters/unidecode
remove text-unidecode dependency
2016-02-25 04:10:45 +11:00
Wolfgang Seeker
4b2297d5d4 add class PseudoProjective for pseudo-projective parsing
PseudoProjective() implements the algorithm from Nivre & Nilsson 2005
using their HEAD decoration scheme.
2016-02-24 11:26:25 +01:00
Henning Peters
12d58a7099 remove text-unidecode dependency 2016-02-24 08:01:59 +01:00
Henning Peters
63deae47fe Update buildbot.json 2016-02-23 13:36:04 +01:00
Matthew Honnibal
92e9134603 * Try new CoNLL tagger method 2016-02-22 22:57:06 +01:00
Matthew Honnibal
3f12fb4191 * Change defaults for character tagger 2016-02-22 22:56:33 +01:00
Matthew Honnibal
422b33838e * Add note explaining parse features 2016-02-22 22:55:52 +01:00
Wolfgang Seeker
8d531c958b replace tests for non-projectivity
- add functions to find non-projective edges
- add test file for non-projectivity functions
2016-02-22 14:40:40 +01:00
Henning Peters
dfd1a1d3a2 Update buildbot.json 2016-02-22 06:13:09 +01:00
Matthew Honnibal
141639ea3a * Fix bug in tokenizer that caused new tokens to be added for affixes 2016-02-21 23:17:47 +00:00
Matthew Honnibal
5f53ef1a43 * Update conll_train for tagger, to use neural network tagger 2016-02-22 00:16:40 +01:00
Matthew Honnibal
c3f334cef1 * Work on character tagger 2016-02-22 00:15:25 +01:00
Matthew Honnibal
7a519ea5af * Add cautionary note to vocab about encoding 2016-02-22 00:13:20 +01:00
Henning Peters
1501ef58e0 Update README.md 2016-02-19 19:36:47 +01:00
Henning Peters
85f94fd314 get rid of pip-clear.py 2016-02-19 18:48:02 +01:00
Henning Peters
37a7020904 move displacy to its own subdomain 2016-02-19 14:03:52 +01:00
Henning Peters
59339d45e5 remove displacy 2016-02-19 13:30:49 +01:00
Henning Peters
0bb05ec7e1 Merge branch 'master' of github.com:spacy-io/spaCy 2016-02-19 13:30:14 +01:00
Henning Peters
d86a2a7a78 Update _installation.jade
with ```pip install -e .``` we don't need to set the PYTHONPATH anymore
also sync build instructions with travis script
2016-02-18 22:54:20 +01:00
Wolfgang Seeker
eae35e9b27 add tokenizer files for German, add/change code to train German pos tagger
- add files to specify rules for German tokenization
- change generate_specials.py to generate from an external file (abbrev.de.tab)
- copy gazetteer.json from lang_data/en/

- init_model.py
	- change doc freq threshold to 0
- add train_german_tagger.py
	- expects conll09-formatted input
2016-02-18 13:24:20 +01:00
Matthew Honnibal
92f62bcb84 * Work on character tagger 2016-02-16 23:55:47 +01:00
Matthew Honnibal
7ae048cf76 * Delete old draft of sense2vec post 2016-02-15 14:50:01 +01:00
Matthew Honnibal
05ec31a134 * Tmp 2016-02-15 14:40:28 +01:00
Matthew Honnibal
2326c5298f * Rename post Sense2Vec with SpaCy 2016-02-15 09:16:58 +01:00
Matthew Honnibal
ceb87e6b14 * Add meta.jade for sense2vec post 2016-02-15 09:16:12 +01:00
Henning Peters
04e1054bfa Merge branch 'master' of github.com:henningpeters/spaCy 2016-02-15 01:34:06 +01:00
Henning Peters
9cc4f8d5b3 avoid shadowing __name__ 2016-02-15 01:33:39 +01:00
Henning Peters
135746947a Update package.json 2016-02-14 20:19:26 +01:00
Henning Peters
4c9e3c7911 upgrade spuntik, enforce data api via model version constraints 2016-02-14 16:03:17 +01:00
Henning Peters
9d8966a2c0 Update test_tokenizer.py 2016-02-10 19:24:37 +01:00
Henning Peters
82c57f21a4 Update requirements.txt 2016-02-10 18:56:21 +01:00
Matthew Honnibal
cc66a63e0a Merge pull request #255 from henningpeters/master
py26 compatibility
2016-02-11 03:24:52 +11:00
Henning Peters
3b5f1e753b py26 compatibility 2016-02-10 14:32:54 +01:00
Henning Peters
5c60847341 remove appveyor 2016-02-10 11:29:29 +01:00
Henning Peters
7e0d1dd8d3 remove appveyor 2016-02-10 11:28:55 +01:00
Henning Peters
73251eddac Update README.md 2016-02-10 11:05:08 +01:00
Henning Peters
66765d4d8f Update .appveyor.yml 2016-02-10 08:04:11 +01:00
Henning Peters
ee1f1ac300 mark test_sentence_space() as model test 2016-02-10 07:49:11 +01:00
Henning Peters
2072120d7d Update package.json 2016-02-09 19:54:52 +01:00
Henning Peters
1c0c2f565b Update .travis.yml 2016-02-09 19:34:24 +01:00
Henning Peters
62a6adf33a Update .travis.yml 2016-02-09 19:29:23 +01:00
Henning Peters
c00dd43fe0 add sun data 2016-02-09 16:42:55 +01:00
Henning Peters
116ec3b849 switch to buildbot.json 2016-02-09 16:11:08 +01:00
Henning Peters
8d3957c5e6 switch to buildbot.json 2016-02-09 15:31:55 +01:00
Matthew Honnibal
bc9a31df3e Merge branch 'master' of ssh://github.com/honnibal/spaCy 2016-02-09 14:43:30 +01:00