Matthew Honnibal
|
eb3057d806
|
* Add updated unsupervised_train script, from the wsd directory
|
2015-07-06 09:33:00 +02:00 |
|
Matthew Honnibal
|
2e4cfe5255
|
* Add script to train the dictionary-supervised supersense tagger
|
2015-07-06 09:06:22 +02:00 |
|
Matthew Honnibal
|
0be251776e
|
* Supply templates as an argument to the parser Config object
|
2015-07-06 08:44:39 +02:00 |
|
Matthew Honnibal
|
8f0fe1a4ea
|
* Note broken sense data in prepare_treebank
|
2015-07-05 21:04:57 +02:00 |
|
Matthew Honnibal
|
96442d9c3e
|
* Put supersenses.json in the wordnet directory, not in a wsd directory
|
2015-07-05 21:03:59 +02:00 |
|
Matthew Honnibal
|
9534d336ed
|
* Ensure word senses are loaded, even if not in probabilities file
|
2015-07-05 11:31:07 +02:00 |
|
Matthew Honnibal
|
149a901ea7
|
* Don't use POS tags in supersense dict
|
2015-07-05 10:50:22 +02:00 |
|
Matthew Honnibal
|
211058f7a6
|
* Load adverb senses
|
2015-07-05 09:13:22 +02:00 |
|
Matthew Honnibal
|
4c6533a019
|
* Write a supersenses.json fil into a wsd directory in init_model
|
2015-07-04 17:24:32 +02:00 |
|
Matthew Honnibal
|
948ea9333a
|
* Fix alignment of supersenses in init_model
|
2015-07-04 12:24:40 +02:00 |
|
Matthew Honnibal
|
ff1f9fe246
|
* Fix init_model to read supersenses from wordnet, not pre-computed supersenses file
|
2015-07-03 13:28:39 +02:00 |
|
Matthew Honnibal
|
c60cc22390
|
* Ignore adjective supersenses
|
2015-07-03 04:46:11 +02:00 |
|
Matthew Honnibal
|
333e414e9f
|
* Hack prepare_treebank script to load wordnet supersenses
|
2015-07-02 08:31:12 +02:00 |
|
Matthew Honnibal
|
041908a272
|
* Merge neuralnet branch into sense-tagger
|
2015-07-01 22:38:22 +02:00 |
|
Matthew Honnibal
|
62cfcd76fe
|
* Add supersense sets to lexemes, from WordNet. Look-up via lemmatization.
|
2015-07-01 18:48:59 +02:00 |
|
Matthew Honnibal
|
31b5e58aeb
|
* Begin reorganizing neuralnet work
|
2015-06-30 14:26:53 +02:00 |
|
Matthew Honnibal
|
1135cfe50a
|
* Tidy nn_train a bit
|
2015-06-29 16:45:14 +02:00 |
|
Matthew Honnibal
|
df8179ca4f
|
* Add separate Param and AdadeltaParam classes. AdadeltaParam seems broken.
|
2015-06-29 16:39:16 +02:00 |
|
Matthew Honnibal
|
1dff04acb5
|
* Apply regularization to the softmax, not the bias
|
2015-06-29 11:45:38 +02:00 |
|
Matthew Honnibal
|
ca30fe1582
|
* Use He initialization trick
|
2015-06-29 10:56:02 +02:00 |
|
Matthew Honnibal
|
fc34e1b6e4
|
* Move Theano functions into nn_train.py script
|
2015-06-29 07:09:16 +02:00 |
|
Matthew Honnibal
|
fe7b24ecef
|
* whitespace
|
2015-06-28 11:37:17 +02:00 |
|
Matthew Honnibal
|
7b8275fcc4
|
* Wire hyperparameters to script interface
|
2015-06-28 11:37:17 +02:00 |
|
Matthew Honnibal
|
897dd0dd0b
|
* Merge changes, and adjust Example to use memoryview
|
2015-06-28 11:36:11 +02:00 |
|
Matthew Honnibal
|
ef97b90833
|
* Fix token scoring
|
2015-06-28 06:22:18 +02:00 |
|
Matthew Honnibal
|
34c0ef2ee8
|
* Don't compile the orig_arc_eager and tree_arc_eager modules used for the EMNLP paper
|
2015-06-23 05:38:17 +02:00 |
|
Matthew Honnibal
|
59e9f9153c
|
* Remove projectivity constraint in train.py, but raise Exception if non-projective sentence is encountered, since we've told GoldParse to projectivize
|
2015-06-23 05:04:46 +02:00 |
|
Matthew Honnibal
|
839e5038b7
|
* Raise exception on non-projective input
|
2015-06-23 00:01:55 +02:00 |
|
Matthew Honnibal
|
4dad4058c3
|
* Uncomment NER training
|
2015-06-16 23:36:54 +02:00 |
|
Matthew Honnibal
|
5699585278
|
* Use tree_arc_eager system as baseline in experiments
|
2015-06-15 08:23:43 +02:00 |
|
Matthew Honnibal
|
4841f8ad5e
|
* Set transition system early
|
2015-06-15 02:54:12 +02:00 |
|
Matthew Honnibal
|
bcfdf126a4
|
* Add toggle for OrigArcEager system
|
2015-06-14 20:28:14 +02:00 |
|
Matthew Honnibal
|
c500d72dc2
|
* Temporarily disable NER, and wire up the verbose flag during training
|
2015-06-14 17:45:31 +02:00 |
|
Matthew Honnibal
|
ac422492cf
|
* Fix write_parses mode of bin/parser/train.py
|
2015-06-07 19:08:48 +02:00 |
|
Matthew Honnibal
|
4073533e28
|
* Upd munge_ewtb for the new json format
|
2015-06-06 02:10:33 +02:00 |
|
Matthew Honnibal
|
6a1341b29e
|
* Add tb pre-process script
|
2015-06-06 01:59:44 +02:00 |
|
Matthew Honnibal
|
1736fc5a67
|
* Add more options to bin/parser/train
|
2015-06-05 23:49:26 +02:00 |
|
Matthew Honnibal
|
362f87dc3a
|
* Update input corruption method to work with lists as well as trings
|
2015-06-05 19:33:32 +02:00 |
|
Matthew Honnibal
|
0aed9c9a33
|
* Fix train.py
|
2015-06-05 15:50:24 +02:00 |
|
Matthew Honnibal
|
8466600add
|
* Clean up train.py, removing unused tag jackknifing code
|
2015-06-05 15:01:28 +02:00 |
|
Matthew Honnibal
|
e772b48dcd
|
* Skip sentences of length 1 in training
|
2015-06-05 02:29:03 +02:00 |
|
Matthew Honnibal
|
e822df0867
|
* Fix bugs in new greedy/beam parser
|
2015-06-02 02:01:33 +02:00 |
|
Matthew Honnibal
|
70a7ad89ca
|
* Removed unused imports from train.py
|
2015-06-02 00:59:09 +02:00 |
|
Matthew Honnibal
|
a3de20118e
|
* Wire up beam-width command line argument
|
2015-06-02 00:54:12 +02:00 |
|
Matthew Honnibal
|
08044ea70c
|
* Remove try/except around parser.train
|
2015-05-31 15:21:56 +02:00 |
|
Matthew Honnibal
|
c8a553fe91
|
* Fix cluster initialization
|
2015-05-31 15:21:28 +02:00 |
|
Matthew Honnibal
|
d7cc2338e7
|
* Fix bug in train.py
|
2015-05-31 06:49:06 +02:00 |
|
Matthew Honnibal
|
c037f80638
|
* Add case expansion to Brown clusters
|
2015-05-31 05:50:50 +02:00 |
|
Matthew Honnibal
|
5ab0f233a1
|
* Ensure words in Brown clusters make it into the vocab, even if they're not in our probs list
|
2015-05-31 05:46:16 +02:00 |
|
Matthew Honnibal
|
d42dda0372
|
* Shuffle docs before doing jackknife partition --- otherwise we'll not get the right genre mixes...
|
2015-05-31 01:25:02 +02:00 |
|