Matthew Honnibal
78aba46530
Update feature/lemmatizer from develop
2019-03-10 02:45:33 +01:00
Matthew Honnibal
5431c47b91
Refactor morphology slightly
2019-03-10 00:59:51 +00:00
Matthew Honnibal
0f12082465
Refactor morphologizer
2019-03-09 22:54:59 +00:00
Matthew Honnibal
41a3016019
Refactor morphologizer class map
2019-03-09 20:55:33 +01:00
Matthew Honnibal
ce1fe8a510
Add comment
2019-03-09 17:51:17 +00:00
Matthew Honnibal
28c26e212d
Fix textcat model for GPU
2019-03-09 17:50:08 +00:00
Ines Montani
610fb306bd
Revert hyphens
2019-03-09 12:51:53 +01:00
Matthew Honnibal
f742900f83
Set pos attribute in morphologizer
2019-03-09 11:51:11 +00:00
Matthew Honnibal
a6d153b0a0
Add UPOS as morphological field in ud_train
2019-03-09 11:50:50 +00:00
Matthew Honnibal
bba5f57f91
Add method to export utf8 array to Doc
2019-03-09 11:50:27 +00:00
Matthew Honnibal
e1a83d15ed
Add support for character features to Tok2Vec
2019-03-09 11:50:08 +00:00
Matthew Honnibal
eae384ebb2
Add POS to morphological fields
2019-03-09 11:49:44 +00:00
Ines Montani
bbabb6aaae
Escape more hyphens
2019-03-09 12:41:05 +01:00
Ines Montani
b8db219850
Auto-format
2019-03-09 12:40:58 +01:00
Ines Montani
a145bfe627
Try escaping hyphens again
2019-03-09 03:06:50 +01:00
Ines Montani
b9c71fc0f0
Fix flags
2019-03-09 02:46:04 +01:00
Ines Montani
ae09b6a6cf
Try fixing unicode inconsistencies on Python 2
2019-03-09 02:37:50 +01:00
Ines Montani
d957d7a697
Auto-format
2019-03-09 02:37:41 +01:00
Ines Montani
65402c3d02
Revert "Experiment with escaping hyphens"
...
This reverts commit 9b42e2d5dd
.
2019-03-09 02:13:00 +01:00
Ines Montani
9b42e2d5dd
Experiment with escaping hyphens
2019-03-09 02:05:26 +01:00
Matthew Honnibal
b6d60d0041
Merge branch 'feature/lemmatizer' of https://github.com/explosion/spaCy into feature/lemmatizer
2019-03-09 00:41:53 +00:00
Matthew Honnibal
4c8730526b
Filter bad retokenizations
2019-03-09 00:41:34 +00:00
Matthew Honnibal
42bc3ad73b
Fix class mapping for morphologizer
2019-03-09 00:20:29 +00:00
Matthew Honnibal
c4df89ab90
Fixes for morphologizer
2019-03-09 00:20:11 +00:00
Ines Montani
76764fcf59
💫 Improve converters and training data file formats ( #3374 )
...
* Populate converter argument info automatically
* Add conversion option for msgpack
* Update docs
* Allow reading training data from JSONL
2019-03-08 23:15:23 +01:00
Matthew Honnibal
cc2b2dba14
Neaten set_morphology option on Tagger
2019-03-08 19:16:02 +01:00
Matthew Honnibal
afa227e25b
Fix setter
2019-03-08 19:10:01 +01:00
Matthew Honnibal
b27bd42613
Fix compile error
2019-03-08 19:06:02 +01:00
Matthew Honnibal
27886d626f
Dont set morphology in Tagger for ud_train
2019-03-08 19:03:31 +01:00
Matthew Honnibal
c91577db02
Add set_morphology cfg option for Tagger
2019-03-08 19:03:17 +01:00
Matthew Honnibal
49cf002ac4
Add missing import
2019-03-08 18:59:25 +01:00
Matthew Honnibal
09b26f5e2e
Fix compile error
2019-03-08 18:58:26 +01:00
Matthew Honnibal
d7ec1d62cb
Fix Morphologizer
2019-03-08 18:54:25 +01:00
Matthew Honnibal
3908911da4
Fix import
2019-03-08 17:04:14 +01:00
Matthew Honnibal
8a9181d95a
Merge __init__
2019-03-08 16:58:42 +01:00
Matthew Honnibal
4cf897e8e1
Update from develop
2019-03-08 16:56:54 +01:00
Ines Montani
ad834be494
Tidy up and auto-format
2019-03-08 13:28:53 +01:00
Ines Montani
d260aa17fd
Merge branch 'develop' into feature/lemmatizer
2019-03-08 13:25:00 +01:00
Ines Montani
296446a1c8
Tidy up and improve docs and docstrings ( #3370 )
...
<!--- Provide a general summary of your changes in the title. -->
## Description
* tidy up and adjust Cython code to code style
* improve docstrings and make calling `help()` nicer
* add URLs to new docs pages to docstrings wherever possible, mostly to user-facing objects
* fix various typos and inconsistencies in docs
### Types of change
enhancement, docs
## Checklist
<!--- Before you submit the PR, go over this checklist and make sure you can
tick off all the boxes. [] -> [x] -->
- [x] I have submitted the spaCy Contributor Agreement.
- [x] I ran the tests, and all new and existing tests passed.
- [x] My changes don't require a change to the documentation, or if they do, I've added all required information.
2019-03-08 11:42:26 +01:00
Matthew Honnibal
19e6b39786
Test morphological features
2019-03-08 01:38:54 +01:00
Matthew Honnibal
9dceb97570
Extend morphanalysis API
2019-03-08 01:38:34 +01:00
Matthew Honnibal
322b64dca0
Allow lookup of morphology by attribute name
2019-03-08 01:38:15 +01:00
Matthew Honnibal
3c32590243
Add test for morph analysis
2019-03-08 00:10:07 +01:00
Matthew Honnibal
3300e3d7ab
Implement more MorphAnalysis API
2019-03-08 00:09:16 +01:00
Matthew Honnibal
9a2d1cc6e0
Add length attribute to MorphAnalysisC
2019-03-08 00:08:57 +01:00
Matthew Honnibal
b5f2b7b454
Add list_features() helper, clean up
2019-03-08 00:08:35 +01:00
Ines Montani
daaeeb7a2b
Merge branch 'master' into develop
2019-03-07 22:07:31 +01:00
Matthew Honnibal
a40d73cb2a
Build out morphological analysis API
2019-03-07 21:59:25 +01:00
Matthew Honnibal
dd9ea478c5
Fix intify_attrs function for obsolete data
2019-03-07 21:59:03 +01:00
Matthew Honnibal
987ee6e884
Fix data reading in morphology
2019-03-07 21:58:43 +01:00