spaCy

mirror of https://github.com/explosion/spaCy.git synced 2025-07-27 08:29:51 +03:00

Author	SHA1	Message	Date
ines	434030e0d0	Fix requirements.txt example (see #1638 )	2017-11-26 15:53:19 +01:00
Søren Lind Kristiansen	ef03e9ea53	Remove unused import.	2017-11-25 13:04:02 +01:00
Søren Lind Kristiansen	b91986b726	Add contributor agreement.	2017-11-24 15:29:54 +01:00
Søren Lind Kristiansen	6aa241bcec	Add day of month tokenizer exceptions for Danish.	2017-11-24 15:03:24 +01:00
Søren Lind Kristiansen	0c276ed020	Add weekday abbreviations and remove abiguous month abbreviations for Danish.	2017-11-24 14:43:29 +01:00
Søren Lind Kristiansen	056547e989	Add multiple tokenizer exceptions for Danish.	2017-11-24 11:51:26 +01:00
Søren Lind Kristiansen	8dc265ac0c	Add test for tokenization of 'i.' for Danish.	2017-11-24 11:29:37 +01:00
Søren Lind Kristiansen	ac8116510d	Fix tokenization of 'i.' for Danish.	2017-11-24 11:16:53 +01:00
Matthew Honnibal	79f11d4f85	Pickle vectors with vocab	2017-11-23 17:19:50 +01:00
ines	726fb2d0b5	Use fewer iterations by default to avoid overfitting on blank model (resolves #1632 )	2017-11-23 15:27:12 +01:00
Matthew Honnibal	f29c3925ee	Fix more efficient nonproj	2017-11-23 12:48:00 +00:00
Matthew Honnibal	e10e9ad2c5	Improve efficiency of Doc.to_array	2017-11-23 12:33:27 +00:00
Matthew Honnibal	2acc907d55	Improve profiling	2017-11-23 12:33:03 +00:00
Matthew Honnibal	fa62427300	Remove lookup-based lemmatization	2017-11-23 12:32:22 +00:00
Matthew Honnibal	fb26b2cb12	Use lookup lemmatizer if lemma unset	2017-11-23 12:31:58 +00:00
Matthew Honnibal	db5c714ad2	Improve efficiency of deprojectivization	2017-11-23 12:31:34 +00:00
Matthew Honnibal	8fec7268eb	Move string cleanup under a setting flag	2017-11-23 12:19:18 +00:00
Matthew Honnibal	5949777b12	Fix misleading multi-threading docstring	2017-11-23 12:18:59 +00:00
Matthew Honnibal	542e6fd4ea	Don't remove entries from specials	2017-11-23 12:17:42 +00:00
Matthew Honnibal	30ba81f881	Merge pull request #1576 from ligser/master Actually reset caches in pipe [wip]	2017-11-23 12:54:48 +01:00
Matthew Honnibal	4988eeb18a	Merge pull request #1631 from markulrich/patch-1 Use local parameter in example MyComponent	2017-11-23 11:47:50 +01:00
Matthew Honnibal	6bc9917a0e	Another small fix to component docs	2017-11-23 11:47:20 +01:00
markulrich	c9b63c0dfc	Use correct local parameter in example MyComponent (and added markulrich.md contributor file)	2017-11-22 15:59:08 -08:00
ines	c90fe92e15	Fix displaCy test	2017-11-22 05:04:39 +01:00
ines	42ceece110	Add Appveyor badge	2017-11-22 04:20:32 +01:00
ines	a6f33ac27d	Fix displaCy test	2017-11-22 04:19:28 +01:00
ines	93b0be611a	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-22 00:28:55 +01:00
ines	60b4915569	Use .pos_ instead of .tags_ in displaCy by default (see #1006 )	2017-11-22 00:28:52 +01:00
Vadim Mazaev	81314f8659	Fixed tokenizer: added char classes; added first lemmatizer and tokenizer tests	2017-11-21 22:23:59 +03:00
Vadim Mazaev	52ee1f9bf9	Updated Russian Language, added lemmatizer, norm exceptions and lex attrs	2017-11-21 11:44:46 +03:00
Ines Montani	ab2342a10e	Merge pull request #1621 from bdewilde/fix-span-orth (resolves #1612 ) Make span.orth_ = span.text, as advertised	2017-11-20 20:53:15 +00:00
Burton DeWilde	a5c6869b2d	Fix bug where span.orth_ != span.text (see #1612 )	2017-11-20 12:05:43 -06:00
Burton DeWilde	635792997c	Add regression test for #1612	2017-11-20 12:05:35 -06:00
Burton DeWilde	833c66c9b2	Add contributor agreement	2017-11-20 11:28:31 -06:00
ines	ec08996000	Add note on tags matching tokenization (see #1613 )	2017-11-20 15:12:47 +01:00
Ines Montani	ac235c0baf	Merge pull request #1620 from cclauss/patch-3 Create cclauss.md	2017-11-20 14:07:36 +00:00
cclauss	31085dcbb6	Create cclauss.md	2017-11-20 14:57:30 +01:00
ines	9a63e32f21	Add noqa to Python 2 compat variables of built-ins (see #1617 )	2017-11-20 14:03:42 +01:00
ines	d70a64d78b	Fix syntax error and formatting in test (see #1617 )	2017-11-20 14:01:25 +01:00
cclauss	2088adb0b7	--exclude=spacy/compat.py,spacy/lang	2017-11-20 14:00:45 +01:00
ines	17849dee4b	Fix French test (see #1617 )	2017-11-20 13:59:59 +01:00
ines	0f6dfb4b81	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-20 13:57:53 +01:00
ines	1a38575de3	Make example Python 2 compatible (see #1617 )	2017-11-20 13:57:51 +01:00
cclauss	6f42b240bc	--exclude=spacy/lang to avoid flake8 infinite loop It is not clear why spacy/lang throws flake8 for a loop.	2017-11-20 12:00:39 +01:00
cclauss	fa7aafd4c7	Use flake8 to look for syntax errors, undefined names	2017-11-20 11:35:28 +01:00
Felix Sonntag	33b0f86de3	Changed tokenizer to add infix when infix_start is offset	2017-11-19 16:32:10 +01:00
Felix Sonntag	8be3392302	Added regression text for 1494	2017-11-19 16:30:35 +01:00
Felix Sonntag	ada4712250	Add contributer aggreement	2017-11-19 16:30:35 +01:00
Ines Montani	9eb5cd0b31	Merge pull request #1608 from tokestermw/bug/fix-span-noun-chunks Fixes error when getting `noun_chunks` from `Span`s. (Issue #1207)	2017-11-18 03:23:32 +00:00
ines	4f7e64e371	Update resources	2017-11-18 02:53:00 +01:00

... 17 18 19 20 21 ...

8727 Commits