spaCy

mirror of https://github.com/explosion/spaCy.git synced 2025-11-11 05:19:52 +03:00

Author	SHA1	Message	Date
Ines Montani	ab2342a10e	Merge pull request #1621 from bdewilde/fix-span-orth (resolves #1612 ) Make span.orth_ = span.text, as advertised	2017-11-20 20:53:15 +00:00
Burton DeWilde	a5c6869b2d	Fix bug where span.orth_ != span.text (see #1612 )	2017-11-20 12:05:43 -06:00
Burton DeWilde	635792997c	Add regression test for #1612	2017-11-20 12:05:35 -06:00
Burton DeWilde	833c66c9b2	Add contributor agreement	2017-11-20 11:28:31 -06:00
ines	ec08996000	Add note on tags matching tokenization (see #1613 )	2017-11-20 15:12:47 +01:00
Ines Montani	ac235c0baf	Merge pull request #1620 from cclauss/patch-3 Create cclauss.md	2017-11-20 14:07:36 +00:00
cclauss	31085dcbb6	Create cclauss.md	2017-11-20 14:57:30 +01:00
ines	9a63e32f21	Add noqa to Python 2 compat variables of built-ins (see #1617 )	2017-11-20 14:03:42 +01:00
ines	d70a64d78b	Fix syntax error and formatting in test (see #1617 )	2017-11-20 14:01:25 +01:00
cclauss	2088adb0b7	--exclude=spacy/compat.py,spacy/lang	2017-11-20 14:00:45 +01:00
ines	17849dee4b	Fix French test (see #1617 )	2017-11-20 13:59:59 +01:00
ines	0f6dfb4b81	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-20 13:57:53 +01:00
ines	1a38575de3	Make example Python 2 compatible (see #1617 )	2017-11-20 13:57:51 +01:00
cclauss	6f42b240bc	--exclude=spacy/lang to avoid flake8 infinite loop It is not clear why spacy/lang throws flake8 for a loop.	2017-11-20 12:00:39 +01:00
cclauss	fa7aafd4c7	Use flake8 to look for syntax errors, undefined names	2017-11-20 11:35:28 +01:00
Ines Montani	9eb5cd0b31	Merge pull request #1608 from tokestermw/bug/fix-span-noun-chunks Fixes error when getting `noun_chunks` from `Span`s. (Issue #1207)	2017-11-18 03:23:32 +00:00
ines	4f7e64e371	Update resources	2017-11-18 02:53:00 +01:00
Motoki Wu	7b5b49eef0	added contributor agreement	2017-11-17 17:27:20 -08:00
Motoki Wu	a52e195a0a	Fixes Issue #1207 where `noun_chunks` of `Span` gives an error. Make sure to reference `self.doc` when getting the noun chunks. Same fix as `9750a0128c`	2017-11-17 17:16:20 -08:00
Motoki Wu	b818afaa0e	Added failing test for Issue #1207 . The noun chunk iterator should work for `Doc` but not for `Span`.	2017-11-17 17:04:27 -08:00
ines	c3051e95f7	Add note on attribute extension defaults (resolves #1587 )	2017-11-17 19:14:29 +01:00
ines	954f8cc6d1	Update syntax theme (should move the modifications out to an extension sometime)	2017-11-17 19:13:53 +01:00
Ines Montani	f1a5c33294	Merge pull request #1604 from raphael0202/patch-1 Fix typo in documentation	2017-11-17 17:21:56 +00:00
Raphaël Bournhonesque	a0793fd4cc	Fix typo	2017-11-17 17:57:55 +01:00
Ines Montani	eee9cc41f4	Merge pull request #1602 from MartinoMensio/master (resolves #1599 ) small typo on docs	2017-11-17 16:19:04 +00:00
Martino Mensio	239a0f391d	added contributor agreement	2017-11-17 16:30:09 +01:00
Martino Mensio	ce1aade41e	small typo on docs	2017-11-17 16:20:22 +01:00
ines	7d5afadf5e	Update vectors_loc description	2017-11-17 14:57:11 +01:00
ines	4187cfe1ea	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-17 14:56:29 +01:00
ines	c57e05bec1	Make sure nr_dim is an int In some languages (e.g. Dutch), the nr_dim is extracted as a byte string, causing an error down the line.	2017-11-17 14:56:27 +01:00
Ines Montani	1e3068ec33	Merge pull request #1594 from pavillet/patch-1 Update _spacy.jade	2017-11-16 23:42:59 +00:00
pavillet	ad2935f0c3	Update _spacy.jade Doc example gives 'object is not subscriptable' error. Correcting as an attribuet	2017-11-17 00:02:20 +01:00
ines	a3d4dd1a5d	Test adding of lots of pipeline components (see #1585 ) Just to make sure that there's no error now or in the future with adding a large number of pipeline components.	2017-11-15 17:28:06 +01:00
Roman Domrachev	61d28d03e4	Try again to do selective remove cache	2017-11-15 19:11:12 +03:00
Roman Domrachev	b3311100c7	Merge branch 'master' of github.com:explosion/spaCy	2017-11-15 18:30:04 +03:00
Matthew Honnibal	b60d92aca8	Increment version	2017-11-15 16:14:46 +01:00
Roman Domrachev	505c6a2f2f	Completely cleanup tokenizer cache Tokenizer cache can have be different keys than string That modification can slow down tokenizer and need to be measured	2017-11-15 17:55:48 +03:00
Matthew Honnibal	cf0be62096	Increment version	2017-11-15 15:00:18 +01:00
Matthew Honnibal	716ccbb71e	Require thinc 6.10.1	2017-11-15 14:59:34 +01:00
ines	97a4f9362b	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-15 14:24:00 +01:00
ines	8e65247886	Fix lex.id if vectors is None	2017-11-15 14:23:58 +01:00
Matthew Honnibal	437ad1a852	Merge pull request #1570 from explosion/feature/fix-beam-leak Fix memory leak in beam parser	2017-11-15 14:15:05 +01:00
Matthew Honnibal	2f169fdb0a	Set lex ID correctly for new tokens in Vocab	2017-11-15 13:58:03 +01:00
Matthew Honnibal	fe3c42a06b	Fix caching in tokenizer	2017-11-15 13:55:46 +01:00
Matthew Honnibal	8d692771f6	Improve profiling	2017-11-15 13:51:25 +01:00
Matthew Honnibal	b797dca977	Merge branch 'master' of https://github.com/explosion/spaCy	2017-11-15 13:11:43 +01:00
Ines Montani	9177c7d7aa	Merge pull request #1583 from yogendrasoni/master (resolves #1582 ) Add rstrip after reading line from vec file #1582	2017-11-15 12:02:48 +00:00
ines	c9d72de0fb	Add dummy serialization methods for Japanese and missing lang getter (resolves #1557 )	2017-11-15 12:44:02 +01:00
yogendrasoni	334ed433b2	rstrip line before rsplit loading english fast text giving error because line contains new line at the end and rsplit is splitting it incorrectly	2017-11-15 13:55:08 +05:30
Matthew Honnibal	d274d3a3b9	Let beam forward use minibatches	2017-11-15 00:51:42 +01:00

1 2 3 4 5 ...

7834 Commits