spaCy

mirror of https://github.com/explosion/spaCy.git synced 2024-11-11 12:18:04 +03:00

Author	SHA1	Message	Date
Renat Shigapov	2e2d0e8701	added spaCyOpenTapioca (#9181 ) * add spaCyOpenTapioca to universe * add agreement * fix misprint in tags	2021-09-11 13:25:25 +09:00
Ines Montani	753149bc88	Update references to contributor agreement [ci skip]	2021-08-31 10:58:22 +02:00
Baltazar	4d85cb88a5	added contribution license	2021-08-19 21:45:18 +02:00
Steele Farnsworth	b18cb1cd2a	Refactor dependencymatcher.pyx to use list comps and enumerate. (#8956 ) * Refactor to use list comps and enumerate. Replace loops that append to a list with a list comprehensions where this does not change the behavior; replace range(len(...)) loops with enumerate. Correct one typo in a comment. Replace a call to set() with a set literal. * Undo double assignment. Expand `tokens_to_key[j] = k = self._get_matcher_key(key, i, j)` to two statements. Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com> * Sign contributors agreement Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com>	2021-08-18 09:55:45 +02:00
Lasse	195e4e48c3	add textdescriptives to universe	2021-08-13 14:35:18 +02:00
Eduard Zorita	439f30faad	Add stub files for main cython classes (#8427 ) * Add stub files for main API classes * Add contributor agreement for ezorita * Update types for ndarray and hash() * Fix __getitem__ and __iter__ * Add attributes of Doc and Token classes * Overload type hints for Span.__getitem__ * Fix type hint overload for Span.__getitem__ Co-authored-by: Luca Dorigo <dorigoluca@gmail.com>	2021-08-07 12:30:03 +02:00
Nick Sorros	0485cdefcc	Add logger debug for project push and pull (#8860 ) * Add logger debug for project push and pull * Sign contributor agreement	2021-08-02 18:13:53 +02:00
Ines Montani	51e5903d6f	Merge pull request #8702 from KennethEnevoldsen/master [ci skip]	2021-07-18 13:18:42 +10:00
Mario Šaško	1ba2e8a646	Add TakeLab/spacy-udpipe to Universe (#8698 ) * Add TakeLab/spacy-udpipe to universe * Add SCA * Sign SCA	2021-07-16 11:15:52 +02:00
jmyerston	993b0fab0e	Added ancient Greek language support (#8606 ) * Add ancient Greek language support Initial commit * Contributor Agreement * grc tokenizer test added and files formatted with black, unnecessary import removed Co-Authored-By: Sofie Van Landeghem <svlandeg@users.noreply.github.com> * Commas in lists fixed. __init__py added to test * Update lex_attrs.py * Update stop_words.py * Update stop_words.py Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com>	2021-07-15 10:27:17 +02:00
KennethEnevoldsen	e5127992a0	added agreement	2021-07-13 10:11:02 +02:00
Edward	8233359225	Fix preservation of spacy package meta (#8663 ) * update package meta with existing_meta and nlp_meta * Add spaCy contributor agreement * Added more info when creating readme	2021-07-12 11:18:52 +02:00
Paul O'Leary McCann	1c70c87daf	Fix autoblack The conditional needs double equals.	2021-07-10 16:02:39 +09:00
Paul O'Leary McCann	b8cdbb4bb6	Make the autoblack job not run on forks The autoblack job is an occasional cleanup job. If it runs on forks and those PRs are accepted the git history will be weird and that doesn't help anyone. The way to make the job not run on forks is a little non-obvious but based on this thread. https://github.com/prisma/prisma/issues/3539	2021-07-10 15:38:20 +09:00
Ines Montani	1c0ed22d1e	Merge pull request #8573 from julien-talkair/code-quality-pre-commit	2021-07-09 23:09:24 +10:00
Sofie Van Landeghem	608fc1d623	avoid msg var impliciteness (#8619 ) * avoid msg var impliciteness * rename local msg * Add CI tests for debug data and train * Adjust debug data CLI test Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com>	2021-07-06 19:08:08 +02:00
Adriane Boyd	5fd0b5207e	Fix vectors check for sourced components (#8559 ) * Fix vectors check for sourced components Since vectors are not loaded when components are sourced, store a hash for the vectors of each sourced component and compare it to the loaded vectors after the vectors are loaded from the `[initialize]` block. * Pop temporary info * Remove stored hash in remove_pipe * Add default for pop * Add additional convert/debug/assemble CLI tests	2021-07-06 12:43:17 +02:00
Yoichiro Hasebe	e541092088	Create yohasebe.md	2021-07-04 08:57:04 +09:00
Ines Montani	c5c4e96597	Fix syntax [ci skip]	2021-07-02 17:46:56 +10:00
Ines Montani	6b905d67df	Try workflow_dispatch and schedule [ci skip]	2021-07-02 17:45:27 +10:00
Ines Montani	70589e348e	Commit as explosion-bot [ci skip]	2021-07-02 17:45:11 +10:00
Ines Montani	dd34a3a433	Try simpler approach [ci skip]	2021-07-02 17:40:49 +10:00
Ines Montani	2898331494	Improve logic [ci skip]	2021-07-02 17:37:35 +10:00
Ines Montani	519a9e29be	Fix git login [ci skip]	2021-07-02 17:30:59 +10:00
Ines Montani	8961f36415	Commit manually in workflow [ci skip]	2021-07-02 17:27:48 +10:00
Ines Montani	2a5cbf1b0c	Test different workflow trigger [ci skip]	2021-07-02 17:22:43 +10:00
Ines Montani	bbbaae0b5e	Update triggers [ci skip]	2021-07-02 17:10:24 +10:00
Ines Montani	cdefb8cf1b	Experimental: add autoblack.yml action [ci skip]	2021-07-02 17:07:05 +10:00
julien-talkair	6b1f9a5be0	add spacy contributor agreement	2021-07-01 17:41:12 +02:00
Ines Montani	88ad41316c	Update issue template [ci skip]	2021-06-28 03:11:37 +02:00
Ines Montani	db6361ab6e	Update issue template [ci skip]	2021-06-28 03:10:52 +02:00
Ines Montani	2e453bda92	Update issue links [ci skip]	2021-06-28 03:09:48 +02:00
Paul O'Leary McCann	0d3caa52a6	Update New Issue choices This uses some new features related to Issue Templates to help direct more people to Discussions. 1. Change the Discussions option to link to Discussions 2. Add a link to the FAQ 3. Disable blank issues	2021-06-27 14:41:33 +09:00
Adrian Zuber	f5aee0bbdf	Raise custom error in EntityLinker when KB is not set (#8442 ) * Raise custom error in EntityLinker when KB is not set * add contributor agreement * Update E1018 error message	2021-06-25 23:04:00 +02:00
Adriane Boyd	172dfec4f2	Test download in CI with ca_core_news_sm (#8493 )	2021-06-24 09:26:30 +02:00
Giovanni Toffoli	19521d525b	Added Italian POS-aware lemmatizer. (#8079 ) * Added Italian POS-aware lemmatizer. Also added the code used to build the lookup tables by POS. * Create gtoffoli.md * Add imports and format * Remove helper script * Use lemma_lookup instead of lemma_lookup_legacy Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com>	2021-06-16 11:14:45 +02:00
Adriane Boyd	33240ed2c5	Temporarily skip model download test	2021-06-16 10:14:42 +02:00
Adriane Boyd	d52ab13b5f	Update CI: update ubuntu image, add download test (#8298 ) * Update CI: update ubuntu image, add download test * Switch instances to `ubuntu-18.04` * Add model download test, currently only for one job with python 3.8 * Fix variable name * Set variables explicitly	2021-06-07 14:46:07 +02:00
Vito De Tullio	3672464e25	applying suggestion to avoid mypy errors (#8265 ) * applying suggestion to avoid mypy errors * sign contributor agreement	2021-06-02 19:25:30 +10:00
Kristian Boda	dc8d8d15d2	Add hmrb to spaCy Universe (#8129 ) * docs: add hmrb to spacy universe * docs: add sentence on spacy versions * docs: update description and images * misc: add spaCy Contributor Agreement	2021-05-31 18:40:48 +10:00
Narayan Acharya	6b79714080	Address missing config overrides post load of models (#8208 )	2021-05-31 18:36:52 +10:00
Julien Salinas	a176d2209a	Sign contributors agreement.	2021-05-14 11:00:27 +02:00
Sevdimali	49aed683cc	Azerbaijani language added (#7911 )	2021-04-28 14:42:02 +02:00
Adriane Boyd	f4080983ea	Extend to cupy 9.0.0 (#7914 )	2021-04-28 10:18:24 +02:00
Janis Klaise	1690595e4d	Update load_lookups return type and docstring (#7907 ) * Update load_lookups return type and docstring * Add contributor agreement	2021-04-27 09:13:39 +02:00
Adriane Boyd	36ecba224e	Set up GPU CI testing (#7293 ) * Set up CI for tests with GPU agent * Update tests for enabled GPU * Fix steps filename * Add parallel build jobs as a setting * Fix test requirements * Fix install test requirements condition * Fix pipeline models test * Reset current ops in prefer/require testing * Fix more tests * Remove separate test_models test * Fix regression 5551 * fix StaticVectors for GPU use * fix vocab tests * Fix regression test 5082 * Move azure steps to .github and reenable default pool jobs * Consolidate/rename azure steps Co-authored-by: svlandeg <sofie.vanlandeghem@gmail.com>	2021-04-22 14:58:29 +02:00
meghanabhange	49ff1126bf	Project Idea : denomme \| Multilingual Name Detection (#7845 ) * Add denomme * spaCy contributor agreement Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com>	2021-04-22 08:48:17 +02:00
Pierre Lison	2f0ef2c9cc	adding skweak to the SpaCy universe	2021-04-22 01:16:34 +02:00
Shantam Raj	6017fcf693	Default code for Setting Entity annotations on the website errors (#7738 ) * the default example for "Setting entity annotations" errors on Binder * updating contributer info * using a new variable to store original entities	2021-04-21 09:16:32 +02:00
broaddeep	ee159b8543	Support match alignments (#7321 ) * Support match alignments * change naming from match_alignments to with_alignments, add conditional flow if with_alignments is given, validate with_alignments, add related test case * remove added errors, utilize bint type, cleanup whitespace * fix no new line in end of file * Minor formatting * Skip alignments processing if as_spans is set * Add with_alignments to Matcher API docs * Update website/docs/api/matcher.md Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com> Co-authored-by: Adriane Boyd <adrianeboyd@gmail.com> Co-authored-by: Sofie Van Landeghem <svlandeg@users.noreply.github.com>	2021-04-08 18:10:14 +10:00

1 2 3 4 5 ...

516 Commits