transformers

Commit Graph

Author	SHA1	Message	Date
Patrick von Platen	73893fc771	[BigBird Pegasus] Make tests faster (#11744 ) * improve tests * remove bogus file * make style Co-authored-by: Patrick von Platen <patrick@huggingface.co>	2021-05-17 06:30:53 -04:00
Michael Benayoun	a0531c8a24	fixed shape issue for T5 tracing (#11742 ) Co-authored-by: Michael Benayoun <michael@huggingface.co>	2021-05-17 06:17:31 -04:00
Julien Chaumond	0fc56df5fb	Add visual + link to Premium Support webpage (#11740 ) * Update README.md * Update index.rst	2021-05-17 05:28:56 -04:00
Julien Chaumond	2f88bd9c4c	Remove tapas model card (#11739 )	2021-05-17 04:42:37 -04:00
Marc van Zee	726e953d44	Improvements to Flax finetuning script (#11727 ) * Add Cloud details to README * Flax script and readme updates * Some simplifications of Flax script	2021-05-17 09:26:33 +01:00
Michael Benayoun	86d5fb0b36	Experimental symbolic tracing feature with torch.fx for BERT, ELECTRA and T5 (#11475 ) Symbolic tracing feature for BERT, ELECTRA and T5 Co-authored-by: Michael Benayoun <michael@huggingface.co> Co-authored-by: Stas Bekman <stas@stason.org> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-05-14 20:57:30 +02:00
Marc van Zee	94a2348706	Add Cloud details to README (#11706 ) * Add Cloud details to README * Flax script and readme updates	2021-05-14 14:51:25 +01:00
Patrick von Platen	113eaa7575	correct example script (#11726 )	2021-05-14 12:02:57 +01:00
Oyvind Tafjord	bd3b599c12	Fix T5 beam search using parallelize (#11717 )	2021-05-14 10:44:03 +01:00
Volodymyr Byno	218d552f30	Fix loading the best model on the last stage of training (#11718 )	2021-05-13 16:11:12 -04:00
Sylvain Gugger	252082001d	Fix v4.6.0 doc	2021-05-13 10:45:28 -04:00
Sylvain Gugger	cbbf49f644	Fix doc deployment	2021-05-13 10:34:14 -04:00
lexhuismans	91cf29153b	[T5] Add 3D attention mask to T5 model (2) (#9643 ) (#11197 ) * Add 3D attention mask to T5 model (#9643) Added code for 3D attention mask in T5 model. Similar to BERT model. * Add test for 3D attention mask Added test for 3D attention mask: test_decoder_model_past_with_3d_attn_mask() 3D attention mask of the shape [Batch_size, Seq_length, Seq_length] both for attention mask and decoder attention mask. Test is passing.	2021-05-13 12:02:27 +01:00
Vasudev Gupta	6ee1a4fd3e	add everything (#11651 )	2021-05-13 11:51:30 +01:00
Patrick von Platen	57b6a80de8	[Flax] Fix BERT initialization & token_type_ids default (#11695 ) * fix some stuff * fix roberta & electra as well * del run bug Co-authored-by: Patrick von Platen <patrick@huggingface.co>	2021-05-13 10:58:19 +01:00
Lysandre Debut	daf0d6a97b	Fix gpt-2 warnings (#11709 )	2021-05-13 03:35:44 -04:00
Philip May	37ed3ab719	Enable option for subword regularization in more tokenizers. (#11417 ) * improve slow class tok usage at xlm rob * add subword regularization for barthez * improve barthez tok. test * fix tokenizer tests * add subword regularization for camembert * add subword regularization for deberta v2 tokenizer * add more doc to deberta v2 tokenizer * add subword regularization for speech to text tok. * fix sp_model_kwargs type in speech 2 text tok. * add subword regularization for M2M100 tok. * add more concrete type hints * fix tests for m2m100 and s2t tok. * add missing Any import * fix syntax error in m2m100 tok. * fix unpickle of m2m100 and s2t tok. * fix test of m2m100 and s2t tok. * improve unpickle of deberta v2 tok. * add test for pickle of barthez & camembert * fix pickle of barthez & camembert * add test for deberta v2 tok. pickle * fix m2m100 tok. pickle * fix s2t tok. pickle * add subword regularization to albert tok. * refactor subword reg. test into TokenizerTesterMixin improve albert tok. test remove sample argument form albert tok. check subword reg. using TokenizerTesterMixin improve tok. tests improve xlm roberta tok. tests improve xlm roberta tok. tests * add subword regularization for big bird t. * improve xlm roberta tok. test * add subword regularization for mbart50 tok. * add subword regularization for pegasus tok. * add subword regularization for reformer tok. * add subword regularization for T5 tok. * fix t5 tok. test formatting * add subword regularization for xlm_proph. tok. * add subword regularization for xlnet tok. * add subword regularization for gert_gen tok. * add typing to tokenizers * add typing to xlm rob. tok * add subword regularization for marian tok. * add reverse tok. test * fix marian tok test * fix marian tok test * fix casing in tok. tests * fix style of tok. common test * fix deberta v2 tok test * add type annotations to tok. tests * add type annotations to tok. __init__ * add typing to kokenizer * add type annotations to tok. __init__ * don't specify the default when it's None * fix barthez tok. doc * move sentencepiece tok. tests to TokenizerTesterMixin * fix unused imports * fix albert tok. test * add comment to sentencepiece test options * fix Any import at big bird tok. * fix Any import at xlm prophetnet tok. * empty commit to trigger CI	2021-05-13 02:44:55 -04:00
NielsRogge	fa84540e98	Vit deit fixes (#11309 ) * Improve docs of DeiT and ViT, add community notebook * Add gitignore for test_samples * Add notebook with Trainer Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-05-12 11:46:02 -04:00
Lysandre	d77eb0cf92	Docs for v4.7.0.dev0	2021-05-12 17:08:35 +02:00
Lysandre	64e78564a5	Release: v4.6.0	2021-05-12 17:03:03 +02:00
Patrick von Platen	fd6204b2a7	[Lazy init] Force fall back to slow init for composite models (#11705 ) * fix encoder-decoder & RAG * finalize * Update src/transformers/models/encoder_decoder/modeling_encoder_decoder.py Co-authored-by: Lysandre Debut <lysandre@huggingface.co> * Update src/transformers/models/rag/modeling_rag.py Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Patrick von Platen <patrick@huggingface.co> Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-05-12 10:52:54 -04:00
Suraj Patil	5c1cda9d3c	fix example in config doc (#11696 )	2021-05-12 09:48:52 -04:00
Philip May	77f4c46b50	remove defaults to None if optional (#11703 )	2021-05-12 09:11:10 -04:00
Marc van Zee	6797cdc077	Updates README and fixes bug (#11701 )	2021-05-12 13:52:52 +01:00
Suraj Patil	f063c56d94	Fix clip docs (#11694 ) * fix doc url * fix example	2021-05-12 15:28:30 +05:30
Suraj Patil	8719afa1ad	CLIP (#11445 ) * begin second draft * fix import, style * add loss * fix embeds, logits_scale, and projection * fix imports * add conversion script * add feature_extractor and processor * style * add tests for tokenizer, extractor and processor * add vision model tests * add weight init * add more tests * fix save_load test * model output, dosstrings, causal mask * config doc * add clip model tests * return dict * bigin integration test * add integration tests * fix-copies * fix init * Clip => CLIP * fix module name * docs * fix doc * output_dim => projection_dim * fix checkpoint names * remoe fast tokenizer file * fix conversion script * fix tests, quality * put causal mask on device * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * fix attribute test * style * address sylvains comments * style * fix docstrings * add qucik_gelu in activations, docstrings * clean-up attention test * fix act fun * fix config * fix torchscript tests * even batch_size * remove comment * fix ouput tu_tuple * fix save load tests * fix add tokens test * add fast tokenizer * update copyright * new processor API * fix docs * docstrings * docs * fix doc * fix doc * fix tokenizer * fix import in doc example * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * check types of config * valhalla => openai * load image using url * fix test * typo Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-05-12 13:48:15 +05:30
Marc van Zee	4ce6bcc310	Adds Flax BERT finetuning example on GLUE (#11564 ) * Adds Flax BERT finetuning example * fix traced jax tensor type * Use Optax losses and learning schedulers * Add 1GPU training results * merge into master & make style * fix input * del file * Fix bug in loss and add torch runs * finish bert flax fine-tune * Update examples/flax/text-classification/README.md * Update examples/flax/text-classification/run_flax_glue.py * add requirements * finalize * finalize Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> Co-authored-by: Patrick von Platen <patrick@huggingface.co>	2021-05-11 19:02:59 +01:00
Sylvain Gugger	f13f1f8fb8	Test checkpointing (#11682 ) * Add test and see where CI is unhappy * Load with strict=False	2021-05-11 12:02:48 -04:00
Julien Plu	d9b286272c	Fix TF Roberta for mixed precision training (#11675 )	2021-05-11 12:01:03 -04:00
Sylvain Gugger	a135f59536	Auto modelcard (#11599 ) * Autogenerate model cards from the Trainer * ModelCard deprecated * Fix test * Style * Apply suggestions from code review Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * Address review comments * Quality * With all metadata * Metadata * Post-merge conflict mess * Data args and all examples * Default license and languages when possible Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2021-05-11 11:30:34 -04:00
Matt	b3429ab678	Grammar and style edits for the frontpage README (#11679 ) * Grammar and style edits for the frontpage README * Going all-in on em-dashes because you only live once * Update README.md Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-05-11 15:49:34 +01:00
nxznm	901153c61e	Fix docstring of description about input_ids (#11672 )	2021-05-11 08:12:02 -04:00
Jonathan Chang	64232bc0df	Add --text_column to run_summarization_no_trainer (#11673 )	2021-05-11 07:58:38 -04:00
Julien Plu	024cd19bb7	Add MacOS TF version (#11674 ) Co-authored-by: Julien Plu <jplu@argos.local>	2021-05-11 05:42:21 -04:00
Pavel Soriano	9120ae7d66	Fixes NoneType exception when topk is larger than one coupled with a small context in the Question-Answering pipeline (#11628 ) * added fix to decode function. added test to qa pipeline tests * completed topk docstring * fixed formatting with black * applied style_doc to fix line length	2021-05-10 13:28:10 -04:00
Patrick von Platen	dcb0e61430	push (#11667 )	2021-05-10 17:38:17 +01:00
Sylvain Gugger	05a930671f	Save scaler state dict when checkpointing (#11663 )	2021-05-10 10:58:30 -04:00
Matt	ef8d32c5ea	Fix suggested by @bhadreshpsavani (#11660 )	2021-05-10 14:28:04 +01:00
Vasudev Gupta	575c979144	Update community.md (#11654 )	2021-05-10 09:48:21 +01:00
Tanmay Laud	f7f872955d	Big Bird Fast Tokenizer implementation (#11075 ) * Added Big Bird Fast Tokenizer initial file * style fixes * flake fixes * Added big bird fast tokenizer to init files * Added big bird fast to Auto tokenization * fix styles * minor quality fixes * Added initial test code * Fix SpmConverter when precompiled_charsmap doesn't exist * fixed post processor * minor style fix * minor fix input names * Actually fix identity normalization * style * Added token type ids to fast tokenizer * style * flake fix * fix copies Co-authored-by: Anthony MOI <m.anthony.moi@gmail.com>	2021-05-10 03:01:23 -04:00
Bhavitvya Malik	80da304a0f	updated user permissions based on umask (#11119 ) * updated user permissions based on umask * updated user permissions based on umask * changes as per suggestions * minor changes	2021-05-10 02:45:29 -04:00
Quentin Lhoest	1a0b41781d	Update requirements.txt (#11634 )	2021-05-10 11:19:52 +05:30
NielsRogge	f785c51692	Update code example (#11631 ) * Update code example * Code review	2021-05-10 11:18:43 +05:30
Tommy Chiang	7e406f4a65	[Examples] Fix invalid links after reorg (#11650 )	2021-05-10 11:16:48 +05:30
Tommy Chiang	f2ffcaf49f	[Examples] Check key exists in datasets first (#11503 )	2021-05-09 15:42:38 -04:00
Stas Bekman	ba0d50f214	[examples] fix sys.path in conftest.py (#11636 ) * restore conftest.py * fix conftest and make copies * remove unneeded parts * remove unwanted files	2021-05-07 14:44:22 -07:00
Stas Bekman	cd9b8d7efe	[self-push CI] sync with self-scheduled (#11637 ) forgot to add the missing `libaio-dev` to this workflow	2021-05-07 14:06:33 -07:00
Lysandre Debut	da37eb8e43	Reduce to 1 worker and set timeout for GPU TF tests (#11633 )	2021-05-07 11:55:20 -04:00
Lysandre Debut	39084ca663	Add the ImageClassificationPipeline (#11598 ) * Add the ImageClassificationPipeline * Code review Co-authored-by: patrickvonplaten <patrick.v.platen@gmail.com> * Have `load_image` at the module level Co-authored-by: patrickvonplaten <patrick.v.platen@gmail.com>	2021-05-07 08:08:40 -04:00
Patrick von Platen	e7bff0aabe	make fix copy (#11627 )	2021-05-07 07:48:51 -04:00

... 5 6 7 8 9 ...

7492 Commits All Branches Search

7492 Commits

All Branches