transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-27 00:09:00 +06:00

Author	SHA1	Message	Date
Stas Bekman	78f5fe1416	[Deepspeed] adapt multiple models, add zero_to_fp32 tests (#12477 ) * zero_to_fp32 tests * args change * remove unnecessary work * use transformers.trainer_utils.get_last_checkpoint * document the new features * cleanup * wip * fix fsmt * add bert * cleanup * add xlm-roberta * electra works * cleanup * sync * split off the model zoo tests * cleanup * cleanup * cleanup * cleanup * reformat * cleanup * casing * deepspeed>=0.4.3 * adjust distilbert * Update docs/source/main_classes/deepspeed.rst Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * style Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-07-13 12:07:32 -07:00
Matt	65bf05cd18	Adding TF translation example (#12667 ) * Adding TF translation example * Fixes and style pass for TF translation example * Remove unused postprocess_text copied from run_summarization * Adding README * Review fixes * Move changes to model.config to after we've initialized the model	2021-07-13 19:08:25 +01:00
Patrick von Platen	cee2d2135f	[Flax Generation] Correct inconsistencies PyTorch/Flax (#12662 ) * fix_torch_device_generate_test * remove @ * correct greedy search * save intertmed * add final logits bias * correct * up * add more tests * fix another bug * finish tests * finish marian tests * up Co-authored-by: Patrick von Platen <patrick@huggingface.co>	2021-07-13 18:53:30 +01:00
Stas Bekman	7a22a02a70	[tokenizer.prepare_seq2seq_batch] change deprecation to be easily actionable (#12669 ) * change deprecation to be easily actionable * Update src/transformers/tokenization_utils_base.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * rework as suggested * one warning together * fix format Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-07-13 09:19:04 -07:00
qqaatw	711d901c49	Fix minor docstring typos. (#12682 )	2021-07-13 12:08:15 -04:00
Sylvain Gugger	90178b0cef	Add option to load a pretrained model with mismatched shapes (#12664 ) * Add option to load a pretrained model with mismatched shapes * Fail at loading when mismatched shapes in Flax * Fix tests * Update src/transformers/modeling_flax_utils.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * Address review comments Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2021-07-13 10:15:15 -04:00
Patrick von Platen	7f6d375029	[Blenderbot] Fix docs (#12227 ) * fix_torch_device_generate_test * remove @ * fix docs	2021-07-13 14:17:31 +01:00
Jeroen Steggink	9519f0cd63	Wrong model is used in example, should be character instead of subword model (#12676 ) * Wrong model is used, should be character instead of subword In the original Google repo for CANINE there was mixup in the model names in the README.md, which was fixed 2 weeks ago. Since this transformer model was created before, it probably resulted in wrong use in this example. s = subword, c = character * canine.rst style fix * Update docs/source/model_doc/canine.rst Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Styling canine.rst * Added links to model cards. * Fixed links to model cards. Co-authored-by: Jeroen Steggink <978411+jsteggink@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-07-13 08:40:27 -04:00
Nick Doiron	5803a2a7ac	Add ByT5 option to example run_t5_mlm_flax.py (#12634 ) * Allow ByT5 type in Flax T5 script * use T5TokenizerFast * change up tokenizer config * model_args * reorder imports * Update run_t5_mlm_flax.py	2021-07-13 13:39:57 +01:00
Lysandre Debut	9da1acaea2	*encode_plus() shouldn't run for W2V2CTC (#12655 ) *encode_plus() shouldn't run for W2V2CTC Typo	2021-07-13 06:31:56 -04:00
Lysandre Debut	a6938c4721	Patch BigBird tokenization test (#12653 )	2021-07-13 02:53:06 -04:00
Omar Sanseviero	c523b241c2	Update timeline for Flax event evaluation	2021-07-12 21:24:58 +02:00
Kevin Canwen Xu	dc06e43580	Fix typo in README_zh-hans.md (#12663 )	2021-07-13 01:50:12 +08:00
Kevin Canwen Xu	9d771c5472	Translate README.md to Simplified Chinese (#12596 ) * README Translation for Chinese (Simplified) * update link * h3->h4 * html refactor * update model list * fix * Add a translation guide * format * update * typo * Refine wording	2021-07-13 01:19:54 +08:00
Philip May	21a81c1e3c	fix typo in modeling_t5.py docstring (#12640 )	2021-07-12 12:24:32 -04:00
Ahmed Khaled	b90d499372	fixed docs (#12646 )	2021-07-12 12:03:13 -04:00
Philipp Schmid	da0e9ee697	remove documentation (#12657 )	2021-07-12 18:02:51 +02:00
Lysandre Debut	b189226e8c	Fix transfo xl integration test (#12652 ) * Cleanup test * Skip TF TransfoXL test	2021-07-12 11:51:35 -04:00
Lysandre Debut	fd41e2daf4	Pipeline should be agnostic (#12656 )	2021-07-12 11:42:59 -04:00
Sylvain Gugger	9b3aab2cce	Pickle auto models (#12654 ) * PoC, it pickles! * Remove old method. * Apply to every auto object	2021-07-12 11:15:54 -04:00
Matt	379f649434	TF summarization example (#12617 ) * Adding a TF summarization example * Style pass * Style fixes * Updates for review comments * Adding README * Style pass * Remove unused import	2021-07-12 15:58:38 +01:00
Sylvain Gugger	0f43e742d9	Fix typo	2021-07-12 10:32:51 -04:00
Sylvain Gugger	9adff7a0f4	Fix syntax in conda file	2021-07-12 09:57:54 -04:00
Sylvain Gugger	ad42054278	Minimum requirement for pyyaml	2021-07-12 09:55:36 -04:00
Lysandre Debut	fb5665b5ad	The extended trainer tests should require torch (#12650 )	2021-07-12 09:47:05 -04:00
Lysandre Debut	0af8579bbe	Skip TestMarian_MT_EN (#12649 ) * Skip TestMarian_MT_EN * Skip EN_ZH and EN_ROMANCE * Skip EN_ROMANCE pipeline	2021-07-12 09:11:32 -04:00
Lewis Bails	a882b9facb	Add tokenizer_file parameter to PreTrainedTokenizerFast docstring (#12624 ) Co-authored-by: Lewis Bails <Lewis.Bails@infomedia.dk>	2021-07-12 07:51:58 -04:00
Suraj Patil	f8f9a679a0	fix type check (#12638 )	2021-07-12 10:48:43 +01:00
Eduardo Gonzalez Ponferrada	2dd9440d08	Point to the right file for hybrid CLIP (#12599 )	2021-07-12 12:16:22 +05:30
Bhadresh Savani	de23ecea36	added test file (#12630 )	2021-07-12 12:15:14 +05:30
Stas Bekman	9ee66adadb	fix anchor (#12620 )	2021-07-09 18:48:28 -07:00
Stas Bekman	0dcc3c86e4	[doc] DP/PP/TP/etc parallelism (#12524 ) * wip * complete the doc * missing img * improve * correction * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-07-09 17:39:09 -07:00
Stas Bekman	4cdbf63c03	[debugging utils] minor doc improvements (#12525 )	2021-07-09 17:38:28 -07:00
Will Rice	fb65f65ea6	Add TFHubertModel (#12206 ) * TFHubert * Update with TFWav2Vec Bug Fixes * Add OOV Error * Feedback changes * Fix kwargs call	2021-07-09 18:55:25 +01:00
Patrick von Platen	934222e3c5	[FLax] Fix marian docs 2 (#12615 ) * fix_torch_device_generate_test * remove @ * up	2021-07-09 18:28:57 +01:00
Patrick von Platen	165606e5b4	[Flax Marian] Add marian flax example (#12614 ) * fix_torch_device_generate_test * remove @ * finish better examples for marian flax	2021-07-09 18:01:58 +01:00
Patrick von Platen	51eb6d3457	[Flax] Fix mt5 auto (#12612 ) * fix_torch_device_generate_test * remove @ * fix mt5 auto	2021-07-09 17:33:04 +01:00
Alex Hedges	e7f33e8cb3	Pass `model_kwargs` when loading a model in `pipeline()` (#12449 ) * Pass model_kwargs when loading a model in pipeline * Add test for model_kwargs parameter of pipeline() * Rewrite test to not download model * Fix failing style checks	2021-07-09 09:24:55 -04:00
Sylvain Gugger	18ca59e1d3	Fix arg count for partial functions (#12609 )	2021-07-09 09:24:42 -04:00
Sylvain Gugger	0cc2dc2456	Simplify unk token (#12582 ) * Base test * More test * Fix mistake * Add a docstring change * Add doc ignore * Simplify logic for unk token in Unigram tokenizers * Remove changes from otehr branch	2021-07-09 09:02:34 -04:00
Patrick von Platen	deecdd4939	[Flax] Fix cur step flax examples (#12608 ) * fix_torch_device_generate_test * remove @ * fix save problem	2021-07-09 13:51:28 +01:00
Patrick von Platen	65e27215ba	[Flax] Add flax marian (#12595 ) * fix_torch_device_generate_test * remove @ * add marian * finish make style * add model * add docs * add test * add integration tests * up * solve bug * correct tests * correct some tests * Apply suggestions from code review Co-authored-by: Suraj Patil <surajp815@gmail.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * correct adapt marian * finish Co-authored-by: Patrick von Platen <patrick@huggingface.co> Co-authored-by: Suraj Patil <surajp815@gmail.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-07-09 11:42:13 +01:00
Nicolas Patry	cc12e1dbf6	This will reduce "Already borrowed error": (#12550 ) * This will reduce "Already borrowed error": Original issue https://github.com/huggingface/tokenizers/issues/537 The original issue is caused by transformers calling many times mutable functions on the rust tokenizers. Rust needs to guarantee that only 1 agent has a mutable reference to memory at a given time (for many reasons which don't need explaining here). Usually, the rust compiler can guarantee that this property is true at compile time. Unfortunately, this is impossible for Python to do that, so PyO3, the bridge between rust and python used by `tokenizers`, will change the compile guarantee for a dynamic guarantee, so if multiple agents try to have multiple mutable borrows at the same time, then the runtime will yell with "Already borrowed". The proposed fix here in transformers, is simply to reduce the actual number of calls that really need mutable borrows. By reducing them, we reduce the risk of running into "Already borrowed" error. The caveat is now we add a call to read the current configuration of the `_tokenizer`, so worst case we have 2 calls instead of 1, and best case we simply have 1 + a Python comparison of a dict (should be negligible). * Adding a test. * trivial error :(. * Update tests/test_tokenization_fast.py Co-authored-by: SaulLu <55560583+SaulLu@users.noreply.github.com> * Adding reference to original issues in the tests. * Update the tests with fast tokenizer. Co-authored-by: SaulLu <55560583+SaulLu@users.noreply.github.com>	2021-07-09 09:36:05 +02:00
Omar Sanseviero	8fe836af5a	Add Flax sprint project evaluation section (#12592 )	2021-07-09 08:52:30 +02:00
Stas Bekman	ce111feed1	[doc] fix broken ref (#12597 )	2021-07-08 14:11:01 -07:00
Stas Bekman	f0dde60127	[model.from_pretrained] raise exception early on failed load (#12574 ) * [model.from_pretrained] raise exception early on failed load Currently if `load` pretrained weights fails in `from_pretrained`, we first print a whole bunch of successful messages and then fail - this PR puts the exception first to avoid all the misleading messages. * style Co-authored-by: Suraj Patil <surajp815@gmail.com>	2021-07-08 08:17:51 -07:00
Sylvain Gugger	75e63dbf70	Fix MT5 init (#12591 )	2021-07-08 11:12:18 -04:00
Nicolas Patry	4da568c152	Fixing the pipeline optimization by reindexing targets (V2) (#12330 ) * Fixing the pipeline optimization by rescaling the logits first. * Add test for target equivalence Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>	2021-07-08 16:58:15 +02:00
Funtowicz Morgan	2aa3cd935d	[RFC] Laying down building stone for more flexible ONNX export capabilities (#11786 ) * Laying down building stone for more flexible ONNX export capabilities * Ability to provide a map of config key to override before exporting. * Makes it possible to export BART with/without past keys. * Supports simple mathematical syntax for OnnxVariable.repeated * Effectively apply value override from onnx config for model * Supports export with additional features such as with-past for seq2seq * Store the output path directly in the args for uniform usage across. * Make BART_ONNX_CONFIG_* constants and fix imports. * Support BERT model. * Use tokenizer for more flexibility in defining the inputs of a model. * Add TODO as remainder to provide the batch/sequence_length as CLI args * Enable optimizations to be done on the model. * Enable GPT2 + past * Improve model validation with outputs containing nested structures * Enable Roberta * Enable Albert * Albert requires opset >= 12 * BERT-like models requires opset >= 12 * Remove double printing. * Enable XLM-Roberta * Enable DistilBERT * Disable optimization by default * Fix missing setattr when applying optimizer_features * Add value field to OnnxVariable to define constant input (not from tokenizers) * Add T5 support. * Simplify model type retrieval * Example exporting token_classification pipeline for DistilBERT. * Refactoring to package `transformers.onnx` * Solve circular dependency & __main__ * Remove unnecessary imports in `__init__` * Licences * Use @Narsil's suggestion to forward the model's configuration to the ONNXConfig to avoid interpolation. * Onnx export v2 fixes (#12388) * Tiny fixes Remove `convert_pytorch` from onnxruntime-less runtimes Correct reference to model * Style * Fix Copied from * LongFormer ONNX config. * Removed optimizations * Remvoe bad merge relicas. * Remove unused constants. * Remove some deleted constants from imports. * Fix unittest to remove usage of PyTorch model for onnx.utils. * Fix distilbert export * Enable ONNX export test for supported model. * Style. * Fix lint. * Enable all supported default models. * GPT2 only has one output * Fix bad property name when overriding config. * Added unittests and docstrings. * Disable with_past tests for now. * Enable outputs validation for default export. * Remove graph opt lvls. * Last commit with on-going past commented. * Style. * Disabled `with_past` for now * Remove unused imports. * Remove framework argument * Remove TFPreTrainedModel reference * Add documentation * Add onnxruntime tests to CircleCI * Add test * Rename `convert_pytorch` to `export` * Use OrderedDict for dummy inputs * WIP Wav2Vec2 * Revert "WIP Wav2Vec2" This reverts commit f665efb04c92525c3530e589029f0ae7afdf603e. * Style * Use OrderedDict for I/O * Style. * Specify OrderedDict documentation. * Style :) Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr> Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-07-08 10:54:42 -04:00
Sylvain Gugger	0085e712dd	Don't stop at num_epochs when using IterableDataset (#12561 )	2021-07-08 07:24:46 -04:00

... 24 25 26 27 28 ...

8821 Commits