transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-31 02:02:21 +06:00

Author	SHA1	Message	Date
Nicolas Patry	d4be498441	Optimizing away the `fill-mask` pipeline. (#12113 ) * Optimizing away the `fill-mask` pipeline. - Don't send anything to the tokenizer unless needed. Vocab check is much faster - Keep BC by sending data to the tokenizer when needed. User handling warning messages will see performance benefits again - Make `targets` and `top_k` work together better `top_k` cannot be higher than `len(targets)` but can be smaller still. - Actually simplify the `target_ids` in case of duplicate (it can happen because we're parsing raw strings) - Removed useless code to fail on empty strings. It works only if empty string is in first position, moved to ignoring them instead. - Changed the related tests as only the tests would fail correctly (having incorrect value in first position) * Make tests compatible for 2 different vocabs... (at the price of a warning). Co-authored-by: @EtaoinWu * ValueError working globally * Update src/transformers/pipelines/fill_mask.py Co-authored-by: Lysandre Debut <lysandre@huggingface.co> * `tokenizer.vocab` -> `tokenizer.get_vocab()` for more compatiblity + fallback. Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-06-23 10:38:04 +02:00
Kevin Canwen Xu	037e466b10	Add CodeCarbon Integration (#12304 ) * Add optional dependency * Add CodeCarbon integration * Add CodeCarbon integration * Add CodeCarbon integration * typo	2021-06-23 14:53:09 +08:00
Stas Bekman	bfd5da8e28	[docs] performance (#12258 ) * initial performance document * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Lysandre Debut <lysandre@huggingface.co> * rewrites based on suggestions * 8x multiple is for AMP only * add contribute section Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-06-22 15:34:19 -07:00
Sylvain Gugger	1562c04e41	FlaxBartPretrainedModel -> FlaxBartPreTrainedModel (#12313 )	2021-06-22 16:37:05 -04:00
Stas Bekman	ebe5413589	[trainer] 2 bug fixes and a rename (#12309 ) * bug fixes and a rename * add extended DDP test	2021-06-22 11:13:23 -07:00
Patrick von Platen	64029abe4c	[Flax] Main doc for event orga (#12305 ) * fix_torch_device_generate_test * remove @ * push * finish * some typos * add more info on communication * add suggestions	2021-06-22 18:02:52 +01:00
Kilian Kluge	032d56a435	Fix and improve documentation for LEDForConditionalGeneration (#12303 ) * Replace conditional generation example (fixes #12268) * Replace model in summarization example with finetuned checkpoint, adapt example text * Fix typo in new summarization example * Fix docstring formatting, add missing import statement to example	2021-06-22 09:58:13 -04:00
Suraj Patil	1498eb9888	add FlaxAutoModelForImageClassification in main init (#12298 )	2021-06-22 18:26:05 +05:30
Stefan Schweter	2affeb2905	trainer_tf: adjust wandb installation command (#12291 )	2021-06-22 08:47:31 -04:00
Hamid Shojanazeri	af6e01c5bc	Fix for the issue of device-id getting hardcoded for token_type_ids during Tracing [WIP] (#11252 ) * registering a buffer for token_type_ids, to pass the error of device-id getting hardcoded when tracing * sytle format * adding persistent flag to the resgitered buffers that prevent from adding them to the state_dict and addresses the Backward compatibility issue * adding the try catch to the fix as persistent flag is only available from PT >1.6 * adding version check * added the condition to only use the token_type_ids buffer when its autogenerated not passed by user * adding comments and making the conidtion where token_type_ids are None to use the registered buffer * taking out position-embeddding from the if block * adding comments * handling the case if buffer for position_ids was not registered * reverted the changes on position_ids, fix the issue with size of token_type_ids buffer, moved the modification for generated token_type_ids to Bertmodel, instead of Embeddings * reverting the token_type_ids in case of None to the previous version * reverting changes on position_ids adding back the if block * changes added by running make fix-copies * changes added by running make fix-copies and added the import version as it was getting used * changes added by running make fix-copies * changes added by running make fix-copies * fixing the import format * fixing the import format * modified to use temp tensor for trimed and expanded token_type_ids buffer * changes made by fix-copies after temp tensor modifications * changes made by fix-copies after temp tensor modifications * changes made by fix-copies after temp tensor modifications * clean up * clean up * clean up * clean up * Nit * Nit * Nit * modified according to support device conversion on traced models * modified according to support device conversion on traced models * modified according to support device conversion on traced models * modified according to support device conversion on traced models * changes based on latest in master * Adapt templates * Add version import Co-authored-by: Ubuntu <ubuntu@ip-172-31-32-81.us-west-2.compute.internal> Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>	2021-06-22 05:21:30 -04:00
Stas Bekman	0d97ba8a98	[tests] multiple improvements (#12294 ) * [tests] multiple improvements * cleanup * style * todo to investigate * fix	2021-06-21 19:51:36 -07:00
Stas Bekman	dad414d5f9	[trainer + examples] set log level from CLI (#12276 ) * set log level from CLI * add log_level_replica + test + extended docs * cleanup * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * rename datasets objects to allow datasets module * improve the doc * style * doc improve Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-06-21 19:30:50 -07:00
Stas Bekman	a4ed074d4b	reset report_to to none, avoid deprecation warning (#12293 )	2021-06-21 16:50:12 -07:00
Patrick von Platen	7ef309ca10	[Flax] Add jax flax to env command (#12251 ) * fix_torch_device_generate_test * remove @ * add commands for flax/jax	2021-06-21 17:12:12 +01:00
Matt	e3cb7a0b60	Tensorflow QA example (#12252 ) * New Tensorflow QA example! * Style pass * Updating README.md for the new example * flake8 fixes * Update examples/tensorflow/question-answering/README.md Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-06-21 16:37:28 +01:00
Patrick von Platen	4e9a6796c7	[Flax] Fix flax test save pretrained (#12256 ) * fix_torch_device_generate_test * remove @ * fix flax save pretrained test	2021-06-21 16:37:13 +01:00
Stas Bekman	b75b5605c9	[DeepSpeed] don't ignore --adafactor (#12257 )	2021-06-21 08:17:00 -07:00
Suraj Patil	eb881674f2	[Flax] [WIP] allow loading head model with base model weights (#12255 ) * boom boom * remove flax clip example * allow loading head model with base model weights * add test * fix imports * disable save, load test for clip * add test_save_load_to_base	2021-06-21 15:56:42 +01:00
Suraj Patil	8d5b7f36e5	[FlaxClip] fix test from/save pretrained test (#12284 ) * boom boom * remove flax clip example * fix from_save_pretrained	2021-06-21 15:54:34 +01:00
Vishal Burman	b53bc55ba9	Fix for making student ProphetNet for Seq2Seq Distillation (#12130 ) * make_student.py: fix to make student ProphetNet * reformat	2021-06-21 09:36:44 -04:00
Lysandre Debut	b76850a808	Better CI feedback (#12279 ) * Better run ID * Only part of CI * Revert "Only part of CI" This reverts commit `29f7f248d2`.	2021-06-21 02:52:12 -04:00
Lysandre	30a5521c0b	Fix the scheduled CI	2021-06-21 08:27:25 +02:00
Stas Bekman	2e5dbdf2db	[t5 doc] make the example work out of the box (#12239 ) * [run_clm.py] restore caching * style * [t5 doc] make the example work out of the box This PR expands the training example to include the correct model type for the example to work, e.g. with `T5Model` this example will break. * Update docs/source/model_doc/t5.rst Co-authored-by: Suraj Patil <surajp815@gmail.com> * expand the other example Co-authored-by: Suraj Patil <surajp815@gmail.com>	2021-06-18 10:00:19 -07:00
Xa9aX ツ	f3558bbcfd	Depreciate pythonic Mish and support PyTorch 1.9 version of Mish (#12240 ) * Moved Mish to Torch 1.9 version * Run black formatting	2021-06-18 09:13:45 -04:00
Suraj Patil	47a9768334	[FlaxBart] few small fixes (#12247 ) * boom boom * remove flax clip example * few small fixes	2021-06-18 10:29:42 +01:00
Suraj Patil	f74655cd9b	[Flax] FlaxAutoModelForSeq2SeqLM (#12228 ) * add FlaxAutoModelForSeq2SeqLM	2021-06-18 13:20:09 +05:30
Bhavitvya Malik	e43e11260f	update desc for map in all examples (#12226 ) * update desc for map in all examples * added plm * suggestions	2021-06-17 15:37:31 -04:00
Sylvain Gugger	adb70eda4d	AutoTokenizer: infer the class from the tokenizer config if possible (#12208 ) * AutoTokenizer: infer the class from the tokenizer config if possible * Add tests * Update src/transformers/models/auto/tokenization_auto.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2021-06-17 12:39:22 -04:00
Lysandre	0daadc1919	Docs for v4.8.0	2021-06-17 18:17:42 +02:00
Lysandre	7a6c9fab8e	Release: v4.7.0	2021-06-17 17:57:42 +02:00
Stas Bekman	d6ea91c96a	fix pt-1.9.0 `add_` deprecation (#12217 ) * fix pt-1.9.0 add_ deprecation * add () for clarity * Trigger CI * require_version(torch	2021-06-17 08:53:59 -07:00
Lysandre Debut	3a960c4857	Support for torch 1.9.0 (#12224 ) * Support for torch 1.9.0 * Torch scatter for 1.9.0 * Github Actions run on 1.9.0	2021-06-17 11:29:01 -04:00
Sylvain Gugger	afdd9e3663	Add link to the course (#12229 )	2021-06-17 11:14:53 -04:00
NielsRogge	29b0aef871	Improve detr (#12147 ) * Remove unused variables * Improve docs * Fix docs of segmentation masks Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-06-17 10:37:54 -04:00
Lysandre Debut	b56848c8c8	Pipeline update & tests (#12207 )	2021-06-17 09:41:16 +02:00
Bhadresh Savani	700cee3446	[Docs] fixed broken link (#12205 ) * fixed broken link * Update docs/source/benchmarks.rst Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Update docs/source/benchmarks.rst Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2021-06-16 15:14:53 -04:00
Sylvain Gugger	255a17a089	Use yaml to create metadata (#12185 ) * Use yaml to create metadata * Fix typo * Remove pin	2021-06-16 13:17:45 -04:00
Nicolas Patry	15ef0dc5c6	Enabling AutoTokenizer for HubertConfig. (#12198 )	2021-06-16 15:28:46 +01:00
Philipp Schmid	afa414d060	updated DLC images and sample notebooks (#12191 )	2021-06-16 07:24:00 -04:00
Patrick von Platen	ccca510276	Hubert (#11889 ) * fix_torch_device_generate_test * remove @ * add hubert * add first test file * more docs * fix bugs * fix bug * finish * finish * finish docstring * fix * fix * finalize * add to ignored * finish * Apply suggestions from code review * correct naming * finish * fix auto config * finish * correct convert script * Apply suggestions from code review Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Suraj Patil <surajp815@gmail.com> * apply suggestions lysandre & suraj Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Suraj Patil <surajp815@gmail.com>	2021-06-16 12:14:12 +01:00
Patrick von Platen	c3c39f7e84	[Flax] Add Beam Search (#12131 ) * fix_torch_device_generate_test * remove @ * push new logit processors * add processors * save first working version * save intermediate * finish * make style * make fix-copies * finish * Update tests/test_modeling_flax_bart.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Apply suggestions from code review Co-authored-by: Suraj Patil <surajp815@gmail.com> Co-authored-by: Patrick von Platen <patrick@huggingface.co> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Suraj Patil <surajp815@gmail.com>	2021-06-16 09:43:54 +01:00
Sylvain Gugger	802ffaff0d	Temporarily deactivate torchhub test (#12184 )	2021-06-15 16:16:51 -04:00
Lysandre Debut	52c7ca0488	Temporarily deactivate torch-scatter while we wait for new release (#12181 ) * Temporarily deactivate torch-scatter while we wait for new release * torch-1.8.1 binary for scatter * Revert to 1.8.0 * Pin torch dependency * torchaudio and torchvision	2021-06-15 16:03:58 -04:00
Sylvain Gugger	7d7ceca396	Model card defaults (#12122 ) * [WIP] Model card defaults * finetuned_from default value * Add all mappings to the mapping file * Be more defensive on finetuned_from arg * Add default task tag * Separate tags from tasks * Edge case for dataset * Apply suggestions from code review Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2021-06-15 16:01:37 -04:00
Stas Bekman	6e7cc5cc51	[testing] ensure concurrent pytest workers use a unique port for torch.dist (#12166 ) * ensure concurrent pytest workers use a unique port for torch.distributed.launch * reword	2021-06-15 11:12:59 -07:00
Amog Kamsetty	b9d66f4c4b	Ray Tune Integration Updates (#12134 ) * fix * fixes * add back to scheduled tests * formatting * Update integrations.py	2021-06-15 14:11:29 -04:00
Kilian Kluge	a79585bbf9	Update AutoModel classes in summarization example (#12178 ) - Convert use of deprecated AutoModelWithLMHead to AutoModelForSeq2SeqLM - Add newly required `truncation=True` to `tokenizer.encode` with `max_length` This silences all warnings.	2021-06-15 10:36:10 -04:00
Sylvain Gugger	d6c929e200	Merge remote-tracking branch 'origin/master'	2021-06-15 09:37:46 -04:00
Sylvain Gugger	a8694b8850	Adjust banner width	2021-06-15 09:37:15 -04:00
kumapo	955b2b97a6	Enable add_prefix_space if model_type is roberta or gpt2 (#12116 )	2021-06-15 09:33:21 -04:00

1 2 3 4 5 ...

7395 Commits