transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-31 02:02:21 +06:00

Author	SHA1	Message	Date
Teven	2a7e8e1608	[Examples] Add automatic dataset splitting in language-modeling examples (#9133 ) * replaced jnp.split + removing textual model inputs + ensuring warmup_steps > 0 * Add automatic dataset splitting in language-modeling examples	2020-12-15 16:02:43 -05:00
Julien Plu	e771749777	Fix add order (#9129 )	2020-12-15 15:16:56 -05:00
Patrick von Platen	18ecd36f65	Fix Bart Shift (#9135 ) * correct mistake in order * fix tensor copy * clone tensor correctly	2020-12-15 19:04:31 +01:00
Patrick von Platen	d018622d8e	correct mistake in order (#9134 )	2020-12-15 23:08:31 +05:30
Patrick von Platen	80bdb9c31a	fix bart loss masking (#9131 )	2020-12-15 18:17:17 +01:00
Manbish	3caba8d35f	Fix typo in trainer_tf.py (#9132 )	2020-12-15 12:12:28 -05:00
Patrick von Platen	abc573f51a	[TF Bart] Refactor TFBart (#9029 ) * reorder file * delete unnecesarry function * make style * save intermediate * fix attention masks * correct tf bart past key values * solve merge conflict bug * correct tensor dims * save intermediate tf * change attn layer * fix typo re-order past * inputs_embeds * make fix copies * finish tests * fix graph mode * appyl lysandres suggestions	2020-12-15 17:31:28 +01:00
sandip	389aba34bf	Added TF OpenAi GPT1 Sequence Classification (#9105 ) * TF OpenAI GPT Sequence Classification * Update src/transformers/models/openai/modeling_tf_openai.py Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2020-12-15 11:27:08 -05:00
Julien Plu	ef2d4cd445	Fix tf2.4 (#9120 ) * Fix tests for TF 2.4 * Remove <2.4 limitation * Add version condition * Update tests/test_optimization_tf.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Update tests/test_optimization_tf.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Update tests/test_optimization_tf.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2020-12-15 10:10:46 -05:00
Lysandre Debut	6ccea0486f	Fix T5 model parallel tes (#9107 ) k	2020-12-15 09:51:12 -05:00
Lysandre Debut	59da3f2700	Fix stack overflow (#9114 )	2020-12-15 09:15:49 -05:00
Stas Bekman	14c79c3e31	native amp leak fix landed in 1.7.1 (#9115 ) update README with good news that the leak fix has been applied to pytorch-1.7.1.	2020-12-15 09:10:41 -05:00
lewtun	ed1845ef4c	Clarify use of TrainingArguments.disable_tqdm in Jupyter Notebooks (#9076 ) * Clarify impact of disable_tqdm on Jupyter Notebooks * Add weblink to argparse * Replace "dev set" with more common "validation set" in do_eval * Tweak prediction_loss_only * Tweak description of Adam hyperparameters * Add weblink to TensorBoard * Capitalise apex * Tweak local_rank description * Add weblink for wandb * Replace nlp with datasets * Tweak grammar in model_parallel * Capitalise apex * Update TensorFlow training args to match PyTorch ones * Fix style * Fix underscore in weblink Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Fix underscore in weblink Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Fix underscore in weblink Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Fix underscore in weblink Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Add obj to datasets.Dataset Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2020-12-15 09:00:19 -05:00
Yoshitomo Matsubara	44c340f45f	fix a bug in eval_batch_retrieval (#9089 )	2020-12-15 14:46:55 +01:00
Stas Bekman	c19d04623e	[finetune_trainer] enhancements and fixes (#9042 ) * trainer and finetune_trainer enhancements and fixes * add fallback default * move the fixing of incorrect keys back into finetune trainer * s/eval/val/ to match the split * trainer can now use a different prefix than eval_ for metrics * document new arg * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * use 'eval' as the default for metric_key_prefix * complete adjust var names + disambiguate * fix logger * add clarifying comment * add clarifying comment * style * Apply suggestions from code review Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * Update src/transformers/trainer.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * complete removal of optional for metric_key_prefix * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2020-12-14 17:45:33 -08:00
Sylvain Gugger	251eb70c97	Also pin TF CPU	2020-12-14 16:17:04 -05:00
Sylvain Gugger	e4ef57a9bb	Pin TF to < 2.4	2020-12-14 16:06:30 -05:00
Julien Plu	df3f4d2aef	Fix T5 and BART for TF (#9063 ) * Fix T5 for graphe compilation+execution * Fix BART * Fix import * Fix naming * fix attribute name * Oops * fix import * fix tests * fix tests * Update test * Add mising import * Address Patrick's comments * Style * Address Patrick's comment	2020-12-14 18:47:00 +01:00
Ahmed Elnaggar	a9c8bff724	Add parallelization support for T5EncoderModel (#9082 ) * add model parallelism to T5EncoderModel add model parallelism to T5EncoderModel * remove decoder from T5EncoderModel parallelize * uodate T5EncoderModel docs * Extend T5ModelTest for T5EncoderModel * fix T5Stask using range for get_device_map * fix style Co-authored-by: Ahmed Elnaggar <elnaggar@rostlab.informatik.tu-muenchen.de>	2020-12-14 12:00:45 -05:00
Stas Bekman	b00eb4fb02	Testing Experimental CI Features (#9070 )	2020-12-14 10:34:59 -05:00
Simon Brandeis	74daf1f954	Fixed a broken link in documentation (#9101 )	2020-12-14 09:12:27 -05:00
Navjot	d6af344c9e	correct var name in TrainingArguments docstring (#9096 )	2020-12-14 09:02:54 -05:00
Patrick von Platen	fa1ddced9e	[RAG, Bart] Align RAG, Bart cache with T5 and other models of transformers (#9098 ) * fix rag * fix slow test * fix past in bart	2020-12-14 12:32:26 +01:00
Lysandre Debut	6587cf9f84	Patch *ForCausalLM model (#9092 )	2020-12-14 00:39:55 -05:00
Julien Plu	51d9c569fa	Fix embeddings resizing in TF models (#8657 ) * Resize the biases in same time than the embeddings * Trigger CI * Biases are not reset anymore * Remove get_output_embeddings + better LM model detection in generation utils * Apply style * First test on BERT * Update docstring + new name * Apply the new resizing logic to all the models * fix tests * Apply style * Update the template * Fix naming * Fix naming * Apply style * Apply style * Remove unused import * Revert get_output_embeddings * Trigger CI * Update num parameters * Restore get_output_embeddings in TFPretrainedModel and add comments * Style * Add decoder resizing * Style * Fix tests * Separate bias and decoder resize * Fix tests * Fix tests * Apply style * Add bias resizing in MPNet * Trigger CI * Apply style	2020-12-13 23:05:24 -05:00
Julien Chaumond	3552d0e0d8	[model_cards] Migrate cards from this repo to model repos on huggingface.co (#9013 ) * rm all model cards * Update the .rst @sgugger it is still not super crystal clear/streamlined so let me know if any ideas to make it simpler * Add a rootlevel README.md with simple instructions/context * Update docs/source/model_sharing.rst Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * make style * rm all model cards Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2020-12-11 18:24:42 -05:00
Sylvain Gugger	29e4597950	Fix min_null_pred in the run_qa script (#9067 )	2020-12-11 16:26:05 -05:00
Patrick von Platen	9cc9f4122e	Make ProphetNetModel really compatible with EncoderDecoder (#9033 ) * improve * finish * upload model * fix lm head * fix test	2020-12-11 16:59:54 +01:00
dependabot[bot]	24f6cdeab6	Bump notebook in /examples/research_projects/movement-pruning/lxmert (#9062 ) Bumps [notebook](https://github.com/jupyter/jupyterhub) from 6.1.4 to 6.1.5. - [Release notes](https://github.com/jupyter/jupyterhub/releases) - [Changelog](https://github.com/jupyterhub/jupyterhub/blob/master/CHECKLIST-Release.md) - [Commits](https://github.com/jupyter/jupyterhub/commits) Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2020-12-11 10:32:43 -05:00
Lysandre Debut	91fa707217	Remove docs only check (#9065 )	2020-12-11 10:27:31 -05:00
Sylvain Gugger	70527ba694	Fix PreTrainedTokenizer.pad when first inputs are empty (#9018 ) * Fix PreTrainedTokenizer.pad when first inputs are empty * Handle empty inputs case	2020-12-11 10:25:00 -05:00
Sylvain Gugger	783d7d2629	Reorganize examples (#9010 ) * Reorganize example folder * Continue reorganization * Change requirements for tests * Final cleanup * Finish regroup with tests all passing * Copyright * Requirements and readme * Make a full link for the documentation * Address review comments * Apply suggestions from code review Co-authored-by: Lysandre Debut <lysandre@huggingface.co> * Add symlink * Reorg again * Apply suggestions from code review Co-authored-by: Thomas Wolf <thomwolf@users.noreply.github.com> * Adapt title * Update to new strucutre * Remove test * Update READMEs Co-authored-by: Lysandre Debut <lysandre@huggingface.co> Co-authored-by: Thomas Wolf <thomwolf@users.noreply.github.com>	2020-12-11 10:07:02 -05:00
Suraj Patil	86896de064	update tatoeba workflow (#9051 )	2020-12-11 20:29:15 +05:30
Ganesh Kharad	7c8f5f6487	Create README.md (#8096 ) * Create README.md * Fix model card Co-authored-by: Julien Chaumond <julien@huggingface.co>	2020-12-11 09:45:12 -05:00
RamonMamon	5527f78721	Create README.md (#8281 ) * Create README.md * Update model_cards/kiri-ai/distiluse-base-multilingual-cased-et/README.md Co-authored-by: Julien Chaumond <chaumond@gmail.com>	2020-12-11 09:41:29 -05:00
joangines	c615df7422	Create README.md (#8751 ) * Create README.md * Update model_cards/Cinnamon/electra-small-japanese-generator/README.md Co-authored-by: Julien Chaumond <chaumond@gmail.com>	2020-12-11 09:40:14 -05:00
Ahmed Abdelali	76df559383	QARiB Arabic and dialects models (#8796 ) * Add QARiB models * fix README.md * Fix README.md * Fix README.md * Fix README.md * Fix QARiB files * add models card for QARiB models 860k, 1790k, and 1970k * try to fix PR * re-add files * links aren't allowed here :) Co-authored-by: Ahmed Abdelali <aabdelali@hbku.edu.qa> Co-authored-by: Julien Chaumond <julien@huggingface.co>	2020-12-11 09:38:38 -05:00
moniquebm	b161f1ae54	Update README.md (#8820 )	2020-12-11 09:24:21 -05:00
Panggi Libersa Jasri Akadol	649d389dab	Initial README for `t5-base-indonesian-summarization-cased` model (#9028 ) * Create README.md Initial README for `t5-base-indonesian-summarization-cased` model * Update README for t5-base-indonesian-summarization-cased Typo in README, change from `small` to `base`	2020-12-11 09:18:16 -05:00
Panggi Libersa Jasri Akadol	5e794b6628	Create README.md (#9030 ) Initial README for `t5-small-indonesian-summarization-cased` model	2020-12-11 09:17:29 -05:00
Cola	935e346959	🎨 Change nn.dropout to layer.Dropout (#9047 )	2020-12-11 10:40:25 +01:00
Julien Plu	b01ddc9577	Remove value error (#8985 ) * Remove value error * Try a fix for parameter ordering * Restore previous behavior * Add documentation * Review the comment	2020-12-10 17:17:19 -05:00
NatLun137	91ab02af28	Fix typo #9012 (#1 ) (#9038 ) There is a tiny typo in the code "transformers/examples/language-modeling/run_mlm_wwm.py" at line 284. [Details.](https://github.com/huggingface/transformers/issues/9012)	2020-12-10 16:41:00 -05:00
Sylvain Gugger	8d4bb02056	Refactor FLAX tests (#9034 )	2020-12-10 15:57:39 -05:00
Sylvain Gugger	1310e1a758	Enforce all objects in the main init are documented (#9014 )	2020-12-10 11:57:12 -05:00
Sylvain Gugger	51e81e5895	MPNet copyright files (#9015 )	2020-12-10 09:29:38 -05:00
Sylvain Gugger	35bffd70e2	Fix documention of book in LayoutLM (#9017 )	2020-12-10 09:28:49 -05:00
Cola	c95de29e31	✏️ Fix typo (#9020 )	2020-12-10 08:22:52 +01:00
Stas Bekman	5e637e6c69	[wip] [ci] doc-job-skip take #4 dry-run (#8980 ) * ci-doc-job-skip-take-4 * wip * wip * wip * wip * skip yaml * wip * wip * wip * wip * wip * wip * wip * wip * wip * wip * wip * wip * ready to test * yet another way * trying with HEAD * trying with head.sha * trying with head.sha fix * trying with head.sha fix wip * undo * try to switch to sha * current branch * current branch * PR number check * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride * joy ride	2020-12-09 15:36:36 -05:00
Patrick von Platen	06971ac4f9	[Bart] Refactor - fix issues, consistency with the library, naming (#8900 ) * remove make on the fly linear embedding * start refactor * big first refactor * save intermediate * save intermediat * correct mask issue * save tests * refactor padding masks * make all tests pass * further refactor * make pegasus test pass * fix bool if * fix leftover tests * continue * bart renaming * delete torchscript test hack * fix imports in tests * correct shift * fix docs and repo cons * re-add fix for FSTM * typo in test * fix typo * fix another typo * continue * hot fix 2 for tf * small fixes * refactor types linting * continue * finish refactor * fix import in tests * better bart names * further refactor and add test * delete hack * apply sylvains and lysandres commens * small perf improv * further perf improv * improv perf * fix typo * make style * small perf improv	2020-12-09 20:55:24 +01:00

1 2 3 4 5 ...

6122 Commits