transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-31 02:02:21 +06:00

Author	SHA1	Message	Date
dependabot[bot]	e49c71fc4c	Bump nbconvert from 6.3.0 to 6.5.1 in /examples/research_projects/lxmert (#18742 ) Bumps [nbconvert](https://github.com/jupyter/nbconvert) from 6.3.0 to 6.5.1. - [Release notes](https://github.com/jupyter/nbconvert/releases) - [Commits](https://github.com/jupyter/nbconvert/compare/6.3.0...6.5.1) --- updated-dependencies: - dependency-name: nbconvert dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com> Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-08-24 06:12:56 -04:00
dependabot[bot]	5b24949669	Bump nbconvert in /examples/research_projects/visual_bert (#18741 ) Bumps [nbconvert](https://github.com/jupyter/nbconvert) from 6.3.0 to 6.5.1. - [Release notes](https://github.com/jupyter/nbconvert/releases) - [Commits](https://github.com/jupyter/nbconvert/compare/6.3.0...6.5.1) --- updated-dependencies: - dependency-name: nbconvert dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com> Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-08-24 06:12:48 -04:00
Daniel Stancl	c72d7d91bf	Add TF implementation of `XGLMModel` (#16543 ) * Add TFXGLM models * Add todo: self.supports_xla_generation = False Co-authored-by: Daniel Stancl <stancld@Daniels-MacBook-Pro.local> Co-authored-by: Daniel Stancl <stancld@daniels-mbp.home> Co-authored-by: Joao Gante <joaofranciscocardosogante@gmail.com> Co-authored-by: Daniel <daniel.stancl@rossum.ai> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2022-08-24 10:51:05 +01:00
Yih-Dar	cecf9f9b27	fix pipeline_tutorial.mdx doctest (#18717 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-24 05:38:03 -04:00
Constantin Hütterer	a442884b87	Add minor doc-string change to include hp_name param in hyperparameter_search (#18700 ) * Add minor doc-string change to include hp_name * fix: missing type-information for kwargs * fix: missing white-space in hyperparameter_search doc-strings	2022-08-24 05:07:17 -04:00
Mishig Davaadorj	c12dbdc246	Update perf_infer_gpu_many.mdx (#18744 )	2022-08-24 10:37:52 +02:00
Joao Gante	6faf283288	CLI: Don't check the model head when there is no model head (#18733 )	2022-08-23 15:38:59 +01:00
SaulLu	438698085c	improve `add_tokens` docstring (#18687 ) * improve add_tokens documentation * format	2022-08-23 07:23:51 -04:00
Nicolas Patry	891704b3c2	Removing warning of model type for `microsoft/tapex-base-finetuned-wtq` (#18711 ) and friends.	2022-08-23 13:17:06 +02:00
Yih-Dar	84beb8a49b	Unpin detectron2 (#18727 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-23 11:10:07 +02:00
Atharva Ingle	d90a36d192	remove check for main process for trackers initialization (#18706 )	2022-08-22 11:16:27 -04:00
tgadeliya	0f257a8774	Add missing tokenizer tests - Longformer (#17677 )	2022-08-22 12:13:20 +02:00
Yih-Dar	3fa45dbd91	Fix Data2VecVision ONNX test (#18587 ) Co-authored-by: lewtun <lewis.c.tunstall@gmail.com> Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-22 11:28:23 +02:00
Yih-Dar	30992ef0d9	[Hotfix] pin detectron2 5aeb252 to avoid test fix (#18701 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-20 00:37:38 +02:00
Patrick von Platen	1f3c2282b5	Temp fix for broken detectron2 import (#18699 ) * add first generation tutorial * [Circle CI] Temporary fix for broken detectron2 import * remove generation	2022-08-19 22:55:33 +02:00
Joao Gante	e95d433d77	Generate: add missing `**model_kwargs` in sample tests (#18696 )	2022-08-19 16:14:27 +01:00
Atharva Ingle	e54a1b49aa	`model.tie_weights()` should be applied after `accelerator.prepare()` (#18676 ) * `model.tie_weights()` should be applied after `accelerator.prepare` Weight tying should be done after the model has been moved to XLA device as mentioned on PyTorch/XLA Troubleshooting guide [here](https://github.com/pytorch/xla/blob/master/TROUBLESHOOTING.md#xla-tensor-quirks) * format code	2022-08-18 13:46:57 -04:00
Loubna Ben Allal	bbbb453e58	Add an examples folder for code downstream tasks (#18679 ) * add examples subfolder * mention examples in codeparrot readme * use Trainer optimizer and scheduler type and add output_dir as argument * add example of text-to-python and python-to-text models * mention the downstream examples in the readme * fix typo	2022-08-18 18:24:24 +02:00
Younes Belkada	a123eee9df	[bnb] Move documentation (#18671 ) * fix bnb documentation - move bnb documentation to `infer_gpu_many` * small refactoring - added text on infer_gpu_one - added a small note on infer_gpu_many - added customized multi gpu example on infer_gpu_many * Update docs/source/en/perf_infer_gpu_many.mdx Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> * apply suggestions Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> * Apply suggestions from code review Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>	2022-08-18 17:34:48 +02:00
Zachary Mueller	358fc18613	Add evaluate to examples requirements (#18666 )	2022-08-18 10:57:39 -04:00
Severin Simmler	d243112b65	Fix breaking change in `onnxruntime` for ONNX quantization (#18336 ) * Fix quantization * Save model * Remove unused comments * Fix formatting	2022-08-18 10:06:16 -04:00
lewtun	5987c637ee	Fix repo consistency (#18682 )	2022-08-18 09:47:50 -04:00
regisss	76454b08c8	Rename second input dimension from "sequence" to "num_channels" for CV models (#17976 )	2022-08-18 15:13:54 +02:00
amyeroberts	780253ce3d	Rename method to avoid clash with property (#18677 )	2022-08-18 12:56:27 +01:00
Yih-Dar	2c947d2939	Ping `detectron2` for CircleCI tests (#18680 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-18 12:57:18 +02:00
Joao Gante	a541d97477	Generate: validate model_kwargs on FLAX (and catch typos in generate arguments) (#18653 )	2022-08-18 10:56:21 +01:00
Patrick von Platen	0ea53822f8	[LongT5] Correct docs long t5 (#18669 ) * add first generation tutorial * [LongT5 Docs] Correct docs * correct expected string * remove incorrect file	2022-08-18 10:03:50 +02:00
Matt	582c537175	Allow users to force TF availability (#18650 ) * Allow users to force TF availability * Correctly name the envvar!	2022-08-18 03:09:09 -04:00
amyeroberts	49e44b216b	Update feature extractor methods to enable type cast before normalize (#18499 ) * Update methods to optionally rescale This is necessary to allow for casting our images / videos to numpy arrays within the feature extractors' call. We want to do this to make sure the behaviour is as expected when flags like are False. If some transformations aren't applied, then the output type can't be unexpected e.g. a list of PIL images instead of numpy arrays. * Cast images to numpy arrays in call to enable consistent behaviour with different configs * Remove accidental clip changes * Update tests to reflect the scaling logic We write a generic function to handle rescaling of our arrays. In order for the API to be intuitive, we take some factor c and rescale the image values by that. This means, the rescaling done in normalize and to_numpy_array are now done with array * (1/255) instead of array / 255. This leads to small differences in the resulting image. When testing, this was in the order of 1e-8, and so deemed OK	2022-08-17 19:57:07 +01:00
Jingya HUANG	86d0b26d6c	Fix matmul inputs dtype (#18585 )	2022-08-17 15:59:43 +02:00
Yih-Dar	c99e984657	Fix Yolos ONNX export test (#18606 ) Co-authored-by: lewtun <lewis.c.tunstall@gmail.com> Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-17 10:04:49 +02:00
Stefan Schweter	358478e729	Examples: add Bloom support for token classification (#18632 ) * examples: add Bloom support for token classification (FLAX, PyTorch and TensorFlow) * examples: remove support for Bloom in token classication (FLAX and TensorFlow currently have no support for it)	2022-08-17 09:50:57 +02:00
Younes Belkada	6d175c1129	[bnb] Minor modifications (#18631 ) * bnb minor modifications - refactor documentation - add troubleshooting README - add PyPi library on DockerFile * Apply suggestions from code review Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> * Apply suggestions from code review * Apply suggestions from code review * Apply suggestions from code review * put in one block - put bash instructions in one block * update readme - refactor a bit hardware requirements * change text a bit * Apply suggestions from code review Co-authored-by: Yih-Dar <2521628+ydshieh@users.noreply.github.com> * apply suggestions Co-authored-by: Yih-Dar <2521628+ydshieh@users.noreply.github.com> * add link to paper * Apply suggestions from code review Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> * Update tests/mixed_int8/README.md * Apply suggestions from code review * refactor a bit * add instructions Turing & Amperer Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> * add A6000 * clarify a bit * remove small part * Update tests/mixed_int8/README.md Co-authored-by: Stas Bekman <stas00@users.noreply.github.com> Co-authored-by: Yih-Dar <2521628+ydshieh@users.noreply.github.com>	2022-08-17 00:48:10 +02:00
zhoutang776	25e651a2de	Update run_translation_no_trainer.py (#18637 ) * Update run_translation_no_trainer.py found an error in selecting `no_decay` parameters and some small modifications when the user continues to train from a checkpoint * fixs `no_decay` and `resume_step` issue 1. change `no_decay` list 2. if use continue to train their model from provided checkpoint, the `resume_step` will not be initialized properly if `args.gradient_accumulation_steps != 1`	2022-08-16 13:25:57 -04:00
flozi00	a27195b1de	Update longt5.mdx (#18634 )	2022-08-16 10:20:46 -05:00
Joao Gante	fd9aa82b07	TF: Fix generation repetition penalty with XLA (#18648 )	2022-08-16 13:30:52 +01:00
Yih-Dar	81ab11124f	Add checks for some workflow jobs (#18583 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-16 13:53:47 +02:00
Yih-Dar	510c2a0b32	Change scheduled CIs to use torch 1.12.1 (#18644 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-16 13:41:37 +02:00
Sourab Mangrulkar	9cf274685a	mac m1 `mps` integration (#18598 ) * mac m1 `mps` integration * Update docs/source/en/main_classes/trainer.mdx Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * addressing comments * Apply suggestions from code review Co-authored-by: Dan Saattrup Nielsen <47701536+saattrupdan@users.noreply.github.com> * resolve comment Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Dan Saattrup Nielsen <47701536+saattrupdan@users.noreply.github.com>	2022-08-16 16:34:51 +05:30
Karim Foda	d6eeb87170	Flax Remat for LongT5 (#17994 ) * [Flax] Add remat (gradient checkpointing) * fix variable naming in test * flip: checkpoint using a method * fix naming * fix class naming * apply PVP's suggestions from code review * add gradient_checkpointing to examples * Add gradient_checkpointing to run_mlm_flax * Add remat to longt5 * Add gradient checkpointing test longt5 * Fix args errors * Fix remaining tests * Make fixup & quality fixes * replace kwargs * remove unecessary kwargs * Make fixup changes * revert long_t5_flax changes * Remove return_dict and copy to LongT5 * Remove test_gradient_checkpointing Co-authored-by: sanchit-gandhi <sanchit@huggingface.co>	2022-08-14 16:27:13 +01:00
Younes Belkada	1ccd2515ed	small change (#18584 )	2022-08-12 20:04:38 +02:00
Stas Bekman	b3ff7c680c	[fsmt] deal with -100 indices in decoder ids (#18592 ) * [fsmt] deal with -100 indices in decoder ids Fixes: https://github.com/huggingface/transformers/issues/17945 decoder ids get the default index -100, which breaks the model - like t5 and many other models add a fix to replace -100 with the correct pad index. For some reason this use case hasn't been used with this model until recently - so this issue was there since the beginning it seems. Any suggestions to how to add a simple test here? or perhaps we have something similar already? user's script is quite massive. * style	2022-08-12 10:50:52 -07:00
Stas Bekman	37c5991843	[doc] fix anchors (#18591 ) the manual anchors end up being duplicated with automatically added anchors and no longer work.	2022-08-12 10:49:59 -07:00
Niklas Muennighoff	56ef0ba447	Update BLOOM parameter counts (#18531 ) * Update BLOOM parameter counts * Update BLOOM parameter counts	2022-08-12 19:36:18 +02:00
NielsRogge	153d1361c7	Fix URLs (#18604 ) Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2022-08-12 18:52:49 +02:00
NielsRogge	2ab790e82d	Add Donut (#18488 ) * First draft * Improve script * Update script * Make conversion work * Add final_layer_norm attribute to Swin's config * Add DonutProcessor * Convert more models * Improve feature extractor and convert base models * Fix bug * Improve integration tests * Improve integration tests and add model to README * Add doc test * Add feature extractor to docs * Fix integration tests * Remove register_buffer * Fix toctree and add missing attribute * Add DonutSwin * Make conversion script work * Improve conversion script * Address comment * Fix bug * Fix another bug * Remove deprecated method from docs * Make Swin and Swinv2 untouched * Fix code examples * Fix processor * Update model_type to donut-swin * Add feature extractor tests, add token2json method, improve feature extractor * Fix failing tests, remove integration test * Add do_thumbnail for consistency * Improve code examples * Add code example for document parsing * Add DonutSwin to MODEL_NAMES_MAPPING * Add model to appropriate place in toctree * Update namespace to appropriate organization Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2022-08-12 16:40:58 +02:00
Younes Belkada	a5ca56ff15	Supporting seq2seq models for `bitsandbytes` integration (#18579 ) * Supporting seq2seq models for `bitsandbytes` integration - `bitsandbytes` integration supports now seq2seq models - check if a model has tied weights as an additional check * small modification - tie the weights before looking at tied weights!	2022-08-12 16:15:09 +02:00
Joao Gante	ed1924e801	Generate: validate `model_kwargs` (and catch typos in generate arguments) (#18261 ) * validate generate model_kwargs * generate tests -- not all models have an attn mask	2022-08-12 14:53:51 +01:00
Yih-Dar	2156619f10	Add `TFAutoModelForSemanticSegmentation` to the main `__init__.py` (#18600 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-08-12 15:10:00 +02:00
Sourab Mangrulkar	4eed2beca0	FSDP bug fix for `load_state_dict` (#18596 )	2022-08-12 08:48:37 -04:00

1 2 3 4 5 ...

10492 Commits