transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-08-03 03:31:05 +06:00

Author	SHA1	Message	Date
Marc Sun	5e11d72d4d	fix_mbart_tied_weights (#26422 ) * fix_mbart_tied_weights * add test	2023-09-28 15:08:35 +02:00
fleance	216dff7549	Do not warn about unexpected decoder weights when loading T5EncoderModel and LongT5EncoderModel (#26211 ) Ignore decoder weights when using T5EncoderModel and LongT5EncoderModel Both T5EncoderModel and LongT5EncoderModel do not have any decoder layers, so loading a pretrained model checkpoint such as t5-small will give warnings about keys found in the model checkpoint that are not in the model itself. To prevent this log warning, r"decoder" has been added to _keys_to_ignore_on_load_unexpected for both T5EncoderModel and LongT5EncoderModel	2023-09-28 11:27:43 +02:00
Younes Belkada	38e96324ef	[`PEFT`] introducing `adapter_kwargs` for loading adapters from different Hub location (`subfolder`, `revision`) than the base model (#26270 ) * make use of adapter_revision * v1 adapter kwargs * fix CI * fix CI * fix CI * fixup * add BC * Update src/transformers/integrations/peft.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * fixup * change it to error * Update src/transformers/modeling_utils.py * Update src/transformers/modeling_utils.py * fixup * change * Update src/transformers/integrations/peft.py --------- Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>	2023-09-28 11:13:03 +02:00
Fakhir Ali	52e2c13da3	[VITS] Fix speaker_embed device mismatch (#26115 ) * [VITS] Fix speaker_embed device mismatch - pass device arg to speaker_id tensor * [VITS] put speaker_embed on device when int * [VITS] device=self.device instead of self.embed_speaker.weight.device * [VITS] make tensor directly on device using torch.full()	2023-09-28 10:56:36 +02:00
Tanishq Abraham	098c3f400c	change mention of decoder_input_ids to input_ids and same with decode_inputs_embeds (#26406 ) * change mention of decoder_input_ids to input_ids and same with decoder_input_embeds * Style --------- Co-authored-by: Lysandre <lysandre@huggingface.co>	2023-09-28 10:15:48 +02:00
Phuc Van Phan	ba47efbfe4	docs: change assert to raise and some small docs (#26232 ) * docs: change assert to raise and some small docs * docs: add rule and some document * fix: fix bug * fix: fix bug * chorse: revert logging * chorse: revert	2023-09-28 10:14:17 +02:00
Yih-Dar	375b4e0935	Fix `cos_sin` device issue in Falcon model (#26448 ) * fix * fix * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-09-28 10:00:15 +02:00
Norm Inui	a7e0ed829c	optimize VRAM for calculating pos_bias in LayoutLM v2, v3 (#26139 ) * optimize layoutv2, v3 for VRAM saving * reformat codes --------- Co-authored-by: NormXU <xunuo@datagrand.com>	2023-09-28 09:55:57 +02:00
Wonhyeong Seo	ab37b801b1	🌐 [i18n-KO] Translated `perf_train_gpu_many.md` to Korean (#26244 ) * dos: ko: perf_train_gpu_many.mdx * feat: chatgpt draft * fix: manual edits * fix: resolve suggestions Change description Follow the glossary Fix discrepancies Co-Authored-By: SeongWooChoi <46990061+nuatmochoi@users.noreply.github.com> Co-Authored-By: 이서정 <97655267+sjlee-wise@users.noreply.github.com> Co-Authored-By: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Hyunho <105839613+hyunhp@users.noreply.github.com> Co-authored-by: SeongWooChoi <46990061+nuatmochoi@users.noreply.github.com> Co-authored-by: 이서정 <97655267+sjlee-wise@users.noreply.github.com> Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2023-09-27 13:51:15 -07:00
Wonhyeong Seo	a0922a538b	🌐 [i18n-KO] Translated `debugging.md` to Korean (#26246 ) * docs:ko:Debugging.md * feat: chatgpt draft * fix: resolve suggestions Co-Authored-By: Sohyun Sim <96299403+sim-so@users.noreply.github.com> Co-Authored-By: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Jang KyuJin <106062329+kj021@users.noreply.github.com> Co-authored-by: Sohyun Sim <96299403+sim-so@users.noreply.github.com> Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2023-09-27 13:47:44 -07:00
Florian Zimmermeister	ef81759e31	[i18n-DE] Complete first toc chapter (#26311 ) * initial * toctree * add tf model * run scripts * peft * llm and agents * Update docs/source/de/peft.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/peft.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/peft.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/run_scripts.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/run_scripts.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/transformers_agents.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update docs/source/de/transformers_agents.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2023-09-27 11:33:05 -07:00
Yih-Dar	6ae71ec836	Update `runs-on` in workflow files (#26435 ) * update * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-09-27 19:25:52 +02:00
Lysandre Debut	78dd120282	Fix failing doctest (#26450 ) * Fix doctest * Adding modeling also for now	2023-09-27 18:47:26 +02:00
Chris Bamford	72958fcd3c	[Mistral] Mistral-7B-v0.1 support (#26447 ) * [Mistral] Mistral-7B-v0.1 support * fixing names * slightly longer test * fixups * not_doctested * wrongly formatted references * make fixuped --------- Co-authored-by: Timothee Lacroix <t@eugen.ai> Co-authored-by: timlacroix <t@mistral.ai>	2023-09-27 18:30:46 +02:00
Younes Belkada	3ca18d6d09	[`PEFT`] Fix PEFT multi adapters support (#26407 ) * fix PEFT multi adapters support * refactor a bit * save pretrained + BC + added tests * Update src/transformers/integrations/peft.py Co-authored-by: Benjamin Bossan <BenjaminBossan@users.noreply.github.com> * add more tests * add suggestion * final changes * adapt a bit * fixup * Update src/transformers/integrations/peft.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * adapt from suggestions --------- Co-authored-by: Benjamin Bossan <BenjaminBossan@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2023-09-27 16:45:31 +02:00
statelesshz	946bac798c	add bf16 mixed precision support for NPU (#26163 ) Co-authored-by: statelesshz <jihuazhong1@huawei.com>	2023-09-27 12:28:40 +02:00
Younes Belkada	153755ee38	[`FA` / `tests`] Add use_cache tests for FA models (#26415 ) * add use_cache tests for FA * fixup	2023-09-27 12:21:54 +02:00
Uri Alon	a0be960dcc	Fixing tokenizer when `transformers` is installed without `tokenizers` (#26236 ) * Fixing tokenizer when tokenizers is not installed * Adding __repr__ function and repr=True in dataclass * Revert "Adding __repr__ function and repr=True in dataclass" This reverts commit `18839505d1`.	2023-09-27 11:58:04 +02:00
Nour Eddine ZEKAOUI	777f2243f5	Update semantic_segmentation.md (#26419 )	2023-09-27 11:51:44 +02:00
Shauray Singh	abd2531034	Fix padding for IDEFICS (#26396 ) * fix * fixup * tests * fixup	2023-09-27 10:56:07 +02:00
Nathan Lambert	408b2b3c50	Add torch `RMSProp` optimizer (#26425 ) add rmsprop	2023-09-26 19:27:09 +02:00
Matt	6ba63ac3a0	[InternLM] Add support for InternLM (#26302 ) * Add config.bias to LLaMA to allow InternLM models to be ported as LLaMA checkpoints * Rename bias -> attention_bias and add docstring	2023-09-26 16:52:19 +01:00
Hugo Laurençon	0ac3875011	Fix DeepSpeed issue with Idefics (#26393 ) Fix deepspeed issue with Idefics	2023-09-26 10:19:00 +02:00
sanjeevk-os	6ce6a5adb9	added support for gradient checkpointing in ESM models (#26386 )	2023-09-26 10:15:53 +02:00
titi	a8531f3bfd	Deleted duplicate sentence (#26394 )	2023-09-26 10:11:28 +02:00
NielsRogge	a09130feee	[ViTMatte] Add resources (#26317 ) Add resource	2023-09-26 07:06:38 +02:00
NielsRogge	ace74d16bd	Add Nougat (#25942 ) * Add conversion script * Add NougatImageProcessor * Add crop margin * More improvements * Add docs, READMEs * Remove print statements * Include model_max_length * Add NougatTokenizerFast * Fix imports * Improve postprocessing * Improve image processor * Fix image processor * Improve normalize method * More improvements * More improvements * Add processor, improve docs * Simplify fast tokenizer * Remove test file * Fix docstrings * Use NougatProcessor in conversion script * Add is_levensthein_available * Add tokenizer tests * More improvements * Use numpy instead of opencv * Add is_cv2_available * Fix cv2_available * Add is_nltk_available * Add image processor tests, improve crop_margin * Add integration tests * Improve integration test * Use do_rescale instead of hacks, thanks Amy * Remove random_padding * Address comments * Address more comments * Add import * Address more comments * Address more comments * Address comment * Address comment * Set max_model_input_sizes * Add tests * Add requires_backends * Add Nougat to exotic tests * Use to_pil_image * Address comment regarding nltk * Add NLTK * Improve variable names, integration test * Add test * refactor, document, and test regexes * remove named capture groups, add comments * format * add non-markdown fixed tokenization * format * correct flakyness of args parse * add regex comments * test functionalities for crop_image, align long axis and expected output * add regex tests * remove cv2 dependency * test crop_margin equality between cv2 and python * refactor table regexes to markdown add newline * change print to log, improve doc * fix high count tables correction * address PR comments: naming, linting, asserts * Address comments * Add copied from * Update conversion script * Update conversion script to convert both small and base versions * Add inference example * Add more info * Fix style * Add require annotators to test * Define all keyword arguments explicitly * Move cv2 annotator * Add tokenizer init method * Transfer checkpoints * Add reference to Donut * Address comments * Skip test * Remove cv2 method * Add copied from statements * Use cached_property * Fix docstring * Add file to not doctested --------- Co-authored-by: Pablo Montalvo <pablo.montalvo.leroux@gmail.com>	2023-09-26 07:06:04 +02:00
Gabriel Yang	5e09af2acd	🌐 [i18n-KO] Translated `audio_classification.mdx` to Korean (#26200 ) * 🌐 [i18n-KO] Translated to Korean * update translation * fix some sentence editing and fixing punctuation * Update docs/source/ko/_toctree.yml Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Apply suggestions from code review Co-authored-by: Hyeonseo Yun <0525yhs@gmail.com> --------- Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> Co-authored-by: Hyeonseo Yun <0525yhs@gmail.com>	2023-09-25 10:24:45 -07:00
qweme32	033ec57c03	Add Russian localization for README (#26208 ) * Add Russian localization * typo * mistake in link * Update README_ru.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Update README_ru.md Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> Co-authored-by: Lysandre Debut <lysandre.debut@reseau.eseo.fr>	2023-09-25 09:42:23 -07:00
Yih-Dar	d9e4bc2895	Update tiny model information and pipeline tests (#26285 ) * Update tiny model summary file * add to pipeline tests * revert * fix import * fix import * fix * fix * update * update * update * fix * remove BarkModelTest * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-09-25 18:08:12 +02:00
Maria Khalusova	546e7679e7	[docs] removed MaskFormerSwin and TimmBackbone from the table on index.md (#26347 ) removed MaskFormerSwin and TimmBackbone from the table	2023-09-25 09:41:59 -04:00
Omar Sanseviero	0ee4590684	Fix MusicGen logging error (#26370 ) * Fix logging error * Update modeling_musicgen.py * Update modeling_musicgen.py	2023-09-25 13:08:25 +02:00
Nino Risteski	6accd5effb	Update add_new_model.md (#26365 ) fixed typos	2023-09-25 12:58:11 +02:00
HanSeokhyeon	5936c8c57c	Fixed unclosed p tags (#26240 )	2023-09-22 11:39:28 -07:00
Phuc Van Phan	910faa3e1f	feat: adding num_proc to load_dataset (#26326 ) * feat: adding num_proc to load_dataset * feat: add add_num_proc for run_mlm_flax * feat: add num_proc for bart and t5 * chorse: remove	2023-09-22 19:22:47 +02:00
LeviVasconcelos	576cd45a57	Add image to image pipeline (#25393 ) * Add image to image pipeline Add image to image pipeline * remove swin2sr from tf auto * make ImageToImage importable * make style make style make style make style * remove tf support * remove nonused imports * fix postprocessing * add important comments; add unit tests * add documentation * remove support for TF * make fixup * fix typehint Image.Image * fix documentation code * address review request; fix unittest type checking * address review request; fix unittest type checking * make fixup * address reviews * Update src/transformers/pipelines/image_to_image.py Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * enhance docs * make style * make style * improve docetest time * improve docetest time * Update tests/pipelines/test_pipelines_image_to_image.py Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> * Update tests/pipelines/test_pipelines_image_to_image.py Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> * make fixup * undo faulty merge * undo faulty merge * add image-to-image to test pipeline mixin * Update src/transformers/pipelines/image_to_image.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update tests/pipelines/test_pipelines_image_to_image.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * improve docs --------- Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>	2023-09-22 19:53:55 +03:00
Sanchit Gandhi	914771cbfe	[TTA Pipeline] Fix MusicGen test (#26348 ) * fix musicgen pipeline test * fix wav2vec2 doctest * revert wav2vec2	2023-09-22 17:55:54 +02:00
Younes Belkada	368a58e61c	[`core` ] Integrate Flash attention 2 in most used models (#25598 ) * v1 * oops * working v1 * fixup * add some TODOs * fixup * padding support + try with module replacement * nit * alternative design * oops * add `use_cache` support for llama * v1 falcon * nit * a bit of refactor * nit * nits nits * add v1 padding support falcon (even though it seemed to work before) * nit * falcon works * fixup * v1 tests * nit * fix generation llama flash * update tests * fix tests + nits * fix copies * fix nit * test- padding mask * stype * add more mem efficient support * Update src/transformers/modeling_utils.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * fixup * nit * fixup * remove it from config when saving * fixup * revert docstring * add more checks * use values * oops * new version * fixup * add same trick for falcon * nit * add another test * change tests * fix issues with GC and also falcon * fixup * oops * Update src/transformers/models/falcon/modeling_falcon.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * add init_rope * updates * fix copies * fixup * fixup * more clarification * fixup * right padding tests * add docs * add FA in docker image * more clarifications * add some figures * add todo * rectify comment * Change to FA2 * Update docs/source/en/perf_infer_gpu_one.md Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * split in two lines * change test name * add more tests * some clean up * remove `rearrange` deps * add more docs * revert changes on dockerfile * Revert "revert changes on dockerfile" This reverts commit `8d72a66b4b`. * revert changes on dockerfile * Apply suggestions from code review Co-authored-by: Lysandre Debut <hi@lysand.re> * address some comments * docs * use inheritance * Update src/transformers/testing_utils.py Co-authored-by: Lysandre Debut <hi@lysand.re> * fixup * Apply suggestions from code review Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update src/transformers/modeling_utils.py * final comments * clean up * style * add cast + warning for PEFT models * fixup --------- Co-authored-by: Felix Marty <9808326+fxmarty@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> Co-authored-by: Lysandre Debut <hi@lysand.re>	2023-09-22 17:42:10 +02:00
Maria Khalusova	dcbfd93d7a	[doc] fixed indices in obj detection example (#26343 ) fixed indexes in obj detection example	2023-09-22 10:29:27 -04:00
Yih-Dar	c3ecf2d95d	Fix doctest CI (#26324 ) fix doc CI Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-09-22 08:58:30 +02:00
Yih-Dar	06ee91aebc	Use CircleCI `store_test_results` (#26223 ) store_test_results Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-09-22 08:56:54 +02:00
Gema Parreño	587b7b16ce	[QUICK FIX LINK] Update trainer.py (#26293 ) * Update trainer.py Fix link * Update src/transformers/trainer.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update trainer.py --------- Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>	2023-09-22 03:33:29 +02:00
Matt	000e52aec8	More error message fixup, plus some linebreaks! (#26296 ) * More error message fixup, plus some linebreaks! * Update src/transformers/dynamic_module_utils.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update src/transformers/dynamic_module_utils.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update src/transformers/dynamic_module_utils.py Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> --------- Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>	2023-09-21 17:36:05 +01:00
Yoach Lacombe	9a30753485	Porting the torchaudio kaldi fbank implementation to audio_utils (#26182 ) * add kaldi fbank * make style * add herz_to_mel_kaldi tests * add mel to hertz kaldi test * integration tests * correct test and remove comment * make style * Apply suggestions from code review Co-authored-by: Sanchit Gandhi <93869735+sanchit-gandhi@users.noreply.github.com> * change parameter name * Apply suggestions from Arthur review Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com> * Update remove_dc_offset description * fix bug + make style * fix error in using np.exp instead of np.power * make style --------- Co-authored-by: Sanchit Gandhi <93869735+sanchit-gandhi@users.noreply.github.com> Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>	2023-09-21 17:52:47 +02:00
Arthur	b132c1703e	update hf hub dependency to be compatible with the new tokenizers (#26301 )	2023-09-21 14:57:36 +02:00
Lysandre Debut	26ba56ccbd	Fix FSMT weight sharing (#26292 )	2023-09-21 14:46:05 +02:00
fxmarty	da971b2271	Keep relevant weights in fp32 when `model._keep_in_fp32_modules` is set even when `accelerate` is not installed (#26225 ) * fix bug where weight would not be kept in fp32 * nit * address review comments * fix test	2023-09-21 19:00:03 +09:00
Shijie Wu	e3a4bd2bee	add custom RMSNorm to `ALL_LAYERNORM_LAYERS` (#26227 ) * add LlamaRMSNorm to ALL_LAYERNORM_LAYERS * fixup * add IdeficsRMSNorm to ALL_LAYERNORM_LAYERS and fixup	2023-09-20 18:51:56 +02:00
Younes Belkada	0b5024ce72	[`Trainer`] Refactor trainer + bnb logic (#26248 ) * refactor trainer + bnb logic * remove logger.info * oops	2023-09-20 17:38:59 +02:00
Arthur	f94c9b3d86	include changes from llama (#26260 ) * include changes from llama * add a test	2023-09-20 17:19:30 +02:00

1 2 3 4 5 ...

14070 Commits