transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-31 02:02:21 +06:00

Author	SHA1	Message	Date
Pengfei Liu	8ad06b7c13	using raw string for regex to search <extra_id> (#21162 ) * using raw string for regex to search <extra_id> * fix the same issue in test file:`tokenization_t5.py`	2023-01-18 09:43:54 -05:00
Wang, Yi	8a17da2f7f	fix the issue that the output dict of jit model could not get [:2] (#21146 ) "TypeError: unhashable type: 'slice'" Signed-off-by: Wang, Yi A <yi.a.wang@intel.com> Signed-off-by: Wang, Yi A <yi.a.wang@intel.com>	2023-01-18 09:41:28 -05:00
Peter Lin	e1ad188641	Fix git model for generate with beam search. (#21071 ) * Fix git model for generate with beam search. * Update comment * Fix bug on multi batch * Add generate tests * Clean up tests * Fix style Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2023-01-18 09:40:24 -05:00
Joao Gante	e15f0d73db	OPT: Fix batched generation with FLAX (#21150 ) * Fix Flax OPT numerical masking * re-enable test * add fix to bart and reintroduce copied from in opt	2023-01-18 14:24:53 +00:00
Jordi Mas	f4786d7f39	Fix typos in documentation (#21160 ) * Fix typos in documentation * Small fix * Fix formatting	2023-01-18 09:05:25 -05:00
Samuel Xu	defdcd2862	Remove Roberta Dependencies from XLM Roberta Flax and Tensorflow models (#21047 ) * Added flax model code * Added tf changes * missed some * Added copy comments * Added style hints * Fixed copy statements * Added suggested fixes * Made some fixes * Style fixup * Added necessary copy statements * Fixing copy statements * Added more copies * Final copy fix * Some bugfixes * Adding imports to init * Fixed up all make fixup errors * Fixed doc errors * Auto model changes	2023-01-18 07:49:39 -05:00
Younes Belkada	023f51fe16	`blip` support for training (#21021 ) * `blip` support for training * remove labels creation * remove unneeded `decoder_input_ids` creation * final changes - add colab link to documentation - reduction = mean for loss * fix nits * update link * clearer error message	2023-01-18 11:24:37 +01:00
Yih-Dar	c8849583ad	Make `test_save_pretrained_signatures` slow test (#21105 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-18 10:43:05 +01:00
Shogo Hida	14154f7238	Add Japanese translation to multilingual.mdx (#21084 ) * Create toctree for Japanese translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Copy English version Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add Japanese translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add Japanese translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> Signed-off-by: Shogo Hida <shogo.hida@gmail.com>	2023-01-18 10:08:18 +01:00
Wonhyeong Seo	30c12301f8	🌐 [i18n-KO] Translated `installation.mdx` to Korean (#20948 ) docs: ko: installation.mdx	2023-01-18 10:05:23 +01:00
layjain	44caf4f6f4	Fixed num_channels!=3 normalization training (#20630 ) * Fixed num_channels!=3 normalization training * empty commit to trigger CI * Empty-Commit for CircleCI * Empty-Commit * Empty Commit try-3: https://discuss.circleci.com/t/github-code-checkout-suddenly-failing/31558 * Empty commit to trigger CI Co-authored-by: Lay Jain <layjain@basil.csail.mit.edu> Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-17 13:06:20 -05:00
Sherman Siu	865da84abb	Add Epsilon- and Eta-Sampling (#21121 ) * Add epsilon- and eta-sampling. Add epsilon- and eta-sampling, following the official code from https://github.com/john-hewitt/truncation-sampling and adapting to be more configurable, as required by Huggingface transformers. * Add unit tests for epsilon- and eta-sampling. * Black: fix code formatting. * Fix docstring spacing. * Clean up newlines. * Fix implementation bugs and their associated tests. * Remove epsilon- and eta-sampling parameters from PretrainedConfig. * Clarify and clean up the documentation. * Remove parameters for PretrainedConfig test.	2023-01-17 13:04:32 -05:00
Maria Khalusova	0248810300	Refactoring of the text generate API docs (#21112 ) * initial commit, refactoring the text generation api reference * removed repetitive code examples * Refactoring the text generation docs to reduce repetition * make style	2023-01-17 12:23:48 -05:00
Maria Khalusova	d386fd646a	Add: An introductory guide for text generation (#21090 ) * Part of the "text generation" rework: adding a high-level overview of the text generation strategies * code samples update via make style * fixed a few formatting issues * Apply suggestions from review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * fixed spaces, and switched two links to markdown * Apply Steven's suggestions from review Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * new lines after headers to fix link rendering * review feedback addressed. added links to image captioning and audio transcription examples * minor capitalization fix * addressed the review feedback * Apply suggestions from review Co-authored-by: Joao Gante <joaofranciscocardosogante@gmail.com> * Applied review suggestions Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> Co-authored-by: Joao Gante <joaofranciscocardosogante@gmail.com>	2023-01-17 12:23:22 -05:00
Maria Khalusova	868d37165f	Add: tensorflow example for image classification task guide (#21038 ) * Added TF example for image classification * Code style polishing * code style polishing * minor polishing * fixed a link in a tip, and a typo in the inference TF content * Apply Amy's suggestions from review Co-authored-by: amyeroberts <22614925+amyeroberts@users.noreply.github.com> * Update docs/source/en/tasks/image_classification.mdx Co-authored-by: amyeroberts <22614925+amyeroberts@users.noreply.github.com> * review feedback addressed * make style * added PushToHubCallback with save_strategy="no" * minor polishing * added PushToHubCallback with save_strategy=no * minor polishing * Update docs/source/en/tasks/image_classification.mdx * added data augmentation Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> * make style Co-authored-by: amyeroberts <22614925+amyeroberts@users.noreply.github.com> Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>	2023-01-17 12:20:08 -05:00
NielsRogge	3a9bd972e2	Add resources (#20872 ) * Add resources * Add more resources * Remove pipeline tag * Add more resources * Add more resources Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2023-01-17 17:42:33 +01:00
Joao Gante	d96098c641	CLI: update hub PR URL (#21154 )	2023-01-17 16:36:47 +00:00
Sayak Paul	f3feaf7f22	Change variable name to prevent shadowing (#21153 ) fix: input -> input_string.	2023-01-17 11:29:23 -05:00
NielsRogge	cf028d0c3d	Add batch of resources (#20647 ) * Add resources * Add more resources * Add more resources * Add TAPAS * Fix pipeline tag * Fix pipeline tags * Remove pipeline tag * Remove depth-estimation tag * Update docs/source/en/model_doc/segformer.mdx Co-authored-by: Maria Khalusova <kafooster@gmail.com> * Apply suggestion * Fix segformer Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local> Co-authored-by: Maria Khalusova <kafooster@gmail.com>	2023-01-17 17:18:56 +01:00
Arthur	bb300ac686	Whisper Timestamp processor and prediction (#20620 ) * add draft logit processor * add template functions * update timesapmt processor parameters * draft script * simplify code * cleanup * fixup and clean * update pipeline * style * clean up previous idea * add tokenization utils * update tokenizer and asr output * fit whisper type * style and update test * clean test * style test * update tests * update error test * udpate code (not based on review yet) * update tokenization * update asr pipeline * update code * cleanup and update test * fmt * remove text verificatino * cleanup * cleanup * add model test * update tests * update code add docstring * update code and add docstring * fix pipeline tests * add draft logit processor add template functions update timesapmt processor parameters draft script simplify code cleanup fixup and clean update pipeline style clean up previous idea add tokenization utils update tokenizer and asr output fit whisper type style and update test clean test style test update tests update error test udpate code (not based on review yet) update tokenization update asr pipeline update code cleanup and update test fmt remove text verificatino cleanup cleanup add model test update tests update code add docstring update code and add docstring fix pipeline tests * Small update. * Fixup. * Tmp. * More support. * Making `forced_decoder_ids` non mandatory for users to set. * update and fix first bug * properly process sequence right after merge if last * tofo * allow list inputs + compute begin index better * start adding tests * add the 3 edge cases * style * format sequences * fixup * update * update * style * test passes, edge cases should be good * update last value * remove Trie * update tests and expec ted values * handle bigger chunk_length * clean tests a bit * refactor chunk iter and clean pipeline * update tests * style * refactor chunk iter and clean pipeline * upade * resolve comments * Apply suggestions from code review Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> * take stride right into account * update test expected values * Update code based on review Co-authored-by: sgugger <sylvain.gugger@gmail.com> Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> Co-authored-by: sgugger <sylvain.gugger@gmail.com>	2023-01-17 15:50:09 +01:00
Nicolas Patry	25ddd91b24	Fixing offline mode for pipeline (when inferring task). (#21113 ) * Fixing offline mode for pipeline (when inferring task). * Update src/transformers/pipelines/__init__.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Updating test to reflect change in exception. * Fixing offline mode. * Clean. Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2023-01-17 15:24:40 +01:00
Sherman Siu	8896ebb9a9	Clarify and add missing typical_p argument docstring. (#21095 ) * Clarify and add missing typical_p docstring. * Make the docstring easier to understand. * Clarify typical_p docstring Accept the suggestion by @stevhliu for paraphrasing the docstring. Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> * Use the same docstring as in GenerationConfig Follow the suggestion suggested by @stevhliu in the pull request conversation. * Fix docstring spacing. Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2023-01-17 09:23:47 -05:00
Sayak Paul	f30bcd5357	feat: add standalone guide on XLA support. (#21141 ) * feat: add standalone guide on XLA support. Co-authored-by: Joao Gante <joaofranciscocardosogante@gmail.com> * Empty commit to trigger CI * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * address PR comments. Co-authored-by: Joao Gante <joaofranciscocardosogante@gmail.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2023-01-17 15:07:59 +01:00
Nick Hill	3bbc2451b1	Small simplification to TopKLogitsWarper (#21130 ) The max of top_k and min_tokens_to_keep performed on every call can just be done once up-front.	2023-01-17 09:06:03 -05:00
amyeroberts	0dde58978a	Rename test_feature_extraction files (#21140 ) * Rename files * Update file names in tests	2023-01-17 14:04:07 +00:00
Joao Gante	7b5e943cb6	Generate: TF contrastive search must pop `use_cache` from `model_kwargs` (#21149 )	2023-01-17 13:42:52 +00:00
Joao Gante	7f3dab39b5	TF: serializable hubert (#20966 ) * serializable hubert	2023-01-17 13:07:37 +00:00
Matt	e5dcceb82c	Fixes to TF collators (#21143 ) * Add num_workers for prepare_tf_dataset * Bugfix in the default collator and change default tensor type * Remove the "num_workers" arg and move it to a new PR	2023-01-17 12:18:56 +00:00
Alara Dirik	2411f0e465	Add Mask2Former (#20792 ) * Adds Mask2Former to transformers Co-authored-by: Shivalika Singh <shivalikasingh95@gmail.com> Co-authored-by: Shivalika Singh <73357305+shivalikasingh95@users.noreply.github.com> Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2023-01-16 20:37:07 +03:00
NielsRogge	9edf375834	[GIT] Fix training (#21133 ) * Fix training * Add test * Fix failing tests Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2023-01-16 15:37:38 +01:00
Yih-Dar	0fb27dc988	Update `TFTapasEmbeddings` (#21107 ) Update TFTapasEmbeddings Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-16 15:29:50 +01:00
Clémentine Fourrier	4bbbabcb2c	Added clefourrier as ref point for graph models in bug reports (#21139 ) * Added clefourrier as ref point for graph models in bug reports * Update PULL_REQUEST_TEMPLATE.md	2023-01-16 15:12:42 +01:00
Yih-Dar	a45914193a	Fix `RealmModelIntegrationTest.test_inference_open_qa` (#21136 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-16 15:09:52 +01:00
Susnato Dhar	a5327c6a9a	Fixed issue #21053 (#21065 ) Co-authored-by: susnato <susnato@tensorflow123456@gmail.com>	2023-01-16 15:06:35 +01:00
Nicolas Patry	488a179ce1	Fixing batching pipelines on single items for ChunkPipeline (#21132 ) * Fixing #20783 * Update src/transformers/pipelines/base.py * Fixing some tests. * Fixup. * Remove ffmpeg dep + a bit more relaxed for bigbird QA precision. * Better dataset. * Prevent failing on TF. * Better condition. We can't use `can_use_iterator` since we cannot use it directly.	2023-01-16 15:04:27 +01:00
Silver	fa906a264b	Add `min_new_tokens` argument in generate() (implementation based on `MinNewTokensLengthLogitsProcessor`) (#21044 ) add a new parameter min_new_tokens for generate()	2023-01-16 15:02:08 +01:00
guillaume-be	125f137562	[LongT5] Remove duplicate encoder_attention_mask default value check (#21124 ) - Remove duplicate encoder_attention_mask default value assignment	2023-01-16 14:26:56 +01:00
NielsRogge	05b8e25fff	[VideoMAE] Fix docstring (#21111 ) Fix docstring Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2023-01-16 09:39:35 +01:00
NielsRogge	4ed89d48ab	Add UperNet (#20648 ) * First draft * More improvements * Add convnext backbone * Add conversion script * Add more improvements * Comment out to_dict * Add to_dict method * Add default config * Fix config * Fix backbone * Fix backbone some more * Add docs, auto mapping, tests * Fix some tests * Fix more tests * Fix more tests * Add conversion script * Improve conversion script * Add support for getting reshaped undownsampled hidden states * Fix forward pass * Add print statements * Comment out set_shift_and_window_size * More improvements * Correct downsampling layers conversion * Fix style * First draft * Fix conversion script * Remove config attribute * Fix more tests * Update READMEs * Update ConvNextBackbone * Fix ConvNext tests * Align ConvNext with Swin * Remove files * Fix index * Improve docs * Add output_attentions to model forward * Add backbone mixin, improve tests * More improvements * Update init_weights * Fix interpolation of logits * Add UperNetImageProcessor * Improve image processor * Fix image processor * Remove print statements * Remove script * Update import * Add image processor tests * Remove print statements * Fix test * Add integration test * Add convnext integration test * Update docstring * Fix README * Simplify config * Apply suggestions * Improve docs * Rename class * Fix test_initialization * Fix import * Address review * Fix confg * Convert all checkpoints * Fix default backbone * Usage same processor as segformer * Apply suggestions * Fix init_weights, update conversion scripts * Improve config * Use Auto API instead of creating a new image processor * Fix docs * Add doctests * Remove ResNetConfig dependency * Add always_partition argument * Fix rebaseé * Improve docs * Convert checkpoints Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local> Co-authored-by: Niels Rogge <nielsrogge@Nielss-MBP.localdomain>	2023-01-16 09:39:13 +01:00
TK Buristrakul	5db9abde43	Fixed typo in docstring (#21115 ) Fixed typo	2023-01-15 11:03:30 +01:00
Yusuke Oda	15adc24208	Use raw string for regex in tokenization_t5_fast.py (#21125 ) Suppress deprecation warning	2023-01-15 10:56:31 +01:00
Arthur	056218dab1	[CI-doc-daily] Remove RobertaPreLayernorm random tests (#20992 ) * Remove random output * remove values * fix copy statements	2023-01-14 19:47:32 +01:00
Sylvain Gugger	c8f35a9ce3	Rework automatic code samples in docstrings (#20757 ) * Rework automatic code samples in docstrings * ImageProcessor->AutoImageProcessor * Add models to fix copies * Last typos * A couple more models * Fix copies	2023-01-14 09:49:36 +01:00
Shogo Hida	7f65d2366a	Add Spanish translation to community.mdx (#21055 ) * Add community to toctree Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Copy English content Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add some translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add some translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add some translations Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Fix position of community Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Fix translation Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add translation Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add translation Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add translation Signed-off-by: Shogo Hida <shogo.hida@gmail.com> * Add translation Signed-off-by: Shogo Hida <shogo.hida@gmail.com> Signed-off-by: Shogo Hida <shogo.hida@gmail.com>	2023-01-14 09:25:05 +01:00
Steven Liu	f58248b824	Update task summary part 1 (#21014 ) * first draft of new task summary * make style * review * apply feedback * apply feedbacks * final touches	2023-01-13 11:01:53 -08:00
Arthur	95f0dd2123	[Tokenizers] Fix a small typo (#21104 ) * typo * change name in `__repr__` * fix my mistake	2023-01-13 16:21:34 +01:00
Yih-Dar	b210c83a78	Fix `torchscript` tests for `AltCLIP` (#21102 ) fix torchscript tests for AltCLIP Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-13 10:03:19 +01:00
Yih-Dar	b3a0aad37d	Fix past CI (#20967 ) * Fix for Past CI * make style * clean up * unindent 2 blocks Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-12 18:04:21 +01:00
Stas Bekman	41b0564b35	[bnb optim] fixing test (#21030 ) * [bnb optim] fixing test * force 1 gpu * fix * fix * fix * finalize * improve commentary * fix * cleanup * more fixes	2023-01-12 08:52:54 -08:00
Yih-Dar	212829ade6	Remove more unused attributes in config classes (#21000 ) * Remove gradient_checkpointing from MarkupLMConfig * Remove predict_special_tokens from OpenAIGPTConfig * Remove enable_cls from RoCBertConfig * Remove batch_size from TrajectoryTransformerConfig * Remove searcher_seq_len from RealmConfig * Remove feat_quantizer_dropout from WavLMConfig * Remove position_biased_input from SEWDConfig * Remove max_source_positions from Speech2Text2Config Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-01-12 13:32:04 +01:00

1 2 3 4 5 ...

11784 Commits