transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-31 02:02:21 +06:00

Author	SHA1	Message	Date
dependabot[bot]	7d0486c106	Bump mako in /examples/research_projects/decision_transformer (#19077 ) Bumps [mako](https://github.com/sqlalchemy/mako) from 1.2.0 to 1.2.2. - [Release notes](https://github.com/sqlalchemy/mako/releases) - [Changelog](https://github.com/sqlalchemy/mako/blob/main/CHANGES) - [Commits](https://github.com/sqlalchemy/mako/commits) --- updated-dependencies: - dependency-name: mako dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com> Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-09-16 22:15:02 +02:00
Lysandre Debut	56c548f17c	Note about developer mode (#19075 )	2022-09-16 22:12:59 +02:00
Sylvain Gugger	9017ba4ca4	Fix tokenizer load from one file (#19073 ) * Fix tokenizer load from one file * Add a test * Style Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>	2022-09-16 16:11:47 -04:00
fxmarty	773314ab80	replace logger.warn by logger.warning (#19068 )	2022-09-16 21:01:57 +02:00
Partho	5e636eee4a	Add type hints for PyTorch UniSpeech, MPNet and Nystromformer (#19039 ) * added type hints pytorch unispeech * added type hints pytorch MPNet * added type hints nystromformer * resolved copy inconsistencies * make fix-copies Co-authored-by: matt <rocketknight1@gmail.com>	2022-09-16 17:59:40 +01:00
Joao Gante	658010c739	TF: tests for (de)serializable models with resized tokens (#19013 ) * resized models that we can actually load * separate embeddings check * add test for embeddings out of bounds * add fake slows	2022-09-16 16:38:08 +01:00
Yih-Dar	70ba10e6d4	Fix `LeViT` checkpoint (#19069 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-16 16:23:58 +02:00
Omar Sanseviero	bc5d0b1046	Automatically tag CLIP repos as zero-shot-image-classification (#19064 ) * Add CLIP to zero-shot-image-classification * Make mapping private as it's not used for AutoClassing	2022-09-16 15:40:38 +02:00
Sylvain Gugger	820cb97a3f	Organize test jobs (#19058 ) * Tests conditional run * Syntax * Deps * Try early exit * Another way * Test with no tests to run * Test all * Typo * Try this way * With tests to run * Mostly finished * Typo * With a modification in one file only * No change, no tests * Final cleanup * Address review comments	2022-09-16 09:19:51 -04:00
Jim Briggs	d63bdf78d4	Add FP32 cast in ConvNext LayerNorm to prevent rounding errors with FP16 input (#18746 ) * Adding cast to fp32 in convnext layernorm to prevent rounding errors in the case of fp16 input * Trigger CI	2022-09-16 08:42:57 -04:00
Tom Aarsen	532ca05079	[doc] Fix link in PreTrainedModel documentation (#19065 )	2022-09-16 07:31:39 -04:00
Michael Benayoun	c603c80f46	FX support for ConvNext, Wav2Vec2 and ResNet (#19053 ) * Support for ConvNext * Support for Wav2Vec2 * Support for Resnet * Fix small issue in test_modeling_convnext	2022-09-16 10:57:41 +02:00
Younes Belkada	c8e40d6fa1	fix `use_cache` (#19060 ) - set `use_cache` to `True` for consistency with other `transformers` models	2022-09-16 09:07:02 +02:00
Colin Dean	0b5c7e4838	Adds package and requirement spec output to version check exception (#18702 ) * Adds package and requirement spec output to version check exception It's difficult to understand what package is affected when `got_ver` here comes back None, so output the requirement and the package. The requirement probably contains the package but let's output both for good measure. Non-exhaustive references for this problem aside from my own encounter: * https://stackoverflow.com/questions/70151167/valueerror-got-ver-is-none-when-importing-tensorflow * https://discuss.huggingface.co/t/valueerror-got-ver-is-none/17465 * https://github.com/UKPLab/sentence-transformers/issues/1186 * https://github.com/huggingface/transformers/issues/13356 I speculate that the root of the error comes from a conflict of conda-managed and pip-managed Python packages but I've not yet proven this. * Combines version presence check and streamlines exception message See also: https://github.com/huggingface/transformers/pull/18702#discussion_r953223275 Co-authored-by: Stas Bekman <stas@stason.org>	2022-09-15 12:53:36 -07:00
Shijie Wu	f3d3863255	fix arg name in BLOOM testing and remove unused arg document (#18843 )	2022-09-15 20:25:32 +02:00
Yih-Dar	16242e1bf0	Run `torchdynamo` tests (#19056 ) * Enable torchdynamo tests * make style Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-15 11:10:16 -07:00
Sylvain Gugger	f7ce4f1ff7	Fix custom tokenizers test (#19052 ) * Fix CI for custom tokenizers * Add nightly tests * Run CI, run! * Fix paths * Typos * Fix test	2022-09-15 11:31:09 -04:00
Nicolas Patry	68bb33d770	Fixing OPT fast tokenizer option. (#18753 ) * Fixing OPT fast tokenizer option. * Remove dependency on `pt`. * Move it to GPT2 tokenization tests. * Added a few tests.	2022-09-15 17:12:58 +02:00
Ekagra Ranjan	578e18e002	🚨🚨🚨 Optimize Top P Sampler and fix edge case (#18984 ) * init PR * optimize top p and add edge case * styling * style * revert tf and flax test * add edge case test for FLAX and TF * update doc with smallest set sampling for top p * make style	2022-09-15 15:50:11 +02:00
Sylvain Gugger	2700ba66d9	Move cache: expand error message (#19051 )	2022-09-15 09:39:59 -04:00
Matt	2322eb8e2f	Update serving signatures and make sure we actually use them (#19034 ) * Override save() to use the serving signature as the default * Replace int32 with int64 in all our serving signatures * Remember one very important line so as not to break every test at once * Dtype fix for TFLED * dtype fix for shift_tokens_right in general * Dtype fixes in mBART and RAG * Fix dtypes for test_unpack_inputs * More dtype fixes * Yet more mBART + RAG dtype fixes * Yet more mBART + RAG dtype fixes * Add a check that the model actually has a serving method	2022-09-15 14:34:22 +01:00
lewtun	9b80a0bc18	Pin minimum PyTorch version for BLOOM ONNX export (#19046 )	2022-09-15 15:22:31 +02:00
Yih-Dar	0a42b61ede	Fix `test_save_load` for `TFViTMAEModelTest` (#19040 ) * Fix test_save_load for TFViTMAEModelTest Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-15 15:21:57 +02:00
amyeroberts	30a28f5227	Update image segmentation pipeline test (#18731 ) * Updated test values The image segmentation pipeline tests - tests/pipelines/test_pipelines_image_segmentation.py - were failing after the merging of #1849 (`49e44b216b`). This was due to the difference in rescaling. Previously the images were rescaled by `image = image / 255`. In the new commit, a `rescale` method was added, and images rescaled using `image = image * scale`. This was known to cause small differences in the processed images (see [PR comment](https://github.com/huggingface/transformers/pull/18499#discussion_r940347575)). Testing locally, changing the `rescale` method to divide by a scale factor (255) resulted in the tests passing. It was therefore decided the test values could be updated, as there was no logic difference between the commits. * Use double quotes, like previous example * Fix up	2022-09-15 07:32:31 -04:00
Younes Belkada	7743caccb9	[bnb] Small improvements on utils (#18646 ) * Small replacement - replace `modules_to_not_convert` by `module_to_not_convert` * refactor a bit - changed variables name - now output a list - change error message * make style * add list * make style * change args name Co-authored-by: stas00 <stas00@users.noreply.github.com> * fix comment * fix typo Co-authored-by: stas00 <stas00@users.noreply.github.com> * Update src/transformers/modeling_utils.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: stas00 <stas00@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2022-09-15 13:01:19 +02:00
Stas Bekman	8edf196310	[doc] debug: fix import (#19042 ) correct the import statement	2022-09-14 16:29:58 -07:00
Hakjin Lee	abca1741cf	Fix a broken link for deepspeed ZeRO inference in the docs (#19001 ) * Fix a broken link for deepspeed ZeRO inference * fix link Co-authored-by: Stas Bekman <stas@stason.org>	2022-09-14 16:21:06 -07:00
Lysandre	16913b3c92	Dev version	2022-09-14 14:58:20 -04:00
Sylvain Gugger	3774010161	Automate check for new pipelines and metadata update (#19029 ) * Automate check for new pipelines and metadata update * Add Datasets to quality extra	2022-09-14 14:06:49 -04:00
SaulLu	0efbb6e93e	fix GPT2 token's `special_tokens_mask` when used with `add_bos_token=True` (#19036 )	2022-09-14 19:32:12 +02:00
Sylvain Gugger	0e24548081	Add safeguards for CUDA kernel load in Deformable DETR (#19037 )	2022-09-14 13:28:40 -04:00
Joao Gante	31be02f14b	TF: tf.debugging assertions without tf.running_eagerly() protection (#19030 )	2022-09-14 18:19:15 +01:00
lewtun	693ba2cc79	Fix GPT-NeoX doc examples (#19033 )	2022-09-14 17:53:42 +02:00
Sylvain Gugger	4eb36f2921	Mark right save_load test as slow (#19031 )	2022-09-14 10:38:39 -04:00
Shinya Otani	f5f430e5c8	Add support for Japanese GPT-NeoX-based model by ABEJA, Inc. (#18814 ) * add gpt-neox-japanese model and tokenizer as new model * Correction to PR's comment for GPT NeoX Japanese - Fix to be able to use gpu - Add comment # Copied... at the top of RotaryEmbedding - Implement nn.Linear instead of original linear class - Add generation test under @slow * fix bias treatment for gpt-neox-japanese * Modidy gpt-neox-japanese following PR - add doc for bias_dropout_add - style change following a PR comment * add document for gpt-neox-japanese * remove unused import from gpt-neox-japanese * fix README for gpt-neox-japanese	2022-09-14 10:17:40 -04:00
Yih-Dar	6a9726ec0e	Fix `DocumentQuestionAnsweringPipelineTests` (#19023 ) * Fix DocumentQuestionAnsweringPipelineTests Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-14 16:13:20 +02:00
Sylvain Gugger	1207deb806	Typo fix	2022-09-14 10:02:14 -04:00
Sylvain Gugger	e1224a2a0f	Making save_load test slow as it times out	2022-09-14 10:01:22 -04:00
Sylvain Gugger	0b567aa430	Add Document QA pipeline metadata (#19028 )	2022-09-14 09:25:15 -04:00
Yih-Dar	77b18783c2	Fix CI for `PegasusX` (#19025 ) * Skip test_torchscript_output_attentions for PegasusXModelTest * fix test_inference_no_head * fix test_inference_head * fix test_seq_to_seq_generation Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-14 14:45:00 +02:00
Partho	77ea35b93a	added type hints (#19015 )	2022-09-14 12:58:05 +01:00
NielsRogge	fc21c9be62	[CookieCutter] Clarify questions (#18959 ) * Clarify cookiecutter questions * Update first question Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>	2022-09-14 13:52:54 +02:00
Sylvain Gugger	6f8f2f6a77	Make AutoProcessor a magic loading class for all modalities (#18963 ) * Make AutoProcessor a magic loading class for all modalities * Quality	2022-09-14 07:36:12 -04:00
Sylvain Gugger	a2a3afbc8d	PyTorch >= 1.7.0 and TensorFlow >= 2.4.0 (#19016 )	2022-09-14 07:19:02 -04:00
Ekagra Ranjan	9f4acd059f	Generate: add missing comments after refactoring of generate() (#18981 )	2022-09-14 11:06:29 +01:00
NielsRogge	59407bbeb3	Add Deformable DETR (#17281 ) * First draft * More improvements * Improve model, add custom CUDA code * Import torch before * Add script that imports custom layer * Add everything in new ops directory * Import custom layer in modeling file * Fix ARCHIVE_MAP typo * Creating the custom kernel on the fly. * Import custom layer in modeling file * More improvements * Fix CUDA loading * More improvements * Improve conversion script * Improve conversion script * Make it work until encoder_outputs * Make forward pass work * More improvements * Make logits match original implementation * Make implementation also support single_scale model * Add support for single_scale and dilation checkpoint * Add support for with_box_refine model * Support also two stage model * Improve tests * Fix more tests * Make more tests pass * Upload all models to the hub * Clean up some code * Improve decoder outputs * Rename intermediate hidden states and reference points * Improve model outputs * Move tests to dedicated folder * Improve model outputs * Fix retain_grad test * Improve docs * Clean up and make test_initialization pass * Improve variable names * Add copied from statements * Improve docs * Fix style * Improve docs * Improve docs, move tests to model folder * Fix rebase * Remove DetrForSegmentation from auto mapping * Apply suggestions from code review * Improve variable names and docstrings * Apply some more suggestions from code review * Apply suggestion from code review * better docs and variables names * hint to num_queries and two_stage confusion * remove asserts and code refactor * add exception if two_stage is True and with_box_refine is False * use f-strings * Improve docs and variable names * Fix code quality * Fix rebase * Add require_torch_gpu decorator * Add pip install ninja to CI jobs * Apply suggestion of @sgugger * Remove DeformableDetrForObjectDetection from auto mapping * Remove DeformableDetrModel from auto mapping * Add model to toctree * Add model back to mappings, skip model in pipeline tests * Apply @sgugger's suggestion * Fix imports in the init * Fix copies * Add CPU implementation * Comment out GPU function * Undo previous change * Apply more suggestions * Remove require_torch_gpu annotator * Fix quality * Add logger.info * Fix logger * Fix variable names * Fix initializaztion * Add missing initialization * Update checkpoint name * Add model to doc tests * Add CPU/GPU equivalence test * Add Deformable DETR to pipeline tests * Skip model for object detection pipeline Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com> Co-authored-by: Nouamane Tazi <nouamane98@gmail.com> Co-authored-by: Sylvain Gugger <Sylvain.gugger@gmail.com>	2022-09-14 11:45:21 +02:00
Ahmed Elnaggar	5a70a77bfa	Add Support to Gradient Checkpointing for LongT5 (#18977 ) FlaxLongT5PreTrainedModel is missing "enable_gradient_checkpointing" function. This gives an error if someone tries to enable gradient checkpointing for longt5. This pull request fixes it.	2022-09-14 09:12:51 +01:00
Joao Gante	4157e3cd7e	new length penalty docstring (#19006 )	2022-09-13 13:16:36 -04:00
Sylvain Gugger	f89f16a51e	Re-add support for single url files in objects download (#19014 )	2022-09-13 13:11:24 -04:00
Yih-Dar	ad5045e3e3	add missing `require_tf` for `TFOPTGenerationTest` (#19010 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-09-13 18:10:11 +02:00

1 2 3 4 5 ...

10680 Commits