transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-16 11:08:23 +06:00

Author	SHA1	Message	Date
dependabot[bot]	7d45a2e81c	Bump numpy in /examples/research_projects/visual_bert (#15367 ) Bumps [numpy](https://github.com/numpy/numpy) from 1.19.2 to 1.21.0. - [Release notes](https://github.com/numpy/numpy/releases) - [Changelog](https://github.com/numpy/numpy/blob/main/doc/HOWTO_RELEASE.rst.txt) - [Commits](https://github.com/numpy/numpy/compare/v1.19.2...v1.21.0) --- updated-dependencies: - dependency-name: numpy dependency-type: direct:production ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2022-01-27 14:45:18 -05:00
Sylvain Gugger	a81fd35524	Fix tests_fetcher (#15376 )	2022-01-27 14:17:48 -05:00
Lysandre	eab338104d	Docs for version v4.16.0	2022-01-27 13:11:51 -05:00
Lysandre	f87db5e412	Release: v4.16.0	2022-01-27 13:06:33 -05:00
Matt	c43749289d	Example script for PushToHubCallback (#15375 ) * Example script for PushToHubCallback * Expanding description slightly	2022-01-27 16:16:24 +00:00
Sylvain Gugger	8f6454bfac	Add proper documentation for Keras callbacks (#15374 ) * Add proper documentation for Keras callbacks * Add dummies	2022-01-27 10:51:38 -05:00
Matt	2de90beeeb	Super-small fix stops us confusing Keras console logging by modifying its logs (#15373 )	2022-01-27 15:43:43 +00:00
Sylvain Gugger	fa6dce250f	Implement fixes for TrainingArguments doc (#15370 ) Co-authored-by: osanseviero <osanseviero@gmail.com> Co-authored-by: osanseviero <osanseviero@gmail.com>	2022-01-27 10:25:43 -05:00
SaulLu	ade7371a41	improve saving strategy of sentencepiece tokenizer (#15328 ) * add new test * add a feature to same the sentencepiece tokenizer model when the init file was deleted * update marian * update m2m_100 * fix marian * update speech to text * override test for layoutxlm * fix saving bartpho * remove harcoded values bartpho * special token string version * finish bartpho * override layoutxml test * add mbart * move special tokens list * format * Revert "format" This reverts commit `37a40df379`. * simplify list of string of special tokens * Re-write `self.fairseq_tokens_to_ids ` initialization logic with special tokens Co-authored-by: Sylvain Gugger <sylvain.gugger@gmail.com> Co-authored-by: Sylvain Gugger <sylvain.gugger@gmail.com>	2022-01-27 16:24:51 +01:00
Anton Lozhkov	196cce6e9b	Add a device argument to the eval script (#15371 ) * Device argument for the eval script * Default to none * isort	2022-01-27 15:58:55 +01:00
Matt	6beae766ee	Fix KerasMetricCallback prediction with generate() and inference of column names (#15351 ) * Fix prediction with generate() and the inference of column names Should now have very few differences with the PyTorch implementation * Minor edit to parent class * Update src/transformers/keras_callbacks.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Explaining the dict conversion * Putting main_input_name back * Fixes to main_input_name Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2022-01-27 14:13:23 +00:00
Sylvain Gugger	da5ef25db9	Push to hub save (#15327 ) * Adapt doc and push at every save * style	2022-01-27 09:00:54 -05:00
Patrick von Platen	9f831bdeaf	[DocTests Speech] Add doc tests for all speech models (#15031 ) * fix_torch_device_generate_test * remove @ * doc tests * up * up * fix doctests * adapt files * finish refactor * up * save intermediate * add more logic * new change * improve * next try * next try * next try * next try * fix final spaces * fix final spaces * improve * renaming * correct more bugs * finish wavlm * add comment * run on test runner * finish all speech models * adapt * finish	2022-01-27 14:29:31 +01:00
Sylvain Gugger	4df69506a8	Fix YosoConfig doc (#15353 )	2022-01-26 21:06:27 +01:00
Stas Bekman	fc8fc400e3	[docs] post-PR merge fix (#15355 ) * [docs] post-PR merge fix * Update docs/source/main_classes/deepspeed.mdx Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2022-01-26 11:23:32 -08:00
novice	99a2771189	Add YOSO (#15091 ) * Add cookiecutter files * Add cuda kernels and cpp files * Update modeling_yoso.py * Add .h files * Update configuration_yoso.py * Updates * Remove tokenizer * Code quality * Update modeling_yoso.py * Update modeling_yoso.py * Fix failing test * Update modeling_yoso.py * Fix code quality * Apply suggestions from code review Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Apply suggestions from code review * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Apply suggestions from code review and fix integration tests * Update src/transformers/models/yoso/modeling_yoso.py Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com> * Apply suggestions from code review * Fix copied from statement * Fix docstring * Fix code quality * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Apply suggestions and fix mask * Apply suggestions from code review * Fix code quality * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Fix docstrings * Fix code quality * Remove trailing whitespace * Update yoso.mdx * Move kernel loading to YosoEncoder * make style * Apply suggestions from code review Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Update src/transformers/models/yoso/modeling_yoso.py Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Add short summary to docs * Update docs/source/model_doc/yoso.mdx Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Update yoso.mdx * Update docs/source/model_doc/yoso.mdx Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Remove CausalLM model and add copied from * Remove autoregressive code * Remove unused imports * add copied from for embeddings * Fix code quality * Update docs/source/model_doc/yoso.mdx Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> * Apply suggestion from code review Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>	2022-01-26 19:18:29 +01:00
Sylvain Gugger	6292532fd1	Update doc writing guide (#15350 )	2022-01-26 12:54:11 -05:00
François REMY	19732cc07a	Fix 'eval_split_name' described as defaulting to 'train' (#15348 ) The default is correct (`test`) but the description is not.	2022-01-26 10:19:38 -05:00
Ngo Quang Huy	5d8b98608c	Fix deepspeed docs (#15346 )	2022-01-26 07:24:33 -05:00
Jacob Deppen	96161ac408	make table into valid Markdown table syntax (#15337 )	2022-01-26 07:10:00 -05:00
Yih-Dar	24e2fa1590	Fix encoder-decoder models when labels is passed (#15172 ) Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-01-26 10:14:46 +01:00
Maciej Pawłowski	e79a0faeae	Added missing code in exemplary notebook - custom datasets fine-tuning (#15300 ) * Added missing code in exemplary notebook - custom datasets fine-tuning Added missing code in tokenize_and_align_labels function in the exemplary notebook on custom datasets - token classification. The missing code concerns adding labels for all but first token in a single word. The added code was taken directly from huggingface official example - this [colab notebook](https://github.com/huggingface/notebooks/blob/master/transformers_doc/custom_datasets.ipynb). * Changes requested in the review - keep the code as simple as possible	2022-01-25 17:26:17 -05:00
Steven Liu	0501beb846	Add 🤗 Accelerate tutorial (#15263 ) * add accelerate tutorial * 🖍 apply feedback from review * 📝 make edits	2022-01-25 13:46:11 -06:00
NielsRogge	637e81752a	[Tests] Fix test (#15324 ) * Fix Swin device * Remove print statement	2022-01-25 15:48:25 +01:00
Sylvain Gugger	e695470794	Avoid using get_list_of_files (#15287 ) * Avoid using get_list_of_files in config * Wip, change tokenizer file getter * Remove call in tokenizer files * Remove last call to get_list_model_files * Better tests * Unit tests for new function * Document bad API	2022-01-25 09:41:21 -05:00
Sylvain Gugger	e65bfc0971	Try without bad instruction	2022-01-24 15:55:29 -05:00
Sylvain Gugger	81156d20cd	Add model like (#14992 ) * Add new model like command * Bad doc-styler * black and doc-styler, stop fighting! * black and doc-styler, stop fighting! * At last * Clean up * Typo * Bad doc-styler * Bad doc-styler * All good maybe? * Use constants * Add doc and type hints * More cleaning * Add doc * Fix Copied from * Doc template * Use typing.Pattern instead * Framework-specific files * Fixes * Select frameworks clean model init * Deal with frameworks in main init * fixes * Last fix * Prompt user for info * Delete exemple config * Last fixes * Add test config * Fix bug with model_type included in each other * Fixes * More fixes * More fixes * Adapt config * Remove print statements * Will fix tokenization later, leave it broken for now * Add test * Quality * Try this way * Debug * Maybe by setting the path? * Let's try another way * It should go better when actually passing the arg... * Remove debug statements and style * Fix config * Add tests * Test require the three backends * intermediate commit * Revamp pattern replacements and start work on feature extractors * Adapt model info * Finalize code for processors * Fix in main init additions * Finish questionnaire for processing classes * Fix file name * Fix for real * Fix patterns * Style * Remove needless warnings * Copied from should work now. * Include Copied form in blocks * Add test * More fixes and tests * Apply suggestions from code review Co-authored-by: Lysandre Debut <lysandre@huggingface.co> * Address review comment Co-authored-by: Lysandre Debut <lysandre@huggingface.co>	2022-01-24 15:25:10 -05:00
Patrick von Platen	457dd4392b	[Examples] Correct run ner label2id for fine-tuned models (#15017 ) * up * up * make style * apply sylvains suggestions * apply changes to accelerate as well * more changes * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2022-01-24 21:18:04 +01:00
Patrick von Platen	8d6acc6c29	[Beam Search] Correct returned beam scores (#14654 ) * better * save intermediate * finish code * up * docs * Apply suggestions from code review * up * add compute transition beam scores function to model and make sure scores are correct with eos * apply nicos comments * Apply suggestions from code review * another fix	2022-01-24 21:13:21 +01:00
novice	e239fc3b0b	Replace NystromformerTokenizer with AutoTokenizer (#15312 )	2022-01-24 16:33:43 +01:00
Patrick von Platen	dcaa5100c9	[LayoutLMV2 Tests] Make sure input is on GPU (#15314 ) * [LayoutLMV2 Tests] Make sure input is on GPU * correct empty line	2022-01-24 15:54:47 +01:00
Yih-Dar	c15bb3fe19	[Fix doc example] fix missing import jnp (#15291 ) * fix missing import jnp * Fix missing jax and k=1 Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-01-24 14:54:23 +01:00
Nicolas Patry	eac4aecc3d	Remove old debug code leftover. (#15306 )	2022-01-24 07:27:45 -05:00
Sylvain Gugger	2390b2cf65	Fix a typo in tag addition (#15286 ) * Fix a typo in tag addition * Put it back again	2022-01-24 07:21:42 -05:00
Kamal Raj	c972433a85	Update CONTRIBUTING.md (#15290 ) Fix typo in doc	2022-01-24 07:21:31 -05:00
Patrick von Platen	4bf97415a4	Update eval.py (#15310 )	2022-01-24 11:46:38 +01:00
Patrick von Platen	b7cb126ccc	[PyTorch-nightly-test] Fix Wav2Vec2 LM & Phoneme tests (#15272 ) * [PyTorch-nightly-test] Fix Wav2Vec2 LM & Phoneme tests * Update .github/workflows/self-nightly-scheduled.yml * change lines * Apply suggestions from code review	2022-01-24 10:53:53 +01:00
Sylvain Gugger	6ac77534bf	Refine errors for pretrained objects (#15261 ) * Refine errors for pretrained objects * PoC to avoid using get_list_of_files * Adapt tests to use new errors * Quality + Fix PoC * Revert "PoC to avoid using get_list_of_files" This reverts commit `cb93b7cae8`. * Revert "Quality + Fix PoC" This reverts commit `3ba6d0d4ca`. * Fix doc * Revert PoC * Add feature extractors * More tests and PT model * Adapt error message * Feature extractor tests * TF model * Flax model and test * Merge flax auto tests * Add tokenization * Fix test	2022-01-21 15:00:09 -05:00
Patrick von Platen	80af1048cf	[Wav2Vec2ProcessorWithLM] improve multi processing (#15247 ) * [Wav2Vec2ProcessorWithLM] improve multi processing * close pool	2022-01-21 18:30:10 +01:00
Sylvain Gugger	4cff3fae11	Second failing test	2022-01-21 12:19:28 -05:00
Sylvain Gugger	f6253147df	Skip failing test	2022-01-21 12:03:21 -05:00
Yih-Dar	7799b6128f	[Fix doc example] TFLayoutLMForTokenClassification: missing import tf (#15268 ) * fix import * remove import torch Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2022-01-21 11:18:11 -05:00
Patrick von Platen	11afb709ec	[Robust Speech Challenge] Add timeline (#15274 )	2022-01-21 17:12:09 +01:00
Evandros	3c3cf17a49	fix link (#15278 )	2022-01-21 09:52:13 -05:00
Ye Wang	95a75a715f	Specify providers explicitly in ORT session initialization (#15235 ) * Specify providers explicitly in ORT session initialization Co-authored-by: Ubuntu <wy@linux-v100.aidmrjtolptuzevavgwhrapqcd.jx.internal.cloudapp.net>	2022-01-21 15:49:29 +01:00
lewtun	833635e259	Move BART + ONNX example to research_projects (#15271 ) * Move BART + ONNX example to research_projects * Add author information	2022-01-21 14:47:34 +01:00
novice	183ce067e0	Fix (#15276 ) * Fix * make style * Remove trailing commas * make style	2022-01-21 08:46:15 -05:00
lewtun	b4ce313e6c	Prepare ONNX export for torch v1.11 (#15270 ) * Prepare ONNX export for torch v1.11	2022-01-21 14:28:19 +01:00
Sylvain Gugger	126bddd1ba	Add module_spec to new model	2022-01-21 08:12:44 -05:00
Jonas Kuball	c962c2adbf	Adds missing module_specs for usages of _LazyModule (#15230 ) * Add missing __spec__ for transformers.models.auto * Moves the __spec__-test to the UnitTest class * Adds module_spec to all instances of _LazyModule * Refactors an old test from pytest to unittest	2022-01-21 07:30:12 -05:00

1 2 3 4 5 ...

8821 Commits