transformers

mirror of https://github.com/huggingface/transformers.git synced 2025-07-18 03:58:25 +06:00

Author	SHA1	Message	Date
Jungnerd	abbc96a214	[i18n-KO] fix: docs: ko: sagemaker anchors and `_toctree.yml` (#22549 ) fix: docs: ko: sagemaker anchors and `_toctree.yml` Co-authored-by: Hyeonseo Yun <0525_hhgus@naver.com> Co-authored-by: Gabriel Yang <gabrielwithhappy@gmail.com> Co-authored-by: Sohyun Sim <96299403+sim-so@users.noreply.github.com> Co-authored-by: Na Yeon Han <nayeon2.han@gmail.com> Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com>	2023-04-17 07:41:52 -04:00
Na Yeon Han	18c894814e	🌐 [i18n-KO] Translated `custom_models.mdx` to Korean (#22534 ) docs: ko: translated `custom_models.mdx` Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> Co-authored-by: Gabriel Yang <gabrielwithhappy@gmail.com> Co-authored-by: Jungnerd <46880056+jungnerd@users.noreply.github.com>	2023-04-17 07:39:53 -04:00
Yih-Dar	76d24f1a83	Fix `test_word_time_stamp_integration` for `Wav2Vec2ProcessorWithLMTest` (#22800 ) * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-17 12:41:55 +02:00
bcol	28f26c107b	Generate: add CJK support to TextStreamer (#22664 )	2023-04-15 10:35:08 +01:00
oscar-garzon	fb3aa06cb6	Move labels to the same device as logits for Whisper (#22779 )	2023-04-14 19:08:41 -04:00
amyeroberts	20e54e49fa	Indexing fix - CLIP checkpoint conversion (#22776 ) * Indexing fix - CLIP checkpoint conversion * Fix up	2023-04-14 19:12:47 +01:00
Joao Gante	895ae3b5c4	Seq2SeqTrainer: Evict decoder_input_ids only when it is created from labels (#22772 )	2023-04-14 17:45:14 +01:00
Mayank Agarwal	daf53241d6	Fix word_ids hyperlink (#22765 ) * Fix word_ids hyperlink * Add suggested fix	2023-04-14 16:18:15 +01:00
Matt	06e737fbaf	Tweak ESM tokenizer for Nucleotide Transformer (#22770 ) * If EOS is None, don't add it to sequences * If EOS is None, don't add it to sequences	2023-04-14 15:18:43 +01:00
Sohyun Sim	c8df3900c8	[WIP]🌐 [i18n-KO] Translated `tutorial/proprecssing.mdx` to Korean (#22578 ) * add ko preprocessing * translate preprocessing.mdx to korean * translate preprocessing.mdx * Update preprocessing.mdx Fixed the line 273 as below: 또한, 특징 추출기에 `sampling_rate` 인자를 추가하여 발생할 수 있는 조용한 오류(silent errors)를 더 잘 디버깅하는 것을 권장합니다. * translate Image part * translated preprocess.mdx * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * Update docs/source/ko/preprocessing.mdx * Update docs/source/ko/preprocessing.mdx * Update docs/source/ko/preprocessing.mdx * Update docs/source/ko/preprocessing.mdx * Update docs/source/ko/preprocessing.mdx * Update docs/source/ko/preprocessing.mdx * fixed translation --------- Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com>	2023-04-14 07:26:44 -04:00
Yih-Dar	53c710d17b	Fix failing torchscript tests for `CpmAnt` model (#22766 ) * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-14 12:53:45 +02:00
Alexander Ljungberg	d2ffc3fc48	Fix a mistake in Llama weight converter log output. (#22764 ) Fixed string format; better tokenizer message. Before: `Saving a {tokenizer_class} to {tokenizer_path}` After: `Saving a LlamaTokenizerFast to outdir.`	2023-04-14 10:26:45 +01:00
Joao Gante	9af845afc2	Generate: pin number of beams in BART test (#22763 )	2023-04-14 09:57:25 +01:00
Joao Gante	66b15efb20	Pix2struct: doctest fix (#22761 )	2023-04-14 09:40:39 +01:00
Sayak Paul	390e121fb5	[Examples] TPU-based training of a language model using TensorFlow (#21657 ) * add: tokenizer training script for TF TPU LM training. * add: script for preparing the TFRecord shards. * add: sequence of execution to readme. * remove limit from the tfrecord shard name. * Add initial train_model.py * Add basic training arguments and model init * Get up to the point of writing the data collator * Pushing progress so far! * Complete first draft of model training code * feat: grouping of texts efficiently. Co-authored-by: Matt <rocketknight1@gmail.com> * Add proper masking collator and get training loop working * fix: things. * Read sample counts from filenames * Read sample counts from filenames * Draft README * Improve TPU warning * Use distribute instead of distribute.experimental * Apply suggestions from code review Co-authored-by: Matt <Rocketknight1@users.noreply.github.com> * Modularize loading and add MLM probability as arg * minor refactoring to better use the cli args. * readme fillup. * include tpu and inference sections in the readme. * table of contents. * parallelize maps. * polish readme. * change script name to run_mlm.py * address PR feedback (round I). --------- Co-authored-by: Matt <rocketknight1@gmail.com> Co-authored-by: Matt <Rocketknight1@users.noreply.github.com>	2023-04-14 10:41:01 +05:30
Hyeonseo Yun	bfb3925fcb	🌐 [i18n-KO] Translated `sequence_classification.mdx` to Korean (#22655 ) * docs: ko: init: tasks/sequence_classification.mdx * docs: ko: revised: change voca in tasks/sequence_classification.mdx * docs: ko: revised: [RE] change voca in tasks/sequence_classification.mdx * docs: ko: revised: spell check and sentence naturally in tasks/sequence_classification.mdx * docs: ko: revised: spell check and consistent vocabulary in tasks/sequence_classification.mdx * docs: ko: revised: Add full stop and change voca in tasks/sequence_classification.mdx * docs: ko: revised: sync first section templates in tasks/sequence_classification.mdx Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com> * fix: revert use of full-stops to colons * colons are used to emphasize the code block that follows * @0525hhgus @wonhyeongseo docs: ko: revised: sync second section templates in tasks/sequence_classification.mdx Co-Authored-By: Wonhyeong Seo <wonhseo@kakao.com> * docs: ko: revised: change 'train', 'finetuning' in tasks/sequence_classification.mdx --------- Co-authored-by: Wonhyeong Seo <wonhseo@kakao.com>	2023-04-13 21:40:36 -04:00
Yih-Dar	a6752a7d3c	Fix `serving_output` for TF composite models (encoder-decoder like models) (#22743 ) * fix * style * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-13 23:45:22 +02:00
Yih-Dar	410b61ad7e	Revert (for now) the change on `Deta` in #22437 (#22750 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-13 21:32:29 +02:00
Joao Gante	9dfd6a4baa	Generate: handle text conditioning with multimodal encoder-decoder models (#22748 )	2023-04-13 19:51:13 +01:00
Ruiyang Sun	90ce374d14	fix(llama): fix LlamaTokenzier (#22746 ) Bug in LlamaTokenizer when #22742	2023-04-13 18:19:38 +01:00
Stas Bekman	d85bf95436	[trainer] update url (#22747 ) * [trainer] update url * style	2023-04-13 09:23:55 -07:00
Yih-Dar	656d41ab4c	Remove `DS_BUILD_AIO=1` (#22741 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-13 18:08:22 +02:00
Yih-Dar	32b08742a5	`DocumentQuestionAnsweringPipeline` only for fast ⚡ tokenizers (#22745 ) * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-13 17:22:59 +02:00
Gabriel Yang	4def2fe969	🌐 [i18n-KO] Translated `training.mdx` to Korean (#22670 ) translate training doc to Korean	2023-04-13 11:04:47 -04:00
Yih-Dar	7df1343292	Change `torch_dtype` to `str` when `saved_model=True` in `save_pretrained` for TF models (#22740 ) * fix --------- Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-13 15:52:16 +02:00
NielsRogge	8eb38f638d	[Pix2struct] Simplify generation (#22527 ) * Add model to doc tests * Remove generate and replace by prepare_inputs_for_generation * More fixes * Remove print statements * Update integration tests * Fix generate * Remove model from auto mapping * Use auto processor * Fix integration tests * Fix test * Add inference code snippet * Remove is_encoder_decoder * Update docs * Remove notebook link	2023-04-13 09:01:14 -04:00
Rinat	95e7057507	Make vilt, switch_transformers compatible with model parallelism (#22703 ) * Update modeling_vilt.py Vilt compatible with model parallelism * Update modeling_switch_transformers.py switch_transformers compatible with model parallelism	2023-04-13 06:50:30 -04:00
Joel Lamy-Poirier	89087597ba	Indexing fix for gpt_bigcode (#22737 ) Fix indexing	2023-04-13 11:00:37 +01:00
Elabonga Atuo	7ade6ef7d4	[Doctest] Add configuration_mvp.py (#22735 ) * added configuration file for mvp model * added configuration_mvp.py line to file	2023-04-13 08:19:18 +02:00
Elabonga Atuo	51007976ec	[Doctest] Add configuration_m2m_100.py (#22733 ) m2m-100-config for doctest	2023-04-13 08:17:07 +02:00
Sylvain Gugger	888c4a2ae0	v4.29.0.dev0	2023-04-12 20:04:29 -04:00
Matt	50f82e1282	Fix docstrings for TF BLIP (#22618 ) * Fix docstrings for TFBLIP * Fix missing line in TF port! * Use values from torch tests now other bugs fixed * Use values from torch tests now other bugs fixed * Fix doctest string	2023-04-12 17:46:41 +01:00
NielsRogge	ce06e4780e	Update warning levels (#22727 ) * Use different level * Remove futurewarning * Use warning_once * Update copies	2023-04-12 17:25:24 +01:00
Arthur	9858195481	add fast support and option (#22724 ) * add fast support and option * update based on review * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * Update src/transformers/models/llama/convert_llama_weights_to_hf.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * nit * add print * fixup --------- Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2023-04-12 18:10:04 +02:00
Michael Benayoun	10fab90fe2	`torch.distributed` group initialization for `torch_neuron` disabled when `optimum-neuron` is installed (#22728 ) * Make the process group initialization not happen if optimum_neuron is installed * Add warning * Remove list and added warning	2023-04-12 17:42:50 +02:00
Stas Bekman	1306b7d3ae	[tests] switch to torchrun (#22712 )	2023-04-12 08:25:45 -07:00
ARKA1112	d87ef00c31	Modify pipeline_tutorial.mdx (#22726 ) generator(model="openai/whisper-large") always returns error. As the error says the generator expects an input, just like the .flac file above. Even the generator object has no parameters called model. While there are parameters which can be passed to generator like 'batch_size' but to pass a model i believe the the parameter has to be passed while instantiating the pipeline and not as a parameter to the instance. I believe the correct term should be: generator = pipeline(model="openai/whisper-large", device=0)	2023-04-12 15:20:25 +01:00
Younes Belkada	370f0ca18c	[`bnb`] Let's make serialization of int8 models possible (#22177 ) * make serialization of int8 models possible * make fixup * add docs * add ability to push to hub and save pretrained * fixes * more addition * more tests * fix issues * change variable * clearer message * adapt from suggestions * few fixes * remove unused function * Update src/transformers/utils/quantization_config.py Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> * address last comments * last warning * clarify doc * protect import * Update src/transformers/modeling_utils.py * Apply suggestions from code review Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com> --------- Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>	2023-04-12 08:01:18 -04:00
pioliverse	523ca4e016	add model resources for CPMAnt (new) (#20906 ) * resolve conflicts * rebase and make style * test * test * test * rebase and make style * rebase and make style * tests * tests * rewrite some functions * rebase and make style * fix load_tf_weights_in_cpmant * reformat some unrelated files * upgrade quality * fix some bugs & docstring * add models and tests * solve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * tests * resolve conflicts * resolve conflicts * fix load_tf_weights_in_cpmant * reformat some unrelated files * upgrade quality * fix some bugs & docstring * save resolution * make style * delete redefinition code * reformat function * reformat * resolve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * tests * resolve conflicts * resolve conflicts * fix load_tf_weights_in_cpmant * reformat some unrelated files * upgrade quality * resolve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * resolve conflicts * fix load_tf_weights_in_cpmant * reformat some unrelated files * upgrade quality * resolve conflicts * make style * fix bugs and refactor * modify docstrings and make style * unify import format in __init__.py * fix import-altclp bug * fix copies to update index.md * fix unused config parameters * fix unused config parameters * fix unused config parameters * update README_ja.md * dummy commit for unit test * fix attention mask * add CPMAntTokenizer&-Fast to auto-mapping * drop redundant changes in README_ko * fix defaults in docstring * fix use_cache and some docstring * add missing args in tokenizer * modify tester inheritance * add is_jieba_available * fix some bugs * make style and fix-copies * add doctests * skip integration tests * add is_jieba_available * fix bugs in common tests * adjust docstrings and make style * add argument docstring * adjust code to some specifications * make style and fix-copies * add fast tokenization test * dummy commit for unit test * dummy commit for unit test * dummy commit for unit test * normalize some comments and names * Bert->CPMAnt * camel names and drop redundant codes * make style and fix-coies * add CpmTokenizerFast _import_structure * drop cpmanttokenizerfast in model_doc * fix some problems * fix CPMAnt tokenization for common test * make style and fixup * fix copies and fixup * fix bugs in tokenization test * dummy commit for connection failure in unittest * fix copies * drop trailing comma * fix decorator in tests * dummy commit for connection failure in unittest --------- Co-authored-by: Gong Baitao <gongbaitao11@gmail.com>	2023-04-12 07:33:20 -04:00
jprivera44	17503b00ea	Added parallel device usage for GPT-J (#22713 )	2023-04-12 07:31:27 -04:00
Arthur	b76e6ebd44	remove wrong doc in readme (#22723 )	2023-04-12 07:11:12 -04:00
amyeroberts	5a71977b8b	Update input values for docstring (#22631 )	2023-04-12 11:44:29 +01:00
Yih-Dar	fe1f5a639d	Fix decorator order (#22708 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-11 17:59:15 +02:00
Sylvain Gugger	1b1867d86b	Replace -100s in predictions by the pad token (#22693 ) * Replace -100s in predictions by the pad token * Style * Try to catch them all	2023-04-11 09:32:20 -04:00
Yih-Dar	ff73deeb0e	Remove 2 failing ONNX conversion tests (#22660 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-11 15:26:32 +02:00
Luc CAILLIAU	06b05d4575	Clarify stride option (#22684 ) * Clarify stride option * formatting	2023-04-11 14:06:54 +01:00
Mayank Agarwal	0224aaf67f	Enable naive Pipeline Parallelism training for Gpt neox japanese and san japanese (#22702 ) Move labels to same device as logits	2023-04-11 09:06:17 -04:00
Sylvain Gugger	28c19ab58d	Make it easier to develop without a dev install (#22697 ) * Make it easier to develop without a dev install * Remove ugly hack that doesn't work anyway	2023-04-11 08:41:53 -04:00
Yih-Dar	4c01231e67	Update some `MarkupLM` tests' expected values (#22667 ) fix Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>	2023-04-11 10:00:34 +02:00
Shahad Mahmud	151425ddb2	Model parallelism: Moving labels to same devices as the logits are (#22691 ) Model parallelism correct labels device	2023-04-10 12:22:53 -04:00

... 48 49 50 51 52 ...

15053 Commits