Thomas Wolf
|
d2f21f08f5
|
Merge pull request #1092 from shijie-wu/xlm-tokenization
Added cleaned configuration properties for tokenizer with serialization - improve tokenization of XLM
|
2019-08-30 23:15:40 +02:00 |
|
Thomas Wolf
|
12b9cc9e26
|
Merge pull request #1110 from huggingface/automodels
Torch.hub now based on AutoModels - Updating AutoModels with AutoModelWithLMHead, Sequence Classification and Question Answering
|
2019-08-30 23:08:57 +02:00 |
|
thomwolf
|
bfe93a5a21
|
fix distilbert in auto tokenizer
|
2019-08-30 22:43:26 +02:00 |
|
thomwolf
|
256086bc69
|
clean up and simplify hubconf
|
2019-08-30 22:34:23 +02:00 |
|
thomwolf
|
80aa87d9a3
|
fix distilbert tokenizer
|
2019-08-30 22:24:23 +02:00 |
|
thomwolf
|
455a4c842c
|
add distilbert tokenizer
|
2019-08-30 22:20:51 +02:00 |
|
thomwolf
|
7a1f174a9d
|
update names of torch.hub to simpler names - update docstring
|
2019-08-30 22:20:44 +02:00 |
|
thomwolf
|
c665e0fcfe
|
Merge branch 'automodels' of https://github.com/huggingface/pytorch-transformers into automodels
|
2019-08-30 21:53:36 +02:00 |
|
LysandreJik
|
9b6e3b34d9
|
Docstrings
|
2019-08-30 14:09:02 -04:00 |
|
LysandreJik
|
dec8f4d6fd
|
Added DistilBERT models to all other AutoModels.
|
2019-08-30 13:52:18 -04:00 |
|
LysandreJik
|
bc29aa67a9
|
HubConf configuration
|
2019-08-30 12:48:55 -04:00 |
|
thomwolf
|
f35f612280
|
updating docstring for AutoModel
|
2019-08-30 12:48:55 -04:00 |
|
LysandreJik
|
7ca9653852
|
Pytorch Hub & AutoModels
|
2019-08-30 12:48:55 -04:00 |
|
LysandreJik
|
25e8389439
|
Tests for added AutoModels
|
2019-08-30 12:48:55 -04:00 |
|
LysandreJik
|
dc43215c01
|
Added multiple AutoModel classes: AutoModelWithLMHead, AutoModelForQuestionAnswering and AutoModelForSequenceClassification
|
2019-08-30 12:48:55 -04:00 |
|
VictorSanh
|
282c276e09
|
typos + file name coherence in distillation README
|
2019-08-30 12:02:29 -04:00 |
|
VictorSanh
|
803c1cc4ea
|
fix relative import bug cf Issue #1140
|
2019-08-30 12:01:27 -04:00 |
|
thomwolf
|
7044ed6b05
|
fix tokenizers serialization
|
2019-08-30 17:36:11 +02:00 |
|
Thomas Wolf
|
cd65c41a83
|
Merge branch 'master' into xlm-tokenization
|
2019-08-30 17:15:16 +02:00 |
|
thomwolf
|
69da972ace
|
added test and debug tokenizer configuration serialization
|
2019-08-30 17:09:36 +02:00 |
|
thomwolf
|
88111de07c
|
saving and reloading tokenizer configurations
|
2019-08-30 16:55:48 +02:00 |
|
Thomas Wolf
|
b66e9b4433
|
Merge pull request #1158 from rabeehk/master
regarding #1026 pull request
|
2019-08-30 16:30:33 +02:00 |
|
Thomas Wolf
|
0a2fecdf90
|
Merge branch 'master' into master
|
2019-08-30 16:30:08 +02:00 |
|
thomwolf
|
3871b8a107
|
adding xlm 17 and 100 models and config on aws
|
2019-08-30 16:28:42 +02:00 |
|
thomwolf
|
8678ff8df5
|
adding 17 and 100 xlm models
|
2019-08-30 16:26:04 +02:00 |
|
LysandreJik
|
e0caab0cf0
|
fix link
|
2019-08-30 10:09:17 -04:00 |
|
LysandreJik
|
a600b30cc3
|
Fix index number in documentation
|
2019-08-30 10:08:14 -04:00 |
|
LysandreJik
|
20c06fa37d
|
Added DistilBERT to documentation index
|
2019-08-30 10:06:51 -04:00 |
|
Rabeeh KARIMI
|
39eb31e11e
|
remove reloading tokenizer in the training, adding it to the evaluation part
|
2019-08-30 15:44:41 +02:00 |
|
Rabeeh KARIMI
|
350bb6bffa
|
updated tokenizer loading for addressing reproducibility issues
|
2019-08-30 15:34:28 +02:00 |
|
thomwolf
|
82462c5cba
|
Added option to setup pretrained tokenizer arguments
|
2019-08-30 15:30:41 +02:00 |
|
Thomas Wolf
|
41f35d0b3d
|
Merge pull request #1089 from dhpollack/dhp/use_pytorch_layernorm
change layernorm code to pytorch's native layer norm
|
2019-08-30 14:49:08 +02:00 |
|
Thomas Wolf
|
01ad55f8cf
|
Merge pull request #1026 from rabeehk/master
loads the tokenizer for each checkpoint, to solve the reproducability…
|
2019-08-30 14:15:36 +02:00 |
|
Thomas Wolf
|
f7978490b2
|
Merge pull request #1148 from huggingface/circleci
Documentation auto-deploy
|
2019-08-30 13:28:16 +02:00 |
|
LysandreJik
|
caf1d116a6
|
Closing bracket in DistilBERT's token count.
|
2019-08-29 15:30:10 -04:00 |
|
LysandreJik
|
e7fba4bef5
|
Documentation auto-deploy
|
2019-08-29 12:14:29 -04:00 |
|
Luis
|
fe8fb10b44
|
Small modification of comment in the run_glue.py example
Add RoBERTa to the comment as it was not explicit that RoBERTa don't use token_type_ids.
|
2019-08-29 14:43:30 +02:00 |
|
LysandreJik
|
bf3dc778b8
|
Changed learning rate for run_squad test
|
2019-08-28 18:24:43 -04:00 |
|
thomwolf
|
0a74c88ac6
|
fix #1131
|
2019-08-28 22:41:42 +02:00 |
|
Thomas Wolf
|
5f297c7be3
|
Merge pull request #1087 from huggingface/fix-warnings
Decode now calls private property instead of public method
|
2019-08-28 22:22:11 +02:00 |
|
Thomas Wolf
|
d9847678b3
|
Merge pull request #1136 from adai183/update_SQuAD_script
swap order of optimizer.step() and scheduler.step()
|
2019-08-28 22:00:52 +02:00 |
|
Thomas Wolf
|
0f8ad89206
|
Merge pull request #1135 from stefan-it/master
distilbert: fix number of hidden_size
|
2019-08-28 22:00:12 +02:00 |
|
LysandreJik
|
9ce42dc540
|
Pretrained models table fix
|
2019-08-28 13:56:28 -04:00 |
|
Andreas Daiminger
|
1d15a7f278
|
swap order of optimizer.step() and scheduler.step()
|
2019-08-28 19:18:27 +02:00 |
|
Stefan Schweter
|
ed2ab1c220
|
distilbert: fix number of hidden_size
|
2019-08-28 18:08:16 +02:00 |
|
Thomas Wolf
|
0ecfd17f49
|
Merge pull request #987 from huggingface/generative-finetuning
Generative finetuning
|
2019-08-28 16:51:50 +02:00 |
|
Thomas Wolf
|
50792dbdcc
|
Merge pull request #1127 from huggingface/dilbert
DilBERT
|
2019-08-28 16:43:09 +02:00 |
|
thomwolf
|
e7706f514b
|
update again
|
2019-08-28 16:37:22 +02:00 |
|
thomwolf
|
b5eb283aaa
|
update credits
|
2019-08-28 16:36:55 +02:00 |
|
LysandreJik
|
f753d4e32b
|
Removed typings for Python 2
|
2019-08-28 10:15:02 -04:00 |
|