Aditya Soni
c356290c8d
typo fix as per Pytorch v1.1+
2019-12-01 14:08:14 +05:30
Rostislav Nedelchev
76c0bc06d5
[XLNet] Changed post-processing of attention w.r.t to target_mapping
...
Whenever target_mapping is provided to the input, XLNet outputs two different attention streams.
Based on that the attention output would be on of the two:
- a list of tensors (usual case for most transformers)
- a list of 2-tuples of tensors, one tesor for each of attention streams
Docs and unit-tests have been updated
2019-11-30 21:01:04 +01:00
Rostislav Nedelchev
b90791e950
fixed XLNet attenttion output for both attention streams
2019-11-30 15:57:51 +01:00
maxvidal
b0ee7c7df3
Added Camembert to available models
2019-11-29 14:17:02 -05:00
Elad Segal
ecf15ebf3b
Add ALBERT to AutoClasses
2019-11-29 11:25:37 -05:00
thomwolf
4a666885b5
reducing my level of enthousiasm
2019-11-29 09:40:50 -05:00
thomwolf
adb5c79ff2
update all tf.shape and tensor.shape to shape_list
2019-11-29 09:40:50 -05:00
Juha Kiili
2421e54f8c
Add link to original source and license to download_glue.data.py
2019-11-29 15:39:28 +02:00
Juha Kiili
41aa0e8003
Refactor logs and fix loss bug
2019-11-29 15:33:25 +02:00
Thomas Wolf
1ab8dc44b3
Merge pull request #1876 from huggingface/mean-fix
...
Mean does not exist in TF2
2019-11-29 09:26:33 +01:00
Thomas Wolf
f0d22b6363
Merge pull request #1873 from stefan-it/distilbert-german
...
German DistilBERT
2019-11-29 09:25:47 +01:00
Lysandre
1e9ac5a7cf
New -> normal
2019-11-28 17:43:47 -05:00
Lysandre
0b84b9fd8a
Add processors to __init__
2019-11-28 17:38:52 -05:00
Lysandre
f671997ef7
Interface with TFDS
2019-11-28 17:17:20 -05:00
Lysandre
bd41e8292a
Cleanup & Evaluation now works
2019-11-28 16:03:56 -05:00
Thomas Wolf
d49c43ff78
Merge pull request #1778 from eukaryote31/patch-2
...
from_pretrained: convert DialoGPT format
2019-11-28 16:08:37 +01:00
Thomas Wolf
91caf2462c
Merge pull request #1770 from huggingface/initi-encoder-mask
...
Only init encoder_attention_mask if stack is decoder
2019-11-28 16:06:55 +01:00
Thomas Wolf
49a69d5b78
Merge pull request #1753 from digantamisra98/patch-1
...
Added Mish Activation Function
2019-11-28 15:24:08 +01:00
Thomas Wolf
96e7ee7238
Merge pull request #1740 from huggingface/fix-ctrl-past
...
Fix CTRL past
2019-11-27 23:28:30 +01:00
thomwolf
8da47b078d
fix merge tests
2019-11-27 23:11:37 +01:00
Stefan Schweter
8c276b9c92
Merge branch 'master' into distilbert-german
2019-11-27 18:11:49 +01:00
Yao Lu
3c28a2daac
add add_special_tokens=True for input examples
2019-11-27 12:05:23 -05:00
Thomas Wolf
a36f981d1b
Merge branch 'master' into fix-ctrl-past
2019-11-27 17:25:46 +01:00
Thomas Wolf
5afca00b47
Merge pull request #1724 from huggingface/fix_encode_plus
...
Fix encode_plus
2019-11-27 17:14:49 +01:00
Thomas Wolf
49108288ba
Merge pull request #1624 from Huawei-MRC-OSI/resumable_http
...
Add support for resumable downloads for HTTP protocol.
2019-11-27 17:11:07 +01:00
Thomas Wolf
5340d1f21f
Merge branch 'master' into resumable_http
2019-11-27 17:10:36 +01:00
VictorSanh
10bd1ddb39
soft launch distilbert multilingual
2019-11-27 11:07:22 -05:00
VictorSanh
d5478b939d
add distilbert + update run_xnli wrt run_glue
2019-11-27 11:07:22 -05:00
VictorSanh
07ab8d7af6
fix bug
2019-11-27 11:07:22 -05:00
VictorSanh
d474022639
cleaning simple_accuracy since not used anymore
2019-11-27 11:07:22 -05:00
VictorSanh
bcd8dc6b48
move xnli_compute_metrics to data/metrics
2019-11-27 11:07:22 -05:00
VictorSanh
73fe2e7385
remove fstrings
2019-11-27 11:07:22 -05:00
VictorSanh
3e7656f7ac
update readme
2019-11-27 11:07:22 -05:00
VictorSanh
abd397e954
uniformize w/ the cache_dir update
2019-11-27 11:07:22 -05:00
VictorSanh
d75d49a51d
add XnliProcessor to doc
2019-11-27 11:07:22 -05:00
VictorSanh
d5910b312f
move xnli processor (and utils) to transformers/data/processors
2019-11-27 11:07:22 -05:00
VictorSanh
289cf4d2b7
change default for XNLI: dev --> test
2019-11-27 11:07:22 -05:00
VictorSanh
cb7b77a8a2
fix some typos
2019-11-27 11:07:22 -05:00
VictorSanh
84a0b522cf
mbert reproducibility results
2019-11-27 11:07:22 -05:00
VictorSanh
c4336ecbbd
xnli - output_mode consistency
2019-11-27 11:07:22 -05:00
VictorSanh
d52e98ff9a
add xnli examples/README.md
2019-11-27 11:07:22 -05:00
VictorSanh
71f71ddb3e
run_xnli + utils_xnli
2019-11-27 11:07:22 -05:00
Julien Chaumond
b5d884d25c
Uniformize #1952
2019-11-27 11:05:55 -05:00
Thomas Wolf
7fd1d42a01
Merge pull request #1592 from watkinsm/do_lower_case
...
Consider do_lower_case in PreTrainedTokenizer
2019-11-27 17:05:18 +01:00
Thomas Wolf
21637d4924
Merge branch 'master' into do_lower_case
2019-11-27 17:04:39 +01:00
Rémi Louf
de2696f68e
suggest to track repo w/ https rather than ssh
2019-11-27 11:02:28 -05:00
root
88b317739f
Fix issue: #1962 , input's shape seem to cause error in 2.2.0 version tf_albert_model
2019-11-27 10:38:10 -05:00
Lysandre
45d767297a
Updated v2.2.0 doc
2019-11-27 10:12:20 -05:00
Lysandre
361620954a
Remove TFBertForPreTraining from ALBERT doc
2019-11-27 10:11:37 -05:00
Lysandre
cc7968227e
Updated v2.2.0 doc
2019-11-26 15:52:25 -05:00