mirror of https://github.com/huggingface/transformers.git synced 2025-07-03 12:50:06 +06:00

History

Alex Hedges 95091e1582 Set `cache_dir` for `evaluate.load()` in example scripts (#28422 ) While using `run_clm.py`,[^1] I noticed that some files were being added to my global cache, not the local cache. I set the `cache_dir` parameter for the one call to `evaluate.load()`, which partially solved the problem. I figured that while I was fixing the one script upstream, I might as well fix the problem in all other example scripts that I could. There are still some files being added to my global cache, but this appears to be a bug in `evaluate` itself. This commit at least moves some of the files into the local cache, which is better than before. To create this PR, I made the following regex-based transformation: `evaluate\.load$(.*?)$` -> `evaluate\.load$$1, cache_dir=model_args.cache_dir$`. After using that, I manually fixed all modified files with `ruff` serving as useful guidance. During the process, I removed one existing usage of the `cache_dir` parameter in a script that did not have a corresponding `--cache-dir` argument declared. [^1]: I specifically used `pytorch/language-modeling/run_clm.py` from v4.34.1 of the library. For the original code, see the following URL: `acc394c4f5/examples/pytorch/language-modeling/run_clm.py`.		2024-01-11 15:38:44 +01:00
..
benchmarking	Apply ruff flake8-comprehensions (#21694 )	2023-02-22 09:14:54 +01:00
contrastive-image-text	Dev version	2023-12-13 18:29:31 +01:00
image-classification	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
language-modeling	Broken links fixed related to datasets docs (#27569 )	2023-11-17 13:44:09 -08:00
language-modeling-tpu	Add an option to reduce compile() console spam (#23938 )	2023-06-02 15:28:52 +01:00
multiple-choice	Dev version	2023-12-13 18:29:31 +01:00
question-answering	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
summarization	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
text-classification	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
token-classification	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
translation	Set `cache_dir` for `evaluate.load()` in example scripts (#28422 )	2024-01-11 15:38:44 +01:00
_tests_requirements.txt	Update the TF pin for 2.15 (#27375 )	2023-11-16 13:47:43 +00:00
README.md	Update README.md (#26198 )	2023-09-19 00:02:50 +02:00
test_tensorflow_examples.py	Proposed fix for TF example now running on safetensors. (#23208 )	2023-05-09 13:04:27 -04:00

README.md

Examples

This folder contains actively maintained examples of the use of 🤗 Transformers organized into different ML tasks. All examples in this folder are TensorFlow examples and are written using native Keras rather than classes like TFTrainer, which we now consider deprecated. If you've previously only used 🤗 Transformers via TFTrainer, we highly recommend taking a look at the new style - we think it's a big improvement!

In addition, all scripts here now support the 🤗 Datasets library - you can grab entire datasets just by changing one command-line argument!

A note on code folding

Most of these examples have been formatted with #region blocks. In IDEs such as PyCharm and VSCode, these blocks mark named regions of code that can be folded for easier viewing. If you find any of these scripts overwhelming or difficult to follow, we highly recommend beginning with all regions folded and then examining regions one at a time!

The Big Table of Tasks

Here is the list of all our examples:

Task	Example datasets
`language-modeling`	WikiText-2
`multiple-choice`	SWAG
`question-answering`	SQuAD
`summarization`	XSum
`text-classification`	GLUE
`token-classification`	CoNLL NER
`translation`	WMT

Coming soon

Colab notebooks to easily run through these scripts!