transformers/docs/source
Jason Phang 0041be5b3d
LLaMA Implementation (#21955)
* LLaMA

* sharding and docs

* tweak

* black

* inits

* ruff

* LLAMA_PRETRAINED_CONFIG_ARCHIVE_MAP

* init

* no checkpoint

* docs

* ruff

* type_vocab_size

* tokenizer fixes

* tokenizer fixes

* Update tokenization_llama.py

* Update tokenization_llama.py

* Update configuration_llama.py

* Update modeling_llama.py

* tokenizer add_bos by default

* licenses

* remove decoder

* norms and mlp

* rope overhaul

* tweaks

* black

* mention OPT implementation

* off-by-one naming

* typo

* fix

* tokenization fix and slicing bug

* padding config

* cleanup

* black

* update tests

* undo typo

* fix vocab caching logic

* ruff

* docbuilder

* attn fix from BlackSamorez

* initial feedback

* typo

* docs

* llama case

* llama case

* load checkpoint docs

* comment about tokenizer

* tokenizer defaults

* clear past_key_values if use_cache=False

* last tweaks

* last tweaks

* last tweaks

* last tweaks

---------

Co-authored-by: Stella Biderman <stellabiderman@gmail.com>
2023-03-16 09:00:53 -04:00
..
de Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
en LLaMA Implementation (#21955) 2023-03-16 09:00:53 -04:00
es Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
fr Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
it Italian Translation of migration.mdx (#22183) 2023-03-16 12:00:07 +00:00
ja Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
ko Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
pt Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
zh Add ConvNeXT V2 (#21679) 2023-03-14 12:08:14 +03:00
_config.py Adding evaluate to the list of libraries required in generated notebooks (#20850) 2022-12-21 14:04:08 +01:00