HuggingFace_transformer

Files

Jason Phang 0041be5b3d LLaMA Implementation (#21955 )

* LLaMA

* sharding and docs

* tweak

* black

* inits

* ruff

* LLAMA_PRETRAINED_CONFIG_ARCHIVE_MAP

* init

* no checkpoint

* docs

* ruff

* type_vocab_size

* tokenizer fixes

* tokenizer fixes

* Update tokenization_llama.py

* Update tokenization_llama.py

* Update configuration_llama.py

* Update modeling_llama.py

* tokenizer add_bos by default

* licenses

* remove decoder

* norms and mlp

* rope overhaul

* tweaks

* black

* mention OPT implementation

* off-by-one naming

* typo

* fix

* tokenization fix and slicing bug

* padding config

* cleanup

* black

* update tests

* undo typo

* fix vocab caching logic

* ruff

* docbuilder

* attn fix from BlackSamorez

* initial feedback

* typo

* docs

* llama case

* llama case

* load checkpoint docs

* comment about tokenizer

* tokenizer defaults

* clear past_key_values if use_cache=False

* last tweaks

* last tweaks

* last tweaks

* last tweaks

---------

Co-authored-by: Stella Biderman <stellabiderman@gmail.com>

2023-03-16 09:00:53 -04:00

asr.mdx

Added "Open in Colab" to task guides (#21729 )

2023-02-22 08:32:35 -05:00

audio_classification.mdx

[Whisper] Add model for audio classification (#21754 )

2023-03-07 16:20:21 +01:00

document_question_answering.mdx

Add: document question answering task guide (#21518 )

2023-02-13 09:24:56 -05:00

image_captioning.mdx

[Tasks] Adds image captioning (#21512 )