CoolFace
Modelpublic

fbaldassarri/Llama-2-70B-Chat-GGUF-tokenizer-legacy

sourceHugging Facellama2updated 3y agoView on Hugging Face
0likes
Model Card

Llama-2-70B-Chat-GGUF-tokenizer-legacy

Tokenizer for llama-2-70b-chat

This repository contains the following files: specialtokensmap.json, tokenizerconfig.json, tokenizer.json, and tokenizer.model. These files are used to load a llama.cpp model as a HuggingFace Transformers model using [llamacppHF loader](https://github.com/oobabooga/text-generation-webui/blob/main/modules/llamacpp_hf.py).

Note: converted using convert_llama_weights_to_hf.py with legacy method.

How to use with oobabooga/text-generation-webui

  1. 1.Download a .gguf file from TheBloke/Llama-2-70B-Chat-GGUF based on your preferred quantization method;
  1. 1.Place your .gguf in a subfolder of models/ along with these 4 files.