Equall/perplexity_evaluation
SaulLM-7B: Pioneering the first Legal Large Language Model Perplexity Analysis This dataset presents the data used in the paper "SaulLM-7B: Pioneering the first Legal Large Language Model" in "6.3 Perplexity Analysis" section. The dataset contains the perplexity scores of SaulLM-7B, Llama2-7B and Mistral-7B across a corpora of recent text. Cleaning We proceeded to standardize the data by removing any special characters using unicodedata… See the full description on the dataset page: https://huggingface.co/datasets/Equall/perplexity_evaluation.
331
Update README.md
Update README.md
Upload train.csv
Upload train.csv
Create README.md
Upload train.csv
initial commit
