CoolFace
Modelpublic

RichardErkhov/juhwanlee_-_gemma-7B-alpaca-case-3-2-gguf

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes1.9kdownloads
Model Card

Quantization made by Richard Erkhov.

Github

Discord

Request more models

gemma-7B-alpaca-case-3-2 - GGUF

  • —Model creator: https://huggingface.co/juhwanlee/
  • —Original model: https://huggingface.co/juhwanlee/gemma-7B-alpaca-case-3-2/

Original model description: --- license: apache-2.0 datasets:

  • —Open-Orca/OpenOrca language:
  • —en ---

Model Details

  • —Model Description: This model is test for data ordering.
  • —Developed by: Juhwan Lee
  • —Model Type: Large Language Model

Model Architecture

This model is based on Gemma-7B. We fine-tuning this model for data ordering task.

Gemma-7B is a transformer model, with the following architecture choices:

  • —Grouped-Query Attention
  • —Sliding-Window Attention
  • —Byte-fallback BPE tokenizer

Dataset

We random sample Open-Orca dataset. (We finetune the 100,000 dataset)

Guthub

https://github.com/trailerAI

License

Apache License 2.0