CoolFace
Modelpublic

ainergy/CodeLlama-SDSAT_L7_13B

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes26downloads
Model Card

Model Card for Model ID

<!-- Provide a quick summary of what the model is/does. -->

The 13B model of "SDSAT: Accelerating LLM Inference through Speculative Decoding with Semantic Adaptive Tokens"

Model Details

Model Description

<!-- Provide a longer summary of what this model is. -->

  • —Developed by: ainergy
  • —Language(s) (NLP): Code
  • —Finetuned from model: CodeLlama-13B

Model Sources

<!-- Provide the basic links for the model. -->

  • —Repository: https://github.com/ainergy-ml/SDSAT
  • —Paper: https://arxiv.org/abs/2403.18647

Evaluation

<!-- This section describes the evaluation protocols and provides the results. -->

Results

image/png

image/png

Walltime improvement

image/png