CoolFace
Modelpublic

Berkelium-ai/BerkeliumGPT2-Coder-3b-MLX

sourceHugging Faceapache-2.0updated 3d agoView on Hugging Face
0likes366downloads
Model Card

BerkeliumGPT2-Coder-3b (MLX & Transformers, bf16)

This is the official full-precision bfloat16 (bf16) model fine-tuned from [Qwen/Qwen2.5-3B](https://huggingface.co/Qwen/Qwen2.5-3B) on the [BerkeliumCoding](https://huggingface.co/datasets/Berkelium-ai/BerkeliumCoding) dataset across 176 permissive repositories with a 4,096 token context window.

Native SafeTensors format ready to run directly in both Apple MLX and Hugging Face Transformers / vLLM.

Model Details

  • —Architecture: Qwen 2.5 (3 Billion Parameters)
  • —Precision: bfloat16 (bf16 full precision, non-quantized)
  • —Context Length: 4,096 tokens
  • —Training Acceleration: Apple Silicon Unified Memory (Metal) via MLX
  • —Format: Native SafeTensors

Run with MLX (Apple Silicon)

1. Installation

bash
pip install mlx-lm

2. Run from CLI

bash
mlx_lm.generate \
  --model Berkelium-ai/BerkeliumGPT2-Coder-3b \
  --prompt "def merge_intervals(intervals: list[list[int]]) -> list[list[int]]:" \
  --max-tokens 256 \
  --temp 0.2

3. Python API

python
from mlx_lm import load, generate

model, tokenizer = load("Berkelium-ai/BerkeliumGPT2-Coder-3b")
response = generate(
    model,
    tokenizer,
    prompt="def merge_intervals(intervals: list[list[int]]) -> list[list[int]]:",
    max_tokens=256,
    verbose=True
)
print(response)

Run with Hugging Face Transformers

python
import torch
from transformers import AutoModelForCausalLM, AutoTokenizer

model_id = "Berkelium-ai/BerkeliumGPT2-Coder-3b"

tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
    model_id,
    torch_dtype=torch.bfloat16,
    device_map="auto"
)

prompt = "def merge_intervals(intervals: list[list[int]]) -> list[list[int]]:"
inputs = tokenizer(prompt, return_tensors="pt").to(model.device)

outputs = model.generate(**inputs, max_new_tokens=256, temperature=0.2)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))