How to use from the
Use from the
Transformers library
# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("text-generation", model="GreenBitAI/Llama-2-7B-layer-mix-bpw-2.2")
# pip install -U transformers accelerate
# Load model directly
from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("GreenBitAI/Llama-2-7B-layer-mix-bpw-2.2")
model = AutoModelForCausalLM.from_pretrained("GreenBitAI/Llama-2-7B-layer-mix-bpw-2.2", device_map="auto")
Quick Links

GreenBit LLMs

This is GreenBitAI's pretrained low-bit LLMs with extreme compression yet still strong performance.

Please refer to our Github page for the code to run the model and more information.

Downloads last month
12
Safetensors
Model size
0.9B params
Tensor type
I32
路
F16
路
I16
路
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support