Llama 3.1 8B Instruct - GGUF

GGUF quantized version of Meta's Llama 3.1 8B Instruct for use with llama.cpp, Ollama, LM Studio, and other GGUF-compatible inference engines.

Available Quantizations

Filename Quant Size Description
Llama-3.1-8B-Instruct-Q6_K.gguf Q6_K ~6.8 GB High quality, recommended for most use cases

Usage

Ollama

ollama run hf.co/beyond-logic-labs/Llama-3.1-8B-Instruct-GGUF:Q6_K

LM Studio

Download the GGUF file and load it directly in LM Studio.

llama.cpp

./llama-cli -m Llama-3.1-8B-Instruct-Q6_K.gguf -p "You are a helpful assistant." -cnv

Model Details

  • Base Model: meta-llama/Llama-3.1-8B-Instruct
  • Context Length: 128K tokens
  • License: Llama 3.1 Community License

About Beyond Logic Labs

We build AI-powered tools for tabletop RPG players and game masters. Check out our other models for specialized RPG assistance.

Downloads last month
5
GGUF
Model size
8B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

6-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for beyond-logic-labs/Llama-3.1-8B-Instruct-GGUF

Quantized
(907)
this model