Llama 3.1 8B Instruct โ€” Solus v1

Meta's 8B instruction-tuned Llama 3.1. This is the size most people land on: capable enough for real work across chat, writing, summarisation, and everyday coding, while still running on a mainstream laptop or desktop.

It supports tool use and eight official languages, and it is the model Solus recommends by default for machines with a modern GPU or a reasonably fast CPU.

Specifications

Parameters 8B
Quantization Q4_K_M
File size 4.58 GB
Minimum RAM 8.00 GB
Minimum VRAM 6.00 GB
Context length 32,768 tokens
SHA-256 7b064f5842bf9532c91456deda288a1b672397a54fa729aa665952863033557c

Single file: Meta-Llama-3.1-8B-Instruct-Q4_K_M.gguf

Quantization

Quantization performed at the Faculty of Engineering, McMaster University.

The GGUF conversion this build is derived from was produced by bartowski, and the weights here are a byte-for-byte copy of that file โ€” the SHA-256 above matches the upstream artifact.

Provenance

Usage

llama-cli -m Meta-Llama-3.1-8B-Instruct-Q4_K_M.gguf -cnv

License

Built with Llama.

Llama 3.1 is licensed under the Llama 3.1 Community License, Copyright (c) Meta Platforms, Inc. All Rights Reserved. Your use of this model is governed by that license and by the Llama 3.1 Acceptable Use Policy.

Downloads last month
9
GGUF
Model size
8B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Fazmin/solus_v1_llama-3.1-8b-instruct-q4

Quantized
(892)
this model