YAML Metadata Warning: The pipeline tag "conversational" is not in the official list: text-classification, token-classification, table-question-answering, question-answering, zero-shot-classification, translation, summarization, feature-extraction, text-generation, text2text-generation, fill-mask, sentence-similarity, text-to-speech, text-to-audio, automatic-speech-recognition, audio-to-audio, audio-classification, audio-text-to-text, voice-activity-detection, depth-estimation, image-classification, object-detection, image-segmentation, text-to-image, image-to-text, image-to-image, image-to-video, unconditional-image-generation, video-classification, reinforcement-learning, robotics, tabular-classification, tabular-regression, tabular-to-text, table-to-text, multiple-choice, text-retrieval, time-series-forecasting, text-to-video, image-text-to-text, visual-question-answering, document-question-answering, zero-shot-image-classification, graph-ml, mask-generation, zero-shot-object-detection, text-to-3d, image-to-3d, image-feature-extraction, video-text-to-text, keypoint-detection, any-to-any, other

Llama.cpp compatible versions of an original 70B model.

Download one of the versions, for example ggml-model-q4_1.gguf.
Download interact_llamacpp.py

How to run:

sudo apt-get install git-lfs
pip install llama-cpp-python fire

python3 interact_llamacpp.py ggml-model-q4_1.gguf

System requirements:

45GB RAM for q4_1

Downloads last month: 103

GGUF

Model size

69B params

Architecture

llama

2-bit

4-bit

Inference Examples

Text Generation

Inference API (serverless) has been turned off for this model.

Datasets used to train IlyaGusev/saiga2_70b_gguf

Spaces using IlyaGusev/saiga2_70b_gguf 5

Collection including IlyaGusev/saiga2_70b_gguf

Saiga GGUF

Collection

LLaMA-based Russian chat model in the GGUF format compatible with llama.cpp • 6 items • Updated Nov 3 • 22