To use with LocalAI this is an example model definition:

name: WizardLM2
backend: transformers
parameters:
  model: fakezeta/Not-WizardLM-2-7B-ov-int8
context_size: 8192
threads: 6
f16: true
type: OVModelForCausalLM
prompt_cache_path: "cache"
prompt_cache_all: true
template:
  use_tokenizer_template: true
Downloads last month
11
Inference Examples
This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Collection including fakezeta/Not-WizardLM-2-7B-ov-int8