YAML Metadata Warning:empty or missing yaml metadata in repo card

Check out the documentation for more information.

tocsa/meridian-sft

Meridian Response (class exercise, Sections 1-4) โ€” SFT model.

Meridian is a fictional humanitarian logistics coordinator. During a flood response, field reports arrive from aid workers, camp coordinators, health volunteers, and protection officers. The model classifies and routes each report so a duty officer can verify instead of reading every report from scratch.

Base model / architecture

nvidia/Mistral-NeMo-Minitron-8B-Instruct (MistralForCausalLM)

Pipeline position

SFT model: a LoRA adapter trained on synthetically generated data, merged into the base model. This is the checkpoint used as the DPO starting point.

Files

  • config.json (702.0B)
  • generation_config.json (111.0B)
  • model.safetensors (15.7GB)
  • special_tokens_map.json (414.0B)
  • tokenizer.json (8.8MB)
  • tokenizer_config.json (173.4KB)

Recorded metrics

No recorded scalar metrics found alongside this artifact.

How to load

from transformers import AutoModelForCausalLM, AutoTokenizer

model = AutoModelForCausalLM.from_pretrained("tocsa/meridian-sft")
tokenizer = AutoTokenizer.from_pretrained("tocsa/meridian-sft")

Course context

Produced in the NVIDIA instructor-led workshop "Adding New Knowledge to LLMs" โ€” Designing and building model-customization pipelines for domain workflows.

The workshop's build-time customization workflow:

Stage What it does Tool
Task definition taxonomy, schema, success criteria
Eval measure target behavior vLLM
Gap evidence baseline + ICL ceiling
Synthetic data targeted examples NeMo Data Designer
Curation quality + dedup NeMo Curator
SFT LoRA task behavior NeMo AutoModel + PEFT
Serve + eval compare model stages vLLM
DPO preference behavior NeMo RL
Model artifact adapter or DPO model

Open-model ecosystem used across the workshop:

Models Data Training Serving
Nemotron NeMo Data Designer NeMo AutoModel vLLM
Open weights NeMo Curator LoRA / PEFT NGC containers
Teacher models Hugging Face NeMo RL Local eval

Training data for this run was synthetically generated (NeMo Data Designer) and/or curated during the workshop exercises. Educational artifact โ€” not reviewed for production use.

Downloads last month
22
Safetensors
Model size
8B params
Tensor type
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for tocsa/meridian-sft

Quantizations
1 model