gemma_2b_curated
Experiment gemma_2b_curated from the Teacher-Free Read/Write Annotation for Simultaneous
Machine Translation project.
Recipe
- Backbone:
google/gemma-4-E2B-it - Corpus:
curated - Annotator:
same_as_backbone - Criterion:
ot(Ï„ = 0.3) - Latencies: ['low', 'low-medium', 'medium', 'medium-high', 'high']
Files in this repo
config.yaml— the exact experiment config that produced this run.manifest.json— git sha, hostname, GPUs, timestamps.logs/— per-stage stdout+stderr frombin/run.eval/— every landed eval-JSON cell (hypothesis, reference, AL, BLEU).- SFT checkpoint (
*.safetensors+ tokenizer).
Reproduce
git clone https://github.com/dipankarsrirag/simt-tor-26.git
cd simt-tor-26
cp .simtrc.example .simtrc # edit paths for your setup
bin/run configs/00_gemma_2b_curated.yaml --ngpus N
Git commit at time of run: unknown
- Downloads last month
- 19
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support