Evaluation Results

#5
by Patelcoder - opened

I tried to compare this model with Gemini-2.5 Flash, and this is the result. I think it needs a little bit more fine-tuning.
[System Latency]
CPG (Local T4 GPU): 31.85 seconds
LLM (Cloud API): 12.15 seconds

[ROUGE-L Score] (Higher = better structural preservation)
CPG: 0.8756
LLM: 0.3632

[BERTScore F1] (Higher = better semantic meaning preservation)
CPG: 0.9658
LLM: 0.8916

Sign up or log in to comment