collaborate, and contribute on compute and better training data

#5
by QyrouNnet-AI - opened
SupraLabs org

Hi, I saw you all use consumer GPUs. I have an RTX 5070. I would like to help y'all make better datasets.

I like synthesizing datasets for fun.
In fact, my favorite thing is to make AI models to make better datasets.

I was looking forward to increasing the context length of the base 50m model with RoPE on 1b tokens, and finetune it on a summarization dataset which I made, and then do some DPO to reinforce it.

I would like to help you all with training data. I was underwhelmed after seeing y'all use training data from a 1.5b model to train the reasoning model.
First of all, 1b models tend to generate bad reasoning chains, and very long reasonign chain which WILL hurt that 1k ctx, and also it looks like the reasoning model hallucinated more than the instruction model. I would like to reduce those hallucinations by a LOT with my data, compute, and expertise.

So... how can I join y'all?

Well, i need LH Tech AI permission, you can talk to him and me, i can't use discord RN, but my discord is axionlab, talk to me later

QyrouNnet-AI changed discussion status to closed

Sign up or log in to comment