English to Canadian French

#7
by KotaNaveen - opened

Hi, I am working on converting the text from English to Canadian french. If I fine-tune 13B model with english to canadian french dataset, what is the possibility of effective translation?

Hi, thanks for the interest in our work!

My suggestion would be firstly trying fine-tuning on Canadian French monolingual data (not necessary to be large, maybe 500K or 1B tokens), and then fine-tune on your en->Canadian fr. I guess it would be very possible to have effective translations.

Thank you for prompt response.

Unfortunately, i have canadian french dataset, but its 60k, but i am not sure the about the diversity of those tokens after tokenizing.

Sign up or log in to comment