Model repeating information and "spitting out" random characters

#12

by brazilianslib - opened Jun 28, 2024

Jun 28, 2024

First of all, congratulations on the launch. Gemma 2 9B is, at least in my tests, the best model for PT-BR. Much better than much larger models.
However, problems are constantly happening, such as:

Repeat information;
"Spit" text infinitely;
Place tags like "</start_of" at the end of your answer.

I am eagerly awaiting a solution.

Once again, I thank the entire Google Gemma team.

ArthurZ

Google org Jun 29, 2024

Would recommend you to use eager attention implementation

yixuantt

Jul 6, 2024

The same error even with eager attention and bf16.

Renu11

Google org Aug 5, 2024

Hi @brazilianslib , Could you please try again by updating the latest transformer version (!pip install -U transformers) and let us know if the issue still persists? Thank you.

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment