4 ms·
Thank you. I've noticed that too, and also that it has a tendency to introduce garbled text when not given a prompt (or a short one). This is using the default
by mishu2 9mo ago
Thank you. I've noticed that too, and also that it has a tendency to introduce garbled text when not given a prompt (or a short one).
This is using the default parameters for the ComfyUI workflow (including a negative prompt written in Chinese), so there is a lot of room for adjustments.
- fouc 9mo agoOh I was wondering why some of the hallucinations introduced Chinese text/visuals, I'm guessing that might be due to the negative prompt.
- mishu2 9mo agoI think the main reason is that the model has a lot of training material with Chinese text in it (I'm assuming, since the research group who released it is from China), but having the negative prompt in Chinese might also play a role. What I've found interesting so far is that sometimes the image plays a big part in the final video, but other times it gets discarded almost immediately after the first few frames. It really depends on the prompt, so prompt engineering is (at least for this model) even more important than I expected. I'm now thinking of adding a 'system' positive prompt and appending the user prompt to it.
- fouc 9mo agoWould be interesting to see how much a good "system"/server-side prompt could improve things. I noticed some animations kept the same sketch style even without specifying that in the prompt.