I have been using video generation models for a long time, especially Seedance 2.5.
I find the understanding of prompt is still very weak even in the strongest model. One most ridiculously simple mistake is Seedance misspelled the words in the video. I think I can fix it with emphasizing it. But it failed again.
LLM may hit a wall now but apparently video generation model has not.
[link] [comments]