You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
hi,
You opened weights saying these weights are same as closed source weights, and yet there are limitations for example i generated a video with this just single text prompt in API and it worked:
"follow text instructions from first frame"
and closed source model generated a video correctly:
And i must say what is this? i mean it contains a large 32b vl encoder and it still fails to understand simple tasks? i am not saying that the models are bad in fact they are the best one so far as opensource, but selling with fake cherry on top of it is not right.
hi,
You opened weights saying these weights are same as closed source weights, and yet there are limitations for example i generated a video with this just single text prompt in API and it worked:
"follow text instructions from first frame"
and closed source model generated a video correctly:
Hailuo_Video_follow.text.instructions.from._541712491691143176.mp4
now look at the output from open weights with same prompt but i even tried different and it never worked:
here the result from both open weights:
https://github.com/user-attachments/assets/20d2975a-4f30-4541-8e8e-0dd126ebaf96
https://github.com/user-attachments/assets/acec3afa-6bdb-4750-aa4a-957cd0ac60a1
one with ref2vid and other fl2vid.
And i must say what is this? i mean it contains a large 32b vl encoder and it still fails to understand simple tasks? i am not saying that the models are bad in fact they are the best one so far as opensource, but selling with fake cherry on top of it is not right.