logoalt Hacker News

_diyartoday at 2:42 PM2 repliesview on HN

I agree. But is that the model or the prompt?

I remember reading a report where people running AI-model instagram account were using insanely long and detailed prompts about the setting, lighting, makeup, pose, disposition, clothing, etc. about their models. Presumably with some reference image of the face / body to remain consistent across images.

It‘s not clear to me whether a sufficiently detailed prompt can generate actually interesting video with a natural ”texture” (for lack of a better word).


Replies

razstertoday at 2:54 PM

That would be the prompt. With the right assistance from Qwen3.5/Ornith I was able to achieve some amazing results. Unfortunately due to their licensing I'm not allowed to use it in the USA, so I had to halt testing.

show 1 reply
fwiptoday at 3:12 PM

I mean, even the demo prompts on the ComfyUI page aren't adhered to by the model. From the first prompt, one of the four lines:

> TRANSITION: a violent WHIP PAN off the rooftop that SMEARS the floating words away with it, motion-streaked —

And the video just didn't do any of that transition at all, it just replaced it with a cut. If you look at the rest of the prompts, you'll find similar lines that are just totally ignored. Except maybe the mouse one, I didn't see anything wrong with that off the bat.