Hey there!
Nice work and thanks for sharing everything :)
Could you please provide additional information how did you train the LoRA and specifically how to construct the prompt for the model given the grid of 4x4 images? I used a prompt template from your paper (attached below), but the results are not as good as yours.
Thanks in advance!
The prompt I used:
This is a four-panel image grid: [TOP-LEFT]: The original source image. [TOP-RIGHT]: An edited version of the [TOP-LEFT] image, transformed to depict a running motion. [BOTTOM-LEFT]: A second original source image. [BOTTOM-RIGHT]: An edited version of the [BOTTOM-LEFT] image, applying the same running motion transformation as used in [TOP-RIGHT].
Hey there!
Nice work and thanks for sharing everything :)
Could you please provide additional information how did you train the LoRA and specifically how to construct the prompt for the model given the grid of 4x4 images? I used a prompt template from your paper (attached below), but the results are not as good as yours.
Thanks in advance!
The prompt I used: