Trained on 256px but inference on 512px?

Author: andreemicCreated May 24, 2023Updated Jan 13, 2026

As far as I understand the training config, IP2P was trained on 256x256 images but images in the paper are generated at 512px.

  1. Doesn't training at a different resolution than inference impact performance?
  2. If so, why did you train on 256px? Did you try training on 512px?

Source: timothybrooks/instruct-pix2pix