The purpose of this model is to allow for editing with multiple input images but only one output image. It works similar to Qwen Image Edit but for Kontext. It uses the keywords "image1" and "image2" to reference the inputs. Here are some examples from the training data to help you understand prompting: Example 1: Inputs: Prompt: Shift the man from image1 into the stance from image2: stand tall,…