Generate a clear base composition
Begin with the intended subject, environment, framing, and visual hierarchy. Test the composition in the target ratio before spending iterations on surface detail.
GPT Image 2 brings text-to-image generation and reference-based image editing into one focused workspace. Start from a written brief or upload as many as 16 source images, choose a composition from square, portrait, landscape, panoramic, and extra-wide ratios, then render at 1K, 2K, or 4K. The model page keeps both workflows together so you can move from an initial concept to a directed revision without changing tools.
Use only the inputs that support the result you need: a clear prompt, optional reference images, a target aspect ratio, and an output resolution.
Describe the subject, environment, composition, lighting, materials, and intended use. GPT Image 2 turns the brief into a new image while the workspace keeps the selected format and resolution visible.
Upload one or more source images when identity, products, layout, or visual direction must come from existing material. Assign a clear purpose to each reference instead of asking every image to control the same detail.
Choose auto sizing or common square, portrait, landscape, wide, and panoramic ratios. Set the final channel first so framing decisions match a product card, poster, thumbnail, banner, or social placement.
Select resolution according to the review stage and delivery need. Use lower resolution for early composition tests, then move to a larger output after subject, layout, and typography directions are stable.
A useful request separates what must remain fixed from what may change. Define the deliverable, organize references, and make the prompt specific enough to review.
Choose the placement, aspect ratio, resolution, audience, and visual goal before writing style details. A banner, product listing, and editorial illustration need different framing and information density.
State the main subject and action first, then scene, camera or viewpoint, lighting, palette, materials, and constraints. Put required text or product details in exact wording and avoid contradictory style instructions.
Identify which upload controls identity, product shape, pose, composition, palette, or texture. Remove redundant or conflicting images and explain which source has priority when details disagree.
Check subject accuracy, hands and edges, text, brand details, spatial relationships, and crop safety. Change one variable per revision so the effect of each edit remains clear.
Begin with the intended subject, environment, framing, and visual hierarchy. Test the composition in the target ratio before spending iterations on surface detail.
Switch to the image-editing mode when the result must preserve a person, object, product, or layout. Upload only relevant references and describe both the required change and the details that must stay untouched.
Inspect the final crop, resolution, small details, embedded text, and consistency with the original brief. Export the version that matches the actual publishing surface rather than a generic master ratio.
Prepare files that the current Kuca workspace can accept, and decide which details each input should control before you submit the request. These practical limits help prevent rejected uploads and make a multi-reference edit easier to direct.
The editing mode accepts a maximum of 16 reference images. Reaching the limit is not a creative goal: use only sources that add necessary information. A concise set for identity, product shape, composition, and styling is usually easier to explain than many nearly identical uploads. Text-to-image mode does not require a reference image.
Each uploaded image can be no larger than 10 MB. Check the actual file size before starting a long brief, especially when working with camera originals or exported design files. If a file is too large, create a clean web-ready copy that preserves the important subject, edges, texture, and readable details without unnecessary metadata.
All three output levels are available in the workspace. Use 1K for early visual direction and quick comparisons, then select 2K or 4K when the layout is stable and the destination needs closer inspection. A higher setting provides a larger output; it does not repair an unclear prompt, conflicting references, or an incorrect composition.
Choose images in which the required person, object, product, or layout is visible enough to guide the edit. Avoid references with incompatible identities, lighting, camera angles, or brand details unless the prompt explicitly resolves the conflict. Do not upload material you lack permission to use, and review the generated result before publishing it.
The model page offers a text workflow and a reference-based editing workflow. Select the mode from the kind of control your task requires, rather than adding uploads to every request by default.
Start with text-to-image when no existing image must be preserved. It works well for exploring subjects, art direction, environments, product-scene ideas, and composition alternatives. Describe the intended output and visual hierarchy first, then add lighting, palette, materials, camera language, and exclusions that meaningfully affect the result.
Choose reference-based editing when a person, object, package, room, pose, or layout should remain recognizable. State the requested change separately from the protected details. For example, identify which background may change while product geometry, label placement, color, and camera position must stay consistent.
If you are unsure which mode fits, keep the aspect ratio, output goal, and core brief consistent while testing a text-only version and a reference-guided version. Compare instruction adherence, preserved details, composition, and revision effort instead of judging only surface style. This produces a more useful model decision for repeat work.
The generator shows the current credit requirement for the selected mode and settings before submission. Treat that live value as the source of truth because model options and costs can change. Review the displayed credits after changing resolution, variant, or other controls, and submit only when the configuration matches your intended test or final render.
Explore product scenes, packaging directions, materials, and controlled backgrounds while keeping required proportions and brand details explicit.
Develop key art and channel-specific crops from one visual brief, then adapt the selected direction to banners, social placements, and promotional layouts.
Change a background, styling, lighting, pose, or composition while telling the model which source details must remain recognizable.
Turn abstract topics, articles, and reports into directed illustrations with a defined point of view, visual metaphor, and publication format.
Create consistent visual checkpoints for characters, locations, props, and key moments before committing to a larger production sequence.
Move an approved composition to 2K or 4K for closer review and delivery when the chosen placement needs more source resolution.
GPT Image 2 supports text-to-image generation and image-to-image editing. You can start with a prompt or upload reference images, select an aspect ratio, and choose 1K, 2K, or 4K output.
The GPT Image 2 editing workflow accepts up to 16 input images. Use fewer, purposeful references when possible, and describe the role of each source so identity, layout, and style instructions do not compete.
The workspace offers auto sizing plus square, portrait, landscape, wide, panoramic, and extra-wide formats, including 1:1, 3:2, 2:3, 16:9, 9:16, 2:1, 1:2, 3:1, 1:3, 21:9, and 9:21.
Yes. Choose the editing variant, upload the source images, describe the requested change, and state what must remain unchanged. This is useful for controlled revisions rather than recreating the entire scene from scratch.
Not always. Early composition tests are easier to compare at a smaller resolution. Move to 2K or 4K after the subject, crop, layout, and visual direction are approved.
Lead with the subject and intended result, then specify environment, composition, viewpoint, lighting, palette, materials, required text, and constraints. For editing, also list what must change and what must be preserved.
Start from a prompt or reference images, then refine the format and resolution in the Kuca image workspace.