Describe with Ideogram 4.0
Describe an image as a structured V4JsonPrompt with Ideogram’s 4.0
image captioner (a fine-tune of the Qwen3-VL vision-language model).
Upload the image as image using multipart/form-data; the
json_prompt is returned directly and can be passed to the
/v1/ideogram-v4/generate endpoints.
Authentication
API key for access control. Use in the header with the name "Api-Key"
Request
A request to describe one image. Send it as multipart/form-data with the image to describe.
The image to describe (max 10MB). JPEG, PNG, and WebP are supported. Multipart requests only.
Whether to include bounding boxes for the subjects and text in the returned json_prompt. Defaults to true, so the prompt preserves the layout of the described image.
Whether to include a free-form style description in the returned json_prompt. Defaults to false.
Response
URL-safe base64 ID of the description that was created.
Structured prompt for Ideogram 4.0 generation. When json_prompt is
supplied, magic-prompt is disabled and the diffusion model consumes
the JSON contract directly. Mutually exclusive with text_prompt
and the legacy prompt field.