Audio (TTS/STT)
POST
/v1/runRun a text-to-speech or speech-to-text model. Asynchronous audio models may include webhook_url for a task callback.
Request body
ImageEditParamsUse gpt-image-2 on this endpoint.
Text description of the requested edit.
One or more source image files. Repeat the multipart field for multiple images.
Optional mask image for inpainting.
Number of edited images to return when supported.
Output dimensions supported by the model.
Output quality supported by the model.
Output background setting.
How closely supported models should preserve source-image details.
Output encoding such as png, webp, or jpeg.
Output compression level from 0 to 100.
Provider-compatible end-user identifier.