Skip to main content
This page lists all available models for Runpod Public Endpoints, as well as the model-specific parameters you can use in your API calls. You can browse and test Public Endpoints using the Runpod console.
Output URLs (image_url, video_url, and audio_url) expire after 7 days. Download and store your generated files immediately if you need to keep them longer.

Available models

The following models are currently available:

Model-specific parameters

Each Public Endpoint accepts a different set of parameters to control the generation process.

Flux Dev

Flux Dev is optimized for high-quality, detailed image generation. The model accepts several parameters to control the generation process:

Flux Schnell

Flux Schnell is optimized for speed and real-time applications:
Flux Schnell is optimized for speed and works best with lower step counts. Using higher values may not improve quality significantly.

IBM Granite-4.0-H-Small

IBM Granite-4.0-H-Small is a 32B parameter long-context instruct model.

Qwen3 32B AWQ

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. The Qwen3 endpoint is also fully compatible with vLLM and the OpenAI API, allowing you to use any of the parameters available in these frameworks. For more details, see Send vLLM requests and the OpenAI API compatibility guide. Here are some examples of how to use the Qwen3 32B AWQ model with the OpenAI API:
You can stream responses from the OpenAI API using the stream and stream_options parameters:

Qwen Image

Qwen Image is an image generation foundation model with advanced text rendering capabilities.

Qwen Image LoRA

Qwen Image with LoRA support allows you to customize generation with fine-tuned LoRA models.

Seedream 3.0

Seedream 3.0 is a native high-resolution bilingual image generation model supporting both Chinese and English prompts.

Seedream 4.0 T2I

Seedream 4.0 is a new-generation image creation model that integrates both generation and editing capabilities.

Nano Banana Edit

Google’s Nano Banana Edit is a state-of-the-art image editing model that combines multiple source images.

Qwen Image Edit

Qwen Image Edit extends the text rendering capabilities to image editing tasks, enabling precise text editing.

Seedream 4.0 Edit

Seedream 4.0 Edit provides advanced image editing capabilities with the same unified architecture as Seedream 4.0 T2I.

InfiniteTalk

InfiniteTalk is an audio-driven video generation model that creates talking or singing videos from a single image and audio input.

Kling v2.1 I2V Pro

Kling 2.1 Pro generates videos from static images with additional control parameters.

Seedance 1.0 Pro

Seedance 1.0 Pro is a high-performance video generation model with multi-shot storytelling capabilities.

SORA 2 I2V

OpenAI’s Sora 2 is a video and audio generation model.

SORA 2 Pro I2V

OpenAI’s Sora 2 Pro is a professional-grade video and audio generation model.

Whisper V3 Large

Whisper V3 Large is a state-of-the-art automatic speech recognition model that transcribes audio to text.

Minimax Speech 02 HD

Minimax Speech 02 HD is a high-definition text-to-speech model with emotional control and voice customization.

Flux Kontext Dev

A 12 billion parameter model for editing images based on text instructions.

WAN 2.5

WAN 2.5 generates videos from static images.

Wan 2.2 I2V 720p LoRA

Wan 2.2 is an open-source video generation model with LoRA support for customized camera movements and effects.

Wan 2.2 I2V 720p

An open-source image-to-video generation model that creates 720p video content from static images.

Wan 2.2 T2V 720p

Open-source model for generating 720p videos from text prompts.

Wan 2.1 I2V 720p

Open-source image-to-video generation model that converts static images into 720p videos.

Wan 2.1 T2V 720p

An open-source video generation model for creating 720p videos from text prompts.