# APIPod Docs ## Docs - [Getting Started with APIPod API](https://docs.apipod.ai/2065583m0.md): ## API Docs - Image Models > GPT Image 2 [GPT Image 2 Text to Image](https://docs.apipod.ai/gpt-image-2/gpt-image-2.md): **GPT Image 2** is a new-generation image model developed by OpenAI, featuring enhanced realism, more refined image editing capabilities, sharper text rendering, and superior product rendering performance. Built specifically for advanced visual workflows, it not only breaks through the limitations of basic text-to-image generation but also fully meets the demands for high-quality creative, commercial, and professional design-level applications. - Image Models > GPT Image 2 [GPT Image 2 Image to Image](https://docs.apipod.ai/gpt-image-2/gpt-image-2-edit.md): **GPT Image 2** is a new-generation image model developed by OpenAI, featuring enhanced realism, more refined image editing capabilities, sharper text rendering, and superior product rendering performance. Built specifically for advanced visual workflows, it not only breaks through the limitations of basic text-to-image generation but also fully meets the demands for high-quality creative, commercial, and professional design-level applications. - Image Models > Nano Banana [Nano Banana 2 Free API Docs - APIPod](https://docs.apipod.ai/nano-banana/nano-banana-2.md): Nano Banana 2 model is Google's latest image generation model offers advanced world knowledge, production-ready specs, subject consistency and more, all at Flash speed. - Image Models > Nano Banana [Nano Banana Pro Free API Docs - APIPod](https://docs.apipod.ai/nano-banana/nano-banana-pro.md): Built upon the robust architecture of Gemini 3 Pro, Nano Banana Pro (Gemini 3 Pro Image Preview) stands as Google’s flagship solution for image generation and editing. - Image Models > Seedream 4.5 [Seedream V4.5 Text to Image Free API - APIPod Docs](https://docs.apipod.ai/seedream/4-5-text-to-image.md): **Seedream V4.5 Text to Image** model is developed by the ByteDance Seed team. It achieves all-round improvement through overall model scaling and supports professional-level text rendering, diverse artistic styles, and accurate semantic understanding. - Image Models > Seedream 4.5 [Seedream V4.5 Image to Image Free API - APIPod Docs](https://docs.apipod.ai/seedream/4-5-image-to-image.md): The **Seedream 4.5 image editing** model has achieved an all-round improvement based on the overall scaling of the model: it can accurately identify and stably lock the main subject in multi-image combinations, and maximize the retention of the original image's features and detailed textures. - Image Models > Seedream 5.0 Lite [Seedream 5.0 Lite Text to Image Free API - APIPod Docs](https://docs.apipod.ai/seedream/5-0-lite-text-to-image.md): **Seedream 5.0 Lite** is a unified multimodal image generation model endowed with deep thinking and online search capabilities, featuring an all-round upgrade in its understanding, reasoning and generation capabilities. - Image Models > Seedream 5.0 Lite [Seedream 5.0 Lite Image to Image Free API - APIPod Docs](https://docs.apipod.ai/seedream/5-0-lite-image-to-image.md): Seedream 5.0 Lite image to image model. - Image Models > WAN 2.7 [ WAN 2.7 Text to Image Free API - APIPod Docs](https://docs.apipod.ai/wan/2-7-text-to-image.md): The WAN 2.7 text-to-image model generates high-quality images from text prompts through its built-in thinking mode, enabling smarter composition and precise alignment with prompt requirements. - Image Models > WAN 2.7 [ WAN 2.7 Image to Image Free API - APIPod Docs](https://docs.apipod.ai/wan/2-7-image-to-image.md): The WAN 2.7 Image Edit model can edit and optimize images through precise and controllable adjustments while maintaining the consistency of their structure and main subjects. - Image Models > WAN 2.7 [WAN 2.7 Text to Image Pro Free API - APIPod Docs](https://docs.apipod.ai/wan/2-7-text-to-image-pro.md): The professional version of the WAN 2.7 Text to Image model supports output with a maximum resolution of 4K (4096×4096), which can meet the needs of print-grade and large-format material production. - Image Models > WAN 2.7 [WAN 2.7 Image to Image Pro Free API - APIPod Docs](https://docs.apipod.ai/wan/2-7-image-to-image-pro.md): The WAN 2.7 Image Edit model can edit and optimize images through precise and controllable adjustments while maintaining the consistency of their structure and main subjects. - Image Models [Query Image Task API - APIPod Docs](https://docs.apipod.ai/query-image-task.md): Query the create image task status and results. This is a unified query interface that works with all images models. - Video Models > Veo 3.1 [Veo 3.1 Fast Free API - APIPod Docs](https://docs.apipod.ai/veo/3-1-fast.md): **Veo 3.1 Fast** for Generation transforms creative ideas into compelling video narratives using Google's advanced video generation model. Veo is capable of generating videos with audio from text prompts, or animating images with textual guidance. - Video Models > Veo 3.1 [Veo 3.1 Fast 4K](https://docs.apipod.ai/veo/3-1-fast-4k.md): **Veo 3.1 Fast** for Generation transforms creative ideas into compelling video narratives using Google's advanced video generation model. Veo is capable of generating videos with audio from text prompts, or animating images with textual guidance. This model version is generated 4K resolution video. - Video Models > Veo 3.1 [Veo 3.1 Fast Reference](https://docs.apipod.ai/veo/3-1-fast-ref.md): **Veo 3.1 Fast Reference** you can provide different reference images (up to 3) to shape character design, lighting style, or color tone, ensuring that the generated video maintains visual consistency in each shot. - Video Models > Veo 3.1 [Veo 3.1 Quality](https://docs.apipod.ai/veo/3-1-quality.md): **Veo 3.1 Quality** for Generation transforms creative ideas into compelling video narratives using Google's advanced video generation model. Create high-quality, 8-second videos with sound using Gooles's state-of-the-art video generation model. - Video Models > Veo 3.1 [Veo 3.1 Quality 4K](https://docs.apipod.ai/veo/3-1-quality-4k.md): **Veo 3.1 Quality** for Generation transforms creative ideas into compelling video narratives using Google's advanced video generation model. Veo is capable of generating videos with audio from text prompts, or animating images with textual guidance. This model version is generated 4K resolution video.Create high-quality, 8-second videos with sound using Gooles's state-of-the-art video generation model. - Video Models > Seedance 2.0 [Seedance 2.0 Text to Video Free API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-text-to-video.md): **Seedance 2.0 text-to-video** model, can generate high-quality cinematic videos based on text prompts, featuring native audio synchronization, realistic physical effects, and multi-shot scene transitions. - Video Models > Seedance 2.0 [Seedance 2.0 Image to Video Free API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-image-to-video.md): **Seedance 2.0's first and last frame** mode can generate smooth video clips from any static image, achieving stable retention of character images, natural and smooth movements, and synchronized audio output. - Video Models > Seedance 2.0 [Seedance 2.0 Reference to Video Free API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-reference-to-video.md): Seedance 2.0 reference material mode. The total number of parameter images and videos must not exceed 5. Please note that the price will be doubled if no reference video is included. - Video Models > Seedance 2.0 [Seedance 2.0 Fast Text to Video Free API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-fast-text-to-video.md): **Seedance 2.0 Fast** is a new-generation multimodal video creation model developed by the Doubao Large Model Team. It inherits the core functions and advantages of the Seedance 2.0 model and features faster generation speed. This is a version of the text-to-video model. - Video Models > Seedance 2.0 [Seedance 2.0 Fast Image to Video API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-fast-image-to-video.md): **Seedance 2.0 Fast** is a new-generation multimodal video creation model developed by the Doubao Large Model Team. It inherits the core functions and advantages of the Seedance 2.0 model and features faster generation speed. This is a version of the image-to-video model. - Video Models > Seedance 2.0 [Seedance 2.0 Fast Reference to Video API - APIPod Docs](https://docs.apipod.ai/seedance/2-0-fast-reference-to-video.md): **Seedance 2.0 Fast** is a new-generation multimodal video creation model developed by the Doubao Large Model Team. It inherits the core functions and advantages of the Seedance 2.0 model and features faster generation speed. This is a version of the reference-to-video model. - Video Models > Grok Imagine [Grok Imageine Text to Video Free API - APIPod Docs](https://docs.apipod.ai/grok-imagine/grok-imagine-t2v.md): **Grok Imagine Text to Video** converts prompts into short AI videos, supporting natural dynamic effects, scene continuity, and synchronized audio, making it ideal for quickly creating visual concepts or short video content. - Video Models > Grok Imagine [Grok Imageine Image to Video Free API - APIPod Docs](https://docs.apipod.ai/grok-imagine/grok-imagine-i2v.md): **Grok Imagine Image to Video** converts a single image into a smooth short video while maintaining the original appearance. It adds dynamic effects, depth, and lighting changes, and is particularly suitable for character animation, product preview, and creative prototype production. - Video Models > Grok Imagine 1.5 Preview [Grok Imagine Video 1.5 Preview](https://docs.apipod.ai/grok-imagine-1-5/grok-imagine-1-5-preview.md): **Grok Imagine Video 1.5 Preview** turns a single still image into fluid, cinematic video. Give it a starting frame and a prompt describing the motion, and it animates the scene, including camera moves, atmosphere, and physics, while staying faithful to your source image. You can generate clips at up to 720p. - Video Models > Sora 2 [Sora 2 Official API](https://docs.apipod.ai/sora-2/sora-2-vip.md): **OpenAI Sora 2 official API** only supports video durations of 4s, 8s, and 12s. It is more stable but has a higher unit price! - Video Models > Google Gemini Omni [Gemini Omni Text to Video](https://docs.apipod.ai/gemini-omni/gemini-omni-t2v.md): **Google Gemini Omni** is a multimodal video generation and editing model. It can convert text, images and reference videos into coherent videos, and delivers stable scene consistency, world understanding capability and natural language control. - Video Models > Google Gemini Omni [Gemini Omni Image to Video](https://docs.apipod.ai/gemini-omni/gemini-omni-i2v.md): **Google Gemini Omni** is a multimodal video generation and editing model. It can convert text, images and reference videos into coherent videos, and delivers stable scene consistency, world understanding capability and natural language control. - Video Models > Google Gemini Omni [Gemini Omni Reference to Video](https://docs.apipod.ai/gemini-omni/gemini-omni-r2v.md): **Google Gemini Omni** is a multimodal video generation and editing model. It can convert text, images and reference videos into coherent videos, and delivers stable scene consistency, world understanding capability and natural language control. - Video Models > Google Gemini Omni [Gemini Omni Extend](https://docs.apipod.ai/gemini-omni/gemini-omni-extend.md): **Google Gemini Omni** is a multimodal video generation and editing model. It can convert text, images and reference videos into coherent videos, and delivers stable scene consistency, world understanding capability and natural language control. - Video Models [Query Video Task API - APIPod Docs](https://docs.apipod.ai/query-video-task.md): Query the create video task status and results. This is a unified query interface that works with all images models. ## Schemas - [NanoBanana](https://docs.apipod.ai/14024404d0.md): - [Seedream4.5](https://docs.apipod.ai/14024943d0.md): - [Seedream5.0](https://docs.apipod.ai/14025414d0.md): - [SubmitTaskResponse](https://docs.apipod.ai/14024675d0.md): - [TaskDetailResponse](https://docs.apipod.ai/14024783d0.md): - [WAN2.7Image](https://docs.apipod.ai/14025470d0.md): - [WAN2.7ImageEdit](https://docs.apipod.ai/14027954d0.md): - [WAN2.7ImagePro](https://docs.apipod.ai/14027955d0.md): - [WAN2.7ImageEditPro](https://docs.apipod.ai/14027966d0.md): - [Veo3.1Fast](https://docs.apipod.ai/14037703d0.md): - [Seedance2.0TextToVideo](https://docs.apipod.ai/14126452d0.md): - [Seedance2.0ImageToVideo](https://docs.apipod.ai/14126895d0.md): - [Seedance2.0ReferenceToVideo](https://docs.apipod.ai/14126909d0.md): - [Seedance2.0FastTextToVideo](https://docs.apipod.ai/14139985d0.md): - [Seedance2.0FastImageToVideo](https://docs.apipod.ai/14139986d0.md): - [Seedance2.0FastReferenceToVideo](https://docs.apipod.ai/14139989d0.md): - [GPTImage2](https://docs.apipod.ai/14454319d0.md): - [GPTImage2Edit](https://docs.apipod.ai/14454323d0.md): - [GrokImagineT2V](https://docs.apipod.ai/14563793d0.md): - [GrokImagineI2V](https://docs.apipod.ai/14563794d0.md): - [Sora2VIP](https://docs.apipod.ai/14667206d0.md): - [NanoBanana2](https://docs.apipod.ai/14807471d0.md): - [GrokImagine15](https://docs.apipod.ai/15589232d0.md): - [Veo3.1 Fast 4K](https://docs.apipod.ai/15747666d0.md): - [Veo3.1 Fast Reference](https://docs.apipod.ai/15747672d0.md): - [Veo3.1 Quality](https://docs.apipod.ai/15747849d0.md): - [Veo3.1 Quality 4K](https://docs.apipod.ai/15747934d0.md): - [Gemini Omni Text to Video](https://docs.apipod.ai/15748097d0.md): - [Gemini Omni Image to Video](https://docs.apipod.ai/15748175d0.md): - [Gemini Omni Reference to Video](https://docs.apipod.ai/15748545d0.md): - [Gemini Omni Extend](https://docs.apipod.ai/15748551d0.md):