Veo 3.1 Lite and fixed per-video pricing
- Added
veo3_litealongsideveo3_faston the existingveo3endpoint. - Fast and Lite support 720p, 1080p, and 4K with 4, 6, or 8-second output.
- Pricing is fixed per video by variant and resolution; duration does not multiply the charge.
- Lite costs 180/210/900 credits and Fast costs 360/390/1080 credits at 720p/1080p/4K.
New Models: Nano Banana 2 Series
Nano Banana 2 (Google)- Up to 14 reference images for image-to-image generation
- Resolution options: 1K, 2K, 4K (default 1K)
- Output format: png or jpg (default jpg)
- Prompt length up to 20000 characters
- Pricing: 1K = 48, 2K = 72, 4K = 108 credits per image
- Budget option at a fixed 24 credits per image
- Up to 10 reference images
- Same 15 aspect ratio options as Nano Banana 2
- No resolution or output format controls
- Use Nano Banana 2 for 2K/4K output, PNG format, or more than 10 references
- Use Lite for high-volume drafts and cost-sensitive workloads
New Model: WAN 2.2 Animate Move
WAN 2.2 Animate Move (Alibaba)- Image + Video animation: Combine an image with a video to create animated content
- No prompt required: Unlike WAN 2.5/2.6, this model is purely image+video driven
- Resolution options: 480p, 580p, 720p
- Input requirements:
- Video: mp4, mov, mkv (max 10MB)
- Image: jpeg, png, webp (max 10MB)
- Fixed 5-second billing
New Model: Seedance 1.5 Pro
Seedance 1.5 Pro (ByteDance)- Audio generation: Generate synchronized audio (voice, sound effects, background music) with
generate_audioparameter - First + Last frame support: Control both start and end frames with up to 2 images in
image_urls - Flexible duration: 4-12 seconds or auto (-1)
- Text-to-video and image-to-video modes
- Resolution options: 480p, 720p (no 1080p)
- Aspect ratios: 16:9, 4:3, 1:1, 3:4, 9:16, 21:9, adaptive
return_last_frameoption for consecutive video generation
When to use Seedance 1.5 Pro vs Seedance Pro Fast:
- Use 1.5 Pro for audio generation and first+last frame control
- Use Pro Fast for 1080p resolution and simpler workflows
New Model: ElevenLabs TTS Turbo 2.5
ElevenLabs Text-to-Speech Turbo 2.5- 50% cheaper than TTS V2
- Faster generation, optimized for low-latency scenarios
- Same 21 voice options and parameters as V2
- Pricing: 40 credits per 1000 characters
New Model: ElevenLabs TTS V2
ElevenLabs Text-to-Speech Multilingual V2- First audio model on the platform
- High-quality multi-language text-to-speech
- 21 voice options (Rachel, Brian, Sarah, etc.)
- Text length: 1-5000 characters
- Voice controls: stability, similarity_boost, style, speed
- Multi-language support with automatic detection
- Pricing: 80 credits per 1000 characters
- Professional/formal: Brian, Daniel, George
- Warm/friendly: Rachel, Sarah, Alice
- Energetic/youthful: Aria, Charlotte, Lily
Model Update: Kling AI V2.6
Kling AI - Text-to-Video & Image-to-Video- NEW: Text-to-video generation support (V2.6)
- NEW: Audio generation with
soundparameter (V2.6 only) - NEW: Aspect ratio control with
aspect_ratioparameter (V2.6 only: 16:9, 9:16, 1:1) - NEW:
kling-v2-6model version - latest and recommended - Supports both
text-to-videoandimage-to-videomodes - Pricing: 65 credits/s (no sound), 130 credits/s (with sound)
New Model: WAN 2.6
WAN 2.6 (Alibaba)- Text-to-video and image-to-video generation
- Extended video duration: up to 15 seconds (vs 10s for WAN 2.5)
- Extended prompt length: up to 5000 characters (vs 800 for WAN 2.5)
- Resolution options: 720p, 1080p
- Pricing: 720p = 80 credits/s, 1080p = 110 credits/s
- Use WAN 2.6 for longer videos (15s), detailed prompts, and improved quality
- Use WAN 2.5 for aspect ratio control, negative prompts, seed control, and prompt expansion
Documentation: Error Codes Reference
Comprehensive Error Codes Documentation- Complete API error codes reference for all HTTP status codes (400-503)
- Detailed error types: authentication, invalid request, quota, permission, rate limit, server errors
- Example responses and solutions for each error code
- Error handling best practices with retry logic examples
- Available in both English and Chinese
New Model: WAN 2.5
WAN 2.5 (Alibaba)- High-quality video generation powered by Alibaba
- Text-to-video and image-to-video modes
- Resolution options: 720p, 1080p
- Video duration: 5s or 10s
- Aspect ratios: 16:9, 9:16, 1:1 (text-to-video only)
- Advanced features: negative prompts, LLM prompt expansion, seed control
- Pricing: 720p = 60 credits/s, 1080p = 100 credits/s
New Models: Flux 2 Series & Nano Banana
Flux 2 Pro & Flux 2 Flex- Megapixel (MP) based pricing for flexible cost optimization
- Flux 2 Pro: 20 credits for first MP + 10 per additional MP
- Flux 2 Flex: 40 credits per MP with advanced controls (guidance & steps)
- Support up to 8 reference images for image-to-image generation
- Total MP = Input MP + Output MP
- Fixed 24 credits per generation - simple and predictable
- Support up to 10 reference images (more than Pro version)
- Smart mode selection automatically chooses text-to-image or image-to-image
- Wide range of aspect ratios: 1:1, 16:9, 9:16, 21:9, and more
New Models: Seedream 4.0 & 4.5
Seedream 4.0 & Seedream 4.5- High-quality image generation powered by ByteDance
- Multi-image fusion: combine up to 14 reference images
- Batch generation: create up to 15 images in one request
- Resolution options: 1K, 2K, 4K
- Seedream 4.0: 30 credits base price
- Seedream 4.5: 40 credits for enhanced quality
New Models: Video Generation Expansion
Kling AI - Image-to-Video- Professional image-to-video generation
- Multiple model versions: v1, v1-5, v1-6, v2-master, v2-1, v2-1-master, v2-5-turbo
- Quality modes: Standard and Pro
- Advanced features: camera controls, dynamic brushes, trajectory animation
- Video duration: 5s or 10s
- Pricing: 10-40 credits (varies by quality mode and duration)
- Fast video generation powered by ByteDance Ark API
- Text-to-video and image-to-video modes
- Resolution options: 480p, 720p, 1080p
- Video duration: 5s or 10s
- Pricing: 10-50 credits/second based on resolution
New Models: Flux Kontext Series
Flux Kontext Pro & Kontext Max- Multi-reference image support (up to 4 images)
- Flexible aspect ratios from 21:9 (cinematic) to 9:21 (portrait)
- Prompt upsampling for enhanced creativity
- Safety tolerance controls (0-6 levels)
- Kontext Pro: 30 credits - fast and cost-effective
- Kontext Max: 60 credits - premium quality output
New Models: Professional Video Generation
Runway Gen-3- Industry-leading video generation technology
- Text-to-video and image-to-video modes
- High-quality output: 720p and 1080p
- Video duration: 5s or 10s
- Aspect ratios: 16:9, 9:16, 1:1
- Pricing: 5-10 credits based on duration and resolution
- Google’s state-of-the-art video generation model
- Three generation modes:
- Text-to-video: pure text prompts
- Image-to-video: first/last frame control
- Reference-to-video: style reference generation
- Multiple aspect ratios supported
- Fast and standard model options
- Pricing: 120 credits per second
Platform Launch
Flux Pro 1.1- High-quality text-to-image generation
- Fast generation speed
- 40 credits per image
- Aspect ratios: 1:1, 16:9, 9:16
- Advanced image generation with resolution control
- Resolution options: 1K, 2K, 4K
- Multi-reference support (up to 8 images)
- Pricing: 10-20 credits based on resolution
- AI-powered virtual garment try-on
- Single garment and combo outfit support
- Model versions: v1 (single item) and v1-5 (top+bottom combo)
- 65 credits per generation
