سایت در حال توسعه است — برخی قابلیت‌ها ممکن است موقتی یا آزمایشی باشند.
IranBrain IranBrain
ورود شروع رایگان
مدل‌ها گالری تعرفه‌ها API مجله

AI MODELS

مدل‌های IranBrain

593 مدل فعال در 6 دسته‌بندی سرویس — مستقیم به ابزارهای داشبورد وصل می‌شوند.

چت هوشمند

گفتگو، نویسندگی و استدلال با مدل‌های زبانی

47 مدل فعال

شروع رایگان
سفارشی 6 اعتبار

ibm-granite: Granite Vision 3.3 2B

Granite-vision-3.3-2b is a compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.

ibm-granite-granite-vision-3.3-2b استفاده ←
سفارشی 8 اعتبار

ibm-granite: Granite 3.2 8B Instruct

Granite-3.2-8B-Instruct is a 8-billion parameter 128K context length language model fine-tuned for reasoning and instruction-following capabilities.

ibm-granite-granite-3.2-8b-instruct استفاده ←
سفارشی 8 اعتبار

ibm-granite: Granite 3.3 8B Instruct

Granite-3.3-8B-Instruct is a 8-billion parameter 128K context length language model fine-tuned for improved reasoning and instruction-following capabilities.

ibm-granite-granite-3.3-8b-instruct استفاده ←
سفارشی 9 اعتبار

ibm-granite: Granite 4.0 H Small

Granite-4.0-H-Small is a 32B parameter long-context instruct model finetuned from Granite-4.0-H-Small-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets.

ibm-granite-granite-4.0-h-small استفاده ←
سفارشی 9 اعتبار

ibm-granite: Granite 4.1 8B

Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from Granite-4.1-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets.

ibm-granite-granite-4.1-8b استفاده ←
سفارشی 9 اعتبار

ibm-granite: Granite Vision 4.1 4B

Granite Vision 4.1 4B is a vision-language model (VLM) that delivers frontier-level performance on structured document extraction tasks — chart extraction, table extraction, and semantic key-value pair extraction — in a compact 4B parameter footprint

ibm-granite-granite-vision-4.1-4b استفاده ←
OpenAI 13 اعتبار

openai: Gpt 5 Nano

Fastest, most cost-effective GPT-5 model from OpenAI

openai-gpt-5-nano استفاده ←
OpenAI 14 اعتبار

openai: Gpt Oss 20B

20b open-weight language model from OpenAI

openai-gpt-oss-20b استفاده ←
سفارشی 15 اعتبار

meta: Llama Guard 4 12B

Replicate

meta-llama-guard-4-12b استفاده ←
OpenAI 15 اعتبار

openai: Gpt 4.1 Nano

Fastest, most cost-effective GPT-4.1 model from OpenAI

openai-gpt-4.1-nano استفاده ←
OpenAI 23 اعتبار

openai: Gpt 4O Mini

Low latency, low cost version of OpenAI's GPT-4o model

openai-gpt-4o-mini استفاده ←
سفارشی 25 اعتبار

meta: Llama 4 Scout Instruct

A 17 billion parameter model with 16 experts

meta-llama-4-scout-instruct استفاده ←
OpenAI 27 اعتبار

openai: Gpt Oss 120B

120b open-weight language model from OpenAI

openai-gpt-oss-120b استفاده ←
سفارشی 36 اعتبار

meta: Llama 4 Maverick Instruct

A 17 billion parameter model with 128 experts

meta-llama-4-maverick-instruct استفاده ←
سفارشی 40 اعتبار

qwen: Qwen3 235B A22B Instruct 2507

Updated Qwen3 model for instruction following

qwen-qwen3-235b-a22b-instruct-2507 استفاده ←
سفارشی 41 اعتبار

qwen: Qwen3 7 Plus

Qwen3.7-Plus is Alibaba's cost-effective multimodal model with vision-language understanding, a 1 million token context window, and strong agentic coding and tool use.

qwen-qwen3-7-plus استفاده ←
OpenAI 60 اعتبار

openai: Gpt 4.1 Mini

Fast, affordable version of GPT-4.1

openai-gpt-4.1-mini استفاده ←
OpenAI 63 اعتبار

openai: Gpt 5 Mini

Faster version of OpenAI's flagship GPT-5 model

openai-gpt-5-mini استفاده ←
Google Gemini 78 اعتبار

google: Gemini 2.5 Flash

Google’s hybrid “thinking” AI model optimized for speed and cost-efficiency

google-gemini-2.5-flash استفاده ←
Alibaba Wan 80 اعتبار

twangodev: Qwenasr

Serve QwenASR speech recognition and alignment.

twangodev-qwenasr استفاده ←
DeepSeek 84 اعتبار

deepseek-ai: Deepseek V3.1

Latest hybrid thinking model from Deepseek

deepseek-ai-deepseek-v3.1 استفاده ←
Google Gemini 100 اعتبار

google: Gemini 3 Flash

Google's most intelligent model built for speed with frontier intelligence, superior search, and grounding

google-gemini-3-flash استفاده ←
DeepSeek 109 اعتبار

deepseek-ai: Deepseek V3

DeepSeek-V3-0324 is the leading non-reasoning model, a milestone for open source

deepseek-ai-deepseek-v3 استفاده ←
Anthropic Claude 175 اعتبار

anthropic: Claude 3.5 Haiku

Anthropic's fastest, most cost-effective model, with a 200K token context window (claude-3-5-haiku-20241022)

anthropic-claude-3.5-haiku استفاده ←
Anthropic Claude 175 اعتبار

anthropic: Claude 4.5 Haiku

Claude Haiku 4.5 gives you similar levels of coding performance but at one-third the cost and more than twice the speed

anthropic-claude-4.5-haiku استفاده ←
OpenAI 200 اعتبار

openai: Gpt 5.6 Luna

OpenAI's GPT-5.6 cost-optimized tier, built for fast, high-volume, latency-sensitive workloads.

openai-gpt-5.6-luna استفاده ←
Google Gemini 300 اعتبار

google: Gemini 3.5 Flash

Google's fast multimodal model with frontier reasoning across agents, coding, and long-context tasks

google-gemini-3.5-flash استفاده ←
OpenAI 300 اعتبار

openai: Gpt 4.1

OpenAI's Flagship GPT model for complex tasks.

openai-gpt-4.1 استفاده ←
OpenAI 313 اعتبار

openai: Gpt 5

OpenAI's new model excelling at coding, writing, and reasoning.

openai-gpt-5 استفاده ←
OpenAI 313 اعتبار

openai: Gpt 5 Structured

GPT-5 with support for structured outputs, web search and custom tools

openai-gpt-5-structured استفاده ←
OpenAI 313 اعتبار

openai: Gpt 5.1

The best model for coding and agentic tasks with configurable reasoning effort.

openai-gpt-5.1 استفاده ←
Anthropic Claude 350 اعتبار

anthropic: Claude Sonnet 5

Anthropic's most agentic Sonnet model, bringing frontier-level coding and tool use at Sonnet's speed and price

anthropic-claude-sonnet-5 استفاده ←
OpenAI 375 اعتبار

openai: Gpt 4O

OpenAI's high-intelligence chat model

openai-gpt-4o استفاده ←
Google Gemini 400 اعتبار

google: Gemini 3.1 Pro

Google's most intelligent model, with improved reasoning and a new medium thinking level

google-gemini-3.1-pro استفاده ←
DeepSeek 438 اعتبار

deepseek-ai: Deepseek R1

A reasoning model trained with reinforcement learning, on par with OpenAI o1

deepseek-ai-deepseek-r1 استفاده ←
OpenAI 438 اعتبار

openai: Gpt 5.2

The best model for coding and agentic tasks across industries

openai-gpt-5.2 استفاده ←
OpenAI 500 اعتبار

openai: Gpt 5.4

OpenAI's most capable frontier model for complex professional work, coding, and multi-step reasoning.

openai-gpt-5.4 استفاده ←
OpenAI 500 اعتبار

openai: Gpt 5.6 Sol

OpenAI's GPT-5.6 flagship tier, built for complex professional work, coding, and deep multi-step reasoning.

openai-gpt-5.6-sol استفاده ←
OpenAI 500 اعتبار

openai: Gpt 5.6 Terra

OpenAI's GPT-5.6 balanced tier, tuned for everyday production work at roughly half the cost of the flagship.

openai-gpt-5.6-terra استفاده ←
Anthropic Claude 525 اعتبار

anthropic: Claude 3.7 Sonnet

The most intelligent Claude model and the first hybrid reasoning model on the market (claude-3-7-sonnet-20250219)

anthropic-claude-3.7-sonnet استفاده ←
Anthropic Claude 525 اعتبار

anthropic: Claude 4 Sonnet

Claude Sonnet 4 is a significant upgrade to 3.7, delivering superior coding and reasoning while responding more precisely to your instructions

anthropic-claude-4-sonnet استفاده ←
Anthropic Claude 525 اعتبار

anthropic: Claude 4.5 Sonnet

Claude Sonnet 4.5 is the best coding model to date, with significant improvements across the entire development lifecycle

anthropic-claude-4.5-sonnet استفاده ←
Anthropic Claude 525 اعتبار

anthropic: Claude Sonnet 4.6

Claude Sonnet 4.6 from Anthropic: a full upgrade to coding, computer use, long-context reasoning, agent planning, knowledge work, and design, with a 1 million token context window in beta.

anthropic-claude-sonnet-4.6 استفاده ←
Anthropic Claude 875 اعتبار

anthropic: Claude Opus 4.6

Anthropic's most intelligent model with state-of-the-art coding, reasoning, and agentic capabilities

anthropic-claude-opus-4.6 استفاده ←
Anthropic Claude 875 اعتبار

anthropic: Claude Opus 4.7

Anthropic's most capable model with a step-change improvement in agentic coding, better vision, and stronger multi-step reasoning

anthropic-claude-opus-4.7 استفاده ←
Anthropic Claude 1,750 اعتبار

anthropic: Claude Fable 5

Claude Fable 5 from Anthropic: the next generation of intelligence for the hardest knowledge work and coding problems.

anthropic-claude-fable-5 استفاده ←
OpenAI 3,750 اعتبار

openai: Gpt 5 Pro

The smartest, fastest, most useful model yet, with built-in thinking that puts expert-level intelligence in everyone’s hands

openai-gpt-5-pro استفاده ←

تصویرساز

تولید و ویرایش تصویر با مدل‌های پیشرفته

182 مدل فعال

شروع رایگان
سفارشی 12 اعتبار

fofr: Color Matcher

Color match and white balance fixes for images

fofr-color-matcher استفاده ←
سفارشی 40 اعتبار

sourceful: Riverflow 2.0 Refsr

Render product images with 100% accuracy and environmental blending

sourceful-riverflow-2.0-refsr استفاده ←
OpenAI 49 اعتبار

openai: Clip

Official CLIP models, generate CLIP (clip-vit-large-patch14) text & image embeddings

openai-clip استفاده ←
سفارشی 50 اعتبار

datalab-to: Marker

Convert PDF to markdown + JSON quickly with high accuracy

datalab-to-marker استفاده ←
سفارشی 65 اعتبار

topazlabs: Image Upscale

Professional-grade image upscaling, from Topaz Labs

topazlabs-image-upscale استفاده ←
سفارشی 67 اعتبار

perceptron-ai-inc: Isaac 0.1

an open-source, 2B-parameter model built for real-world applications

perceptron-ai-inc-isaac-0.1 استفاده ←
سفارشی 70 اعتبار

leonardoai: Phoenix 1.0

Leonardo AI’s first foundational model produces images up to 5 megapixels (fast, quality and ultra modes)

leonardoai-phoenix-1.0 استفاده ←
Black Forest Labs (FLUX) 75 اعتبار

black-forest-labs: Flux 2 Klein 4B

Very fast image generation and editing model. 4 steps distilled, sub-second inference for production and near real-time applications.

black-forest-labs-flux-2-klein-4b استفاده ←
سفارشی 75 اعتبار

sourceful: Riverflow 2.0 Fast

Agentic image model optimized for high-quality, fast generations supporting font control

sourceful-riverflow-2.0-fast استفاده ←
OpenAI 90 اعتبار

openai: Gpt Image 1

A multimodal image generation model that creates high-quality images. You need to bring your own verified OpenAI key to use this model. Your OpenAI account will be charged for usage.

openai-gpt-image-1 استفاده ←
سفارشی 101 اعتبار

moonshotai: Kimi K2 Thinking

Kimi K2 Thinking is the latest, most capable version of an open-source thinking model.

moonshotai-kimi-k2-thinking استفاده ←
سفارشی 120 اعتبار

topazlabs: Image Colorization

Image colorization model from Topaz Labs

topazlabs-image-colorization استفاده ←
سفارشی 126 اعتبار

moonshotai: Kimi K2.5

Moonshot AI's latest open model. It unifies vision and text, thinking and non-thinking modes, and single-agent and multi-agent execution into one model

moonshotai-kimi-k2.5 استفاده ←
سفارشی 148 اعتبار

moonshotai: Kimi K2.6

Moonshot AI's frontier open model, built for long-horizon coding, agent swarms, and autonomous software engineering. 1 trillion parameters, 262k context window, vision and tool use.

moonshotai-kimi-k2.6 استفاده ←
OpenAI 150 اعتبار

openai: O4 Mini

OpenAI's fast, lightweight reasoning model

openai-o4-mini استفاده ←
OpenAI 165 اعتبار

openai: O1 Mini

A small model alternative to o1

openai-o1-mini استفاده ←
سفارشی 175 اعتبار

topazlabs: Dust And Scratch V2

Remove dust and scratches from old photos

topazlabs-dust-and-scratch-v2 استفاده ←
OpenAI 210 اعتبار

openai: Gpt Image 1 Mini

A cost-efficient version of GPT Image 1

openai-gpt-image-1-mini استفاده ←
سفارشی 250 اعتبار

sourceful: Riverflow V2.5 Fast

Speed-optimized variant of Riverflow 2.5 for production and latency-sensitive workflows

sourceful-riverflow-v2.5-fast استفاده ←
سفارشی 250 اعتبار

sourceful: Riverflow V2.5 Pro

Top-quality agentic image model with multi-step reasoning, candidate scoring, and adjustable thinking effort

sourceful-riverflow-v2.5-pro استفاده ←
سفارشی 275 اعتبار

philz1337x: Clarity Pro Upscaler

The first creative upscaler which keeps identity. Stunning photorealistic results, realistic skin, and full creative control.

philz1337x-clarity-pro-upscaler استفاده ←
سفارشی 345 اعتبار

prunaai: P Image Try On

Virtual try-on. Put one or more garments onto a person photo while keeping their face, pose, and body.

prunaai-p-image-try-on استفاده ←
سفارشی 390 اعتبار

ibm-granite: Granite Embedding Small English R2

Granite-embedding-small-english-r2 is a 47M parameter dense biencoder embedding model from the Granite Embeddings collection that can be used to generate high quality text embeddings.

ibm-granite-granite-embedding-small-english-r2 استفاده ←
Google Gemini 400 اعتبار

google: Gemini 3 Pro

Google's most advanced reasoning Gemini model

google-gemini-3-pro استفاده ←
سفارشی 500 اعتبار

minimax: Image 01

Minimax's first image model, with character reference support

minimax-image-01 استفاده ←
Black Forest Labs (FLUX) 500 اعتبار

prunaai: Flux Kontext Fast

Ultra fast flux kontext endpoint

prunaai-flux-kontext-fast استفاده ←
سفارشی 500 اعتبار

prunaai: P Image Edit

A sub 1 second 0.01$ multi-image editing model built for production use cases. For image generation, check out p-image here: https://replicate.com/prunaai/p-image

prunaai-p-image-edit استفاده ←
Ideogram 500 اعتبار

prunaai: P Image Ideogram

P-Image-Ideogram is Pruna AI’s text-to-image model starting from 0,003$ per generation.

prunaai-p-image-ideogram استفاده ←
سفارشی 500 اعتبار

prunaai: P Image Upscale

Fastest image upscaler in the world (<1s) supporting outputs up to 128 MP. contact Pruna AI for dedicated endpoints.

prunaai-p-image-upscale استفاده ←
سفارشی 500 اعتبار

prunaai: Z Image Turbo

Z-Image Turbo is a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.

prunaai-z-image-turbo استفاده ←
Recraft 500 اعتبار

recraft-ai: Recraft Remove Background

Automated background removal for images. Tuned for AI-generated content, product photos, portraits, and design workflows

recraft-ai-recraft-remove-background استفاده ←
Recraft 500 اعتبار

recraft-ai: Recraft Vectorize

Convert raster images to high-quality SVG format with precision and clean vector paths, perfect for logos, icons, and scalable graphics.

recraft-ai-recraft-vectorize استفاده ←
سفارشی 500 اعتبار

reve: Edit Fast

Reve's fast image edit model at only $0.01 per edit

reve-edit-fast استفاده ←
Black Forest Labs (FLUX) 550 اعتبار

black-forest-labs: Flux 2 Klein 9B Base

Un-distilled version of FLUX.2 [klein]. A foundation model for maximum flexibility and control

black-forest-labs-flux-2-klein-9b-base استفاده ←
Black Forest Labs (FLUX) 600 اعتبار

black-forest-labs: Flux 2 Dev

Quality image generation and editing with support for reference images

black-forest-labs-flux-2-dev استفاده ←
OpenAI 600 اعتبار

openai: Gpt Image 2

OpenAI's state-of-the-art image generation model. Create and edit images from text with strong instruction following, sharp text rendering, and detailed editing.

openai-gpt-image-2 استفاده ←
OpenAI 650 اعتبار

openai: Gpt Image 1.5

OpenAI's latest image generation model with better instruction following and adherence to prompts

openai-gpt-image-1.5 استفاده ←
Black Forest Labs (FLUX) 750 اعتبار

black-forest-labs: Flux 2 Klein 9B

4 step distilled version of FLUX.2 [klein]. A foundation model for maximum flexibility and control

black-forest-labs-flux-2-klein-9b استفاده ←
Black Forest Labs (FLUX) 750 اعتبار

black-forest-labs: Flux 2 Pro

High-quality image generation and editing with support for eight reference images

black-forest-labs-flux-2-pro استفاده ←
سفارشی 835 اعتبار

leonardoai: Lucid Origin

Artistic and high-quality visuals with improved prompt adherence, diversity, and definition

leonardoai-lucid-origin استفاده ←
سفارشی 850 اعتبار

retro-diffusion: Rd Fast

Fast pixel art image generation

retro-diffusion-rd-fast استفاده ←
سفارشی 850 اعتبار

sourceful: Riverflow 2.0 Pro

Agentic image model optimized for robust, high-precision generations supporting font control

sourceful-riverflow-2.0-pro استفاده ←
سفارشی 900 اعتبار

bria: Remove Background

Bria AI's remove background model

bria-remove-background استفاده ←
Black Forest Labs (FLUX) 950 اعتبار

black-forest-labs: Flux 2 Klein 4B Base Lora

A version of FLUX.2 [klein] 4B-base that supports fast fine-tuned lora inference

black-forest-labs-flux-2-klein-4b-base-lora استفاده ←
Black Forest Labs (FLUX) 1,000 اعتبار

black-forest-labs: Flux Schnell Lora

The fastest image generation model tailored for fine-tuned use

black-forest-labs-flux-schnell-lora استفاده ←
Google (Veo / Nano Banana / Lyria) 1,000 اعتبار

google: Imagen 4 Fast

Use this fast version of Imagen 4 when speed and cost are more important than quality

google-imagen-4-fast استفاده ←
Google (Veo / Nano Banana / Lyria) 1,000 اعتبار

google: Upscaler

Upscale images 2x or 4x times

google-upscaler استفاده ←
OpenAI 1,000 اعتبار

openai: Dall E 2

The original classic DALLᐧE 2

openai-dall-e-2 استفاده ←
Alibaba Wan 1,000 اعتبار

prunaai: Wan 2.2 Image

This model generates beautiful cinematic 2 megapixel images in 3-4 seconds and is derived from the Wan 2.2 model through optimisation techniques from the pruna package

prunaai-wan-2.2-image استفاده ←
سفارشی 1,000 اعتبار

qwen: Qwen Image 2512

Qwen Image 2512 is an improved version of Qwen Image with more realistic human generation, finer textures, and stronger text rendering

qwen-qwen-image-2512 استفاده ←
سفارشی 1,000 اعتبار

tencent: Hunyuan Image 2.1

Generate high-quality 2K resolution images from text prompts

tencent-hunyuan-image-2.1 استفاده ←
Black Forest Labs (FLUX) 1,050 اعتبار

black-forest-labs: Flux 2 Klein 9B Base Lora

A version of FLUX.2 [klein] 9B-base that supports fast fine-tuned lora inference

black-forest-labs-flux-2-klein-9b-base-lora استفاده ←
Recraft 1,100 اعتبار

recraft-ai: Recraft 20B

Affordable and fast images

recraft-ai-recraft-20b استفاده ←
سفارشی 1,200 اعتبار

retro-diffusion: Rd Plus

High quality and authentic pixel art image generation

retro-diffusion-rd-plus استفاده ←
سفارشی 1,200 اعتبار

retro-diffusion: Rd Tile

All the tools you need for generating pixel art tilesets

retro-diffusion-rd-tile استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Canny Dev

Open-weight edge-guided image generation. Control structure and composition using Canny edge detection.

black-forest-labs-flux-canny-dev استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Depth Dev

Open-weight depth-aware image generation. Edit images while preserving spatial relationships.

black-forest-labs-flux-depth-dev استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Dev

A 12 billion parameter rectified flow transformer capable of generating images from text descriptions

black-forest-labs-flux-dev استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Kontext Dev

Open-weight version of FLUX.1 Kontext

black-forest-labs-flux-kontext-dev استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Krea Dev

An opinionated text-to-image model from Black Forest Labs in collaboration with Krea that excels in photorealism. Creates images that avoid the oversaturated "AI look".

black-forest-labs-flux-krea-dev استفاده ←
Black Forest Labs (FLUX) 1,250 اعتبار

black-forest-labs: Flux Redux Dev

Open-weight image variation model. Create new versions while preserving key elements of your original.

black-forest-labs-flux-redux-dev استفاده ←
Ideogram 1,250 اعتبار

ideogram-ai: Ideogram V2A Turbo

Like Ideogram v2 turbo, but now faster and cheaper

ideogram-ai-ideogram-v2a-turbo استفاده ←
سفارشی 1,250 اعتبار

qwen: Qwen Image

An image generation foundation model in the Qwen series that achieves significant advances in complex text rendering.

qwen-qwen-image استفاده ←
سفارشی 1,250 اعتبار

reve: Create

Image generation model from Reve

reve-create استفاده ←
سفارشی 1,250 اعتبار

wavespeedai: Qwen Image

A 20B MMDiT model for next-gen text-to-image generation

wavespeedai-qwen-image استفاده ←
Black Forest Labs (FLUX) 1,500 اعتبار

black-forest-labs: Flux 2 Max

The highest fidelity image model from Black Forest Labs

black-forest-labs-flux-2-max استفاده ←
سفارشی 1,500 اعتبار

bytedance: Dreamina 3.1

4MP text-to-image generation with enhanced cinematic-quality image generation with precise style control, improved text rendering, and commercial design optimization.

bytedance-dreamina-3.1 استفاده ←
ByteDance Seedream 1,500 اعتبار

bytedance: Seedream 3

A text-to-image model with support for native high-resolution (2K) image generation

bytedance-seedream-3 استفاده ←
ByteDance Seedream 1,500 اعتبار

bytedance: Seedream 4

Unified text-to-image generation and precise single-sentence editing at up to 4K resolution

bytedance-seedream-4 استفاده ←
Ideogram 1,500 اعتبار

ideogram-ai: Ideogram V3 Turbo

Turbo is the fastest and cheapest Ideogram v3. v3 creates images with stunning realism, creative designs, and consistent styles

ideogram-ai-ideogram-v3-turbo استفاده ←
سفارشی 1,500 اعتبار

krea: Krea 2 Medium

Foundation image model from Krea, tuned for expressive illustration, anime, and painterly styles. Fast and consistent across artistic directions.

krea-krea-2-medium استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Fusion

Fusion – Product/object blending that fixes perspective and lighting so the subject melts into a new background via the Fusion LoRA.

qwen-edit-apps-qwen-image-edit-plus-lora-fusion استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Next Scene

Next Scene – “Next beat” cinematic edits that keep subject identity while steering to the next camera move via the Next Scene LoRA

qwen-edit-apps-qwen-image-edit-plus-lora-next-scene استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Photo To Anime

Photo to Anime – Stylized conversion that turns photos into crisp cel-shaded anime frames using the Photo-to-Anime LoRA.

qwen-edit-apps-qwen-image-edit-plus-lora-photo-to-anime استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Relight

Relight – Soft, curtain-filtered relighting that repaints the scene with golden-hour or moody tones using the Relight LoRA.

qwen-edit-apps-qwen-image-edit-plus-lora-relight استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Skin

Skin – Natural beauty retouch that enhances pores and tonal variation (no plastic skin) via the Skin LoRA.

qwen-edit-apps-qwen-image-edit-plus-lora-skin استفاده ←
سفارشی 1,500 اعتبار

qwen-edit-apps: Qwen Image Edit Plus Lora Upscale

Upscale – Detail-loving upscale/restore pass that sharpens textures and color fidelity with the Upscale LoRA.

qwen-edit-apps-qwen-image-edit-plus-lora-upscale استفاده ←
سفارشی 1,500 اعتبار

qwen: Qwen Edit Multiangle

Camera-aware edits for Qwen/Qwen-Image-Edit-2509 with Lightning + multi-angle LoRA

qwen-qwen-edit-multiangle استفاده ←
سفارشی 1,500 اعتبار

qwen: Qwen Image Edit

Edit images using a prompt. This model extends Qwen-Image’s unique text rendering capabilities to image editing tasks, enabling precise text editing

qwen-qwen-image-edit استفاده ←
سفارشی 1,500 اعتبار

qwen: Qwen Image Edit 2511

An enhanced version over Qwen-Image-Edit-2509, featuring multiple improvements including notably better consistency

qwen-qwen-image-edit-2511 استفاده ←
سفارشی 1,500 اعتبار

qwen: Qwen Image Edit Plus

The latest Qwen-Image’s iteration with improved multi-image editing, single-image consistency, and native support for ControlNet

qwen-qwen-image-edit-plus استفاده ←
سفارشی 1,500 اعتبار

qwen: Qwen Image Edit Plus Lora

Qwen Image Edit 2509 LoRA explorer, uses HuggingFace URLs to load any safetensor

qwen-qwen-image-edit-plus-lora استفاده ←
Black Forest Labs (FLUX) 1,600 اعتبار

black-forest-labs: Flux Dev Lora

A version of flux-dev, a text to image model, that supports fast fine-tuned lora inference

black-forest-labs-flux-dev-lora استفاده ←
Black Forest Labs (FLUX) 1,600 اعتبار

black-forest-labs: Flux Kontext Dev Lora

FLUX.1 Kontext[dev] image editing model for running lora finetunes

black-forest-labs-flux-kontext-dev-lora استفاده ←
Google (Veo / Nano Banana / Lyria) 1,700 اعتبار

google: Nano Banana 2 Lite

Google's fastest image generation model — the lightweight, low-cost version of Nano Banana 2, for rapid creation and editing

google-nano-banana-2-lite استفاده ←
Google (Veo / Nano Banana / Lyria) 1,750 اعتبار

google: Nano Banana Pro

Google's state of the art image generation and editing model 🍌🍌

google-nano-banana-pro استفاده ←
سفارشی 1,750 اعتبار

qwen: Qwen Image 2

A next-generation image generation and editing model from Alibaba's Qwen team. Supports text-to-image and image editing with strong text rendering, especially for Chinese.

qwen-qwen-image-2 استفاده ←
سفارشی 1,750 اعتبار

stability-ai: Stable Diffusion 3.5 Medium

2.5 billion parameter image model with improved MMDiT-X architecture

stability-ai-stable-diffusion-3.5-medium استفاده ←
Google Gemini 1,950 اعتبار

google: Gemini 2.5 Flash Image

Google's latest image generation model in Gemini 2.5

google-gemini-2.5-flash-image استفاده ←
Google (Veo / Nano Banana / Lyria) 1,950 اعتبار

google: Nano Banana

Google's latest image editing model in Gemini 2.5

google-nano-banana استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

black-forest-labs: Flux 1.1 Pro

Faster, better FLUX Pro. Text-to-image model with excellent image quality, prompt adherence, and output diversity.

black-forest-labs-flux-1.1-pro استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

black-forest-labs: Flux Fill Dev

Open-weight inpainting model for editing and extending images. Guidance-distilled from FLUX.1 Fill [pro].

black-forest-labs-flux-fill-dev استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

black-forest-labs: Flux Kontext Pro

A state-of-the-art text-based image editing model that delivers high-quality outputs with excellent prompt following and consistent results for transforming images through natural language

black-forest-labs-flux-kontext-pro استفاده ←
سفارشی 2,000 اعتبار

bria: Eraser

SOTA Object removal, enables precise removal of unwanted objects from images while maintaining high-quality outputs. Trained exclusively on licensed data for safe and risk-free commercial use

bria-eraser استفاده ←
سفارشی 2,000 اعتبار

bria: Fibo

SOTA Open source model trained on licensed data, transforming intent into structured control for precise, high-quality AI image generation in enterprise and agentic workflows.

bria-fibo استفاده ←
سفارشی 2,000 اعتبار

bria: Fibo Edit

FIBO-Edit brings the power of structured prompt generation to image editing

bria-fibo-edit استفاده ←
سفارشی 2,000 اعتبار

bria: Generate Background

Bria Background Generation allows for efficient swapping of backgrounds in images via text prompts or reference image, delivering realistic and polished results. Trained exclusively on licensed data for safe and risk-free commercial use

bria-generate-background استفاده ←
سفارشی 2,000 اعتبار

bria: Genfill

Bria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use.

bria-genfill استفاده ←
سفارشی 2,000 اعتبار

bria: Image 3.2

Commercial-ready, trained entirely on licensed data, text-to-image model. With only 4B parameters provides exceptional aesthetics and text rendering. Evaluated to be on par to other leading models in the market

bria-image-3.2 استفاده ←
سفارشی 2,000 اعتبار

bria: Increase Resolution

Bria Increase resolution upscales the resolution of any image. It increases resolution using a dedicated upscaling method that preserves the original image content without regeneration.

bria-increase-resolution استفاده ←
سفارشی 2,000 اعتبار

bria: Product Cutout

Precise AI-powered product cutout with 256-level transparency for eCommerce

bria-product-cutout استفاده ←
سفارشی 2,000 اعتبار

bria: Product Packshot

Transform any product photo into professional 2000x2000px packshots with optimal positioning

bria-product-packshot استفاده ←
سفارشی 2,000 اعتبار

bria: Product Shadow

Add consistent, customizable shadows to product cutouts for enhanced visual appeal

bria-product-shadow استفاده ←
ByteDance Seedream 2,000 اعتبار

bytedance: Seedream 4.5

Seedream 4.5: Upgraded Bytedance image model with stronger spatial understanding and world knowledge

bytedance-seedream-4.5 استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Cartoonify

Turn your image into a cartoon with FLUX.1 Kontext [pro]

flux-kontext-apps-cartoonify استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Change Haircut

Quickly change someone's hair style and hair color, powered by FLUX.1 Kontext [pro]

flux-kontext-apps-change-haircut استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Depth Of Field

Bring your subjects into focus with FLUX.1 Kontext [pro]

flux-kontext-apps-depth-of-field استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Face To Many Kontext

Become a character, in style

flux-kontext-apps-face-to-many-kontext استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Filters

Add simple filters to your images

flux-kontext-apps-filters استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Iconic Locations

Put yourself in an iconic location around the world from a single image

flux-kontext-apps-iconic-locations استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Impossible Scenarios

Experience impossible adventures and extreme scenarios from a single image

flux-kontext-apps-impossible-scenarios استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Multi Image Kontext Pro

An experimental model with FLUX Kontext Pro that can combine two input images

flux-kontext-apps-multi-image-kontext-pro استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Portrait Series

Create a series of portrait photos from a single image

flux-kontext-apps-portrait-series استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Professional Headshot

Create a professional headshot photo from any single image

flux-kontext-apps-professional-headshot استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Restore Image

Use FLUX Kontext to restore, fix scratches and damage, and colorize old photos

flux-kontext-apps-restore-image استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Text Removal

Remove all text from an image with FLUX.1 Kontext

flux-kontext-apps-text-removal استفاده ←
Ideogram 2,000 اعتبار

ideogram-ai: Ideogram V2A

Like Ideogram v2, but faster and cheaper

ideogram-ai-ideogram-v2a استفاده ←
Recraft 2,000 اعتبار

recraft-ai: Recraft V3

Recraft V3 (code-named red_panda) is a text-to-image model with the ability to generate long texts, and images in a wide list of styles. As of today, it is SOTA in image generation, proven by the Text-to-Image Benchmark by Artificial Analysis

recraft-ai-recraft-v3 استفاده ←
Recraft 2,000 اعتبار

recraft-ai: Recraft V4

Recraft's latest image generation model, built around design taste. Strong prompt accuracy, art-directed composition, and integrated text rendering. Fast and cost-efficient at standard resolution.

recraft-ai-recraft-v4 استفاده ←
Recraft 2,000 اعتبار

recraft-ai: Recraft V4.1

Recraft's latest image generation model, built around design taste. Strong prompt accuracy, art-directed composition, and integrated text rendering. Fast and cost-efficient at standard resolution.

recraft-ai-recraft-v4.1 استفاده ←
Recraft 2,000 اعتبار

recraft-ai: Recraft V4.1 Svg

Generate production-ready SVG vector images from text prompts. Recraft V4.1's design taste applied to vector output — clean geometry, structured layers, and editable paths.

recraft-ai-recraft-v4.1-svg استفاده ←
Recraft 2,000 اعتبار

recraft-ai: Recraft V4.1 Utility

A faster, lighter Recraft image generation model optimized for high-volume and production pipelines. Same design taste as V4.1, built for speed and throughput.

recraft-ai-recraft-v4.1-utility استفاده ←
سفارشی 2,000 اعتبار

reve: Edit

Image editing model from Reve

reve-edit استفاده ←
سفارشی 2,000 اعتبار

reve: Remix

Image generation model from Reve which handles multiple input reference images

reve-remix استفاده ←
سفارشی 2,000 اعتبار

stability-ai: Stable Diffusion 3.5 Large Turbo

A text-to-image model that generates high-resolution images with fine details. It supports various artistic styles and produces diverse outputs from the same prompt, with a focus on fewer inference steps

stability-ai-stable-diffusion-3.5-large-turbo استفاده ←
سفارشی 2,000 اعتبار

xai: Grok Imagine Image 2

xAI's Grok Imagine Image 2.0 — text-to-image generation and editing with a quality control and output up to 2k

xai-grok-imagine-image-2 استفاده ←
Recraft 2,200 اعتبار

recraft-ai: Recraft 20B Svg

Affordable and fast vector images

recraft-ai-recraft-20b-svg استفاده ←
ByteDance Seedream 2,250 اعتبار

bytedance: Seedream 5 Pro

ByteDance's flagship text-to-image and image editing model, generating sharp 1K and 2K images from text or up to 10 reference images

bytedance-seedream-5-pro استفاده ←
OpenAI 2,250 اعتبار

openai: O1

OpenAI's first o-series reasoning model

openai-o1 استفاده ←
Black Forest Labs (FLUX) 2,500 اعتبار

black-forest-labs: Flux Canny Pro

Professional edge-guided image generation. Control structure and composition using Canny edge detection

black-forest-labs-flux-canny-pro استفاده ←
Black Forest Labs (FLUX) 2,500 اعتبار

black-forest-labs: Flux Depth Pro

Professional depth-aware image generation. Edit images while preserving spatial relationships.

black-forest-labs-flux-depth-pro استفاده ←
Black Forest Labs (FLUX) 2,500 اعتبار

black-forest-labs: Flux Fill Pro

Professional inpainting and outpainting model with state-of-the-art performance. Edit or extend images with natural, seamless results.

black-forest-labs-flux-fill-pro استفاده ←
سفارشی 2,500 اعتبار

easel: Ai Avatars

Use one or two face images to create AI avatars

easel-ai-avatars استفاده ←
Ideogram 2,500 اعتبار

ideogram-ai: Ideogram V2 Turbo

A fast image model with state of the art inpainting, prompt comprehension and text rendering.

ideogram-ai-ideogram-v2-turbo استفاده ←
سفارشی 2,500 اعتبار

philz1337x: Crystal Upscaler

High-precision image upscaler optimized for portraits, faces and products. One of the upscale modes powered by Clarity AI. X:https://x.com/philz1337x

philz1337x-crystal-upscaler استفاده ←
سفارشی 2,500 اعتبار

xai: Grok Imagine Image Quality

xAI's higher-quality image model with sharper details, better text rendering, and 2k output

xai-grok-imagine-image-quality استفاده ←
Black Forest Labs (FLUX) 2,750 اعتبار

black-forest-labs: Flux Pro

State-of-the-art image generation with top of the line prompt following, visual quality, image detail and output diversity.

black-forest-labs-flux-pro استفاده ←
Black Forest Labs (FLUX) 3,000 اعتبار

black-forest-labs: Flux 1.1 Pro Ultra

FLUX1.1 [pro] in ultra and raw modes. Images are up to 4 megapixels. Use raw mode for realism.

black-forest-labs-flux-1.1-pro-ultra استفاده ←
Black Forest Labs (FLUX) 3,000 اعتبار

black-forest-labs: Flux 2 Flex

Max-quality image generation and editing with support for ten reference images

black-forest-labs-flux-2-flex استفاده ←
Black Forest Labs (FLUX) 3,000 اعتبار

black-forest-labs: Flux Pro Finetuned

Inference model for FLUX.1 [pro] using custom `finetune_id`

black-forest-labs-flux-pro-finetuned استفاده ←
Google (Veo / Nano Banana / Lyria) 3,000 اعتبار

google: Imagen 4 Ultra

Use this ultra version of Imagen 4 when quality matters more than speed and cost

google-imagen-4-ultra استفاده ←
Ideogram 3,000 اعتبار

ideogram-ai: Ideogram V3 Balanced

Balance speed, quality and cost. Ideogram v3 creates images with stunning realism, creative designs, and consistent styles

ideogram-ai-ideogram-v3-balanced استفاده ←
Ideogram 3,000 اعتبار

ideogram-ai: Ideogram V4 Balanced

Balance speed, quality and cost. Ideogram v4 creates images with stunning realism, creative designs, and consistent styles

ideogram-ai-ideogram-v4-balanced استفاده ←
سفارشی 3,000 اعتبار

krea: Krea 2 Large

Krea's flagship foundation image model. Larger and more flexible than Krea 2 Medium, with particular strength in photorealism and expressive artistic styles.

krea-krea-2-large استفاده ←
سفارشی 3,250 اعتبار

stability-ai: Stable Diffusion 3.5 Large

A text-to-image model that generates high-resolution images with fine details. It supports various artistic styles and produces diverse outputs from the same prompt, thanks to Query-Key Normalization.

stability-ai-stable-diffusion-3.5-large استفاده ←
Google (Veo / Nano Banana / Lyria) 3,350 اعتبار

google: Nano Banana 2

Google's fast image generation model with conversational editing, multi-image fusion, and character consistency

google-nano-banana-2 استفاده ←
Black Forest Labs (FLUX) 3,500 اعتبار

black-forest-labs: Flux 1.1 Pro Ultra Finetuned

Inference model for FLUX 1.1 [pro] Ultra using custom `finetune_id`. Supports 4MP images and raw mode for realism

black-forest-labs-flux-1.1-pro-ultra-finetuned استفاده ←
سفارشی 3,500 اعتبار

ibm-granite: Granite Embedding 278M Multilingual

Granite-Embedding-278M-Multilingual is a 278M parameter model from the Granite Embeddings suite that can be used to generate high quality text embeddings

ibm-granite-granite-embedding-278m-multilingual استفاده ←
سفارشی 3,500 اعتبار

retro-diffusion: Rd Animation

Style consistent animated pixel art sprite generation

retro-diffusion-rd-animation استفاده ←
سفارشی 3,750 اعتبار

qwen: Qwen Image 2 Pro

The pro version of Qwen Image 2 from Alibaba's Qwen team. Enhanced text rendering, realism, and semantic adherence for high-quality image generation and editing.

qwen-qwen-image-2-pro استفاده ←
Black Forest Labs (FLUX) 4,000 اعتبار

black-forest-labs: Flux Kontext Max

A premium text-based image editing model that delivers maximum performance and improved typography generation for transforming images through natural language prompts

black-forest-labs-flux-kontext-max استفاده ←
Black Forest Labs (FLUX) 4,000 اعتبار

flux-kontext-apps: Multi Image Kontext Max

An experimental FLUX Kontext model that can combine two input images

flux-kontext-apps-multi-image-kontext-max استفاده ←
Black Forest Labs (FLUX) 4,000 اعتبار

flux-kontext-apps: Multi Image List

FLUX Kontext max with list input for multiple images

flux-kontext-apps-multi-image-list استفاده ←
Ideogram 4,000 اعتبار

ideogram-ai: Ideogram V2

An excellent image model with state of the art inpainting, prompt comprehension and text rendering

ideogram-ai-ideogram-v2 استفاده ←
Recraft 4,000 اعتبار

recraft-ai: Recraft V3 Svg

Recraft V3 SVG (code-named red_panda) is a text-to-image model with the ability to generate high quality SVG images including logotypes, and icons. The model supports a wide list of styles.

recraft-ai-recraft-v3-svg استفاده ←
سفارشی 4,000 اعتبار

tencent: Hunyuan Image 3

A powerful native multimodal model for image generation (PrunaAI squeezed)

tencent-hunyuan-image-3 استفاده ←
Ideogram 4,500 اعتبار

ideogram-ai: Ideogram V3 Quality

The highest quality Ideogram v3 model. v3 creates images with stunning realism, creative designs, and consistent styles

ideogram-ai-ideogram-v3-quality استفاده ←
Ideogram 4,500 اعتبار

ideogram-ai: Layerize

Take a flat graphic, remove text, and get structured text layers back for editing and recomposing

ideogram-ai-layerize استفاده ←
Ideogram 5,000 اعتبار

ideogram-ai: Ideogram Character

Generate consistent characters from a single reference image. Outputs can be in many styles. You can also use inpainting to add your character to an existing image.

ideogram-ai-ideogram-character استفاده ←
Ideogram 5,000 اعتبار

ideogram-ai: Ideogram V4 Quality

The highest quality Ideogram v4 model. v4 creates images with stunning realism, creative designs, and consistent styles

ideogram-ai-ideogram-v4-quality استفاده ←
OpenAI 6,000 اعتبار

openai: Dall E 3

An AI system that can create realistic images and art from a description in natural language.

openai-dall-e-3 استفاده ←
Black Forest Labs (FLUX) 7,000 اعتبار

black-forest-labs: Flux 2 Klein 4B Base

Un-distilled version of FLUX.2 [klein]. Optimized for fine-tuning, customization, and post-training workflows

black-forest-labs-flux-2-klein-4b-base استفاده ←
سفارشی 8,500 اعتبار

qwen: Qwen Image Lora Trainer Legacy

Fine-tunable Qwen Image model with exceptional composition abilities - train custom LoRAs for any style or subject

qwen-qwen-image-lora-trainer-legacy استفاده ←
سفارشی 10,000 اعتبار

reve: Reve 2.1

Generate and edit images from text and reference images with Reve 2.1

reve-reve-2.1 استفاده ←
Recraft 12,500 اعتبار

recraft-ai: Recraft V4 Pro

Recraft's latest image generation model at ~2048px resolution. Same design taste and prompt accuracy as V4, with higher resolution for print-ready and large-scale work.

recraft-ai-recraft-v4-pro استفاده ←
Recraft 12,500 اعتبار

recraft-ai: Recraft V4.1 Pro

Recraft's latest image generation model at ~2048px resolution. Same design taste and prompt accuracy as V4.1, with higher resolution for print-ready and large-scale work.

recraft-ai-recraft-v4.1-pro استفاده ←
Recraft 12,500 اعتبار

recraft-ai: Recraft V4.1 Pro Svg

Generate detailed SVG vector graphics from text prompts. Recraft V4.1 Pro's design taste with more geometric detail and finer paths — clean layers, editable output, and scalable to any size.

recraft-ai-recraft-v4.1-pro-svg استفاده ←
Recraft 12,500 اعتبار

recraft-ai: Recraft V4.1 Utility Pro

A faster, lighter Recraft image generation model at ~2048px resolution, optimized for high-volume production. Design taste and prompt accuracy at high resolution with better throughput.

recraft-ai-recraft-v4.1-utility-pro استفاده ←
سفارشی 12,500 اعتبار

sync: Lipsync 2

Generate realistic lipsyncs with Sync Labs' 2.0 model

sync-lipsync-2 استفاده ←
Recraft 15,000 اعتبار

recraft-ai: Recraft Creative Upscale

Creative Upscale focuses on enhancing details and refining complex elements in the image. It doesn’t just increase resolution but adds depth by improving textures, fine details, and facial features.

recraft-ai-recraft-creative-upscale استفاده ←
Recraft 15,000 اعتبار

recraft-ai: Recraft V4 Pro Svg

Generate detailed SVG vector graphics from text prompts. Recraft V4 Pro's design taste with more geometric detail and finer paths — clean layers, editable output, and scalable to any size.

recraft-ai-recraft-v4-pro-svg استفاده ←
سفارشی 20,813 اعتبار

sync: Lipsync 2 Pro

Studio-grade lipsync in minutes, not weeks

sync-lipsync-2-pro استفاده ←
سفارشی 41,750 اعتبار

sync: React 1

Realistic lipsync with refined human emotion capabilities

sync-react-1 استفاده ←
سفارشی 50,000 اعتبار

intelligent-utilities: Html To Image

Replicate

intelligent-utilities-html-to-image استفاده ←
سفارشی 100,000 اعتبار

nightmareai: Real Esrgan

Real-ESRGAN with optional face correction and adjustable upscale

nightmareai-real-esrgan استفاده ←
Black Forest Labs (FLUX) 150,000 اعتبار

black-forest-labs: Flux Redux Schnell

Fast, efficient image variation model for rapid iteration and experimentation.

black-forest-labs-flux-redux-schnell استفاده ←
Black Forest Labs (FLUX) 150,000 اعتبار

black-forest-labs: Flux Schnell

The fastest image generation model tailored for local development and personal use

black-forest-labs-flux-schnell استفاده ←
Black Forest Labs (FLUX) 250,000 اعتبار

prunaai: Flux Fast

This is the fastest Flux endpoint in the world.

prunaai-flux-fast استفاده ←
سفارشی 250,000 اعتبار

prunaai: Hidream L1 Fast

This is an optimised version of the hidream-l1 model using the pruna ai optimisation toolkit!

prunaai-hidream-l1-fast استفاده ←
سفارشی 250,000 اعتبار

prunaai: P Image

A sub 1 second text-to-image model built for production use cases.

prunaai-p-image استفاده ←
سفارشی 250,000 اعتبار

prunaai: P Image Lora

Use trained LoRAs from the https://replicate.com/prunaai/p-image-trainer. Find or contribute LoRAs here https://huggingface.co/collections/PrunaAI/p-image-loras

prunaai-p-image-lora استفاده ←
Recraft 300,000 اعتبار

recraft-ai: Recraft Crisp Upscale

Designed to make images sharper and cleaner, Crisp Upscale increases overall quality, making visuals suitable for web use or print-ready materials.

recraft-ai-recraft-crisp-upscale استفاده ←

ویدیوساز

ساخت ویدیو از متن، تصویر و رفرنس

293 مدل فعال

شروع رایگان
سفارشی 5 اعتبار

lucataco: Frame Extractor

Extract the first or last frame from any video file as a high-quality image

lucataco-frame-extractor استفاده ←
Udio 8 اعتبار

lucataco: Video Audio Merge

merge a video and an audio file

lucataco-video-audio-merge استفاده ←
Udio 17 اعتبار

emersimeon: Image Audio Video

Add an image and a song and generate a video

emersimeon-image-audio-video استفاده ←
سفارشی 25 اعتبار

bitflow: Kinetic Captions

Generate dynamic, stylish captions and hard-burn them into videos using ASS subtitles and FFmpeg.

bitflow-kinetic-captions استفاده ←
سفارشی 25 اعتبار

bitflow: Video Color Filter Lut

This LUT-based color filter is ideal for color grading user-generated AI videos or short videos shot on smartphones.

bitflow-video-color-filter-lut استفاده ←
سفارشی 25 اعتبار

chamelaion: Lipsync

Lip sync any video in up to Full HD resolution at 1080p to any audio track, including videos with multiple speakers. Active speaker detection is built in, while speaking style preservation ensures natural mouth movements.

chamelaion-lipsync استفاده ←
سفارشی 25 اعتبار

craftfulcharles: Waveform

Generate Waveform videos using an audio file and an image.

craftfulcharles-waveform استفاده ←
سفارشی 25 اعتبار

foixasoftware: Ffmpeg

merge videos

foixasoftware-ffmpeg استفاده ←
سفارشی 25 اعتبار

hautechai: Json To Media

Replicate

hautechai-json-to-media استفاده ←
سفارشی 25 اعتبار

idan054: Sarra Video Maker V1

This Adds Background Music & Prefect subtitles from json file

idan054-sarra-video-maker-v1 استفاده ←
سفارشی 25 اعتبار

lex2029: Vlogme Avatar Bridge

Create a vertical talking-avatar video from a centered photo and speech audio.

lex2029-vlogme-avatar-bridge استفاده ←
سفارشی 25 اعتبار

littletry: Media Metadata

Extract normalized metadata from video, audio, and image files.

littletry-media-metadata استفاده ←
سفارشی 25 اعتبار

littletry: Video Frame Custom

Extract the first, last, middle, specific, or evenly spaced frames from a video.

littletry-video-frame-custom استفاده ←
سفارشی 25 اعتبار

lucataco: Dotted Waveform Visualizer

Create a dotted waveform video from an audio file

lucataco-dotted-waveform-visualizer استفاده ←
سفارشی 25 اعتبار

lucataco: Split Screen Video

Combines two videos into a single split-screen layout

lucataco-split-screen-video استفاده ←
سفارشی 25 اعتبار

lucataco: Vid2Webp

Convert your video into webp format (with looping)

lucataco-vid2webp استفاده ←
سفارشی 25 اعتبار

onemadgeek: Frames To Video Merger

Convert a set of image frames (JPG or PNG) into a high-quality MP4 video. Automatically handles sorting and frame order for smooth playback.

onemadgeek-frames-to-video-merger استفاده ←
سفارشی 25 اعتبار

peelss: Ai Video Subtitle

Add animated captions to your video

peelss-ai-video-subtitle استفاده ←
سفارشی 25 اعتبار

smokeyang: Transition Generator

Cinematic photo or video transition generator featuring 4 premium motion styles driven by Stable Diffusion.

smokeyang-transition-generator استفاده ←
سفارشی 25 اعتبار

thecmdrunner: Feather 1

AI-powered Text-To-Video Generation for Animated Motion Graphics

thecmdrunner-feather-1 استفاده ←
سفارشی 25 اعتبار

zhyjiong: Sora Watermark Remover

Remove 'Sora' watermark on videos generated by sora

zhyjiong-sora-watermark-remover استفاده ←
سفارشی 38 اعتبار

lucataco: Image To Video Slideshow

Transform a collection of images into a video slideshow

lucataco-image-to-video-slideshow استفاده ←
سفارشی 44 اعتبار

lucataco: Dotted Video

Converts a video into a black and white dotted video effect

lucataco-dotted-video استفاده ←
Udio 45 اعتبار

nightowlstudio-11: Clipforge

Add Text & Music to Videos

nightowlstudio-11-clipforge استفاده ←
سفارشی 46 اعتبار

hjunior29: Video Text Remover

Clean videos by automatically removing text overlays

hjunior29-video-text-remover استفاده ←
سفارشی 49 اعتبار

mattsays: Sam3 Image

A unified foundation model for prompt-based segmentation in images and videos

mattsays-sam3-image استفاده ←
سفارشی 56 اعتبار

ghostljj: Sora2 Watermark Remover Fix

Removes the watermark from Sora 2 videos using a trained model and IOpaint (Fixed watermark detection leaks and production errors.)

ghostljj-sora2-watermark-remover-fix استفاده ←
سفارشی 56 اعتبار

hcolde: Bmv2

Alpha-only video matting for 5-10 second white-background videos using BackgroundMattingV2 MobileNetV2 TorchScript.

hcolde-bmv2 استفاده ←
سفارشی 56 اعتبار

hcolde: Rvm

RobustVideoMatting on Replicate: input mp4 video, output black-and-white alpha-mask.mp4.

hcolde-rvm استفاده ←
سفارشی 56 اعتبار

nelsonjchen: Op Replay Clipper

GPU accelerated replay renderer / video data clipper for comma.ai connect's openpilot route data. SEE README.

nelsonjchen-op-replay-clipper استفاده ←
سفارشی 75 اعتبار

mirelo: Video To Sfx V1

Generate synced sounds for any video, and return it with its new sound track

mirelo-video-to-sfx-v1 استفاده ←
سفارشی 80 اعتبار

mirelo: Video To Sfx V1.5

Generate synced sounds for any video and return it with its new soundtrack - now enhanced in version 1.5 for improved sound synchronization and realism

mirelo-video-to-sfx-v1.5 استفاده ←
سفارشی 85 اعتبار

mptamilselvan: Download Media

Download videos or extract audio from popular social media platforms quickly and easily. This tool supports links from platforms like Facebook, Instagram, and YouTube, allowing users to save content for offline viewing or personal use.

mptamilselvan-download-media استفاده ←
سفارشی 95 اعتبار

megacubos: Video Watermark

Add a customizable watermark to any video with a simple API call. Used in the video AI creator at UGCMade.com and brought to you by your friends at Megacubos.com.

megacubos-video-watermark استفاده ←
سفارشی 110 اعتبار

bfirsh: Concatenate Videos

Stitches videos together

bfirsh-concatenate-videos استفاده ←
سفارشی 115 اعتبار

nicolascoutureau: Video Utils

Replicate

nicolascoutureau-video-utils استفاده ←
سفارشی 115 اعتبار

pixverse: Pixverse V4

Quickly generate smooth 5s or 8s videos at 540p, 720p or 1080p

pixverse-pixverse-v4 استفاده ←
سفارشی 125 اعتبار

pixverse: Pixverse V3.5

Create videos in as little as 10 seconds. 5s or 8s videos at 360p, 540p, 720p or 1080p.

pixverse-pixverse-v3.5 استفاده ←
سفارشی 135 اعتبار

lucataco: Video Merge

Simple tool to merge together separate video snippets

lucataco-video-merge استفاده ←
سفارشی 170 اعتبار

onemadgeek: Video To Frames Extractor

Extract frames from videos at custom frame rates

onemadgeek-video-to-frames-extractor استفاده ←
سفارشی 180 اعتبار

topazlabs: Video Upscale

Video Upscaling from Topaz Labs

topazlabs-video-upscale استفاده ←
سفارشی 200 اعتبار

csviverdeia: Locateanything 3B H100

LocateAnything-3B em H100 (Hopper) — visual grounding imagem+video (boxes/points). Detecção densa, pointing, tracker none/sort/reid. PRINCIPAL (rápido).

csviverdeia-locateanything-3b-h100 استفاده ←
سفارشی 220 اعتبار

pixverse: Pixverse V5

Create 5s-8s videos with enhanced character movement, visual effects, and exclusive 1080p-8s support. Optimized for anime characters and complex actions

pixverse-pixverse-v5 استفاده ←
سفارشی 230 اعتبار

bitflow: Video Concat Safe Pro

When combining short AI videos encoded with different tools, codecs, and frame rates, this process standardizes them to ensure they are merged safely so that the final combined video remains intact.

bitflow-video-concat-safe-pro استفاده ←
Udio 230 اعتبار

zsxkib: Mmaudio

Add sound to video using the MMAudio V2 model. An advanced AI model that synthesizes high-quality audio from video content, enabling seamless video-to-audio transformation.

zsxkib-mmaudio استفاده ←
Udio 244 اعتبار

acappemin: Deepaudio V1

DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation

acappemin-deepaudio-v1 استفاده ←
Udio 244 اعتبار

acappemin: Video To Audio And Piano

Enhance Generation Quality of Flow Matching V2A Model via Multi-Step CoT-Like Guidance and Combined Preference Optimization

acappemin-video-to-audio-and-piano استفاده ←
سفارشی 244 اعتبار

ailoov: Video Rembg

Replicate

ailoov-video-rembg استفاده ←
سفارشی 244 اعتبار

bytedance: Sa2Va 4B Video

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-4b-video استفاده ←
سفارشی 244 اعتبار

colinhughes2121: Thumbnail Clickbait Enhancer

Boost any photo into a high-CTR YouTube thumbnail. Saturated colors, dramatic lighting, eye-catching contrast. For creators, video editors, content agencies. · api.gocreativeai.com — AI agent APIs

colinhughes2121-thumbnail-clickbait-enhancer استفاده ←
سفارشی 244 اعتبار

hcolde: Bgogo Feno

BiRefNet + Cutie video segmentation with stacked output video

hcolde-bgogo-feno استفاده ←
سفارشی 244 اعتبار

jkukul: Minimax Remover

Remove any object from video - fast.

jkukul-minimax-remover استفاده ←
Google (Veo / Nano Banana / Lyria) 244 اعتبار

lucataco: Interactiveomni 8B

A unified omni-modal model that can simultaneously receive inputs such as images, audio, text, and video and directly generate coherent text and speech

lucataco-interactiveomni-8b استفاده ←
سفارشی 244 اعتبار

lucataco: Motif Video

Motif-Video-2B: a 2B-parameter text-to-video diffusion transformer

lucataco-motif-video استفاده ←
سفارشی 244 اعتبار

prakhar-bhartiya: Meta Tribev2 Social Media Content Signal

Predicts virality of short videos using a Meta TRIBE v2 brain-encoding model

prakhar-bhartiya-meta-tribev2-social-media-content-signal استفاده ←
سفارشی 244 اعتبار

zsxkib: Create Video Dataset

Easily create video datasets with auto-captioning for Hunyuan-Video LoRA finetuning

zsxkib-create-video-dataset استفاده ←
سفارشی 244 اعتبار

zsxkib: Keep

🧼Upscales faces in videos look to be clearer and better using KEEP, Kalman-Inspired Feature Propagation for Video Face Super-Resolution🫟

zsxkib-keep استفاده ←
سفارشی 250 اعتبار

lucataco: Trim Video

Simple tool to quickly trim a video or audio file

lucataco-trim-video استفاده ←
سفارشی 260 اعتبار

idan054: Better Video Merge

Fix Diffrent Sizes for each clip. Fork of lucataco/cog-video-merge.git

idan054-better-video-merge استفاده ←
سفارشی 265 اعتبار

lucataco: Qwen2.5 Omni 7B

Qwen2.5-Omni is an end-to-end multimodal model designed to perceive diverse modalities, including text, images, audio, and video, while simultaneously generating text and natural speech responses in a streaming manner.

lucataco-qwen2.5-omni-7b استفاده ←
سفارشی 315 اعتبار

zsxkib: Thinksound

Generate contextual audio from video using step-by-step reasoning🎶

zsxkib-thinksound استفاده ←
سفارشی 345 اعتبار

ddvinh1: Video Faceswap Gpu

Replicate

ddvinh1-video-faceswap-gpu استفاده ←
Alibaba Wan 350 اعتبار

atonamy: Wan Alpha

Wan-Alpha text-to-video generation with alpha channel, optimized for Replicate deployment. WebM/WebP output with transparency support.

atonamy-wan-alpha استفاده ←
Alibaba Wan 350 اعتبار

jichengdu: Wan I2V

i2v-14B-720p-2.1

jichengdu-wan-i2v استفاده ←
سفارشی 350 اعتبار

paullux: Framepack Runner

FramePack video generation with image + motion prompt. Based on Stanford's 2025 model.

paullux-framepack-runner استفاده ←
سفارشی 350 اعتبار

subhash25rawat: Venhance

Enhance your video quality with AI

subhash25rawat-venhance استفاده ←
سفارشی 350 اعتبار

zsxkib: Memo

MEMO is a state-of-the-art open-weight model for audio-driven talking video generation.

zsxkib-memo استفاده ←
OpenAI 381 اعتبار

biggpt1: Pana Video

Replicate

biggpt1-pana-video استفاده ←
سفارشی 381 اعتبار

civicvideo: Billy Portrait

Replicate

civicvideo-billy-portrait استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Blade Runner

Hunyuan-Video model finetuned on Blade Runner (1982). Trigger word is "BLDRN". Use "A video in the style of BLDRN, BLDRN" at the beginning of your prompt for best results.

deepfates-hunyuan-blade-runner استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Blade Runner 2049

Hunyuan-Video model finetuned on Blade Runner 2049 (2017). Trigger word is "BLDRN". Use "A video in the style of BLDRN, BLDRN" at the beginning of your prompt for best results.

deepfates-hunyuan-blade-runner-2049 استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Game Of Thrones

Hunyuan-Video model finetuned on Game of Thrones (2011). Trigger word is "GMFTH". Use "A video in the style of GMFTH, GMFTH" at the beginning of your prompt for best results

deepfates-hunyuan-game-of-thrones استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Her

Hunyuan-Video model finetuned on Her (2013). Trigger word is "HR". Use "A video in the style of HR, HR" at the beginning of your prompt for best results.

deepfates-hunyuan-her استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Inception

Hunyuan-Video model finetuned on Inception (2010). Trigger word is "NCPTN". Use "A video in the style of NCPTN, NCPTN" at the beginning of your prompt for best results.

deepfates-hunyuan-inception استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Indiana Jones

Hunyuan-Video model finetuned on Indiana Jones Series (1981). Trigger word is "NDNJN". Use "A video in the style of NDNJN, NDNJN" at the beginning of your prompt for best results.

deepfates-hunyuan-indiana-jones استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Joker

Hunyuan-Video model finetuned on Joker (2019). Trigger word is "JKR". Use "A video in the style of JKR, JKR" at the beginning of your prompt for best results.

deepfates-hunyuan-joker استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Once Upon A Time In Hollywood

Hunyuan-Video model finetuned on Once Upon a Time in Hollywood (2019). Trigger word is "NCPNT". Use "A video in the style of NCPNT, NCPNT" at the beginning of your prompt for best results.

deepfates-hunyuan-once-upon-a-time-in-hollywood استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Rrr

Hunyuan-Video model finetuned on RRR (2022). Trigger word is "RRR". Use "A video in the style of RRR, RRR" at the beginning of your prompt for best results.

deepfates-hunyuan-rrr استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan The Lord Of The Rings

Hunyuan-Video model finetuned on The Lord of the Rings Trilogy (2001). Trigger word is "THLRD". Use "A video in the style of THLRD, THLRD" at the beginning of your prompt for best results.

deepfates-hunyuan-the-lord-of-the-rings استفاده ←
سفارشی 381 اعتبار

deepfates: Hunyuan Westworld

Hunyuan-Video model finetuned on Westworld (2016). Trigger word is "WSTWR". Use "A video in the style of WSTWR, WSTWR" at the beginning of your prompt for best results.

deepfates-hunyuan-westworld استفاده ←
سفارشی 381 اعتبار

derickson: Dn2Trl 1

Dune 2 Trailer LoRA for hunyuan-video-lora. Trigger word is DN2TRL

derickson-dn2trl-1 استفاده ←
سفارشی 381 اعتبار

derickson: Neomacao 1

attempt at a video lora

derickson-neomacao-1 استفاده ←
سفارشی 381 اعتبار

dfischer: Tokmeshmetal 4 Epoch

Replicate

dfischer-tokmeshmetal-4-epoch استفاده ←
سفارشی 381 اعتبار

dfischer: Toksci 2 Epoch

Replicate

dfischer-toksci-2-epoch استفاده ←
سفارشی 381 اعتبار

foxelli-group-video: Amigurumi V2

Replicate

foxelli-group-video-amigurumi-v2 استفاده ←
سفارشی 381 اعتبار

itsheadkrack: Videome

Replicate

itsheadkrack-videome استفاده ←
سفارشی 381 اعتبار

jimothyjohn: Superslomo

Slow down choppy videos up to 10x slower. It's basically 40x more slo mo for a small price..

jimothyjohn-superslomo استفاده ←
سفارشی 381 اعتبار

lucataco: Longcat Video

Replicate

lucataco-longcat-video استفاده ←
سفارشی 381 اعتبار

lucataco: Ltx Video Iclora

LTX Video 0.9.7 Distilled with ICLoRAs

lucataco-ltx-video-iclora استفاده ←
سفارشی 381 اعتبار

lucataco: Stable Avatar

End-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing

lucataco-stable-avatar استفاده ←
سفارشی 381 اعتبار

mushroomfleet: Fxf Tokyo Meet

JDM style Car Meetup

mushroomfleet-fxf-tokyo-meet استفاده ←
سفارشی 381 اعتبار

mushroomfleet: Xjx Tokyoracer

Clasic JDM highway racing

mushroomfleet-xjx-tokyoracer استفاده ←
سفارشی 381 اعتبار

nanoukader: Video

Replicate

nanoukader-video استفاده ←
سفارشی 381 اعتبار

papina: Seedvr2

🔥 SeedVR2: one-step video & image restoration with 7B and Adjustable Resolution

papina-seedvr2 استفاده ←
سفارشی 381 اعتبار

replicate: Ltxvideo 2.3 Lora

LTX 2.3 with popular community LoRA tasks like Ingredients or camera control

replicate-ltxvideo-2.3-lora استفاده ←
سفارشی 381 اعتبار

shridharathi: Ghibli Vid

Make a video of anything in Studio Ghibli style

shridharathi-ghibli-vid استفاده ←
سفارشی 381 اعتبار

shridharathi: Van Gogh Vid

Make your videos van gogh-esque

shridharathi-van-gogh-vid استفاده ←
Luma Dream Machine 381 اعتبار

twinstarcreatives: Luma Portrait

Replicate

twinstarcreatives-luma-portrait استفاده ←
سفارشی 415 اعتبار

bytedance: Sa2Va 8B Video

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-8b-video استفاده ←
سفارشی 430 اعتبار

sai88uk: Minicpm V 45 V9

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding

sai88uk-minicpm-v-45-v9 استفاده ←
سفارشی 488 اعتبار

andreasjansson: Video Stitcher

Fast GPU-powered concatenation of multiple videos, with short audio crossfades

andreasjansson-video-stitcher استفاده ←
Luma Dream Machine 500 اعتبار

luma: Reframe Image

Change the aspect ratio of any photo using AI (not cropping)

luma-reframe-image استفاده ←
Luma Dream Machine 550 اعتبار

luma: Modify Video

Modify a video with style transfer and prompt-based editing

luma-modify-video استفاده ←
سفارشی 600 اعتبار

lucataco: Videollama3 7B

VideoLLaMA 3: Frontier Multimodal Foundation Models for Video Understanding

lucataco-videollama3-7b استفاده ←
Alibaba Wan 625 اعتبار

wan-video: Wan 2.2 5B Fast

The fastest Wan 2.2 text-to-image and image-to-video model

wan-video-wan-2.2-5b-fast استفاده ←
Black Forest Labs (FLUX) 650 اعتبار

beancakes: Flux Cyberpunk People

Flux lora inspired by the video game, Cyberpunk 2077, for generating people in futuristic and dystopian backgrounds

beancakes-flux-cyberpunk-people استفاده ←
سفارشی 650 اعتبار

lucataco: Nsfw Video Detection

FalconAIs NSFW detection model, extended for videos

lucataco-nsfw-video-detection استفاده ←
سفارشی 750 اعتبار

bytedance: Sa2Va 26B Image

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-26b-image استفاده ←
سفارشی 800 اعتبار

cjwbw: Controlvideo

Training-free Controllable Text-to-Video Generation

cjwbw-controlvideo استفاده ←
سفارشی 1,250 اعتبار

prunaai: P Video

Fast video generation with built-in draft mode for rapid creative iteration. Text-to-video, image-to-video, and audio-to-video in a single endpoint.

prunaai-p-video استفاده ←
سفارشی 1,250 اعتبار

sprited: Birefnet Video

BiRefNet video background removal — every variant + ToonOut, applied per frame. Transparent WebM output by default (also MP4/MOV/GIF). MIT-licensed for commercial use.

sprited-birefnet-video استفاده ←
Alibaba Wan 1,400 اعتبار

andreasjansson: Wan 1.3B Inpaint

Inpainting and video2video experiments with Wan 2.1

andreasjansson-wan-1.3b-inpaint استفاده ←
سفارشی 1,400 اعتبار

lucataco: Sam3 Video

A unified foundation model for prompt-based segmentation in images and videos

lucataco-sam3-video استفاده ←
سفارشی 1,500 اعتبار

lucataco: Hotshot Xl

😊 Hotshot-XL is an AI text-to-GIF model trained to work alongside Stable Diffusion XL

lucataco-hotshot-xl استفاده ←
Runway 1,500 اعتبار

runwayml: Gen4 Image Turbo

Gen-4 Image Turbo is cheaper and 2.5x faster than Gen-4 Image. An image model with references, use up to 3 reference images to create the exact image you need. Capture every angle.

runwayml-gen4-image-turbo استفاده ←
Alibaba Wan 1,500 اعتبار

wan-video: Wan 2.7 Image

Generate and edit images with Alibaba's Wan 2.7

wan-video-wan-2.7-image استفاده ←
Alibaba Wan 1,500 اعتبار

wan-video: Wan 2.7 Image Pro

Generate and edit high-quality images with Alibaba's Wan 2.7 Pro with 4K output, thinking mode, text-to-image, multi-image editing, and image set generation

wan-video-wan-2.7-image-pro استفاده ←
سفارشی 1,700 اعتبار

bytedance: Sa2Va 4B Image

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-4b-image استفاده ←
سفارشی 1,800 اعتبار

fofr: Kontext Ps1

FLUX Kontext fine-tune that let's you restyle any image as a PS1 or PS2 video game

fofr-kontext-ps1 استفاده ←
سفارشی 1,800 اعتبار

uglyrobot: Sora2 Watermark Remover

Removes the watermark from Sora 2 videos using a trained model and IOpaint

uglyrobot-sora2-watermark-remover استفاده ←
سفارشی 1,850 اعتبار

meta: Sam 2 Video

SAM 2: Segment Anything v2 (for videos)

meta-sam-2-video استفاده ←
Black Forest Labs (FLUX) 2,000 اعتبار

flux-kontext-apps: Restyle Video Frame

Use flux-kontext-pro to change the first or last frame of a video. Useful to use as inputs for restyling an entire video in a certain way

flux-kontext-apps-restyle-video-frame استفاده ←
سفارشی 2,250 اعتبار

zsxkib: Bsrgan

Upscale videos + images with BSRGAN

zsxkib-bsrgan استفاده ←
سفارشی 2,300 اعتبار

arielreplicate: Robust Video Matting

extract foreground of a video

arielreplicate-robust-video-matting استفاده ←
سفارشی 2,350 اعتبار

bytedance: Sa2Va 8B Image

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-8b-image استفاده ←
سفارشی 2,500 اعتبار

cjwbw: Damo Text To Video

Multi-stage text-to-video generation

cjwbw-damo-text-to-video استفاده ←
Alibaba Wan 2,500 اعتبار

wan-video: Wan 2.2 I2V Fast

A very fast and cheap PrunaAI optimized version of Wan 2.2 A14B image-to-video

wan-video-wan-2.2-i2v-fast استفاده ←
Alibaba Wan 2,500 اعتبار

wan-video: Wan 2.2 T2V Fast

A very fast and cheap PrunaAI optimized version of Wan 2.2 A14B text-to-video

wan-video-wan-2.2-t2v-fast استفاده ←
سفارشی 2,600 اعتبار

arielreplicate: Deoldify Video

Add colours to old video footage.

arielreplicate-deoldify-video استفاده ←
سفارشی 2,850 اعتبار

hjunior29: Video Text Generator

Generate animated captions and text overlays for social media videos

hjunior29-video-text-generator استفاده ←
سفارشی 2,900 اعتبار

tencent: Hunyuanvideo Foley

Text-Video-to-Audio Synthesis: Generate realistic audio from video and text descriptions

tencent-hunyuanvideo-foley استفاده ←
سفارشی 2,950 اعتبار

pixverse: Pixverse V4.5

Quickly make 5s or 8s videos at 540p, 720p or 1080p. It has enhanced motion, prompt coherence and handles complex actions well.

pixverse-pixverse-v4.5 استفاده ←
سفارشی 3,050 اعتبار

bytedance: Sa2Va 26B Video

Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

bytedance-sa2va-26b-video استفاده ←
سفارشی 3,100 اعتبار

tmappdev: Change Video Bg

Change or Replace Video Background with any Image

tmappdev-change-video-bg استفاده ←
سفارشی 3,300 اعتبار

fofr: Tooncrafter

Create videos from illustrated input images

fofr-tooncrafter استفاده ←
سفارشی 3,400 اعتبار

lightricks: Ltx Video 0.9.7

DiT-based 13b video generation model, creating 30fps video

lightricks-ltx-video-0.9.7 استفاده ←
سفارشی 3,443 اعتبار

bytedance: Video Upscaler

Upscale and enhance video up to 4K at 60fps, with scene-aware presets for AI-generated content, short dramas, UGC, and film restoration.

bytedance-video-upscaler استفاده ←
Kling AI 3,500 اعتبار

kwaivgi: Kling Lip Sync

Add lip-sync to any video with an audio file or text

kwaivgi-kling-lip-sync استفاده ←
سفارشی 3,500 اعتبار

open-mmlab: Pia

Personalized Image Animator

open-mmlab-pia استفاده ←
Udio 3,500 اعتبار

zsxkib: Mmaudio T4

Cost-optimized MMAudio V2 (T4 GPU): Add sound to video using this version running on T4 hardware for lower cost. Synthesizes high-quality audio from video content.

zsxkib-mmaudio-t4 استفاده ←
ByteDance Seedance 3,750 اعتبار

bytedance: Seedance 1 Pro Fast

A faster and cheaper version of Seedance 1 Pro

bytedance-seedance-1-pro-fast استفاده ←
سفارشی 3,950 اعتبار

zsxkib: Animate Diff

🎨 AnimateDiff (w/ MotionLoRAs for Panning, Zooming, etc): Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

zsxkib-animate-diff استفاده ←
سفارشی 4,100 اعتبار

anotherjesse: Zeroscope V2 Xl

Zeroscope V2 XL & 576w

anotherjesse-zeroscope-v2-xl استفاده ←
سفارشی 4,150 اعتبار

lightricks: Ltx Video

LTX-Video is the first DiT-based video generation model capable of generating high-quality videos in real-time. It produces 24 FPS videos at a 768x512 resolution faster than they can be watched.

lightricks-ltx-video استفاده ←
ByteDance Seedance 4,500 اعتبار

bytedance: Seedance 1 Lite

A video generation model that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 720p resolution

bytedance-seedance-1-lite استفاده ←
سفارشی 5,000 اعتبار

lightricks: Ltx 2 Distilled

LTX-2: The first open source audio-video model

lightricks-ltx-2-distilled استفاده ←
سفارشی 5,000 اعتبار

lucataco: Animate Diff

Animate Your Personalized Text-to-Image Diffusion Models

lucataco-animate-diff استفاده ←
سفارشی 5,000 اعتبار

minimax: Hailuo 02

Hailuo 2 is a text-to-video and image-to-video model that can make 6s or 10s videos at 768p (standard) or 1080p (pro). It excels at real world physics.

minimax-hailuo-02 استفاده ←
سفارشی 5,000 اعتبار

minimax: Hailuo 02 Fast

A low cost and fast version of Hailuo 02. Generate 6s and 10s videos in 512p

minimax-hailuo-02-fast استفاده ←
سفارشی 5,000 اعتبار

nvidia: Nemotron Nano V2 12B Vl

A multi-modal AI model for visual Q&A, summarization, and data extraction, supporting text, images, and video.

nvidia-nemotron-nano-v2-12b-vl استفاده ←
Alibaba Wan 5,000 اعتبار

wan-video: Wan 2.2 Animate Replace

Use Wan 2.2 Animate to replace a character in a video scene

wan-video-wan-2.2-animate-replace استفاده ←
Alibaba Wan 5,000 اعتبار

wan-video: Wan 2.2 S2V

Generate a video from an audio clip and a reference image

wan-video-wan-2.2-s2v استفاده ←
سفارشی 5,500 اعتبار

ayushunleashed: Minimax Remover

Remove any object from video - fast.

ayushunleashed-minimax-remover استفاده ←
سفارشی 5,500 اعتبار

zsxkib: Seedvr2

🔥 SeedVR2: one-step video & image restoration with 3B/7B hot‑swap and optional color fix 🎬✨

zsxkib-seedvr2 استفاده ←
سفارشی 6,000 اعتبار

andreasjansson: Tile Morph

Create tileable animations with seamless transitions

andreasjansson-tile-morph استفاده ←
سفارشی 6,000 اعتبار

fofr: Video Morpher

Generate a video that morphs between subjects, with an optional style

fofr-video-morpher استفاده ←
سفارشی 6,000 اعتبار

zsxkib: Animatediff Illusions

Monster Labs' Controlnet QR Code Monster v2 For SD-1.5 on top of AnimateDiff Prompt Travel (Motion Module SD 1.5 v2)

zsxkib-animatediff-illusions استفاده ←
ByteDance Seedance 6,250 اعتبار

bytedance: Seedance 1.5 Pro

A joint audio-video model that accurately follows complex instructions.

bytedance-seedance-1.5-pro استفاده ←
سفارشی 6,250 اعتبار

prunaai: P Video Avatar

p-video-avatar is the fastest and cheapest avatar/lipsync video model on the market.

prunaai-p-video-avatar استفاده ←
سفارشی 6,500 اعتبار

arielreplicate: Stable Diffusion Infinite Zoom

Use Runway's Stable-diffusion inpainting model to create an infinite loop video

arielreplicate-stable-diffusion-infinite-zoom استفاده ←
سفارشی 6,500 اعتبار

cjwbw: Videocrafter

VideoCrafter2: Text-to-Video and Image-to-Video Generation and Editing

cjwbw-videocrafter استفاده ←
سفارشی 6,500 اعتبار

lightricks: Ltx Video 0.9.7 Distilled

Faster slight quality reduction compared to LTX-Video 13b

lightricks-ltx-video-0.9.7-distilled استفاده ←
سفارشی 7,000 اعتبار

cjwbw: Text2Video Zero

Text-to-Image Diffusion Models are Zero-Shot Video Generators

cjwbw-text2video-zero استفاده ←
ByteDance Seedance 7,500 اعتبار

bytedance: Seedance 1 Pro

A pro version of Seedance that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 1080p resolution

bytedance-seedance-1-pro استفاده ←
سفارشی 7,500 اعتبار

lightricks: Ltx 2.5 Fast

Fast video generation with text-to-video and image-to-video, portrait and landscape support, synchronized audio, and frame interpolation. Up to 20 seconds at 1080p, and 4K resolution.

lightricks-ltx-2.5-fast استفاده ←
Luma Dream Machine 7,500 اعتبار

luma: Ray 3.2

Luma's reasoning video model. Generates cinematic 5s or 10s video from text or images, with native HDR and EXR export for professional production pipelines.

luma-ray-3.2 استفاده ←
سفارشی 7,500 اعتبار

prunaai: P Video Animate

p-video-animate animates a reference image with the motion and audio of a source video. Optimized for speed and cost — 5.24s per 1s of video.

prunaai-p-video-animate استفاده ←
سفارشی 7,500 اعتبار

prunaai: P Video Replace

p-video-replace swaps the person in a video with one from a reference image, keeping motion, timing, camera, and scene exactly as they were. 3.58s per 1s of video generated.

prunaai-p-video-replace استفاده ←
Luma Dream Machine 8,250 اعتبار

luma: Ray Flash 2 540P

Generate 5s and 9s 540p videos, faster and cheaper than Ray 2

luma-ray-flash-2-540p استفاده ←
سفارشی 8,325 اعتبار

heygen: Lipsync Speed

Fast lip-sync: replace or dub audio on any video with quick audio-driven lip sync

heygen-lipsync-speed استفاده ←
سفارشی 8,325 اعتبار

heygen: Video Agent

Turn a text prompt into a complete, polished video with AI-generated script, avatar presenter, voiceover, visuals, and editing.

heygen-video-agent استفاده ←
سفارشی 8,325 اعتبار

heygen: Video Translate

Translate videos into over 150 languages

heygen-video-translate استفاده ←
Alibaba Wan 8,750 اعتبار

alibaba: Wan 3

Wan 3.0 generates video from a text prompt, with cinematic motion and support for 480p, 720p, and 1080p output.

alibaba-wan-3 استفاده ←
سفارشی 9,500 اعتبار

bitflow: Video Super Resolution Rife Pro

Super video quality enhancement featuring fast upscaling with TensorRT and frame interpolation with RIFE.

bitflow-video-super-resolution-rife-pro استفاده ←
ByteDance Seedance 10,000 اعتبار

bytedance: Seedance 2.0 Mini

A lower-cost variant of Seedance 2.0 for high-volume video generation with multimodal inputs and native audio.

bytedance-seedance-2.0-mini استفاده ←
سفارشی 10,000 اعتبار

character-ai: Ovi I2V

Ovi: generate videos with audio from image and text inputs

character-ai-ovi-i2v استفاده ←
سفارشی 10,000 اعتبار

decart: Lucy Edit 2

Edit and transform videos with text prompts and reference images. Style transfers, object replacement, character transformation, and more.

decart-lucy-edit-2 استفاده ←
سفارشی 10,000 اعتبار

lightricks: Ltx 2 Fast

Ideal for rapid ideation and mobile workflows. Perfect for creators who need instant feedback, real-time previews, or high-throughput content.

lightricks-ltx-2-fast استفاده ←
Alibaba Wan 10,000 اعتبار

lucataco: Wan 2.1 1.3B Vid2Vid

Wan 2.1 1.3b Video to Video. Wan is a powerful visual generation model developed by Tongyi Lab of Alibaba Group

lucataco-wan-2.1-1.3b-vid2vid استفاده ←
سفارشی 10,000 اعتبار

pixverse: Lipsync

Generate realistic lipsync animations from audio for high-quality synchronization

pixverse-lipsync استفاده ←
سفارشی 10,000 اعتبار

vidu: Q3 Turbo

Fast video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.

vidu-q3-turbo استفاده ←
Alibaba Wan 10,000 اعتبار

wan-video: Wan 2.1 1.3B

Generate 5s 480p videos. Wan is an advanced and powerful visual generation model developed by Tongyi Lab of Alibaba Group

wan-video-wan-2.1-1.3b استفاده ←
سفارشی 11,500 اعتبار

ali-vilab: I2Vgen Xl

RESEARCH/NON-COMMERCIAL USE ONLY: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

ali-vilab-i2vgen-xl استفاده ←
سفارشی 12,000 اعتبار

andreasjansson: Stable Diffusion Animation

Animate Stable Diffusion by interpolating between two prompts

andreasjansson-stable-diffusion-animation استفاده ←
سفارشی 12,500 اعتبار

bria: Video Erase Object

A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency

bria-video-erase-object استفاده ←
سفارشی 12,500 اعتبار

bytedance: Dreamactor M2.0

Animate any character, humans, cartoons, animals, even non-humans, from a single image + driving video

bytedance-dreamactor-m2.0 استفاده ←
Google (Veo / Nano Banana / Lyria) 12,500 اعتبار

google: Veo 3.1 Lite

Google's cost-efficient video generation model with native audio, optimized for high-volume applications

google-veo-3.1-lite استفاده ←
Kling AI 12,500 اعتبار

kwaivgi: Kling V1.5 Standard

Generate 5s and 10s videos in 720p resolution at 30fps

kwaivgi-kling-v1.5-standard استفاده ←
Kling AI 12,500 اعتبار

kwaivgi: Kling V1.6 Standard

Generate 5s and 10s videos in 720p resolution at 30fps

kwaivgi-kling-v1.6-standard استفاده ←
Kling AI 12,500 اعتبار

kwaivgi: Kling V2.1

Use Kling v2.1 to generate 5s and 10s videos in 720p and 1080p resolution from a starting image (image-to-video)

kwaivgi-kling-v2.1 استفاده ←
سفارشی 12,500 اعتبار

pixverse: Pixverse V6

PixVerse's flagship video generation model. Generate cinematic videos with synchronized audio, multi-shot sequences, and precise camera control.

pixverse-pixverse-v6 استفاده ←
سفارشی 12,500 اعتبار

pollinations: Real Basicvsr Video Superresolution

RealBasicVSR: Investigating Tradeoffs in Real-World Video Super-Resolution

pollinations-real-basicvsr-video-superresolution استفاده ←
Runway 12,500 اعتبار

runwayml: Gen4 Turbo

Generate 5s and 10s 720p videos fast

runwayml-gen4-turbo استفاده ←
Alibaba Wan 12,500 اعتبار

wan-video: Wan 2.5 I2V

Alibaba Wan 2.5 Image to video generation with background audio

wan-video-wan-2.5-i2v استفاده ←
Alibaba Wan 12,500 اعتبار

wan-video: Wan 2.5 T2V

Alibaba Wan 2.5 text to video generation model

wan-video-wan-2.5-t2v استفاده ←
Alibaba Wan 12,500 اعتبار

wan-video: Wan2.6 I2V Flash

Image-to-video generation with optional audio, multi-shot narrative support, and faster inference

wan-video-wan2.6-i2v-flash استفاده ←
سفارشی 12,500 اعتبار

xai: Grok Imagine R2V

Generate videos guided by reference images using xAI's Grok Imagine Video model

xai-grok-imagine-r2v استفاده ←
سفارشی 12,500 اعتبار

xai: Grok Imagine Video

Generate videos using xAI's Grok Imagine Video model

xai-grok-imagine-video استفاده ←
سفارشی 12,500 اعتبار

xai: Grok Imagine Video Extension

Extend videos with xAI's Grok Imagine Video model. Provide a source video and describe what happens next.

xai-grok-imagine-video-extension استفاده ←
سفارشی 13,000 اعتبار

zsxkib: Hunyuan Video Lora

Hunyuan-Video LoRA Explorer + Trainer

zsxkib-hunyuan-video-lora استفاده ←
Kling AI 14,000 اعتبار

kwaivgi: Kling Avatar V2

Create avatar videos with realistic humans, animals, cartoons, or stylized characters

kwaivgi-kling-avatar-v2 استفاده ←
سفارشی 14,000 اعتبار

minimax: Hailuo 2.3

A high-fidelity video generation model optimized for realistic human motion, cinematic VFX, expressive characters, and strong prompt and style adherence across both text-to-video and image-to-video workflows

minimax-hailuo-2.3 استفاده ←
سفارشی 14,500 اعتبار

lucataco: Ltx Video 0.9.8 Distilled

Generate native long-form video, with controllability

lucataco-ltx-video-0.9.8-distilled استفاده ←
ElevenLabs 14,667 اعتبار

elevenlabs: Dubbing

Translate audio and video into 90+ languages while preserving each speaker's voice, emotion, and timing

elevenlabs-dubbing استفاده ←
Black Forest Labs (FLUX) 15,000 اعتبار

black-forest-labs: Flux 3

Generate video with synchronized audio from text, images, or video. FLUX 3 is Black Forest Labs' multimodal model (early access preview).

black-forest-labs-flux-3 استفاده ←
سفارشی 15,000 اعتبار

leonardoai: Motion 2.0

Create 5s 480p videos from a text prompt

leonardoai-motion-2.0 استفاده ←
سفارشی 15,000 اعتبار

lightricks: Ltx 2 Pro

Delivers high visual fidelity with fast turnaround. Great for daily content creation, marketing teams, and iterative creative workflows.

lightricks-ltx-2-pro استفاده ←
سفارشی 15,000 اعتبار

lightricks: Ltx 2.3 Fast

Lightning-fast video generation with portrait support, camera controls, and synchronized audio. Up to 20 seconds at 1080p, 4K at 50 FPS.

lightricks-ltx-2.3-fast استفاده ←
Luma Dream Machine 15,000 اعتبار

luma: Ray Flash 2 720P

Generate 5s and 9s 720p videos, faster and cheaper than Ray 2

luma-ray-flash-2-720p استفاده ←
Luma Dream Machine 15,000 اعتبار

luma: Reframe Video

Change the aspect ratio of any video up to 30 seconds long, outputs will be 720p

luma-reframe-video استفاده ←
سفارشی 16,500 اعتبار

deforum: Deforum Stable Diffusion

Animating prompts with stable diffusion

deforum-deforum-stable-diffusion استفاده ←
سفارشی 16,675 اعتبار

heygen: Lipsync Precision

High-accuracy lip-sync: replace or dub audio on any video with avatar-inference lip sync

heygen-lipsync-precision استفاده ←
Alibaba Wan 17,000 اعتبار

alibaba: Wan 3 Prime

Generate videos from text prompts using Alibaba's Wan 3.0 Prime model. Up to 1080p and 30 seconds, with 480p, 720p, and 1080p output.

alibaba-wan-3-prime استفاده ←
Alibaba Wan 17,000 اعتبار

wan-video: Wan 2.5 I2V Fast

Wan 2.5 image-to-video, optimized for speed

wan-video-wan-2.5-i2v-fast استفاده ←
Alibaba Wan 17,000 اعتبار

wan-video: Wan 2.5 T2V Fast

Wan 2.5 text-to-video, optimized for speed

wan-video-wan-2.5-t2v-fast استفاده ←
ByteDance Seedance 17,500 اعتبار

bytedance: Seedance 2.0 Fast

A faster variant of Seedance 2.0 for quicker video generation with multimodal inputs and native audio.

bytedance-seedance-2.0-fast استفاده ←
Kling AI 17,500 اعتبار

kwaivgi: Kling V2.5 Turbo Pro

Kling 2.5 Turbo Pro: Unlock pro-level text-to-video and image-to-video creation with smooth motion, cinematic depth, and remarkable prompt adherence.

kwaivgi-kling-v2.5-turbo-pro استفاده ←
Kling AI 17,500 اعتبار

kwaivgi: Kling V2.6

Kling 2.6 Pro: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation

kwaivgi-kling-v2.6 استفاده ←
Kling AI 17,500 اعتبار

kwaivgi: Kling V2.6 Motion Control

Enables precise control of character actions and expressions from a reference image.

kwaivgi-kling-v2.6-motion-control استفاده ←
Kling AI 17,500 اعتبار

kwaivgi: Kling V3 Motion Control

Kling 3.0 motion control: transfer motion from a reference video to any character image with improved consistency and quality.

kwaivgi-kling-v3-motion-control استفاده ←
سفارشی 17,500 اعتبار

pixverse: Pixverse V5.6

Latest video model from Pixverse with astonishing physics

pixverse-pixverse-v5.6 استفاده ←
سفارشی 17,500 اعتبار

vidu: Q3 Pro

High-fidelity video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.

vidu-q3-pro استفاده ←
Alibaba Wan 17,500 اعتبار

wavespeedai: Wan 2.1 T2V 480P

Accelerated inference for Wan 2.1 14B text to video, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.

wavespeedai-wan-2.1-t2v-480p استفاده ←
سفارشی 18,750 اعتبار

heygen: Avatar Iv

Create realistic talking avatar videos from text with HeyGen's Avatar IV engine

heygen-avatar-iv استفاده ←
سفارشی 18,750 اعتبار

heygen: Avatar V

Create realistic talking avatar videos from text with HeyGen's Avatar V engine — the newest, highest-quality avatar engine with cross-reference-driven animation.

heygen-avatar-v استفاده ←
ByteDance Seedance 20,000 اعتبار

bytedance: Seedance 2.0

ByteDance's multimodal video generation model with native audio, multimodal reference inputs, and intelligent duration control.

bytedance-seedance-2.0 استفاده ←
سفارشی 20,000 اعتبار

lightricks: Ltx 2.3 Pro

High-fidelity video generation with portrait support, audio-to-video, retake, and extend. Text, image, and audio-driven creation up to 4K at 50 FPS.

lightricks-ltx-2.3-pro استفاده ←
سفارشی 20,000 اعتبار

veed: Fabric 1.0

VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video

veed-fabric-1.0 استفاده ←
Alibaba Wan 20,000 اعتبار

wan-video: Wan 2.2 I2V A14B

Image-to-video at 720p and 480p with Wan 2.2 A14B

wan-video-wan-2.2-i2v-a14b استفاده ←
سفارشی 20,000 اعتبار

xai: Grok Imagine Video 1.5

Image-to-video with synchronized audio using xAI's Grok Imagine Video 1.5 preview model

xai-grok-imagine-video-1.5 استفاده ←
سفارشی 21,000 اعتبار

genmoai: Mochi 1

Mochi 1 preview is an open video generation model with high-fidelity motion and strong prompt adherence in preliminary evaluation

genmoai-mochi-1 استفاده ←
سفارشی 22,000 اعتبار

zsxkib: Pyramid Flow

Text-to-Video + Image-to-Video: Pyramid Flow Autoregressive Video Generation method based on Flow Matching

zsxkib-pyramid-flow استفاده ←
Alibaba Wan 22,500 اعتبار

wavespeedai: Wan 2.1 I2V 480P

Accelerated inference for Wan 2.1 14B image to video, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.

wavespeedai-wan-2.1-i2v-480p استفاده ←
Alibaba Wan 23,500 اعتبار

lucataco: Wan2.1 I2V Lora

Wan2.1 14B 480p LoRA inference via Diffusers (Work in progress)

lucataco-wan2.1-i2v-lora استفاده ←
Kling AI 23,750 اعتبار

kwaivgi: Kling V1.5 Pro

Generate 5s and 10s videos in 1080p resolution at 30fps

kwaivgi-kling-v1.5-pro استفاده ←
Kling AI 23,750 اعتبار

kwaivgi: Kling V1.6 Pro

Generate 5s and 10s videos in 1080p resolution

kwaivgi-kling-v1.6-pro استفاده ←
Alibaba Wan 24,500 اعتبار

wan-video: Wan 2.2 Animate Animation

Use Wan 2.2 Animate to copy the motion of a video to another scene

wan-video-wan-2.2-animate-animation استفاده ←
سفارشی 25,000 اعتبار

cuuupid: Cogvideox 5B

Generate high quality videos from a prompt

cuuupid-cogvideox-5b استفاده ←
Udio 25,000 اعتبار

lightricks: Audio To Video

Use audio input with an image or prompt to generate videos

lightricks-audio-to-video استفاده ←
سفارشی 25,000 اعتبار

lightricks: Ltx 2 Retake

Take any shot and edit specific sections. Rephrase, change the action, camera angles and more

lightricks-ltx-2-retake استفاده ←
Luma Dream Machine 25,000 اعتبار

luma: Ray 2 540P

Generate 5s and 9s 540p videos

luma-ray-2-540p استفاده ←
سفارشی 25,000 اعتبار

minimax: Video 01

Generate 6s videos with prompts or images. (Also known as Hailuo). Use a subject reference to make a video with a character and the S2V-01 model.

minimax-video-01 استفاده ←
سفارشی 25,000 اعتبار

minimax: Video 01 Director

Generate videos with specific camera movements

minimax-video-01-director استفاده ←
سفارشی 25,000 اعتبار

minimax: Video 01 Live

An image-to-video (I2V) model specifically trained for Live2D and general animation use cases

minimax-video-01-live استفاده ←
OpenAI 25,000 اعتبار

openai: Sora 2

OpenAI's Flagship video generation with synced audio

openai-sora-2 استفاده ←
سفارشی 25,000 اعتبار

philz1337x: Crystal Video Upscaler

High-precision video upscaler optimized for portraits, faces and products. One of the upscale modes powered by Clarity AI. X:https://x.com/philz1337x

philz1337x-crystal-video-upscaler استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.6 I2V

Alibaba Wan 2.6 image to video generation model

wan-video-wan-2.6-i2v استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.6 T2V

Alibaba Wan 2.6 text to video generation model

wan-video-wan-2.6-t2v استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.7 I2V

Generate videos from images, with support for first-and-last-frame control, clip continuation, and audio synchronization using Alibaba's Wan 2.7 model

wan-video-wan-2.7-i2v استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.7 R2V

Generate videos from reference images or clips while preserving subject identity using Alibaba's Wan 2.7 reference-to-video model

wan-video-wan-2.7-r2v استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.7 T2V

Generate videos with audio from text prompts using Alibaba's Wan 2.7 model. 1080p, up to 15 seconds, with audio synchronization.

wan-video-wan-2.7-t2v استفاده ←
Alibaba Wan 25,000 اعتبار

wan-video: Wan 2.7 Videoedit

Edit videos with natural language instructions using Alibaba's Wan 2.7 VideoEdit model

wan-video-wan-2.7-videoedit استفاده ←
سفارشی 26,500 اعتبار

zsxkib: Framepack

🕹️FramePack: video diffusion that feels like image diffusion🎥

zsxkib-framepack استفاده ←
سفارشی 28,000 اعتبار

lucataco: Real Esrgan Video

Real-ESRGAN Video Upscaler

lucataco-real-esrgan-video استفاده ←
سفارشی 29,500 اعتبار

zsxkib: Multitalk

Audio-driven multi-person conversational video generation - Upload audio files and a reference image to create realistic conversations between multiple people

zsxkib-multitalk استفاده ←
سفارشی 31,500 اعتبار

deepfates: Hunyuan Spiderverse

Hunyuan-Video model finetuned on Spider-Man: Into the Spider-Verse (2018). Trigger word is "SPDRV". Use "A video in the style of SPDRV, SPDRV" at the beginning of your prompt for best results.

deepfates-hunyuan-spiderverse استفاده ←
سفارشی 32,000 اعتبار

deepfates: Hunyuan The Matrix Trilogy

Hunyuan-Video model finetuned on The Matrix Trilogy (1999). Trigger word is "THMTR". Use "A video in the style of THMTR, THMTR" at the beginning of your prompt for best results.

deepfates-hunyuan-the-matrix-trilogy استفاده ←
سفارشی 32,500 اعتبار

zsxkib: Hunyuan Video2Video

A state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions

zsxkib-hunyuan-video2video استفاده ←
سفارشی 35,000 اعتبار

alibaba: Happyhorse 1.0

Alibaba's Happy Horse 1.0 generates videos from text prompts or animates a single image into video. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.

alibaba-happyhorse-1.0 استفاده ←
سفارشی 35,000 اعتبار

alibaba: Happyhorse 1.1

Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.

alibaba-happyhorse-1.1 استفاده ←
سفارشی 35,000 اعتبار

bytedance: Omni Human

Turns your audio/video/images into professional-quality animated videos

bytedance-omni-human استفاده ←
Google (Veo / Nano Banana / Lyria) 37,500 اعتبار

google: Veo 3 Fast

A faster and cheaper version of Google’s Veo 3 video model, with audio

google-veo-3-fast استفاده ←
Google (Veo / Nano Banana / Lyria) 37,500 اعتبار

google: Veo 3.1 Fast

New and improved version of Veo 3 Fast, with higher-fidelity video, context-aware audio and last frame support

google-veo-3.1-fast استفاده ←
سفارشی 40,000 اعتبار

bytedance: Omni Human 1.5

A film-grade digital human model that generates realistic video from a single image, audio clip, and optional text prompt.

bytedance-omni-human-1.5 استفاده ←
Kling AI 42,000 اعتبار

kwaivgi: Kling V3 Omni Video

Kling Video 3.0 Omni: Unified multimodal video generation with reference images, video editing, native audio, and multi-shot control

kwaivgi-kling-v3-omni-video استفاده ←
Kling AI 42,000 اعتبار

kwaivgi: Kling V3 Video

Kling Video 3.0: Generate cinematic videos up to 15 seconds with multi-shot control, native audio, and improved consistency

kwaivgi-kling-v3-video استفاده ←
Luma Dream Machine 45,000 اعتبار

luma: Ray 2 720P

Generate 5s and 9s 720p videos

luma-ray-2-720p استفاده ←
سفارشی 50,000 اعتبار

wavespeedai: Hunyuan Video Fast

Accelerated inference for HunyuanVideo with high resolution (1280x720), a state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions

wavespeedai-hunyuan-video-fast استفاده ←
ByteDance Seedance 57,800 اعتبار

bytedance: Seedance 2.5

ByteDance's flagship multimodal video model with native audio, native 30-second generation, and large multimodal reference sets.

bytedance-seedance-2.5 استفاده ←
Alibaba Wan 60,000 اعتبار

wavespeedai: Wan 2.1 T2V 720P

Accelerated inference for Wan 2.1 14B text to video with high resolution, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.

wavespeedai-wan-2.1-t2v-720p استفاده ←
سفارشی 61,000 اعتبار

deepfates: Hunyuan Arcane

Hunyuan-Video model finetuned on Arcane (2021). Trigger word is "RCN". Use "A video in the style of RCN, RCN" at the beginning of your prompt for best results.

deepfates-hunyuan-arcane استفاده ←
Alibaba Wan 62,500 اعتبار

wavespeedai: Wan 2.1 I2V 720P

Accelerated inference for Wan 2.1 14B image to video with high resolution, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.

wavespeedai-wan-2.1-i2v-720p استفاده ←
Alibaba Wan 65,000 اعتبار

wan-video: Wan2.1 With Lora

Run Wan2.1 14b or 1.3b with a lora

wan-video-wan2.1-with-lora استفاده ←
سفارشی 68,000 اعتبار

deepfates: Hunyuan Dune

Hunyuan-Video model finetuned on Dune (2021). Trigger word is "DN". Use "A video in the style of DN, DN" at the beginning of your prompt for best results.

deepfates-hunyuan-dune استفاده ←
Kling AI 70,000 اعتبار

kwaivgi: Kling V2.0

Generate 5s and 10s videos in 720p resolution

kwaivgi-kling-v2.0 استفاده ←
Kling AI 70,000 اعتبار

kwaivgi: Kling V2.1 Master

A premium version of Kling v2.1 with superb dynamics and prompt adherence. Generate 1080p 5s and 10s videos from text or an image

kwaivgi-kling-v2.1-master استفاده ←
سفارشی 74,000 اعتبار

zsxkib: Star

STAR Video Upscaler: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-Resolution

zsxkib-star استفاده ←
OpenAI 75,000 اعتبار

openai: Sora 2 Pro

OpenAI's Most advanced synced-audio video generation

openai-sora-2-pro استفاده ←
ByteDance Seedance 78,750 اعتبار

Bytedance Seedance 2.5

KIE

bytedance-seedance-2-5 استفاده ←
سفارشی 83,000 اعتبار

zsxkib: Animatediff Prompt Travel

🎨AnimateDiff Prompt Travel🧭 Seamlessly Navigate and Animate Between Text-to-Image Prompts for Dynamic Visual Narratives

zsxkib-animatediff-prompt-travel استفاده ←
Runway 84,000 اعتبار

runwayml: Aleph 2

Edit one frame to update an entire video. Aleph 2.0 is Runway's in-context video editor: longer clips (up to 30s), multi-shot edits, and image-level precision via keyframe references.

runwayml-aleph-2 استفاده ←
Google (Veo / Nano Banana / Lyria) 100,000 اعتبار

google: Veo 3

Sound on: Google’s flagship Veo 3 text to video model, with audio

google-veo-3 استفاده ←
Google (Veo / Nano Banana / Lyria) 100,000 اعتبار

google: Veo 3.1

New and improved version of Veo 3, with higher-fidelity video, context-aware audio, reference image and last frame support

google-veo-3.1 استفاده ←
سفارشی 110,000 اعتبار

deepfates: Hunyuan Cowboy Bebop

Hunyuan-Video model finetuned on Cowboy Bebop (1998). Trigger word is "CWBYB". Use "A video in the style of CWBYB, CWBYB" at the beginning of your prompt for best results.

deepfates-hunyuan-cowboy-bebop استفاده ←
سفارشی 120,000 اعتبار

deepfates: Hunyuan Pixar

Hunyuan-Video model finetuned on Pixar Films (1995). Trigger word is "PXR". Use "A video in the style of PXR, PXR" at the beginning of your prompt for best results.

deepfates-hunyuan-pixar استفاده ←
Google (Veo / Nano Banana / Lyria) 125,000 اعتبار

google: Veo 2

State of the art video generation model. Veo 2 can faithfully follow simple and complex instructions, and convincingly simulates real-world physics as well as a wide range of visual styles.

google-veo-2 استفاده ←
سفارشی 127,500 اعتبار

tencent: Hunyuan Video

A state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions

tencent-hunyuan-video استفاده ←
سفارشی 130,500 اعتبار

zsxkib: Step Video T2V

Generate high-quality videos from text prompts using StepVideo

zsxkib-step-video-t2v استفاده ←
سفارشی 392,000 اعتبار

deepfates: Hunyuan The Grand Budapest Hotel

Hunyuan-Video model finetuned on The Grand Budapest Hotel (2014). Trigger word is "THGRN". Use "A video in the style of THGRN, THGRN" at the beginning of your prompt for best results.

deepfates-hunyuan-the-grand-budapest-hotel استفاده ←
سفارشی 403,500 اعتبار

deepfates: Hunyuan Spider Man Into The Spider Verse

Hunyuan-Video model finetuned on Spider-Man: Into the Spider-Verse (2018). Trigger word is "SPDRM". Use "A video in the style of SPDRM, SPDRM" at the beginning of your prompt for best results.

deepfates-hunyuan-spider-man-into-the-spider-verse استفاده ←
سفارشی 439,500 اعتبار

deepfates: Hunyuan La La Land

Hunyuan-Video model finetuned on La La Land (2016). Trigger word is "LLLND". Use "A video in the style of LLLND, LLLND" at the beginning of your prompt for best results.

deepfates-hunyuan-la-la-land استفاده ←
سفارشی 452,000 اعتبار

deepfates: Hunyuan Twin Peaks

Hunyuan-Video model finetuned on Twin Peaks (1990). Trigger word is "TWNPK". Use "A video in the style of TWNPK, TWNPK" at the beginning of your prompt for best results.

deepfates-hunyuan-twin-peaks استفاده ←
سفارشی 476,500 اعتبار

deepfates: Hunyuan Pulp Fiction

Hunyuan-Video model finetuned on Pulp Fiction (1994). Trigger word is "PLPFC". Use "A video in the style of PLPFC, PLPFC" at the beginning of your prompt for best results.

deepfates-hunyuan-pulp-fiction استفاده ←

موسیقی

آهنگ‌سازی و تولید موسیقی هوشمند

10 مدل فعال

شروع رایگان
Google (Veo / Nano Banana / Lyria) 800 اعتبار

google: Lyria 2

Lyria 2 is a music generation model that produces 48kHz stereo audio through text-based prompts

google-lyria-2 استفاده ←
سفارشی 1,500 اعتبار

minimax: Music 1.5

Music-1.5: Full-length songs (up to 4 mins) with natural vocals & rich instrumentation

minimax-music-1.5 استفاده ←
سفارشی 1,750 اعتبار

minimax: Music 01

Quickly generate up to 1 minute of music with lyrics and vocals in the style of a reference track

minimax-music-01 استفاده ←
Google (Veo / Nano Banana / Lyria) 2,000 اعتبار

google: Lyria 3

Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model

google-lyria-3 استفاده ←
ElevenLabs 3,320 اعتبار

elevenlabs: Music

Compose a song from a prompt or a composition plan

elevenlabs-music استفاده ←
Google (Veo / Nano Banana / Lyria) 4,000 اعتبار

google: Lyria 3 Pro

Generate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model

google-lyria-3-pro استفاده ←
سفارشی 7,500 اعتبار

minimax: Music 2.5

Generate full-length songs with vocals, lyrics, and rich instrumentation from a text prompt

minimax-music-2.5 استفاده ←
سفارشی 7,500 اعتبار

minimax: Music 2.6

Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics

minimax-music-2.6 استفاده ←
Udio 10,000 اعتبار

stability-ai: Stable Audio 2.5

Generate high-quality music and sound from text prompts

stability-ai-stable-audio-2.5 استفاده ←
سفارشی 15,000 اعتبار

minimax: Music Cover

Reimagine any song in a different style — change voice, instruments, genre, and arrangement while keeping the original melody

minimax-music-cover استفاده ←

صدا و گفتار

تبدیل متن به گفتار و رونویسی صدا

25 مدل فعال

شروع رایگان
سفارشی 3 اعتبار

ibm-granite: Granite Speech 4.1 2B

Granite Speech 4.1 2B is a compact and efficient speech-language model, specifically designed for multilingual automatic speech recognition (ASR) and bidirectional automatic speech translation (AST) for English, French, German, Spanish, Portuguese and Jap

ibm-granite-granite-speech-4.1-2b استفاده ←
ElevenLabs 5 اعتبار

elevenlabs: Flash V2.5

ElevenLabs's fastest speech synthesis model

elevenlabs-flash-v2.5 استفاده ←
ElevenLabs 5 اعتبار

elevenlabs: V3

The most expressive Text to Speech model

elevenlabs-v3 استفاده ←
سفارشی 5 اعتبار

inworld: Realtime Tts 1.5 Mini

Ultra-fast, cost-efficient realtime text-to-speech with ~120ms latency and 15-language support

inworld-realtime-tts-1.5-mini استفاده ←
سفارشی 5 اعتبار

inworld: Realtime Tts 2

Most expressive text-to-speech model from Inworld, with natural-language steering, real-time latency, and multilingual support across 100+ languages.

inworld-realtime-tts-2 استفاده ←
سفارشی 6 اعتبار

inworld: Realtime Tts 1.5 Max

Highest-quality realtime text-to-speech with <200ms latency, emotion control, and 15-language support

inworld-realtime-tts-1.5-max استفاده ←
ElevenLabs 8 اعتبار

elevenlabs: Turbo V2.5

High quality, low latency text to speech in 32 languages

elevenlabs-turbo-v2.5 استفاده ←
ElevenLabs 12 اعتبار

elevenlabs: Scribe V2

Transcribe speech with ElevenLabs Scribe v2. 90+ languages, word-level timestamps, speaker diarization for up to 32 speakers, audio event tagging, and keyterm biasing. Files up to 3 GB and 10 hours.

elevenlabs-scribe-v2 استفاده ←
سفارشی 13 اعتبار

resemble-ai: Chatterbox Pro

Generate expressive, natural speech with Resemble AI's Chatterbox.

resemble-ai-chatterbox-pro استفاده ←
ElevenLabs 17 اعتبار

elevenlabs: V2 Multilingual

Generate multilingual text-to-speech audio in over 30 languages

elevenlabs-v2-multilingual استفاده ←
سفارشی 21 اعتبار

xai: Grok Text To Speech

Convert text to natural-sounding speech with xAI's Grok TTS. 5 voices, 20 languages, expressive speech tags, and high-fidelity MP3 / WAV / telephony audio output.

xai-grok-text-to-speech استفاده ←
سفارشی 49 اعتبار

ibm-granite: Granite Speech 3.3 8B

Granite-speech-3.3-8b is a compact and efficient speech-language model, specifically designed for automatic speech recognition (ASR) and automatic speech translation (AST).

ibm-granite-granite-speech-3.3-8b استفاده ←
سفارشی 55 اعتبار

datalab-to: Ocr

Detect and transcribe text in images with accurate bounding boxes, layout analysis, reding order, and table recognition, in 90 languages

datalab-to-ocr استفاده ←
OpenAI 62 اعتبار

openai: Gpt 4O Mini Transcribe

A speech-to-text model that uses GPT-4o mini to transcribe audio

openai-gpt-4o-mini-transcribe استفاده ←
سفارشی 83 اعتبار

xai: Grok Speech To Text

Transcribe audio to text with xAI's Grok. Handles 25 languages, word-level timestamps, speaker diarization, multichannel audio, and files up to 500 MB.

xai-grok-speech-to-text استفاده ←
سفارشی 85 اعتبار

resemble-ai: Chatterbox Turbo

The fastest open source TTS model without sacrificing quality.

resemble-ai-chatterbox-turbo استفاده ←
OpenAI 124 اعتبار

openai: Gpt 4O Transcribe

A speech-to-text model that uses GPT-4o to transcribe audio

openai-gpt-4o-transcribe استفاده ←
سفارشی 240 اعتبار

minimax: Speech 02 Turbo

Text-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Designed for real-time applications with low latency

minimax-speech-02-turbo استفاده ←
سفارشی 240 اعتبار

minimax: Speech 2.6 Turbo

Low‑latency MiniMax Speech 2.6 Turbo brings multilingual, emotional text-to-speech to Replicate with 300+ voices and real-time friendly pricing

minimax-speech-2.6-turbo استفاده ←
سفارشی 240 اعتبار

minimax: Speech 2.8 Turbo

Minimax Speech 2.8 Turbo: Turn text into natural, expressive speech with voice cloning, emotion control, and support for 40+ languages

minimax-speech-2.8-turbo استفاده ←
سفارشی 400 اعتبار

minimax: Speech 02 Hd

Text-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Optimized for high-fidelity applications like voiceovers and audiobooks.

minimax-speech-02-hd استفاده ←
سفارشی 400 اعتبار

minimax: Speech 2.6 Hd

MiniMax Speech 2.6 HD delivers studio-quality multilingual text-to-audio on Replicate with nuanced prosody, subtitle export, and premium voices

minimax-speech-2.6-hd استفاده ←
سفارشی 400 اعتبار

minimax: Speech 2.8 Hd

Minimax Speech 2.8 HD focuses on high-fidelity audio generation with features like studio-grade quality, flexible emotion control, multilingual support, and voice cloning capabilities

minimax-speech-2.8-hd استفاده ←
Google Gemini 408 اعتبار

google: Gemini 3.1 Flash Tts

Google's fast, expressive text-to-speech model with 30 voices and 70+ language support

google-gemini-3.1-flash-tts استفاده ←
سفارشی 600 اعتبار

resemble-ai: Chatterbox

Generate expressive, natural speech. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.

resemble-ai-chatterbox استفاده ←

مدل‌ساز ۳بعدی

ساخت مدل سه‌بعدی از متن یا تصویر

36 مدل فعال

شروع رایگان
سفارشی 49 اعتبار

vufinder: Map Anything

Universal Feed-Forward Metric 3D Reconstruction

vufinder-map-anything استفاده ←
سفارشی 49 اعتبار

vufinder: Vggt 1B Depth

Feed-forward neural network that directly infers all key 3D attributes of a scene.

vufinder-vggt-1b-depth استفاده ←
سفارشی 390 اعتبار

colinhughes2121: Ai Avatar 3D

Turn any photo into a 3D-rendered character avatar. Roblox/Fortnite/Discord-style. For gamers, creators, virtual worlds, profile pics. · api.gocreativeai.com — AI agent APIs

colinhughes2121-ai-avatar-3d استفاده ←
سفارشی 390 اعتبار

colinhughes2121: Pixar Character Stylizer

Convert any photo into a 3D animated Pixar-style character. Big expressive eyes, soft rendering, family-movie aesthetic. · api.gocreativeai.com — AI agent APIs

colinhughes2121-pixar-character-stylizer استفاده ←
سفارشی 390 اعتبار

fire: Trellis

For the paper "Structured 3D Latents for Scalable and Versatile 3D Generation".

fire-trellis استفاده ←
سفارشی 390 اعتبار

fishwowater: Trellis Image

SOTA image-to-3D generator TRELLIS. Equipped with **All image-condition types** you need: (1) single-image (2) multi-images (3) different images for geometry and texture generation (4) mesh+images for detail variation(texture painting)

fishwowater-trellis-image استفاده ←
سفارشی 390 اعتبار

hcl14: Direct3D S2

Direct3D‑S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention

hcl14-direct3d-s2 استفاده ←
سفارشی 390 اعتبار

hcl14: Omnipart

OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion

hcl14-omnipart استفاده ←
سفارشی 390 اعتبار

subhash25rawat: Morphix3D

Transform Images & Text into 3D Models with AI

subhash25rawat-morphix3d استفاده ←
سفارشی 390 اعتبار

vufinder: Map Anything Pano

Universal Feed-Forward Metric 3D Reconstruction from panoramic images

vufinder-map-anything-pano استفاده ←
سفارشی 390 اعتبار

vufinder: Map Anything Pi3

A feed-forward neural network that offers a novel approach to visual geometry reconstruction.

vufinder-map-anything-pi3 استفاده ←
سفارشی 390 اعتبار

vufinder: Map Anything Pi3X

A feed-forward neural network that offers a novel approach to visual geometry reconstruction.

vufinder-map-anything-pi3x استفاده ←
سفارشی 390 اعتبار

vufinder: Map Anything Vggt 1B

Feed-forward neural network that directly infers all key 3D attributes of a scene.

vufinder-map-anything-vggt-1b استفاده ←
Alibaba Wan 390 اعتبار

wanjon: Ardy2Vrm

Generate text-driven 3D motion with NVIDIA ARDY and export animated previews, NPZ data, and VRM Animation files for Three.js and Three-VRM.

wanjon-ardy2vrm استفاده ←
سفارشی 390 اعتبار

zylim0702: Yun 3D 2.1

Replicate

zylim0702-yun-3d-2.1 استفاده ←
سفارشی 560 اعتبار

ayushunleashed: Partpacker

Part-level 3D object generation from single-view images

ayushunleashed-partpacker استفاده ←
سفارشی 560 اعتبار

fire: Part Crafter

PartCrafter is a structured 3D mesh generation model that creates multiple parts and objects from a single RGB image.

fire-part-crafter استفاده ←
سفارشی 560 اعتبار

hcl14: Hunyuan P3 Sam

Replicate

hcl14-hunyuan-p3-sam استفاده ←
سفارشی 610 اعتبار

aliadhami: 3D Sheet

Generates minimal, consistent 3-view design for objects in the style of technical 3D reference sheets. Outputs include front, back, and side views with consistent proportions against a white background.

aliadhami-3d-sheet استفاده ←
سفارشی 950 اعتبار

valllllex: Cartoonia 3D

Restyling the classic cartoon and comic book characters into a soft painterly 3D style. FLUX Kontext Lora

valllllex-cartoonia-3d استفاده ←
سفارشی 1,500 اعتبار

vufinder: Vggt 1B

Feed-forward neural network that directly infers all key 3D attributes of a scene.

vufinder-vggt-1b استفاده ←
سفارشی 1,700 اعتبار

firtoz: Trellis

A powerful 3D asset generation model

firtoz-trellis استفاده ←
سفارشی 1,900 اعتبار

vufinder: Vggt 1B Point

Feed-forward neural network that directly infers all key 3D attributes of a scene.

vufinder-vggt-1b-point استفاده ←
سفارشی 2,150 اعتبار

kfarr: Sharp Ml

Apple's SHARP model — single image to 3D Gaussian splats

kfarr-sharp-ml استفاده ←
سفارشی 2,250 اعتبار

tencent: Hunyuan3D 2

Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

tencent-hunyuan3d-2 استفاده ←
سفارشی 5,000 اعتبار

tencent: Hunyuan3D 2Mv

Hunyuan3D-2mv is finetuned from Hunyuan3D-2 to support multiview controlled shape generation.

tencent-hunyuan3d-2mv استفاده ←
سفارشی 5,500 اعتبار

ndreca: Hunyuan3D 2

[Turbo Mode] Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

ndreca-hunyuan3d-2 استفاده ←
سفارشی 10,000 اعتبار

uthana: Create Character V1

Rig any 3D bipedal character mesh

uthana-create-character-v1 استفاده ←
سفارشی 11,000 اعتبار

prunaai: Hunyuan3D 2

hunyuan3d-2 optimised with the pruna toolkit: https://github.com/PrunaAI/pruna

prunaai-hunyuan3d-2 استفاده ←
سفارشی 13,000 اعتبار

ndreca: Hunyuan3D 2.1

[Quality Mode] Scaling Diffusion Models for High Resolution Textured 3D Assets Generation

ndreca-hunyuan3d-2.1 استفاده ←
سفارشی 20,000 اعتبار

hyper3d: Rodin

Generate complex 3D models from images with Rodin Gen-2

hyper3d-rodin استفاده ←
سفارشی 20,000 اعتبار

tencent: Hunyuan 3D 3.1

3D models with texture fidelity and geometry precision

tencent-hunyuan-3d-3.1 استفاده ←
سفارشی 23,000 اعتبار

aaronjmars: Unirig Ai

One Model to Rig Them All: Diverse Skeleton Rigging with UniRig

aaronjmars-unirig-ai استفاده ←
سفارشی 25,000 اعتبار

uthana: Text To Motion Diffusion V2

Generate 3D character animation data from a text prompt

uthana-text-to-motion-diffusion-v2 استفاده ←
سفارشی 25,000 اعتبار

uthana: Text To Motion Vqvae V1

Generate 3D character animation data from a text prompt

uthana-text-to-motion-vqvae-v1 استفاده ←
سفارشی 58,000 اعتبار

fishwowater: Trellis2

TRELLIS.2: Native and Compact Structured Latents for 3D Generation

fishwowater-trellis2 استفاده ←

آماده شروعی؟

همه این مدل‌ها داخل داشبورد در دسترس‌اند. یک حساب بسازید و از سرویس مورد نظرتان استفاده کنید.

ثبت‌نام و شروع رایگان