AI MODELS
مدلهای IranBrain
593 مدل فعال در 6 دستهبندی سرویس — مستقیم به ابزارهای داشبورد وصل میشوند.
چت هوشمند
گفتگو، نویسندگی و استدلال با مدلهای زبانی
47 مدل فعال
ibm-granite: Granite Vision 3.3 2B
Granite-vision-3.3-2b is a compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.
ibm-granite: Granite 3.2 8B Instruct
Granite-3.2-8B-Instruct is a 8-billion parameter 128K context length language model fine-tuned for reasoning and instruction-following capabilities.
ibm-granite: Granite 3.3 8B Instruct
Granite-3.3-8B-Instruct is a 8-billion parameter 128K context length language model fine-tuned for improved reasoning and instruction-following capabilities.
ibm-granite: Granite 4.0 H Small
Granite-4.0-H-Small is a 32B parameter long-context instruct model finetuned from Granite-4.0-H-Small-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets.
ibm-granite: Granite 4.1 8B
Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from Granite-4.1-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets.
ibm-granite: Granite Vision 4.1 4B
Granite Vision 4.1 4B is a vision-language model (VLM) that delivers frontier-level performance on structured document extraction tasks — chart extraction, table extraction, and semantic key-value pair extraction — in a compact 4B parameter footprint
openai: Gpt 5 Nano
Fastest, most cost-effective GPT-5 model from OpenAI
openai: Gpt Oss 20B
20b open-weight language model from OpenAI
meta: Llama Guard 4 12B
Replicate
openai: Gpt 4.1 Nano
Fastest, most cost-effective GPT-4.1 model from OpenAI
openai: Gpt 4O Mini
Low latency, low cost version of OpenAI's GPT-4o model
meta: Llama 4 Scout Instruct
A 17 billion parameter model with 16 experts
openai: Gpt Oss 120B
120b open-weight language model from OpenAI
meta: Llama 4 Maverick Instruct
A 17 billion parameter model with 128 experts
qwen: Qwen3 235B A22B Instruct 2507
Updated Qwen3 model for instruction following
qwen: Qwen3 7 Plus
Qwen3.7-Plus is Alibaba's cost-effective multimodal model with vision-language understanding, a 1 million token context window, and strong agentic coding and tool use.
openai: Gpt 4.1 Mini
Fast, affordable version of GPT-4.1
openai: Gpt 5 Mini
Faster version of OpenAI's flagship GPT-5 model
google: Gemini 2.5 Flash
Google’s hybrid “thinking” AI model optimized for speed and cost-efficiency
twangodev: Qwenasr
Serve QwenASR speech recognition and alignment.
deepseek-ai: Deepseek V3.1
Latest hybrid thinking model from Deepseek
google: Gemini 3 Flash
Google's most intelligent model built for speed with frontier intelligence, superior search, and grounding
deepseek-ai: Deepseek V3
DeepSeek-V3-0324 is the leading non-reasoning model, a milestone for open source
anthropic: Claude 3.5 Haiku
Anthropic's fastest, most cost-effective model, with a 200K token context window (claude-3-5-haiku-20241022)
anthropic: Claude 4.5 Haiku
Claude Haiku 4.5 gives you similar levels of coding performance but at one-third the cost and more than twice the speed
openai: Gpt 5.6 Luna
OpenAI's GPT-5.6 cost-optimized tier, built for fast, high-volume, latency-sensitive workloads.
google: Gemini 3.5 Flash
Google's fast multimodal model with frontier reasoning across agents, coding, and long-context tasks
openai: Gpt 4.1
OpenAI's Flagship GPT model for complex tasks.
openai: Gpt 5
OpenAI's new model excelling at coding, writing, and reasoning.
openai: Gpt 5 Structured
GPT-5 with support for structured outputs, web search and custom tools
openai: Gpt 5.1
The best model for coding and agentic tasks with configurable reasoning effort.
anthropic: Claude Sonnet 5
Anthropic's most agentic Sonnet model, bringing frontier-level coding and tool use at Sonnet's speed and price
openai: Gpt 4O
OpenAI's high-intelligence chat model
google: Gemini 3.1 Pro
Google's most intelligent model, with improved reasoning and a new medium thinking level
deepseek-ai: Deepseek R1
A reasoning model trained with reinforcement learning, on par with OpenAI o1
openai: Gpt 5.2
The best model for coding and agentic tasks across industries
openai: Gpt 5.4
OpenAI's most capable frontier model for complex professional work, coding, and multi-step reasoning.
openai: Gpt 5.6 Sol
OpenAI's GPT-5.6 flagship tier, built for complex professional work, coding, and deep multi-step reasoning.
openai: Gpt 5.6 Terra
OpenAI's GPT-5.6 balanced tier, tuned for everyday production work at roughly half the cost of the flagship.
anthropic: Claude 3.7 Sonnet
The most intelligent Claude model and the first hybrid reasoning model on the market (claude-3-7-sonnet-20250219)
anthropic: Claude 4 Sonnet
Claude Sonnet 4 is a significant upgrade to 3.7, delivering superior coding and reasoning while responding more precisely to your instructions
anthropic: Claude 4.5 Sonnet
Claude Sonnet 4.5 is the best coding model to date, with significant improvements across the entire development lifecycle
anthropic: Claude Sonnet 4.6
Claude Sonnet 4.6 from Anthropic: a full upgrade to coding, computer use, long-context reasoning, agent planning, knowledge work, and design, with a 1 million token context window in beta.
anthropic: Claude Opus 4.6
Anthropic's most intelligent model with state-of-the-art coding, reasoning, and agentic capabilities
anthropic: Claude Opus 4.7
Anthropic's most capable model with a step-change improvement in agentic coding, better vision, and stronger multi-step reasoning
anthropic: Claude Fable 5
Claude Fable 5 from Anthropic: the next generation of intelligence for the hardest knowledge work and coding problems.
openai: Gpt 5 Pro
The smartest, fastest, most useful model yet, with built-in thinking that puts expert-level intelligence in everyone’s hands
تصویرساز
تولید و ویرایش تصویر با مدلهای پیشرفته
182 مدل فعال
fofr: Color Matcher
Color match and white balance fixes for images
sourceful: Riverflow 2.0 Refsr
Render product images with 100% accuracy and environmental blending
openai: Clip
Official CLIP models, generate CLIP (clip-vit-large-patch14) text & image embeddings
datalab-to: Marker
Convert PDF to markdown + JSON quickly with high accuracy
topazlabs: Image Upscale
Professional-grade image upscaling, from Topaz Labs
perceptron-ai-inc: Isaac 0.1
an open-source, 2B-parameter model built for real-world applications
leonardoai: Phoenix 1.0
Leonardo AI’s first foundational model produces images up to 5 megapixels (fast, quality and ultra modes)
black-forest-labs: Flux 2 Klein 4B
Very fast image generation and editing model. 4 steps distilled, sub-second inference for production and near real-time applications.
sourceful: Riverflow 2.0 Fast
Agentic image model optimized for high-quality, fast generations supporting font control
openai: Gpt Image 1
A multimodal image generation model that creates high-quality images. You need to bring your own verified OpenAI key to use this model. Your OpenAI account will be charged for usage.
moonshotai: Kimi K2 Thinking
Kimi K2 Thinking is the latest, most capable version of an open-source thinking model.
topazlabs: Image Colorization
Image colorization model from Topaz Labs
moonshotai: Kimi K2.5
Moonshot AI's latest open model. It unifies vision and text, thinking and non-thinking modes, and single-agent and multi-agent execution into one model
moonshotai: Kimi K2.6
Moonshot AI's frontier open model, built for long-horizon coding, agent swarms, and autonomous software engineering. 1 trillion parameters, 262k context window, vision and tool use.
openai: O4 Mini
OpenAI's fast, lightweight reasoning model
openai: O1 Mini
A small model alternative to o1
topazlabs: Dust And Scratch V2
Remove dust and scratches from old photos
openai: Gpt Image 1 Mini
A cost-efficient version of GPT Image 1
sourceful: Riverflow V2.5 Fast
Speed-optimized variant of Riverflow 2.5 for production and latency-sensitive workflows
sourceful: Riverflow V2.5 Pro
Top-quality agentic image model with multi-step reasoning, candidate scoring, and adjustable thinking effort
philz1337x: Clarity Pro Upscaler
The first creative upscaler which keeps identity. Stunning photorealistic results, realistic skin, and full creative control.
prunaai: P Image Try On
Virtual try-on. Put one or more garments onto a person photo while keeping their face, pose, and body.
ibm-granite: Granite Embedding Small English R2
Granite-embedding-small-english-r2 is a 47M parameter dense biencoder embedding model from the Granite Embeddings collection that can be used to generate high quality text embeddings.
google: Gemini 3 Pro
Google's most advanced reasoning Gemini model
minimax: Image 01
Minimax's first image model, with character reference support
prunaai: Flux Kontext Fast
Ultra fast flux kontext endpoint
prunaai: P Image Edit
A sub 1 second 0.01$ multi-image editing model built for production use cases. For image generation, check out p-image here: https://replicate.com/prunaai/p-image
prunaai: P Image Ideogram
P-Image-Ideogram is Pruna AI’s text-to-image model starting from 0,003$ per generation.
prunaai: P Image Upscale
Fastest image upscaler in the world (<1s) supporting outputs up to 128 MP. contact Pruna AI for dedicated endpoints.
prunaai: Z Image Turbo
Z-Image Turbo is a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.
recraft-ai: Recraft Remove Background
Automated background removal for images. Tuned for AI-generated content, product photos, portraits, and design workflows
recraft-ai: Recraft Vectorize
Convert raster images to high-quality SVG format with precision and clean vector paths, perfect for logos, icons, and scalable graphics.
reve: Edit Fast
Reve's fast image edit model at only $0.01 per edit
black-forest-labs: Flux 2 Klein 9B Base
Un-distilled version of FLUX.2 [klein]. A foundation model for maximum flexibility and control
black-forest-labs: Flux 2 Dev
Quality image generation and editing with support for reference images
openai: Gpt Image 2
OpenAI's state-of-the-art image generation model. Create and edit images from text with strong instruction following, sharp text rendering, and detailed editing.
openai: Gpt Image 1.5
OpenAI's latest image generation model with better instruction following and adherence to prompts
black-forest-labs: Flux 2 Klein 9B
4 step distilled version of FLUX.2 [klein]. A foundation model for maximum flexibility and control
black-forest-labs: Flux 2 Pro
High-quality image generation and editing with support for eight reference images
leonardoai: Lucid Origin
Artistic and high-quality visuals with improved prompt adherence, diversity, and definition
retro-diffusion: Rd Fast
Fast pixel art image generation
sourceful: Riverflow 2.0 Pro
Agentic image model optimized for robust, high-precision generations supporting font control
bria: Remove Background
Bria AI's remove background model
black-forest-labs: Flux 2 Klein 4B Base Lora
A version of FLUX.2 [klein] 4B-base that supports fast fine-tuned lora inference
black-forest-labs: Flux Schnell Lora
The fastest image generation model tailored for fine-tuned use
google: Imagen 4 Fast
Use this fast version of Imagen 4 when speed and cost are more important than quality
google: Upscaler
Upscale images 2x or 4x times
openai: Dall E 2
The original classic DALLᐧE 2
prunaai: Wan 2.2 Image
This model generates beautiful cinematic 2 megapixel images in 3-4 seconds and is derived from the Wan 2.2 model through optimisation techniques from the pruna package
qwen: Qwen Image 2512
Qwen Image 2512 is an improved version of Qwen Image with more realistic human generation, finer textures, and stronger text rendering
tencent: Hunyuan Image 2.1
Generate high-quality 2K resolution images from text prompts
black-forest-labs: Flux 2 Klein 9B Base Lora
A version of FLUX.2 [klein] 9B-base that supports fast fine-tuned lora inference
recraft-ai: Recraft 20B
Affordable and fast images
retro-diffusion: Rd Plus
High quality and authentic pixel art image generation
retro-diffusion: Rd Tile
All the tools you need for generating pixel art tilesets
black-forest-labs: Flux Canny Dev
Open-weight edge-guided image generation. Control structure and composition using Canny edge detection.
black-forest-labs: Flux Depth Dev
Open-weight depth-aware image generation. Edit images while preserving spatial relationships.
black-forest-labs: Flux Dev
A 12 billion parameter rectified flow transformer capable of generating images from text descriptions
black-forest-labs: Flux Kontext Dev
Open-weight version of FLUX.1 Kontext
black-forest-labs: Flux Krea Dev
An opinionated text-to-image model from Black Forest Labs in collaboration with Krea that excels in photorealism. Creates images that avoid the oversaturated "AI look".
black-forest-labs: Flux Redux Dev
Open-weight image variation model. Create new versions while preserving key elements of your original.
ideogram-ai: Ideogram V2A Turbo
Like Ideogram v2 turbo, but now faster and cheaper
qwen: Qwen Image
An image generation foundation model in the Qwen series that achieves significant advances in complex text rendering.
reve: Create
Image generation model from Reve
wavespeedai: Qwen Image
A 20B MMDiT model for next-gen text-to-image generation
black-forest-labs: Flux 2 Max
The highest fidelity image model from Black Forest Labs
bytedance: Dreamina 3.1
4MP text-to-image generation with enhanced cinematic-quality image generation with precise style control, improved text rendering, and commercial design optimization.
bytedance: Seedream 3
A text-to-image model with support for native high-resolution (2K) image generation
bytedance: Seedream 4
Unified text-to-image generation and precise single-sentence editing at up to 4K resolution
ideogram-ai: Ideogram V3 Turbo
Turbo is the fastest and cheapest Ideogram v3. v3 creates images with stunning realism, creative designs, and consistent styles
krea: Krea 2 Medium
Foundation image model from Krea, tuned for expressive illustration, anime, and painterly styles. Fast and consistent across artistic directions.
qwen-edit-apps: Qwen Image Edit Plus Lora Fusion
Fusion – Product/object blending that fixes perspective and lighting so the subject melts into a new background via the Fusion LoRA.
qwen-edit-apps: Qwen Image Edit Plus Lora Next Scene
Next Scene – “Next beat” cinematic edits that keep subject identity while steering to the next camera move via the Next Scene LoRA
qwen-edit-apps: Qwen Image Edit Plus Lora Photo To Anime
Photo to Anime – Stylized conversion that turns photos into crisp cel-shaded anime frames using the Photo-to-Anime LoRA.
qwen-edit-apps: Qwen Image Edit Plus Lora Relight
Relight – Soft, curtain-filtered relighting that repaints the scene with golden-hour or moody tones using the Relight LoRA.
qwen-edit-apps: Qwen Image Edit Plus Lora Skin
Skin – Natural beauty retouch that enhances pores and tonal variation (no plastic skin) via the Skin LoRA.
qwen-edit-apps: Qwen Image Edit Plus Lora Upscale
Upscale – Detail-loving upscale/restore pass that sharpens textures and color fidelity with the Upscale LoRA.
qwen: Qwen Edit Multiangle
Camera-aware edits for Qwen/Qwen-Image-Edit-2509 with Lightning + multi-angle LoRA
qwen: Qwen Image Edit
Edit images using a prompt. This model extends Qwen-Image’s unique text rendering capabilities to image editing tasks, enabling precise text editing
qwen: Qwen Image Edit 2511
An enhanced version over Qwen-Image-Edit-2509, featuring multiple improvements including notably better consistency
qwen: Qwen Image Edit Plus
The latest Qwen-Image’s iteration with improved multi-image editing, single-image consistency, and native support for ControlNet
qwen: Qwen Image Edit Plus Lora
Qwen Image Edit 2509 LoRA explorer, uses HuggingFace URLs to load any safetensor
black-forest-labs: Flux Dev Lora
A version of flux-dev, a text to image model, that supports fast fine-tuned lora inference
black-forest-labs: Flux Kontext Dev Lora
FLUX.1 Kontext[dev] image editing model for running lora finetunes
google: Nano Banana 2 Lite
Google's fastest image generation model — the lightweight, low-cost version of Nano Banana 2, for rapid creation and editing
google: Nano Banana Pro
Google's state of the art image generation and editing model 🍌🍌
qwen: Qwen Image 2
A next-generation image generation and editing model from Alibaba's Qwen team. Supports text-to-image and image editing with strong text rendering, especially for Chinese.
stability-ai: Stable Diffusion 3.5 Medium
2.5 billion parameter image model with improved MMDiT-X architecture
google: Gemini 2.5 Flash Image
Google's latest image generation model in Gemini 2.5
google: Nano Banana
Google's latest image editing model in Gemini 2.5
black-forest-labs: Flux 1.1 Pro
Faster, better FLUX Pro. Text-to-image model with excellent image quality, prompt adherence, and output diversity.
black-forest-labs: Flux Fill Dev
Open-weight inpainting model for editing and extending images. Guidance-distilled from FLUX.1 Fill [pro].
black-forest-labs: Flux Kontext Pro
A state-of-the-art text-based image editing model that delivers high-quality outputs with excellent prompt following and consistent results for transforming images through natural language
bria: Eraser
SOTA Object removal, enables precise removal of unwanted objects from images while maintaining high-quality outputs. Trained exclusively on licensed data for safe and risk-free commercial use
bria: Fibo
SOTA Open source model trained on licensed data, transforming intent into structured control for precise, high-quality AI image generation in enterprise and agentic workflows.
bria: Fibo Edit
FIBO-Edit brings the power of structured prompt generation to image editing
bria: Generate Background
Bria Background Generation allows for efficient swapping of backgrounds in images via text prompts or reference image, delivering realistic and polished results. Trained exclusively on licensed data for safe and risk-free commercial use
bria: Genfill
Bria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use.
bria: Image 3.2
Commercial-ready, trained entirely on licensed data, text-to-image model. With only 4B parameters provides exceptional aesthetics and text rendering. Evaluated to be on par to other leading models in the market
bria: Increase Resolution
Bria Increase resolution upscales the resolution of any image. It increases resolution using a dedicated upscaling method that preserves the original image content without regeneration.
bria: Product Cutout
Precise AI-powered product cutout with 256-level transparency for eCommerce
bria: Product Packshot
Transform any product photo into professional 2000x2000px packshots with optimal positioning
bria: Product Shadow
Add consistent, customizable shadows to product cutouts for enhanced visual appeal
bytedance: Seedream 4.5
Seedream 4.5: Upgraded Bytedance image model with stronger spatial understanding and world knowledge
flux-kontext-apps: Cartoonify
Turn your image into a cartoon with FLUX.1 Kontext [pro]
flux-kontext-apps: Change Haircut
Quickly change someone's hair style and hair color, powered by FLUX.1 Kontext [pro]
flux-kontext-apps: Depth Of Field
Bring your subjects into focus with FLUX.1 Kontext [pro]
flux-kontext-apps: Face To Many Kontext
Become a character, in style
flux-kontext-apps: Filters
Add simple filters to your images
flux-kontext-apps: Iconic Locations
Put yourself in an iconic location around the world from a single image
flux-kontext-apps: Impossible Scenarios
Experience impossible adventures and extreme scenarios from a single image
flux-kontext-apps: Multi Image Kontext Pro
An experimental model with FLUX Kontext Pro that can combine two input images
flux-kontext-apps: Portrait Series
Create a series of portrait photos from a single image
flux-kontext-apps: Professional Headshot
Create a professional headshot photo from any single image
flux-kontext-apps: Restore Image
Use FLUX Kontext to restore, fix scratches and damage, and colorize old photos
flux-kontext-apps: Text Removal
Remove all text from an image with FLUX.1 Kontext
ideogram-ai: Ideogram V2A
Like Ideogram v2, but faster and cheaper
recraft-ai: Recraft V3
Recraft V3 (code-named red_panda) is a text-to-image model with the ability to generate long texts, and images in a wide list of styles. As of today, it is SOTA in image generation, proven by the Text-to-Image Benchmark by Artificial Analysis
recraft-ai: Recraft V4
Recraft's latest image generation model, built around design taste. Strong prompt accuracy, art-directed composition, and integrated text rendering. Fast and cost-efficient at standard resolution.
recraft-ai: Recraft V4.1
Recraft's latest image generation model, built around design taste. Strong prompt accuracy, art-directed composition, and integrated text rendering. Fast and cost-efficient at standard resolution.
recraft-ai: Recraft V4.1 Svg
Generate production-ready SVG vector images from text prompts. Recraft V4.1's design taste applied to vector output — clean geometry, structured layers, and editable paths.
recraft-ai: Recraft V4.1 Utility
A faster, lighter Recraft image generation model optimized for high-volume and production pipelines. Same design taste as V4.1, built for speed and throughput.
reve: Edit
Image editing model from Reve
reve: Remix
Image generation model from Reve which handles multiple input reference images
stability-ai: Stable Diffusion 3.5 Large Turbo
A text-to-image model that generates high-resolution images with fine details. It supports various artistic styles and produces diverse outputs from the same prompt, with a focus on fewer inference steps
xai: Grok Imagine Image 2
xAI's Grok Imagine Image 2.0 — text-to-image generation and editing with a quality control and output up to 2k
recraft-ai: Recraft 20B Svg
Affordable and fast vector images
bytedance: Seedream 5 Pro
ByteDance's flagship text-to-image and image editing model, generating sharp 1K and 2K images from text or up to 10 reference images
openai: O1
OpenAI's first o-series reasoning model
black-forest-labs: Flux Canny Pro
Professional edge-guided image generation. Control structure and composition using Canny edge detection
black-forest-labs: Flux Depth Pro
Professional depth-aware image generation. Edit images while preserving spatial relationships.
black-forest-labs: Flux Fill Pro
Professional inpainting and outpainting model with state-of-the-art performance. Edit or extend images with natural, seamless results.
easel: Ai Avatars
Use one or two face images to create AI avatars
ideogram-ai: Ideogram V2 Turbo
A fast image model with state of the art inpainting, prompt comprehension and text rendering.
philz1337x: Crystal Upscaler
High-precision image upscaler optimized for portraits, faces and products. One of the upscale modes powered by Clarity AI. X:https://x.com/philz1337x
xai: Grok Imagine Image Quality
xAI's higher-quality image model with sharper details, better text rendering, and 2k output
black-forest-labs: Flux Pro
State-of-the-art image generation with top of the line prompt following, visual quality, image detail and output diversity.
black-forest-labs: Flux 1.1 Pro Ultra
FLUX1.1 [pro] in ultra and raw modes. Images are up to 4 megapixels. Use raw mode for realism.
black-forest-labs: Flux 2 Flex
Max-quality image generation and editing with support for ten reference images
black-forest-labs: Flux Pro Finetuned
Inference model for FLUX.1 [pro] using custom `finetune_id`
google: Imagen 4 Ultra
Use this ultra version of Imagen 4 when quality matters more than speed and cost
ideogram-ai: Ideogram V3 Balanced
Balance speed, quality and cost. Ideogram v3 creates images with stunning realism, creative designs, and consistent styles
ideogram-ai: Ideogram V4 Balanced
Balance speed, quality and cost. Ideogram v4 creates images with stunning realism, creative designs, and consistent styles
krea: Krea 2 Large
Krea's flagship foundation image model. Larger and more flexible than Krea 2 Medium, with particular strength in photorealism and expressive artistic styles.
stability-ai: Stable Diffusion 3.5 Large
A text-to-image model that generates high-resolution images with fine details. It supports various artistic styles and produces diverse outputs from the same prompt, thanks to Query-Key Normalization.
google: Nano Banana 2
Google's fast image generation model with conversational editing, multi-image fusion, and character consistency
black-forest-labs: Flux 1.1 Pro Ultra Finetuned
Inference model for FLUX 1.1 [pro] Ultra using custom `finetune_id`. Supports 4MP images and raw mode for realism
ibm-granite: Granite Embedding 278M Multilingual
Granite-Embedding-278M-Multilingual is a 278M parameter model from the Granite Embeddings suite that can be used to generate high quality text embeddings
retro-diffusion: Rd Animation
Style consistent animated pixel art sprite generation
qwen: Qwen Image 2 Pro
The pro version of Qwen Image 2 from Alibaba's Qwen team. Enhanced text rendering, realism, and semantic adherence for high-quality image generation and editing.
black-forest-labs: Flux Kontext Max
A premium text-based image editing model that delivers maximum performance and improved typography generation for transforming images through natural language prompts
flux-kontext-apps: Multi Image Kontext Max
An experimental FLUX Kontext model that can combine two input images
flux-kontext-apps: Multi Image List
FLUX Kontext max with list input for multiple images
ideogram-ai: Ideogram V2
An excellent image model with state of the art inpainting, prompt comprehension and text rendering
recraft-ai: Recraft V3 Svg
Recraft V3 SVG (code-named red_panda) is a text-to-image model with the ability to generate high quality SVG images including logotypes, and icons. The model supports a wide list of styles.
tencent: Hunyuan Image 3
A powerful native multimodal model for image generation (PrunaAI squeezed)
ideogram-ai: Ideogram V3 Quality
The highest quality Ideogram v3 model. v3 creates images with stunning realism, creative designs, and consistent styles
ideogram-ai: Layerize
Take a flat graphic, remove text, and get structured text layers back for editing and recomposing
ideogram-ai: Ideogram Character
Generate consistent characters from a single reference image. Outputs can be in many styles. You can also use inpainting to add your character to an existing image.
ideogram-ai: Ideogram V4 Quality
The highest quality Ideogram v4 model. v4 creates images with stunning realism, creative designs, and consistent styles
openai: Dall E 3
An AI system that can create realistic images and art from a description in natural language.
black-forest-labs: Flux 2 Klein 4B Base
Un-distilled version of FLUX.2 [klein]. Optimized for fine-tuning, customization, and post-training workflows
qwen: Qwen Image Lora Trainer Legacy
Fine-tunable Qwen Image model with exceptional composition abilities - train custom LoRAs for any style or subject
reve: Reve 2.1
Generate and edit images from text and reference images with Reve 2.1
recraft-ai: Recraft V4 Pro
Recraft's latest image generation model at ~2048px resolution. Same design taste and prompt accuracy as V4, with higher resolution for print-ready and large-scale work.
recraft-ai: Recraft V4.1 Pro
Recraft's latest image generation model at ~2048px resolution. Same design taste and prompt accuracy as V4.1, with higher resolution for print-ready and large-scale work.
recraft-ai: Recraft V4.1 Pro Svg
Generate detailed SVG vector graphics from text prompts. Recraft V4.1 Pro's design taste with more geometric detail and finer paths — clean layers, editable output, and scalable to any size.
recraft-ai: Recraft V4.1 Utility Pro
A faster, lighter Recraft image generation model at ~2048px resolution, optimized for high-volume production. Design taste and prompt accuracy at high resolution with better throughput.
sync: Lipsync 2
Generate realistic lipsyncs with Sync Labs' 2.0 model
recraft-ai: Recraft Creative Upscale
Creative Upscale focuses on enhancing details and refining complex elements in the image. It doesn’t just increase resolution but adds depth by improving textures, fine details, and facial features.
recraft-ai: Recraft V4 Pro Svg
Generate detailed SVG vector graphics from text prompts. Recraft V4 Pro's design taste with more geometric detail and finer paths — clean layers, editable output, and scalable to any size.
sync: Lipsync 2 Pro
Studio-grade lipsync in minutes, not weeks
sync: React 1
Realistic lipsync with refined human emotion capabilities
intelligent-utilities: Html To Image
Replicate
nightmareai: Real Esrgan
Real-ESRGAN with optional face correction and adjustable upscale
black-forest-labs: Flux Redux Schnell
Fast, efficient image variation model for rapid iteration and experimentation.
black-forest-labs: Flux Schnell
The fastest image generation model tailored for local development and personal use
prunaai: Flux Fast
This is the fastest Flux endpoint in the world.
prunaai: Hidream L1 Fast
This is an optimised version of the hidream-l1 model using the pruna ai optimisation toolkit!
prunaai: P Image
A sub 1 second text-to-image model built for production use cases.
prunaai: P Image Lora
Use trained LoRAs from the https://replicate.com/prunaai/p-image-trainer. Find or contribute LoRAs here https://huggingface.co/collections/PrunaAI/p-image-loras
recraft-ai: Recraft Crisp Upscale
Designed to make images sharper and cleaner, Crisp Upscale increases overall quality, making visuals suitable for web use or print-ready materials.
ویدیوساز
ساخت ویدیو از متن، تصویر و رفرنس
293 مدل فعال
lucataco: Frame Extractor
Extract the first or last frame from any video file as a high-quality image
lucataco: Video Audio Merge
merge a video and an audio file
emersimeon: Image Audio Video
Add an image and a song and generate a video
bitflow: Kinetic Captions
Generate dynamic, stylish captions and hard-burn them into videos using ASS subtitles and FFmpeg.
bitflow: Video Color Filter Lut
This LUT-based color filter is ideal for color grading user-generated AI videos or short videos shot on smartphones.
chamelaion: Lipsync
Lip sync any video in up to Full HD resolution at 1080p to any audio track, including videos with multiple speakers. Active speaker detection is built in, while speaking style preservation ensures natural mouth movements.
craftfulcharles: Waveform
Generate Waveform videos using an audio file and an image.
foixasoftware: Ffmpeg
merge videos
hautechai: Json To Media
Replicate
idan054: Sarra Video Maker V1
This Adds Background Music & Prefect subtitles from json file
lex2029: Vlogme Avatar Bridge
Create a vertical talking-avatar video from a centered photo and speech audio.
littletry: Media Metadata
Extract normalized metadata from video, audio, and image files.
littletry: Video Frame Custom
Extract the first, last, middle, specific, or evenly spaced frames from a video.
lucataco: Dotted Waveform Visualizer
Create a dotted waveform video from an audio file
lucataco: Split Screen Video
Combines two videos into a single split-screen layout
lucataco: Vid2Webp
Convert your video into webp format (with looping)
onemadgeek: Frames To Video Merger
Convert a set of image frames (JPG or PNG) into a high-quality MP4 video. Automatically handles sorting and frame order for smooth playback.
peelss: Ai Video Subtitle
Add animated captions to your video
smokeyang: Transition Generator
Cinematic photo or video transition generator featuring 4 premium motion styles driven by Stable Diffusion.
thecmdrunner: Feather 1
AI-powered Text-To-Video Generation for Animated Motion Graphics
zhyjiong: Sora Watermark Remover
Remove 'Sora' watermark on videos generated by sora
lucataco: Image To Video Slideshow
Transform a collection of images into a video slideshow
lucataco: Dotted Video
Converts a video into a black and white dotted video effect
nightowlstudio-11: Clipforge
Add Text & Music to Videos
hjunior29: Video Text Remover
Clean videos by automatically removing text overlays
mattsays: Sam3 Image
A unified foundation model for prompt-based segmentation in images and videos
ghostljj: Sora2 Watermark Remover Fix
Removes the watermark from Sora 2 videos using a trained model and IOpaint (Fixed watermark detection leaks and production errors.)
hcolde: Bmv2
Alpha-only video matting for 5-10 second white-background videos using BackgroundMattingV2 MobileNetV2 TorchScript.
hcolde: Rvm
RobustVideoMatting on Replicate: input mp4 video, output black-and-white alpha-mask.mp4.
nelsonjchen: Op Replay Clipper
GPU accelerated replay renderer / video data clipper for comma.ai connect's openpilot route data. SEE README.
mirelo: Video To Sfx V1
Generate synced sounds for any video, and return it with its new sound track
mirelo: Video To Sfx V1.5
Generate synced sounds for any video and return it with its new soundtrack - now enhanced in version 1.5 for improved sound synchronization and realism
mptamilselvan: Download Media
Download videos or extract audio from popular social media platforms quickly and easily. This tool supports links from platforms like Facebook, Instagram, and YouTube, allowing users to save content for offline viewing or personal use.
megacubos: Video Watermark
Add a customizable watermark to any video with a simple API call. Used in the video AI creator at UGCMade.com and brought to you by your friends at Megacubos.com.
bfirsh: Concatenate Videos
Stitches videos together
nicolascoutureau: Video Utils
Replicate
pixverse: Pixverse V4
Quickly generate smooth 5s or 8s videos at 540p, 720p or 1080p
pixverse: Pixverse V3.5
Create videos in as little as 10 seconds. 5s or 8s videos at 360p, 540p, 720p or 1080p.
lucataco: Video Merge
Simple tool to merge together separate video snippets
onemadgeek: Video To Frames Extractor
Extract frames from videos at custom frame rates
topazlabs: Video Upscale
Video Upscaling from Topaz Labs
csviverdeia: Locateanything 3B H100
LocateAnything-3B em H100 (Hopper) — visual grounding imagem+video (boxes/points). Detecção densa, pointing, tracker none/sort/reid. PRINCIPAL (rápido).
pixverse: Pixverse V5
Create 5s-8s videos with enhanced character movement, visual effects, and exclusive 1080p-8s support. Optimized for anime characters and complex actions
bitflow: Video Concat Safe Pro
When combining short AI videos encoded with different tools, codecs, and frame rates, this process standardizes them to ensure they are merged safely so that the final combined video remains intact.
zsxkib: Mmaudio
Add sound to video using the MMAudio V2 model. An advanced AI model that synthesizes high-quality audio from video content, enabling seamless video-to-audio transformation.
acappemin: Deepaudio V1
DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation
acappemin: Video To Audio And Piano
Enhance Generation Quality of Flow Matching V2A Model via Multi-Step CoT-Like Guidance and Combined Preference Optimization
ailoov: Video Rembg
Replicate
bytedance: Sa2Va 4B Video
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
colinhughes2121: Thumbnail Clickbait Enhancer
Boost any photo into a high-CTR YouTube thumbnail. Saturated colors, dramatic lighting, eye-catching contrast. For creators, video editors, content agencies. · api.gocreativeai.com — AI agent APIs
hcolde: Bgogo Feno
BiRefNet + Cutie video segmentation with stacked output video
jkukul: Minimax Remover
Remove any object from video - fast.
lucataco: Interactiveomni 8B
A unified omni-modal model that can simultaneously receive inputs such as images, audio, text, and video and directly generate coherent text and speech
lucataco: Motif Video
Motif-Video-2B: a 2B-parameter text-to-video diffusion transformer
prakhar-bhartiya: Meta Tribev2 Social Media Content Signal
Predicts virality of short videos using a Meta TRIBE v2 brain-encoding model
zsxkib: Create Video Dataset
Easily create video datasets with auto-captioning for Hunyuan-Video LoRA finetuning
zsxkib: Keep
🧼Upscales faces in videos look to be clearer and better using KEEP, Kalman-Inspired Feature Propagation for Video Face Super-Resolution
lucataco: Trim Video
Simple tool to quickly trim a video or audio file
idan054: Better Video Merge
Fix Diffrent Sizes for each clip. Fork of lucataco/cog-video-merge.git
lucataco: Qwen2.5 Omni 7B
Qwen2.5-Omni is an end-to-end multimodal model designed to perceive diverse modalities, including text, images, audio, and video, while simultaneously generating text and natural speech responses in a streaming manner.
zsxkib: Thinksound
Generate contextual audio from video using step-by-step reasoning🎶
ddvinh1: Video Faceswap Gpu
Replicate
atonamy: Wan Alpha
Wan-Alpha text-to-video generation with alpha channel, optimized for Replicate deployment. WebM/WebP output with transparency support.
jichengdu: Wan I2V
i2v-14B-720p-2.1
paullux: Framepack Runner
FramePack video generation with image + motion prompt. Based on Stanford's 2025 model.
subhash25rawat: Venhance
Enhance your video quality with AI
zsxkib: Memo
MEMO is a state-of-the-art open-weight model for audio-driven talking video generation.
biggpt1: Pana Video
Replicate
civicvideo: Billy Portrait
Replicate
deepfates: Hunyuan Blade Runner
Hunyuan-Video model finetuned on Blade Runner (1982). Trigger word is "BLDRN". Use "A video in the style of BLDRN, BLDRN" at the beginning of your prompt for best results.
deepfates: Hunyuan Blade Runner 2049
Hunyuan-Video model finetuned on Blade Runner 2049 (2017). Trigger word is "BLDRN". Use "A video in the style of BLDRN, BLDRN" at the beginning of your prompt for best results.
deepfates: Hunyuan Game Of Thrones
Hunyuan-Video model finetuned on Game of Thrones (2011). Trigger word is "GMFTH". Use "A video in the style of GMFTH, GMFTH" at the beginning of your prompt for best results
deepfates: Hunyuan Her
Hunyuan-Video model finetuned on Her (2013). Trigger word is "HR". Use "A video in the style of HR, HR" at the beginning of your prompt for best results.
deepfates: Hunyuan Inception
Hunyuan-Video model finetuned on Inception (2010). Trigger word is "NCPTN". Use "A video in the style of NCPTN, NCPTN" at the beginning of your prompt for best results.
deepfates: Hunyuan Indiana Jones
Hunyuan-Video model finetuned on Indiana Jones Series (1981). Trigger word is "NDNJN". Use "A video in the style of NDNJN, NDNJN" at the beginning of your prompt for best results.
deepfates: Hunyuan Joker
Hunyuan-Video model finetuned on Joker (2019). Trigger word is "JKR". Use "A video in the style of JKR, JKR" at the beginning of your prompt for best results.
deepfates: Hunyuan Once Upon A Time In Hollywood
Hunyuan-Video model finetuned on Once Upon a Time in Hollywood (2019). Trigger word is "NCPNT". Use "A video in the style of NCPNT, NCPNT" at the beginning of your prompt for best results.
deepfates: Hunyuan Rrr
Hunyuan-Video model finetuned on RRR (2022). Trigger word is "RRR". Use "A video in the style of RRR, RRR" at the beginning of your prompt for best results.
deepfates: Hunyuan The Lord Of The Rings
Hunyuan-Video model finetuned on The Lord of the Rings Trilogy (2001). Trigger word is "THLRD". Use "A video in the style of THLRD, THLRD" at the beginning of your prompt for best results.
deepfates: Hunyuan Westworld
Hunyuan-Video model finetuned on Westworld (2016). Trigger word is "WSTWR". Use "A video in the style of WSTWR, WSTWR" at the beginning of your prompt for best results.
derickson: Dn2Trl 1
Dune 2 Trailer LoRA for hunyuan-video-lora. Trigger word is DN2TRL
derickson: Neomacao 1
attempt at a video lora
dfischer: Tokmeshmetal 4 Epoch
Replicate
dfischer: Toksci 2 Epoch
Replicate
foxelli-group-video: Amigurumi V2
Replicate
itsheadkrack: Videome
Replicate
jimothyjohn: Superslomo
Slow down choppy videos up to 10x slower. It's basically 40x more slo mo for a small price..
lucataco: Longcat Video
Replicate
lucataco: Ltx Video Iclora
LTX Video 0.9.7 Distilled with ICLoRAs
lucataco: Stable Avatar
End-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing
mushroomfleet: Fxf Tokyo Meet
JDM style Car Meetup
mushroomfleet: Xjx Tokyoracer
Clasic JDM highway racing
nanoukader: Video
Replicate
papina: Seedvr2
🔥 SeedVR2: one-step video & image restoration with 7B and Adjustable Resolution
replicate: Ltxvideo 2.3 Lora
LTX 2.3 with popular community LoRA tasks like Ingredients or camera control
shridharathi: Ghibli Vid
Make a video of anything in Studio Ghibli style
shridharathi: Van Gogh Vid
Make your videos van gogh-esque
twinstarcreatives: Luma Portrait
Replicate
bytedance: Sa2Va 8B Video
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
sai88uk: Minicpm V 45 V9
A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding
andreasjansson: Video Stitcher
Fast GPU-powered concatenation of multiple videos, with short audio crossfades
luma: Reframe Image
Change the aspect ratio of any photo using AI (not cropping)
luma: Modify Video
Modify a video with style transfer and prompt-based editing
lucataco: Videollama3 7B
VideoLLaMA 3: Frontier Multimodal Foundation Models for Video Understanding
wan-video: Wan 2.2 5B Fast
The fastest Wan 2.2 text-to-image and image-to-video model
beancakes: Flux Cyberpunk People
Flux lora inspired by the video game, Cyberpunk 2077, for generating people in futuristic and dystopian backgrounds
lucataco: Nsfw Video Detection
FalconAIs NSFW detection model, extended for videos
bytedance: Sa2Va 26B Image
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
cjwbw: Controlvideo
Training-free Controllable Text-to-Video Generation
prunaai: P Video
Fast video generation with built-in draft mode for rapid creative iteration. Text-to-video, image-to-video, and audio-to-video in a single endpoint.
sprited: Birefnet Video
BiRefNet video background removal — every variant + ToonOut, applied per frame. Transparent WebM output by default (also MP4/MOV/GIF). MIT-licensed for commercial use.
andreasjansson: Wan 1.3B Inpaint
Inpainting and video2video experiments with Wan 2.1
lucataco: Sam3 Video
A unified foundation model for prompt-based segmentation in images and videos
lucataco: Hotshot Xl
😊 Hotshot-XL is an AI text-to-GIF model trained to work alongside Stable Diffusion XL
runwayml: Gen4 Image Turbo
Gen-4 Image Turbo is cheaper and 2.5x faster than Gen-4 Image. An image model with references, use up to 3 reference images to create the exact image you need. Capture every angle.
wan-video: Wan 2.7 Image
Generate and edit images with Alibaba's Wan 2.7
wan-video: Wan 2.7 Image Pro
Generate and edit high-quality images with Alibaba's Wan 2.7 Pro with 4K output, thinking mode, text-to-image, multi-image editing, and image set generation
bytedance: Sa2Va 4B Image
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
fofr: Kontext Ps1
FLUX Kontext fine-tune that let's you restyle any image as a PS1 or PS2 video game
uglyrobot: Sora2 Watermark Remover
Removes the watermark from Sora 2 videos using a trained model and IOpaint
meta: Sam 2 Video
SAM 2: Segment Anything v2 (for videos)
flux-kontext-apps: Restyle Video Frame
Use flux-kontext-pro to change the first or last frame of a video. Useful to use as inputs for restyling an entire video in a certain way
zsxkib: Bsrgan
Upscale videos + images with BSRGAN
arielreplicate: Robust Video Matting
extract foreground of a video
bytedance: Sa2Va 8B Image
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
cjwbw: Damo Text To Video
Multi-stage text-to-video generation
wan-video: Wan 2.2 I2V Fast
A very fast and cheap PrunaAI optimized version of Wan 2.2 A14B image-to-video
wan-video: Wan 2.2 T2V Fast
A very fast and cheap PrunaAI optimized version of Wan 2.2 A14B text-to-video
arielreplicate: Deoldify Video
Add colours to old video footage.
hjunior29: Video Text Generator
Generate animated captions and text overlays for social media videos
tencent: Hunyuanvideo Foley
Text-Video-to-Audio Synthesis: Generate realistic audio from video and text descriptions
pixverse: Pixverse V4.5
Quickly make 5s or 8s videos at 540p, 720p or 1080p. It has enhanced motion, prompt coherence and handles complex actions well.
bytedance: Sa2Va 26B Video
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
tmappdev: Change Video Bg
Change or Replace Video Background with any Image
fofr: Tooncrafter
Create videos from illustrated input images
lightricks: Ltx Video 0.9.7
DiT-based 13b video generation model, creating 30fps video
bytedance: Video Upscaler
Upscale and enhance video up to 4K at 60fps, with scene-aware presets for AI-generated content, short dramas, UGC, and film restoration.
kwaivgi: Kling Lip Sync
Add lip-sync to any video with an audio file or text
open-mmlab: Pia
Personalized Image Animator
zsxkib: Mmaudio T4
Cost-optimized MMAudio V2 (T4 GPU): Add sound to video using this version running on T4 hardware for lower cost. Synthesizes high-quality audio from video content.
bytedance: Seedance 1 Pro Fast
A faster and cheaper version of Seedance 1 Pro
zsxkib: Animate Diff
🎨 AnimateDiff (w/ MotionLoRAs for Panning, Zooming, etc): Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
anotherjesse: Zeroscope V2 Xl
Zeroscope V2 XL & 576w
lightricks: Ltx Video
LTX-Video is the first DiT-based video generation model capable of generating high-quality videos in real-time. It produces 24 FPS videos at a 768x512 resolution faster than they can be watched.
bytedance: Seedance 1 Lite
A video generation model that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 720p resolution
lightricks: Ltx 2 Distilled
LTX-2: The first open source audio-video model
lucataco: Animate Diff
Animate Your Personalized Text-to-Image Diffusion Models
minimax: Hailuo 02
Hailuo 2 is a text-to-video and image-to-video model that can make 6s or 10s videos at 768p (standard) or 1080p (pro). It excels at real world physics.
minimax: Hailuo 02 Fast
A low cost and fast version of Hailuo 02. Generate 6s and 10s videos in 512p
nvidia: Nemotron Nano V2 12B Vl
A multi-modal AI model for visual Q&A, summarization, and data extraction, supporting text, images, and video.
wan-video: Wan 2.2 Animate Replace
Use Wan 2.2 Animate to replace a character in a video scene
wan-video: Wan 2.2 S2V
Generate a video from an audio clip and a reference image
ayushunleashed: Minimax Remover
Remove any object from video - fast.
zsxkib: Seedvr2
🔥 SeedVR2: one-step video & image restoration with 3B/7B hot‑swap and optional color fix 🎬✨
andreasjansson: Tile Morph
Create tileable animations with seamless transitions
fofr: Video Morpher
Generate a video that morphs between subjects, with an optional style
zsxkib: Animatediff Illusions
Monster Labs' Controlnet QR Code Monster v2 For SD-1.5 on top of AnimateDiff Prompt Travel (Motion Module SD 1.5 v2)
bytedance: Seedance 1.5 Pro
A joint audio-video model that accurately follows complex instructions.
prunaai: P Video Avatar
p-video-avatar is the fastest and cheapest avatar/lipsync video model on the market.
arielreplicate: Stable Diffusion Infinite Zoom
Use Runway's Stable-diffusion inpainting model to create an infinite loop video
cjwbw: Videocrafter
VideoCrafter2: Text-to-Video and Image-to-Video Generation and Editing
lightricks: Ltx Video 0.9.7 Distilled
Faster slight quality reduction compared to LTX-Video 13b
cjwbw: Text2Video Zero
Text-to-Image Diffusion Models are Zero-Shot Video Generators
bytedance: Seedance 1 Pro
A pro version of Seedance that offers text-to-video and image-to-video support for 5s or 10s videos, at 480p and 1080p resolution
lightricks: Ltx 2.5 Fast
Fast video generation with text-to-video and image-to-video, portrait and landscape support, synchronized audio, and frame interpolation. Up to 20 seconds at 1080p, and 4K resolution.
luma: Ray 3.2
Luma's reasoning video model. Generates cinematic 5s or 10s video from text or images, with native HDR and EXR export for professional production pipelines.
prunaai: P Video Animate
p-video-animate animates a reference image with the motion and audio of a source video. Optimized for speed and cost — 5.24s per 1s of video.
prunaai: P Video Replace
p-video-replace swaps the person in a video with one from a reference image, keeping motion, timing, camera, and scene exactly as they were. 3.58s per 1s of video generated.
luma: Ray Flash 2 540P
Generate 5s and 9s 540p videos, faster and cheaper than Ray 2
heygen: Lipsync Speed
Fast lip-sync: replace or dub audio on any video with quick audio-driven lip sync
heygen: Video Agent
Turn a text prompt into a complete, polished video with AI-generated script, avatar presenter, voiceover, visuals, and editing.
heygen: Video Translate
Translate videos into over 150 languages
alibaba: Wan 3
Wan 3.0 generates video from a text prompt, with cinematic motion and support for 480p, 720p, and 1080p output.
bitflow: Video Super Resolution Rife Pro
Super video quality enhancement featuring fast upscaling with TensorRT and frame interpolation with RIFE.
bytedance: Seedance 2.0 Mini
A lower-cost variant of Seedance 2.0 for high-volume video generation with multimodal inputs and native audio.
character-ai: Ovi I2V
Ovi: generate videos with audio from image and text inputs
decart: Lucy Edit 2
Edit and transform videos with text prompts and reference images. Style transfers, object replacement, character transformation, and more.
lightricks: Ltx 2 Fast
Ideal for rapid ideation and mobile workflows. Perfect for creators who need instant feedback, real-time previews, or high-throughput content.
lucataco: Wan 2.1 1.3B Vid2Vid
Wan 2.1 1.3b Video to Video. Wan is a powerful visual generation model developed by Tongyi Lab of Alibaba Group
pixverse: Lipsync
Generate realistic lipsync animations from audio for high-quality synchronization
vidu: Q3 Turbo
Fast video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.
wan-video: Wan 2.1 1.3B
Generate 5s 480p videos. Wan is an advanced and powerful visual generation model developed by Tongyi Lab of Alibaba Group
ali-vilab: I2Vgen Xl
RESEARCH/NON-COMMERCIAL USE ONLY: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models
andreasjansson: Stable Diffusion Animation
Animate Stable Diffusion by interpolating between two prompts
bria: Video Erase Object
A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency
bytedance: Dreamactor M2.0
Animate any character, humans, cartoons, animals, even non-humans, from a single image + driving video
google: Veo 3.1 Lite
Google's cost-efficient video generation model with native audio, optimized for high-volume applications
kwaivgi: Kling V1.5 Standard
Generate 5s and 10s videos in 720p resolution at 30fps
kwaivgi: Kling V1.6 Standard
Generate 5s and 10s videos in 720p resolution at 30fps
kwaivgi: Kling V2.1
Use Kling v2.1 to generate 5s and 10s videos in 720p and 1080p resolution from a starting image (image-to-video)
pixverse: Pixverse V6
PixVerse's flagship video generation model. Generate cinematic videos with synchronized audio, multi-shot sequences, and precise camera control.
pollinations: Real Basicvsr Video Superresolution
RealBasicVSR: Investigating Tradeoffs in Real-World Video Super-Resolution
runwayml: Gen4 Turbo
Generate 5s and 10s 720p videos fast
wan-video: Wan 2.5 I2V
Alibaba Wan 2.5 Image to video generation with background audio
wan-video: Wan 2.5 T2V
Alibaba Wan 2.5 text to video generation model
wan-video: Wan2.6 I2V Flash
Image-to-video generation with optional audio, multi-shot narrative support, and faster inference
xai: Grok Imagine R2V
Generate videos guided by reference images using xAI's Grok Imagine Video model
xai: Grok Imagine Video
Generate videos using xAI's Grok Imagine Video model
xai: Grok Imagine Video Extension
Extend videos with xAI's Grok Imagine Video model. Provide a source video and describe what happens next.
zsxkib: Hunyuan Video Lora
Hunyuan-Video LoRA Explorer + Trainer
kwaivgi: Kling Avatar V2
Create avatar videos with realistic humans, animals, cartoons, or stylized characters
minimax: Hailuo 2.3
A high-fidelity video generation model optimized for realistic human motion, cinematic VFX, expressive characters, and strong prompt and style adherence across both text-to-video and image-to-video workflows
lucataco: Ltx Video 0.9.8 Distilled
Generate native long-form video, with controllability
elevenlabs: Dubbing
Translate audio and video into 90+ languages while preserving each speaker's voice, emotion, and timing
black-forest-labs: Flux 3
Generate video with synchronized audio from text, images, or video. FLUX 3 is Black Forest Labs' multimodal model (early access preview).
leonardoai: Motion 2.0
Create 5s 480p videos from a text prompt
lightricks: Ltx 2 Pro
Delivers high visual fidelity with fast turnaround. Great for daily content creation, marketing teams, and iterative creative workflows.
lightricks: Ltx 2.3 Fast
Lightning-fast video generation with portrait support, camera controls, and synchronized audio. Up to 20 seconds at 1080p, 4K at 50 FPS.
luma: Ray Flash 2 720P
Generate 5s and 9s 720p videos, faster and cheaper than Ray 2
luma: Reframe Video
Change the aspect ratio of any video up to 30 seconds long, outputs will be 720p
deforum: Deforum Stable Diffusion
Animating prompts with stable diffusion
heygen: Lipsync Precision
High-accuracy lip-sync: replace or dub audio on any video with avatar-inference lip sync
alibaba: Wan 3 Prime
Generate videos from text prompts using Alibaba's Wan 3.0 Prime model. Up to 1080p and 30 seconds, with 480p, 720p, and 1080p output.
wan-video: Wan 2.5 I2V Fast
Wan 2.5 image-to-video, optimized for speed
wan-video: Wan 2.5 T2V Fast
Wan 2.5 text-to-video, optimized for speed
bytedance: Seedance 2.0 Fast
A faster variant of Seedance 2.0 for quicker video generation with multimodal inputs and native audio.
kwaivgi: Kling V2.5 Turbo Pro
Kling 2.5 Turbo Pro: Unlock pro-level text-to-video and image-to-video creation with smooth motion, cinematic depth, and remarkable prompt adherence.
kwaivgi: Kling V2.6
Kling 2.6 Pro: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation
kwaivgi: Kling V2.6 Motion Control
Enables precise control of character actions and expressions from a reference image.
kwaivgi: Kling V3 Motion Control
Kling 3.0 motion control: transfer motion from a reference video to any character image with improved consistency and quality.
pixverse: Pixverse V5.6
Latest video model from Pixverse with astonishing physics
vidu: Q3 Pro
High-fidelity video generation with text-to-video, image-to-video, and start-end-to-video modes. Up to 16 seconds at 1080p with synchronized audio.
wavespeedai: Wan 2.1 T2V 480P
Accelerated inference for Wan 2.1 14B text to video, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.
heygen: Avatar Iv
Create realistic talking avatar videos from text with HeyGen's Avatar IV engine
heygen: Avatar V
Create realistic talking avatar videos from text with HeyGen's Avatar V engine — the newest, highest-quality avatar engine with cross-reference-driven animation.
bytedance: Seedance 2.0
ByteDance's multimodal video generation model with native audio, multimodal reference inputs, and intelligent duration control.
lightricks: Ltx 2.3 Pro
High-fidelity video generation with portrait support, audio-to-video, retake, and extend. Text, image, and audio-driven creation up to 4K at 50 FPS.
veed: Fabric 1.0
VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video
wan-video: Wan 2.2 I2V A14B
Image-to-video at 720p and 480p with Wan 2.2 A14B
xai: Grok Imagine Video 1.5
Image-to-video with synchronized audio using xAI's Grok Imagine Video 1.5 preview model
genmoai: Mochi 1
Mochi 1 preview is an open video generation model with high-fidelity motion and strong prompt adherence in preliminary evaluation
zsxkib: Pyramid Flow
Text-to-Video + Image-to-Video: Pyramid Flow Autoregressive Video Generation method based on Flow Matching
wavespeedai: Wan 2.1 I2V 480P
Accelerated inference for Wan 2.1 14B image to video, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.
lucataco: Wan2.1 I2V Lora
Wan2.1 14B 480p LoRA inference via Diffusers (Work in progress)
kwaivgi: Kling V1.5 Pro
Generate 5s and 10s videos in 1080p resolution at 30fps
kwaivgi: Kling V1.6 Pro
Generate 5s and 10s videos in 1080p resolution
wan-video: Wan 2.2 Animate Animation
Use Wan 2.2 Animate to copy the motion of a video to another scene
cuuupid: Cogvideox 5B
Generate high quality videos from a prompt
lightricks: Audio To Video
Use audio input with an image or prompt to generate videos
lightricks: Ltx 2 Retake
Take any shot and edit specific sections. Rephrase, change the action, camera angles and more
luma: Ray 2 540P
Generate 5s and 9s 540p videos
minimax: Video 01
Generate 6s videos with prompts or images. (Also known as Hailuo). Use a subject reference to make a video with a character and the S2V-01 model.
minimax: Video 01 Director
Generate videos with specific camera movements
minimax: Video 01 Live
An image-to-video (I2V) model specifically trained for Live2D and general animation use cases
openai: Sora 2
OpenAI's Flagship video generation with synced audio
philz1337x: Crystal Video Upscaler
High-precision video upscaler optimized for portraits, faces and products. One of the upscale modes powered by Clarity AI. X:https://x.com/philz1337x
wan-video: Wan 2.6 I2V
Alibaba Wan 2.6 image to video generation model
wan-video: Wan 2.6 T2V
Alibaba Wan 2.6 text to video generation model
wan-video: Wan 2.7 I2V
Generate videos from images, with support for first-and-last-frame control, clip continuation, and audio synchronization using Alibaba's Wan 2.7 model
wan-video: Wan 2.7 R2V
Generate videos from reference images or clips while preserving subject identity using Alibaba's Wan 2.7 reference-to-video model
wan-video: Wan 2.7 T2V
Generate videos with audio from text prompts using Alibaba's Wan 2.7 model. 1080p, up to 15 seconds, with audio synchronization.
wan-video: Wan 2.7 Videoedit
Edit videos with natural language instructions using Alibaba's Wan 2.7 VideoEdit model
zsxkib: Framepack
🕹️FramePack: video diffusion that feels like image diffusion🎥
lucataco: Real Esrgan Video
Real-ESRGAN Video Upscaler
zsxkib: Multitalk
Audio-driven multi-person conversational video generation - Upload audio files and a reference image to create realistic conversations between multiple people
deepfates: Hunyuan Spiderverse
Hunyuan-Video model finetuned on Spider-Man: Into the Spider-Verse (2018). Trigger word is "SPDRV". Use "A video in the style of SPDRV, SPDRV" at the beginning of your prompt for best results.
deepfates: Hunyuan The Matrix Trilogy
Hunyuan-Video model finetuned on The Matrix Trilogy (1999). Trigger word is "THMTR". Use "A video in the style of THMTR, THMTR" at the beginning of your prompt for best results.
zsxkib: Hunyuan Video2Video
A state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions
alibaba: Happyhorse 1.0
Alibaba's Happy Horse 1.0 generates videos from text prompts or animates a single image into video. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.
alibaba: Happyhorse 1.1
Alibaba's Happy Horse 1.1 generates videos from text, animates a single image, or builds a video from multiple reference images. Supports 720p and 1080p, 3-15 second durations, and five aspect ratios.
bytedance: Omni Human
Turns your audio/video/images into professional-quality animated videos
google: Veo 3 Fast
A faster and cheaper version of Google’s Veo 3 video model, with audio
google: Veo 3.1 Fast
New and improved version of Veo 3 Fast, with higher-fidelity video, context-aware audio and last frame support
bytedance: Omni Human 1.5
A film-grade digital human model that generates realistic video from a single image, audio clip, and optional text prompt.
kwaivgi: Kling V3 Omni Video
Kling Video 3.0 Omni: Unified multimodal video generation with reference images, video editing, native audio, and multi-shot control
kwaivgi: Kling V3 Video
Kling Video 3.0: Generate cinematic videos up to 15 seconds with multi-shot control, native audio, and improved consistency
luma: Ray 2 720P
Generate 5s and 9s 720p videos
wavespeedai: Hunyuan Video Fast
Accelerated inference for HunyuanVideo with high resolution (1280x720), a state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions
bytedance: Seedance 2.5
ByteDance's flagship multimodal video model with native audio, native 30-second generation, and large multimodal reference sets.
wavespeedai: Wan 2.1 T2V 720P
Accelerated inference for Wan 2.1 14B text to video with high resolution, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.
deepfates: Hunyuan Arcane
Hunyuan-Video model finetuned on Arcane (2021). Trigger word is "RCN". Use "A video in the style of RCN, RCN" at the beginning of your prompt for best results.
wavespeedai: Wan 2.1 I2V 720P
Accelerated inference for Wan 2.1 14B image to video with high resolution, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation.
wan-video: Wan2.1 With Lora
Run Wan2.1 14b or 1.3b with a lora
deepfates: Hunyuan Dune
Hunyuan-Video model finetuned on Dune (2021). Trigger word is "DN". Use "A video in the style of DN, DN" at the beginning of your prompt for best results.
kwaivgi: Kling V2.0
Generate 5s and 10s videos in 720p resolution
kwaivgi: Kling V2.1 Master
A premium version of Kling v2.1 with superb dynamics and prompt adherence. Generate 1080p 5s and 10s videos from text or an image
zsxkib: Star
STAR Video Upscaler: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-Resolution
openai: Sora 2 Pro
OpenAI's Most advanced synced-audio video generation
Bytedance Seedance 2.5
KIE
zsxkib: Animatediff Prompt Travel
🎨AnimateDiff Prompt Travel🧭 Seamlessly Navigate and Animate Between Text-to-Image Prompts for Dynamic Visual Narratives
runwayml: Aleph 2
Edit one frame to update an entire video. Aleph 2.0 is Runway's in-context video editor: longer clips (up to 30s), multi-shot edits, and image-level precision via keyframe references.
google: Veo 3
Sound on: Google’s flagship Veo 3 text to video model, with audio
google: Veo 3.1
New and improved version of Veo 3, with higher-fidelity video, context-aware audio, reference image and last frame support
deepfates: Hunyuan Cowboy Bebop
Hunyuan-Video model finetuned on Cowboy Bebop (1998). Trigger word is "CWBYB". Use "A video in the style of CWBYB, CWBYB" at the beginning of your prompt for best results.
deepfates: Hunyuan Pixar
Hunyuan-Video model finetuned on Pixar Films (1995). Trigger word is "PXR". Use "A video in the style of PXR, PXR" at the beginning of your prompt for best results.
google: Veo 2
State of the art video generation model. Veo 2 can faithfully follow simple and complex instructions, and convincingly simulates real-world physics as well as a wide range of visual styles.
tencent: Hunyuan Video
A state-of-the-art text-to-video generation model capable of creating high-quality videos with realistic motion from text descriptions
zsxkib: Step Video T2V
Generate high-quality videos from text prompts using StepVideo
deepfates: Hunyuan The Grand Budapest Hotel
Hunyuan-Video model finetuned on The Grand Budapest Hotel (2014). Trigger word is "THGRN". Use "A video in the style of THGRN, THGRN" at the beginning of your prompt for best results.
deepfates: Hunyuan Spider Man Into The Spider Verse
Hunyuan-Video model finetuned on Spider-Man: Into the Spider-Verse (2018). Trigger word is "SPDRM". Use "A video in the style of SPDRM, SPDRM" at the beginning of your prompt for best results.
deepfates: Hunyuan La La Land
Hunyuan-Video model finetuned on La La Land (2016). Trigger word is "LLLND". Use "A video in the style of LLLND, LLLND" at the beginning of your prompt for best results.
deepfates: Hunyuan Twin Peaks
Hunyuan-Video model finetuned on Twin Peaks (1990). Trigger word is "TWNPK". Use "A video in the style of TWNPK, TWNPK" at the beginning of your prompt for best results.
deepfates: Hunyuan Pulp Fiction
Hunyuan-Video model finetuned on Pulp Fiction (1994). Trigger word is "PLPFC". Use "A video in the style of PLPFC, PLPFC" at the beginning of your prompt for best results.
موسیقی
آهنگسازی و تولید موسیقی هوشمند
10 مدل فعال
google: Lyria 2
Lyria 2 is a music generation model that produces 48kHz stereo audio through text-based prompts
minimax: Music 1.5
Music-1.5: Full-length songs (up to 4 mins) with natural vocals & rich instrumentation
minimax: Music 01
Quickly generate up to 1 minute of music with lyrics and vocals in the style of a reference track
google: Lyria 3
Generate 30-second music clips from text prompts or images with Lyria 3, Google's music generation model
elevenlabs: Music
Compose a song from a prompt or a composition plan
google: Lyria 3 Pro
Generate full-length songs up to 3 minutes from text prompts or images with Lyria 3 Pro, Google's most capable music generation model
minimax: Music 2.5
Generate full-length songs with vocals, lyrics, and rich instrumentation from a text prompt
minimax: Music 2.6
Generate full-length songs or instrumentals from a text prompt, with optional auto-generated lyrics
stability-ai: Stable Audio 2.5
Generate high-quality music and sound from text prompts
minimax: Music Cover
Reimagine any song in a different style — change voice, instruments, genre, and arrangement while keeping the original melody
صدا و گفتار
تبدیل متن به گفتار و رونویسی صدا
25 مدل فعال
ibm-granite: Granite Speech 4.1 2B
Granite Speech 4.1 2B is a compact and efficient speech-language model, specifically designed for multilingual automatic speech recognition (ASR) and bidirectional automatic speech translation (AST) for English, French, German, Spanish, Portuguese and Jap
elevenlabs: Flash V2.5
ElevenLabs's fastest speech synthesis model
elevenlabs: V3
The most expressive Text to Speech model
inworld: Realtime Tts 1.5 Mini
Ultra-fast, cost-efficient realtime text-to-speech with ~120ms latency and 15-language support
inworld: Realtime Tts 2
Most expressive text-to-speech model from Inworld, with natural-language steering, real-time latency, and multilingual support across 100+ languages.
inworld: Realtime Tts 1.5 Max
Highest-quality realtime text-to-speech with <200ms latency, emotion control, and 15-language support
elevenlabs: Turbo V2.5
High quality, low latency text to speech in 32 languages
elevenlabs: Scribe V2
Transcribe speech with ElevenLabs Scribe v2. 90+ languages, word-level timestamps, speaker diarization for up to 32 speakers, audio event tagging, and keyterm biasing. Files up to 3 GB and 10 hours.
resemble-ai: Chatterbox Pro
Generate expressive, natural speech with Resemble AI's Chatterbox.
elevenlabs: V2 Multilingual
Generate multilingual text-to-speech audio in over 30 languages
xai: Grok Text To Speech
Convert text to natural-sounding speech with xAI's Grok TTS. 5 voices, 20 languages, expressive speech tags, and high-fidelity MP3 / WAV / telephony audio output.
ibm-granite: Granite Speech 3.3 8B
Granite-speech-3.3-8b is a compact and efficient speech-language model, specifically designed for automatic speech recognition (ASR) and automatic speech translation (AST).
datalab-to: Ocr
Detect and transcribe text in images with accurate bounding boxes, layout analysis, reding order, and table recognition, in 90 languages
openai: Gpt 4O Mini Transcribe
A speech-to-text model that uses GPT-4o mini to transcribe audio
xai: Grok Speech To Text
Transcribe audio to text with xAI's Grok. Handles 25 languages, word-level timestamps, speaker diarization, multichannel audio, and files up to 500 MB.
resemble-ai: Chatterbox Turbo
The fastest open source TTS model without sacrificing quality.
openai: Gpt 4O Transcribe
A speech-to-text model that uses GPT-4o to transcribe audio
minimax: Speech 02 Turbo
Text-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Designed for real-time applications with low latency
minimax: Speech 2.6 Turbo
Low‑latency MiniMax Speech 2.6 Turbo brings multilingual, emotional text-to-speech to Replicate with 300+ voices and real-time friendly pricing
minimax: Speech 2.8 Turbo
Minimax Speech 2.8 Turbo: Turn text into natural, expressive speech with voice cloning, emotion control, and support for 40+ languages
minimax: Speech 02 Hd
Text-to-Audio (T2A) that offers voice synthesis, emotional expression, and multilingual capabilities. Optimized for high-fidelity applications like voiceovers and audiobooks.
minimax: Speech 2.6 Hd
MiniMax Speech 2.6 HD delivers studio-quality multilingual text-to-audio on Replicate with nuanced prosody, subtitle export, and premium voices
minimax: Speech 2.8 Hd
Minimax Speech 2.8 HD focuses on high-fidelity audio generation with features like studio-grade quality, flexible emotion control, multilingual support, and voice cloning capabilities
google: Gemini 3.1 Flash Tts
Google's fast, expressive text-to-speech model with 30 voices and 70+ language support
resemble-ai: Chatterbox
Generate expressive, natural speech. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.
مدلساز ۳بعدی
ساخت مدل سهبعدی از متن یا تصویر
36 مدل فعال
vufinder: Map Anything
Universal Feed-Forward Metric 3D Reconstruction
vufinder: Vggt 1B Depth
Feed-forward neural network that directly infers all key 3D attributes of a scene.
colinhughes2121: Ai Avatar 3D
Turn any photo into a 3D-rendered character avatar. Roblox/Fortnite/Discord-style. For gamers, creators, virtual worlds, profile pics. · api.gocreativeai.com — AI agent APIs
colinhughes2121: Pixar Character Stylizer
Convert any photo into a 3D animated Pixar-style character. Big expressive eyes, soft rendering, family-movie aesthetic. · api.gocreativeai.com — AI agent APIs
fire: Trellis
For the paper "Structured 3D Latents for Scalable and Versatile 3D Generation".
fishwowater: Trellis Image
SOTA image-to-3D generator TRELLIS. Equipped with **All image-condition types** you need: (1) single-image (2) multi-images (3) different images for geometry and texture generation (4) mesh+images for detail variation(texture painting)
hcl14: Direct3D S2
Direct3D‑S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention
hcl14: Omnipart
OmniPart: Part-Aware 3D Generation with Semantic Decoupling and Structural Cohesion
subhash25rawat: Morphix3D
Transform Images & Text into 3D Models with AI
vufinder: Map Anything Pano
Universal Feed-Forward Metric 3D Reconstruction from panoramic images
vufinder: Map Anything Pi3
A feed-forward neural network that offers a novel approach to visual geometry reconstruction.
vufinder: Map Anything Pi3X
A feed-forward neural network that offers a novel approach to visual geometry reconstruction.
vufinder: Map Anything Vggt 1B
Feed-forward neural network that directly infers all key 3D attributes of a scene.
wanjon: Ardy2Vrm
Generate text-driven 3D motion with NVIDIA ARDY and export animated previews, NPZ data, and VRM Animation files for Three.js and Three-VRM.
zylim0702: Yun 3D 2.1
Replicate
ayushunleashed: Partpacker
Part-level 3D object generation from single-view images
fire: Part Crafter
PartCrafter is a structured 3D mesh generation model that creates multiple parts and objects from a single RGB image.
hcl14: Hunyuan P3 Sam
Replicate
aliadhami: 3D Sheet
Generates minimal, consistent 3-view design for objects in the style of technical 3D reference sheets. Outputs include front, back, and side views with consistent proportions against a white background.
valllllex: Cartoonia 3D
Restyling the classic cartoon and comic book characters into a soft painterly 3D style. FLUX Kontext Lora
vufinder: Vggt 1B
Feed-forward neural network that directly infers all key 3D attributes of a scene.
firtoz: Trellis
A powerful 3D asset generation model
vufinder: Vggt 1B Point
Feed-forward neural network that directly infers all key 3D attributes of a scene.
kfarr: Sharp Ml
Apple's SHARP model — single image to 3D Gaussian splats
tencent: Hunyuan3D 2
Scaling Diffusion Models for High Resolution Textured 3D Assets Generation
tencent: Hunyuan3D 2Mv
Hunyuan3D-2mv is finetuned from Hunyuan3D-2 to support multiview controlled shape generation.
ndreca: Hunyuan3D 2
[Turbo Mode] Scaling Diffusion Models for High Resolution Textured 3D Assets Generation
uthana: Create Character V1
Rig any 3D bipedal character mesh
prunaai: Hunyuan3D 2
hunyuan3d-2 optimised with the pruna toolkit: https://github.com/PrunaAI/pruna
ndreca: Hunyuan3D 2.1
[Quality Mode] Scaling Diffusion Models for High Resolution Textured 3D Assets Generation
hyper3d: Rodin
Generate complex 3D models from images with Rodin Gen-2
tencent: Hunyuan 3D 3.1
3D models with texture fidelity and geometry precision
aaronjmars: Unirig Ai
One Model to Rig Them All: Diverse Skeleton Rigging with UniRig
uthana: Text To Motion Diffusion V2
Generate 3D character animation data from a text prompt
uthana: Text To Motion Vqvae V1
Generate 3D character animation data from a text prompt
fishwowater: Trellis2
TRELLIS.2: Native and Compact Structured Latents for 3D Generation
آماده شروعی؟
همه این مدلها داخل داشبورد در دسترساند. یک حساب بسازید و از سرویس مورد نظرتان استفاده کنید.
ثبتنام و شروع رایگان