# AGENTS
Source: https://docs.imagine.art/AGENTS
> **First-time setup**: Customize this file for your project. Prompt the user to customize this file for their project.
> For Mintlify product knowledge (components, configuration, writing standards),
> install the Mintlify skill: `npx skills add https://mintlify.com/docs`
# Documentation project instructions
## About this project
* This is a documentation site built on [Mintlify](https://mintlify.com)
* Pages are MDX files with YAML frontmatter
* Configuration lives in `docs.json`
* Run `mint dev` to preview locally
* Run `mint broken-links` to check links
## Terminology
## Style preferences
* Use active voice and second person ("you")
* Keep sentences concise — one idea per sentence
* Use sentence case for headings
* Bold for UI elements: Click **Settings**
* Code formatting for file names, commands, paths, and code references
## Content boundaries
# Cancel Subscription
Source: https://docs.imagine.art/account/cancel-subscription
How to cancel your ImagineArt subscription and what happens to your credits and access afterward.
You can cancel your ImagineArt subscription at any time. Cancellation stops future billing but does not immediately remove your access.
Cancelling your subscription removes access to all paid features — including Pro models and team collaboration — at the end of your current billing period. Make sure to export or save any work you want to keep before your access changes.
## How to cancel
Go to [imagine.art](https://www.imagine.art) and sign in.
Click your **profile picture** in the top-right corner and select **Billing & Subscription**.
Scroll to your active plan and click **Cancel subscription**. You may be shown a confirmation prompt or a brief retention flow — read through it and confirm your cancellation.
You should receive a confirmation email at your account email address. Your subscription status in **Billing & Subscription** will update to show the date your paid access ends.
## What happens after you cancel?
### Access and features
Your paid plan remains active through the **end of your current billing period**. During this time you continue to have access to all paid features, including Pro models and your current team seats.
After the billing period ends, your account reverts to the Free plan:
* You receive **100 free credits per day** (no monthly subscription credits)
* **Pro models** are no longer accessible
* Team collaboration features and additional seats are removed
### Subscription credits
Unused subscription credits are **forfeited** when your paid period ends. Subscription credits do not roll over and cannot be refunded.
### Top-up credits
Top-up credits never expire. If you purchased top-up credits (also called Tokens in the app), they remain on your account after cancellation. You can spend them on features available under the Free plan.
## Resubscribing
You can resubscribe at any time by going to **Billing & Subscription** and selecting a new plan. Your previous content and settings are preserved on your account.
## Need help?
If you have trouble cancelling or have a question about your billing, contact [billing@imagine.art](mailto:billing@imagine.art) with your account email and a description of the issue.
## Related
* [Subscription plans](/account/subscription-plans)
* [Refund policy](/policies/refund-policy)
* [Update billing information](/account/updating-billing)
# Delete Account
Source: https://docs.imagine.art/account/delete-account
How to submit an ImagineArt account deletion request and what happens to your data.
If you want to permanently delete your ImagineArt account, you can submit a deletion request through the app or by contacting support. Account deletion is irreversible.
Account deletion is permanent. Once your account is deleted, all your data — including your gallery, generation history, saved settings, and any remaining credits — is removed and cannot be recovered. Make sure to download any content you want to keep before submitting a request.
## How to request account deletion
Before submitting a deletion request, download any images or videos from your gallery that you want to keep. After deletion, this content cannot be retrieved.
If you have an active paid subscription, cancel it first to stop future billing. You can do this from **Billing & Subscription** in your account settings. See [Cancel subscription](/account/cancel-subscription) for step-by-step instructions.
You can request deletion in one of two ways:
**From within the app:**
1. Click your **profile picture** in the top-right corner.
2. Go to **Account Settings**.
3. Scroll to the **Danger zone** or **Delete account** section and follow the prompts.
**By email:**
Email [support@imagine.art](mailto:support@imagine.art) from the address associated with your account. Include the subject line "Account deletion request" and confirm that you want your account and all associated data permanently deleted.
For security, ImagineArt may ask you to verify your identity before processing the deletion. Respond promptly to any verification request from the support team.
Account deletion is typically processed within a few business days. You will receive an email confirmation when your account has been deleted.
## What data is deleted?
When your account is deleted, ImagineArt removes:
* Your profile information (name, email, sign-in connections)
* Your gallery and all generated content stored on the platform
* Your generation history and usage analytics
* Any saved prompts, preferences, and settings
* Remaining subscription credits and top-up credits
## Data that may be retained
ImagineArt may retain certain data for a limited period as required by law or legitimate business purposes, such as billing records needed for tax compliance. This is detailed in the [Privacy Policy](/policies/privacy-policy).
## Account Deletion Policy
For the full policy governing account deletion, including data retention timelines, see the [Account Deletion Policy](https://www.imagine.art) on the ImagineArt website.
## Related
* [Privacy policy](/policies/privacy-policy)
* [Cancel subscription](/account/cancel-subscription)
# Mobile App
Source: https://docs.imagine.art/account/mobile-app
Download and use ImagineArt on iOS and Android to create images and videos from your phone or tablet.
ImagineArt is available on both iOS and Android. This page covers where to download the app, how to get started, and how to fix common issues on mobile.
## Download the app
| Platform | Link | Requirements |
| ----------- | ----------------------------------------------------------------------------------------------------- | -------------------- |
| **iOS** | [Download on the App Store](https://apps.apple.com/us/app/imagineart-ai-video-generator/id1664121419) | iOS 18.2 or later |
| **Android** | [Get it on Google Play](https://play.google.com/store/search?q=imagineart\&c=apps) | Android 8.0 or later |
## What you can do on mobile
The ImagineArt mobile app gives you access to the core creative tools:
* **Generate images** — write a prompt and generate images using the same models available on web
* **Generate videos** — create videos from text or from an existing image
* **Browse your gallery** — view, download, and share your past generations
* **Manage credits** — check your credit balance and purchase top-ups
* **Account settings** — update your profile and manage your subscription
Some advanced features — such as Workflows (Imagine Flow) and certain editing tools — are currently optimized for the web experience at [imagine.art](https://www.imagine.art). Check the app for the latest list of supported features, as new tools are added regularly.
## Signing in on mobile
The mobile app supports the same sign-in methods as the web:
* Google
* Facebook
* Discord
* Email and password
Your account, credits, and gallery are shared across web and mobile — sign in with the same account on both platforms to access your content everywhere.
## Credits on mobile
Your credit balance is shared between web and mobile. Credits you earn or spend on one platform are reflected on the other in real time. If you need to buy top-up credits (Tokens), you can do so from within the mobile app.
## Common issues on iOS
**My plan hasn't activated after subscribing**
This is usually an App Store sync delay. Go to Settings → your name → Subscriptions and confirm ImagineArt shows as active. Then open the app and look for a "Restore purchases" option. If it still doesn't activate, contact us at [support@imagine.art](mailto:support@imagine.art) with your Apple receipt.
**The app is crashing on launch**
Force-close the app and reopen it. If the issue continues, uninstall and reinstall from the App Store. If it keeps crashing after a reinstall, reach out to [support@imagine.art](mailto:support@imagine.art) with your device model and iOS version.
**My images aren't saving to my gallery**
Go to Settings → ImagineArt → Photos and make sure photo library access is enabled. If it's already on, a reinstall usually fixes it.
**How do I cancel my iOS subscription?**
Go to Settings → your name → Subscriptions → ImagineArt → Cancel Subscription. Cancellations apply to the next billing cycle. For billing disputes, contact Apple Support directly.
## Common issues on Android
**Google Play charged me but my plan isn't active**
Open the Play Store, go to Subscriptions, and confirm ImagineArt is listed as active. Then reopen the app. If the plan still doesn't reflect after a few minutes, email us at [support@imagine.art](mailto:support@imagine.art) with your Google Play Order ID.
**The app is stuck on the loading screen**
Go to Settings → Apps → ImagineArt → Storage → Clear Cache, then reopen the app. If that doesn't work, try a full reinstall from the Play Store. If you're using a VPN, try disabling it as this can sometimes interfere with the login process.
**I'm stuck in a login loop**
This is usually caused by an expired session token. Fully log out, clear the app cache (Settings → Apps → ImagineArt → Storage → Clear Cache), then log back in. Disabling any active VPN before logging in can also help.
**How do I cancel my Android subscription?**
Open the Play Store, tap your profile icon, then go to Payments & subscriptions → Subscriptions → ImagineArt → Cancel subscription. Cancellations apply to the next billing cycle. For billing disputes, contact Google Support directly.
## Issues on both platforms
**Credits were deducted but no image was generated**
Check your generation history in the app — failed or interrupted generations can still deduct credits. If the history doesn't account for the drop, email [support@imagine.art](mailto:support@imagine.art) with your account details and we'll look into it.
**My web subscription isn't showing on the app (or vice versa)**
Mobile and web subscriptions are now synchronized. To link them, log in at imagine.art using the same email address associated with your Google Play Store or App Store account and your plan will reflect automatically. If it still isn't showing after logging in, contact us at [support@imagine.art](mailto:support@imagine.art) and we'll get it sorted.
Still having trouble? Contact us at **[support@imagine.art](mailto:support@imagine.art)** and include your device model, OS version, and a description of the issue.
# Password Recovery
Source: https://docs.imagine.art/account/password-recovery
How to recover your ImagineArt account if you forgot your password or can't remember which sign-in method you used.
ImagineArt supports four sign-in methods: **Google**, **Facebook**, **Discord**, and **Email**. If you are locked out of your account, follow the guidance below for your method.
## Forgot your password (email sign-in)
Visit [imagine.art](https://www.imagine.art) and click **Sign in**.
On the email sign-in form, click the **Forgot password?** link below the password field.
Type the email address associated with your ImagineArt account and click **Send reset link**.
You will receive a password reset email within a few minutes. Check your spam or junk folder if it doesn't arrive promptly.
Click the link in the email. You will be taken to a page where you can enter and confirm a new password. The link expires after a short time, so complete this step promptly.
Return to [imagine.art](https://www.imagine.art) and sign in using your email and new password.
## Not sure which sign-in method you used?
If you can't sign in, you may be trying the wrong method. Here are tips for each option:
Try signing in with each social option one at a time. ImagineArt will tell you if no account is linked to a given method, which helps you narrow down which one you originally used.
### Google
Click **Continue with Google** and select the Google account you used when you signed up. If you have multiple Google accounts, try each one.
### Facebook
Click **Continue with Facebook** and log in with your Facebook credentials. Make sure you are using the same Facebook account you connected at registration.
### Discord
Click **Continue with Discord**. You will be redirected to Discord to authorize the connection. Ensure you are signed in to the correct Discord account.
### Email
If you registered with an email and password, use the **Forgot password?** flow described above. Check all email addresses you might have used when signing up.
If you signed up with a social method (Google, Facebook, or Discord), there is no password to reset — your sign-in is handled entirely by that provider. Use the provider's own account recovery flow if you are locked out of the social account itself.
## Still can't access your account?
If none of the above works, email [support@imagine.art](mailto:support@imagine.art) with:
* The email address you believe is associated with your account
* The sign-in method you think you used
* Any additional context (when you last successfully signed in, etc.)
The support team will help you recover access.
## Related
* [Troubleshooting your account](/account/troubleshooting)
* [Account deletion](/account/account-deletion)
# Sign Up and Sign In
Source: https://docs.imagine.art/account/sign-up-sign-in
How to create an ImagineArt account and sign in — covering all four authentication methods, troubleshooting, and support contacts.
ImagineArt supports four authentication methods: **Google**, **Discord**, **Email**, and **Facebook**. You must use the same method every time — accounts are tied to the method you chose at registration.
Use **Continue with Google** for the easiest and quickest access. Google sign-in requires no password to manage and is the most commonly used method.
## Creating a new account
Visit [imagine.art](https://www.imagine.art) and click **Sign Up**.
Select one of the four options below. Your choice here becomes your permanent sign-in method.
* **Google** — Links your existing Google account. Requires a verified Google account with accessible credentials.
* **Discord** — Links your Discord account. Your Discord login credentials will serve as your ImagineArt credentials. Account must be verified.
* **Email** — Creates a traditional account with a unique password. A verification email will be sent to confirm your address.
* **Facebook** — Links your Facebook account. You will use the same Facebook credentials to sign in to ImagineArt.
Follow the prompts for your chosen method. For Email sign-up, check your inbox (and spam folder) for a verification email and click the link to confirm your address.
Once verified, you are signed in and ready to use ImagineArt. New accounts receive **100 free credits** that refresh every 24 hours.
## Signing in to an existing account
Visit [imagine.art](https://www.imagine.art) and click **Log In**.
Choose the same method you used when you created your account:
* **Google** — Select your Google account and authorize access.
* **Discord** — Enter your Discord credentials and authorize platform access.
* **Email** — Enter your email address and password.
* **Facebook** — Enter your Facebook credentials and authorize access.
## Troubleshooting
Try each method one at a time — Google, Facebook, Discord, then Email. ImagineArt will indicate when no account is linked to a given method, which helps you narrow it down. Google is the most commonly used method and is a good starting point.
This means an account with your email is already registered. Instead of signing up again, click **Log In** and use the method you originally registered with.
1. Check your spam or junk folder.
2. Verify you entered the correct email address.
3. Wait a few minutes, then request a new verification email.
4. If the email still doesn't arrive, contact [support@imagine.art](mailto:support@imagine.art).
* Confirm you are signed in to the correct Discord account.
* Check that your Discord account email is verified.
* Clear your browser cache or try an incognito window.
* Re-authorize the connection between Discord and ImagineArt.
* Confirm you are selecting the correct Google account (you may have multiple).
* Disable any pop-up blockers that might prevent the Google authorization window from opening.
* Try signing in from an incognito window with extensions disabled.
* Clear your browser cache and cookies, then try again.
On the sign-in page, click the **Forgot password?** link. Enter your email address and follow the instructions in the reset email. If you signed up with Google, Facebook, or Discord, there is no ImagineArt password — use your provider's own account recovery flow instead.
See [Password Recovery](/account/password-recovery) for the full step-by-step process.
## Contact support
| Issue | Contact |
| -------------------------------- | ----------------------------------------------------------------- |
| Account access, sign-in problems | [support@imagine.art](mailto:support@imagine.art) |
| Billing and payment issues | [billing.support@imagine.art](mailto:billing.support@imagine.art) |
| Community help | [ImagineArt Discord server](https://discord.gg/imagineart) |
## Related
* [Password Recovery](/account/password-recovery)
* [Subscription Plans](/account/subscription-plans)
* [Troubleshooting your account](/account/troubleshooting)
# Subscription Plans
Source: https://docs.imagine.art/account/subscription-plans
Compare ImagineArt's Free, Basic, Standard, Ultimate, Creator, and Enterprise plans, pricing, credits, billing options, and features.
ImagineArt uses **credits** as the currency of the platform. You spend credits to explore different tools, image, video, workflows, apps, and more. All paid plans are available for **Individuals** and **Teams**, and you can choose **Monthly**, **Quarterly**, or **Yearly** billing.
Discounts and promotions can change over time. Always confirm the final prices on the live [Subscription page](https://www.imagine.art/subscription).
## Credits & Refresh Rules
Credits are assigned monthly and **never roll over**.
* Credits refresh **every month**, regardless of which billing term you're on.
* At the **end of each month**, unused credits **expire** and are replaced by a fresh allotment.
* If you run short during the month, you can **upgrade your plan** or **buy top-up credits**.
### Billing term vs. credit refresh
| Billing term | How you're charged | Credits |
| ------------- | --------------------------- | ------------------------------------------------------ |
| **Monthly** | Billed every month | Refresh each month; unused credits expire at month-end |
| **Quarterly** | Billed once every 3 months | Still refresh monthly — 3 refreshes per quarter |
| **Yearly** | Billed once every 12 months | Still refresh monthly — 12 refreshes per year |
***
## Plan Overview
**100 credits per day**, refreshed daily.
* Access to standard models
* No premium or Pro models
* Great for trying the platform
No payment required.
**\$13 / month** — first step beyond Free.
* **3,000 credits** per month
* Access to premium models
* Faster generation
**\$30 / month** — most popular plan.
* **8,000 credits** per month
* Premium models
* **3 team seats**
* Private generations, prompt enhancer, upscale, 1080p exports
**\$50 / month** — for power users.
* **16,000 credits** per month
* Unlimited video and image generation
* **6 team seats**
* 4K video, up to 14 image references
* Priority queue
**\$250 / month** — for larger creative teams.
* **100,000 credits** per month
* Unlimited generation, unlimited personalizations
* **20 team seats**
* 5 concurrent video generations
* Highest priority queue
**Custom pricing** for organizations.
* Custom credit allocation
* Custom seat count
* SSO, SLAs, and tailored configuration
[Contact sales](https://www.imagine.art/subscription) to get started.
***
## Pricing Tables
### Monthly
| Plan | Price / Month | Credits / Month | Image Gens\* | Video Gens\* | Team Seats | Priority Queue |
| -------------- | ------------- | --------------- | ------------ | ------------ | ---------- | -------------- |
| **Free** | Free | 100 / day | Limited | Limited | — | No |
| **Basic** | \$13 | 3,000 | \~600 / mo | \~97 / mo | 1 | No |
| **Standard** | \$30 | 8,000 | \~1,000 / mo | \~125 / mo | 3 | No |
| **Ultimate** | \$50 | 16,000 | \~3,000 / mo | \~375 / mo | 6 | Yes |
| **Creator** | \$250 | 100,000 | \~8,000 / mo | \~1,000 / mo | 20 | Highest |
| **Enterprise** | Custom | Custom | Custom | Custom | Custom | Custom |
### Quarterly (15% Discount)
| Plan | Effective Price / Month | Billed Quarterly | Credits / Month |
| ------------ | ----------------------- | ---------------- | --------------- |
| **Basic** | \$11 | \$33 | 3,000 |
| **Standard** | \$25 | \$75 | 8,000 |
| **Ultimate** | \$41 | \$125 | 16,000 |
| **Creator** | \$213 | \$640 | 100,000 |
### Yearly (30% Discount)
| Plan | Effective Price / Month | Billed Yearly | Credits / Month |
| ------------ | ----------------------- | ------------- | --------------- |
| **Basic** | \$9 | \$108 | 3,000 |
| **Standard** | \$20 | \$240 | 8,000 |
| **Ultimate** | \$34 | \$410 | 16,000 |
| **Creator** | \$175 | \$2,100 | 100,000 |
Credits **still refresh monthly** even when billed quarterly or yearly. The 15% and 30% discounts may vary by location and are subject to change — confirm current pricing on the [Subscription page](https://www.imagine.art/subscription).
*\*Approximate usage based on typical credit costs; actual numbers vary by model and settings.*
***
## Features by Plan
| Feature | Free | Basic | Standard | Ultimate | Creator |
| ------------------------------ | ------- | ------- | -------- | ------------ | ------------ |
| **Monthly credits** | 100/day | 3,000 | 8,000 | 16,000 | 100,000 |
| **Premium models** | No | Yes | Yes | Yes | Yes |
| **Veo models** | No | No | No | Yes | Yes |
| **Private generations** | No | No | Yes | Yes | Yes |
| **Video length** | — | 3–6s | 3–10s | 3–20s | 3–20s |
| **Video resolution** | — | 768p | 1080p | 4K | 4K |
| **Concurrent image gens** | 1 | 4 | 8 | 12 | 16 |
| **Concurrent video gens** | 1 | 2 | 3 | 4 | 5 |
| **Unlimited video generation** | No | No | No | Yes | Yes |
| **Unlimited image generation** | No | No | No | Yes | Yes |
| **Priority queue** | No | No | No | Yes | Highest |
| **Prompt Enhancer** | No | No | Yes | Yes | Yes |
| **Upscale** | No | Yes | Yes | Yes | Yes |
| **Last Frame** | No | No | Yes | Yes | Yes |
| **Exports** | 720p | 720p | 1080p | Unlimited 4K | Unlimited 4K |
| **Personalize** | 1 | Up to 3 | Up to 5 | Up to 30 | Unlimited |
| **Multi-image reference** | — | 2 | 4 | 14 | 14 |
| **Team seats** | — | — | 3 | 6 | 20 |
| **Support** | — | General | General | Priority | Priority |
For a complete breakdown, see the [Pricing and Features](https://help.imagine.art/subscription-plans-and-pricing-1/subscription-plans/pricing-and-features) reference.
***
## Enterprise
Enterprise is designed for organizations that need more than what standard plans offer — custom scale, security controls, and a dedicated support relationship.
Define exactly how many credits and seats your organization needs. No hard caps — everything is negotiated to fit your actual usage.
SSO/SAML integration, advanced permission controls, and data handling arrangements to meet your organization's requirements.
Service level agreements with guaranteed response times and escalation paths for production-critical workflows.
A direct line to the ImagineArt team — onboarding assistance, training, and ongoing account management.
### Who Enterprise is for
* Studios and agencies running high-volume production pipelines
* Companies that need centralized billing across many users or departments
* Organizations with SSO, compliance, or audit requirements
* Teams that need more than 20 seats or 100,000 credits per month
### How to get started
Visit [imagine.art/subscription](https://www.imagine.art/subscription) and use the **Enterprise contact** option, or email [billing@imagine.art](mailto:billing@imagine.art) to speak with the sales team. Come prepared with your approximate monthly generation volume, team size, and any compliance or integration requirements.
***
## Individuals vs. Teams
The plan tables above apply to both **Individual** and **Team** accounts. The key difference is how seats work:
* **Individual** accounts use a single login.
* **Team** accounts share the same credit pool and subscription across all members, up to the seat limit included in your plan (3 / 6 / 20 for Standard / Ultimate / Creator).
Teams can centralize billing, add members, and manage access from a shared workspace. If you need more seats than your plan includes, contact sales for an Enterprise arrangement.
***
## Legacy Plans
If you purchased a plan **before 27 September 2025**, you are on a **legacy plan**. Legacy plans may have different prices and credit amounts compared to the tables above.
Contact [support@imagine.art](mailto:support@imagine.art) or [billing@imagine.art](mailto:billing@imagine.art) to confirm your current entitlements.
***
## Running Out of Credits
If you hit your credit limit before the month ends:
* **Upgrade your plan** — move to a higher tier for a larger monthly allotment.
* **Buy top-up credits** — one-time purchases that follow the credit usage policy and are consumed after your subscription credits run out.
Top-ups are separate from your recurring subscription. In the app, top-ups may appear as **Tokens**.
View and manage your subscription at [imagine.art/subscription](https://www.imagine.art/subscription).
***
## Payment Methods
ImagineArt accepts the following payment methods, all processed with bank-grade encryption:
* **Credit card** — all major cards accepted
* **Debit card** — must be enabled for online transactions
* **Link Payment** — secure checkout linked directly to your bank account
Credits are issued only after payment is successfully received. Failed or reversed transactions may result in a temporary pause to account access until resolved. All sales are subject to the [refund policy](https://help.imagine.art/terms-and-policies/refund-policy).
***
## Related
* [Sign up and sign in](/account/sign-up-sign-in)
* [Upgrade or downgrade your plan](/account/upgrade-downgrade)
* [Cancel your subscription](/account/cancel-subscription)
* [Update billing information](/account/updating-billing)
* [Understanding credits](/overview/understanding-credits)
# Troubleshooting
Source: https://docs.imagine.art/account/troubleshooting
Solutions to common issues with your ImagineArt web account, including login problems, locked models, failed generations, and credit issues.
Find solutions to the most common ImagineArt account issues below. If your problem isn't listed here, contact [support@imagine.art](mailto:support@imagine.art).
**Check which sign-in method you used.** ImagineArt supports Google, Facebook, Discord, and Email. Trying the wrong method is the most common cause of login failures.
Try each option on the sign-in screen. If you registered with a social provider (Google, Facebook, or Discord), there is no email/password combination to enter — use that provider's button directly.
**If you use email sign-in and forgot your password:**
1. Click **Forgot password?** on the sign-in form.
2. Enter your account email and click **Send reset link**.
3. Check your inbox (and spam folder) for the reset email.
4. Follow the link and set a new password.
**Other things to try:**
* Clear your browser cache and cookies, then try again.
* Try a different browser or an incognito/private window.
* Disable any browser extensions that might interfere with authentication (ad blockers, privacy shields).
* Make sure your device clock is set to the correct time — authentication tokens can fail if the time is skewed.
If you still cannot sign in, email [support@imagine.art](mailto:support@imagine.art) with your account email and a description of the error you see.
After upgrading, Pro models should unlock immediately. If they are still showing as locked:
1. **Refresh the page** — the UI sometimes needs a manual refresh to reflect the plan change.
2. **Sign out and sign back in** — this forces a fresh session with your updated plan status.
3. **Check your billing status** — go to **Billing & Subscription** and confirm your upgrade completed successfully. If you see a payment error, resolve it first.
4. **Wait a few minutes** — in rare cases, plan activation takes up to 5 minutes to propagate.
If Pro models remain locked after all of the above, email [support@imagine.art](mailto:support@imagine.art) and include your account email and the name of the plan you upgraded to.
Generation failures can happen for several reasons:
* **Temporary server load** — try submitting your generation again. Most transient failures resolve on a second attempt.
* **Prompt content** — if your prompt contains content that violates ImagineArt's usage policies, the generation will be blocked. Review the [Terms and Conditions](/policies/terms-and-conditions) for content guidelines.
* **Unsupported settings combination** — some models have constraints on certain setting combinations (e.g., resolution or aspect ratio). Try adjusting your settings and regenerating.
* **Browser issues** — clear your cache, try a different browser, or disable extensions.
If a generation fails, credits are generally not deducted for the failed job. If you believe you were charged for a failed generation, contact [support@imagine.art](mailto:support@imagine.art) with the date and approximate time of the failure.
Generation time depends on the model, settings, and current platform load. Typical ranges:
* **Images:** 4–15 seconds
* **Videos:** 30–60 seconds
If you are consistently experiencing much longer wait times:
* **Check platform status** — there may be a temporary increase in demand. Try again after a few minutes.
* **Try a different model** — some models are computationally heavier and naturally take longer.
* **Reduce output settings** — lowering resolution, output count, or video duration will generally speed up generation.
* **Check your internet connection** — a slow or unstable connection can delay results from reaching your browser.
If slow generations persist across multiple sessions and models, email [support@imagine.art](mailto:support@imagine.art).
**Subscription credits** refresh on your monthly billing anniversary — not on the 1st of each month. To check your renewal date:
1. Click your **profile picture** in the top-right corner.
2. Select **Billing & Subscription**.
3. Look for your **next renewal date**.
If the renewal date has passed and your credits have not refreshed:
* Check that your payment method is valid and your subscription is still active (a failed payment can pause your subscription).
* Sign out and sign back in to force a session refresh.
**Free credits** reset every 24 hours from when they were last issued. If your free credits haven't refreshed after 24 hours, sign out and back in.
If the issue persists, contact [support@imagine.art](mailto:support@imagine.art) with your account email and a screenshot of your credits modal.
All your past generations are stored in your **Assets**. To access it:
1. Click [Assets](https://www.imagine.art/asset/all) in the left nav bar.
If you are logged in as the correct account and still cannot find a specific generation, it may have been deleted. ImagineArt does have a recently deleted section accessible via assets screen. Go to Assets and click on "recently deleted" folder accessible via the right side of the screen
If you believe the content disappeared unexpectedly, email [support@imagine.art](mailto:support@imagine.art) with the approximate date of the generation and your account email.
## Contact support
If none of the above resolves your issue, reach out directly:
* **General support:** [support@imagine.art](mailto:support@imagine.art)
* **Billing issues:** [billing@imagine.art](mailto:billing@imagine.art)
When contacting support, include your account email address, a description of the issue, and any error messages or screenshots you have.
# Update Billing
Source: https://docs.imagine.art/account/updating-billing
How to update your payment method or billing address on your ImagineArt account.
You can update your payment method and billing address at any time from the billing settings in your account. Changes apply to your next renewal charge.
## Update your payment method
Go to [imagine.art](https://www.imagine.art) and sign in.
Click your **profile picture** in the top-right corner and select **Billing & Subscription**.
Scroll to the **Payment method** section. You will see your current card on file (last four digits and expiry date).
Click **Update payment method** (or **Add payment method** if none is saved). Enter your new card details in the form and click **Save**.
Your new payment method is now active. It will be charged at your next billing date.
## Update your billing address
From your **profile picture** menu, select **Billing & Subscription**.
Scroll to the **Billing address** section.
Click **Edit**, update the address fields, and click **Save**.
Updating your billing address affects future invoices only. Previously issued invoices will not be retroactively updated.
## View past invoices
Your billing history, including past invoices and payment dates, is available in the **Billing & Subscription** section under **Invoice history**. You can download any invoice as a PDF.
## Billing contact
If you encounter a billing error, a failed payment, or need help with your account charges, email [billing@imagine.art](mailto:billing@imagine.art) and include:
* Your account email address
* A description of the issue
* Any relevant invoice numbers or transaction IDs
## Related
* [Subscription plans](/account/subscription-plans)
* [Upgrade or downgrade your plan](/account/upgrade-downgrade)
* [Cancel your subscription](/account/cancel-subscription)
* [Refund policy](/policies/refund-policy)
# Upgrade or Downgrade
Source: https://docs.imagine.art/account/upgrade-downgrade
Change your ImagineArt subscription plan and understand how the transition affects your credits and team seats.
You can switch your subscription plan at any time from your account settings. The process differs slightly depending on whether you are moving to a higher or lower tier.
## Upgrading your plan
Sign in to [imagine.art](https://www.imagine.art) and click your **profile picture** in the top-right corner. Select **Billing & Subscription**.
Review the available plans. Click **Upgrade** next to the plan you want. You can compare plans on the [pricing page](https://www.imagine.art/subscription).
Review the prorated charge summary and confirm your payment method. Click **Confirm upgrade** to complete the change.
After the upgrade, your new credit allocation appears immediately in **Subscription Credits Left** on your profile. You now have access to all features included in your new plan, including Pro models if you were previously on the Free plan.
### When does an upgrade take effect?
Upgrades take effect **immediately**. You are charged a prorated amount for the remainder of your current billing period, and your new monthly rate applies from the next renewal date.
If you upgraded but Pro models are still locked, allow a few minutes for the change to propagate, then refresh the page. If the issue persists, email [support@imagine.art](mailto:support@imagine.art) with your account email and plan name.
## Downgrading your plan
Sign in to [imagine.art](https://www.imagine.art) and click your **profile picture** in the top-right corner. Select **Billing & Subscription**.
Click **Change plan** and select the plan you want to move to. Review the feature differences before confirming.
Click **Confirm downgrade**. Your current plan remains active through the end of the billing period you have already paid for.
At your next renewal date, your plan switches to the lower tier, your credit allocation adjusts to match the new plan, and you are billed at the lower rate going forward.
### What happens to unused credits?
* **Subscription credits** you earned under the higher plan are available until your renewal date. Unused credits do not roll over when the plan switches.
* **Top-up credits** are not affected by a plan change — they never expire and remain on your account.
### Seat management when downgrading a Teams plan
If you downgrade to a plan with fewer seats (for example, from Creator with 20 seats to Standard with 3 seats), you need to reduce your team size before the downgrade takes effect:
1. Go to **Team Settings** and remove members until you are within the seat limit of the new plan.
2. Removed members lose access to the workspace at your next renewal date.
3. If you do not reduce seats before the renewal, the downgrade may be blocked until the seat count is within limit.
Downgrading to the Free plan removes access to Pro models and team collaboration features. Make sure all team members are aware before you proceed.
## Related
* [Subscription plans](/account/subscription-plans)
* [Cancel your subscription](/account/cancel-subscription)
* [Update billing information](/account/updating-billing)
# Ad Maker
Source: https://docs.imagine.art/ad-maker
## Summary
The Ad Maker Node generates a complete ad video from a single prompt, a product, and an avatar. Instead of chaining together separate generation and editing nodes, this one node handles product placement, an on-camera presenter, and the ad's creative direction in a single run — a fast path to a UGC-style or hook-driven ad clip without hand-assembling a workflow for it.
It's a standalone node — a fast-path shortcut (keyboard shortcut **A**) at the top of the node picker, not filed under Image, Video, Audio, or Text.
## How to Use
Click the Add (+) button and select **Ad Maker**, or use the shortcut (A).
Both are required. Use the **+Product** and **+Avatar** pickers on the node to select or upload each.
Describe the advertisement in the prompt field. Type **@** to pull in a specific character or product from your library by name instead of describing it from scratch.
Choose one of three toggles: **UGC**, **Hook**, or **Scene** — these steer the creative direction of the generated ad (casual creator-style content, an attention-grabbing opening moment, or a fuller scene composition).
In the Properties panel, set **Aspect Ratio** (e.g. 16:9), **Resolution** (720p by default), and **Duration** (a slider, defaulting to 8 seconds).
Click Run. The node outputs a single finished ad video.
## Inputs and outputs
| Handle | Required | Notes |
| --------------- | -------- | --------------------------------- |
| Product | Yes | The product featured in the ad |
| Avatar | Yes | The on-camera presenter |
| Reference Image | No | Optional visual reference |
| Reference Video | No | Optional motion/style reference |
| Reference Audio | No | Optional voice or music reference |
| Video (output) | — | The finished ad clip |
## Sample Use Cases
Add a product and an avatar, switch to **UGC** mode, and describe the pitch. Ad Maker composes a casual, creator-style ad clip in one run instead of separately generating a scene, an avatar performance, and an edit.
Use **Hook** mode when you need a strong first few seconds to stop the scroll before the rest of the ad plays out — describe the hook moment directly in the prompt.
Feed a Reference Image or Reference Video from an earlier node into Ad Maker to keep a consistent product or visual style, then pass its output into downstream edit or export nodes.
For a guided, multi-step version of this same idea with a dedicated product/avatar/format/hook pipeline, see [Ad Studio](/ad-studio/what-is-ad-studio) — Ad Maker is the single-node, workflow-canvas equivalent.
# Ad Reference
Source: https://docs.imagine.art/ad-studio/ad-reference
Take a viral ad and give it your twist — keep the same vibe and energy, but make it all about your product.
**Ad Reference** is for when you've seen an ad that's clearly working — a format, a rhythm, an energy — and want that same feeling, but for your own product. Instead of building a look from scratch, you hand Ad Studio a reference video and it keeps the vibe while swapping in your product.
## When to use Ad Reference
* You've spotted a viral ad in your feed or a competitor's campaign that's clearly landing, and want to borrow its structure honestly (not clone it exactly) for your own product.
* You want a faster starting point than picking a format, hook, and setting individually.
* You have your own past ad you liked and want to riff on it again.
## How it works
In the left sidebar under **Tools**, click **Ad Reference**.
Upload your own reference (MP4/MOV/WEBM, up to 50MB) or pick one from **Latest Viral Ads** — a curated, regularly refreshed feed of real viral ad clips.
Click **Continue**, then bring in your own product the same way you would from the main workspace.
Borrow the energy, not the exact content. The point of a viral ad is that it already earned attention with its rhythm and hook — your job is to make your product the one being featured, not to reproduce someone else's ad beat for beat.
# Building an Ad From Scratch
Source: https://docs.imagine.art/ad-studio/building-from-scratch
Full creative control — prompt, format, hook, setting, product, and avatar.
Building from scratch is where you take full creative control. You write the prompt, pick the format, choose the hook, set the scene, and specify the product and avatar. This is the workflow for when URL to Ad isn't enough — or when you want to test very specific creative ideas.
The workflow has six stages, and we'll walk through each in order.
## 6.1 — The control panel and the @ shortcut
The workspace is built around a set of structured cards — **Template**, **Product**, **Avatar**, **Hook**, **Background** — plus an optional **Describe your advertisement** text field below them. You don't need to write a full brief in prose; picking the right combination of cards does most of the work, and the text field is for anything the cards don't cover.
At the top, **Medium** switches the whole panel between **Image** and **Video** output.
### The @ shortcut
Typing **@** in the **Describe your advertisement** field opens a dropdown of your saved avatars and products to mention by name — the system then uses your actual uploaded assets instead of guessing from a text description. This matters more than it sounds: it's the difference between a generic ad and an ad that actually shows your product.
**Try this** — Add one product and one avatar first, then type **@** in the description field and pick them by name. You'll see how much more specific the output becomes when the system has actual assets to work with.
## 6.2 — Choosing the right format
**Template** is the format decision — it determines the entire visual and structural language of the ad. Click **Change** on the Template card to open **Choose a format**. Ad Studio currently offers **8 formats**:
| Format | What it is | Best for |
| ---------------------- | ------------------------------------------------------------------------------------- | -------------------------------------------------------------------------------------------- |
| **Football TVC** | A top-tier TVC-director style football/soccer commercial. | Sports brands, World Cup and football-season campaigns. |
| **UGC** | Hyper-realistic Instagram/TikTok talking-head content. | Mass-market consumer products, social-first brands, anywhere trust matters more than polish. |
| **Unboxing** | Premium product unboxing and showcase videos. | Physical products with strong packaging, gifts, gadgets, premium goods. |
| **Tutorial** | Instagram-style talking-head tutorial ads, a real person walking through the product. | Complex products, tech, beauty regimens — anything that needs explaining. |
| **Virtual Try ON** | Instagram Reels-style fashion try-on hauls, talking-head UGC. | Fashion, eyewear, accessories, makeup — anything you wear or apply. |
| **Pro Virtual Try ON** | Higher-fidelity try-on with a pre-generation setup pass. | Premium fashion, luxury accessories, jewelry, high-end beauty. |
| **Asmr** | Senior-UGC-director-style sensory, sound-focused content. | Beauty, skincare, food, candles, sensory goods. |
| **Hyper-Motion** | Fast-cut, motion-heavy hyper-motion product ad. | Tech, sports, automotive, athletic wear — anything dynamic. |
Don't pick a format because it looks cool — pick it because it matches your product and audience. A luxury watch in UGC will feel cheap; a \$9.99 hair clip in Pro Virtual Try On will feel like you're trying too hard. **Match the format to the product's price point and emotional category.**
## 6.3 — Picking a hook
Click the **Hook** card to open **Pick your hook**. Rather than a set of psychological "archetypes," Ad Studio's hooks are literal, named moments you can drop into the first few seconds of your ad — filterable by **All**, **Stunt**, **Subtle**, and a seasonal **World Cup** tab.
A sample of what's in there:
| Hook | What it does |
| --------------------- | ----------------------------------------------------------------------- |
| **Product Hit** | An object flies into frame and hits the subject. |
| **Slap** | A hand swings in and slaps the subject's cheek. |
| **Random Object Mic** | During a casual vlog, a random object becomes an absurd microphone. |
| **Product Drop** | A slow-mo product drop from above. |
| **Interview** | The avatar holds the product, an interviewer off-camera asks questions. |
| **Blizzard** | A cozy scene is suddenly hit by a blizzard. |
| **Car Hit** | A person walking on the road, a fast car passes close by. |
| **Bird Hit** | The avatar opens their mouth to start talking — interrupted by a bird. |
There are roughly 30 hooks in total, split across the Stunt and Subtle tabs — browse the picker itself for the full set, since new ones (like the World Cup tab) get added over time.
Test hooks aggressively. The same ad with a different hook can swing performance by **3–5×**. Generate the same ad with a few different hooks, run all of them for a small budget, and double down on the winner. Cycle every 7–10 days as fatigue sets in.
## 6.4 — Setting the scene
Click the **Background** card to open **Set the scene**. Settings are filterable by **All**, **Unrealistic**, **Realistic**, and a seasonal **World Cup** tab, with dozens of named, literal locations rather than broad categories:
| Setting | Description shown in the picker |
| ---------------------- | ----------------------------------------------------------- |
| **Airplane** | Person sits on an airplane wing mid-flight. |
| **Volcano** | Person sits on an active volcano rim. |
| **Glacier** | Person stands on calving glacier ice. |
| **Temple** | Person sits on a rain-slicked temple step. |
| **Gym** | Gym floor, locker room, or post-workout setting. |
| **City** | Walking on a sidewalk or standing in a city street. |
| **Disco Club** | A disco club with music and other dancers. |
| **Abandoned Car Park** | An indoor, spooky abandoned car park. |
| **Pandora** | Person stands at the edge of a fantastical alien landscape. |
| **Tribal People** | A tropical beach with tribal-styled figures. |
Settings carry hidden assumptions about your customer. A kitchen assumes the viewer cooks at home; a co-working space assumes flexible work. Pick a setting that flatters **the way your customer wants to see themselves** — not just one that fits the product.
## 6.5 — Output settings (aspect ratio, quality, duration)
Below the description field, three pills control aspect ratio, quality, and duration.
### Aspect ratio
**16:9, 9:16, 3:4, 4:3, 1:1, 21:9.** 9:16 is the default and the dominant ratio for social-first paid media (TikTok, Reels, Shorts); 1:1 is the safest bet when you don't know exactly where an ad will run; 16:9 and 21:9 suit wide-screen placements.
### Quality
Four tiers: **Max, High, Medium, Low.** Use **Low** or **Medium** while you're iterating on several variations; switch to **High** or **Max** for the version you're actually shipping.
### Duration
An exact per-second dropdown from **5s to 15s** (default 10s) — not a set of preset ranges. Shorter durations suit attention-only hooks; use the full 15s when the ad needs a small narrative arc (problem → product → payoff).
## 6.6 — Adding products and avatars
The **Product** and **Avatar** cards are where you load in the two assets that make an ad recognizably yours.
### Adding a product
Click **Product** to open a picker with a real preset library — roughly 55 items across 10 categories (Streetwear, Fashion, Wellness, Skincare, Accessories, Home, Sports, Beverages, Tech, Menswear) — plus a **Saved** tab for anything you've added yourself.
Beyond the presets, you have two ways to add your own:
* **Paste a product URL** to scrape a live product page.
* **Create Manually** — upload up to 9 images (JPG/PNG/WEBP, ≤25MB), give it a **Name** (used for @-mentions), and optionally add **Additional Instructions**. A single image is enough; multiple angles aren't required, though they help.
URL scraping isn't always reliable. Sites that block bots (Amazon is a known example) can silently return the wrong product — a blank thumbnail or a garbled, unrelated title — with no warning that the pull is wrong. Always check the pulled result before generating; don't assume a "successful" pull actually got your product.
### Adding an avatar
Click **Avatar** to open a similar picker — roughly 50 named preset avatars across 8 categories (including a seasonal **World Cup** category), plus **Saved**.
Two ways to add your own:
* **Generate Avatar** — a single free-text description field (no gender/ethnicity/age filters), which generates 4 options to choose from.
* **Create Manually** — upload up to 10 photos (JPG/PNG/WEBP, ≤25MB) or select from your past creations, give it a **Name**, pick **Male** or **Female**, and optionally add **Additional Instructions**.
### Attaching a voiceover or music reference
Below the Product/Avatar/Hook/Background cards, an optional **Upload media** row (images, videos, or audio) opens **Add References** — filterable by **All**, **Image**, **Video**, and **Audio**, accepting JPG/PNG/WEBP/MP4/WEBM/MP3/WAV/OGG up to 30MB. This is how you attach a real voiceover or music track as a reference for the generation, separate from the Product and Avatar cards.
There's no built-in captions or subtitles toggle — several format templates explicitly instruct the model to avoid on-screen text overlays. If your platform needs captions, add them afterward with another tool rather than expecting Ad Studio to burn them in.
### Casting your avatar
Pick the avatar that matches your target customer — **not** the one that looks the most attractive. A 35-year-old skincare brand doesn't sell better with a 20-year-old face; it sells better with a 35-year-old face. Your audience needs to see themselves in the ad, not aspire to be someone else.
Once you find an avatar that performs, reuse it across the entire campaign. Consistency builds recognition; recognition builds trust; trust drives clicks. The same avatar across 10 ads is far more powerful than 10 different avatars across 10 ads.
# Getting Started: Your First Ad
Source: https://docs.imagine.art/ad-studio/getting-started
Follow this section once and the rest of the guide will make a lot more sense.
This section walks you through opening Ad Studio and producing your first ad. Follow it once and the rest of the guide will make a lot more sense.
## Step 1 — Open Ad Studio
Open your browser and go to **imagine.art/ad-studio**.
Sign in if you are not already logged in.
You'll land on the Ad Studio home view, with featured format examples in the main panel and your projects listed on the left.
## Step 2 — Start a new project
Click the **"+"** button next to "Projects" on the left. A new project is created immediately — there's no separate naming step first.
New projects start out as "Untitled Project." Rename it to something descriptive (for example, "Skincare Q4 Launch") once you're inside it.
Treat each campaign or product as its own project. This makes it much easier to find your variations later, especially when you're A/B testing — and you will be A/B testing.
## Step 3 — Pick your starting point
Ad Studio gives you several ways to begin. Pick the one that fits your situation:
* **URL to Ad.** You have a live product page. Click "URL to Ad" in the Tools section, paste the link, pick a style, and generate. Best for: e-commerce sellers with existing product pages.
* **Motion Design.** Same product picker as URL to Ad, but generates a storyboard of independently editable scenes instead of a single clip — see [Motion Design](/ad-studio/motion-design).
* **Ad Reference.** You've seen a viral ad that's clearly working and want that energy for your own product — see [Ad Reference](/ad-studio/ad-reference).
* **Build from scratch.** You work directly in the main control panel — Template, Product, Avatar, Hook, and Background cards. Best for: full creative control, or when you don't have a product page yet.
If you're unsure, start with **URL to Ad**. It's the fastest way to see what Ad Studio can do with your actual product.
# Glossary
Source: https://docs.imagine.art/ad-studio/glossary
A quick reference to the terms used throughout this guide and inside Ad Studio.
A quick reference to the terms used throughout this guide and inside Ad Studio.
| Term | Meaning |
| ------------------------ | ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| **A/B testing** | Running multiple variations of an ad to learn which performs best. |
| **Ad Reference** | A tool that lets you remix an existing viral ad — keeping its vibe and energy — with your own product. See [What is Ad Studio?](/ad-studio/what-is-ad-studio). |
| **Aspect ratio** | The width-to-height ratio of the ad. Real options: 16:9, 9:16, 3:4, 4:3, 1:1, 21:9. |
| **Avatar** | The person, animal, or mascot (real or AI-generated) who appears in the ad. |
| **Background** | The card that opens **Set the scene** — where the ad takes place, filterable by All/Unrealistic/Realistic/World Cup. |
| **Call to action (CTA)** | The line that tells the viewer what to do next — "Shop now," "Use code," etc. |
| **Create Manually** | The upload-your-own path in the Product and Avatar pickers — images/photos, a name, and optional instructions, rather than a preset or URL pull. |
| **Creative** | Industry term for the actual ad — the video, image, copy, and design assets. |
| **Creative fatigue** | When an audience has seen an ad too many times and stops responding. Solved by refreshing creative. |
| **DTC** | Direct-to-consumer — brands that sell directly through their own storefronts. |
| **Format** | The structural style of the ad, chosen via the **Template** card. Ad Studio offers 8 (Football TVC, UGC, Unboxing, Tutorial, Virtual Try ON, Pro Virtual Try ON, Asmr, Hyper-Motion). |
| **Hook** | A named moment (e.g. "Product Hit," "Slap") placed in the first few seconds of an ad to stop the scroll, filterable by All/Stunt/Subtle/World Cup — not a psychological archetype. |
| **Medium** | The Image/Video toggle at the top of the control panel. |
| **Motion Design** | A landing-page tool using the same product picker as URL to Ad, aimed at motion-graphics-style ads. |
| **Performance ad** | An ad designed to drive a measurable action (click, install, purchase), not just brand awareness. |
| **Quality** | Output fidelity tier: Max, High, Medium, or Low. |
| **Stop-scroll** | The action of pausing on an ad mid-feed. The first job of any hook. |
| **Template** | The card that opens **Choose a format** — Ad Studio's term for the format picker. |
| **UGC** | User-generated content. A format that mimics casual, real-person content rather than polished production. |
| **URL to Ad** | Ad Studio's tool for generating ads directly from a product page URL — no product-details review step, and prone to silent bad pulls on scraper-blocking sites. |
| **@ shortcut** | Typing @ in the description field to pull in specific products or avatars from your library. |
# Motion Design
Source: https://docs.imagine.art/ad-studio/motion-design
Turn a product into a storyboard-driven motion graphics ad — pick a style, generate scene-by-scene, then refine each scene independently.
Motion Design turns a product photo into a motion-graphics-style ad built from a storyboard: a set of individually generated scenes, not one single video generation. It shares its product picker with [URL to Ad](/ad-studio/url-to-ad), then branches into its own storyboard flow once you've picked a style.
## How Motion Design works
In the left sidebar under **Tools**, click **Motion Design**. This opens the same product picker as URL to Ad — paste a URL, upload images, or pick an existing product from your library or the preset categories.
On **Choose a style**, describe the video, the product, and the vibe, then pick a style template — Product Showcase, SaaS Explainer, Mobile App Launch, Brand Campaign, Location Promotion, and more.
Motion Design is scoped to ByteDance's Seedance lineup only — a narrower, purpose-built set compared to the much larger model list in the general Video tool.
| Model | Credits | Resolution | Duration | Notes |
| ----------------- | ------- | ---------- | -------- | ----------------------------------------------------------- |
| Seedance 2 | 730 | 4K–1080p | 4–15s | ByteDance's latest video model |
| Seedance 2.5 | 530 | 480p–1080p | 4–30s | ByteDance's latest video model, longest max duration |
| Seedance 2.0 Mini | 175 | 480p–720p | 5–15s | Lightweight, fast, built for multi-shot video — the default |
All three support Start/End frame control and audio.
Click **Create Storyboard**. This kicks off a multi-stage background job — computing grid geometry, assembling a "mega-prompt," generating the storyboard, then receiving split frames.
Generation is genuinely asynchronous, not just a progress bar. Click the minimize button on the generation screen and keep working anywhere else in Ad Studio — the **Motion Design** entry in the left sidebar shows a live spinner until the storyboard is ready, confirming it keeps running in the background rather than blocking the rest of the app.
**This can genuinely take longer than "several minutes."** The generation screen itself says to expect several minutes, but in testing a single storyboard was still generating well past the 8-minute mark. Don't assume a stall — check back later rather than retrying.
## What you get back
Motion Design doesn't generate one clip — it produces a **storyboard of individual scenes**, based on real, first-hand tester feedback from the team (not yet independently re-verified end-to-end in this pass, since a real generation was still running past 8 minutes during this documentation session):
* **Per-scene editing.** Each scene in the storyboard can be regenerated on its own without affecting the others — useful for fixing one weak beat in an otherwise good sequence.
* **Working background generation.** The product photo carries through into a real generated background per scene, not a placeholder.
* **Premium output.** The finished export includes typography and transitions applied automatically, not just raw clips stitched together.
**Not the same as Film Studio's Storyboard mode.** Both features use the word "storyboard," but they're different mechanisms. [Film Studio's Storyboard mode](/film-studio/create-image#storyboard-mode) generates a single contact-sheet grid image that you manually split into frames. Motion Design's storyboard is an automated multi-scene pipeline purpose-built for ads — you don't do the splitting yourself.
## When to use Motion Design
* You want a polished, template-driven motion graphics ad rather than a single straight-through video generation.
* You're open to a longer, asynchronous generation in exchange for a multi-scene result with independently editable beats.
* You want typography and transitions applied automatically rather than adding them yourself afterward.
If you want the fastest possible single clip instead, use [URL to Ad](/ad-studio/url-to-ad) or the main workspace's video generation.
# Pro Tips & Best Practices
Source: https://docs.imagine.art/ad-studio/pro-tips
Habits that lift your output from 'acceptable' to 'performing.'
Once you've shipped a few ads, these habits will lift your output from "acceptable" to "performing."
## Strategy
* Write down your customer, their problem, and the platform before you generate anything. Every creative decision flows from those three answers.
* Pick one job per ad. "Build awareness" and "drive a sale" are different jobs; trying to do both usually does neither.
* Match production quality to product price. UGC for impulse buys, Pro Virtual Try On for premium. Mismatched quality kills credibility.
## Creative
* Hook in the first second. Not the first three — the first one. If a viewer doesn't feel a reason to watch by frame 30, they're gone.
* Show the product within the first 3 seconds, even when the format is story-driven. Brand recognition compounds across impressions.
* End with a clear call to action. "Shop now," "Use code," "Tap the link" — pick one and say it explicitly.
* Subtitles matter — most paid social plays muted by default, and Ad Studio doesn't burn in captions itself (several formats explicitly avoid on-screen text). Add subtitles in a separate editing pass before you publish. In the meantime, make sure the ad works with sound off on its own.
## Testing
* Generate three variations of every ad. Different hooks, same product. Run them in parallel and learn which hook style fits your audience.
* Refresh creative every 7–14 days. Ad fatigue is real — the same ad shown to the same audience loses performance fast.
* Keep a winners file. When something performs, save the prompt, the format, the hook, and the avatar. That combination is repeatable.
## Production rhythm
* Use URL to Ad for first drafts, the main workspace for refinement. Don't try to perfect inside URL to Ad.
* Reuse avatars and products across a campaign. Consistency drives recognition; recognition drives trust; trust drives clicks.
* Name your variations clearly. "Ad V3 — Confession hook" beats "Untitled 7."
## Common mistakes to avoid
Picking a format because it looks cool, not because it fits the product.
Using a luxury setting for a budget product, or vice versa. Setting signals price; a mismatch confuses the buyer.
Casting an avatar who looks like the dream customer instead of the actual customer.
Shipping one ad and waiting to see what happens. You need three to learn anything; you need ten to learn enough.
Treating Ad Studio output as final on the first generation. It's a draft — iterate.
# Quick Reference Card
Source: https://docs.imagine.art/ad-studio/quick-reference
One page to return to whenever you need a fast reminder.
One page to return to whenever you need a fast reminder.
## A · Three ways to start
* **URL to Ad** → paste product link or pick existing → Choose a style → Create. Fast path for e-commerce (heads up: no review step, and scraper-blocked sites can silently pull wrong data).
* **Motion Design** → same product picker, but generates a storyboard of independently editable scenes (Seedance models only, can take well past the stated "several minutes") — see [Motion Design](/ad-studio/motion-design).
* **Ad Reference** → upload or pick a viral ad → remix it with your product.
* **Main workspace** → Template / Product / Avatar / Hook / Background cards, optional description → Generate.
## B · The @ shortcut
Type **@** in the description field → pick a specific product or avatar from your library → the system uses your **actual assets**.
## C · The eight formats
| | | | |
| -------------- | ------------------ | -------- | ------------ |
| Football TVC | UGC | Unboxing | Tutorial |
| Virtual Try ON | Pro Virtual Try ON | Asmr | Hyper-Motion |
## D · Hooks — named moments, not archetypes
Filterable by **All / Stunt / Subtle / World Cup** — roughly 30 total. Examples: Product Hit, Slap, Product Drop, Interview, Blizzard, Car Hit, Product Bump, Bird Hit. Browse the picker for the full set.
## E · Setting categories
Filterable by **All / Unrealistic / Realistic / World Cup** — dozens of named locations, not broad categories. Examples: Airplane, Volcano, Glacier, Temple, Gym, City, Disco Club, Abandoned Car Park, Pandora, Tribal People.
## F · Output cheat sheet
* **Aspect:** 16:9 · 9:16 (default) · 3:4 · 4:3 · 1:1 · 21:9.
* **Duration:** exact per-second, 5s–15s (default 10s) — not preset ranges.
* **Quality:** Max · High · Medium · Low (default). Use Low/Medium for iteration, High/Max for shipping.
* **Audio:** Add Audio / Add References → voiceover or music, filterable by Image/Video/Audio. No built-in captions — add those separately.
## G · The discipline that pays off
* **Three variations per ad** — same product, different hook.
* **Refresh creative every 7–14 days** to beat ad fatigue.
* **Reuse avatars across a campaign** for recognition.
* **Keep a winners file:** prompt + format + hook + avatar combinations that perform.
**The final word** — Ad Studio is most powerful when you generate three, ship the best, learn, and iterate. The brands that win on paid social aren't the ones with the prettiest ad — they're the ones who **learn the fastest**. This tool exists to make learning fast.
# Recipes for High-Converting Ads
Source: https://docs.imagine.art/ad-studio/recipes
Six proven format + hook + setting combinations that consistently perform across categories.
Below are combinations that consistently perform across categories. Treat them as starting points — copy the combination, swap in your product and avatar, and tune from there. Each recipe pairs a format with a hook and a setting from the real pickers — see [Building an Ad From Scratch](/ad-studio/building-from-scratch) for the full lists.
| Recipe | Format | Hook | Setting | Best for |
| --------------------- | ------------ | -------------------------------------------------------------------- | ------------------------------------------- | -------------------------------------------------------- |
| **The Honest Friend** | UGC | **Interview** (avatar holds the product, interviewer asks questions) | A **Realistic** everyday location | Beauty, wellness — where trust matters more than polish. |
| **The Reveal** | Unboxing | **Product Drop** (slow-mo product drop from above) | An **Unrealistic** or aspirational location | Premium consumer goods, gifts, gadgets. |
| **The Trigger** | Asmr | **Product Bump** (POV walk-in, camera taps the avatar) | **Realistic**, close and intimate | Beauty products, candles, food, sensory goods. |
| **The Hero** | Hyper-Motion | **Product Hit** (an object flies into frame and hits the subject) | Any high-energy **Unrealistic** setting | Tech, sports, automotive, premium athletic wear. |
The Hook and Background pickers have roughly 30 and dozens of entries respectively, well beyond what fits in a table here. Browse the **Stunt**/**Subtle** hook tabs and the **Realistic**/**Unrealistic** setting tabs directly in the picker — the four recipes above are proven starting points, not the full set of good combinations.
## How to use these recipes
Pick the recipe that matches your product category.
Load your real product and a matching avatar.
Set duration to **10–15 seconds** and aspect ratio to **9:16** for social-first delivery.
Generate two or three variations from the same recipe — the system will produce different interpretations.
Pick your strongest, ship it, watch the data, then iterate.
Don't change three things at once. When you iterate, change the hook **OR** the setting **OR** the avatar — never all three. That way you actually learn which lever moved the result.
# The Fast Path: URL to Ad
Source: https://docs.imagine.art/ad-studio/url-to-ad
URL to Ad is the single fastest way to produce a finished ad — paste a product page, get an ad.
URL to Ad is the single fastest way to produce a finished ad. You paste a live product page, Ad Studio pulls the relevant product information — images, name, description, branding — and generates a complete ad based on it. For e-commerce sellers, this is the killer feature.
## When to use URL to Ad
* You sell on Shopify, Amazon, Etsy, your own storefront, or any e-commerce platform with public product pages.
* You want a first creative draft fast — to test a concept, fill a campaign deadline, or generate volume for A/B testing.
* You have a catalog of products and want ads for many of them without setting up each one manually.
* You want a baseline ad you'll then refine in the main workspace.
## How URL to Ad works
In the left sidebar under **Tools**, click **URL to Ad**. This opens a product picker — paste a URL, upload images, or pick an existing product from your library or the preset categories (Streetwear, Fashion, Wellness, and more).
Paste the full URL of your product page and press enter, or click **Create Manually** to upload your own images instead. Use the live public URL — not a draft page that requires login.
A new tile appears in the grid while the pull runs, then resolves to a thumbnail and name, or a **Failed** state. Select the tile and click **Continue**.
This opens **Choose a style to start** — pick a format, with aspect ratio, quality, and duration already defaulted (9:16, Low, 10s). Adjust if needed, then click **Create**.
**There's no product-details review step.** Unlike a form you fill in and check, URL to Ad never shows you the pulled name, description, or images for editing before you generate — what you see in the product tile is what you get. If a pull looks wrong, delete it and try again rather than assuming it'll self-correct downstream.
**A failed pull isn't the only risk — a "successful" one can be silently wrong.** Some sites block automated scraping (Amazon is a known case) and return an unrelated page that gets scraped as if it were legitimate — producing a real product tile with a blank or garbled thumbnail and a nonsense name, with no error or warning shown. Always look at what was actually pulled before generating from it.
## Getting better URL to Ad results
URL to Ad works best when your product page is clean: good product photos, a clear product name, an accurate description. If your page is cluttered or your photos are weak, fix the page first — then generate. **The output is only as good as the input.**
* Make sure your product page has at least **three clean photos**: hero, lifestyle/in-use, and detail.
* Use a clear, descriptive product name. "Hydrating Vitamin C Serum" generates better than "Glow Drops V2."
* If your page has a strong tagline or unique selling point in the description, the system will use it. If it doesn't, edit the pulled description before generating.
* Generate two or three variations and compare. Pick the strongest and run it; refine the others in the main workspace.
Have a catalog? Run several products through URL to Ad back-to-back to build a library of baseline drafts, then spend your creative time only on the ones worth refining.
## What to do with the output
Treat the URL-to-Ad output as a **draft, not a final**. It will get you 70–80% of the way to a good ad. The remaining 20–30% — picking the perfect hook, tuning the avatar, dialing in the setting — is where the main workspace comes in. The next section walks you through that workflow.
# Welcome to Ad Studio
Source: https://docs.imagine.art/ad-studio/what-is-ad-studio
A production environment for performance ads — from a product URL to a high-converting ad.
Ad Studio is a production environment for performance ads. It is built for a specific job: turning a product into a finished, platform-ready video ad — fast, repeatedly, and at the quality that actually performs in a paid feed.
If Film Studio is for directing movies, Ad Studio is for shipping campaigns. The tagline says it best: ads, ready when you are. You can paste a product URL and have an ad in minutes, remix an existing viral ad with [Ad Reference](/ad-studio/ad-reference), or build one from scratch with full control over format, hook, setting, and avatar. Either way, you stay out of the production rabbit hole — no shoot day, no model casting, no editing suite — and you get to test creative variations at a speed traditional ad production cannot match.
## Who this guide is for
This guide is written for the people who actually ship ads:
* **E-commerce sellers** on Shopify, Amazon, Etsy, or any DTC storefront.
* **Brand and performance marketers** who run paid social campaigns.
* **Agencies** producing creative for multiple clients at once.
* **Founders** running ads themselves before they hire a marketing team.
* **Anyone testing creative iteratively** — UGC variants, hook A/B tests, format tests.
## How this guide is organized
The guide is structured around two workflows: the **fast path** (URL to Ad) for when you have a product page and need an ad now, and the **build-from-scratch path** for when you want full creative control. We start with orientation, then walk through both paths, then give you recipes and best practices.
* **Sections 1–4** cover orientation: what Ad Studio is, what you need, and how the workspace is laid out.
* **Section 5** covers the URL-to-Ad fast path.
* **Section 6** covers building an ad from scratch — formats, hooks, settings, and assets.
* **Section 7** gives you proven recipes for the most common ad types.
* **Sections 8–10** are reference material: tips, glossary, and a one-page cheat sheet.
## How to use this guide
Throughout the guide you will see three kinds of callouts. Each marks a different kind of advice:
* **Tip** — a non-obvious technique that saves time or money.
* **Try this** — a quick exercise to lock in what you just learned.
* **Heads up** — a limit, a constraint, or a common mistake to avoid.
You'll also meet **Pro Tip**, **Template**, and **Read this first** boxes. They follow the same color language used across the rest of this document — purple for guidance, amber for caution, green for hands-on practice.
## What you'll be able to do by the end
* Open Ad Studio, create a project, and produce a finished ad from a product URL in minutes.
* Build an ad from scratch with the right format for your product — UGC, unboxing, review, virtual try-on, ASMR, or hyper-motion.
* Pick a hook that earns the first three seconds of attention.
* Choose a scene setting that matches your product and audience.
* Configure aspect ratio, quality, and duration for the platforms you run on.
* Add product images and avatars — uploaded, generated, or from the library — to lock in a consistent visual identity.
* Apply proven recipes for common ad types, and test variations quickly enough to learn what actually performs.
**Read this first** — If you only have ten minutes, read [Getting Started](/ad-studio/getting-started) and [The Fast Path: URL to Ad](/ad-studio/url-to-ad). That alone is enough to produce your first ad. Come back to the rest when you want to tune the creative.
## Before you begin
Ad Studio runs entirely in your web browser. There is nothing to install. Before you start, take a minute to make sure the basics are in place — it will save you frustration later.
### What you need
Chrome, Edge, Safari, or Firefox, kept reasonably up to date.
Generation runs on our servers, so a steady connection means a smoother experience.
Sign up or log in at imagine.art before opening Ad Studio.
Either a live product URL (Shopify, Amazon, your storefront) or clean product photos you can upload.
TikTok, Instagram Reels, Meta in-feed, YouTube Shorts. The platform determines aspect ratio, duration, and pacing.
### A useful mindset
Ads are not films. Ads do one job: **stop the scroll, sell the click.** Every choice in Ad Studio — format, hook, setting, duration — should be made in service of that job. A beautifully directed ad that loses the viewer in the first second has failed; an imperfect ad that holds attention and drives the click has won.
This guide gives you frameworks and recipes, but the most important habit is **iteration**. Ship a version, look at the data, refine, ship again. Ad Studio's value is that you can do this in hours instead of weeks.
Before you generate anything, write down three things on a sticky note: **who the customer is**, **what problem the product solves** for them, and **where they are going to see this ad**. Every creative decision flows from those three answers.
# Understanding the Workspace
Source: https://docs.imagine.art/ad-studio/workspace
Three areas where you'll spend almost all your time — learn them once and every later section gets faster.
When you open Ad Studio, you'll see three areas where you'll spend almost all your time. Knowing what each one does will make every later section faster to follow.
The URL-to-Ad tool and all your projects.
Format examples and the prompt bar where ads are made.
The core control surface: format, hook, setting, assets.
Navigation, search, upgrades, and your profile.
## The left sidebar
The left sidebar has two important sections:
* **Tools.** Three fast-path entry points: **[URL to Ad](/ad-studio/url-to-ad)** (generate from a product link), **[Motion Design](/ad-studio/motion-design)** (the same product picker, but a storyboard of independently editable scenes instead of one clip), and **[Ad Reference](/ad-studio/ad-reference)** (remix an existing viral ad with your own product).
* **Projects.** Every ad you build lives inside a project. Use the search box to find existing projects, or click "+" to start a new one.
## The center workspace
The center is where ad creation happens. By default it shows a grid of format examples under "Start from preset" — filterable by categories like Lifestyle UGC, Motion Studio, Creative Studio, TVC, Environments, and Cinematic — that double as inspiration and a quick-start gallery. Opening a project shows the control panel, built around a set of cards:
* **Medium.** Switches the whole panel between **Image** and **Video** output.
* **Template.** Shows your currently selected format, with a **Change** button to switch between the 8 real ad formats.
* **Product.** Add the product that will be featured.
* **Avatar.** Add the person, animal, or mascot who appears in the ad.
* **Hook.** Pick from a library of named hook effects, filterable by Stunt, Subtle, and seasonal categories.
* **Background.** Choose where the ad takes place.
* **Describe your advertisement (optional).** Free text for anything the cards above don't cover. Use the **@** symbol to pull in specific products or avatars by name.
* **Output pills.** Aspect ratio, quality, and duration, shown as three separate pills below the description field.
* **Generate.** The big purple button that generates your ad.
See [Building an Ad From Scratch](/ad-studio/building-from-scratch) for the full depth of each of these.
## The top bar
Project navigation, search, contact sales, plan upgrades, notifications, and your profile menu. Standard top-bar functionality — nothing creative happens here.
**Try this** — Before you read further, open a project in Ad Studio and click each of the control cards (Template, Hook, Background) plus the output pills below the description field. Just notice what options appear — you'll get more from the next sections if you've already seen the menus.
# AI Copilot
Source: https://docs.imagine.art/ai-copilot
## Summary
The AI Copilot Node leverages the power of LLMs to process and generate high-quality text outputs. Whether you're working with text, images, or videos as inputs, this node provides advanced capabilities to analyze content, extract contextual information, and generate detailed and coherent text.

### Key Features
* **Advanced Text Processing:** Utilize LLMs to process and refine text inputs, whether they come from images, videos, or raw text.
* **Contextual Understanding:** The node not only generates text but also extracts context, personas, and identifiers from the inputs, offering detailed outputs.
* **Multi-Input Integration:** Combine text, images, and videos as inputs to produce more informed and contextual outputs.
* **High-Quality Output:** Generate rich, contextually accurate text that can be used for scripts, descriptions, stories, and more.
## How to Use
From the workflow canvas, click the Add (+) button. Select AI Copilot from the Text node category (it sits at the top level, not inside Text Utilities), or use shortcut (T) to add the node on canvas.
You can input any type of file: text, images, or videos. The AI Copilot node will analyze these inputs to extract context and information. For example, if you input a video, the node will identify key details and generate a description based on the visual and contextual elements present in the video.
Once the node processes the input, it will generate a detailed output, whether it's a summary, an analysis, or a creative piece of text like a script or blog post. You can then feed this output into other nodes, such as Image or Video, for further use.
### Example: Marketing Campaign Script
Let's say you are working on a marketing campaign for a new product launch. You upload a video ad of the product, and the AI Copilot node will analyze the video's tone, the emotions conveyed, and the visuals. It will then generate a script that matches the mood and message of the video, ensuring consistency between the video and the campaign's marketing text.

## Sample Use Cases
"Write a 5-minute YouTube video script about 'The Future of Artificial Intelligence in Content Creation.' The script should be engaging, informative, and tailored for an audience of young professionals in the content creation space. Include a captivating intro, main points, and a call-to-action at the end."

"Write a short story set in a futuristic city where technology controls every aspect of life. The protagonist, a young rebel, discovers a hidden secret about the city's artificial intelligence systems that could change everything. Explore themes of freedom, surveillance, and the nature of control."

"Create 5 short ad copy variations for a Google ad campaign promoting an eco-friendly fashion brand. Each ad should emphasize sustainability, stylish design, and comfort. Keep the copy concise (up to 200 characters) and focused on driving conversions."

"Write a brand story for a luxury leather goods company that prides itself on craftsmanship and sustainability. The story should evoke emotions, highlight the artisanship behind each product, and emphasize the brand's commitment to eco-friendly practices."

# Choosing the Right Image Model
Source: https://docs.imagine.art/ai-models/image-models
Find the best ImagineArt image model for your project — use case guidance, side-by-side comparisons, and direct links to every model page.
ImagineArt gives you access to 30 image generation models, each optimized for a different creative or professional workflow. Use this guide to find the right one for your project.
## Quick picks by use case
**ImagineArt 2.0** — ImagineArt's flagship proprietary model. Best photorealism, comprehension, and composition control.
**Ideogram v3** — \~90–95% text accuracy. Best for posters, signage, and branded layouts.
**Recraft v4** — #1 HuggingFace Arena (ELO 1172). Only model with native SVG vector output.
**Nano Banana 2** — \~4× faster than Nano Banana Pro, up to 14 references, 4K output.
## By what you're building
**Priority: resolution + composition + text rendering**
For print-ready output where resolution and typographic accuracy both matter, [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) is a strong choice — native 4K generation with enhanced composition control and reliable text rendering. For typography-dominant designs, [Ideogram v3](/ai-models/image/ideogram-v3) achieves \~90–95% text accuracy with 58 style presets.
[Recraft v4 Pro](/ai-models/image/recraft-v4-pro) generates at 4-megapixel with native SVG vector output — the best option when you need editable, scalable assets alongside raster output. [Seedream v4.5](/ai-models/image/seedream-v4-5) and [Flux.2 Pro](/ai-models/image/flux-2-pro) both deliver production-grade text rendering at 4K.
| Model | Resolution | Text rendering | Best for |
| --------------------------------------------------------- | --------------- | ----------------- | ----------------------------------- |
| [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) | Native 4K | Strong | Posters, editorial, product visuals |
| [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) | 4MP (2048×2048) | Excellent | Print + SVG vector |
| [Ideogram v3](/ai-models/image/ideogram-v3) | Up to 1536px | \~90–95% | Typography-dominant designs |
| [Flux.2 Pro](/ai-models/image/flux-2-pro) | Up to 4MP | Production-grade | Branded print campaigns |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Up to 8K | Excellent (dense) | Large-format, multi-language |
**Priority: consistency + accuracy + material rendering**
[Nano Banana Pro](/ai-models/image/nano-banana-pro) (Google Gemini 3 Pro Image) is built for production product work — preserves the identity of up to 5 subjects for consistent brand asset creation, with Google Search grounding for accurate product rendering. [Nano Banana 2](/ai-models/image/nano-banana-2) handles up to 14 references at near-real-time speed for high-volume batches.
[Seedream v4.5](/ai-models/image/seedream-v4-5) supports up to 14 references at 4K — ideal for multi-product campaign sets. [Flux.2 Max](/ai-models/image/flux-2-max) accepts up to 10 references with real-time web grounding.
| Model | Reference images | Output quality | Best for |
| --------------------------------------------------- | ---------------- | -------------- | ------------------------------ |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | 5 subjects | 4K | Complex product composites |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | Up to 14 | 4K | High-volume batch product |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Up to 14 | Up to 8K | Multi-product campaign sets |
| [Flux.2 Max](/ai-models/image/flux-2-max) | Up to 10 | Up to 4MP | Premium product with grounding |
**Priority: skin fidelity + lighting + anatomy**
[ImagineArt 2.0](/ai-models/image/imagineart-2-0) is ImagineArt's highest-capability proprietary model — the best choice for premium portraits and lifestyle photography where the output needs to stand alongside actual photography. [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) is the strong middle option at 4K native. [ImagineArt 1.5](/ai-models/image/imagineart-1-5) is cost-efficient for draft exploration.
[Minimax Image](/ai-models/image/minimax-image) also excels at lifelike skin rendering with cinematic lighting inherited from MiniMax's video model lineage.
| Model | Portrait quality | Speed | Cost tier |
| --------------------------------------------------------- | ---------------- | ----- | --------- |
| [ImagineArt 2.0](/ai-models/image/imagineart-2-0) | Flagship | Fast | Premium |
| [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) | Excellent | Fast | Standard |
| [Minimax Image](/ai-models/image/minimax-image) | Excellent | Fast | — |
| [ImagineArt 1.5](/ai-models/image/imagineart-1-5) | Good | Fast | Lower |
**Priority: spelling accuracy + layout control**
[Ideogram v3](/ai-models/image/ideogram-v3) is the clear leader — \~90–95% text accuracy with support for stylized lettering, complex multi-line layouts, and multilingual text in six languages. Use the DESIGN style type for layout-aware graphic work.
[GPT Image 2](/ai-models/image/chatgpt-image-2) delivers 99%+ text accuracy across multilingual scripts including CJK and Indic — the strongest option when text accuracy and language coverage both matter. [ChatGPT Image 1.5](/ai-models/image/chatgpt-image-1-5) (and the original [ChatGPT Image](/ai-models/image/chatgpt-image)) are best when your text references real-world knowledge — logos, flags, diagrams. [Recraft v4](/ai-models/image/recraft-v4) delivers production-quality text with native SVG output. [Z Image Turbo](/ai-models/image/z-image-turbo) and [Qwen Image](/ai-models/image/qwen-image) both excel at bilingual Chinese-English layouts.
| Model | Text accuracy | Multilingual | Best for |
| ------------------------------------------------------- | ---------------------- | -------------------- | ------------------------------------------------------ |
| [GPT Image 2](/ai-models/image/chatgpt-image-2) | 99%+ | CJK, Indic, and more | Multilingual layouts, infographics, knowledge-grounded |
| [Ideogram v3](/ai-models/image/ideogram-v3) | \~90–95% | 6 languages | Posters, packaging, brand typography |
| [ChatGPT Image 1.5](/ai-models/image/chatgpt-image-1-5) | Superior (dense) | Limited | Infographics, fast knowledge-grounded |
| [Recraft v4](/ai-models/image/recraft-v4) | Production-grade | Limited | Design, SVG, signage |
| [Qwen Image](/ai-models/image/qwen-image) | Excellent (complex) | ZH + EN | Bilingual layouts, East Asian markets |
| [Z Image Turbo](/ai-models/image/z-image-turbo) | Excellent (lowest WER) | EN + ZH | Bilingual advertising, fast iteration |
| [Flux.2 Pro](/ai-models/image/flux-2-pro) | Production-grade | Limited | Commercial campaigns with text |
**Priority: identity locking + cross-scene variation**
[Nano Banana Pro](/ai-models/image/nano-banana-pro) preserves up to 5 distinct subject identities simultaneously — the strongest option for mascot or character consistency across scenes. [Seedream v4.5](/ai-models/image/seedream-v4-5) supports up to 14 reference images for multi-character compositions. [Flux.2 Max](/ai-models/image/flux-2-max) accepts up to 10 references.
| Model | Reference images | Consistency | Training needed |
| --------------------------------------------------- | ---------------- | ----------- | --------------- |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | 5 subjects | Excellent | None |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Up to 14 | Excellent | None |
| [Flux.2 Max](/ai-models/image/flux-2-max) | Up to 10 | Strong | None |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | Up to 14 | Good | None |
**Priority: lighting depth + mood + stylistic range**
[Dreamina Image 3.1](/ai-models/image/dreamina-3-1) is designed for cinematic aesthetics — rich lighting, atmospheric depth, and vibrant color palettes across a style range from photorealism to anime. [Midjourney V7](/ai-models/image/midjourney-v7) delivers richer textures and improved anatomical accuracy through its completely rebuilt architecture.
| Model | Cinematic quality | Style range | Best for |
| --------------------------------------------------- | ----------------- | -------------- | --------------------------------- |
| [Midjourney V7](/ai-models/image/midjourney-v7) | Excellent | Wide | Artistic, textured, atmospheric |
| [Dreamina Image 3.1](/ai-models/image/dreamina-3-1) | Excellent | Photo to anime | Cinematic portraits, stylized art |
| [Seedream v4](/ai-models/image/seedream-4) | Strong | Wide | Concept boards, branded imagery |
**Priority: photorealism + scalability + consistency**
[ImagineArt 2.0](/ai-models/image/imagineart-2-0) is the flagship for high-stakes commercial imagery. [Seedream v4.5](/ai-models/image/seedream-v4-5) handles up to 14 references and generates at 4K in seconds — ideal for consistent multi-output campaigns. [Flux.2 Max](/ai-models/image/flux-2-max) with real-time web grounding is best when you need current-subject accuracy.
| Model | Photorealism | Speed | Consistency | Best for |
| --------------------------------------------------- | ------------ | ----- | ----------- | -------------------------------- |
| [ImagineArt 2.0](/ai-models/image/imagineart-2-0) | Flagship | Fast | Excellent | Hero shots, premium commercial |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Excellent | 8–14s | Excellent | Multi-output campaign sets |
| [Flux.2 Max](/ai-models/image/flux-2-max) | Excellent | 4–10s | Strong | Grounded commercial imagery |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | High | Fast | Excellent | High-volume consistent campaigns |
**Priority: style control + detail + creative range**
[Qwen Image](/ai-models/image/qwen-image) is ranked #1 on the AI Arena leaderboard for stylized illustration — exceptional for detailed concept art, scientific illustrations, and complex text-in-image layouts. [Dreamina Image 3.1](/ai-models/image/dreamina-3-1) supports styles from anime to Baroque oil painting. [Ideogram v3](/ai-models/image/ideogram-v3) has 58 style presets including Oil Painting, Cyberpunk, Art Deco, and more.
| Model | Illustration quality | Style presets | Best for |
| --------------------------------------------------- | -------------------- | ------------- | ------------------------------------- |
| [Qwen Image](/ai-models/image/qwen-image) | Excellent | Open-ended | Concept art, scientific illustrations |
| [Dreamina Image 3.1](/ai-models/image/dreamina-3-1) | Excellent | Wide | Anime, fantasy, cinematic stills |
| [Ideogram v3](/ai-models/image/ideogram-v3) | Strong | 58 presets | Styled graphic design, mood art |
**Priority: speed + low cost**
[Nano Banana 2](/ai-models/image/nano-banana-2) is the fastest at \~4× the speed of Nano Banana Pro. [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) generates in 3–5 seconds with intelligent prompt interpretation. [Z Image Turbo](/ai-models/image/z-image-turbo) delivers \~4× faster than FLUX with bilingual capability.
A recommended workflow: iterate on [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) or [Nano Banana 2](/ai-models/image/nano-banana-2) to validate direction, then move to [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro), [Seedream v4.5](/ai-models/image/seedream-v4-5), or [ImagineArt 2.0](/ai-models/image/imagineart-2-0) for the final output.
| Model | Speed | Best for |
| ----------------------------------------------------- | --------------------- | ---------------------------------- |
| [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) | 3–5 seconds | Exploration, intelligent prompting |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | \~4× faster than Pro | High-volume, reference-based |
| [Z Image Turbo](/ai-models/image/z-image-turbo) | \~4× faster than FLUX | Open-source, bilingual drafts |
| [ImagineArt 1.5](/ai-models/image/imagineart-1-5) | Fast | Lower-cost photorealistic drafts |
**Priority: pixel count + detail at scale**
[Seedream v4.5](/ai-models/image/seedream-v4-5) supports up to 8192×8192 — the highest raw resolution available. [Flux.2 Max](/ai-models/image/flux-2-max) and [Flux 1.1 Ultra](/ai-models/image/flux-ultra) both generate at 4MP. [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) delivers 4MP with native SVG vector alongside raster output.
| Model | Max resolution | Key advantage |
| --------------------------------------------------------- | --------------- | -------------------------------- |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Up to 8192×8192 | Highest raw resolution + 14 refs |
| [Flux.2 Max](/ai-models/image/flux-2-max) | 4MP | Web grounding + 10 refs |
| [Flux 1.1 Ultra](/ai-models/image/flux-ultra) | 4MP | Raw mode, architectural |
| [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) | 4MP | SVG vector + raster at 4MP |
| [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) | Native 4K | Composition + text at 4K |
## Full model comparison
| Model | Resolution | Text | Multi-ref | Speed | Best for |
| --------------------------------------------------------------------------- | ------------ | ------------------ | --------- | --------------------- | ----------------------------------------------------- |
| [ImagineArt 2.0](/ai-models/image/imagineart-2-0) | 2K | Good | — | Fast | Flagship photorealism, commercial hero |
| [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) | Native 4K | Strong | — | Fast | Posters, product visuals, professional |
| [ImagineArt 1.5](/ai-models/image/imagineart-1-5) | 2K | Good | — | Fast | Cost-efficient photorealistic drafts |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | Up to 4K | Good | 14 | Ultra-fast | High-volume, rapid iteration |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | Up to 4K | Excellent | 14 | Fast | Complex compositions, production |
| [Nano Banana](/ai-models/image/nano-banana) | Up to 2K | Strong | 4 | Near real-time | Quick edits, product swaps |
| [Nano Banana - Lite](/ai-models/image/nano-banana-lite) | Unconfirmed | Unconfirmed | — | Unconfirmed | Newest Nano Banana tier — specs pending |
| [Seedream V5 pro](/ai-models/image/seedream-v5-pro) | Unconfirmed | Unconfirmed | — | Unconfirmed | Newest Seedream tier — specs pending |
| [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) | Up to 4K | Good (titles) | 14 | 3–5s | Rapid exploration, intelligent prompting |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Native 4K | Excellent (dense) | 4 | 8–14s | Production deliverables, large-format |
| [Seedream v4](/ai-models/image/seedream-4) | Native 4K | Multilingual | 4 | Near real-time | Commercial campaigns |
| [Recraft v4.1](/ai-models/image/recraft-v4-1) | Unconfirmed | Unconfirmed | — | Unconfirmed | Newest Recraft tier — specs pending |
| [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) | View tooltip | Excellent | — | \~30s | Print design, SVG vector |
| [Recraft v4](/ai-models/image/recraft-v4) | View tooltip | Excellent | — | \~10s | Design, branding, SVG |
| [GPT Image 2](/ai-models/image/chatgpt-image-2) | Up to 4K | 99%+, multilingual | 1 | Fastest | Multilingual text, reasoning, character consistency |
| [ChatGPT Image 1.5](/ai-models/image/chatgpt-image-1-5) | 1536×1024 | Superior (dense) | 4 | Fast (4× v1) | Infographics, knowledge-grounded |
| [ChatGPT Image](/ai-models/image/chatgpt-image) | 1536×1024 | Best-in-class | 4 | Up to 2 min | Complex knowledge-grounded |
| [Midjourney V7](/ai-models/image/midjourney-v7) | View tooltip | Good | — | Fast / Draft | Artistic, cinematic, atmospheric |
| [xAI Grok Imagine](/ai-models/image/grok-imagine) | 1K | Good | 3 | Fast | Real-entity accuracy, diverse styles |
| [Flux.2 Max](/ai-models/image/flux-2-max) | View tooltip | Best-in-class | 4 | 4–10s | Maximum quality, web grounding |
| [Flux.2 Pro](/ai-models/image/flux-2-pro) | View tooltip | Production-grade | 4 | Fast | Commercial campaigns with text |
| [Flux 1.1 Ultra](/ai-models/image/flux-ultra) | 2K | Good | — | \~10s | High-res commercial, Raw mode |
| [Krea Flux 1](/ai-models/image/krea-flux-1) | Unconfirmed | Unconfirmed | — | Unconfirmed | Photorealistic, distinctive aesthetic — specs pending |
| [Z Image Turbo](/ai-models/image/z-image-turbo) | View tooltip | Excellent (EN+ZH) | — | \~4× faster than FLUX | Rapid generation, bilingual advertising |
| [Dreamina Image 3.1](/ai-models/image/dreamina-3-1) | 2K | EN + ZH | — | Under 20s | Cinematic portraits, stylized art |
| [Ideogram v4](/ai-models/image/ideogram-v4) | Unconfirmed | Unconfirmed | — | Unconfirmed | General-purpose, "highly aesthetic" — specs pending |
| [Ideogram v3](/ai-models/image/ideogram-v3) | 1K | \~90–95% | — | Flash to Quality | Typography, branding, posters |
| [Minimax Image](/ai-models/image/minimax-image) | 1K | Limited | — | Fast | Portraits, product shots |
| [Qwen Image](/ai-models/image/qwen-image) | 1K | Excellent (ZH+EN) | — | Fast | Illustrations, bilingual, concept art |
| [Stable Diffusion 3.5 Medium](/ai-models/image/stable-diffusion-3-5-medium) | Unconfirmed | Unconfirmed | — | Unconfirmed | New provider (Stability AI) — specs pending |
# Chatgpt image
Source: https://docs.imagine.art/ai-models/image/chatgpt-image
IMAGE MODEL
by OpenAI
gpt-image-1
ChatGPT Image
OpenAI's native image generation built into GPT-4o — grounded in world knowledge for accurate logos, diagrams, and text. Best-in-class for prompt-accurate, text-heavy, and multi-reference visual work.
Resolutions
1024×1024, 1536×1024, 1024×1536
Input refs
Up to 10 images
Editing
Mask-based inpainting
## What makes ChatGPT Image different
ChatGPT Image is built natively into GPT-4o's architecture — not a separate model bolted on. This means it draws on GPT-4o's full knowledge base when composing images: it can accurately render national flags, company logos, scientific diagrams, maps, and UI mockups that most models would get wrong. It's also the best model for infographics and text-heavy layouts where both visual composition and textual accuracy matter.
## Capabilities
Significantly more precise than DALL-E 3. Follows multi-element, multi-constraint prompts with high accuracy.
Accepts up to 10 reference images for editing — combine subjects, backgrounds, products, and styles in a single generation.
Refine images through natural chat context — maintains consistency and intent across multiple iterative edits.
## Specifications
| Feature | Details |
| -------------------------- | ------------------------------------------------- |
| **Model API name** | `gpt-image-1` |
| **Resolutions** | 1024×1024 (1:1), 1536×1024 (3:2), 1024×1536 (2:3) |
| **Quality tiers** | Low, Medium, High |
| **Output formats** | PNG, JPEG, WebP |
| **Transparent background** | Yes (PNG and WebP) |
| **Max reference images** | 10 (for editing workflows) |
| **Released** | March 25, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **ChatGPT Image**.
Write a detailed, structured prompt. ChatGPT Image handles complex multi-element instructions well — be specific about all required components.
Upload up to 10 reference images for compositing or style guidance.
Generate your image. Use follow-up prompts to refine specific elements while maintaining overall composition.
## Prompting tips
* **Describe text content precisely** — Include exact wording, font style, and placement. Example: *"A poster with the title 'Sale Ends Friday' in large bold red sans-serif text at the top."*
* **Use it for knowledge-dependent visuals** — Prompts referencing specific brands, flags, maps, or scientific concepts will produce more accurate results than other models.
* **Multi-step editing** — Generate a base image, then use follow-up instructions to modify specific elements: *"Change the background to a sunset"*, *"Make the text white"*.
* **Be explicit with layout** — For infographics: *"Three-column layout, icons on the left, text on the right of each icon"*.
### Example prompts
> A clean infographic showing the water cycle: evaporation, condensation, precipitation, and collection. Labeled with arrows, minimal design, blue and white color palette.
> A product label for "Alpine Spring Water" with mountain imagery, clean typography, and a blue gradient background. Professional, minimal design.
> A social media post graphic for a coffee shop: warm brown tones, a latte art photo, text reading "Good Morning, Seattle" in serif font, minimal modern layout.
## Compare models
| Model | Text rendering | World knowledge | References | Best for |
| --------------------------------------------------- | --------------------- | --------------- | --------------- | ------------------------------------------------------------ |
| [ChatGPT Image 2](/ai-models/image/chatgpt-image-2) | 99%+, multilingual | Yes (GPT-5.4) | Up to 10 | Multilingual text, reasoning, 4K output |
| **ChatGPT Image** | Best-in-class | Yes (GPT-4o) | Up to 10 | Infographics, text-heavy layouts, knowledge-grounded visuals |
| **Ideogram v3** | Excellent | No | Up to 3 (style) | Typography, posters, brand design |
| **Nano Banana** | Strong | No | Up to 4 | E-commerce, product compositing |
| **Seedream 4.0** | Strong (multilingual) | No | Up to 6 | Commercial campaigns, multilingual markets |
ChatGPT Image uses GPT-4o's architecture to ground image generation in world knowledge, making it particularly effective for prompts that reference specific real-world objects, brands, or concepts that other models typically misrepresent.
# Chatgpt image 1 5
Source: https://docs.imagine.art/ai-models/image/chatgpt-image-1-5
IMAGE MODEL
by OpenAI
gpt-image-1.5
ChatGPT Image 1.5
OpenAI's gpt-image-1.5 — the updated generation of ChatGPT's image model. 4× faster than gpt-image-1, with more reliable editing precision, superior rendering of small and dense text, and a 20% lower cost. The same GPT-4o world-knowledge grounding, now faster and sharper.
Text rendering
Superior (dense + small)
## A faster, sharper ChatGPT image model
gpt-image-1.5 builds on the native multimodal architecture of its predecessor — text and images processed together in a unified neural network — but delivers the output 4× faster at 20% lower cost. The most visible improvements are in editing precision and small text rendering. Edits now preserve lighting, composition, likeness, and appearance while applying changes exactly as instructed.
All the world-knowledge grounding from GPT-4o carries over: accurate depictions of logos, flags, landmarks, scientific diagrams, and knowledge-rich infographics.
## Capabilities
More reliable instruction following during edits preserves lighting, composition, subject likeness, while applying the specified change.
GPT-4o's knowledge enables accurate rendering of real-world entities: brand logos, national flags, diagrams, landmarks.
Accepts up to 10 reference images in a single generation combine products, people, backgrounds, and styles for complex composite outputs.
/
## Specifications
| Feature | Details |
| ------------------------- | ---------------------------------------- |
| **Model ID** | gpt-image-1.5 |
| **Speed vs. gpt-image-1** | 4× faster |
| **Cost vs. gpt-image-1** | 20% cheaper |
| **Reference images** | Up to 10 |
| **Architecture** | Native multimodal (unified text + image) |
| **Released** | December 16, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **ChatGPT 1.5**.
Include real-world references, specific knowledge, or detailed text requirements in your prompt. gpt-image-1.5 excels when your prompt requires world knowledge.
Upload up to 10 reference images for compositing or style anchoring.
Click **Generate**. At 4× the speed of the original, even complex knowledge-grounded generations are fast.
## Prompting tips
* **Name real-world subjects directly** — "A poster celebrating the Apollo 11 mission, with an accurate depiction of the lunar module" leverages GPT-4o's world knowledge for accuracy.
* **Use it for infographic copy** — Dense, labeled infographics with specific text content render cleanly. Describe exact headings, labels, and callout text you want.
* **For editing, be specific about what stays** — "Change the background to a forest but keep the subject's lighting, pose, and outfit exactly as-is."
### Example prompts
> A detailed infographic explaining photosynthesis with labeled diagrams, clean white background, educational illustration style, accurate scientific labels.
> A World War II era propaganda poster aesthetic with bold typography reading "INNOVATION NEVER SLEEPS", dark navy and gold, vintage texture.
> A modern smartphone mockup showing a fitness app dashboard with charts, clean UI design, soft shadow.
## Compare models
| Model | Speed | Text rendering | World knowledge | Best for |
| --------------------------------------------------- | ------------ | ------------------------ | --------------- | --------------------------------------- |
| [ChatGPT Image 2](/ai-models/image/chatgpt-image-2) | Fastest | 99%+, multilingual | Yes (GPT-5.4) | Multilingual text, reasoning, 4K output |
| **ChatGPT Image 1.5** | Fast (4× v1) | Superior (dense + small) | Excellent | Infographics, fast knowledge-grounded |
| [ChatGPT Image](/ai-models/image/chatgpt-image) | Up to 2 min | Best-in-class | Excellent | Complex knowledge-grounded, 10 refs |
| [Ideogram v3](/ai-models/image/ideogram-v3) | Fast | \~90–95% | Limited | Typography, brand design |
ChatGPT Image 1.5 is the updated version of gpt-image-1. If you need the absolute maximum in knowledge-grounded accuracy and can wait for longer generation times, the original [ChatGPT Image](/ai-models/image/chatgpt-image) is still available.
# Chatgpt image 2
Source: https://docs.imagine.art/ai-models/image/chatgpt-image-2
IMAGE MODEL
by OpenAI
gpt-image-2
GPT Image 2
OpenAI's most advanced image model — powered by GPT-5.4. Near-perfect text rendering in any language, reasoning-driven generation, and consistent multi-image output across a single prompt. The benchmark leader for complex, knowledge-grounded, and multilingual visual work.
Text rendering
99%+ accuracy, multilingual
Input refs
Up to 10 images
## What makes GPT Image 2 different
GPT Image 2 is OpenAI's first image model built on GPT-5.4 — their most capable reasoning architecture. Unlike previous image models, gpt-image-2 actively *thinks* before generating: it plans composition, resolves spatial relationships, and interprets multi-part instructions before a single pixel is produced.
The result is near-perfect in-image text accuracy (99%+) across dozens of languages including Chinese, Japanese, Korean, Hindi, and Bengali, comprehensive prompt fidelity for complex multi-element scenes, and character consistency across batches of up to 10 images. It ranks #1 on all Image Arena leaderboards with a +242 point lead at launch.
## Capabilities
99%+ accuracy for in-image text including multilingual scripts — CJK (Chinese, Japanese, Korean), Indic (Hindi, Bengali), and more. The strongest model for infographics, posters, and text-heavy layouts.
Powered by GPT-5.4's reasoning capabilities. The model plans composition, resolves spatial relationships, and interprets complex multi-element prompts before generating — yielding higher instruction fidelity than any prior model.
Generates up to 10 images per prompt while maintaining consistent facial features, clothing, expressions, and visual identity across different scenes and poses.
GPT-5.4's knowledge base enables accurate rendering of logos, national flags, landmarks, scientific diagrams, and UI mockups that other models typically misrepresent.
Describe changes in plain English — the model applies them without requiring manual mask drawing. Also supports mask-based inpainting and outpainting for precise region-level control.
Accepts up to 10 reference images for editing — combine subjects, backgrounds, products, and styles in a single generation with accurate spatial and stylistic coherence.
## Specifications
| Feature | Details |
| -------------------------- | ------------------------------------ |
| **Model API name** | `gpt-image-2` |
| **Max resolution** | Up to 4K |
| **Aspect ratios** | 1:1, 3:4, 4:3, 9:16, 16:9, 3:2, 21:9 |
| **Quality tiers** | Low, Medium, High |
| **Output formats** | PNG, JPEG, WebP |
| **Transparent background** | No |
| **Max reference images** | 10 (for editing workflows) |
| **Architecture** | Native GPT-5.4 multimodal |
| **Released** | April 21, 2026 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **GPT Image 2**.
Write a detailed, structured prompt. GPT Image 2 excels at multi-element instructions — describe text content, spatial relationships, style, and real-world references explicitly.
Upload up to 10 reference images for compositing, style guidance, or character consistency.
Generate your image. Use follow-up prompts to refine specific elements — the model maintains composition intent and subject identity across iterative edits.
## Prompting tips
* **Name text content explicitly** — Include exact wording, language, font style, and placement. Example: *"A poster with the Japanese title '春の祭り' in bold brushstroke style at the top."*
* **Use it for knowledge-dependent visuals** — Prompts referencing specific brands, flags, scientific concepts, or real-world diagrams produce accurate results that other models get wrong.
* **Leverage reasoning for complex scenes** — Describe spatial relationships, layering, and composition constraints directly: *"Three-column infographic: icons left, data center, footnotes right."*
* **For editing, specify what to preserve** — *"Change the background to a night city skyline but keep the subject's lighting, pose, and outfit exactly as-is."*
* **Multi-image consistency** — To generate scene variations, describe all scenes in a single prompt. The model will maintain visual identity across all outputs.
### Example prompts
> A bilingual product packaging label for "Alpine Spring Water" — English headline at top, Japanese subtitle 天然湧水 below, mountain waterfall illustration, clean minimal design, blue and white palette.
> A six-panel manga page: a samurai confronts a dragon in a bamboo forest. Consistent character design, bold linework, speech bubbles with legible Japanese text, dramatic panel transitions.
> A scientific infographic illustrating CRISPR gene editing — labeled molecular diagrams, step-by-step breakdown, clean white background, accurate scientific notation, sans-serif type throughout.
> A social media post for a coffee shop grand opening: warm amber tones, latte art, bold text reading "Now Open — Shibuya, Tokyo" in English and Japanese, minimal modern layout.
## Compare models
| Model | Text rendering | Speed | World knowledge | Best for |
| ------------------------------------------------------- | ------------------------ | ---------------- | ------------------ | ----------------------------------------------------------- |
| **GPT Image 2** | 99%+, multilingual | Fastest | Yes (GPT-5.4) | Multilingual text, complex reasoning, character consistency |
| [ChatGPT Image 1.5](/ai-models/image/chatgpt-image-1-5) | Superior (dense + small) | Fast (4× v1) | Excellent (GPT-4o) | Fast knowledge-grounded infographics |
| [ChatGPT Image](/ai-models/image/chatgpt-image) | Best-in-class | Up to 2 min | Excellent (GPT-4o) | Complex multi-reference compositing |
| [Ideogram v3](/ai-models/image/ideogram-v3) | \~90–95% | Flash to Quality | Limited | Typography, posters, brand design |
GPT Image 2 does not support transparent background output. For images requiring a transparent PNG or WebP with alpha channel, use [ChatGPT Image](/ai-models/image/chatgpt-image) or [ChatGPT Image 1.5](/ai-models/image/chatgpt-image-1-5).
# Dreamina 3 1
Source: https://docs.imagine.art/ai-models/image/dreamina-3-1
IMAGE MODEL
by ByteDance
dreamina-3.1
Dreamina Image 3.1
ByteDance's cinematic image generation model — known for nuanced lighting and reflections, improved human anatomy, and a broad style range from photorealism to anime, illustration, and Baroque-inspired art.
Max dimension
\~1328px/side
Dreamina Image 3.1 is part of ByteDance's Seedream model family. For the latest version with native 4K and multi-reference support, see [Seedream 4.0](/ai-models/image/seedream-4).
## Capabilities
Adapts to documentary, fashion, commercial photography, Baroque, Cubism, ink wash, anime characters, and fantasy scenes.
Place text in quotes within your prompt for improved rendering. Supports English and Chinese text in images.
Claimed 98% prompt adherence for style and composition instructions — reliable for specific creative briefs.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Dreamina Image 3.1**.
Describe the subject, environment, style, and lighting. For text in images, put the exact text in quotation marks within your prompt.
Choose from 1:1, 16:9, 9:16, 4:3, 3:4, 3:2, or 2:3.
Click **Generate** and review the output.
## Prompting tips
* **For text rendering** — Put the exact text you want in the image in quotes: *A poster with "SPRING COLLECTION 2025" in bold serif text.*
* **For cinematic style** — Use descriptive lighting language: *"volumetric lighting"*, *"Rembrandt lighting"*, *"neon-lit rain reflections"*, *"golden hour backlighting"*.
* **For anime or illustration** — Specify art style explicitly: *"anime key visual"*, *"ink wash painting"*, *"Baroque oil painting style"*.
### Example prompts
> A woman in a crimson ball gown standing in an abandoned baroque ballroom, shafts of moonlight streaming through broken windows, atmospheric dust particles, cinematic composition.
> A stylized anime character with silver hair and glowing eyes in a futuristic cityscape, neon lights, detailed illustration style.
## Compare models
| Model | Cinematic style | Anatomy | Multi-reference | Best for |
| ---------------------- | --------------- | -------- | --------------- | --------------------------------- |
| **Dreamina Image 3.1** | Excellent | Improved | No | Cinematic portraits, stylized art |
| **Seedream 4.0** | Excellent | Strong | Up to 6 | Commercial campaigns, 4K output |
| **Seedream v3** | Strong | Good | No | Fast branded imagery |
| **ImagineArt 1.5 Pro** | Excellent | Strong | Multi-ref | Posters, product visuals |
# Flux 2 max
Source: https://docs.imagine.art/ai-models/image/flux-2-max
IMAGE MODEL
by Black Forest Labs
FLUX.2 family
Flux.2 Max
Black Forest Labs' highest-tier FLUX.2 model — 32 billion parameters with the strongest prompt following and style consistency in the family. Supports up to 10 reference images, real-time web search grounding for trending subjects, and 4-megapixel output. The definitive choice when you need maximum quality from the FLUX.2 generation.
Reference Images
Up to 10
## The flagship of FLUX.2
FLUX.2 \[max] combines the 32-billion parameter rectified flow transformer architecture with a Mistral-3 24B vision-language model, creating a 46,864-token context window that enables dramatically more nuanced prompt understanding than any previous FLUX model. The result is stronger adherence to complex multi-element prompts, more consistent character and style identity across references, and grounded generation that can visualize real-world trending subjects.
FLUX.2 Max is the top tier in the FLUX.2 family — above FLUX.2 Pro — for workflows where maximum quality and prompt fidelity are the priority.
## Capabilities
Real-time web search integration enables accurate visualization of trending products, current events, and recently released subjects without detailed descriptions.
Generates up to 4MP (approximately 2752×1536) — print-ready resolution for commercial, editorial, and large-format work.
Exceptionally large context window enables nuanced, multi-paragraph creative briefs and highly specific compositional instructions.
## Specifications
| Feature | Details |
| -------------------- | -------------------------------------------------- |
| **Architecture** | 32B rectified flow transformer + Mistral-3 24B VLM |
| **Context window** | 46,864 tokens |
| **Resolution** | Up to 4MP |
| **Reference images** | Up to 10 |
| **Web grounding** | Yes (real-time search) |
| **Generation time** | 8–10s (4MP), 4–6s (1MP) |
| **Released** | November 25, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Flux.2 Max**.
Be as detailed as needed. With 46,864 tokens of context, Flux.2 Max can handle exhaustive creative briefs — describe spatial relationships, lighting, materials, mood, and style precisely.
Upload up to 10 reference images for character consistency or style anchoring.
Click **Generate**. For most prompts, 1MP generates in 4–6 seconds; 4MP in 8–10 seconds.
## Prompting tips
* **Maximize the context window** — Unlike smaller models, Flux.2 Max won't drop details from long prompts. Use the space to be precise about lighting, materials, spatial composition, and style.
* **Use reference images for consistency** — For campaigns or series, passing reference images maintains visual identity far better than text descriptions alone.
* **Web grounding works for current subjects** — Reference recently launched products or current events by name; grounded generation handles the visual accuracy.
### Example prompts
> A high-end residential kitchen with Calacatta marble countertops, brushed brass fixtures, integrated appliances, warm pendant lighting over a central island, architectural photography style, shot at dusk.
> A futuristic electric sports car (based on \[reference image]) in a wind-tunnel test environment, dramatic directional lighting, photorealistic CGI quality.
## Compare models
| Model | Parameters | Resolution | References | Best for |
| --------------------------------------------- | ---------- | ---------- | -------------- | ---------------------------------- |
| **Flux.2 Max** | 32B | Up to 4MP | 10 | Maximum quality, complex campaigns |
| [Flux.2 Pro](/ai-models/image/flux-2-pro) | 32B | Up to 4MP | 10 | Production-grade, text rendering |
| [Flux 1.1 Ultra](/ai-models/image/flux-ultra) | — | 4MP | Image-to-image | Raw mode, high-res commercial |
| [Flux Dev](/ai-models/image/flux-dev) | 12B | \~1MP | — | Open-weight, creative exploration |
Flux.2 Max sits at the top of the FLUX.2 lineup. For production-grade output with strong text rendering where the absolute ceiling on prompt adherence isn't needed, [Flux.2 Pro](/ai-models/image/flux-2-pro) is a slightly faster and more cost-efficient option.
# Flux 2 pro
Source: https://docs.imagine.art/ai-models/image/flux-2-pro
IMAGE MODEL
by Black Forest Labs
FLUX.2 family
Flux.2 Pro
Black Forest Labs' production-grade FLUX.2 model — 32 billion parameters built for reliable commercial output. The major leap over FLUX.1: legible text in complex typography, true multi-reference character consistency, and real-world grounded spatial logic. Up to 4-megapixel output for print-ready deliverables.
Text rendering
Production-grade
## FLUX rebuilt for production
FLUX.2 Pro is a ground-up 32B rebuild — not a fine-tune of FLUX.1. The model addresses the three most significant FLUX.1 limitations: text-in-image generation that previously failed in production contexts now works reliably for typography, infographics, and UI mockups; character consistency using multiple references is meaningfully stronger; and spatial logic (objects interacting with physical realism, lighting aligned to real-world behavior) is noticeably more accurate.
FLUX.2 Pro sits in the middle of the FLUX.2 family — above FLUX.1 Ultra in the generational hierarchy, and below FLUX.2 Max in the current lineup.
## Capabilities
Objects interact with physical realism — surfaces, shadows, reflections, and spatial relationships align with how light and objects behave in the real world.
Up to 4MP resolution for print-ready commercial deliverables — editorial, packaging, advertising, and large-format output.
More reliable adherence to detailed prompts compared to FLUX.1 — complex spatial instructions, specific material descriptions, and multi-element compositions are followed accurately.
## Specifications
| Feature | Details |
| -------------------- | ------------------------------ |
| **Architecture** | 32B rectified flow transformer |
| **Resolution** | Up to 4MP |
| **Reference images** | Up to 10 |
| **Text rendering** | Production-grade |
| **Released** | November 25, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Flux.2 Pro**.
Include text exactly as you want it to appear in the image — Flux.2 Pro will render it accurately in most configurations.
Upload up to 10 reference images for character or style consistency across a campaign.
Click **Generate** and assess the output at your chosen resolution.
## Prompting tips
* **Include exact text content** — Wrap the text you want rendered in quotes within the prompt: `"A product label with the text 'ARCTIC BLEND' in bold condensed sans-serif, cool blue tones."` The model renders quoted text with high accuracy.
* **Describe lighting physics explicitly** — Flux.2 Pro's improved spatial logic responds well to precise lighting descriptions: "soft diffused overhead light with warm fill from the left, no harsh shadows."
* **Reference images for consistency** — For serial campaigns, attaching reference images produces more reliable identity preservation than relying on text description alone.
### Example prompts
> A premium whiskey bottle label design with the text "OLD HIGHLAND RESERVE" in embossed serif lettering, dark navy background, gold foil accents, traditional distillery aesthetic.
> A fashion editorial spread: a model in a structured white blazer standing in a sun-drenched courtyard, architectural background, high-fashion photography, natural midday light.
## Compare models
| Model | Text rendering | Parameters | Resolution | Best for |
| --------------------------------------------- | ---------------- | ---------- | ---------- | -------------------------------- |
| **Flux.2 Pro** | Production-grade | 32B | Up to 4MP | Commercial campaigns, text-heavy |
| [Flux.2 Max](/ai-models/image/flux-2-max) | Best-in-class | 32B | Up to 4MP | Maximum quality, web grounding |
| [Flux 1.1 Ultra](/ai-models/image/flux-ultra) | Good | — | 4MP | Raw mode, architectural |
| [Flux Dev](/ai-models/image/flux-dev) | Decent | 12B | \~1MP | Open-weight, creative |
Flux.2 Pro is the practical production workhorse in the FLUX.2 lineup. For the absolute maximum in prompt adherence and web-grounded generation, step up to [Flux.2 Max](/ai-models/image/flux-2-max). For open-weight creative work, [Flux Dev](/ai-models/image/flux-dev) remains the right choice.
# Flux ultra
Source: https://docs.imagine.art/ai-models/image/flux-ultra
IMAGE MODEL
by Black Forest Labs
FLUX1.1 \[pro] Ultra
Flux Ultra 1.1
Black Forest Labs' highest-resolution model — up to 4 megapixels of output, with full prompt adherence preserved at scale. Includes Raw mode for authentic, candid photography aesthetics.
Dimensions
Up to 2752×1536
## Two modes
**Ultra Mode** generates images at up to 4 megapixels — four times the resolution of standard FLUX1.1 \[pro] — with full prompt fidelity preserved at the higher resolution. No quality degradation.
**Raw Mode** produces a less synthetic, more natural aesthetic: candid photography style, increased diversity in human subjects, and enhanced realism for nature and documentary scenes. Enable it when you want output that looks like it was captured, not generated.
## Capabilities
Produces candid, photojournalism-style images with authentic lighting variation, more diverse human subjects, and reduced synthetic look.
Accepts a base image as input with adjustable influence strength (0.0–1.0) for style-guided or composition-directed generation.
Optional prompt enhancement automatically expands short prompts into richer descriptions for better results.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Flux Ultra 1.1**.
Select standard **Ultra Mode** for maximum resolution, or enable **Raw Mode** if you want a more candid, photojournalistic aesthetic.
Write a detailed prompt. For Raw mode, describe candid scenes, environmental portraits, or nature photography subjects for the most authentic results.
Choose the aspect ratio. Landscape (16:9) or portrait (9:16) recommended for maximum pixel dimensions.
Click **Generate**. Flux Ultra delivers results in approximately 10 seconds.
## Prompting tips
* **Ultra mode** works best with highly descriptive prompts — include lighting, composition, texture, and atmosphere details to take advantage of the 4MP fidelity.
* **Raw mode** excels with scene-based prompts: *"Street photographer capturing a busy market"*, *"Wildlife portrait of an otter in natural habitat"*, *"Candid portrait, soft natural window light"*.
* **Product photography** — Use Ultra mode with precise material and lighting descriptions: *"Levitating wireless mouse, studio soft box lighting, shadow gradient, minimalist white background"*.
### Example prompts
> Macro photograph of bioluminescent mushrooms glowing in a dark forest, ethereal blue-green light, mist, shallow depth of field, National Geographic style.
> Aerial view of mountain terrain at golden hour, dramatic shadow and light across ridgelines, landscape photography, ultra-detailed.
## Compare models
| Model | Max resolution | Key differentiator | Best for |
| ---------------------- | ----------------- | ----------------------------------- | ---------------------------------------- |
| **Flux Ultra 1.1** | 4MP (2752×1536) | Raw mode, 4× resolution | Print, editorial, commercial photography |
| **Flux Dev** | \~1MP (1024×1024) | Open weights, fine-tuning | Research, experimentation, LoRA base |
| **ImagineArt 1.5 Pro** | Native 4K | Composition control, text rendering | Posters, branding, typography |
| **Seedream 4.0** | Native 4K | Up to 6 references | Multi-reference campaigns |
# Grok imagine
Source: https://docs.imagine.art/ai-models/image/grok-imagine
IMAGE MODEL
by xAI
Aurora architecture
xAI Grok Imagine
xAI's Aurora image generation engine — an autoregressive mixture-of-experts model trained on billions of text-image pairs. Delivers photorealistic output with precise rendering of real-world entities, logos, and text, across styles from cinematic photography to anime and oil painting. Supports prompts up to 10,000 characters.
Architecture
Autoregressive MoE
Prompt Length
Up to 10,000 chars
## Aurora: xAI's image generation engine
Grok Imagine is powered by **Aurora** — xAI's autoregressive mixture-of-experts image generation model. Unlike diffusion models, Aurora generates images autoregressively, which enables strong handling of real-world knowledge: logos, text, brand identities, and named entities render accurately because the model reasons about them rather than interpolating from patterns.
Aurora supports multimodal input (text and existing images) and generates across a wide style range — from photorealistic portraiture to cinematic stills, anime, oil paintings, and pencil sketches.
## Capabilities
Output at 1K (1024×1024) resolution.
Accepts existing images alongside text prompts for editing, style transfer, image-to-image.
Handles extremely detailed, multi-paragraph prompts without truncation
## Specifications
| Feature | Details |
| --------------------- | --------------------------------------- |
| **Underlying model** | Aurora (xAI) |
| **Architecture** | Autoregressive Mixture-of-Experts (MoE) |
| **Resolution** | 1K (1024×1024) |
| **Max prompt length** | 10,000 characters |
| **Input types** | Text + image (multimodal) |
| **Released** | December 2024 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **xAI Grok Imagine**.
Name real-world subjects directly when accuracy matters. Aurora handles known entities, logos, and brands well — you don't need to describe what is widely known.
For image-to-image generation or style transfer, attach your reference image before generating.
Click **Generate** and review the output.
## Prompting tips
* **Name entities explicitly** — "An Aston Martin DB5 on a wet London street at night" is more effective than describing a generic luxury car.
* **Detailed prompts are welcomed** — Up to 10,000 characters means you can include full scene descriptions, mood boards in text, and multi-element briefs without worrying about truncation.
* **For style variety** — Append style directives at the end: "oil painting style", "anime illustration", "pencil sketch on cream paper", "cinematic 35mm film photography."
### Example prompts
> An Airbus A380 flying over the Swiss Alps at golden hour, dramatic clouds, photorealistic aviation photography, sharp detail.
> A samurai warrior standing in a bamboo forest at dawn, morning mist, dramatic back-lighting, anime illustration style, vibrant color palette.
> A vintage travel poster for Kyoto, Japan — Mount Fuji in background, cherry blossom trees, art deco typography, rich color blocks.
Grok Imagine is powered by Aurora, xAI's in-house autoregressive MoE model. Its distinctive strength is accuracy for real-world subjects — logos, brands, named entities, and recognizable objects render with higher fidelity than diffusion models on the same prompts.
# Ideogram v3
Source: https://docs.imagine.art/ai-models/image/ideogram-v3
IMAGE MODEL
by Ideogram AI
Released March 2025
Ideogram v3
The leading model for typography and graphic design — approximately 90–95% text accuracy versus Midjourney's \~30–40%. Combines strong photorealism, flexible styles, and multilingual text rendering for posters, branding, and commercial design.
Style refs
Up to 3 images
Speed options
Flash, Turbo, Default, Quality
Languages
EN, ES, IT, FR, ZH, AR
## Capabilities
Upload up to 3 reference images to define the visual style. Over 4 billion style presets and reusable Style Codes for consistent outputs.
Supports text rendering in Spanish, Italian, French, Chinese, and Arabic alongside English — essential for international campaigns.
Choose from presets including Oil Painting, Watercolor, Pop Art, Cyberpunk, Art Deco, Minimalist, JAPANDI\_FUSION, and more.
## Style types
| Style | Best for |
| ------------- | ------------------------------------------------ |
| **AUTO** | Let the model choose based on prompt content |
| **GENERAL** | Broad creative work, illustrations |
| **REALISTIC** | Photorealistic portraits, commercial photography |
| **DESIGN** | Graphic design, marketing, branding |
| **FICTION** | Fantasy, sci-fi, concept art |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Ideogram v3**.
Select a style type (General, Realistic, Design, or Fiction) to guide the visual direction.
Write your prompt. For text-heavy designs, place the exact text you want rendered in quotes within the prompt.
Upload up to 3 reference images to define the visual style or composition.
Click **Generate**. For quick drafts, use Flash or Turbo. For final deliverables, use Quality.
## Prompting tips
* **Put text content in quotes** — Explicitly quote the text you want rendered: *A poster with "SUMMER SALE" in bold yellow lettering at the top.*
* **Specify font style** — *"serif"*, *"handwritten"*, *"bold sans-serif"*, *"neon glow"* all produce very different text treatments.
* **Use Design style type** for commercial work — activating the DESIGN style type significantly improves layout awareness and graphic structure.
* **Keep multilingual text short** — Shorter words and phrases render more accurately than full sentences, especially in non-Latin alphabets.
### Example prompts
> A vintage circus poster with "THE GRAND SHOW" in large ornate typography, illustrated animals in the background, warm sepia tones, Art Deco style.
> A modern product label for a craft gin bottle reading "Northern Lights" with aurora imagery and minimalist Scandinavian design.
## Compare models
| Model | Text accuracy | Style refs | Photorealism | Best for |
| ---------------------- | --------------------- | ---------------- | ------------ | ---------------------------------------- |
| **Ideogram v3** | \~90–95% | Up to 3 | Strong | Typography, posters, brand design |
| **ChatGPT Image** | Best-in-class | Up to 10 | Very strong | Knowledge-grounded visuals, infographics |
| **ImagineArt 1.5 Pro** | Strong | Reference images | Excellent | Poster design, product visuals |
| **Seedream 4.0** | Strong (multilingual) | Up to 6 | Excellent | Commercial campaigns |
# Ideogram v4
Source: https://docs.imagine.art/ai-models/image/ideogram-v4
## Ideogram v4
Ideogram's newest model, live in the Image tool's model picker alongside [Ideogram v3](/ai-models/image/ideogram-v3) — this is an addition to the lineup, not a replacement. In-app it's described simply as a "Highly aesthetic, general-purpose model," a broader positioning than v3's typography-first specialization.
This page covers what's confirmed live in the product so far: the model exists, is badged "new," with the tagline above. Detailed specs (text accuracy, style presets, resolution) haven't been independently verified against v3's documented \~90–95% text accuracy and 58 style presets — treat this as a placeholder pending a fuller pass.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Ideogram v4**.
Write your prompt and generate as normal.
## Compare models
| Model | Positioning | Notes |
| ------------------------------------------- | ----------------------------------- | ------------------------------------------------- |
| **Ideogram v4** | General-purpose, "highly aesthetic" | Newest in the lineup — specs pending verification |
| [Ideogram v3](/ai-models/image/ideogram-v3) | Typography-focused | \~90–95% text accuracy, 58 style presets |
# Imagineart 1 5
Source: https://docs.imagine.art/ai-models/image/imagineart-1-5
IMAGE MODEL
by ImagineArt
ImagineArt 1.5
ImagineArt's accessible photorealistic generation model — built on the same generation as ImagineArt 1.5 Pro at a lower cost tier. High-quality true-to-life imagery for portraits, lifestyle photography, and realistic scenes. The right choice when you need reliable photorealism without the 4K overhead of the Pro variant.
Best for
Portraits, lifestyle, scenes
## Reliable photorealism at lower cost
ImagineArt 1.5 shares the same architectural generation as ImagineArt 1.5 Pro — the same photorealistic quality pipeline — at a standard resolution output and lower credit cost. It's the ideal choice when you need dependable photorealism for portraits, lifestyle images, and realistic scenes without requiring the 4K native output that the Pro variant delivers.
A recommended workflow: iterate on ImagineArt 1.5 to validate composition, lighting, and subject, then move to [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) for the final high-resolution deliverable.
## Capabilities
Natural environments, urban scenes, and architectural compositions with realistic depth, atmospheric perspective, and color accuracy.
Proper light behavior: natural shadows, ambient occlusion, golden hour gradients, and studio lighting setups that read as genuinely photographic.
More cost-efficient than ImagineArt 1.5 Pro — generate multiple variations to explore direction before committing to a full-resolution final.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **ImagineArt 1.5**.
Use descriptive, photography-style prompts — lighting conditions, environment, subject details, and any stylistic references.
Choose the aspect ratio for your subject — portrait for people, landscape for environments and wide scenes.
Click **Generate**. Use the lower cost to explore multiple directions before moving to 1.5 Pro for the final output.
## Prompting tips
* **Describe the photography style** — "Natural light portrait", "golden hour landscape", "studio product photography" guide the model toward specific photorealistic aesthetics.
* **Camera details improve results** — "85mm portrait lens", "wide-angle establishing shot", "macro close-up" influence framing and perspective significantly.
* **Avoid overly stylized requests** — ImagineArt 1.5 is optimized for photorealism. For illustrations or artistic styles, use a different model.
### Example prompts
> A 30-something man with natural stubble, warm hazel eyes, photographed outdoors in soft overcast light, shallow depth of field, lifestyle portrait photography.
> A coastal village at sunset, warm amber light on whitewashed buildings, gentle waves, Mediterranean atmosphere, landscape photography.
## When to use 1.5 vs other ImagineArt models
| Need | Recommended model |
| ---------------------------- | ---------------------- |
| Quick photorealistic draft | **ImagineArt 1.5** |
| Final 4K output with text | **ImagineArt 1.5 Pro** |
| Latest flagship capabilities | **ImagineArt 2.0** |
ImagineArt 1.5 is the most cost-efficient entry in the ImagineArt model family. Use it for rapid exploration and iteration, then step up to [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) or [ImagineArt 2.0](/ai-models/image/imagineart-2-0) for final production output.
# Imagineart 1 5 pro
Source: https://docs.imagine.art/ai-models/image/imagineart-1-5-pro
IMAGE MODEL
by ImagineArt
ImagineArt 1.5 Pro
ImagineArt's flagship image model — native 4K output, enhanced realism, accurate text rendering, and composition intelligence for professional-grade creative work.
## Capabilities
Generate images directly at 4K resolution. Fine details, textures, and typography stay crisp — ideal for posters, banners, and large-format visuals.
Refined lighting, depth, and material handling produce cohesive visuals with natural shadows, accurate skin tones, and realistic reflective surfaces.
Maintains spatial relationships and visual hierarchy across complex multi-element compositions, keeping designs balanced and structured.
Sharp, correctly spelled text at 4K resolution. Ideal for posters, advertisements, and typography-forward editorial layouts.
Colors apply accurately to intended objects or regions — critical for brand designs and product concepts where color fidelity matters.
Refined facial expressions and emotional nuance make it well-suited for portraits, cinematic stills, and narrative-driven artwork.
## How to use
Navigate to the **AI Image Generator** in ImagineArt.
Open the model dropdown and choose **ImagineArt 1.5 Pro**.
Write a clear, descriptive prompt. Include style cues, text instructions, or upload reference images to guide the model toward your intended result.
Apply visual styles, effects, or color palettes to match your brand or creative direction.
Choose the aspect ratio that fits your project — 1:1 for social, 16:9 for widescreen, 4:3 for editorial, and so on.
Click **Generate** and let the model produce your image.
## Use cases
Native 4K support and strong composition handling make ImagineArt 1.5 Pro ideal for posters, banners, and complex visual layouts. The model's spatial awareness keeps multi-element designs structured and balanced.
Accurate lighting, materials, and realism make it well-suited for product mockups and concept designs where photographic quality is essential.
Clear, accurate text rendering at 4K makes it the right choice when typography is the focal point — advertisements, editorial layouts, and branded text overlays.
Enhanced realism and emotional detail make it effective for cinematic stills, storytelling-driven projects, and portrait work requiring nuanced facial expression.
## Compare models
| Model | Visual style | Strengths | Best for |
| ---------------------- | ------------------------ | -------------------------------------------- | ------------------------------------ |
| **ImagineArt 1.5 Pro** | Realistic, cinematic | Native 4K, strong composition, accurate text | Posters, product visuals, typography |
| **Flux 2** | Stylized, artistic | Strong color interpretation, fast generation | Concept art, abstract scenes |
| **Google Imagen 4** | Photorealism, clean | Sharp detail, accurate text, fast generation | Marketing, editorial work |
| **Seedream 4.0** | Photorealistic, adaptive | Multi-reference (up to 6), fast generation | Branded imagery, complex campaigns |
| **Nano Banana** | Photorealistic, adaptive | One-shot accuracy, multi-image editing | Product mockups, e-commerce |
# Imagineart 2 0
Source: https://docs.imagine.art/ai-models/image/imagineart-2-0
IMAGE MODEL
by ImagineArt
Next generation
ImagineArt 2.0
ImagineArt's most advanced proprietary image model — the next generation beyond the 1.5 family. Delivers measurably higher detail fidelity, better understanding of complex prompts, enhanced composition control, and more naturalistic photorealism. The flagship for professional creative and commercial work on the ImagineArt platform.
Generation
2.0 (Next gen)
Best for
Professional, commercial
## ImagineArt's most capable model
ImagineArt 2.0 represents the next generation of ImagineArt's in-house model development — a step change in capability beyond the 1.5 family. The model delivers noticeably higher detail fidelity across all subject types, more reliable interpretation of complex and nuanced prompts, and photorealism that sits closer to actual photography in lighting behavior, texture accuracy, and spatial depth.
For work where the ImagineArt 1.5 Pro quality ceiling isn't enough — high-stakes commercial campaigns, hero shots, and premium editorial — ImagineArt 2.0 is the model to use.
## Capabilities
Higher detail fidelity than the 1.5 family — surfaces, materials, lighting, and depth render with a closer resemblance to actual photography.
Better understanding of complex, multi-element prompts — nuanced descriptions of mood, lighting, composition, and subject relationships are followed more precisely.
Stronger compositional logic — subject placement, framing, depth of field, and spatial relationships are handled with more professional-grade intent.
The most accurate rendering of human subjects in ImagineArt's lineup — skin, micro-expressions, and body proportions with exceptional detail.
Suitable for high-stakes commercial work — campaigns, product launches, editorial features, and hero imagery that needs to hold up at full size.
Excels across portraits, landscapes, product photography, architectural shots, and lifestyle imagery — a versatile flagship for all photorealistic use cases.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **ImagineArt 2.0**.
Be specific about lighting, mood, subject, and style. ImagineArt 2.0's enhanced comprehension means even nuanced details in the prompt will be reflected in the output.
Choose the output dimensions appropriate for your deliverable.
Click **Generate** and review the output. ImagineArt 2.0 is built for final-quality deliverables.
## Prompting tips
* **Be specific about mood and atmosphere** — "Warm, late-afternoon light casting long shadows across the subject, soft golden haze, candid portrait aesthetic" produces richer results with ImagineArt 2.0 than with previous models.
* **Camera-language prompts work well** — f/1.4 aperture, 50mm lens, natural window light — the model interprets photographic vocabulary accurately.
* **Use it for the final deliverable** — Draft with ImagineArt 1.5 or 1.5 Pro, then run the finalized prompt on ImagineArt 2.0 for the production output.
### Example prompts
> A fashion editorial portrait of a woman in a structured burgundy overcoat, cold overcast light, urban background slightly out of focus, high-fashion photography, Vogue aesthetic.
> An architectural exterior shot of a contemporary home at dusk, warm interior lights contrasting cool twilight sky, lush garden, professional architectural photography.
> A premium skincare product hero shot: a glass serum bottle on a wet black stone surface, macro lens, dramatic directional light with reflections, commercial photography.
## ImagineArt model comparison
| Model | Quality tier | Best for |
| --------------------------------------------------------- | ------------ | ---------------------------------- |
| **ImagineArt 2.0** | Flagship | High-stakes commercial, hero shots |
| [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) | Professional | 4K output, composition control |
| [ImagineArt 1.5](/ai-models/image/imagineart-1-5) | Standard | Cost-efficient photorealism |
ImagineArt 2.0 is ImagineArt's highest-capability proprietary model. For most photorealistic work, [ImagineArt 1.5 Pro](/ai-models/image/imagineart-1-5-pro) or [ImagineArt 1.5](/ai-models/image/imagineart-1-5) deliver excellent results at lower cost — use 2.0 when you need the platform's absolute best.
# Krea flux 1
Source: https://docs.imagine.art/ai-models/image/krea-flux-1
## Krea Flux 1
A new model from Krea (built on Black Forest Labs' FLUX architecture), live in the Image tool's model picker. In-app it's described as "Photorealistic AI image generation with a distinctive aesthetic" — no "new" badge was observed, but it wasn't documented before this pass.
This page covers what's confirmed live in the product so far: the model exists, with the tagline above. Detailed specs (resolution, speed, reference support) haven't been independently verified — treat this as a placeholder pending a fuller pass.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Krea Flux 1**.
Write your prompt and generate as normal.
# Midjourney v7
Source: https://docs.imagine.art/ai-models/image/midjourney-v7
IMAGE MODEL
by Midjourney
Complete architectural rebuild
Midjourney V7
Midjourney's most ambitious release — a completely rebuilt architecture delivering richer textures, dramatically improved human anatomy and hands, precise prompt adherence, and personalized output that learns your aesthetic. Draft Mode generates 10× faster at half the cost for rapid exploration.
Draft Mode
10× faster, half cost
Architecture
Complete rebuild
Personalization
Built-in, default on
## A completely different architecture
Midjourney V7 is not an incremental update — it's a ground-up rebuild. CEO David Holz described it as "a totally different architecture." The result is visibly richer textures, more coherent spatial compositions, and significantly improved human anatomy (especially hands and bodies) that have historically been a weak point for generative image models.
Prompt adherence is the headline improvement: V7 follows complex, multi-element instructions with far fewer dropped or altered elements compared to V6.
## Capabilities
Default personalization learns your aesthetic preferences from your generation history. A brief rating session calibrates the model to your taste.
Unified reference system for style, subject, and character references. Use --sref codes to explore curated style galleries or lock a specific look.
Markedly improved material rendering — fabric, skin, stone, and organic surfaces look more convincing. Human hands and body proportions are significantly more accurate.
## Key parameters
| Parameter | Range | Effect |
| --------------------- | ------ | ----------------------------------------------------- |
| `--s` (stylization) | 0–1000 | Low = prompt-faithful; High = artistic interpretation |
| `--c` (chaos) | 0–100 | Low = consistent; High = unusual, varied results |
| `--exp` | 0–100 | Experimental aesthetics; suggested 5, 10, 25, 50, 100 |
| `--sw` (style weight) | 0–1000 | Strength of `--sref` style reference (default 100) |
| Draft Mode | On/Off | 10× faster at half cost, reduced quality |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Midjourney V7**.
Be specific about subject, mood, lighting, and style. V7 rewards precise prompts — the more detail you include, the more faithfully the output matches your intent.
Enable Draft Mode to explore compositions quickly. Switch to full quality once you've found the direction you want.
Adjust `--s` for artistic interpretation, `--exp` for experimental aesthetics, or `--sref` to anchor to a specific visual style.
## Prompting tips
* **Draft Mode first, always** — At 10× speed and half cost, Draft Mode is your ideation layer. Lock composition and mood before committing credits to full quality.
* **Use `--exp 10–25` for creative work** — Low `--exp` values push the model into more unusual aesthetic territory without losing coherence.
* **Personalization makes a visible difference** — If you've rated images or generated extensively in Midjourney, V7's personalization will skew outputs toward your established preferences automatically.
* **Be explicit about anatomy** — Despite improvements, specifying "correct hand anatomy" or "natural posture" in prompts reinforces V7's improved but not flawless anatomical rendering.
### Example prompts
> A lone lighthouse on a rocky Atlantic coast at dusk, crashing waves, deep blue and orange sky, cinematic composition, photorealistic.
> A 1970s Tokyo neon-lit street at night, wet cobblestones reflecting signs, rain, cinematic film photography aesthetic, rich shadow detail.
> An architect's desk covered in blueprints and scale models, warm desk lamp, afternoon window light, documentary photography.
Midjourney V7 became the default model on June 17, 2025. For the most experimental outputs, try `--exp 25` or higher alongside a reduced `--s` value to balance creative expression with prompt fidelity.
# Minimax image
Source: https://docs.imagine.art/ai-models/image/minimax-image
IMAGE MODEL
by MiniMax AI
Image-01
Minimax Image
MiniMax's photorealistic image model — known for lifelike human subjects, cinematic lighting, and complex environmental scenes. Features subject reference support for consistent character generation across multiple outputs.
Aspect ratios
1:1, 16:9, 4:3, 9:16, 21:9 +
## Capabilities
Accepts a subject reference image to maintain consistent character appearance across multiple generated images.
Generate from 512px to 2048px on each side independently — supports 1:1, 16:9, 4:3, 21:9, 9:16, and more.
Handles intricate product details, fabric textures, and complex material surfaces with high fidelity.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Minimax Image**.
Describe the subject, environment, lighting, and style. Minimax Image responds well to cinematic and photographic language.
To maintain character consistency across multiple images, upload a reference photo using the subject reference input.
Choose the aspect ratio and pixel dimensions for your output.
Click **Generate** and review the output.
## Prompting tips
* **Describe the lighting explicitly** — *"soft natural window light"*, *"dramatic rim lighting"*, *"overcast studio"* significantly influence the photorealistic output quality.
* **Use subject reference for campaigns** — For consistent character appearances across multiple images (product shots, brand campaigns), always upload a subject reference.
* **Specify materials for product work** — *"brushed stainless steel"*, *"glossy ceramic"*, *"matte leather"* push the model toward more accurate material rendering.
### Example prompts
> A woman in a red silk dress standing in a rain-soaked alley at night, neon reflections on the wet pavement, cinematic composition, shallow depth of field.
> A luxury perfume bottle on a marble surface, soft diffused studio lighting, white background, product photography, ultra-detailed.
## Compare models
| Model | Portraits | Products | Subject consistency | Best for |
| ---------------------- | --------- | --------- | ----------------------- | ----------------------------------- |
| **Minimax Image** | Excellent | Excellent | Yes (subject reference) | Portraits, product shots, campaigns |
| **Nano Banana** | Strong | Excellent | Yes (up to 4 refs) | E-commerce, multi-image editing |
| **ImagineArt 1.5 Pro** | Strong | Excellent | Multi-ref support | Posters, branding, typography |
| **Google Imagen 4** | Excellent | Excellent | Limited | Marketing, editorial |
# Nano banana
Source: https://docs.imagine.art/ai-models/image/nano-banana
IMAGE MODEL
by Google DeepMind
Gemini 2.5 Flash Image
Nano Banana
Google's precision image model — known for one-shot accuracy, multi-image editing up to 4 references, and near real-time speed. Built for branding, e-commerce, and campaign workflows.
Nano Banana is available exclusively through ImagineArt, integrated into both the AI Image Generator and the Image Editor.
## Capabilities
Maintains subject identity, lighting, and style coherence across multiple outputs — ideal for brand campaigns and product series.
Delivers high-quality results at near real-time speeds — efficient for quick iteration and large-scale projects alike.
Adapts to text-to-image generation, guided image editing, and multi-reference design creation across different workflows.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Google Nano Banana**.
Write your text prompt and upload up to 4 reference images for richer guidance.
Click **Generate** and wait for the model to process your request.
Go to the **ImagineArt Image Editor**.
Click **Settings** and select **Google Nano Banana** from the model dropdown.
Apply edits or transformations directly to your uploaded visuals.
Use **Visual Prompts** to make further adjustments, add annotations, or overlay images on your base image.
Go to the **ImagineArt AI Video Generator**.
Choose **Kling 2.1** or your desired video model from the dropdown.
Set the resolution to **1080p**.
Use **Google Nano Banana** to edit and refine the start and end images before generating the video.
## Prompting tips
Nano Banana is optimized for one-shot accuracy — precise, structured prompts produce the best results.
1. **Be specific** — Define the subject, background, product, and style in one prompt. Example: *"A man in a modern kitchen holding a coffee mug, styled like a product ad."*
2. **Use multiple references** — Upload images of the subject, product, and setting together to maintain consistency across outputs.
3. **Guide the mood** — Include terms like *"cinematic"*, *"flat lay"*, or *"studio lighting"* to shape composition and tone.
### Example prompts
> A fluid gender albino person inside a vintage car with bright yellow chrysanthemums spilling out through the windows, holding a baby blue handbag. The scene blends nature and fashion with a dreamlike atmosphere, shot in a cinematic and analog photo style.
> A sleek modern kitchen with stainless steel appliances and a marble countertop, featuring a woman in a red dress holding a glass of wine. The scene should reflect contemporary elegance, with soft lighting highlighting the product and surroundings.
## Compare models
| Model | Visual style | Strengths | Best for |
| ---------------------- | -------------------------- | -------------------------------------- | ----------------------------------- |
| **Nano Banana** | Photorealistic, adaptive | One-shot accuracy, multi-image editing | Image editing, branding, e-commerce |
| **Flux Kontext** | Artistic, flexible | Strong stylization, precision control | Creative edits, experimental work |
| **Seedream 4.0** | Photorealistic, consistent | Up to 6 references, native 4K | Complex campaigns, high-res assets |
| **ImagineArt 1.5 Pro** | Realistic, cinematic | Native 4K, composition control | Posters, typography-based designs |
# Nano banana 2
Source: https://docs.imagine.art/ai-models/image/nano-banana-2
IMAGE MODEL
by Google DeepMind
Gemini 3.1 Flash Image
Nano Banana 2
Google's Gemini 3.1 Flash Image model — the fastest in the Nano Banana lineup. Approximately 4× faster than Nano Banana Pro, with support for up to 14 reference images, real-time web search grounding, and up to 4K output. Built for high-volume workflows and rapid iteration at scale.
Speed
\~4× faster than Pro
Reference Images
Up to 14
## The fastest Nano Banana
Nano Banana 2 is powered by **Gemini 3.1 Flash Image** — Google's most efficient image generation model. It delivers approximately 4× faster generation than Nano Banana Pro at roughly half the cost, while maintaining high output quality. At 1K resolution, images generate in 5–15 seconds; 4K takes 15–40 seconds.
The model integrates **real-time web search grounding**, enabling accurate rendering of current products, events, and entities without requiring detailed prompts.
## Capabilities
Integrated Google Search grounding renders current products, logos, and real-world entities accurately — no need to over-describe known subjects.
Supports resolutions from 512px up to 4096px on either axis, with flexible aspect ratios across standard and custom dimensions.
Accurate text rendering and data visualization capabilities — suitable for infographic drafts, labeled diagrams, and product callouts.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------- |
| **Underlying model** | Gemini 3.1 Flash Image |
| **Resolution range** | 512px – 4096px (per axis) |
| **Reference images** | Up to 14 per generation |
| **Generation speed** | 5–15s (1K), 15–40s (4K) |
| **Context window** | 131,072 tokens |
| **Web grounding** | Yes (real-time Google Search) |
| **Released** | February 26, 2026 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Nano Banana 2**.
Write a focused prompt. For reference-based generation, attach your reference images before generating.
Upload up to 14 reference images for product consistency, character matching, or style anchoring.
Click **Generate**. Nano Banana 2 is the fastest way to explore concepts before committing to a higher-resolution model.
## Prompting tips
* **Use it for exploration first** — Nano Banana 2's speed makes it ideal for validating compositions before upgrading to Nano Banana Pro or ImagineArt 2.0 for final output.
* **Reference images do the heavy lifting** — Rather than describing every detail in text, pass reference images. The model excels at compositing from visual references.
* **Let web grounding handle known subjects** — For branded products, logos, or real-world locations, name them directly instead of describing them. The grounding handles the rest.
### Example prompts
> A sleek wireless speaker on a white marble surface, soft diffused studio lighting, product photography, crisp shadow.
> A woman in a yellow raincoat standing on a mountain ridge, dramatic storm clouds, cinematic photography.
## Compare models
| Model | Speed | Reference images | Resolution | Best for |
| --------------------------------------------------- | -------------------- | ---------------- | ---------- | -------------------------------- |
| **Nano Banana 2** | \~4× faster than Pro | Up to 14 | Up to 4K | Rapid iteration, high-volume |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | Fast | 5 subjects | Up to 4K | Complex compositions, production |
| [Nano Banana](/ai-models/image/nano-banana) | Near real-time | Up to 4 | Up to 2K | Quick edits, product swaps |
Nano Banana 2 is the best starting point for any high-volume workflow. Generate at 1K to validate direction, then switch to Nano Banana Pro or ImagineArt 1.5 Pro for final 4K output.
# Nano banana lite
Source: https://docs.imagine.art/ai-models/image/nano-banana-lite
## Nano Banana - Lite
A new, lighter tier in Google's Nano Banana lineup, live in the Image tool's model picker alongside [Nano Banana](/ai-models/image/nano-banana), [Nano Banana 2](/ai-models/image/nano-banana-2), and [Nano Banana Pro](/ai-models/image/nano-banana-pro) — an addition to the lineup, not a replacement for any of them. In-app it shares the same tagline as Nano Banana 2 ("Next-level multi-image generation and editing"), so its actual differentiation from the rest of the lineup isn't yet clear from the product UI alone.
This page covers what's confirmed live in the product so far: the model exists, with the tagline above (no "new" badge observed). Detailed specs and how it differs from the other three Nano Banana tiers haven't been independently verified — treat this as a placeholder pending a fuller pass.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Nano Banana - Lite**.
Write your prompt and generate as normal.
## Compare models
| Model | Speed | Reference images | Resolution |
| --------------------------------------------------- | -------------------- | ---------------- | ----------- |
| **Nano Banana - Lite** | Unconfirmed | Unconfirmed | Unconfirmed |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | \~4× faster than Pro | Up to 14 | Up to 4K |
| [Nano Banana Pro](/ai-models/image/nano-banana-pro) | Fast | 5 subjects | Up to 4K |
| [Nano Banana](/ai-models/image/nano-banana) | Near real-time | Up to 4 | Up to 2K |
# Nano banana pro
Source: https://docs.imagine.art/ai-models/image/nano-banana-pro
IMAGE MODEL
by Google DeepMind
Gemini 3 Pro Image
Nano Banana Pro
Google's Gemini 3 Pro Image model — built for production-ready complex compositions. Supports up to 14 reference images, localized edits and camera transformations, and uses Google Search grounding for real-world accuracy. The professional tier in the Nano Banana lineup.
Identity Preservation
Up to 14
Context Window
65,536 tokens
## Built for complex, production-grade work
Nano Banana Pro is powered by **Gemini 3 Pro Image** — Google's most capable image generation model. Where Nano Banana 2 excels at speed and volume, Nano Banana Pro excels at depth: multi-subject identity preservation, precise localized edits, camera transformations, and highly detailed prompt adherence. It's the right choice when your output needs to hold up to client scrutiny.
## Capabilities
Real-world grounding via Google Search produces accurate depictions of real products, brands, places, and entities without overprompting.
Industry-leading text quality, including long passages, multilingual layouts, data visualizations, and typographic compositions.
Blend and composite multiple reference images while maintaining consistent lighting, perspective, and style across the final output.
## Specifications
| Feature | Details |
| -------------------- | ---------------------------------- |
| **Underlying model** | Gemini 3 Pro Image |
| **Resolution** | Up to 4K (flexible aspect ratios) |
| **Reference images** | Up to 14 |
| **Context window** | 65,536 input tokens |
| **Web grounding** | Yes (Google Search) |
| **Editing** | Localized, lighting, focus, camera |
| **Released** | November 20, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Nano Banana Pro**.
Be specific about subjects, lighting, and composition. Nano Banana Pro follows complex multi-part instructions reliably.
For multi-subject consistency, upload your reference images. The model preserves identity across all subjects.
Review the output and use localized edit instructions to refine specific areas without regenerating the whole image.
## Prompting tips
* **Name your subjects explicitly** — "A woman in a red blazer (Subject A) and a man in a navy suit (Subject B) shaking hands" gives the model clear anchors for identity preservation.
* **Use localized edit instructions** — After generating, follow up with targeted edits: "Change the background to a blurred office interior" or "Shift the lighting to golden hour."
* **Let Google grounding handle real-world references** — For known brands, locations, or products, name them directly. Nano Banana Pro will render them accurately.
### Example prompts
> A product shoot featuring a cream-colored luxury skincare bottle (Subject A) and a matching moisturizer jar (Subject B) on a pale stone surface, dramatic side lighting, high-end editorial photography.
> A tech CEO (Subject A) being interviewed on stage at a large conference, blurred audience in background, professional event photography, warm stage lighting.
## Compare models
| Model | Best for | Identity preservation | Speed |
| ----------------------------------------------- | -------------------------------- | --------------------- | -------------- |
| **Nano Banana Pro** | Complex compositions, production | Up to 5 subjects | Fast |
| [Nano Banana 2](/ai-models/image/nano-banana-2) | Rapid iteration, high volume | Up to 14 refs | \~4× faster |
| [Nano Banana](/ai-models/image/nano-banana) | Quick edits, product swaps | Up to 4 refs | Near real-time |
Nano Banana Pro is the recommended choice when final output needs to meet professional quality standards. Use [Nano Banana 2](/ai-models/image/nano-banana-2) for early-stage exploration, then switch to Pro for the polished deliverable.
# Qwen image
Source: https://docs.imagine.art/ai-models/image/qwen-image
IMAGE MODEL
by Alibaba
Apache 2.0
Qwen Image
Alibaba's open-source image generation model — ranked #1 on the AI Arena leaderboard for text-to-image and image editing. Exceptional for complex text rendering, bilingual layouts, and stylized illustrations with enhanced lighting and texture detail.
Resolution
Native 2K (2048×2048)
Languages
EN + Chinese (bilingual)
License
Apache 2.0 (Open source)
## Capabilities
The unified Qwen-Image-2.0 model handles generation and editing in one — style transfer, object insertion/removal, text editing within images, pose manipulation.
Detailed facial features, skin textures, pores, and natural lighting for high-quality portrait and lifestyle photography.
Generates at 2048×2048 natively — no upscaling. Supports custom dimensions from 512px to 2048px on each side.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Qwen Image**.
Write a detailed prompt. For text-in-image generation, include the exact text you want rendered and describe its layout and style.
Choose your output dimensions (512px–2048px per side) and aspect ratio.
Click **Generate** and review the output.
## Prompting tips
* **For text-in-image** — Describe exactly what text should appear and how: *"A book cover with the title 'The Last Algorithm' in large silver metallic lettering, sci-fi aesthetic."*
* **For bilingual layouts** — Specify both the English and Chinese text content and their relative positions for accurate rendering.
* **For illustrations** — Use rich descriptive vocabulary: *"hyperdetailed concept art"*, *"scientific illustration style"*, *"editorial infographic"* push the model toward its strongest outputs.
* **For editing** — Describe changes clearly: *"change the sky to a sunset"*, *"add a coffee cup on the desk"*, *"make the text red"*.
### Example prompts
> A detailed scientific diagram of a human cell with labeled organelles — nucleus, mitochondria, Golgi apparatus, endoplasmic reticulum — clean white background, educational illustration style.
> A Chinese New Year poster with bold red typography reading "新年快乐" (Happy New Year) and "2026" in gold, traditional patterns, fireworks illustration.
> A book cover for a fantasy novel: "The Iron Sky" in large ornate metallic lettering, dramatic storm clouds, a lone figure silhouetted on a rocky cliff.
## Compare models
| Model | Text rendering | Bilingual | Image editing | Open source | Best for |
| ----------------- | --------------------------- | --------------------- | ---------------- | ---------------- | -------------------------------------------- |
| **Qwen Image** | Excellent (complex layouts) | Yes (ZH + EN) | Yes (unified) | Yes (Apache 2.0) | Infographics, bilingual design, illustration |
| **Ideogram v3** | \~90–95% | Limited (6 languages) | Limited | No | Typography, brand design |
| **ChatGPT Image** | Best-in-class | Limited | Yes (mask-based) | No | Knowledge-grounded visuals |
| **Seedream 4.0** | Strong (multilingual) | Yes | No | No | Commercial campaigns |
# Recraft v4
Source: https://docs.imagine.art/ai-models/image/recraft-v4
IMAGE MODEL
by Recraft AI
#1 HuggingFace Arena
Recraft v4
Recraft AI's ground-up rebuilt image model — ranked #1 on the HuggingFace Text-to-Image Arena with an ELO of 1172 and a 72% win rate against Midjourney, DALL-E 3, and FLUX. Built with professional designers to prioritize design taste: composition, lighting, material realism, and precise text rendering. The only model in this lineup with native SVG vector output.
Arena Ranking
#1 (ELO 1172)
## Design taste as a first principle
Recraft v4 was built from scratch with professional designers — not just trained on datasets, but built alongside the people who use these outputs professionally. The result is a model that treats composition, color relationships, material realism, and typography as first-class concerns rather than afterthoughts. It outperforms Midjourney V8 and DALL-E 3 by 188 ELO points in head-to-head arena evaluations.
The standout capability: **native SVG vector output** — the only generative model that produces editable, layered vector files directly from a text prompt.
/
## Capabilities
Reliable, sharp text-in-image generation for infographics, menus, signage, packaging, and UI mockups — with accurate multi-line layout support.
Trained with professional designers to produce images with strong compositional logic, cohesive color relationships, and authentic material depth.
Handles multi-element instructions with minimal element dropout — subject relationships, spatial positioning, and stylistic directives are followed accurately.
## Specifications
| Feature | Details |
| --------------------- | --------------------------- |
| **Resolution** | 1024×1024 (1MP) |
| **Generation time** | \~10 seconds |
| **Vector output** | Yes (native SVG) |
| **Aspect ratios** | 1:1, 16:9, 9:16, 4:3, 3:4 |
| **HuggingFace Arena** | #1 (ELO 1172, 72% win rate) |
| **Released** | February 2026 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Recraft v4**.
Describe your design intent specifically — style, layout, color palette, subject, and mood. Recraft v4 rewards design-language prompts.
Select raster (PNG) for standard output, or vector (SVG) for scalable, editable files.
Click **Generate**. At \~10 seconds, Recraft v4 is well-suited for rapid design iteration.
## Prompting tips
* **Use design vocabulary** — "flat illustration with muted earthy tones", "editorial minimalism", "bold typographic composition" produce stronger results than generic descriptions.
* **Text placement instructions work** — "Large bold sans-serif text 'OPEN NOW' centered, white on navy background, menu design" produces clean, accurate results.
* **For SVG vector output** — Simpler geometric compositions with clear subject separation produce the cleanest editable vector files. Complex photorealistic scenes work better as rasters.
### Example prompts
> A minimalist coffee shop menu board with "Morning Specials" as the header, items listed below in clean sans-serif, dark charcoal background, chalk-style illustration accents.
> A tech startup logo concept: abstract interlocking geometric shapes, electric blue and white, clean vector style.
> A flat-style world map illustration with major cities marked, soft muted palette, clean editorial design.
## Compare models
| Model | Resolution | Vector output | Best for |
| ------------------------------------------------- | ------------ | ------------- | ------------------------------- |
| **Recraft v4** | 1024×1024 | Yes (SVG) | Design, branding, typography |
| [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) | 2048×2048 | Yes (SVG) | Print-ready design work |
| [Ideogram v3](/ai-models/image/ideogram-v3) | Up to 1536px | No | Stylized typography, 58 presets |
Recraft v4 holds the #1 position on the HuggingFace Text-to-Image Arena — outranking Midjourney, DALL-E 3, and FLUX in design-oriented head-to-head evaluations. For print-resolution output, see [Recraft v4 Pro](/ai-models/image/recraft-v4-pro).
# Recraft v4 1
Source: https://docs.imagine.art/ai-models/image/recraft-v4-1
## Recraft v4.1
Recraft's newest generation model, live in the Image tool's model picker alongside [Recraft v4](/ai-models/image/recraft-v4) and [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) — this is an addition to the lineup, not a replacement for either.
This page covers what's confirmed live in the product so far: the model exists, is badged "new," and is described in-app as "Recraft's latest gen model." Detailed specs (resolution, generation time, what's changed versus v4) haven't been independently verified yet — treat this as a placeholder pending a fuller pass, and don't assume it simply supersedes v4/v4 Pro's documented specs.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Recraft v4.1**.
Write your prompt and generate as normal.
## Compare models
| Model | Vector output | Notes |
| ------------------------------------------------- | ------------- | ------------------------------------------------- |
| **Recraft v4.1** | Unconfirmed | Newest in the lineup — specs pending verification |
| [Recraft v4](/ai-models/image/recraft-v4) | Yes (SVG) | Design, branding, typography |
| [Recraft v4 Pro](/ai-models/image/recraft-v4-pro) | Yes (SVG) | Print-ready design work |
# Recraft v4 pro
Source: https://docs.imagine.art/ai-models/image/recraft-v4-pro
IMAGE MODEL
by Recraft AI
4-megapixel output
Recraft v4 Pro
All the design-taste intelligence of Recraft v4 — the #1 ranked image model on HuggingFace Arena — at 2048×2048 resolution (4 megapixels). The same SVG vector output, precise text rendering, and compositional control, now at print-ready quality for large-scale deliverables.
Resolution
2048×2048 (4MP)
Generation time
\~28–30 seconds
Vector output
Yes (native SVG)
## Recraft v4 at print scale
Recraft v4 Pro is not a separate model — it's the same V4 architecture running at 2048×2048 (4-megapixel) output. Everything that makes Recraft v4 the #1 model on HuggingFace Arena carries over: design-taste composition, reliable text rendering, native SVG generation, and precise prompt adherence. The difference is resolution and generation time (\~28–30 seconds vs. \~10 seconds for v4 standard).
Use Recraft v4 (standard) to iterate quickly. Use Recraft v4 Pro when you're ready for the final deliverable.
## Capabilities
Same compositional intelligence, design-taste training, and text rendering accuracy as Recraft v4 — just at 4× the resolution.
Sufficient resolution and detail for large-format print: posters, banners, packaging, and editorial spreads.
At 2048px, text elements hold sharpness and readability at scale.
## Specifications
| Feature | Details |
| -------------------- | ------------------------- |
| **Resolution** | 2048×2048 (4MP) |
| **Generation time** | \~28–30 seconds |
| **Vector output** | Yes (native SVG) |
| **Aspect ratios** | 1:1, 16:9, 9:16, 4:3, 3:4 |
| **Underlying model** | Recraft v4 |
| **Released** | February 2026 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Recraft v4 Pro**.
Use [Recraft v4](/ai-models/image/recraft-v4) (standard) to iterate on composition and style at 1K. At \~10 seconds per image, it's 3× faster for exploration.
Once you have the direction locked, switch to Recraft v4 Pro for 4MP output.
Select PNG for raster or SVG for editable vector output — both are available at full 4MP quality.
## When to use v4 vs v4 Pro
| Scenario | Recommendation |
| --------------------------------- | ---------------------------------- |
| Exploring concepts and directions | **Recraft v4** (faster, cheaper) |
| Final deliverable for digital use | **Recraft v4** (1MP is sufficient) |
| Print-ready posters and banners | **Recraft v4 Pro** |
| Editable vector for large-format | **Recraft v4 Pro** |
| Product packaging at scale | **Recraft v4 Pro** |
A good workflow: iterate quickly on [Recraft v4](/ai-models/image/recraft-v4) until you have the exact composition, typography, and style you want — then run the final prompt on Recraft v4 Pro for the production-ready 4MP output.
# Seedream 4
Source: https://docs.imagine.art/ai-models/image/seedream-4
IMAGE MODEL
by ByteDance
Seedream 4.0
ByteDance's most capable image model — native 4K output, up to 4 reference images, multilingual typography, and near real-time speed. Built for high-end campaigns, commercial design, and print-ready assets.
Seedream 4.0 is developed by ByteDance, the parent company of TikTok, as part of their SEED AI research program.
## Capabilities
Complex edits — including detailed product or character transformations — completed in one generation, reducing iteration time.
Maintains character identity, product appearance, and lighting across multiple variations and batch generations for large-scale projects.
Renders text in multiple languages with improved layout placement, ideal for branding across international markets.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Seedream 4.0**.
Write your text prompt and upload up to **4 reference images** for detailed guidance.
Click **Generate** and wait for the model to process your image.
Go to the **ImagineArt Image Editor**.
Click **Settings** and choose **Seedream 4.0** from the model dropdown.
Upload your base image and apply edits or transformations.
Use **visual prompts** to add, replace, or remove elements and make precise adjustments to the composition.
## Prompting tips
* **Be specific** — Define subject, environment, and style in detail. Example: *"A vintage car parked on a sunny street with a dog, shot in a retro photography style."*
* **Use multiple references** — Upload images of your subject, background, and product together for consistent outputs.
* **Guide mood and style** — Add descriptors like *"cinematic lighting"*, *"flat lay"*, or *"ink illustration"* to shape composition.
* **Use 4K mode for final work** — Select **4K mode** for print-ready visuals or large-scale campaigns. Use 2K for drafts.
### Example prompts
> A red "Glaze Cherry" cosmetic bottle placed on top of two shiny fresh cherries, with a pink and white gradient background. Bold red text saying "New" is displayed on the bottle. The scene should reflect beauty product photography, glossy surfaces, and ultra-realistic lighting.
> A close-up portrait of a teenage boy with curly dark hair and freckles, wearing a navy tracksuit jacket, standing in front of a yellow wall. Looking intensely into the camera with a serious expression, in a candid street photography style with natural emotion.
> A Black person leans on a city overpass railing, gazing seriously at the camera. They wear a sharp white shirt with a tie and a skirt. The urban setting below features a road with passing cars, contrasting the subject's poised and editorial stance.
## Seedream 4.0 vs. Nano Banana
| Feature | Seedream 4.0 | Nano Banana (Gemini 2.5 Flash Image) |
| -------------------------- | ------------------------------------------------- | ------------------------------------------- |
| **Developer** | ByteDance | Google DeepMind |
| **Multi-image references** | Up to 6 | Up to 4 |
| **Resolution** | 2K and 4K | Up to 2K |
| **Speed** | Near real-time | Near real-time |
| **One-shot accuracy** | Yes | Yes |
| **Typography** | Strong, multilingual | Strong (English and Chinese) |
| **Best for** | Complex campaigns, branding, multilingual markets | Quick edits, fashion try-ons, product swaps |
Generating 4K images with Seedream 4.0 uses more credits than standard resolution. Use 2K mode for draft iterations before committing to 4K for final assets.
# Seedream v4 5
Source: https://docs.imagine.art/ai-models/image/seedream-v4-5
IMAGE MODEL
by ByteDance
Seedream family
Seedream v4.5
ByteDance's production-standard image model — a major update over Seedream v4 with 30–40% faster generation, significantly improved text rendering for small and dense type, stronger subject consistency, and multi-image editing support with up to 4 references. Native resolution up to 4K, with support up to 8192×8192 for large-format work.
Max Resolution
Up to 8192×8192
Speed vs. v4
30–40% faster
## The production standard in the Seedream family
Seedream v4.5 is a significant upgrade over v4 — not a minor patch. Generation is 30–40% faster (8–14 seconds depending on resolution and complexity), text rendering is substantially improved for small and dense type (logos, menus, data tables), and multi-image editing supports up to 4 reference images with accurate subject identification and reference detail preservation.
Where [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) is optimized for speed and exploration, v4.5 is the right choice for the final 20–30% of a production workflow: the outputs that will actually ship to clients.
## Capabilities
Native 4K output with support up to 8192×8192 for large-format print, billboard, and ultra-high-resolution deliverables.
Meaningful speed improvement at equivalent quality — more iterations per session without sacrificing the detail that v4 established.
Firmer grasp on subjects across generation variations — facial features, lighting, and color tone are preserved more reliably across a session.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------------------------------- |
| **Architecture** | Diffusion Transformer (DiT) with high-compression VAE |
| **Resolution** | Up to 8192×8192 (native 4K standard) |
| **Reference images** | Up to 4 |
| **Generation speed** | 8–14 seconds |
| **Speed vs. v4** | 30–40% faster |
| **Released** | December 3, 2025 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Seedream v4.5**.
For text-in-image, include the exact text you want rendered. For multi-reference composites, describe subjects clearly.
Upload up to 4 reference images for consistent multi-product or multi-character outputs.
Click **Generate**. At 8–14 seconds, Seedream v4.5 is fast enough for efficient iteration at production quality.
## Prompting tips
* **For dense text layouts** — Describe the full text hierarchy: headline, subheads, body copy, and any labels or callouts. v4.5 renders dense layouts accurately at production size.
* **Use reference images for product consistency** — Attach your product photography references and describe the scenario. The model preserves product details (colors, textures, shapes) precisely.
* **For large-format output** — Select custom resolution and push above 4K for billboard or large-format print. The high-compression VAE maintains quality at extreme sizes.
### Example prompts
> A premium restaurant menu spread with "TASTING MENU" as the header, five courses listed with descriptions in clean serif type, cream background, elegant layout, no imagery.
> Three luxury fragrance bottles (from references) arranged on a marble surface with fresh botanicals, editorial product photography, soft natural light.
## Compare models
| Model | Speed | Resolution | Text rendering | References | Best for |
| ----------------------------------------------------- | -------- | ---------- | ----------------- | ---------- | ---------------------------- |
| **Seedream v4.5** | 8–14s | Up to 8K | Excellent (dense) | 14 | Production deliverables |
| [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) | 3–5s | Up to 4K | Good (titles) | — | Rapid iteration, exploration |
| [Seedream v4](/ai-models/image/seedream-4) | Moderate | 4K | Good | Up to 6 | Commercial imagery |
Use [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) for the ideation phase — it's 2–4× faster at lower cost. Switch to Seedream v4.5 for the final production output, especially when you need dense text accuracy or 14-reference compositing.
# Seedream v5 lite
Source: https://docs.imagine.art/ai-models/image/seedream-v5-lite
IMAGE MODEL
by ByteDance
Deep thinking
Seedream v5 Lite
ByteDance's fastest Seedream model — generates in 3–5 seconds with deep thinking reasoning and real-time web search grounding. A compressed version of Seedream 5.0 that infers intent intelligently from short or vague prompts, making it the fastest way to explore creative directions before committing to production quality in v4.5.
Generation speed
3–5 seconds
Reasoning
Chain-of-thought
## Fast generation, intelligent interpretation
Seedream v5 Lite introduces something new to the Seedream family: **deep thinking** — a multi-step chain-of-thought reasoning capability that infers creative intent from short, abstract, or ambiguous prompts like a human designer would. Combined with optional **real-time web search**, the model can ground generations in current information, trending subjects, and real-world accuracy.
At 3–5 seconds per image, it's the fastest way to explore the full creative space before committing to Seedream v4.5 for the final production output.
## Capabilities
Toggle-able internet access grounds generations in current information, trending products, and real-world subjects without detailed descriptions.
Generates from 1440px to 4096px on any axis — supports standard and custom aspect ratios for web, social, and intermediate-quality print work.
Handles logic puzzles, assembly sequences, and information visualizations — the reasoning capability enables outputs that require step-by-step compositional thinking.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------- |
| **Base model** | Compressed Seedream 5.0 |
| **Resolution** | 1440px–4096px (any axis) |
| **Generation speed** | 3–5 seconds |
| **Reasoning** | Chain-of-thought (multi-step) |
| **Web search** | Yes (toggle-able) |
| **Speed vs. v4.5** | 2–4× faster |
| **Released** | February 13, 2026 |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Seedream v5 Lite**.
Short, directional prompts work well — the deep thinking capability will expand them intelligently. Enable web search for prompts involving current subjects or real-world references.
At 3–5 seconds per image, generate multiple variations to explore your creative direction. The cost is low enough to generate at volume.
Once you've locked a composition and style, run the same prompt on [Seedream v4.5](/ai-models/image/seedream-v4-5) for the production-quality deliverable.
## Prompting tips
* **Short prompts are fine** — The reasoning capability handles vague inputs. "A campaign hero shot for a premium running shoe, energetic, outdoors" will be interpreted and expanded automatically.
* **Enable web search for branded work** — For prompts referencing specific products, events, or current subjects, web search grounding produces more accurate results.
* **Use it before v4.5** — The standard workflow: iterate on v5 Lite to nail composition and style, then generate the final on v4.5.
### Example prompts
> A startup pitch deck hero image: ambitious, forward-looking, modern, technology theme.
> Product launch visual for a new wireless earbud — minimal, white background, premium feel.
## Compare models
| Model | Speed | Text rendering | Best for |
| ----------------------------------------------- | -------- | ----------------- | ---------------------------- |
| **Seedream v5 Lite** | 3–5s | Good (titles) | Rapid iteration, exploration |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | 8–14s | Excellent (dense) | Production deliverables |
| [Seedream v4](/ai-models/image/seedream-4) | Moderate | Good | Standard commercial |
Seedream v5 Lite is designed to be your first step, not your last. Use it to explore. Use [Seedream v4.5](/ai-models/image/seedream-v4-5) to deliver.
# Seedream v5 pro
Source: https://docs.imagine.art/ai-models/image/seedream-v5-pro
## Seedream V5 pro
ByteDance's newest Seedream model, live in the Image tool's model picker alongside [Seedream v5 Lite](/ai-models/image/seedream-v5-lite), [Seedream v4.5](/ai-models/image/seedream-v4-5), and [Seedream v4](/ai-models/image/seedream-4) — an addition to the lineup, not a replacement. In-app it's described as a "New fast, high-quality model from ByteDance."
This page covers what's confirmed live in the product so far: the model exists, is badged "new," with the tagline above. Detailed specs (resolution, reference-image support, speed relative to v5 Lite) haven't been independently verified — treat this as a placeholder pending a fuller pass.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Seedream V5 pro**.
Write your prompt and generate as normal.
## Compare models
| Model | Notes |
| ----------------------------------------------------- | ------------------------------------------------- |
| **Seedream V5 pro** | Newest in the lineup — specs pending verification |
| [Seedream v5 Lite](/ai-models/image/seedream-v5-lite) | 3–5 seconds, intelligent prompt interpretation |
| [Seedream v4.5](/ai-models/image/seedream-v4-5) | Up to 8192×8192, 14 references |
| [Seedream v4](/ai-models/image/seedream-4) | Near real-time, multilingual |
# Stable diffusion 3 5 medium
Source: https://docs.imagine.art/ai-models/image/stable-diffusion-3-5-medium
## Stable Diffusion 3.5 Medium
A new model from Stability AI, live in the Image tool's model picker — the first Stable Diffusion model in ImagineArt's lineup and a new provider for this product. In-app it's described as "Efficient, high-quality text-to-image generation from Stability AI" and badged "new."
This page covers what's confirmed live in the product so far: the model exists, is badged "new," with the tagline above. Detailed specs (resolution, speed, licensing notes relevant to this integration) haven't been independently verified — treat this as a placeholder pending a fuller pass.
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Stable Diffusion 3.5 Medium**.
Write your prompt and generate as normal.
# Z image turbo
Source: https://docs.imagine.art/ai-models/image/z-image-turbo
IMAGE MODEL
by Alibaba Tongyi Lab
Apache 2.0
#1 open-source
Z Image Turbo
Alibaba's ultra-fast, open-source image model — 8-step distilled generation, bilingual English and Chinese text rendering, and photorealistic quality at approximately 4× the speed of FLUX. Ranked #1 among open-source models on the Artificial Analysis Text-to-Image Leaderboard.
Resolution
Up to 2048×2048
## Built for speed without sacrificing quality
Z Image Turbo is built on **S3-DiT** (Scalable Single-Stream Diffusion Transformer) — a unified architecture where text, visual semantic, and image tokens are processed in a single stream rather than dual-stream models like FLUX. Combined with **Decoupled-DMD distillation**, generation is compressed to just 8 steps with no classifier-free guidance required, delivering results approximately 4× faster than FLUX.2 Dev at comparable or better quality.
## Capabilities
Photography-grade quality with accurate lighting, shadows, and fine material detail. Performs at or above models with 5× more parameters.
Accurately renders named landmarks, cultural references, and recognizable figures — drawing on Alibaba's Qwen3-4B text encoder for depth.
Built-in structured reasoning chains expand and refine prompts automatically for richer, more coherent outputs from short instructions.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------------------------------- |
| **Architecture** | S3-DiT (Scalable Single-Stream Diffusion Transformer) |
| **Text encoder** | Qwen3-4B |
| **Parameters** | 6.15 billion |
| **Inference steps** | 8 (distilled via Decoupled-DMD) |
| **CFG guidance** | Not required (scale: 0.0) |
| **Resolution** | 512×512 to 2048×2048 |
| **VRAM requirement** | 16 GB (fits RTX 3080 Ti, 4080, Mac M-series) |
| **License** | Apache 2.0 |
| **Released** | November 26, 2025 |
## Benchmarks
Z Image Turbo was evaluated against leading proprietary and open-source models:
| Benchmark | Z Image Turbo | Ranking |
| --------------------------------------------- | ---------------------- | ------------------------------- |
| Artificial Analysis Text-to-Image Leaderboard | Elo 1025, 45% win rate | **#1 open-source**, 4th overall |
| CVTG-2K text rendering (word accuracy) | 0.8585 | Top tier |
| LongText-Bench English | 0.917 | Top tier |
| LongText-Bench Chinese | 0.926 | Top tier |
| Speed vs. FLUX.2 Dev (100 imgs @ 1024×1024) | 279s vs. 1,152s | **\~4× faster** |
## How to use
Go to the **ImagineArt AI Image Generator**.
From the model dropdown, choose **Z Image Turbo**.
Write a clear, focused prompt. Z Image Turbo responds best to precise, concise descriptions — overly long prompts can add noise rather than detail.
Choose from 512×512 up to 2048×2048. The model performs consistently across the full resolution range.
Click **Generate**. At 8 steps, results arrive significantly faster than most other models.
## Prompting tips
* **Keep prompts concise and specific** — Z Image Turbo is optimized for structured, precise prompts. Dense, paragraph-length prompts can reduce coherence rather than improve it.
* **For bilingual text in images** — Include both the English and Chinese text you want rendered, with explicit placement: *"A product banner with bold red text reading 'Summer Sale' and '夏季特卖' below it."*
* **Avoid high CFG values** — The model was trained at guidance scale 0.0. Using high CFG in manual configurations introduces artifacts. Leave guidance at default.
* **Use prompt enhancement** — Enable the built-in prompt enhancer for short or abstract prompts. It applies Alibaba's structured reasoning to expand your intent into richer descriptions.
### Example prompts
> A Japanese ramen shop at night, warm amber light spilling from the windows onto rain-wet cobblestones, steam rising from bowls inside, photorealistic, cinematic composition.
> A product flatlay of a wireless speaker on brushed concrete, minimalist studio lighting, crisp shadow, commercial photography style.
> A bold event poster with "OPEN MIC NIGHT" in large neon-style lettering and "每周五 / Every Friday" beneath it, dark urban background.
## Compare models
| Model | Speed | Text rendering | Parameters | License | Best for |
| ------------------------------------------------- | --------------------- | ---------------------------- | ---------- | -------------- | ----------------------------------------- |
| **Z Image Turbo** | \~4× faster than FLUX | Bilingual (EN + ZH), low WER | 6.15B | Apache 2.0 | Rapid generation, bilingual, photorealism |
| [Flux Dev](/ai-models/image/flux-dev) | Moderate (\~7–18s) | Decent | 12B | Non-commercial | Fine-tuning base, creative research |
| [Qwen Image](/ai-models/image/qwen-image) | Fast | Excellent (EN + ZH) | 7B (2.0) | Apache 2.0 | Illustrations, bilingual, complex layouts |
| [Seedream v3](/ai-models/image/seedream-1) | Seconds | EN + ZH | 12B | Commercial | Fast branded imagery |
| [ImagineArt 1.0](/ai-models/image/imagineart-1-0) | Industry-leading | Good | — | Commercial | Photorealistic portraits |
Z Image Turbo is developed by Alibaba's Tongyi Lab under the Apache 2.0 license. It outperforms models with 5× more parameters — including FLUX.2 Dev (32B) — on several benchmarks, making it one of the most efficient high-quality image models available.
# Choosing the Right Video Model
Source: https://docs.imagine.art/ai-models/video-models
Find the best ImagineArt video model for your project — use case guidance, side-by-side comparisons, and direct links to all 44 model pages.
ImagineArt gives you access to 44 AI video generation models across every major provider. Use this guide to find the right one for your project — by use case, capability, or output format.
## Quick picks by use case
**Seedance 2** — ByteDance's flagship. Native audio, 4 generation modes, and the most comprehensive reference system available (9 img + 3 vid + 3 audio).
**Kling 3.0 Pro** — 1080p at 60 FPS, up to 15 seconds, Omni Native Audio with multilingual lip-sync, 6-shot storytelling.
**Sora 2 Pro** — Up to 25 seconds with integrated audio and physics-aware rendering. The longest single generation available.
**xAI Grok Video** — \~17-second generation with native audio and up to 7 reference images. The fastest AI video model available.
## By what you're building
**Priority: audio quality + lip-sync precision + language support**
For multilingual dialogue with millisecond-precision lip-sync across 8+ languages, [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) is the specialist. [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) delivers Omni Native Audio with EN, JP, KO, and ES lip-sync at 1080p. For English and Chinese including singing, [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) generates at 48 FPS with simultaneous A/V in a single pass.
For broadcast-quality audio at the highest visual fidelity, [Google Veo 3.1](/ai-models/video/google-veo-3-1) generates 48kHz stereo audio alongside 4–8 second clips at up to 4K. [Sora 2 Pro](/ai-models/video/sora-2-pro) pairs physics-aware motion with integrated dialogue and effects up to 25 seconds. For the fastest audio-enabled generation, [xAI Grok Video](/ai-models/video/grok-video) delivers in \~17 seconds.
| Model | Audio type | Lip-sync | Duration |
| ----------------------------------------------------- | -------------------------- | ------------ | --------- |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | Dialogue, SFX | 8+ languages | 4–12s |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Omni Native — EN/JP/KO/ES | Yes | Up to 15s |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | Dialogue, SFX, singing | EN + Chinese | Up to 10s |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | 48kHz stereo dialogue, SFX | — | 4–8s |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | Dialogue, SFX, ambient | — | Up to 25s |
| [xAI Grok Video](/ai-models/video/grok-video) | Music, SFX, ambient | — | 6 or 10s |
**Priority: resolution + frame rate + visual fidelity**
Four models offer 4K output. [Kling 3.0 4K](/ai-models/video/kling-3-0-4k) is Kling AI's 4K tier of the 3.0 family — native audio, first-and-last-frame control, and clips up to 15 seconds. [Google Veo 3.1](/ai-models/video/google-veo-3-1) reaches 4K with 4–8 second clips and 48kHz stereo audio. [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) provides 4K with selectable 4, 6, or 8-second clips at lower cost.
| Model | Resolution | Audio | Duration |
| ----------------------------------------------------------- | ---------- | ----- | -------- |
| [Kling 3.0 4K](/ai-models/video/kling-3-0-4k) | 4K | Yes | 3–15s |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | Up to 4K | Yes | 4–8s |
| [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | Up to 4K | No | 4–8s |
**Priority: duration + narrative continuity + rendering stability**
[Sora 2 Pro](/ai-models/video/sora-2-pro) supports up to 25 seconds — the longest single generation available, with integrated audio and physics-aware rendering. [Google Veo 3.1](/ai-models/video/google-veo-3-1) generates 4–8 second clips with 4K output and broadcast-quality 48kHz stereo audio.
At 15 seconds: [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro), [Kling O3](/ai-models/video/kling-o3), [Seedance 2](/ai-models/video/seedance-2), [Seedance 2 Fast](/ai-models/video/seedance-2-fast), [Wan 2.6](/ai-models/video/wan-2-6), and [PixVerse v6](/ai-models/video/pixverse-v6).
| Model | Max duration | Audio | Resolution |
| ------------------------------------------------- | ------------ | ----- | ---------- |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | 4–8s | Yes | Up to 4K |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | 25s | Yes | 1080p |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 15s | Yes | 1080p |
| [Seedance 2](/ai-models/video/seedance-2) | 15s | Yes | 720p |
| [Wan 2.6](/ai-models/video/wan-2-6) | 15s | Yes | 1080p |
**Priority: turnaround time + iteration speed**
[xAI Grok Video](/ai-models/video/grok-video) generates a 6-second video in approximately 17 seconds — the fastest in the lineup by a significant margin, powered by Aurora's autoregressive sequential frame prediction. [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) is 5× faster than standard Runway Gen 4, generating 10-second clips in \~30 seconds. [PixVerse v5](/ai-models/video/pixverse-v5) and [PixVerse v5.5](/ai-models/video/pixverse-v5-5) also generate in approximately 30 seconds at 1080p.
| Model | Generation time | Duration | Audio |
| --------------------------------------------------------- | --------------- | --------- | ----- |
| [xAI Grok Video](/ai-models/video/grok-video) | \~17s | 6 or 10s | Yes |
| [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) | \~30s | 10s | No |
| [PixVerse v5](/ai-models/video/pixverse-v5) | \~30s | Up to 15s | No |
| [PixVerse v5.5](/ai-models/video/pixverse-v5-5) | \~30s | Up to 10s | Yes |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | Under 60s | 5 or 10s | No |
**Priority: object interaction + material behavior + environmental dynamics**
[Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) delivers industry-leading physics simulation — the strongest model for realistic fluid dynamics, collision physics, and material deformation. [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) offers the same NCR architecture at lower cost. [Kling O3](/ai-models/video/kling-o3) includes a purpose-built physics engine covering gravity, collision, inertia, deformation, and fluid dynamics alongside native 4K and audio.
| Model | Physics tier | Resolution | Cost tier |
| ----------------------------------------------- | ---------------- | ---------- | --------- |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | Industry-leading | 1080p | Pro |
| [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) | Strong | 1080p | Standard |
| [Kling O3](/ai-models/video/kling-o3) | Advanced engine | 4K | Pro |
**Priority: art style fidelity + stylization coherence + micro-expression quality**
[Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) delivers the highest quality for anime, illustration, ink-wash, and game-CG styles — with enhanced micro-expression rendering and physics-aware stylized scenes. [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) offers the same style range at lower cost. [PixVerse v5](/ai-models/video/pixverse-v5) is also strong for anime and game character consistency, particularly for complex movement.
| Model | Style quality | Physics | Cost tier |
| ------------------------------------------------- | ------------- | -------- | --------- |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | Maximum | High | Pro |
| [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) | High | Good | Standard |
| [PixVerse v5](/ai-models/video/pixverse-v5) | Good | Standard | Standard |
**Priority: identity preservation + cross-scene consistency + reference fidelity**
[Wan 2.6](/ai-models/video/wan-2-6) is built specifically for this — its R2V (Reference-to-Video) mode inserts a character's appearance and voice from a reference image across any generated scene with consistent identity preservation. [Kling O3](/ai-models/video/kling-o3) accepts 10+ references across 6 generation modes. [Seedance 2](/ai-models/video/seedance-2) accepts 9 images + 3 video + 3 audio clips simultaneously. [xAI Grok Video](/ai-models/video/grok-video) accepts up to 7 reference images for identity preservation at speed.
| Model | Reference capacity | Voice input | Best for |
| --------------------------------------------- | ----------------------- | ----------- | --------------------------------------- |
| [Wan 2.6](/ai-models/video/wan-2-6) | Appearance + voice | Yes | Character identity + voice in any scene |
| [Kling O3](/ai-models/video/kling-o3) | 10+ images | No | Multi-reference 4K |
| [Seedance 2](/ai-models/video/seedance-2) | 9 img + 3 vid + 3 audio | Yes (audio) | Full multimodal references |
| [xAI Grok Video](/ai-models/video/grok-video) | Up to 7 images | No | Fast identity preservation |
**Priority: explicit trajectory control + camera movement accuracy**
[Wan 2.2](/ai-models/video/wan-2-2) offers the most explicit camera control in the lineup — the VACE (Video Animation Control Engine) provides programmatic camera trajectory input with subject locking, background stabilization, and precise pans, zooms, and focus pulls. LoRA-based style adaptation (10–20 images) also sets it apart. [Runway 4.5](/ai-models/video/runway-4-5) and [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) are strong for cinematic camera-precise output from natural language prompts.
| Model | Camera control type | Style adaptation | Best for |
| --------------------------------------------------------- | ------------------------------ | ----------------- | --------------------------------- |
| [Wan 2.2](/ai-models/video/wan-2-2) | VACE trajectory (programmatic) | LoRA (10–20 imgs) | Exact camera paths, custom styles |
| [Runway 4.5](/ai-models/video/runway-4-5) | Prompt-based cinematic | — | Cinematic camera, final renders |
| [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) | Prompt-based cinematic | — | Fast cinematic iteration |
**Priority: transition precision + motion interpolation between defined states**
[Pika 2.2](/ai-models/video/pika-2-2) is purpose-built for this — Pikaframes lets you define the exact opening and closing frame of any clip, with Pika generating the motion between them. [Kling 2.1 Pro](/ai-models/video/kling-2-1-pro) and [Kling O1](/ai-models/video/kling-o1) support first-and-last-frame conditioning with advanced motion interpolation. [Seedance 2](/ai-models/video/seedance-2) includes First and Last Frame as one of its four generation modes alongside full audio.
| Model | Keyframe mode | Audio | Best for |
| ----------------------------------------------- | ------------------------- | ----- | -------------------------------- |
| [Pika 2.2](/ai-models/video/pika-2-2) | Pikaframes (dedicated) | No | Precise start-to-end transitions |
| [Kling 2.1 Pro](/ai-models/video/kling-2-1-pro) | First + last frame | No | Image animation, HD |
| [Kling O1](/ai-models/video/kling-o1) | First + last frame | No | Unified creation + editing |
| [Seedance 2](/ai-models/video/seedance-2) | First and Last Frame mode | Yes | Transitions with audio |
**Priority: low cost per generation + speed + acceptable quality floor**
[Seedance 2.0 Mini](/ai-models/video/seedance-2-0-mini) is ByteDance's lightweight, low-cost model — fast inference at 480p–720p, built for multi-shot video. [Google Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) delivers Veo 3.1 architecture at less than 50% of the Fast tier cost. [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) and [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) both have fast variants that reduce cost by 50%. [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) is the speed-optimized tier of Seedance 1.0 Pro.
| Model | Cost tier | Speed | Quality floor |
| ----------------------------------------------------------- | ------------- | ---------------------- | ------------------------ |
| [Seedance 2.0 Mini](/ai-models/video/seedance-2-0-mini) | Lowest | Fast | Good (480p–720p) |
| [Google Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | \<50% of Fast | Fast | Veo 3.1 quality at 1080p |
| [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) | Standard | Fast variant available | Stylized, 768p/1080p |
| [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) | Standard | Fast variant available | Physics-capable, 1080p |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | Standard | 30–60% faster than Pro | Cinematic, 480p–1080p |
## Full model comparison
| Model | Provider | Resolution | Duration | Audio | Best for |
| --------------------------------------------------------------- | ----------------- | ----------- | -------- | ----- | ------------------------------------------------- |
| [Happy Horse](/ai-models/video/happy-horse) | Alibaba | 720p–1080p | 3–15s | Yes | Fluid lifelike motion with native audio |
| [Kling 3.0 4K](/ai-models/video/kling-3-0-4k) | Kling AI | 4K | 3–15s | Yes | Maximum resolution, 4K delivery |
| [Lucy](/ai-models/video/lucy) | Decart | 720p | — | No | Fast image animation, social content |
| [Seedance 2 Fast](/ai-models/video/seedance-2-fast) | ByteDance | 720p | 4–15s | Yes | Fast production pipelines, iteration |
| [Seedance 2](/ai-models/video/seedance-2) | ByteDance | 720p-1080p | 4–15s | Yes | Max quality multimodal, full references |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Kling AI | 1080p | 3-15s | Yes | Multi-shot storytelling, 60 FPS |
| [Runway 4.5](/ai-models/video/runway-4-5) | Runway | 720p | 5–10s | No | Cinematic camera control, final renders |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | ByteDance | 480p-720p | 4–12s | Yes | Multilingual dialogue, 8+ languages |
| [Pika 2.2](/ai-models/video/pika-2-2) | Pika Labs | 720p-1080p | 5-10s | Yes | Keyframe control (Pikaframes) |
| [Kling O3](/ai-models/video/kling-o3) | Kling AI | 1080p | 5-10s | No | Advanced physics, 6 modes, 4K |
| [PixVerse v6](/ai-models/video/pixverse-v6) | PixVerse | 540p-1080p | 5-10s | No | Cinematic lens control, 20+ optical params |
| [Google Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | Google | 720p-1080p | 8s | No | Cost-efficient, high-volume generation |
| [Luma Ray 2](/ai-models/video/luma-ray-2) | Luma AI | 540p-720p | 5–9s | No | Photorealistic motion, natural movement |
| [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) | MiniMax | 768P | 6-10s | No | Physics realism, cost-efficient |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | MiniMax | 1080p | 6s | No | Industry-leading physics simulation |
| [Kling 2.1 Pro](/ai-models/video/kling-2-1-pro) | Kling AI | 1080p | 5–10s | No | First + last frame image animation |
| [Seedance 2.0 Mini](/ai-models/video/seedance-2-0-mini) | ByteDance | 480p–720p | 5-15s | Yes | Fast, lightweight, multi-shot video |
| [Seedance 2.5](/ai-models/video/seedance-2-5) | ByteDance | 480p-1080p | 4-30s | Yes | Longest Seedance duration ceiling |
| [Seedance 1.0 Pro](/ai-models/video/seedance-1-0-pro) | ByteDance | 480p-1080p | 3-12s | No | Cinematic storytelling, camera work |
| [Wan 2.2](/ai-models/video/wan-2-2) | Alibaba | 720p | 5s | No | VACE camera control, LoRA style adapt |
| [PixVerse v5](/ai-models/video/pixverse-v5) | PixVerse | 540p-720p | 5-8s | No | Complex movement, anime, game characters |
| [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) | Runway | 720p | 5-10s | No | Rapid iteration, 5× faster than Gen 4 |
| [Kling 2.5 Pro](/ai-models/video/kling-2-5-pro) | Kling AI | 1080p | 5-10s | No | Cost-efficient HD, sports + physics |
| [Wan 2.5](/ai-models/video/wan-2-5) | Alibaba | 480p–1080p | 5–10s | Yes | Audio-visual sync, lip-sync |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | OpenAI | 720p-1080p | 4-20s | Yes | Final production, physics-aware |
| [Kling O1](/ai-models/video/kling-o1) | Kling AI | 1080p | 5–10s | No | Unified create + edit workflows |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | Kling AI | 1080p | 5-10s | Yes | Audio-synced, EN/Chinese, 48 FPS |
| [PixVerse v5.5](/ai-models/video/pixverse-v5-5) | PixVerse | 540p-1080p | 5-8s | Yes | Script-first narrated multi-shot |
| [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | Google | 720p-4k | 4/6/8s | No | Balanced speed + quality + audio |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | Google | 720p-4k | 4/6/8s | Yes | Broadcast, commercial, 4K |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | ByteDance | 480p-1080p | 3-12s | No | Speed-optimized Seedance Pro |
| [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) | MiniMax | 768p | 6-10s | No | Stylized — anime, illustration, game-CG |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | MiniMax | 1080p | 6s | No | Stylized + physics, pro quality |
| [Wan 2.6](/ai-models/video/wan-2-6) | Alibaba | 720-1080p | 5-15s | Yes | Character reference-to-video, R2V |
| [xAI Grok Video](/ai-models/video/grok-video) | xAI | 480-720p | 6-15s | Yes | Fastest generation (\~17 seconds) |
| [xAI Grok 1.5](/ai-models/video/xai-grok-1-5) | xAI | 480-720p | 5-15s | Yes | Newest xAI video model |
| [MiniMax Hailuo H3](/ai-models/video/minimax-hailuo-h3) | MiniMax | 2K | 5-15s | Yes | Frontier 2K, subject reference + camera direction |
| [MiniMax Hailuo H3 Max](/ai-models/video/minimax-hailuo-h3-max) | MiniMax | 480p-768p | 5-15s | Yes | Larger H3 variant |
| [Wan 3](/ai-models/video/wan-3) | Alibaba | 480p-1080p | 5-30s | Yes | Newest Wan model, longest duration |
| [Kling 3.0 Turbo Pro](/ai-models/video/kling-3-0-turbo-pro) | Kling AI | 1080p | 5-15s | Yes | Turbo tier of Kling 3.0 |
| [Kling 3.0 Turbo SD](/ai-models/video/kling-3-0-turbo-sd) | Kling AI | 720p | 5-15s | Yes | Lower-cost Turbo tier |
| [Google Omni Flash](/ai-models/video/google-omni-flash) | Google | 720p | 3-10s | Yes | Faster, lighter than Veo 3.1 |
| [Flux 3](/ai-models/video/flux-3) | Black Forest Labs | 720p-1080p | 5-20s | Yes | New provider for video |
| [LTX 2.3](/ai-models/video/ltx-2-3) | Lightricks | 1080p-2160p | 6-10s | Yes | New provider, highest resolution ceiling (4K) |
# Flux 3
Source: https://docs.imagine.art/ai-models/video/flux-3
## Flux 3
A new video model from Black Forest Labs — a new provider in ImagineArt's video lineup (Black Forest Labs previously only appeared under image models, as Flux).
## Specifications
| Feature | Details |
| ----------------- | ----------------- |
| **Developer** | Black Forest Labs |
| **Resolution** | 720p–1080p |
| **Duration** | 5–20 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Flux 3**.
Write your prompt (and optionally set start/end frames) and generate as normal.
# Google omni flash
Source: https://docs.imagine.art/ai-models/video/google-omni-flash
## Google Omni Flash
A new, faster Google video model, live alongside the [Veo 3.1](/ai-models/video/google-veo-3-1) family. Positioned as a lighter, quicker option than the Veo 3.1 tiers based on its shorter default duration range.
## Specifications
| Feature | Details |
| ----------------- | ------------ |
| **Developer** | Google |
| **Resolution** | 720p |
| **Duration** | 3–10 seconds |
| **Audio** | Yes |
| **Frame control** | Start frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Google Omni Flash**.
Write your prompt and generate as normal.
## Google video model comparison
| Model | Resolution | Duration | Audio |
| ----------------------------------------------------------- | ---------- | -------- | ----- |
| **Google Omni Flash** | 720p | 3–10s | Yes |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | Up to 4K | 4–8s | Yes |
| [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | Up to 4K | 4/6/8s | No |
| [Google Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | 720p–1080p | 8s | No |
# Google veo 3 1
Source: https://docs.imagine.art/ai-models/video/google-veo-3-1
VIDEO MODEL
by Google DeepMind
Released October 2025
Google Veo 3.1
Google DeepMind's flagship video generation model — 4–8 seconds of output, native 48kHz stereo audio including dialogue and voiceover, 4K resolution with selectable frame rates (24, 30, or 60 FPS), and the highest visual fidelity in the Veo family for broadcast-quality production.
Frame rates
24, 30, or 60 FPS
## Google's most capable video model
Google Veo 3.1, released October 2025, is the production flagship of the Veo 3.1 family — generating 4–8 seconds of content per generation. Combined with broadcast-quality audio at 48kHz stereo (including dialogue, voiceover, and sound effects) and 4K resolution, Veo 3.1 is positioned for broadcast, commercial, and high-end production.
Selectable frame rates (24 FPS for cinematic, 30 FPS for standard, 60 FPS for sports and fast motion) give the model flexibility across production contexts that no other model in the lineup matches.
## Capabilities
Choose from 4, 6, or 8 second clips — enough for complete commercial spots, narrative sequences, and high-quality short-form content.
Broadcast-quality audio at 48kHz stereo — includes dialogue, voiceover, sound effects, and ambient soundscapes generated natively.
4K output with the Transformer backbone's spatial fidelity — production-ready for broadcast, cinema, and large-format display.
Choose 24 FPS (cinematic standard), 30 FPS (broadcast/digital), or 60 FPS (sports, action, smooth motion) per generation.
Maximum quality configuration of the Veo 3.1 Transformer backbone — richer texture, more stable scene composition, and more precise prompt adherence than Fast and Lite tiers.
Responds accurately to cinematographic language — camera movements, lighting conditions, scene pacing, and narrative cues all translate to the output.
## Specifications
| Feature | Details |
| ----------------- | ------------------------------------------------ |
| **Developer** | Google DeepMind |
| **Released** | October 2025 |
| **Resolution** | 720p, 1080p, 4K |
| **Duration** | 4–8 seconds (4, 6, or 8s selectable) |
| **Frame rates** | 24 FPS, 30 FPS, 60 FPS (selectable) |
| **Audio** | 48kHz stereo — dialogue, voiceover, SFX, ambient |
| **Aspect ratios** | 16:9, 9:16 |
| **Architecture** | Transformer backbone, spatio-temporal patches |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Google Veo 3.1** from the model dropdown.
For long-form content, structure your prompt as a scene description with clear narrative progression — describe the arc of what happens over the full duration.
Choose your desired length — 4, 6, or 8 seconds.
Choose 24 FPS for cinematic, 30 FPS for standard delivery, or 60 FPS for smooth high-speed action.
Choose 720p, 1080p, or 4K based on delivery requirements.
Click **Generate**. 4K generations take longer than 720p or 1080p.
## Prompting tips
* **Structure prompts as narratives** — "The scene opens on... then transitions to... building to a conclusion where..." — Veo 3.1 understands temporal narrative structure.
* **Specify audio timing** — "The music builds from quiet to full orchestral by the halfway point" — temporal audio cues are interpreted accurately.
* **Match FPS to content** — 24 FPS for drama and cinema; 30 FPS for documentary and commercial; 60 FPS for sports or fluid slow-motion effect.
* **Use 4K for large-format or print-quality moments** — Product launches, hero brand shots, and cinematic sequences benefit most from 4K.
### Example prompts
> A nature documentary segment: A wolf pack hunts across a frozen tundra at dusk. The alpha leads, the pack follows in formation. Wide shot establishing the landscape, intercutting with close-ups of paw prints, breath visible in the cold air. Narration: "In the far north, survival depends on the pack." Ambient wind and distant howling. 8 seconds, 4K.
> A coffee brand spot: Product sits on a warm kitchen table at sunrise. A pair of hands wrap around the mug. Steam rises in close-up. Brand voice: "Start every morning right." Warm, cinematic. 6 seconds.
## Compare models
| Model | Duration | Audio | Resolution | FPS options | Best for |
| ---------------------------------------------------- | -------- | -------------- | ---------- | ----------- | ------------------------------- |
| **Veo 3.1** | 4–8s | 48kHz stereo | Up to 4K | 24/30/60 | Broadcast, commercial, high-end |
| [Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | 4/6/8s | SFX + ambient | Up to 4K | 24 | Short-form production, 4K |
| [Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | 8s | No | 1080p | — | Cost-efficient, no audio |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | 25s | Dialogue + SFX | 1080p | — | Long-form A/V, physics |
Choose Veo 3.1 when you need the highest visual fidelity in the Veo family with full audio (dialogue, voiceover, SFX) and up to 4K. For faster generation at lower cost, [Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) offers the same 4–8s and 4K with a reduced credit cost.
# Google veo 3 1 fast
Source: https://docs.imagine.art/ai-models/video/google-veo-3-1-fast
VIDEO MODEL
by Google DeepMind
Veo 3.1 family
Google Veo 3.1 Fast
Google DeepMind's balanced Veo 3.1 variant — up to 4K resolution, 4, 6, or 8-second clips with 3 reference image support, and faster generation times than the flagship Veo 3.1 for production workflows.
Duration
4, 6, or 8 seconds
References
Up to 3 images
## Balanced speed and quality
Veo 3.1 Fast sits in the middle of the Veo 3.1 family — faster than the flagship Veo 3.1 with more capability than Veo 3.1 Lite. Multi-reference image input (up to 3 images) and 4K resolution support are included.
Generation times are 60–90 seconds for 720p and 90–120 seconds for 1080p, making it practical for production workflows where quality and speed need to be balanced. The Transformer backbone with spatio-temporal patches is shared across the Veo 3.1 family.
## Capabilities
Supports 720p, 1080p, and 4K output — choose the resolution tier that fits your delivery requirements.
Multi-reference input with up to 3 images for subject appearance, visual style, and scene composition anchoring.
Selectable clip length — choose 4, 6, or 8 seconds depending on content needs and credit budget.
Supports image-to-video with natural, physically plausible motion from a reference starting frame.
Shorter generation times than Veo 3.1 — 60–120 seconds at 720p–1080p for production-pace workflows.
## Veo 3.1 family comparison
| Model | Audio | Duration | Max res | Speed | Cost |
| ---------------------------------------------------- | ----- | -------- | ------- | -------- | ------- |
| [Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | No | 8s | 1080p | Fast | Lowest |
| **Veo 3.1 Fast** | Yes | 4/6/8s | 4K | Balanced | Medium |
| [Veo 3.1](/ai-models/video/google-veo-3-1) | Yes | 4/6/8s | 4K | Slower | Highest |
## Specifications
| Feature | Details |
| -------------------- | ------------------------------------------------- |
| **Developer** | Google DeepMind |
| **Resolution** | 720p, 1080p, 4K |
| **Duration** | 4, 6, or 8 seconds (selectable) |
| **Frame rate** | 24 FPS |
| **Audio** | No native audio |
| **Reference images** | Up to 3 |
| **Aspect ratios** | 16:9, 9:16 |
| **Generation time** | \~60–90s (720p), \~90–120s (1080p), \~2–3min (4K) |
| **Architecture** | Transformer backbone, spatio-temporal patches |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Google Veo 3.1 Fast** from the model dropdown.
Include scene description, subject behavior, camera movement, and audio environment details.
Add up to 3 reference images for character appearance or visual style anchoring.
Choose 720p, 1080p, or 4K depending on your output requirements and credit budget.
Click **Generate** and receive your output.
## Prompting tips
* **Describe the visual scene in detail** — "A waterfall cascades in the background, mist rising into the air" sets a vivid visual scene.
* **Use reference images for product or character consistency** — Upload a product shot or character photo as a reference to anchor the visual in your generated clip.
* **Be specific about camera framing** — "Tight close-up," "wide establishing shot," or "over-the-shoulder angle" guide Veo 3.1 Fast's framing decisions.
### Example prompts
> A barista steams milk in an artisan coffee shop. Close-up on the steam wand, foam forming. Warm, cozy lighting. 8 seconds, 1080p.
> A coastal drone shot at sunrise. Wide angle, slow forward movement over calm ocean. Golden light, misty horizon. 8 seconds, 4K.
## Compare models
| Model | Audio | Duration | Resolution | Speed | Best for |
| ---------------------------------------------------- | ----- | -------- | ---------- | -------- | ------------------------------- |
| **Veo 3.1 Fast** | Yes | 4/6/8s | Up to 4K | Balanced | Audio-visual production, 4K |
| [Veo 3.1 Lite](/ai-models/video/google-veo-3-1-lite) | No | 8s | 1080p | Fastest | Cost-efficient, no audio |
| [Veo 3.1](/ai-models/video/google-veo-3-1) | Yes | 4–8s | Up to 4K | Slowest | Max fidelity, broadcast quality |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | Yes | 25s | 1080p | Standard | Long-form A/V, physics |
Veo 3.1 Fast is the practical default choice in the Veo 3.1 family — it includes audio, supports 4K, and generates faster than the flagship. Move up to Veo 3.1 when you need the highest visual fidelity and full audio including dialogue and voiceover.
# Google veo 3 1 lite
Source: https://docs.imagine.art/ai-models/video/google-veo-3-1-lite
VIDEO MODEL
by Google DeepMind
Veo 3.1 family
Google Veo 3.1 Lite
Google DeepMind's cost-efficient Veo 3.1 variant — equal speed to Veo 3.1 Fast at less than 50% of the cost. Built for high-volume generation, developer applications, and use cases where output throughput matters as much as quality.
Cost
\<50% of Veo 3.1 Fast
Veo 3.1 Lite does not include native audio generation. If you need synchronized audio, use [Google Veo 3.1](/ai-models/video/google-veo-3-1) or [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) instead.
## Google's most cost-efficient video model
Google Veo 3.1 Lite, released March 31, 2026, completes the three-tier Veo 3.1 lineup — sitting below Veo 3.1 Fast in cost while matching it in inference speed. It uses the same Transformer backbone with spatio-temporal patches as the broader Veo 3.1 family, but is optimized for cost-efficient deployment in developer pipelines, high-volume generation, and use cases where output quantity matters alongside quality.
Duration is fixed at 8 seconds per clip.
## Capabilities
Less than 50% of the credit cost of Veo 3.1 Fast — the most affordable way to access the Veo 3.1 architecture on ImagineArt.
Consistent 8-second generation length — predictable credit consumption and straightforward pipeline integration.
Generates at 720p–1080p resolution, maintaining visual quality suitable for most digital delivery formats.
Shares the same Transformer backbone with spatio-temporal patches as Veo 3.1 Fast and Veo 3.1 — consistent scene coherence and prompt adherence.
Same generation speed as Veo 3.1 Fast — rapid turnaround without paying for the higher-cost tier.
Supports text-to-video and image-to-video workflows for creative flexibility.
## Veo 3.1 family comparison
| Model | Audio | Duration | Resolution | Relative cost | Best for |
| ---------------------------------------------------- | ----- | -------- | ---------- | ------------- | ------------------------------- |
| **Veo 3.1 Lite** | No | 8s | 1080p | Lowest | High-volume, cost-sensitive |
| [Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | Yes | 4/6/8s | Up to 4K | Medium | Balanced speed + quality |
| [Veo 3.1](/ai-models/video/google-veo-3-1) | Yes | 4/6/8s | Up to 4K | Highest | Broadcast quality, max fidelity |
## Specifications
| Feature | Details |
| ----------------- | --------------------------------------------- |
| **Developer** | Google DeepMind |
| **Released** | March 31, 2026 |
| **Resolution** | 720p–1080p |
| **Duration** | 8 seconds |
| **Aspect ratios** | 16:9, 9:16 |
| **Audio** | No |
| **Architecture** | Transformer backbone, spatio-temporal patches |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Google Veo 3.1 Lite** from the model dropdown.
Describe the scene, subject, camera movement, and mood. Veo 3.1 Lite responds well to cinematographic language.
Duration is fixed at 8 seconds.
Select 16:9 for landscape or 9:16 for vertical/mobile delivery.
Click **Generate** and review the 1080p output.
## Prompting tips
* **Be specific and concise** — Veo 3.1 Lite performs best with clear, direct prompts. Focus on the key visual elements: subject, action, setting, and lighting.
* **Plan for 8 seconds** — Structure your prompt as a complete 8-second scene with clear subject, action, and setting.
* **Use cinematographic language** — "Wide establishing shot," "tight close-up," "slow pan right" all guide the visual framing effectively.
### Example prompts
> A hummingbird hovers near a red flower in a sunlit garden. Close-up, natural light, green bokeh background. 8 seconds, 16:9.
> An architect reviews blueprints spread across a large table. Overhead shot, warm studio lighting, slight camera drift. 8 seconds.
For high-volume video generation pipelines where cost efficiency is a priority, Veo 3.1 Lite delivers the Veo 3.1 architecture at the lowest credit cost. If your use case requires audio or selectable clip length (4, 6, or 8s), step up to [Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) or [Veo 3.1](/ai-models/video/google-veo-3-1).
# Grok video
Source: https://docs.imagine.art/ai-models/video/grok-video
VIDEO MODEL
by xAI
Aurora architecture
xAI Grok Video
xAI's Aurora autoregressive video model — generates video in approximately 17 seconds with native audio including background music, sound effects, and ambient audio. Accepts up to 7 reference images for identity and style preservation, with text-to-video, image-to-video, and reference-to-video modes. Supports clips from 6 to 15 seconds.
Generation time
\~17 seconds
Audio
Music + SFX + Ambient
References
Up to 7 images
## The fastest AI video generation available
xAI's Grok Video is built on Aurora — an autoregressive architecture that predicts video frames sequentially rather than through the diffusion process used by most other models. This fundamental difference is what enables Aurora's \~17-second generation time, making it the fastest AI video model available on ImagineArt by a significant margin.
Despite the speed advantage, Grok Video delivers native audio (background music, sound effects, and ambient audio synchronized with the video), identity preservation with up to 7 reference images, and smooth natural motion. The reference-to-video mode is particularly strong: character identity, style, and visual consistency are preserved across the generation with minimal drift.
## Capabilities
Approximately 17 seconds per clip — the fastest generation time in the lineup. Enables rapid iteration at a pace no diffusion model can match.
Background music, sound effects, and ambient audio generated natively with the video — synchronized without post-production.
Identity and style preservation with up to 7 reference images — characters and visual styles are maintained consistently throughout the generated video.
Sequential frame prediction rather than diffusion — produces smooth, coherent motion with natural temporal consistency between frames.
Strong identity preservation in reference-based generation — character appearance, style, and smooth natural movement preserved from reference inputs.
Text-to-video, image-to-video, and reference-to-video — flexible workflow support from any starting point.
## Aurora vs. diffusion architecture
| Feature | **Grok Video (Aurora)** | Diffusion models |
| -------------------- | --------------------------- | ---------------------------- |
| Architecture | Autoregressive (sequential) | Diffusion (iterative) |
| Generation speed | \~17 seconds | 30 seconds – several minutes |
| Temporal consistency | Strong (sequential) | Variable |
| Output resolution | 720p | Up to 4K |
| Audio | Native | Varies |
## Specifications
| Feature | Details |
| -------------------- | ---------------------------------------------------- |
| **Developer** | xAI |
| **Architecture** | Aurora (autoregressive, sequential frame prediction) |
| **Resolution** | 720p |
| **Duration** | 6–15 seconds |
| **Frame rate** | — |
| **Audio** | Background music, SFX, ambient (native) |
| **Reference images** | Up to 7 |
| **Aspect ratios** | 16:9, 9:16, 1:1 |
| **Input modes** | Text-to-video, image-to-video, reference-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **xAI Grok Video** from the model dropdown.
Select text-to-video, image-to-video, or reference-to-video.
For reference-to-video, upload up to 7 reference images to anchor identity and visual style.
Describe the scene, motion, and audio environment. Include sound cues explicitly for the audio generation.
Click **Generate** — expect results in approximately 17 seconds.
## Prompting tips
* **Use it for rapid iteration** — 17-second generation means you can test 10–15 variations in the time it takes other models to produce 2 or 3. Explore directions aggressively before committing.
* **Audio cues work naturally** — "With upbeat jazz playing in the background" or "the sound of waves crashing" integrate naturally into Grok Video's audio generation.
* **Reference-to-video for consistent characters** — Upload multiple reference angles of a character (front, side, 3/4 view) to improve identity consistency across different generated scenes.
* **Keep prompts focused** — Aurora's sequential architecture produces the most coherent motion when the prompt describes a single, clear visual sequence rather than a complex multi-event narrative.
### Example prompts
> A golden retriever puppy plays in a field of daisies, tail wagging. Upbeat acoustic guitar music. Bright afternoon sunlight, slow motion on the playful moments. 6 seconds, 16:9.
> A barista writes a customer's name on a coffee cup with a marker. Soft café ambient sounds, quiet chatter in background. Close-up, handheld feel. 6 seconds.
## Compare models
| Model | Speed | Audio | References | Best for |
| --------------------------------------------------------- | --------- | ----- | ----------- | ------------------------ |
| **xAI Grok Video** | \~17s | Yes | Up to 7 | Maximum speed + audio |
| [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) | \~30s | No | — | Fast cinematic, no audio |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | Under 60s | No | Image input | Fast Seedance quality |
| [PixVerse v5](/ai-models/video/pixverse-v5) | \~30s | No | — | Fast character animation |
xAI Grok Video is the right choice when generation speed is a priority — especially for clients needing fast previews, high-volume production pipelines, or exploratory rapid iteration with audio. For maximum resolution or longer clips, other models in the lineup offer higher output specifications.
# Hailuo 02 pro
Source: https://docs.imagine.art/ai-models/video/hailuo-02-pro
VIDEO MODEL
by MiniMax
Hailuo 02 family
Hailuo 02 Pro
MiniMax's premium Hailuo 02 model — industry-leading physics-driven scene support, native 1080p output without upscaling, and the Noise-aware Compute Redistribution architecture at its maximum quality configuration for film, commercial, and high-fidelity production.
Architecture
NCR (Pro tier)
## Maximum physics fidelity
Hailuo 02 Pro is the premium configuration of MiniMax's Noise-aware Compute Redistribution (NCR) architecture — the same fundamental model as Hailuo 02 SD, tuned for maximum quality rather than cost efficiency. The defining differentiator is physics simulation: Hailuo 02 Pro delivers industry-leading physics-driven scene support, producing accurate simulations of fluid dynamics, rigid body collisions, deformation, and gravity that rival what's possible with physics engines in post-production VFX.
The no-upscaling 1080p output ensures that the resolution reflects genuine generative quality — not a post-processed enlargement.
## Capabilities
The most physically accurate video simulation available — fluid dynamics, rigid body collisions, deformation, and gravity rendered at a level that competes with VFX post-production.
True native 1080p — every pixel is generated at full resolution, preserving fine detail and avoiding the softness of upscaled video.
Noise-aware Compute Redistribution at the Pro quality tier — more compute allocated to complex scene areas for richer, more accurate output.
Supports both input modes with consistent physics behavior regardless of starting point — text prompts or reference images both feed the same physics engine.
Output quality suitable for film production, advertising, educational content, and high-profile social media where physical realism matters.
Characters and objects maintain their visual identity throughout the clip, with accurate interaction physics between subjects.
## Specifications
| Feature | Details |
| ---------------- | ---------------------------------------------- |
| **Developer** | MiniMax |
| **Architecture** | Noise-aware Compute Redistribution (NCR) — Pro |
| **Resolution** | Native 1080p (no upscaling) |
| **Duration** | 6 seconds |
| **Frame rate** | 24–30 FPS |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
| **Physics** | Industry-leading simulation |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Hailuo 02 Pro** from the model dropdown.
For maximum physics quality, describe physical interactions explicitly — material types, forces acting on objects, and the expected behavior you want to see.
For image-to-video, upload your scene reference and describe the motion and physical events.
Duration is fixed at 6 seconds.
Click **Generate** for native 1080p output with maximum physics fidelity.
## Prompting tips
* **Name the physical interaction explicitly** — "A ceramic mug falls from a table and shatters on a tile floor" will trigger accurate rigid body simulation. The more specifically you describe the physical event, the better the physics engine responds.
* **Describe material properties** — "Heavy cast iron," "thin glass," "wet sand" — material descriptions activate different physical simulation parameters.
* **Slow motion emphasizes physics** — Slow-motion prompts reveal the physics simulation detail more clearly, especially for liquid and impact scenes.
* **Camera framing matters** — Close-ups on physical interactions let the model dedicate rendering resources to the physics detail rather than wide-scene environment.
### Example prompts
> A champagne bottle is uncorked in extreme slow motion. Cork flies out, foam bubbles surge upward, liquid cascades down the neck. Black background, dramatic backlight. Native 1080p, 6 seconds.
> A skateboarder performs a trick and lands hard. Realistic impact forces — board flex, body weight transfer. Concrete texture, urban setting. 6 seconds, 16:9.
## Compare models
| Model | Physics | Quality tier | Audio | Best for |
| ------------------------------------------------- | ---------------- | ------------ | ----- | ------------------------------ |
| **Hailuo 02 Pro** | Industry-leading | Pro | No | Maximum physics fidelity |
| [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) | High | Standard | No | Cost-efficient physics content |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | High + stylized | Pro | No | Stylized + physics |
| [Kling O3](/ai-models/video/kling-o3) | Advanced | Pro | Yes | Physics + audio + 4K |
Hailuo 02 Pro is the right choice when the physical realism of your scene is the primary concern and cost is secondary. For stylized content (anime, illustration) with strong physics, consider [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro).
# Hailuo 02 sd
Source: https://docs.imagine.art/ai-models/video/hailuo-02-sd
VIDEO MODEL
by MiniMax
Hailuo 02 family
Hailuo 02 SD
MiniMax's standard-tier Hailuo 02 model — built on the novel Noise-aware Compute Redistribution (NCR) architecture with 3× more parameters and 4× more training data than its predecessor. 768p output, 2.5× faster inference, and ranked #2 globally on Artificial Analysis benchmarks.
Benchmark rank
#2 globally
## A new architecture for physics mastery
Hailuo 02 SD is built on MiniMax's Noise-aware Compute Redistribution (NCR) architecture — a novel approach that reallocates compute resources based on noise level during the diffusion process, directing more processing power where it's needed most. The result is a 2.5× speed improvement over the previous generation alongside significantly better output quality.
With 3× more parameters and 4× more training data than Hailuo 01, Hailuo 02 SD achieves notably better physical scene simulation — object interactions, fluid dynamics, and gravity-driven motion are rendered with accuracy that earned it a #2 global ranking on Artificial Analysis video benchmarks.
## Capabilities
Object interactions, fluid dynamics, and gravity-driven motion are rendered with high physical accuracy — a defining strength of the NCR architecture.
Generates at 768p resolution for efficient, high-quality output.
Noise-aware Compute Redistribution dynamically allocates processing power during diffusion, producing better results faster than conventional diffusion approaches.
Significantly reduced inference time versus Hailuo 01 — practical for production pipelines where generation speed matters.
Supports text-to-video and image-to-video workflows with consistent subject rendering across the generated clip.
Choose between 6-second or 10-second generation lengths for flexible content creation.
## Hailuo 02 SD vs Pro
| Feature | **Hailuo 02 SD** | [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) |
| ------------------ | ------------------------- | ----------------------------------------------- |
| Architecture | NCR | NCR |
| Resolution | 768p | 1080p |
| Physics simulation | High | Industry-leading |
| Duration | 6 or 10s | 6s |
| Cost tier | Standard | Pro |
| Best for | Cost-efficient production | Maximum physics fidelity |
## Specifications
| Feature | Details |
| ----------------- | ---------------------------------------- |
| **Developer** | MiniMax |
| **Architecture** | Noise-aware Compute Redistribution (NCR) |
| **Resolution** | 768p |
| **Duration** | 6 or 10 seconds |
| **Frame rate** | 24–30 FPS |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
| **Parameters** | 3× more than Hailuo 01 |
| **Training data** | 4× more than Hailuo 01 |
| **Benchmark** | #2 globally, Artificial Analysis |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Hailuo 02 SD** from the model dropdown.
For text-to-video, describe the scene, physics-heavy action, and visual style. For image-to-video, upload your reference and describe the motion.
Select 6 or 10 seconds depending on the complexity of your scene.
Click **Generate** and review the 768p output.
## Prompting tips
* **Lean into physics** — Hailuo 02 SD handles physical interactions exceptionally well. Prompts involving water, falling objects, explosions, or impact scenes will produce notably accurate results.
* **Cinematic camera moves** — "Slow pan left," "push-in shot," and "crane shot" are all well-supported camera movements.
* **Be specific about materials** — "Water splashing on stone," "smoke billowing from a chimney," or "glass shattering" will trigger the physics simulation accurately.
### Example prompts
> A basketball player dunks in slow motion. Ball impacts the rim and net with realistic physics. Arena lights, crowd noise suggested by visual cues. 6 seconds, 768p.
> Waves crash against rocky cliffs at sunset. Wide shot, dramatic foam and spray, physically realistic water simulation. 10 seconds, 16:9.
## Compare models
| Model | Physics | Resolution | Cost | Best for |
| ----------------------------------------------- | ---------------- | ------------ | -------- | -------------------------------- |
| **Hailuo 02 SD** | High | 1080p native | Standard | Physics-accurate, cost-efficient |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | Industry-leading | 1080p native | Pro | Maximum fidelity physics |
| [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) | High + stylized | 768p/1080p | Standard | Stylized + physics |
| [Wan 2.2](/ai-models/video/wan-2-2) | Good | 1080p | Standard | Open-source, camera control |
Hailuo 02 SD is the standard-quality entry point for the Hailuo 02 architecture. For scenes requiring the highest possible physics fidelity, upgrade to [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro).
# Hailuo 2 3 pro
Source: https://docs.imagine.art/ai-models/video/hailuo-2-3-pro
VIDEO MODEL
by MiniMax
Released October 2025
Hailuo 2.3 Pro
MiniMax's premium Hailuo 2.3 configuration — maximum physics fidelity, enhanced facial micro-expression rendering, anime, illustration, ink-wash, and game-CG stylization options, and professional-grade output for demanding art direction and animation production.
## Maximum quality stylized video
Hailuo 2.3 Pro is the quality-priority configuration of MiniMax's Hailuo 2.3 model — the same architecture as Hailuo 2.3 SD, tuned for maximum output quality rather than cost efficiency. The improvements over Hailuo 2.3 SD are most visible in complex scenes: physics-driven interactions render with higher accuracy, character expressions carry more nuanced detail, and stylized art styles (anime, illustration, ink-wash, game-CG) produce more polished, consistent results.
A fast variant is available at 50% reduced cost for volume generation at the same quality profile.
## Capabilities
The highest quality output for anime, illustration, ink-wash, and game-CG styles — more consistent linework, richer color depth, and better artistic coherence than the SD tier.
More detailed facial expression modeling — subtle emotional nuances, natural eye animation, and realistic lip movement are all rendered with higher precision.
Improved physics accuracy for object interactions, material behavior, and environment dynamics in complex stylized scenes.
Output quality suitable for commercial animation, game cinematics, and professional art-directed video production.
Pro quality at 50% lower cost via the fast variant — same quality profile, reduced generation cost for higher-volume projects.
Full text-to-video and image-to-video support with style consistency across both input modes.
## Specifications
| Feature | Details |
| ---------------- | ------------------------------------------------------ |
| **Developer** | MiniMax |
| **Released** | October 2025 |
| **Resolution** | 1080p |
| **Duration** | 6 seconds |
| **Audio** | No native audio |
| **Styles** | Anime, illustration, ink-wash, game-CG, photorealistic |
| **Input modes** | Text-to-video, image-to-video |
| **Fast variant** | Yes (50% lower cost) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Hailuo 2.3 Pro** from the model dropdown.
Include the art style explicitly: "anime style," "game-CG cinematic," "traditional ink-wash," or "detailed illustration."
For the best use of the enhanced micro-expression system, describe the emotional state and specific facial nuances you want.
Output is 1080p at 6 seconds.
Click **Generate** for professional-quality stylized output.
## Prompting tips
* **Push stylization further** — Hailuo 2.3 Pro handles more complex style descriptions than the SD tier. "Studio Ghibli-inspired watercolor animation with a soft, hand-painted texture" will produce coherent results.
* **Combine styles with physics** — "A ninja flips through falling cherry blossom petals in anime style" benefits from both the stylization and physics systems simultaneously.
* **Expression + micro-cue descriptions** — "Her eyes well with tears, a single drop forming at the corner — she blinks slowly and turns away" will be rendered with the enhanced micro-expression system.
* **For game cinematics** — "Unreal Engine 5-style game cinematic, photorealistic character models, dramatic cinematic lighting, high-production value" produces strong game-CG outputs.
### Example prompts
> Studio Ghibli-style animation: a young girl stands at the top of a hill, wind blowing through her hair. The countryside stretches below. Soft afternoon light, watercolor sky. Her expression is serene, eyes distant. 6 seconds, 1080p.
> Game-CG cinematic: two knights clash swords in a rain-soaked castle courtyard. High-detail armor, realistic water splashing underfoot, dramatic lightning in the background. 6 seconds, 1080p.
## Compare models
| Model | Style quality | Physics | Resolution | Best for |
| ----------------------------------------------- | ------------- | ---------------- | ---------- | ------------------------------- |
| **Hailuo 2.3 Pro** | Maximum | High | 768p/1080p | Stylized + physics, pro quality |
| [Hailuo 2.3 SD](/ai-models/video/hailuo-2-3-sd) | High | Good | 768p/1080p | Stylized, cost-efficient |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | Standard | Industry-leading | 1080p | Maximum physics, realistic |
| [PixVerse v6](/ai-models/video/pixverse-v6) | Good | Standard | 1080p | Lens control, audio |
Hailuo 2.3 Pro delivers the most polished stylized video output available — particularly for anime, illustration, and game cinematics. For purely realistic footage with maximum physics accuracy, [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) remains the stronger choice.
# Hailuo 2 3 sd
Source: https://docs.imagine.art/ai-models/video/hailuo-2-3-sd
VIDEO MODEL
by MiniMax
Released October 2025
Hailuo 2.3 SD
MiniMax's standard-tier Hailuo 2.3 — significantly enhanced body movement accuracy, facial micro-expression detail, expanded stylization support including anime, illustration, ink-wash, and game-CG art styles, and improved physics versus Hailuo 2.0.
Styles
Anime, illustration, game-CG
## Style-first video generation
Hailuo 2.3 SD, released October 2025, is MiniMax's stylization-focused upgrade to the Hailuo generation. Where Hailuo 02 focused on physics realism for naturalistic footage, Hailuo 2.3 expands the stylization palette — anime, illustration, ink-wash painting, and game-CG art styles all produce high-quality results with strong visual consistency.
Facial micro-expression modeling is a notable improvement: subtle emotional cues — raised eyebrows, slight lip curls, eye movements — are captured with greater detail, making character animation more expressive and lifelike.
## Capabilities
Strong results across anime, illustration, ink-wash painting, and game-CG art styles — one of the most versatile stylization models in the lineup.
Improved micro-expression modeling — subtle emotional cues like brow movements, eye direction, and lip nuances are captured with higher fidelity.
Improved accuracy for human body movement — natural gait, gesture flow, and posture transitions versus Hailuo 2.0.
Improved physics rendering over Hailuo 2.0 — better handling of gravity, material interactions, and environmental physics.
Available in standard quality and a fast variant that reduces generation cost by 50% — choose based on quality vs. throughput needs.
Supports text-to-video and image-to-video with consistent style and subject rendering.
## Hailuo 2.3 tier comparison
| Feature | **Hailuo 2.3 SD** | [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) |
| ---------------------- | -------------------------------------- | ------------------------------------------------- |
| Style support | Anime, illustration, ink-wash, game-CG | Same + more options |
| Physics quality | High | Higher |
| Expression detail | Enhanced | Enhanced |
| Cost tier | Standard | Pro |
| Fast variant available | Yes (50% cost reduction) | Yes |
## Specifications
| Feature | Details |
| ---------------- | ------------------------------------------------------ |
| **Developer** | MiniMax |
| **Released** | October 2025 |
| **Resolution** | 768p |
| **Duration** | 6 or 10 seconds |
| **Audio** | No native audio |
| **Styles** | Anime, illustration, ink-wash, game-CG, photorealistic |
| **Input modes** | Text-to-video, image-to-video |
| **Fast variant** | Yes (50% lower cost) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Hailuo 2.3 SD** from the model dropdown.
Include the art style explicitly in your prompt: "anime style," "ink-wash painting," "game-CG render," or "illustration style" to activate the model's stylization system.
Choose 6 or 10 seconds at 768p.
Click **Generate** and review the stylized output.
## Prompting tips
* **Name the art style explicitly** — "Anime-style," "traditional Chinese ink-wash painting style," or "JRPG game-CG cinematic style" activate significantly different visual rendering.
* **Describe expressions in detail** — For character-focused content, "a subtle smile forming slowly, eyes crinkling slightly at the corners" will be rendered with the improved micro-expression system.
* **Anime prompt tips** — Include "cel-shading," "vibrant saturated colors," or "hand-drawn animation aesthetic" for the most authentic anime output.
* **Ink-wash tips** — "Black ink on white paper, minimalist brush strokes, traditional Chinese landscape" works well for the ink-wash style.
### Example prompts
> Anime-style heroine sits beneath a cherry blossom tree. Petals drift in the breeze. Close-up on her face — a gentle smile, eyes reflecting the falling blossoms. Vibrant colors, cel-shaded. 6 seconds.
> Traditional Chinese ink-wash painting: a lone crane wades in still water at dawn. Minimal brush strokes, misty mountains in the background. Slow, contemplative movement. 10 seconds.
## Compare models
| Model | Stylization | Physics | Expression | Best for |
| ------------------------------------------------- | ----------- | -------- | ---------- | -------------------------------- |
| **Hailuo 2.3 SD** | Excellent | High | Enhanced | Stylized content, cost-efficient |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | Excellent | Higher | Enhanced | Stylized + maximum physics |
| [Hailuo 02 SD](/ai-models/video/hailuo-02-sd) | Standard | High | Standard | Physics realism |
| [PixVerse v5](/ai-models/video/pixverse-v5) | Good | Standard | Standard | Character animation, anime |
Hailuo 2.3 SD is the best model for stylized video content — especially anime, illustration, and game-CG styles. For natural/realistic footage with maximum physics accuracy, [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) is the better choice.
# Happy horse
Source: https://docs.imagine.art/ai-models/video/happy-horse
VIDEO MODEL
by Alibaba
Happy Horse
Alibaba's flagship video model — built for fluid, lifelike motion with native audio generation, selectable durations from 3 to 15 seconds, and output up to 1080p.
## Fluid, lifelike motion from Alibaba
Happy Horse is Alibaba's best video model, engineered specifically for natural, physics-consistent motion. It generates videos up to 15 seconds long at resolutions between 720p and 1080p, with native audio output — dialogue, ambient sound, and environmental effects — generated alongside the video in a single pass.
The model excels at scenes requiring believable organic movement: human motion, natural environments, animals, and fluid dynamics all render with a level of realism that makes the output feel grounded rather than synthetic. Native audio completes the picture by matching the generated soundscape to the visual content without post-processing.
## Capabilities
Engineered for natural movement — human motion, environmental dynamics, and organic subjects render with realistic physics and consistent body mechanics.
Generates audio alongside video in a single pass — ambient sound, environmental effects, and dialogue without requiring separate post-processing.
Selectable resolution between 720p and 1080p for flexible delivery across social, web, and production pipelines.
Generate clips from 3 to 15 seconds — enough length for full narrative beats, product demonstrations, or scene-level storytelling.
Provide a reference image as the opening frame to anchor the model's visual output to a specific subject, composition, or environment.
Handles complex visual scenes — crowd motion, environmental weather, lighting changes — with temporal consistency across the full clip.
## Specifications
| Feature | Details |
| ---------------- | ---------------------------- |
| **Developer** | Alibaba |
| **Resolution** | 720p–1080p |
| **Duration** | 3–15 seconds |
| **Audio** | Native audio generation |
| **Input** | Start frame (image-to-video) |
| **Base credits** | 252 |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Happy Horse** from the model dropdown.
Upload an image to anchor the opening composition. If skipped, the model generates from the text prompt alone.
Describe the scene, motion, atmosphere, and any audio direction. Be specific about how subjects and the environment should move.
Choose a clip length between 3 and 15 seconds depending on your content needs.
Click **Generate**. Happy Horse produces a video with synchronized native audio.
## Prompting tips
* **Describe motion specifically** — Happy Horse rewards precise motion language. "The subject walks slowly across the frame" produces more consistent results than "someone moving."
* **Include audio direction** — Since audio is generated natively, describe what you want to hear: "light rain on pavement," "crowd murmur in background," or "ambient wind."
* **Use the start frame for subject anchoring** — If your scene has a specific character or environment, upload a reference image. The model will maintain its appearance throughout the clip.
* **Match duration to content** — Simple motion reads well at 3–5 seconds. Multi-beat scenes or longer narratives benefit from 8–15 seconds.
### Example prompts
> A woman walks through a sunlit park in slow motion, leaves drifting around her. Soft ambient birdsong and gentle wind. 1080p, 10 seconds.
> A tiger moves through tall grass at dusk, each step deliberate. Low ambient hum of insects, distant thunder. Wide shot. 15 seconds.
> Ocean waves crash against rocky cliffs at golden hour. Spray catches the light. Deep resonant sound of water against stone. 8 seconds.
## Compare models
| Model | Resolution | Audio | Duration | Best for |
| ----------------------------------------------- | ---------- | ----- | -------- | --------------------------------------- |
| **Happy Horse** | 720p–1080p | Yes | 3–15s | Fluid lifelike motion with native audio |
| [Wan 2.6](/ai-models/video/wan-2-6) | 720p–1080p | Yes | 5–15s | Character reference-to-video, R2V |
| [Wan 2.5](/ai-models/video/wan-2-5) | 480p–1080p | Yes | 5–10s | Audio-visual sync, lip-sync |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 1080p | Yes | 3–15s | Multi-shot storytelling, 60 FPS |
| [Seedance 2](/ai-models/video/seedance-2) | 720p–1080p | Yes | 4–15s | Multimodal references, full production |
Happy Horse is the right choice when natural, physics-consistent motion is the priority and you want native audio included without extra steps. For multi-shot storyboarding, compare with Kling 3.0 Pro. For character identity across scenes, compare with Wan 2.6.
# Kling 2 1 pro
Source: https://docs.imagine.art/ai-models/video/kling-2-1-pro
VIDEO MODEL
by Kling AI
Kling 2.1 family
Kling 2.1 Pro
Kling AI's professional-tier 2.1 model — first-and-last-frame conditioning for precise transition control, enhanced sharpness with realistic lighting, advanced interpolation for crisp motion detail, and 1080p professional-grade visual fidelity.
Frame control
First + last frame
## Precision image animation
Kling 2.1 Pro is the professional-tier configuration of the Kling 2.1 generation — focused on high-fidelity image animation with first-and-last-frame conditioning. This means you can specify both the opening and closing image of a clip, giving you deterministic control over what the video looks like at the start and end while the model handles the transition.
Enhanced sharpness and advanced interpolation deliver crisp motion at every frame — a step up from the standard 2.1 tier for commercial and professional content creation.
## Capabilities
Define the exact opening and closing frames — Kling 2.1 Pro generates the motion between your specified images for precise transition control.
Improved rendering for fine detail, realistic lighting, and surface clarity versus the standard 2.1 tier.
Crisp, detailed motion between frames — no blurry or artifact-heavy transitions in fast-moving sequences.
1080p output with color accuracy and visual consistency suitable for commercial, advertising, and content production.
Animate any reference image with natural motion, accurate lighting continuity, and subject consistency.
Generates directly from text prompts with strong prompt adherence for scene composition, subject behavior, and camera style.
## Specifications
| Feature | Details |
| ----------------- | --------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Resolution** | Up to 1080p |
| **Duration** | 5–10 seconds |
| **Frame control** | First and last frame conditioning |
| **Aspect ratios** | Multiple supported |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling 2.1 Pro** from the model dropdown.
Upload the starting frame — the image you want to animate.
Upload a second image to use as the ending frame for a controlled transition.
Write a prompt describing how the scene should move — camera behavior, subject action, atmosphere.
Click **Generate** for pro-tier fidelity output.
Go to the **AI Video Generator** and select **Kling 2.1 Pro**.
Describe the scene, subject movement, camera style, and lighting. Be specific about visual quality expectations.
Choose 5 or 10 seconds and your preferred resolution up to 1080p.
Click **Generate** and review the output.
## Prompting tips
* **For first-and-last-frame: let the images do the visual work** — Your prompt should focus on motion style, pacing, and atmosphere rather than describing what's visible in the frames.
* **Specify motion timing** — "Slowly" vs "quickly" significantly changes the feel. "The character turns in 2 seconds with a graceful, measured movement" is more useful than just "the character turns."
* **Use lighting continuity** — If your reference image has specific lighting, describe it in the prompt so the model maintains consistency through the motion.
### Example prompts
> A luxury watch sits on a velvet surface. The camera slowly orbits around it, revealing the face and side profile. Soft studio lighting, no harsh shadows. 8 seconds.
> A cityscape transitions from dusk to night. Time-lapse style, lights flickering on across the skyline, smooth motion. 10 seconds, 16:9.
## Compare models
| Model | Frame control | Resolution | Audio | Best for |
| ----------------------------------------------- | ------------------ | ---------- | ----- | -------------------------- |
| **Kling 2.1 Pro** | First + last frame | 1080p | No | Controlled image animation |
| [Kling 2.5 Pro](/ai-models/video/kling-2-5-pro) | No | 1080p | No | Fast, cost-efficient HD |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | No | 1080p | Yes | Audio-synced production |
| [Pika 2.2](/ai-models/video/pika-2-2) | Pikaframes | 1080p | No | Keyframe precision |
Kling 2.1 Pro is ideal when you need to animate from a known start image to a known end image. For audio-synchronized content, step up to [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro).
# Kling 2 5 pro
Source: https://docs.imagine.art/ai-models/video/kling-2-5-pro
VIDEO MODEL
by Kling AI
Kling 2.5 family
Kling 2.5 Pro
Kling AI's cost-efficient professional model — 3× faster generation than previous Kling models, a state-of-the-art physics engine, 30% lower cost, multi-reference support with up to 4 input images, and 1080p HD output.
References
Up to 4 images
## Speed, physics, and value
Kling 2.5 Pro was designed around three principles: speed, physics accuracy, and cost efficiency. At 3× faster than previous Kling models and 30% lower in cost, it's the most practical HD Kling model for volume production workflows and budget-conscious professional content.
The state-of-the-art physics engine handles realistic motion — gravity, momentum, and material behavior all render accurately, making Kling 2.5 Pro particularly strong for sports content, action scenes, and any footage where physical plausibility is important. Multi-reference support (up to 4 input images) enables character and style consistency across a generation.
## Capabilities
Generates 3× faster than previous Kling Pro models — significantly reduces time-to-result for production workflows.
Gravity, momentum, and material behavior render accurately — noted as best-in-class for AI sports video due to realistic motion physics.
30% more affordable than comparable previous Kling models — practical for high-volume and budget-sensitive production.
Accepts up to 4 reference images to anchor character appearance, visual style, and scene consistency.
Full HD 1080p resolution at 5 or 10 seconds for professional deliverables.
Supports both text-to-video and image-to-video workflows with strong prompt adherence.
## Specifications
| Feature | Details |
| -------------------- | ---------------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Resolution** | 1080p HD |
| **Duration** | 5 or 10 seconds |
| **Speed** | 3× faster than previous Kling Pro models |
| **Cost** | 30% lower than previous Kling models |
| **Reference images** | Up to 4 |
| **Aspect ratios** | 16:9, 9:16, 1:1 |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling 2.5 Pro** from the model dropdown.
Upload up to 4 reference images for character appearance or visual style anchoring.
Describe the scene, motion physics, camera behavior, and atmosphere. For sports or action content, describe the specific physical action explicitly.
Select 5 or 10 seconds and your preferred aspect ratio.
Click **Generate** for fast 1080p output.
## Prompting tips
* **Describe physics explicitly** — "A soccer player kicks the ball with full force, realistic ball trajectory, natural spin" — the physics engine responds accurately to described physical events.
* **Use references for consistent characters** — For brand content or recurring characters, upload reference images to anchor visual identity across generations.
* **Sports and action work well** — Kling 2.5 Pro's physics accuracy makes it particularly strong for sports footage, action sequences, and any physically demanding motion.
### Example prompts
> A tennis player serves a ball at full power. Slow-motion capture of the ball impact, realistic ball deformation and spin. Courtside, natural daylight. 5 seconds, 16:9.
> Two fencers spar in an elegant dojo. Steel blades catch the light on each clash. Realistic weight and momentum in every movement. 10 seconds.
## Kling Pro family comparison
| Model | Speed | Cost | Audio | Resolution | Best for |
| ----------------------------------------------- | --------- | --------- | ----- | ---------- | -------------------------- |
| **Kling 2.5 Pro** | 3× faster | 30% lower | No | 1080p | Cost-efficient HD, sports |
| [Kling 2.1 Pro](/ai-models/video/kling-2-1-pro) | Standard | Standard | No | 1080p | First + last frame control |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | Standard | Standard | Yes | 1080p | Audio-synced production |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Standard | Higher | Yes | 4K | Cinematic 4K, multi-shot |
Kling 2.5 Pro offers the best value-to-quality ratio in the Kling lineup for standard professional content. For audio-synchronized output, step up to [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro).
# Kling 2 6 pro
Source: https://docs.imagine.art/ai-models/video/kling-2-6-pro
VIDEO MODEL
by Kling AI
Released December 2025
Kling 2.6 Pro
Kling AI's audio-synchronized flagship — simultaneous audio and video generation from a single pass, English and Chinese lip-sync with tone-accurate singing, 48 FPS smooth output, and enhanced full-body motion fidelity for martial arts, dance, and high-speed action sequences.
Audio
Dialogue + SFX + Singing
## Simultaneous audio and video
Kling 2.6 Pro, released December 2025, introduced simultaneous audio-visual generation to the Kling Pro lineup — audio is not added after video generation but produced in a single pass alongside the visuals. This ensures tight synchronization between lip movements, dialogue, sound effects, and ambient audio.
The lip-sync system supports both English and Chinese dialogue, narration, and singing — with accurate tone production for singing content, not just spoken words. At 48 FPS, motion sequences — particularly martial arts, dance, and fast physical action — are rendered with the smoothness typically associated with high-frame-rate broadcast and sports content.
## Capabilities
Audio and video generated in a single pass — tight synchronization between dialogue, lip movements, sound effects, and ambient audio.
Accurate lip-sync for English and Chinese dialogue and narration — tone-accurate singing in both languages.
High frame rate output at 48 FPS — smooth motion for dance, martial arts, sports, and fast action sequences.
Improved fidelity for fast, intricate full-body movements — martial arts, dance, gymnastics — with no ghosting or body part distortion.
Accepts motion reference clips (3–30 seconds) to anchor specific movement patterns and action sequences.
Native sound effects and ambient noise generation — footsteps, environment sounds, impact effects — synchronized to the visual action.
## Specifications
| Feature | Details |
| -------------------- | -------------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Released** | December 2025 |
| **Resolution** | 1080p |
| **Frame rate** | 48 FPS |
| **Duration** | Up to 10 seconds |
| **Audio** | Dialogue, SFX, ambient sounds, singing |
| **Languages** | English, Chinese |
| **Lip-sync** | Yes — including singing |
| **Motion reference** | 3–30 seconds |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling 2.6 Pro** from the model dropdown.
Include explicit audio cues in your prompt: dialogue lines, sound effect descriptions, music style, and ambient environment.
Choose up to 10 seconds.
Click **Generate** for 1080p output at 48 FPS with synchronized audio.
Go to the **AI Video Generator** and select **Kling 2.6 Pro**.
Upload a motion reference clip (3–30 seconds) to anchor the movement pattern you want.
Describe the character, setting, and any audio elements to generate around the referenced motion.
Kling 2.6 Pro applies the motion pattern to the generated character with 48 FPS smoothness.
## Prompting tips
* **Include dialogue in quotes** — "A character says 'Welcome home' warmly" — quoted text is interpreted as a lip-sync target for the audio generation.
* **Specify Chinese or English explicitly** — "The character speaks in Mandarin Chinese" or "narration in English" ensures accurate phoneme production.
* **Singing works** — "A singer performs a pop chorus, upbeat tempo, clear pronunciation" will produce tone-accurate singing with synchronized lip movements.
* **48 FPS rewards fast motion** — Prompts involving dance, martial arts, and sports produce their best results at 48 FPS. Describe the full action to benefit from the frame rate.
### Example prompts
> A pop singer performs on stage under colorful spotlights. The camera slowly circles. The singer sings in English with clear enunciation. Upbeat music, crowd cheering in the background. 10 seconds, 1080p.
> A martial artist performs a high-speed combination — three kicks and a spinning strike. 48 FPS, smooth motion, dojo setting, impact sound effects synchronized to each strike.
## Compare models
| Model | Audio | Lip-sync | FPS | Motion fidelity | Best for |
| ----------------------------------------------------- | ----- | ------------- | --- | --------------- | --------------------------- |
| **Kling 2.6 Pro** | Yes | EN + Chinese | 48 | Enhanced | Audio-synced, fast motion |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Yes | Multilingual | 60 | Strong | 4K cinematic multi-shot |
| [Kling O3](/ai-models/video/kling-o3) | Yes | 10+ languages | 60 | Advanced | Physics + audio, 4K |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | Yes | 8+ languages | — | Good | Multilingual dialogue focus |
Kling 2.6 Pro is the best model for content where EN/Chinese lip-sync and high-frame-rate motion quality are both priorities — brand films, K-pop style content, dialogue-driven action, and singing videos.
# Kling 3 0 4k
Source: https://docs.imagine.art/ai-models/video/kling-3-0-4k
VIDEO MODEL
by Kling AI
Kling 3 family
Kling 3.0 4K
Kling AI's latest 3.0 model in 4K — delivers the full Kling 3.0 feature set at native 4K resolution with native audio, first-and-last-frame control, and clips up to 15 seconds.
Duration
Up to 15 seconds
## Kling 3.0 at native 4K
Kling 3.0 4K is the 4K-output tier of Kling AI's 3.0 model family — bringing the same MVL (Multi-modal Visual Language) architecture as Kling 3.0 Pro to native 4K resolution. It supports first-and-last-frame conditioning, native audio generation, and clips up to 15 seconds, making it the highest-resolution option in the Kling lineup.
Where Kling 3.0 Pro targets cinematic storytelling at 1080p/60 FPS, Kling 3.0 4K prioritizes maximum output resolution for productions that require the finest pixel detail — broadcast delivery, large-format display, or post-production workflows where 4K source material is mandatory.
## Capabilities
Generates at native 4K resolution — the highest output in the Kling 3.0 family, suited for broadcast, large-format, and post-production workflows.
Audio is generated alongside video in a single pass — ambient sound, dialogue, and environmental effects without separate post-processing.
Define the opening and closing frame of the clip. The model generates the motion between them, giving you precise control over transitions.
Generate clips from 3 to 15 seconds — enough for full narrative beats and cinematic sequences at 4K.
Multi-modal Visual Language architecture processes text, images, and audio as unified inputs for coherent multimodal output at 4K.
Handles fast physical actions — sports, dance, environmental dynamics — with consistent motion fidelity at the full 4K resolution.
## Specifications
| Feature | Details |
| ---------------- | --------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Base credits** | 755 |
| **Resolution** | 4K |
| **Duration** | 3–15 seconds |
| **Audio** | Native audio generation |
| **Input** | Start frame, end frame, or both |
| **Architecture** | Multi-modal Visual Language (MVL) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling 3.0 4K** from the model dropdown.
Upload a start frame to anchor the opening composition, an end frame to define where the clip ends, or both to control the full transition.
Describe the scene, motion, camera direction, and audio you want. Include any details about lighting, atmosphere, and subject behavior.
Choose a clip length between 3 and 15 seconds.
Click **Generate**. Kling 3.0 4K produces a 4K clip with synchronized native audio.
## Prompting tips
* **Use first-and-last-frame for controlled transitions** — Upload both a start and end frame when you need precise control over how a scene opens and closes. The model handles the motion between them.
* **Include audio direction in the prompt** — Native audio responds to prompt language: "ambient rain," "orchestral underscore," or "crowd applause" all meaningfully shape the audio output.
* **Technical camera language works well** — "Slow push in," "aerial pull-back," "rack focus from foreground to background" all produce distinct cinematic results at 4K.
* **Reserve 4K for final delivery** — For iteration and drafting, use [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) at lower credit cost. Move to 4K for final output when resolution is the priority.
### Example prompts
> A mountain range at sunrise, mist filling the valleys. The camera slowly pulls back from a tight close-up of frost on pine needles to a sweeping wide shot. Ambient wind and birdsong. 4K, 10 seconds.
> A product on a rotating pedestal, crisp studio lighting with subtle shadow movement. Clean ambient tone, no music. 4K, 5 seconds.
> A couple walks along a rain-soaked street at night, neon reflections in puddles. Soft jazz in the background, rain on pavement. 4K, 15 seconds.
## Compare models
| Model | Resolution | Audio | Duration | Best for |
| ----------------------------------------------------------- | ---------- | ----- | -------- | ------------------------------------ |
| **Kling 3.0 4K** | 4K | Yes | 3–15s | Maximum resolution, 4K delivery |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 1080p | Yes | 3–15s | Multi-shot storytelling, 60 FPS |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | Up to 4K | Yes | 4–8s | Broadcast 4K with 48kHz stereo audio |
| [Google Veo 3.1 Fast](/ai-models/video/google-veo-3-1-fast) | Up to 4K | No | 4–8s | Cost-efficient 4K |
| [Kling O3](/ai-models/video/kling-o3) | 1080p | No | 5–10s | Advanced physics, 6 generation modes |
Kling 3.0 4K is the right choice when you need native 4K resolution with audio and the longest available durations in the Kling 3.0 family. For 60 FPS multi-shot storytelling at lower credit cost, use Kling 3.0 Pro.
# Kling 3 0 pro
Source: https://docs.imagine.art/ai-models/video/kling-3-0-pro
VIDEO MODEL
by Kling AI
Kling 3 family
Kling 3.0 Pro
Kling AI's most advanced video model — 1080p at 60 FPS, Omni Native Audio with multilingual dialogue and environmental soundscapes, and the ability to generate up to 6 distinct shots in a single 15-second output.
Duration
Up to 15 seconds
Shots per generation
Up to 6
## Kling 3.0 Pro
Kling 3.0 Pro marks Kling AI's most significant architectural leap — 1080p output at 60 frames per second with Omni Native Audio and multi-shot storyboarding in a single generation.
The Multi-modal Visual Language (MVL) architecture unifies text, image, video, and audio inputs into a single model, enabling true multi-shot storyboarding — up to 6 distinct shots, each with specified duration, shot size, perspective, narrative, and camera movement, all generated from one prompt.
## Capabilities
Generates 1080p video at 60 frames per second — smooth, high frame-rate output for cinematic and action-heavy content.
Multilingual audio generation including English, Japanese, Korean, Spanish, and environmental soundscapes — generated natively alongside the video.
Specify up to 6 shots in a single 15-second generation — each with its own duration, shot size, perspective, camera movement, and narrative.
Multi-modal Visual Language architecture natively processes text, images, video, and audio as unified inputs for coherent multimodal output.
Accepts up to 10 reference images for subject appearance, style, and composition anchoring across a multi-shot sequence.
Handles fast, intricate physical actions — martial arts, dance, sports — with consistent body mechanics and no ghosting artifacts.
## Specifications
| Feature | Details |
| ------------------------ | ----------------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Base credits** | 300 |
| **Resolution** | 1080p |
| **Frame rate** | 60 FPS |
| **Duration** | Up to 15 seconds |
| **Shots per generation** | Up to 6 |
| **Audio** | Omni Native Audio — dialogue, SFX, music |
| **Languages** | English, Japanese, Korean, Spanish + more |
| **Max reference images** | 10 |
| **Architecture** | Multi-modal Visual Language (MVL) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling 3.0 Pro** from the model dropdown.
For multi-shot output, describe each shot with explicit transitions: "SHOT 1 (3s, wide, establishing): ... SHOT 2 (2s, close-up): ..." Kling 3.0 Pro interprets these cues to generate distinct cinematographic cuts.
Upload up to 10 reference images for character appearance, environment style, or composition guidance.
Describe the audio landscape — dialogue lines, ambient environment, music style — within the prompt for Omni Native Audio.
Click **Generate**. Kling 3.0 Pro produces a 1080p, 60 FPS output with synchronized audio.
## Prompting tips
* **Structure shots explicitly** — "SHOT 1: wide establishing exterior, 3 seconds, slow pan right. SHOT 2: medium close-up on protagonist, 2 seconds, static camera." Kling 3.0 Pro follows cinematographic structure in prompts.
* **Specify language for dialogue** — If your scene requires characters speaking a specific language, state it clearly: "The character speaks in Japanese with a formal tone."
* **Reference images anchor identity** — For character consistency across shots, upload a reference image and describe the character consistently in each shot description.
* **Use technical camera terms** — "Shallow depth of field," "Dutch angle," "rack focus," and "tracking shot" all meaningfully influence the cinematic output.
### Example prompts
> SHOT 1 (4s, wide, cinematic): A samurai stands at the edge of a misty forest at dawn. Slow pan left, revealing a village in the distance. Traditional Japanese ambient sounds. SHOT 2 (3s, close-up): The samurai's hand grips a sword hilt. Rain begins to fall. SHOT 3 (3s, medium): The samurai turns and walks into the mist.
> A professional basketball player dribbles through defenders and dunks. Wide angle, 60 FPS, 5 seconds. Arena crowd roaring in the background, sneakers squeaking on hardwood.
## Compare models
| Model | Resolution | FPS | Audio | Shots | Best for |
| ----------------------------------------------- | ---------- | --- | ----------- | ------- | ------------------------------------ |
| **Kling 3.0 Pro** | 1080p | 60 | Omni Native | Up to 6 | Multi-shot storytelling, 60 FPS |
| [Kling O3](/ai-models/video/kling-o3) | 4K | 60 | Yes | Up to 6 | Advanced physics, 6 generation modes |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | 1080p | 48 | Lip-sync | — | Audio-synced content, fast motion |
| [Kling 2.5 Pro](/ai-models/video/kling-2-5-pro) | 1080p | — | No | — | Cost-efficient HD production |
Kling 3.0 Pro is the right choice when you need structured multi-shot storytelling at 60 FPS with native audio in a single generation. For 4K output, compare with [Kling O3](/ai-models/video/kling-o3).
# Kling 3 0 turbo pro
Source: https://docs.imagine.art/ai-models/video/kling-3-0-turbo-pro
## Kling 3.0 Turbo Pro
A new "Turbo" tier in the Kling 3.0 lineup, live alongside [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) and [Kling 3.0 4K](/ai-models/video/kling-3-0-4k).
## Specifications
| Feature | Details |
| ----------------- | ------------ |
| **Developer** | Kling AI |
| **Resolution** | 1080p |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Kling 3.0 Turbo Pro**.
Write your prompt and generate as normal.
## Kling 3.0 family comparison
| Model | Resolution | Duration | Audio |
| --------------------------------------------------------- | ---------- | -------- | ----- |
| **Kling 3.0 Turbo Pro** | 1080p | 5–15s | Yes |
| [Kling 3.0 Turbo SD](/ai-models/video/kling-3-0-turbo-sd) | 720p | 5–15s | Yes |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 1080p | 3–15s | Yes |
| [Kling 3.0 4K](/ai-models/video/kling-3-0-4k) | 4K | 3–15s | Yes |
# Kling 3 0 turbo sd
Source: https://docs.imagine.art/ai-models/video/kling-3-0-turbo-sd
## Kling 3.0 Turbo SD
A new, lower-cost "Turbo" tier in the Kling 3.0 lineup, live alongside [Kling 3.0 Turbo Pro](/ai-models/video/kling-3-0-turbo-pro), [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro), and [Kling 3.0 4K](/ai-models/video/kling-3-0-4k).
## Specifications
| Feature | Details |
| ----------------- | ------------ |
| **Developer** | Kling AI |
| **Resolution** | 720p |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Kling 3.0 Turbo SD**.
Write your prompt and generate as normal.
## Kling 3.0 family comparison
| Model | Resolution | Duration | Audio |
| ----------------------------------------------------------- | ---------- | -------- | ----- |
| **Kling 3.0 Turbo SD** | 720p | 5–15s | Yes |
| [Kling 3.0 Turbo Pro](/ai-models/video/kling-3-0-turbo-pro) | 1080p | 5–15s | Yes |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 1080p | 3–15s | Yes |
| [Kling 3.0 4K](/ai-models/video/kling-3-0-4k) | 4K | 3–15s | Yes |
# Kling o1
Source: https://docs.imagine.art/ai-models/video/kling-o1
VIDEO MODEL
by Kling AI
Kling O series
Kling O1
Kling AI's unified video creation and editing model — the world's first multimodal video model to unify generation and editing in a single system. Accepts text, images, keyframes, reference videos, and motion inputs, with up to 7 reference images and 6 camera cuts per generation.
## Unified creation and editing
Kling O1 is the first video model to unify generation and editing in a single system — you can create a new video from scratch and then edit specific sections, restyle footage, extend shots, or swap elements within the same model, without exporting to a separate editing tool.
The Multi-modal Visual Language (MVL) architecture accepts six input types simultaneously: text, images, keyframes, reference videos, motion references, and video editing instructions. This makes Kling O1 uniquely capable for production pipelines that need a single model to handle multiple stages.
## Capabilities
The first model to handle both video creation and video editing in one system — generate footage and edit it within the same generation pipeline.
Accepts text, images, keyframes, reference videos, motion references, and editing instructions as simultaneous inputs.
Anchor character appearance, visual style, and scene composition with up to 7 reference images in a single generation.
Generates up to 6 distinct shots per generation — structured multi-shot output from a single model invocation.
Transform the visual style of existing footage — apply new aesthetics, change time of day, or retheme content while preserving the underlying motion.
Extend existing shots seamlessly — continue the motion and scene from the end of an existing clip.
## Input types supported
| Input | Use |
| ------------------------ | ---------------------------------------------------------- |
| **Text** | Scene description, style direction, audio cues |
| **Images (up to 7)** | Subject appearance, visual style, composition anchoring |
| **Keyframes** | Define start, middle, or end frames for transition control |
| **Reference videos** | Motion and style reference from existing footage |
| **Motion references** | Camera trajectory and subject movement patterns |
| **Editing instructions** | Targeted edits to specific elements in existing video |
## Specifications
| Feature | Details |
| -------------------- | --------------------------------------------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Architecture** | Multi-modal Visual Language (MVL) |
| **Resolution** | Up to 1080p |
| **Duration** | 5–10 seconds |
| **Reference images** | Up to 7 |
| **Camera cuts** | Up to 6 per generation |
| **Audio** | No native audio |
| **Input modes** | 6 (text, image, keyframe, ref video, motion ref, editing) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling O1** from the model dropdown.
Select the combination of input types that fits your use case — text only, text + images, keyframes + motion reference, or video editing mode.
Upload up to 7 reference images, a reference video, or motion reference as needed.
For multi-shot output, structure your prompt with explicit shot descriptions — up to 6 shots per generation.
Click **Generate**. Generation typically completes in 1–2 minutes for complex multi-input requests.
## Prompting tips
* **Describe edit targets precisely** — In editing mode: "Change the background from day to night while keeping the subject unchanged" is more accurate than "make it darker."
* **Use keyframes for transitions** — Define your start and end keyframes; let Kling O1 fill in the motion between them consistently.
* **Combine input types** — "Based on this reference image \[image], in this visual style \[image 2], with this camera movement \[motion ref]..." — the MVL architecture processes all inputs cohesively.
### Example prompts
> SHOT 1 (wide, 3s): A detective walks into a rain-soaked alley at night. SHOT 2 (close-up, 2s): Detective looks at a clue on the ground, rain drops visible. SHOT 3 (medium, 3s): Detective turns and exits the alley. Reference image for detective character appearance attached.
> Restyle the provided footage to a vintage 1970s Super 8 film look. Keep all motion and subjects identical; change only the visual aesthetic.
## Compare models
| Model | Edit support | Input types | Camera cuts | Audio | Best for |
| ----------------------------------------------- | ----------------------- | ----------- | ----------- | ----- | ---------------------------- |
| **Kling O1** | Yes (unified) | 6 | Up to 6 | No | Create + edit workflows |
| [Kling O3](/ai-models/video/kling-o3) | Partial | 6 | Up to 6 | Yes | Max capability + audio |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | No | 2 | Up to 6 | Yes | 4K cinematic, multi-shot |
| [Pika 2.2](/ai-models/video/pika-2-2) | Partial (swaps, scenes) | 2 | No | No | Creative effects + keyframes |
Kling O1 is the strongest model when your workflow requires both creating new footage and editing or transforming existing video within the same pipeline. For maximum capability with audio, consider [Kling O3](/ai-models/video/kling-o3).
# Kling o3
Source: https://docs.imagine.art/ai-models/video/kling-o3
VIDEO MODEL
by Kling AI
Kling O series
Kling O3
Kling AI's most advanced video model — 1080p at 60 FPS, advanced physics engine simulating gravity, collision, and fluid dynamics, 6 generation modes, and up to 6 distinct shots in a single generation.
## Six modes, one model
Kling O3 is Kling AI's most capable unified model — designed to handle every stage of a production workflow in a single system. Six distinct generation modes (text-to-video, image-to-video, video-to-video, frames-to-video, motion control, and reference-to-video) eliminate the need to switch between models for different tasks.
The advanced physics engine is the headline capability that separates Kling O3 from earlier models: gravity, balance, deformation, collision, and inertia are all simulated accurately. Characters fall with real weight, water flows with physical plausibility, and rigid objects interact with correct momentum.
## Capabilities
Simulates gravity, balance, deformation, collision, and inertia — objects fall, splash, and interact with physical accuracy rarely seen in generative video.
1080p resolution at 60 frames per second — high-quality output without upscaling.
Text-to-video, image-to-video, video-to-video, frames-to-video, motion control, and reference-to-video — a complete production toolkit in one model.
Generate up to 6 distinct cinematographic shots within a single 15-second output — structured multi-shot storytelling from one prompt.
Accepts more than 10 reference images for character appearance, style anchoring, and multi-subject scene construction.
## Generation modes
| Mode | Description |
| ---------------------- | -------------------------------------------------------- |
| **Text-to-video** | Generate video directly from a text prompt |
| **Image-to-video** | Animate a reference image with described motion |
| **Video-to-video** | Restyle or transform an existing video |
| **Frames-to-video** | Specify start, middle, and/or end frames |
| **Motion control** | Apply specific camera trajectories and subject movements |
| **Reference-to-video** | Anchor generation to reference subjects and styles |
## Specifications
| Feature | Details |
| ------------------------ | ------------------- |
| **Developer** | Kling AI (Kuaishou) |
| **Resolution** | 1080p |
| **Frame rate** | 60 FPS |
| **Duration** | Up to 15 seconds |
| **Shots per generation** | Up to 6 |
| **Generation modes** | 6 |
| **Audio** | No native audio |
| **Max reference images** | 10+ |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Kling O3** from the model dropdown.
Select the mode that fits your workflow — text-to-video, image-to-video, video-to-video, frames-to-video, motion control, or reference-to-video.
For multi-shot output, describe each shot with explicit shot size, camera movement, duration, and scene content.
Upload reference images, video clips, or motion references to anchor characters, styles, or movement patterns.
Click **Generate** for 1080p, 60 FPS output.
## Prompting tips
* **Leverage physics descriptions** — "A glass falls off a table and shatters, water splashing realistically" will be rendered with physically accurate simulation. Be explicit about the physical behavior you want.
* **Use mode strategically** — For restyling existing footage, use video-to-video. For maximum control over a sequence, use frames-to-video with defined start and end frames.
* **Multi-shot structure** — "SHOT 1 (wide, 3s): ... SHOT 2 (close-up, 2s): ..." is interpreted as discrete cinematographic cuts by Kling O3.
* **Use references for character identity** — Upload multiple reference angles of a character to improve identity consistency across shots.
### Example prompts
> A professional boxer trains in a gym. SHOT 1 (4s, wide): Boxer shadowboxes in the ring, motion blur on fast punches. SHOT 2 (3s, close-up): Sweat flies off the boxer's face as they throw a right hook. SHOT 3 (3s, medium): The boxer catches their breath, hands on the rope. 1080p, 60 FPS.
> A raindrop falls onto a still pond surface. Extreme close-up macro shot. Water ripples expand outward with physically accurate fluid dynamics. Slow motion, 60 FPS, 5 seconds.
## Compare models
| Model | Resolution | FPS | Physics | Modes | Best for |
| ----------------------------------------------- | ---------- | --- | -------- | ------------ | ------------------------------ |
| **Kling O3** | 4K | 60 | Advanced | 6 | Max capability, physics-heavy |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | 4K | 60 | Standard | Text + Image | Cinematic 4K, multi-shot |
| [Kling O1](/ai-models/video/kling-o1) | 1080p | — | Standard | Unified | Editing + generation workflows |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | 1080p | 48 | Standard | Text + Image | Audio-synced, fast motion |
Kling O3 is the right choice when physical realism — fluid dynamics, object collisions, realistic falls — is central to your scene. For 4K cinematic output without the physics complexity, [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) covers most professional needs.
# Ltx 2 3
Source: https://docs.imagine.art/ai-models/video/ltx-2-3
## LTX 2.3
A new video model from Lightricks — a new provider in ImagineArt's video lineup. It has the highest resolution ceiling of any video model currently offered (up to 2160p/4K).
## Specifications
| Feature | Details |
| ----------------- | ------------ |
| **Developer** | Lightricks |
| **Resolution** | 1080p–2160p |
| **Duration** | 6–10 seconds |
| **Audio** | Yes |
| **Frame control** | Start frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **LTX 2.3**.
Write your prompt and generate as normal.
# Lucy
Source: https://docs.imagine.art/ai-models/video/lucy
VIDEO MODEL
by Decart
14B
Lucy
Decart's fast image-to-video model — animate any still image with physics-aware motion, cinematic quality, and lightning-fast generation. Lucy transforms a single frame and a text prompt into a 10-second 720p video clip with consistent motion and no visual artifacts.
## Fast image-to-video animation
Lucy is Decart's 14-billion parameter image-to-video model, built on a diffusion-transformer architecture that delivers cinematic motion from a single still image. You provide a start frame — a photo, an illustration, a rendered scene — write a prompt describing the motion, and Lucy generates a smooth 720p clip in seconds.
What sets Lucy apart is its physics-aware generation. The model learns the structure of the world implicitly — understanding how fabric drapes, how liquids move, how surfaces react to light — without relying on depth maps, green screens, or 3D meshes. The result is motion that looks natural rather than interpolated: subjects move as they would in the real world, and the visual quality holds consistently frame to frame without flickering or drift.
Lucy is the fastest option on ImagineArt for creators who need high-quality animated clips from existing imagery — product shots, character art, concept illustrations, or photos — without the overhead of higher-credit models.
## Capabilities
Upload any still image as the starting frame — photo, illustration, or rendered scene — and Lucy animates it with natural, physics-consistent motion.
The model understands world structure implicitly: fabric moves like fabric, liquids flow realistically, and surfaces respond to light correctly — no 3D rigs or depth data required.
Optimized for low-latency inference — Lucy generates clips significantly faster than comparable quality models, making iteration quick and affordable.
Describe the motion, camera angle, atmosphere, and style in natural language. Lucy follows complex multi-part instructions and applies them to the starting frame.
Frame-to-frame consistency is maintained throughout the clip — eliminating the flickering, morphing, and temporal artifacts common in lower-quality video models.
Accepts JPG, JPEG, PNG, WebP, GIF, and AVIF input files — compatible with the full range of image formats used in creative and production workflows.
## Specifications
| Feature | Details |
| -------------------- | ------------------------------- |
| **Developer** | Decart |
| **Model size** | 14B parameters |
| **Resolution** | 720p |
| **Aspect ratios** | 16:9, 9:16 |
| **Input** | Start frame (image-to-video) |
| **Accepted formats** | JPG, JPEG, PNG, WebP, GIF, AVIF |
| **Output format** | MP4 (H.264) |
| **Base credits** | 240 |
| **Audio** | No native audio generation |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Lucy** from the model dropdown.
Upload the image you want to animate. Lucy uses this as the first frame of the generated clip. Supported formats: JPG, JPEG, PNG, WebP, GIF, AVIF.
Describe what you want to happen in the clip — the motion, camera movement, atmosphere, and any stylistic direction. Be specific about how things should move.
Choose **16:9** for landscape/widescreen or **9:16** for vertical/portrait content.
Click **Generate**. Lucy animates your start frame and delivers a 720p MP4 clip.
## Prompting tips
* **Describe motion explicitly** — Lucy needs to know what moves and how. "The camera slowly pushes in," "the subject turns their head to look left," "leaves drift downward" are all more actionable than vague scene descriptions.
* **Reference the starting image indirectly** — Lucy already knows what the scene looks like from your start frame. Focus your prompt on motion, camera behavior, and atmosphere rather than restating visual elements.
* **Use physics language** — Phrases like "fabric ripples in the breeze," "water surface shimmers," or "steam rises from the cup" take advantage of Lucy's physics-aware generation to produce natural results.
* **Aspect ratio determines framing** — Choose 16:9 for landscape subjects (scenes, cityscapes, wide shots) and 9:16 for portrait subjects (people, vertical compositions, social media delivery).
* **Keep motion achievable in 10 seconds** — A single, clear action or camera move tends to produce better results than a complex sequence. Save multi-shot narratives for models like Seedance or Wan.
### Example prompts
> The subject slowly turns toward the camera, hair catching a gentle breeze. Soft afternoon light. Natural movement, cinematic. 16:9.
> A steaming mug of coffee sits on a wooden desk. Gentle wisps of steam rise and curl. Shallow depth of field, warm tones, static camera. 16:9.
> A city street at night. Rain begins to fall softly, droplets catching the neon reflections on wet pavement. The camera holds still. 16:9.
> A fashion model stands against a white backdrop. Fabric of the dress moves gently as if caught in a slow breeze. 9:16, vertical format.
## Compare models
| Model | Input | Resolution | Speed | Best for |
| ----------------------------------------------- | ------------- | ----------- | -------- | --------------------------------------- |
| **Lucy** | Start frame | 720p | Fast | Quick image animation, social content |
| [Luma Ray 2](/ai-models/video/luma-ray-2) | Text or image | Up to 1080p | Moderate | Photorealistic textures, natural motion |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | Text or image | 1080p | Moderate | Cinematic quality, audio sync |
| [Seedance 2](/ai-models/video/seedance-2) | Text or image | 1080p | Moderate | High-fidelity multi-modal generation |
| [Runway 4.5](/ai-models/video/runway-4-5) | Text or image | 720p | Moderate | Camera-precise cinematic output |
Lucy is the best choice when you already have a strong still image and want to animate it quickly without high credit cost. For productions that need higher resolution, audio, or complex multi-shot control, consider stepping up to Kling, Seedance, or a Wan model.
# Luma ray 2
Source: https://docs.imagine.art/ai-models/video/luma-ray-2
VIDEO MODEL
by Luma AI
Dream Machine Ray 2
Luma Ray 2
Luma AI's most powerful video model — a multimodal architecture with 10× the compute power of Ray 1. Ray 2 delivers fast, coherent motion, ultra-realistic texture detail, and logically sequenced events that make AI-generated footage look genuinely natural.
## Photorealism at scale
Luma Ray 2 is a ground-up rebuild of Luma AI's Dream Machine video generation platform. With 10× the compute of Ray 1, the model achieves a level of photorealism — realistic texture rendering, accurate material properties, and physically plausible motion — that makes it the strongest option for footage that needs to pass as real-world video.
The model excels at natural motion: human movement flows without the jitter or drift common in earlier video models, and event sequences follow a logical order rather than hallucinating random intermediate states. Both text-to-video and image-to-video workflows are supported.
## Capabilities
Material surfaces — skin, fabric, metal, stone, liquid — render with photorealistic fidelity, accurate to real-world lighting and surface properties.
Movement is fast and coherent, following the logical physical sequence of events — no random mid-motion artifacts or unnatural interpolation.
A multimodal architecture with 10× the compute capacity of its predecessor — noticeably sharper outputs and more detailed scene rendering.
Animate any still image with natural, physics-plausible motion — the strongest image-animation workflow in Luma's lineup.
Native 24 FPS output for the standard cinematic frame rate — consistent with professional film and commercial production standards.
Supports 16:9, 9:16, 4:3, 3:4, 21:9, and 9:21 for flexible delivery across platforms.
## Specifications
| Feature | Details |
| ----------------- | -------------------------------- |
| **Developer** | Luma AI |
| **Resolution** | 540p–720p |
| **Duration** | 5–9 seconds |
| **Frame rate** | 24 FPS |
| **Aspect ratios** | 16:9, 9:16, 4:3, 3:4, 21:9, 9:21 |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
| **Compute** | 10× Ray 1 |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Luma Ray 2** from the model dropdown.
For text-to-video, write a detailed prompt. For image-to-video, upload your reference image and describe the motion.
Select your clip length (5–9 seconds) and aspect ratio.
Click **Generate** and review the photorealistic output.
## Prompting tips
* **Describe the real-world material properties** — "Silk fabric catching the breeze," "rain-wet cobblestones under warm streetlights," "frost forming on a glass surface" — Ray 2 renders material physics convincingly.
* **Specify motion type** — "A slow, controlled pan," "subject walks naturally from left to right, maintaining eye contact with camera."
* **Image-to-video: upload high-resolution references** — Ray 2 preserves fine detail from the input image; a higher quality starting frame produces a higher quality animation.
* **Use 16:9 for cinematic output** — The 24 FPS cinematic standard and 16:9 ratio combination produces the most professional-looking results for film and commercial use.
### Example prompts
> A slow close-up of honey dripping from a wooden spoon into a glass jar. Warm backlighting, macro lens, photorealistic golden tones. 5 seconds, 16:9.
> A fashion model walks toward the camera on a rain-slicked Parisian street at night. Neon reflections in puddles, natural movement, cinematic. 8 seconds.
## Compare models
| Model | Photorealism | Motion quality | Audio | Best for |
| ----------------------------------------------- | ------------ | ------------------- | ----- | -------------------------------------- |
| **Luma Ray 2** | Excellent | Coherent, natural | No | Photorealistic footage, natural motion |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | Very high | Physics-aware | Yes | Physics + audio integration |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | High | Cinematic | Yes | Audio-synced cinematic |
| [Runway 4.5](/ai-models/video/runway-4-5) | High | Precise, controlled | No | Camera-controlled cinematic |
Luma Ray 2 is the best choice when photorealistic texture and material rendering is the top priority. For productions that also need synchronized audio, pair the visual output with a model like [Wan 2.5](/ai-models/video/wan-2-5) or add audio in post-production.
# Minimax hailuo h3
Source: https://docs.imagine.art/ai-models/video/minimax-hailuo-h3
## MiniMax Hailuo H3
MiniMax's newest video model, sitting above the existing [Hailuo 02](/ai-models/video/hailuo-02-pro) and [Hailuo 2.3](/ai-models/video/hailuo-2-3-pro) lines. In-app it's described as "MiniMax's frontier 2K video model with subject reference and camera direction." Requested by name in the team's Slack and confirmed shipped the same day.
## Specifications
| Feature | Details |
| ----------------- | --------------- |
| **Developer** | MiniMax |
| **Resolution** | 2K |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **MiniMax Hailuo H3**.
Write your prompt and generate as normal.
## Hailuo family comparison
| Model | Resolution | Duration | Audio |
| --------------------------------------------------------------- | ---------- | -------- | ----- |
| **MiniMax Hailuo H3** | 2K | 5–15s | Yes |
| [MiniMax Hailuo H3 Max](/ai-models/video/minimax-hailuo-h3-max) | 480p–768p | 5–15s | Yes |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | 1080p | 6s | No |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | 1080p | 6s | No |
# Minimax hailuo h3 max
Source: https://docs.imagine.art/ai-models/video/minimax-hailuo-h3-max
## MiniMax Hailuo H3 Max
The larger variant of [MiniMax Hailuo H3](/ai-models/video/minimax-hailuo-h3), MiniMax's newest video model — sitting above the existing [Hailuo 02](/ai-models/video/hailuo-02-pro) and [Hailuo 2.3](/ai-models/video/hailuo-2-3-pro) lines.
## Specifications
| Feature | Details |
| ----------------- | --------------- |
| **Developer** | MiniMax |
| **Resolution** | 480p–768p |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **MiniMax Hailuo H3 Max**.
Write your prompt and generate as normal.
## Hailuo family comparison
| Model | Resolution | Duration | Audio |
| ------------------------------------------------------- | ---------- | -------- | ----- |
| **MiniMax Hailuo H3 Max** | 480p–768p | 5–15s | Yes |
| [MiniMax Hailuo H3](/ai-models/video/minimax-hailuo-h3) | 2K | 5–15s | Yes |
| [Hailuo 02 Pro](/ai-models/video/hailuo-02-pro) | 1080p | 6s | No |
| [Hailuo 2.3 Pro](/ai-models/video/hailuo-2-3-pro) | 1080p | 6s | No |
# Pika 2 2
Source: https://docs.imagine.art/ai-models/video/pika-2-2
VIDEO MODEL
by Pika Labs
Pika 2.2
Pika Labs' quality-first video model — Pikaframes keyframe control for precise start and end frame specification, native 1080p output, 10-second clips, and a suite of scene manipulation tools including Pikascenes, Pikaffects, Pikadditions, and Pikaswaps.
Duration
Up to 10 seconds
Keyframe control
Pikaframes
## Quality over speed
Pika 2.2 is Pika Labs' most refined video generation model — built around the Pikaframes system, which allows you to define both the starting and ending frames of a 1–10 second transition, with the model generating the motion between them. This makes Pika 2.2 especially powerful for controlled, intentional video sequences where the visual result matters more than generation speed.
Beyond Pikaframes, Pika 2.2 includes a full toolkit of scene manipulation tools — Pikaffects (dynamic visual effects), Pikascenes (environment transformations), Pikadditions (adding elements to scenes), and Pikaswaps (swapping objects or subjects) — making it a versatile choice for creative production and marketing content.
## Capabilities
Specify the starting frame and ending frame of a transition — Pika 2.2 generates the motion between them with high precision and visual quality.
Generates at full 1080p resolution without upscaling — production-ready quality for commercial and social deliverables.
Apply dynamic visual effects to scenes — explosions, weather, magical transformations, and stylized visual treatments.
Transform scene environments — change settings, backgrounds, time of day, and atmosphere while keeping the subject consistent.
Add new elements, objects, or characters to an existing scene with natural integration and consistent lighting.
Swap objects, subjects, or materials in a scene — replace one element with another while maintaining visual coherence.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------------------------- |
| **Developer** | Pika Labs |
| **Resolution** | 1080p (native) |
| **Duration** | Up to 10 seconds |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:5, 5:4, 3:2, 2:3 |
| **Keyframe control** | Pikaframes (start + end frame) |
| **Audio** | No native audio generation |
| **Scene tools** | Pikaffects, Pikascenes, Pikadditions, Pikaswaps |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Pika 2.2** from the model dropdown.
Describe the scene, motion, subject, style, and atmosphere in your text prompt.
Choose your clip length (up to 10 seconds) and select from 7 available aspect ratios.
Click **Generate** and review the 1080p output.
Open the **AI Video Generator**, select **Pika 2.2**, and enable the **Pikaframes** option.
Upload the image you want as the first frame and the image you want as the last frame of the clip.
Choose the transition duration between 1 and 10 seconds.
Add a descriptive prompt to guide the motion and style of the transition.
Click **Generate**. Pika 2.2 creates the motion between your defined start and end frames.
## Prompting tips
* **For Pikaframes, let the frames do the work** — Your start and end images already define the visual range; the prompt should focus on the motion style and atmosphere, not restating what's in the frames.
* **Describe motion dynamics** — "Slow, graceful drift," "snappy cut with energy," or "smooth continuous movement" guide the pacing.
* **Use Pikaffects for dramatic moments** — Effects like fire, lightning, water, or explosions are added naturally when referenced explicitly in the prompt.
* **Match aspect ratio to delivery** — 16:9 for cinema/YouTube, 9:16 for vertical social, 1:1 for feeds, 4:5 for Instagram.
### Example prompts
> A butterfly rests on a flower petal in golden afternoon light. A gentle breeze causes the petals to sway softly. Macro lens, shallow depth of field. 5 seconds, 16:9.
> A glass of water shatters in slow motion, water droplets frozen mid-air, black background, cinematic high-speed photography style. 3 seconds, 1:1.
## Compare models
| Model | Resolution | Keyframe control | Audio | Best for |
| ----------------------------------------------- | ---------- | ---------------- | ----- | ------------------------------------- |
| **Pika 2.2** | 1080p | Yes (Pikaframes) | No | Controlled transitions, quality-first |
| [PixVerse v6](/ai-models/video/pixverse-v6) | 1080p | No | Yes | Audio-visual, cinematic lens control |
| [PixVerse v5.5](/ai-models/video/pixverse-v5-5) | 1080p | No | Yes | Script-first multi-shot |
| [Runway 4.5](/ai-models/video/runway-4-5) | 720p | No | No | Camera-precise cinematic output |
Pika 2.2 is the best choice when you need to control exactly what the video starts and ends on. The Pikaframes system gives you a level of deterministic visual control that most generative video models don't offer.
# Pixverse v5
Source: https://docs.imagine.art/ai-models/video/pixverse-v5
VIDEO MODEL
by PixVerse
PixVerse v5 family
PixVerse v5
PixVerse's fast character animation model — fast generation in approximately 30 seconds, excellent performance on complex character movements including gymnastics, parkour, and martial arts, strong anime and game character consistency, and 15+ creative visual effects.
Generation time
\~30 seconds
## Fast generation with strong character animation
PixVerse v5 generates video with a generation time of approximately 30 seconds. This speed-quality combination made it one of the most practical choices for commercial production on launch.
The model's standout capability is complex character movement: gymnastics, parkour, martial arts, and dance sequences all render with accurate body mechanics, smooth motion transitions, and no major anatomical artifacts. For anime and game character content, v5 maintains strong visual consistency across frames.
## Capabilities
Gymnastics, parkour, martial arts, and dance render with accurate body mechanics and smooth transitions — a category-leading capability in PixVerse v5.
Strong frame-to-frame visual consistency for anime and game character styles — identity and style maintained throughout the clip.
Approximately 30 seconds per clip — rapid turnaround without sacrificing output quality.
A library of over 15 stylized visual effects — fire, explosions, magic, glitch, neon glow, and more — applied as native generation properties.
Focused generation window for punchy, character-driven content.
Supports 16:9, 9:16, 1:1, and other ratios for flexible platform delivery.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------- |
| **Developer** | PixVerse |
| **Resolution** | 540p–720p |
| **Duration** | 5–8 seconds |
| **Generation speed** | \~30 seconds |
| **Visual effects** | 15+ |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
| **Aspect ratios** | 16:9, 9:16, 1:1, and more |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **PixVerse v5** from the model dropdown.
For character movement content, describe the specific action type, character style, and environment in detail. For effects content, name the visual effect explicitly in the prompt.
Choose your clip length (5–8 seconds) and target aspect ratio.
Click **Generate** — 1080p output is typically ready in approximately 30 seconds.
## Prompting tips
* **Name complex movements explicitly** — "Backflip," "roundhouse kick," "breakdance freezes" — PixVerse v5 has strong training on complex human movement vocabulary.
* **Specify anime or game style** — "Anime-style," "game-CG style," "hand-drawn animation" all activate the model's stylized character rendering.
* **Name visual effects directly** — "With a magical energy burst effect," "neon glow trails," "fire explosion" — effects are applied more accurately when named explicitly rather than described abstractly.
* **Use 9:16 for character close-ups** — Vertical format is especially effective for character-focused content and social media delivery.
### Example prompts
> A martial artist performs a spinning hook kick in a dojo. Slow motion, cinematic close-up on the impact. Clean background, dramatic lighting. 8 seconds, 16:9.
> Anime-style heroine launches into the air, unleashing a magical energy burst. Dynamic action pose, bright particle effects, stylized motion blur. 6 seconds, 9:16.
## Compare models
| Model | Character movement | Anime/game | Audio | Speed | Best for |
| ----------------------------------------------- | ------------------ | ---------- | ----- | -------- | ------------------------------------- |
| **PixVerse v5** | Excellent | Strong | No | \~30s | Complex movement, character animation |
| [PixVerse v5.5](/ai-models/video/pixverse-v5-5) | Strong | Strong | Yes | \~30s | Multi-shot, audio-synced |
| [PixVerse v6](/ai-models/video/pixverse-v6) | Very strong | Strong | Yes | Standard | Cinematic lens control, A/V |
| [Kling O3](/ai-models/video/kling-o3) | Excellent | Standard | Yes | Slower | Physics + complex action, 4K |
PixVerse v5 is the best entry point in the PixVerse family for complex character movement and anime-style content. For audio-synchronized multi-shot narratives, step up to [PixVerse v5.5](/ai-models/video/pixverse-v5-5) or [PixVerse v6](/ai-models/video/pixverse-v6).
# Pixverse v5 5
Source: https://docs.imagine.art/ai-models/video/pixverse-v5-5
VIDEO MODEL
by PixVerse
PixVerse v5 family
PixVerse v5.5
PixVerse's audio-enabled multi-shot model — native audio generation with accurate A/V sync and automatic lip-sync, script-first content creation where a single sentence is broken into structured shots with voiceover and ambient sound, and output in approximately 30 seconds.
Audio
Native A/V + Lip-sync
Generation time
\~30 seconds
## Script-first video creation
PixVerse v5.5 is the audio-enabled evolution of the v5 architecture — the same core generation quality and speed, now with native audio-video synchronization and a script-first workflow. Type a sentence, and v5.5 automatically breaks it into structured shots, adds voiceover, and layers ambient sound. The result is complete, production-ready content from a minimal text input.
The automatic lip-sync system animates character mouths in sync with the generated voiceover, making v5.5 well-suited for narrative content, character-driven clips, and social media storytelling without separate audio post-production.
## Capabilities
Type a single sentence or paragraph — v5.5 automatically structures it into shots, adds voiceover narration, and generates synchronized ambient sound.
Audio and video generated simultaneously with accurate A/V synchronization — dialogue, ambient sounds, and voiceover all timed to the visual content.
Characters' lip movements are automatically synchronized to the generated voiceover — no manual lip-sync post-processing needed.
Generates structured multi-shot sequences from narrative prompts — scene cuts, transitions, and story beats handled automatically.
Generation in approximately 30 seconds — same speed advantage as PixVerse v5 with the addition of audio.
Maintains subject and visual style consistency across shots — strong for recurring characters in multi-shot sequences.
## Specifications
| Feature | Details |
| -------------------- | ------------------------------------------ |
| **Developer** | PixVerse |
| **Resolution** | 540p–1080p |
| **Duration** | 5–8 seconds |
| **Generation speed** | \~30 seconds at 1080p |
| **Audio** | Native — voiceover, SFX, ambient |
| **Lip-sync** | Automatic |
| **Multi-shot** | Yes |
| **Architecture** | Diffusion backbone with Transformer layers |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **PixVerse v5.5** from the model dropdown.
Write a sentence or paragraph describing your story — v5.5 will break it into shots automatically with voiceover and ambient sound.
For more control, use "SHOT 1: ... SHOT 2: ..." structure with explicit scene, audio, and camera descriptions per shot.
Click **Generate** for output with synchronized audio in approximately 30 seconds.
## Prompting tips
* **The script-first approach works well for narrated content** — "A documentary about deep-sea creatures begins with a wide shot of the ocean surface. Narrator says: 'Beneath the waves lies a world unseen.'" produces a complete narrated clip.
* **Name audio elements explicitly for ambient control** — "Quiet jazz playing in the background," "rain pattering on the roof" — ambient audio follows explicit cues.
* **Use character references for consistent lip-sync** — Upload a character reference image for more accurate and consistent lip animation across the clip.
### Example prompts
> A travel documentary opens in Tokyo at night. Wide shot of neon-lit streets. Narrator voice: "Tokyo never sleeps." CUT TO medium shot of street food vendor preparing ramen. Ambient street sounds. 10 seconds.
> A product advertisement: SHOT 1 — a skincare bottle on a marble surface, dramatic lighting. SHOT 2 — close-up of product label. Voiceover: "Natural ingredients. Visible results." Soft background music. 8 seconds.
## Compare models
| Model | Audio | Lip-sync | Multi-shot | Speed | Best for |
| ------------------------------------------- | ----- | -------- | ---------- | -------- | ----------------------------- |
| **PixVerse v5.5** | Yes | Auto | Yes | \~30s | Script-first narrated content |
| [PixVerse v5](/ai-models/video/pixverse-v5) | No | No | No | \~30s | Character animation, effects |
| [PixVerse v6](/ai-models/video/pixverse-v6) | Yes | Yes | Yes | Standard | Cinematic lens control, A/V |
| [Wan 2.5](/ai-models/video/wan-2-5) | Yes | Yes | No | Standard | Flexible A/V production |
PixVerse v5.5 is the fastest path from a text idea to a complete video with narration and ambient sound. For precise optical control and longer clips, use [PixVerse v6](/ai-models/video/pixverse-v6).
# Pixverse v6
Source: https://docs.imagine.art/ai-models/video/pixverse-v6
VIDEO MODEL
by PixVerse
Released March 2026
PixVerse v6
PixVerse's flagship model — 20+ cinematic lens controls including focal length, aperture, and chromatic aberration, multi-shot storytelling engine, up to 1080p at 5–10 seconds, and multilingual text rendered within frames.
## The most lens-controlled AI video model
PixVerse v6, released March 30, 2026, is PixVerse's most comprehensive video generation model. The headline feature is its 20+ cinematic lens control system — parameters including focal length, aperture, depth of field, lens distortion, chromatic aberration, and vignetting that are typically exclusive to real camera setups. Combined with a multi-shot storytelling engine, v6 bridges the gap between AI generation and professional cinematography.
Multilingual text rendering within frames (subtitles, labels, titles) enables localized global content production without a separate post-production step.
## Capabilities
Focal length, aperture, depth of field, lens distortion, chromatic aberration, vignetting, and more — precise optical simulation for cinematic results.
Structured multi-shot storytelling within a single generation — scene cuts, transitions, and narrative arcs handled automatically from your prompt.
Renders accurate text in multiple languages directly within the video frame — titles, subtitles, labels, and signage in your target locale.
Up to 1080p output for 5–10 second clips — flexible delivery across platforms.
Supports video extension and scene transition generation — seamlessly continue or connect scenes without re-generating from scratch.
## Specifications
| Feature | Details |
| ----------------- | ------------------------------------------------------------- |
| **Developer** | PixVerse |
| **Released** | March 30, 2026 |
| **Resolution** | 540p–1080p |
| **Duration** | 5–10 seconds |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:3, 3:4 |
| **Audio** | No native audio |
| **Lens controls** | 20+ (focal length, aperture, DoF, distortion, CA, vignetting) |
| **Text in-frame** | Yes, multilingual |
| **Multi-shot** | Yes |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **PixVerse v6** from the model dropdown.
Include cinematic language — lens type, aperture mood, audio cues, and scene progression. v6 interprets optical and cinematographic vocabulary directly.
Use advanced settings to tune specific optical parameters — focal length for compression/expansion, aperture for depth of field, vignetting for atmosphere.
Click **Generate**. PixVerse v6 produces a video with the specified optical characteristics.
## Prompting tips
* **Use optical terminology** — "Shot on an 85mm lens with a wide-open aperture, shallow depth of field" will be interpreted and applied as actual optical rendering properties.
* **Use descriptive scene details** — "With the visual mood of waves crashing and seagulls" helps set the visual atmosphere of the scene.
* **Leverage chromatic aberration for mood** — A subtle chromatic aberration setting adds a cinematic, slightly analog feel to otherwise perfect digital footage.
* **Multi-shot: use transition cues** — "CUT TO:" or "THEN:" in prompts cue the multi-shot engine to generate discrete scene transitions.
### Example prompts
> A couple walks on a sunset beach, shot on a 35mm lens, f/1.8, golden bokeh in the background. Slight vignette. Cinematic, 10 seconds.
> A product launch promo: SHOT 1 — smartphone spinning on a white surface, dramatic lighting. CUT TO SHOT 2 — close-up of screen with "NEW ERA" text in bold. 10 seconds, 16:9.
## Compare models
| Model | Audio | Lens controls | Multi-shot | Duration | Best for |
| ----------------------------------------------- | ----- | ------------- | ------------- | -------- | ------------------------------- |
| **PixVerse v6** | Yes | 20+ | Yes | 15s | Cinematic optical control, A/V |
| [PixVerse v5.5](/ai-models/video/pixverse-v5-5) | Yes | Limited | Yes | 10s | Script-first, multi-shot |
| [PixVerse v5](/ai-models/video/pixverse-v5) | No | None | No | 15s | Fast 1080p, character animation |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Yes | No | Up to 6 shots | 15s | 4K cinematic storytelling |
PixVerse v6 is the top choice when cinematographic lens behavior matters — bokeh quality, optical distortion, and depth of field are rendered as optical simulation, not post-processing effects.
# Runway 4 5
Source: https://docs.imagine.art/ai-models/video/runway-4-5
VIDEO MODEL
by Runway
Runway 4.5
Runway's flagship full-quality video generation model — precise camera controls, consistent subject rendering, and cinematic motion for professional productions. Transformer-based architecture with strong spatial awareness and prompt adherence.
## Runway's full-quality model
Runway 4.5 is Runway's flagship video generation model, built on a Transformer-based architecture that prioritizes spatial coherence, subject consistency, and cinematographic control. Where [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) is optimized for speed, Runway 4.5 delivers maximum quality — finer detail, more stable motion, and stronger adherence to complex camera and subject instructions.
The model handles both text-to-video and image-to-video workflows, with standout performance on controlled camera movements, consistent character behavior across a clip, and physically accurate scene composition.
## Capabilities
Responsive to cinematographic prompts — dolly shots, pans, tilts, tracking shots, and static frames all render reliably with the specified movement.
Maintains consistent appearance of characters and objects throughout a clip, reducing drift or morphing common in earlier video models.
Generates high-quality video directly from detailed text prompts, with strong spatial and narrative coherence.
Animates a static reference image with specified motion, lighting changes, or camera movement.
Supports 16:9, 9:16, 1:1, 4:3, 3:4, and 21:9 for flexible delivery across platforms and formats.
Transformer architecture delivers smooth, physically plausible motion — natural inertia, realistic object interaction, and stable scene composition.
## Specifications
| Feature | Details |
| ----------------- | ------------------------------- |
| **Developer** | Runway |
| **Resolution** | 720p (4K upscaling available) |
| **Duration** | 5–10 seconds |
| **Frame rate** | 24 FPS |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:3, 3:4, 21:9 |
| **Input modes** | Text-to-video, image-to-video |
| **Audio** | No native audio |
| **Architecture** | Transformer-based |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Runway 4.5** from the model dropdown.
Describe the scene with clear camera direction, subject behavior, and mood. Include cinematographic language for the best results.
Choose your desired duration (5 or 10 seconds) and aspect ratio for your target platform.
Click **Generate** and review the output.
Go to the **AI Video Generator** and select **Runway 4.5**.
Upload the starting frame image you want to animate.
Write a prompt that describes how the image should come to life — camera movement, subject action, environmental changes.
Click **Generate**. Runway 4.5 animates the image with the described motion.
## Prompting tips
* **Use camera terms precisely** — "Slow dolly forward," "handheld close-up tracking the subject," and "static wide shot with subtle camera shake" all produce reliably different cinematic results.
* **Describe subject behavior** — "A woman slowly turns to face the camera," "a car accelerates from a standstill" — specific actions translate to consistent motion.
* **Specify lighting** — "Soft diffused window light," "high-contrast backlit silhouette," or "golden hour side lighting" guide the visual mood.
* **Image-to-video: match your prompt to your image** — The reference image anchors the scene; the prompt should describe the motion and changes, not restate what's already visible.
### Example prompts
> A lone astronaut stands on the surface of Mars, slowly turning to survey a vast red canyon. Cinematic wide shot, dust particles catching the sunlight, no sound. 24 FPS.
> Close-up on a steaming cup of coffee being placed on a marble table. Slow-motion, soft natural window light, shallow depth of field, bokeh background.
## Compare models
| Model | Quality | Speed | Audio | Best for |
| --------------------------------------------------------- | --------------- | --------- | ----- | ------------------------------------------ |
| **Runway 4.5** | Maximum quality | Standard | No | Cinematic productions, final deliverables |
| [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo) | High quality | 5× faster | No | Rapid iteration, cost-efficient production |
| [Sora 2 Pro](/ai-models/video/sora-2-pro) | Cinematic | Standard | Yes | Physics-aware content with audio |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | 1080p cinematic | Standard | Yes | Audio-synced film content |
For fast iterations and drafts, use [Runway Gen 4 Turbo](/ai-models/video/runway-gen-4-turbo). Switch to Runway 4.5 for final-quality renders where the extra generation time is worth the result.
# Runway gen 4 turbo
Source: https://docs.imagine.art/ai-models/video/runway-gen-4-turbo
VIDEO MODEL
by Runway
Released April 2025
Runway Gen 4 Turbo
Runway's speed-optimized video generation model — 5× faster than standard Runway Gen 4, lower credit cost, and approximately 30 seconds per clip. Ideal for rapid iteration, drafting, and image-to-video workflows where turnaround speed matters.
Generation time
\~30 seconds (10s clip)
Runway Gen 4 Turbo is the speed-optimized version of [Runway 4.5](/ai-models/video/runway-4-5). Use Gen 4 Turbo for rapid iteration and drafts; use Runway 4.5 for final-quality production output.
## Speed at every stage
Runway Gen 4 Turbo, released April 2025, is built for iteration speed. At 5× faster than standard Runway Gen 4 with a reduced credit cost, it makes rapid visual exploration practical — generate 10 different directions for a scene in the time it would take to produce 2 with the full-quality model. Supports 5 or 10 second clip lengths.
Image-to-video is a particular strength: uploading a reference image and animating it with a motion prompt produces results in approximately 30 seconds, making Turbo ideal for quick concepting, client previews, and pipeline drafts.
## Capabilities
Generates video 5× faster than standard Runway Gen 4 — enables rapid creative exploration and faster production pipelines.
Excellent image animation performance — strong visual fidelity when animating a reference starting frame with a motion prompt.
5 to 10 seconds per generation — sufficient for most social media clips, brand content, and narrative sequences.
Supports 16:9, 9:16, 1:1, 4:3, 3:4, and 21:9 for all platform formats.
Responsive to cinematographic prompting — camera movements, subject behavior, and scene composition all follow prompt instructions accurately.
Hardware-accelerated Transformer architecture with the same spatial coherence principles as the full-quality model, optimized for speed.
## Specifications
| Feature | Details |
| ------------------- | --------------------------------- |
| **Developer** | Runway |
| **Released** | April 2025 |
| **Resolution** | 720p (4K upscaling available) |
| **Duration** | 5–10 seconds |
| **Frame rate** | Optimized for speed |
| **Speed** | 5× faster than Runway Gen 4 |
| **Generation time** | \~30 seconds for a 10-second clip |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:3, 3:4, 21:9 |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Runway Gen 4 Turbo** from the model dropdown.
Write a text prompt for text-to-video, or upload a reference image for image animation.
Include explicit camera direction and subject movement in your prompt for the most controlled results.
Click **Generate**. Your clip will be ready in approximately 30 seconds.
## Prompting tips
* **Use Turbo for direction exploration** — Generate 5–10 variations of a scene quickly, pick the best direction, then render the final in [Runway 4.5](/ai-models/video/runway-4-5).
* **Image prompts + motion description** — Upload a product image and write "the camera slowly circles clockwise around the object" for fast product video.
* **Keep prompts direct** — For rapid iteration, focused prompts ("wide shot, camera pans right, golden hour lighting") produce faster and more predictable results than long descriptive paragraphs.
### Example prompts
> A car drives along a coastal highway at sunset. Tracking shot from the side, maintaining pace with the car. Warm light, slight motion blur. 10 seconds.
> A coffee shop interior in the morning. Camera slowly drifts forward through the space, customers in background, warm light through windows. Handheld feel.
## Compare models
| Model | Speed | Quality | Duration | Audio | Best for |
| ------------------------------------------------------- | ---------- | ------- | -------- | ----- | --------------------------------- |
| **Runway Gen 4 Turbo** | 5× faster | High | 10s | No | Rapid iteration, draft generation |
| [Runway 4.5](/ai-models/video/runway-4-5) | Standard | Maximum | 5–10s | No | Final production output |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | Fast | High | 5 or 10s | No | Fast Seedance generation |
| [xAI Grok Video](/ai-models/video/grok-video) | Ultra-fast | High | 6 or 10s | Yes | Fastest overall + audio |
Run Runway Gen 4 Turbo first to validate your direction quickly, then switch to [Runway 4.5](/ai-models/video/runway-4-5) for the final render. The Turbo model's output is representative enough to make confident decisions about composition, camera, and pacing.
# Seedance 1 0 pro
Source: https://docs.imagine.art/ai-models/video/seedance-1-0-pro
VIDEO MODEL
by ByteDance
Seedance 1 family
Seedance 1.0 Pro
ByteDance's professional Seedance model — multi-shot storytelling with consistent subjects and visual style, complex camera movements including dolly, zoom, and tracking shots, advanced semantic understanding, and 24 FPS cinematic output up to 1080p.
Camera control
Dolly, zoom, pan, track
## Professional multi-shot storytelling
Seedance 1.0 Pro is the full-quality production model in ByteDance's original Seedance lineup. Its defining characteristic is multi-shot coherence — subjects and visual styles remain consistent across cuts and transitions, making it the most reliable model for narrative video sequences that span multiple scenes.
The camera movement library is comprehensive: dolly shots, zooms, pans, and tracking shots are all supported with cinematic accuracy. Combined with advanced semantic understanding of prompts and 24 FPS output at up to 1080p, Seedance 1.0 Pro is suited for professional narrative content, brand films, and complex storyboards.
## Capabilities
Subject appearance, visual style, and scene coherence are maintained across shots and cuts — essential for narrative video production.
Dolly in/out, zoom, pan, and tracking shots all render with cinematographic accuracy at 24 FPS.
Responds accurately to nuanced prompt instructions — emotional tone, atmospheric details, lighting conditions, and compositional descriptions.
Standard cinematic frame rate for consistent professional-quality footage.
Full HD output for commercial, broadcast, and high-quality digital delivery.
Supports text-to-video and image-to-video workflows with consistent visual quality across both modes.
## Specifications
| Feature | Details |
| -------------------- | ----------------------------- |
| **Developer** | ByteDance |
| **Resolution** | 480p, 720p, 1080p |
| **Frame rate** | 24 FPS |
| **Duration** | 3–12 seconds |
| **Camera movements** | Dolly, zoom, pan, tracking |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Seedance 1.0 Pro** from the model dropdown.
For best results, include the camera movement, subject action, setting details, lighting, and emotional tone. Seedance 1.0 Pro's semantic understanding makes detailed prompts especially effective.
Choose 480p, 720p, or 1080p and a duration between 3 and 12 seconds.
Click **Generate**. For a 5-second clip at 1080p, expect approximately 40 seconds of generation time.
## Prompting tips
* **Specify camera movement explicitly** — "Slow dolly forward toward the subject," "tracking shot following from behind," or "static wide with subtle camera breathing" all produce reliably different results.
* **Include lighting details** — "Backlit by a setting sun," "soft diffused morning light," or "dramatic single-source key light from above" all influence the cinematic feel.
* **Name your multi-shot structure** — "SCENE 1: ... then CUT TO SCENE 2: ..." helps Seedance 1.0 Pro maintain subject consistency across the transition.
### Example prompts
> A documentary filmmaker walks through a dense jungle, camera tracking behind her at shoulder level. Dappled sunlight through the canopy, natural ambient sounds suggested by visuals. Slow dolly forward. 10 seconds, 1080p.
> A chef prepares sushi in a clean, minimalist kitchen. Overhead close-up of hands shaping rice, then CUT TO a medium wide shot of the chef presenting the completed dish. Consistent subject, natural lighting. 10 seconds.
## Seedance family comparison
| Model | Quality | Speed | Duration | Camera control | Best for |
| ------------------------------------------------------- | -------- | -------- | -------- | -------------- | --------------------------------- |
| **Seedance 1.0 Pro** | High | Standard | 5 or 10s | Full | Narrative, cinematic storytelling |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | High | Faster | 5 or 10s | Full | Speed + quality balance |
| [Seedance Lite](/ai-models/video/seedance-lite) | Standard | Fast | 5 or 10s | Basic | Quick clips, social media |
| [Seedance 2](/ai-models/video/seedance-2) | Advanced | Standard | 4–15s | Full | Full multimodal with audio |
Seedance 1.0 Pro is the best Seedance 1 model for professional narrative content and complex camera work. For faster iteration on the same quality level, try [Seedance Pro Fast](/ai-models/video/seedance-pro-fast).
# Seedance 1 5 pro
Source: https://docs.imagine.art/ai-models/video/seedance-1-5-pro
VIDEO MODEL
by ByteDance
Seedance 1 family
Seedance 1.5 Pro
ByteDance's 4.5-billion-parameter video model — millisecond-precision lip-sync, native support for 8+ languages including English, Mandarin, Japanese, Korean, and Spanish, up to 1080p resolution, and 10× faster inference than its predecessor.
## Built for multilingual dialogue and lip-sync
Seedance 1.5 Pro is ByteDance's purpose-built model for dialogue-heavy and multilingual video content. The 4.5-billion-parameter Dual-Branch Diffusion Transformer (DB-DiT) architecture achieves millisecond-precision lip-sync — character mouth movements align exactly with the audio, across 8 languages and regional dialects including English, Mandarin, Japanese, Korean, Spanish, Portuguese, Indonesian, and Cantonese.
At 10× faster inference than its predecessor, Seedance 1.5 Pro is viable for production workflows that require consistent talking-head or dialogue-scene generation at scale.
## Capabilities
Character lip movements align precisely with generated audio at the millisecond level — across 8 languages and regional dialects.
Native dialogue generation in English, Mandarin, Japanese, Korean, Spanish, Portuguese, Indonesian, Cantonese, and Sichuanese.
A 4.5-billion-parameter Dual-Branch Diffusion Transformer — capable of nuanced character expressions, complex scene compositions, and consistent identity.
Full HD output for production-ready talking-head videos, interviews, and dialogue-driven scenes.
Runs 10× faster than the previous generation — practical for batch content creation and localized video production pipelines.
Maintains subject appearance, expression nuance, and visual identity across scenes within a generation.
## Specifications
| Feature | Details |
| ------------------- | ------------------------------------------------------------------------------------------- |
| **Developer** | ByteDance |
| **Parameters** | 4.5 billion |
| **Architecture** | Dual-Branch Diffusion Transformer (DB-DiT) |
| **Resolution** | Up to 1080p |
| **Duration** | 4–12 seconds |
| **Languages** | English, Mandarin, Japanese, Korean, Spanish, Portuguese, Indonesian, Cantonese, Sichuanese |
| **Lip-sync** | Millisecond-precision |
| **Audio** | Native dialogue with lip-sync |
| **Inference speed** | 10× faster than predecessor |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Seedance 1.5 Pro** from the model dropdown.
For talking-head or character dialogue scenes, upload a reference image of the character whose lips you want to animate.
Describe the dialogue scene, specify the language if relevant, and include any visual context — setting, lighting, emotion.
Choose your clip length (up to 12 seconds) and resolution (up to 1080p).
Click **Generate**. Seedance 1.5 Pro produces the video with synchronized dialogue and lip movement.
## Prompting tips
* **Specify the language explicitly** — "A character speaking in formal Japanese" or "conversational Cantonese dialogue" helps the model produce accurate phoneme-to-mouth mapping.
* **Describe emotional tone** — "Excited," "calm and measured," "whispering urgently" all influence both the audio generation and facial expressions.
* **Use a clear reference image** — For best lip-sync accuracy, use a front-facing or slightly angled reference image where the character's mouth is clearly visible.
* **Keep dialogue clips concise** — For maximum coherence, target 5–8 second clips per generation and stitch together longer sequences.
### Example prompts
> A news anchor speaks directly to camera in formal English. Well-lit studio background, professional broadcast style, neutral expression. 8 seconds, 1080p.
> A young woman laughs and responds excitedly in Mandarin during a casual conversation. Warm indoor lighting, natural expressions, slight camera movement. 6 seconds.
## Compare models
| Model | Lip-sync | Languages | Resolution | Best for |
| ----------------------------------------------- | --------------------- | ------------ | ---------- | ----------------------------------- |
| **Seedance 1.5 Pro** | Millisecond precision | 8+ | 1080p | Multilingual dialogue, talking-head |
| [Seedance 2](/ai-models/video/seedance-2) | Native | — | 720p | Multi-reference, full multimodal |
| [Wan 2.5](/ai-models/video/wan-2-5) | Yes | Limited | 1080p | Audio-synced general content |
| [Kling 2.6 Pro](/ai-models/video/kling-2-6-pro) | Yes | EN + Chinese | 1080p | EN/Chinese audio-synced production |
Seedance 1.5 Pro is the strongest model for multilingual dialogue content and precise lip-sync across non-English languages. For full multimodal production with video and audio references, step up to [Seedance 2](/ai-models/video/seedance-2).
# Seedance 2
Source: https://docs.imagine.art/ai-models/video/seedance-2
VIDEO MODEL
by ByteDance
Released February 2026
Seedance 2
ByteDance's most advanced video model — native audio-video joint generation, four generation modes including first-and-last-frame and full references mode, up to 9 reference images, 3 reference videos, and 3 audio clips, with up to 15 seconds of output at 720p–1080p.
References
9 img + 3 vid + 3 audio
Audio
Dialogue + SFX + Music
Seedance 2 is available in a [Fast variant](/ai-models/video/seedance-2-fast) with the same architecture but lower latency — use Fast for rapid iteration and Seedance 2 for maximum quality final renders.
## ByteDance's most capable video model
Seedance 2, released February 10, 2026, is built on the Dual-Branch Diffusion Transformer (DB-DiT) architecture — a significant advancement over the Seedance 1 generation. The model generates audio and video jointly in a single pass, with audio (dialogue, sound effects, music) synchronized at the frame level with the visual output.
The references system is the most expansive in the Seedance lineup: up to 9 reference images, 3 reference video clips, and 3 reference audio clips can be provided simultaneously, giving exhaustive creative control over visual style, character appearance, motion patterns, and audio atmosphere.
## Generation modes
Generate video directly from a text prompt. Describe scene, motion, camera behavior, and audio environment — Seedance 2 generates the complete audio-visual output.
Animate a reference image with described motion. Camera behavior, lighting changes, and audio elements are all added in generation.
Define both the opening and closing frames — Seedance 2 generates the motion, lighting, and audio between them for precise transition control.
Use up to 9 images, 3 video clips, and 3 audio clips as simultaneous references for maximum creative direction over every aspect of the output.
## Capabilities
Audio and video generated in a single pass — dialogue, sound effects, and music synchronized at the frame level without post-processing.
Maintains subject identity, visual style, and scene logic across shots and transitions within a single generation.
9 reference images + 3 reference videos + 3 reference audio clips — the most comprehensive reference input system in the lineup.
Complex camera movements including dolly, zoom, pan, tracking, and crane shots with cinematographic accuracy.
Extended generation window at 720p–1080p — suitable for narrative sequences, commercial spots, and music video segments.
Dual-Branch Diffusion Transformer processes visual and audio branches simultaneously for coherent joint generation.
## Specifications
| Feature | Details |
| ------------------------ | ------------------------------------------ |
| **Developer** | ByteDance |
| **Released** | February 10, 2026 |
| **Architecture** | Dual-Branch Diffusion Transformer (DB-DiT) |
| **Resolution** | 720p–1080p |
| **Duration** | 4–15 seconds |
| **Aspect ratios** | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 |
| **Audio** | Dialogue, SFX, music (native) |
| **Max reference images** | 9 |
| **Max reference videos** | 3 |
| **Max reference audio** | 3 |
| **Generation modes** | 4 |
## Availability and requirements
| Requirement | Details |
| ---------------------- | ------------------------------------- |
| **Plan** | Creator plan or above |
| **Email verification** | Business domain verification required |
## How to use
Before accessing Seedance 2, complete business domain email verification in your account settings.
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Seedance 2** from the model dropdown. Confirm your plan is Creator or above.
Select **Text to Video**, **Image to Video**, **First and Last Frame**, or **References** depending on your workflow.
In References mode, upload up to 9 images, 3 video clips, and 3 audio clips to guide the output.
Describe the scene, subject, motion, camera behavior, and audio atmosphere.
Click **Generate** to create your video with synchronized audio.
## Prompting tips
* **Describe audio explicitly** — "With the sound of a violin playing softly in the background" or "city traffic noise in the distance" directly influences the audio generation.
* **Use audio references for music style** — Upload a short audio clip in References mode to anchor the musical style and tempo of the generated audio.
* **First-and-Last-Frame for precise transitions** — Define your opening and closing images; write the prompt around motion style and atmosphere rather than restating what's in the frames.
* **Multi-shot: use transition cues** — "THEN CUT TO:" or "The camera pulls back to reveal..." helps Seedance 2 understand shot structure.
### Example prompts
> A musician plays acoustic guitar on a rooftop at sunset. The camera slowly orbits around them. Warm orange light, city skyline in background. Guitar melody generated naturally with the visuals. 10 seconds.
> FIRST FRAME: woman standing at a window looking out at rain. LAST FRAME: woman smiling, holding a warm mug. Generate the transition — mood shift from pensive to content. Soft piano music.
## Seedance family comparison
| Model | Audio | References | Duration | Speed | Best for |
| ----------------------------------------------------- | -------------- | ----------------------- | -------- | -------- | ---------------------------- |
| **Seedance 2** | Yes | 9 img + 3 vid + 3 audio | 4–15s | Standard | Max quality, full multimodal |
| [Seedance 2 Fast](/ai-models/video/seedance-2-fast) | Yes | 9 img + 3 vid + 3 audio | 4–15s | Fast | Rapid iteration, pipelines |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | Yes (lip-sync) | Image input | 4–12s | Standard | Multilingual dialogue |
| [Seedance 1.0 Pro](/ai-models/video/seedance-1-0-pro) | No | Image input | 5–10s | Standard | Cinematic storytelling |
# Seedance 2 0 mini
Source: https://docs.imagine.art/ai-models/video/seedance-2-0-mini
## Seedance 2.0 Mini
ByteDance's lightweight video model, built for fast, multi-shot video generation. It's the default model in [Ad Studio's Motion Design](/ad-studio/motion-design) tool, where it's the cheapest of the three Seedance variants offered there.
## Specifications
| Feature | Details |
| ----------------- | --------------- |
| **Developer** | ByteDance |
| **Resolution** | 480p–720p |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Seedance 2.0 Mini**.
Write your prompt and generate as normal.
## Seedance family comparison
| Model | Resolution | Duration | Audio |
| ----------------------------------------------- | ---------- | -------- | ----- |
| **Seedance 2.0 Mini** | 480p–720p | 5–15s | Yes |
| [Seedance 2](/ai-models/video/seedance-2) | 720p–1080p | 4–15s | Yes |
| [Seedance 2.5](/ai-models/video/seedance-2-5) | 480p–1080p | 4–30s | Yes |
| [Seedance Lite](/ai-models/video/seedance-lite) | 480p–720p | 3–12s | No |
# Seedance 2 5
Source: https://docs.imagine.art/ai-models/video/seedance-2-5
## Seedance 2.5
ByteDance's newest video model — now the default in the Video tool's model picker, live alongside [Seedance 2](/ai-models/video/seedance-2) (both exist simultaneously; this isn't a replacement). It has the longest duration ceiling in the whole Seedance lineup.
## Specifications
| Feature | Details |
| ----------------- | --------------- |
| **Developer** | ByteDance |
| **Resolution** | 480p–1080p |
| **Duration** | 4–30 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
Credit cost fluctuated between checks during a live "Unlimited Seedance 2.5" promotional period — treat any specific credit number as temporary, not a stable spec.
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Seedance 2.5**.
Write your prompt (and optionally set start/end frames) and generate as normal.
## Seedance family comparison
| Model | Resolution | Duration | Audio |
| --------------------------------------------------- | ---------- | -------- | ----- |
| **Seedance 2.5** | 480p–1080p | 4–30s | Yes |
| [Seedance 2](/ai-models/video/seedance-2) | 720p–1080p | 4–15s | Yes |
| [Seedance 2 Fast](/ai-models/video/seedance-2-fast) | 720p | 4–15s | Yes |
| Seedance 2.0 Mini | 480p–720p | 5–15s | Yes |
# Seedance 2 fast
Source: https://docs.imagine.art/ai-models/video/seedance-2-fast
VIDEO MODEL
by ByteDance
Seedance 2 family
Seedance 2 Fast
ByteDance's fast-tier variant of Seedance 2.0 — the same Dual-Branch Diffusion Transformer architecture with native audio-video generation, up to 15 seconds, multimodal references, and significantly lower latency for production and high-volume workflows.
Audio
Dialogue + SFX + Music
References
9 images + 3 video + 3 audio
Seedance 2 Fast uses the same underlying model as Seedance 2 but is optimized for lower latency. Choose Seedance 2 Fast for rapid iteration and production pipelines where speed matters; choose [Seedance 2](/ai-models/video/seedance-2) when maximum quality is the priority.
## Fast-tier Seedance 2
Seedance 2 Fast is ByteDance's production-optimized endpoint for the Seedance 2.0 architecture — released February 10, 2026 alongside the standard model. The underlying Dual-Branch Diffusion Transformer is identical; the Fast variant trades a small margin of peak quality for meaningfully lower inference times, making it the practical choice for iterative workflows, A/B testing, and high-frequency generation pipelines.
Native audio-video joint generation is preserved in the Fast variant — dialogue, sound effects, and music are generated simultaneously with the video, synchronized at the frame level.
## Capabilities
Generates dialogue, sound effects, and music synchronized with the video in a single pass — no post-production audio required.
Supports generation lengths from 4 to 15 seconds, covering short social clips through extended narrative sequences.
Accepts up to 9 reference images, 3 reference video clips, and 3 audio clips simultaneously for maximum creative direction.
Generates coherent multi-shot sequences from a single prompt — scene transitions, subject consistency, and style maintained across cuts.
Supports 21:9, 16:9, 4:3, 1:1, 3:4, and 9:16 aspect ratios for any platform or format.
Optimized for lower latency — ideal for rapid iteration, pipeline integrations, and high-volume production.
## Specifications
| Feature | Details |
| ------------------------ | ------------------------------------------ |
| **Developer** | ByteDance |
| **Released** | February 10, 2026 |
| **Resolution** | 720p |
| **Duration** | 4–15 seconds |
| **Aspect ratios** | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 |
| **Audio** | Dialogue, SFX, music (native) |
| **Max reference images** | 9 |
| **Max reference videos** | 3 |
| **Max reference audio** | 3 |
| **Architecture** | Dual-Branch Diffusion Transformer (DB-DiT) |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Seedance 2 Fast** from the model dropdown.
Select text-to-video, image-to-video, or references mode depending on your creative needs.
Upload up to 9 reference images, 3 video clips, and 3 audio clips to guide the output style, subject, and sound.
Describe the scene, subjects, motion, camera behavior, and audio atmosphere in your prompt.
Click **Generate**. Seedance 2 Fast will produce a video with synchronized audio faster than the standard variant.
## Prompting tips
* **Describe audio explicitly** — Include what you want to hear: "with the sound of rain pattering on a window and a soft piano melody in the background."
* **Specify camera movement** — "Slow dolly forward," "static wide shot," or "handheld tracking shot" all meaningfully influence the output.
* **Use reference audio for tone** — Uploading a reference audio clip helps anchor the musical style and ambient mood of the generated video.
* **Keep multi-shot prompts structured** — For sequences, describe each shot with a clear transition cue: "SHOT 1: ... CUT TO SHOT 2: ..."
### Example prompts
> A chef in a professional kitchen carefully plates a dish under warm overhead lighting. Close-up on hands arranging microgreens. Ambient kitchen sounds — sizzling pans, light chatter in the background. Cinematic, handheld camera.
> A timelapse of a city square from empty early morning through bustling midday. Wide establishing shot. Birds chirping at dawn, building to the hum of traffic and crowd noise by noon.
## Compare models
| Model | Speed | Max duration | Audio | References | Best for |
| ------------------------------------------------------- | -------- | ------------ | --------------- | ----------------------- | ------------------------------- |
| **Seedance 2 Fast** | Fast | 15s | Native | 9 img + 3 vid + 3 audio | Production pipelines, iteration |
| [Seedance 2](/ai-models/video/seedance-2) | Standard | 15s | Native | 9 img + 3 vid + 3 audio | Maximum quality output |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | Standard | 12s | Native lip-sync | Image input | Dialogue, multilingual |
| [Seedance Pro Fast](/ai-models/video/seedance-pro-fast) | Fast | 10s | No | Image input | Quick clips without audio |
If your workflow requires high throughput or rapid iteration, Seedance 2 Fast is the right default. Switch to the standard [Seedance 2](/ai-models/video/seedance-2) when you need the absolute best quality for a final deliverable.
# Seedance pro fast
Source: https://docs.imagine.art/ai-models/video/seedance-pro-fast
VIDEO MODEL
by ByteDance
Seedance 1 family
Seedance Pro Fast
ByteDance's speed-optimized Seedance Pro variant — 30–60% faster inference than standard Seedance 1.0 Pro, with the same multi-shot coherence, complex camera movements, and up to 1080p visual quality. Built for rapid iteration, ad variation generation, and high-frequency production pipelines.
Generation time
Under 60 seconds
Seedance Pro Fast uses the same underlying model as [Seedance 1.0 Pro](/ai-models/video/seedance-1-0-pro) but with a speed-optimized inference configuration. Output quality is very close to the standard Pro with significantly lower generation time.
## Production-pace Pro generation
Seedance Pro Fast is the speed-optimized variant of Seedance 1.0 Pro — designed for workflows where the Pro model's quality is needed but the generation time of the standard model is a bottleneck. At 30–60% faster inference and under 60 seconds per clip, it's practical for rapid iteration cycles, client preview workflows, ad variation production, and any pipeline that generates Seedance Pro content at volume.
The multi-shot coherence, semantic prompt understanding, and complex camera movement support from Seedance 1.0 Pro are all preserved in the Fast configuration.
## Capabilities
Significantly reduced generation time versus standard Seedance 1.0 Pro — sub-60-second generation for most clip configurations.
Maintains the visual fidelity, multi-shot coherence, and semantic understanding of Seedance 1.0 Pro with minimal quality trade-off.
Full support for dolly, zoom, pan, and tracking shots — the complete Seedance Pro camera vocabulary.
Subject and visual style consistency maintained across cuts and transitions — core Seedance architecture strength preserved.
Full HD output at 480p, 720p, or 1080p depending on delivery requirements.
Designed for ad variation pipelines, client preview generation, and any workflow requiring Seedance Pro quality at higher frequency.
## Specifications
| Feature | Details |
| ------------------- | ----------------------------------- |
| **Developer** | ByteDance |
| **Resolution** | 480p, 720p, 1080p |
| **Duration** | 3–12 seconds |
| **Speed** | 30–60% faster than Seedance 1.0 Pro |
| **Generation time** | Under 60 seconds |
| **Audio** | No native audio |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Seedance Pro Fast** from the model dropdown.
Use the same prompting approach as Seedance 1.0 Pro — detailed scene descriptions with camera direction, subject behavior, and lighting.
Select 480p, 720p, or 1080p and a duration between 3 and 12 seconds.
Click **Generate** — expect under 60 seconds for most configurations.
## Prompting tips
* Same as Seedance 1.0 Pro — Seedance Pro Fast uses the same model and responds to the same prompt patterns.
* **For A/B testing**: Generate multiple prompt variations quickly in Fast mode, then produce the winning direction in standard Seedance 1.0 Pro for maximum quality.
* **Iterative refinement**: Use Fast for quick compositional adjustments; the results are close enough to the standard Pro to make confident creative decisions.
### Example prompts
> An explorer discovers an ancient temple hidden in a jungle. Tracking shot following from behind as she approaches the entrance. Dappled light filtering through the canopy. 10 seconds, 1080p.
> A fashion model poses on a rooftop at golden hour. Camera slowly orbits counterclockwise. Wind in hair, city skyline in background. 5 seconds.
## Seedance family comparison
| Model | Quality | Speed | Duration | Audio | Best for |
| ----------------------------------------------------- | -------- | -------- | -------- | ----- | ----------------------------- |
| **Seedance Pro Fast** | High | Fast | 5 or 10s | No | Speed + quality balance |
| [Seedance 1.0 Pro](/ai-models/video/seedance-1-0-pro) | Highest | Standard | 5 or 10s | No | Maximum quality output |
| [Seedance Lite](/ai-models/video/seedance-lite) | Standard | Fastest | 5 or 10s | No | Quick social/e-commerce |
| [Seedance 2 Fast](/ai-models/video/seedance-2-fast) | Advanced | Fast | 4–15s | Yes | Audio-visual, fast production |
Use Seedance Pro Fast as your default for Seedance Pro generation. Switch to [Seedance 1.0 Pro](/ai-models/video/seedance-1-0-pro) only when you need the maximum possible quality for a final deliverable.
# Sora 2 pro
Source: https://docs.imagine.art/ai-models/video/sora-2-pro
VIDEO MODEL
by OpenAI
MM-DiT architecture
Sora 2 Pro
OpenAI's physics-aware flagship video model — 4–20 seconds at 1080p with integrated dialogue, sound effects, and ambient audio generated in a single pass. Built for final production output where physical accuracy, prompt fidelity, and long-form narrative matter most.
Audio
Dialogue + SFX + Ambient
A standard [Sora 2](/ai-models/video/sora-2) variant is also available for rapid iteration and exploration. Sora 2 Pro delivers higher final quality, more stable rendering in complex scenes, and better adherence to nuanced prompts — use it for final production output.
## OpenAI's final-production video model
Sora 2 Pro is built on OpenAI's Multimodal Diffusion Transformer (MM-DiT) architecture and generates video at up to 1080p for 4–20 seconds. Audio (dialogue, sound effects, ambient) is generated in a single pass alongside the video, synchronized at the frame level without post-production.
The Pro tier offers meaningfully higher quality over standard Sora 2 in the scenarios where it counts most: complex multi-element scenes with accurate physics, nuanced prompt instructions, and long-form narratives where rendering stability matters across the full clip duration.
## Capabilities
Dialogue, sound effects, and ambient audio generated in a single pass — precisely synchronized with the visual output without post-editing.
Understands gravity, collisions, and spatial relationships naturally — better object stability, realistic material behavior, and fewer visual artifacts in complex scenes.
A generous generation window — suitable for narrative sequences, commercial spots, and multi-beat storytelling.
Responds accurately to instructions for camera movements, emotional tone, lighting, pacing, and scene transitions — including nuanced multi-part instructions.
Accepts text prompts alone, an uploaded image as a starting frame, or a combination of both for greater control over visual consistency.
Multimodal Diffusion Transformer processes visual and audio branches with joint attention — coherent audio-visual output from a single generation pass.
## Specifications
| Feature | Details |
| ----------------- | ----------------------------------------- |
| **Developer** | OpenAI |
| **Architecture** | Multimodal Diffusion Transformer (MM-DiT) |
| **Resolution** | 1080p |
| **Duration** | 4–20 seconds |
| **Aspect ratios** | 16:9 (1280×720), 9:16 (720×1280) |
| **Audio** | Dialogue, SFX, ambient (native) |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Sora 2 Pro** from the model dropdown.
Write a text prompt, upload an image as a starting frame, or combine both. Include explicit audio cues in your prompt for synchronized sound.
Set the video **duration** (4–20 seconds) and **aspect ratio** based on your project needs.
Click **Generate** to produce the video with integrated audio.
Preview the output and refine your prompt or parameters before downloading.
## Prompting tips
* **Include explicit audio cues** — "With the sound of rain on glass" or "soft jazz playing in the background" directly influences the audio generation alongside the visual.
* **Use the full duration for narratives** — Describe a beginning, middle, and resolution. Sora 2 Pro maintains rendering stability and character consistency across the full duration.
* **Specify camera behavior precisely** — "The camera slowly orbits around the subject" or "cut to a close-up on the hands" gives Sora 2 Pro clear direction for camera motion.
* **Describe physics interactions explicitly** — "A glass tips over and water spills across the table" or "leaves scatter in a gust of wind" benefit from the physics-aware rendering.
* **For image-to-video** — Make sure the reference image style matches the aesthetic in your text prompt to avoid visual inconsistency in the generation.
### Example prompts
> Wide shot: two figures stand in the foreground, gazing at a majestic waterfall cascading into a river below. The camera slowly pans left to reveal the full expanse of the waterfall, capturing the lush greenery and dramatic sky. The roar of the water fills the audio. 15 seconds.
> A barista carefully prepares a latte, steaming the milk with practiced precision. Soft café ambient sounds, quiet chatter in the background. Close-up on the hands, slow rack focus to the finished drink. 10 seconds.
> POV shot: a mountain biker navigates a muddy trail in a dense forest during a rainstorm. The camera tracks forward, capturing mud splashes and rain. The sound of the storm and bike tires on wet ground. 20 seconds.
## Compare models
| Model | Duration | Audio | Physics | Best for |
| ------------------------------------------------- | --------- | ----- | ------- | --------------------------------------------- |
| **Sora 2 Pro** | Up to 25s | Yes | Yes | Final production, long-form, physics-accurate |
| [Sora 2](/ai-models/video/sora-2) | Up to 25s | Yes | Yes | Rapid iteration, exploration |
| [Google Veo 3.1](/ai-models/video/google-veo-3-1) | Up to 60s | Yes | — | Longest clips, broadcast quality |
| [Kling 3.0 Pro](/ai-models/video/kling-3-0-pro) | Up to 15s | Yes | — | 4K, multilingual audio, multi-shot |
| [Seedance 2](/ai-models/video/seedance-2) | Up to 15s | Yes | — | Max references, multimodal |
Sora 2 Pro is the right choice when physical accuracy, audio coherence, and long-form narrative stability matter more than generation speed. For the fastest OpenAI output, use [Sora 2](/ai-models/video/sora-2) for iteration before committing to a final Pro render.
# Wan 2 2
Source: https://docs.imagine.art/ai-models/video/wan-2-2
VIDEO MODEL
by Alibaba
Wan family
Wan 2.2
Alibaba's Mixture of Experts video model — the Video Animation Control Engine (VACE) with camera trajectory controls, subject locking, and background stabilization, plus a few-shot LoRA pipeline for custom style adaptation using just 10–20 images. 720p at 30 FPS.
Architecture
MoE (\~10B params)
Camera control
VACE engine
## Advanced camera control and style adaptation
Wan 2.2 is built on a Mixture of Experts (MoE) diffusion architecture — approximately 10 billion parameters arranged as specialized experts for high-noise and low-noise diffusion stages, producing more efficient and higher-quality output than the 14B monolithic architecture of Wan 2.1.
The Video Animation Control Engine (VACE) is the central differentiating feature: explicit camera trajectory inputs, subject locking (keeping a defined subject stationary relative to camera movement), and background stabilization for controlled scene composition. Combined with a few-shot LoRA pipeline that adapts the model to a custom visual style using only 10–20 reference images, Wan 2.2 is the most configurable model in the Wan family.
Licensed under Apache 2.0, the underlying model is open for commercial use.
## Capabilities
Explicit camera trajectory inputs — pans, zooms, focus pulls, and custom paths — for precise cinematographic control over the generated video.
Keep a defined subject visually stable relative to camera movement — useful for product showcase, character focus, and controlled scene composition.
Stabilize the background while the subject moves, or vice versa — independent control over foreground and background motion.
Adapt the model to a custom visual style using just 10–20 reference images via a few-shot LoRA pipeline — style consistency across generations.
\~10B parameter Mixture of Experts model — more efficient than 14B monolithic architectures, with specialized processing for different noise levels.
720p output at 30 frames per second.
## Specifications
| Feature | Details |
| -------------------- | --------------------------------------------- |
| **Developer** | Alibaba (Wan Video) |
| **Architecture** | Mixture of Experts (MoE), \~10B parameters |
| **Resolution** | 720p |
| **Frame rate** | 30 FPS |
| **Duration** | Up to 5 seconds (T2V); multiple for I2V |
| **Camera control** | VACE — pans, zooms, focus pulls, custom paths |
| **Style adaptation** | Few-shot LoRA (10–20 images) |
| **Audio** | No native audio |
| **License** | Apache 2.0 |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Wan 2.2** from the model dropdown.
Use VACE parameters to set your camera trajectory — specify pans, zoom direction, focus pull behavior, and any subject locking requirements.
Describe the scene with detail on subjects, setting, style, and motion. Wan 2.2's semantic understanding handles complex compositional instructions.
Click **Generate** for 720p output at 30 FPS.
## Prompting tips
* **Be explicit with camera trajectory** — "Camera slowly pans right while maintaining focus on the stationary subject" is processed as a VACE instruction, not just a suggestion.
* **Describe background vs. foreground separately** — "Background blurs softly while the subject remains sharp and stationary in frame center" activates subject locking and background stabilization.
* **Use style references for LoRA** — For brand content requiring a specific visual style, the few-shot LoRA pipeline can adapt the model to your reference aesthetic.
### Example prompts
> A product bottle sits on a frosted glass surface. The camera orbits slowly around it 180 degrees, keeping the bottle perfectly centered. Studio lighting, clean white background. 5 seconds, 720p.
> A wide shot of a mountain valley. The camera slowly pushes forward through a field of wildflowers, foreground flowers in focus, mountains slightly blurred in background. Golden hour light. 5 seconds.
## Compare models
| Model | Camera control | Style adapt | Audio | Architecture | Best for |
| ----------------------------------------------- | -------------- | ----------------- | ----- | ------------ | --------------------------- |
| **Wan 2.2** | VACE explicit | LoRA (10–20 imgs) | No | MoE 10B | Camera-precise, brand style |
| [Wan 2.5](/ai-models/video/wan-2-5) | Prompt-based | No | Yes | — | Audio-visual sync |
| [Wan 2.6](/ai-models/video/wan-2-6) | Prompt-based | No | Yes | — | Character reference, audio |
| [Kling 2.5 Pro](/ai-models/video/kling-2-5-pro) | Prompt-based | No | No | — | Fast, affordable 1080p |
Wan 2.2 is the best choice when precise camera control and brand-style consistency matter. The VACE system gives you a level of programmatic camera control not available in prompt-only models.
# Wan 2 5
Source: https://docs.imagine.art/ai-models/video/wan-2-5
VIDEO MODEL
by Alibaba
Wan family
Wan 2.5
Alibaba's audio-visual sync model — generates ambient sounds, sound effects, and voice with precise lip-sync alongside the video in a single pass. Supports 480p to 1080p at 5 or 10 seconds with flexible aspect ratios and text or image input.
Audio
Ambient + SFX + Voice
## Audio-visual synchronization in a single pass
Wan 2.5 is Alibaba's dedicated audio-visual synchronization model in the Wan family. Its primary strength is the one-pass A/V generation system — ambient sounds, sound effects, and voice are generated simultaneously with the video, synchronized at the frame level without post-production. Lip-sync support makes it particularly well-suited for content where characters speak, sing, or react expressively to audio.
For reference-to-video with character insertion and voice reference support, see [Wan 2.6](/ai-models/video/wan-2-6) — the successor model with expanded capabilities. Wan 2.5 is the audio-capable general-purpose member of the Wan family for standard A/V production.
## Capabilities
Ambient sounds, sound effects, and voice generated simultaneously with the video — no separate audio editing or syncing required.
Character lip movements synchronized accurately with generated audio — suitable for dialogue, narration, and character-driven clips.
Consistent subject movement, natural transitions, and fluid camera behavior across the full clip duration.
480p, 720p, or 1080p — select based on quality requirements and credit budget.
Supports text prompts, uploaded reference images, or a combination of both for broader creative control.
16:9, 9:16, 1:1, 4:3, and 3:4 — flexible framing for social, cinematic, and standard formats.
## Specifications
| Feature | Details |
| ----------------- | -------------------------------------- |
| **Developer** | Alibaba (Wan Video) |
| **Resolution** | 480p, 720p, 1080p |
| **Duration** | 5 or 10 seconds |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:3, 3:4 |
| **Audio** | Ambient, SFX, voice (native, one-pass) |
| **Lip-sync** | Yes |
| **Input modes** | Text-to-video, image-to-video |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Wan 2.5** from the model dropdown.
Write a text prompt or upload a reference image. Include explicit motion, mood, and audio cues for best results.
Choose **5 or 10 seconds** and your preferred resolution (480p, 720p, or 1080p).
Click **Generate** to produce the video with synchronized audio.
Preview the clip, adjust your prompt or settings as needed, and download.
## Prompting tips
* **Include audio cues explicitly** — "Rain in the background," "distant city traffic," or "soft piano music" feed directly into the audio generation alongside the visual.
* **Describe motion and mood** — Be specific about how subjects move and the atmosphere you want. "Slow pan," "bustling city energy," or "tense stillness" all guide the model.
* **Use camera terminology** — "Overhead shot," "wide establishing shot," and "slow zoom in" give clear directional cues.
* **Specify lighting** — "Golden hour," "low-key studio lighting," or "overcast afternoon" guide the visual output alongside the audio.
* **For lip-sync** — Describe your character's speech or emotional reaction explicitly to anchor the lip movement generation.
### Example prompts
> Close-up shot: a woman in a vintage suit sits pensively at a café table. The camera slowly zooms in on her thoughtful expression as she speaks softly. Warm, ambient café sounds — quiet chatter, distant music. 10 seconds, 16:9.
> A young man carefully unpacks a pair of headphones in a modern apartment. Smooth dolly shot, slow zoom in on his focused expression. City ambient sounds through open windows in the background. 10 seconds, 1080p.
## Compare models
| Model | Audio | Lip-sync | Duration | R2V | Best for |
| ----------------------------------------------------- | ----- | ------------ | -------- | --- | ------------------------------------ |
| **Wan 2.5** | Yes | Yes | 10s | No | General A/V, lip-sync |
| [Wan 2.6](/ai-models/video/wan-2-6) | Yes | Yes | 15s | Yes | Character reference, voice insertion |
| [Wan 2.2](/ai-models/video/wan-2-2) | No | No | 5s | No | Camera control, LoRA style |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | Yes | Multilingual | 12s | No | Multilingual precision lip-sync |
Use Wan 2.5 when your project needs both visual impact and audio coherence in a single generation. For character identity preservation with voice reference input, upgrade to [Wan 2.6](/ai-models/video/wan-2-6) — the R2V successor model.
# Wan 2 6
Source: https://docs.imagine.art/ai-models/video/wan-2-6
VIDEO MODEL
by Alibaba
Wan family
Wan 2.6
Alibaba's reference-to-video model — insert a character's appearance and voice from a reference input, generate multi-shot narratives with synchronized audio, and produce up to 15 seconds at 1080p with precise lip-sync. Built for character-centric, multilingual, and audio-synchronized production.
Duration
Up to 15 seconds
Audio
SFX + Music + Lip-sync
## Reference-to-video: put real characters in any scene
Wan 2.6's headline capability is its R2V (Reference-to-Video) mode — upload a reference image of a character and Wan 2.6 inserts that character's appearance consistently into a generated scene. Combined with voice reference input, both the character's face and voice can be preserved in the generated video, making Wan 2.6 uniquely capable for creator-centric workflows where personal or brand character identity needs to appear in generated content.
The model also introduces comprehensive upgrades across text-to-video, image-to-video, and audio-to-video generation with one-pass A/V synchronization and precise lip-sync.
## Capabilities
Insert a character's appearance from a reference image — and optionally their voice — into any generated scene with consistent identity preservation.
Audio and video generated in a single pass — synchronized sound effects, music, and voice generated with the video without post-production.
Character lip movements synchronized accurately with generated or reference audio — suitable for dialogue-driven content.
Generates coherent multi-shot sequences from simple prompts — scene transitions, character continuity, and narrative flow maintained automatically.
One of the longer generation windows in the lineup — supports more developed narrative sequences at 5, 10, or 15-second intervals.
Text-to-video, image-to-video, audio-to-video, and reference-to-video all supported in a single model.
## Generation modes
| Mode | Description |
| ---------------------------- | --------------------------------------------------------------- |
| **Text-to-video** | Generate video from text prompt with A/V sync |
| **Image-to-video** | Animate a reference image with motion and audio |
| **Reference-to-video (R2V)** | Insert a character's appearance and voice from reference inputs |
| **Audio-to-video** | Generate matching visuals from an audio reference |
## Specifications
| Feature | Details |
| ----------------- | -------------------------------------- |
| **Developer** | Alibaba (Wan Video) |
| **Resolution** | 720p, 1080p |
| **Duration** | 5, 10, or 15 seconds |
| **Frame rate** | 24 FPS |
| **Aspect ratios** | 16:9, 9:16, 1:1, 4:3, 3:4 |
| **Audio** | SFX, music, synchronized, lip-sync |
| **R2V** | Character appearance + voice insertion |
## How to use
Log into ImagineArt and go to the **AI Video Generator**.
Choose **Wan 2.6** from the model dropdown.
Choose the **Reference-to-Video** generation mode.
Upload a reference image of the character to use. Optionally, upload a voice reference audio clip.
Write a prompt describing the scene, environment, action, and audio atmosphere around your character.
Click **Generate**. Wan 2.6 places your referenced character into the generated scene with synchronized audio.
Go to the **AI Video Generator** and select **Wan 2.6**.
Describe scene, subjects, motion, and audio cues. Include any multi-shot structure with transition cues.
Choose 5, 10, or 15 seconds at your target resolution.
Click **Generate** for audio-synced video.
## Prompting tips
* **R2V: describe the scene, not the character** — The reference image provides the character; your prompt should focus on the setting, action, camera, and audio environment.
* **Include audio cues for one-pass sync** — "A jazz trio plays softly in the background" or "footsteps echo on the marble floor" integrate directly into the audio generation.
* **Multi-shot: use transition language** — "THEN CUT TO:" or "The camera pulls back to reveal..." cues structured multi-shot generation.
* **15-second clips for narratives** — Use the full 15-second window for storylines that need a beginning, middle, and resolution within one generation.
### Example prompts
> \[R2V mode] Reference character appears as a chef in a busy restaurant kitchen. The chef plates a dish confidently, a soft smile as they look at the camera. Warm kitchen sounds, sizzling in background. 10 seconds.
> A multilingual brand video: a young woman introduces a product in front of a clean white background. She speaks naturally, hands gesturing. Confident, friendly. 10 seconds, 1080p.
## Compare models
| Model | R2V | Audio | Lip-sync | Duration | Best for |
| ----------------------------------------------------- | --- | ----- | ------------ | -------- | -------------------------- |
| **Wan 2.6** | Yes | Yes | Yes | 15s | Character reference, A/V |
| [Wan 2.5](/ai-models/video/wan-2-5) | No | Yes | Yes | 10s | General A/V production |
| [Wan 2.2](/ai-models/video/wan-2-2) | No | No | No | 5s | Camera control, style LoRA |
| [Seedance 1.5 Pro](/ai-models/video/seedance-1-5-pro) | No | Yes | Multilingual | 12s | Multilingual precision |
Wan 2.6 is the best choice when a specific character needs to appear consistently in generated video — the R2V system provides character identity preservation that other models can't match from a simple image reference alone.
# Wan 3
Source: https://docs.imagine.art/ai-models/video/wan-3
## Wan 3
Alibaba's newest Wan video model, live alongside [Wan 2.6](/ai-models/video/wan-2-6), [Wan 2.5](/ai-models/video/wan-2-5), and [Wan 2.2](/ai-models/video/wan-2-2) — an addition to the lineup, not a replacement. It has the longest duration ceiling of any Wan model so far.
## Specifications
| Feature | Details |
| ----------------- | --------------- |
| **Developer** | Alibaba |
| **Resolution** | 480p–1080p |
| **Duration** | 5–30 seconds |
| **Audio** | Yes |
| **Frame control** | Start/End frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **Wan 3**.
Write your prompt and generate as normal.
## Wan family comparison
| Model | Resolution | Duration | Audio |
| ----------------------------------- | ---------- | -------- | ----- |
| **Wan 3** | 480p–1080p | 5–30s | Yes |
| [Wan 2.6](/ai-models/video/wan-2-6) | 720–1080p | 5–15s | Yes |
| [Wan 2.5](/ai-models/video/wan-2-5) | 480p–1080p | 5–10s | Yes |
| [Wan 2.2](/ai-models/video/wan-2-2) | 720p | 5s | No |
# Xai grok 1 5
Source: https://docs.imagine.art/ai-models/video/xai-grok-1-5
## xAI Grok 1.5
A new xAI video model, live alongside [xAI Grok Video](/ai-models/video/grok-video).
## Specifications
| Feature | Details |
| ----------------- | ------------ |
| **Developer** | xAI |
| **Resolution** | 480p–720p |
| **Duration** | 5–15 seconds |
| **Audio** | Yes |
| **Frame control** | Start frame |
## How to use
Go to the **ImagineArt AI Video Generator**.
From the model dropdown, choose **xAI Grok 1.5**.
Write your prompt and generate as normal.
## xAI video model comparison
| Model | Resolution | Duration | Audio |
| --------------------------------------------- | ---------- | -------- | ----- |
| **xAI Grok 1.5** | 480p–720p | 5–15s | Yes |
| [xAI Grok Video](/ai-models/video/grok-video) | 480–720p | 6–15s | Yes |
# AI Resize
Source: https://docs.imagine.art/ai-resize
## Summary
The AI Resize Node intelligently resizes your images to different aspect ratios while preserving the visual integrity of the content. Unlike basic cropping or stretching, this node uses AI to adapt your composition—extending backgrounds, repositioning elements, and filling in new areas seamlessly. It's ideal for repurposing a single design asset like a poster, ad, or banner into multiple format variations.
### How to Use
Click the Add (+) button and select AI Resize from the Image node category.
Connect an image from another node (such as Generate Image or Edit Image), or upload one directly.
Choose your desired output ratio from the Aspect Ratio dropdown (e.g., 1:1, 16:9, 4:3, 9:16).
Click Run, and the AI will resize your image to the selected aspect ratio, intelligently adapting the composition.
### Choosing the Right Settings
| Setting | Type | Impact on Output |
| ------------ | ------------------------------------- | ----------------------------------------------------------------------------------------------------------------------------------------------- |
| Aspect Ratio | Dropdown (1:1, 16:9, 4:3, 9:16, etc.) | Defines the target dimensions for the resized image. The AI adapts the composition to fit the new ratio without awkward cropping or stretching. |
### Sample Use Cases
Take a single ad creative and resize it for every platform in one workflow—1:1 for Instagram, 9:16 for Stories and TikTok, 16:9 for YouTube thumbnails, and 4:3 for Facebook feeds.
Resize a vertical movie or event poster into a horizontal banner for web headers or a square format for social media—without losing the key visual elements.
Convert a single hero product image into multiple banner sizes for your storefront, email headers, and marketplace listings.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Imagine Apps
Source: https://docs.imagine.art/apps/imagine-apps
Specialized AI-powered tools for creating CGI ads, editing images, and generating music — all with just a few clicks and no prior editing experience required.
Imagine Apps is a suite of specialized AI tools designed to produce ready-to-share creative content with minimal effort. Each app focuses on a specific creative output — from cinematic product commercials to original music tracks — and handles the heavy lifting so you can focus on the creative direction.
You can access all Imagine Apps at [imagine.art/apps](https://www.imagine.art/apps).
## How Imagine Apps works
Every app in the suite follows the same simple pattern:
Browse 150+ ready-to-use templates or select a specific app mode tailored to your creative goal. Each template is designed for a particular type of content or effect.
Upload your images, enter a text prompt, or select a genre — depending on the app. The AI uses your inputs as creative direction to generate the final output.
Configure options such as aspect ratio, style, or genre to match your project requirements.
Click **Generate** and let the AI produce your content. No manual editing required.
## Accessing Imagine Apps
Navigate to **Imagine Apps** from the main ImagineArt navigation, or go directly to [imagine.art/apps](https://www.imagine.art/apps). Each app is accessible from the apps dashboard, where you can browse by category or search for a specific template type.
Imagine Apps uses your ImagineArt credits for generations. Check your current credit balance before starting a project, especially for high-resolution outputs.
## Google Drive Integration
Connect your Google Drive to pull assets directly into any Workflow and push generated outputs back to your Drive — without leaving the canvas.
**To connect Google Drive:**
In Workflows, click your profile or workspace settings and navigate to **Integrations**.
Select **Google Drive** and click **Connect**. You'll be redirected to Google's OAuth flow to authorize access.
Once connected, use the **Import** node on the canvas and select **Google Drive** as the source. Browse your Drive folders and select the files you want to use.
Use the **Export** node and select **Google Drive** as the destination. Generated outputs will be saved to the Drive folder you choose.
You can import and export in the same workflow — pull a reference image from Drive, generate a video, and push the result back, all in one run.
# What are Audio Tools
Source: https://docs.imagine.art/audio-tools
Generate speech, music, and cloned voices using AI.
Audio Tools let you create original music tracks, convert text into natural-sounding speech, and generate audio in cloned voices — all from simple text inputs.
Convert any script into natural-sounding speech. Choose from multiple voices, emotions, and languages, with fine-grained control over speed, pitch, and volume.
Generate original full-length music tracks from a text description. Pick a style and let the AI compose an instrumental or vocal track tailored to your mood.
Produce speech in the voice of a wide range of celebrity and original AI voices. Adjust stability to control how consistent or expressive the delivery sounds.
# Carousel Maker
Source: https://docs.imagine.art/carousel-maker
## Summary
The Carousel Maker Node generates a multi-slide social media carousel from a single prompt. Describe the content once, and the node produces a full set of slides sized and sequenced for a platform like Instagram — useful for turning one idea into a ready-to-post carousel without designing each slide individually.
## How to Use
Click the Add (+) button and select **Carousel Maker** from the Image node category.
Describe the carousel's content and theme. This input is required.
Connect or upload a reference image to guide the carousel's visual style.
In the Properties panel, choose **Dimensions** (defaults to "Instagram Carousel 4:5"), **Language** (defaults to English), and **No. of Slides** (defaults to 5, adjustable with a stepper).
Click Run. The node generates the full set of slides as its output.
## Settings
| Setting | Type | Impact on Output |
| ------------- | -------- | ----------------------------------------------------------------------------------------- |
| Dimensions | Dropdown | Sets the slide format/aspect ratio — defaults to a standard Instagram Carousel (4:5) size |
| Language | Dropdown | Sets the language used in the generated slide text |
| No. of Slides | Stepper | Controls how many slides the carousel contains — defaults to 5 |
## Sample Use Cases
Describe a product's key features in the prompt and Carousel Maker breaks it into a multi-slide sequence, one idea per slide.
Write a prompt describing a numbered list or step-by-step explainer — the node paces it out across the generated slides.
# Combine Text
Source: https://docs.imagine.art/combine-text
The Combine Text Node allows you to merge multiple text inputs into one cohesive output. It's ideal for projects that require synthesizing different ideas, prompts, or pieces of content into a single, unified text.

## How to Use
Click the Add (+) button and select Combine Text from the Text Utilities node category.
Connect multiple text inputs (such as paragraphs, prompts, or content) into the node.
The node will automatically merge the provided inputs into a single, well-structured output.
## Sample Use Cases
You can have multiple inputs like research data, key quotes from experts, and your original thoughts. The Combine Text Node will merge all of these into a smooth, structured blog post, ensuring the content flows naturally.

You might have various sections of a script—an introduction, product features, and a call to action. The Combine Text Node will merge these sections into a smooth, professional video script.

You may have separate bullet points describing the features, benefits, and usage of a product. The Combine Text Node will merge them into a single, engaging product description for your e-commerce platform.

## When to Use
* **Content Merging**: When you need to merge various pieces of content into a single, cohesive output.
* **Scriptwriting**: Combine different sections of a script (intro, body, conclusion) into one seamless narrative.
* **Product Descriptions**: Merge individual product features into a polished, engaging description for e-commerce.
* **Marketing Materials**: Combine promotional copy, calls to action, and product details into a unified marketing message.
# Combine Videos and Audios
Source: https://docs.imagine.art/combine-videos-and-audios
Merge a video with an audio track to create a single file with embedded audio. Perfect for adding voiceovers, music, sound effects, or dialogue to your video.
Join multiple video clips sequentially into one continuous video. Assemble segments in the order you want them to play.
## Combine Audio and Video
The Combine Audio & Video Node takes a video input and an audio input and merges them into a single video file with the audio track embedded. Both inputs are required.
### How to Use
Click the Add (+) button and select **Combine Audio & Video** from the Video node category.
Link a video via the Video input handle (marked in green). This is the visual content.
Link an audio file via the Audio input handle (marked in pink). This is the audio that will be layered onto the video.
Click **Run**, and the node outputs a single video file with both the visual and audio tracks combined.
## Combine Videos
The Combine Videos Node joins multiple video clips together sequentially, outputting a single video that plays each clip one after another in the order they are connected.
### How to Use
Click the Add (+) button and select **Combine Videos** from the Video node category.
Link multiple video inputs in the order you want them to play. Each connected clip will be stitched sequentially—first in, first played.
Click **Run**, and the node outputs a single video with all clips joined together in sequence.
## Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Monetization
Source: https://docs.imagine.art/community/monetization
Earn credits every time someone downloads your assets on ImagineArt.
## Monetization on ImagineArt Community
ImagineArt offers an exciting opportunity for creators to **monetize their assets** and earn rewards for their work. By participating in the community, you can generate income each time someone downloads your asset. Whether you're sharing art, videos, or other AI-generated creations, this page will guide you through the monetization process and explain how credit transactions work.
### How Monetization Works?
As a creator on ImagineArt, you can **earn credits** every time a user downloads your asset in **2K** or **4K resolution**. These credits can be used for purchases within the platform, saved up for rewards, or withdrawn.
**Credit Breakdown:**
* **2K Quality**: 10 credits per download.
* **4K Quality**: 20 credits per download.
**Transaction Process:**
When a user downloads your asset:
1. The **credits** (10 for 2K, 20 for 4K) are deducted from the user's account.
2. The same amount of credits is added to your **Community Credits** balance.
You can earn credits on every download in either resolution, allowing you to generate revenue as your assets gain popularity.
#### Community Challenges and Additional Rewards
ImagineArt hosts **monthly challenges** where creators can compete for cash prizes of up to **\$11,000**. These challenges often focus on metrics such as:
* **Most liked asset** of the month.
* **Most views** on a profile.
* **Most downloads** of an asset.
* **Best-curated profile**, and more.
Participating in these challenges is an excellent way to gain more exposure, earn additional credits, and compete for cash rewards.
#### Managing Your Credits
Your earned credits will be displayed in your **Community Credits** section, accessible through the drop-down menu at the top-right of your profile. The **Community Credits** are a special balance that accumulates whenever your assets are downloaded in high-quality resolutions (2K/4K).
**Where You Can See Your Credits:**
* **Community Credits**: This shows your total credits earned from asset downloads.
* **Subscription Credits**: Your subscription-based credits, which may include special offers or promotional credits.
* **Top-Up Credits**: Additional credits purchased or added manually to your account.
These credits can be used to unlock assets from other creators, redeem rewards, or saved for future purchases within the community.
#### How to Track and Manage Your Earnings
To keep track of your earnings and manage your credits:
1. Go to your **Profile** and check the **Community Credits** section for a detailed breakdown of all credits earned and spent.
2. Under your profile's **drop-down menu**, you'll find your **Total Credits** as well as breakdowns for **Subscription Credits** and **Top-Up Credits**.
As you continue to share, engage, and participate in community events, your credits will build up over time, giving you more ways to explore and enhance your experience.
#### Next Steps for Creators
1. **Publish Your Work**: Start by uploading and publishing your creations. The more downloads your assets get, the more credits you'll earn.
2. **Promote Your Assets**: Share your profile and asset links to drive traffic and encourage others to download your work.
3. **Participate in Challenges**: Join monthly challenges to increase your exposure and earn cash prizes.
Monetizing your assets on ImagineArt gives you the power to turn your creative work into tangible rewards. Whether you're creating digital art, AI-generated images, or videos, our platform makes it easy to earn credits and grow your artistic career.
# Create your Profile
Source: https://docs.imagine.art/community/profile
Set up and manage your ImagineArt community profile to showcase your work, grow your audience, and unlock monetization opportunities.
## Setting up your Profile on ImagineArt
Welcome to the first step of your creative journey on ImagineArt! By creating your profile, you'll join a dynamic community of artists and creators, where you can showcase your work, gain exposure, and interact with other members. Your profile is the perfect space to build your portfolio, grow your presence, and start sharing your AI-generated creations with the world.
Your profile is your portfolio. It's how potential clients, collaborators, and fans discover your work.
#### Why Create a Profile?
Your profile is more than just a collection of assets—it's a personal showcase where you can:
* **Showcase Your Work**: Display your top assets and create a curated portfolio that highlights your best work.
* **Gain Exposure**: Get noticed by millions of users browsing the community for inspiration. The more engagement your assets receive, the higher your visibility.
* **Track Your Progress**: Monitor the performance of your assets, including likes, downloads, and views, all in one place.
* **Monetize Your Creations**: Start earning credits when others download your assets, and unlock opportunities to participate in monthly community challenges for cash rewards.
### How to Create Your Profile
To start, you need to sign in to your ImagineArt account. If you don't have one yet, simply create an account by entering your details.
Once logged in, navigate to the **"View Your Profile"** section from the top right drop down by clicking on your profile picture. This is where you'll build your personal space on ImagineArt.
Add a profile picture, cover photo, and a brief bio that represents you and your creative style. This is your chance to introduce yourself to the community.
Now that your profile is ready to set up, start adding images and videos to your profile to showcase your work! There are multiple ways to go about this:
1. **Add from my creations** in your profile when you are starting from scratch.
2. **Click on "Share your creations"** on the main community page.
3. If you already have some assets published on your profile, simply click the **publish button** on your profile page that will open all your unpublished assets. Here you can multi select your assets and bulk publish!
#### What to Include in Your Profile
* **Profile Picture**: Choose a picture that represents your artistic identity or a logo that reflects your brand.
* **Cover Image**: Add a visually appealing cover photo to make your profile stand out.
* **Bio**: Write a short description about yourself, your creative journey, and what inspires you. This is your opportunity to introduce yourself to the community.
* **Featured Assets**: Curate a selection of your top creations. Select the pieces that best represent your style and talent.
* **Social Links**: If you have any other platforms or social media accounts where you showcase your work, link them to your profile to grow your network.
#### Why It's Important to Build a Strong Profile
* **Build Your Portfolio**: Your profile is your portfolio. It's how potential clients, collaborators, and fans discover your work.
* **Engage with the Community**: With a complete profile, you can engage more effectively with the community, participate in discussions, and share feedback.
* **Gain Recognition**: The more interaction your assets receive—likes, views, downloads—the more likely they are to be featured in the community's main feed, increasing your visibility.
**Curate Your Best Work**: Don't upload everything—select your best assets to showcase your talent.
#### Next Steps After Creating Your Profile
* **Publish More Assets**: Keep uploading your work to stay active and visible in the community.
* **Participate in Challenges**: Each month, new community challenges are launched. Upload your best creations and stand a chance to win cash prizes up to **\$11,000**.
* **Promote Your Profile**: Share your profile link across social media or with your network to bring more attention to your creations.
* **Monetize Your Work**: As your assets are downloaded, you'll start earning credits. Use these credits to further explore the community or redeem rewards.
Creating your profile on ImagineArt is the first step in building your artistic career and connecting with a global community of creators. Start now and join a platform designed to inspire, engage, and reward.
# Showcase Your Work
Source: https://docs.imagine.art/community/showcase
Publish and share your AI-generated images and videos in the ImagineArt community to build your portfolio and grow your audience.
## Get more visible
Showcasing your work on ImagineArt allows you to gain visibility, build your portfolio, and connect with a global audience. Whether you're a professional artist or a hobbyist, this platform is your space to display your best creations and inspire others.
### How to Showcase Your Work
Upload your best images, videos, or AI-generated content directly to your profile. Choose between **2K** or **4K** resolution for higher quality.
Select your top assets to highlight on your profile. Curate your work to reflect your style and expertise.
Interact with other creators by liking, commenting, and sharing feedback on their work. Build connections and inspire others.
#### What Happens After You Publish?
Once you publish your assets:
* **Get Feedback**: Community members can like, download, add to collection and follow your work.
* **Increase Engagement**: More interaction can lead to higher exposure and more downloads.
* **Monetize**: Earn credits every time someone downloads your asset in 2K or 4K quality.
Showcasing your work on ImagineArt is an essential step in building your presence, gaining recognition, and engaging with a global creative community. Start uploading and share your creations with the world today!
# Top Creators
Source: https://docs.imagine.art/community/top-creators
Learn how ImagineArt selects Top Creators and how to work toward earning the designation.
## Top Creators on ImagineArt Community
The **Top Creators** page highlights the most engaged and successful artists on ImagineArt. These creators are featured based on their activity, community involvement, and the popularity of their assets.
Being a **Top Creator** not only increases your visibility but also allows you to gain recognition among millions of users.
### How to Become a Top Creator?
Regularly publish your best assets to maintain visibility.
Interact with other creators, comment, and participate in challenges.
Share your profile and assets on social media to increase traffic and engagement.
### How Are Top Creators Selected?
Top Creators are chosen based on:
* **Engagement**: Likes, views, and downloads of your assets.
* **Profile Activity**: Consistent uploads, interaction with other creators, and community participation.
* **Quality of Work**: High-quality, well-curated assets that gain significant attention from the community.
Being a **Top Creator** on ImagineArt gives you the chance to grow your reputation, reach a wider audience, and gain recognition for your work.
# ImagineArt Community
Source: https://docs.imagine.art/community/what-is-community
A vibrant ecosystem for creators to showcase, share, and monetize their AI-generated assets.
## What is Community?
Welcome to the **ImagineArt Community**, a vibrant ecosystem that empowers creators to showcase, share, and monetize their AI-generated assets. Our platform is designed to foster creativity, inspire users, and help creators build a professional portfolio. Whether you're here to discover the best AI art, learn from top creators, or start your own journey as a creator, this is the space to do it.
### How to Get Started
To start your journey in the ImagineArt community, here's how you can get involved:
Build your portfolio, introduce yourself to the community, and start publishing your creations.
Upload and curate your best assets to gain visibility and inspire others.
Earn credits every time someone downloads your assets in 2K or 4K resolution.
Learn how to earn recognition as one of ImagineArt's featured top creators.
### Our Goal: Empowering Creators and Users
ImagineArt's community aims to:
* **Empower Creators**: By providing a platform for creators to generate, curate, and monetize their digital assets.
* **Inspire Users**: By offering a curated, high-quality feed of AI-generated art that serves as a reliable source of inspiration.
* **Foster Learning & Discovery**: By allowing users to discover top creators, learn from their techniques, and engage with their work.
#### Explore the Community
ImagineArt is more than just a platform; it's a thriving community where creators support each other, share knowledge, and collaborate. Whether you're a seasoned artist or just starting your creative journey, this is the place to expand your reach, learn, and earn from your creations.
Start exploring the ImagineArt Community today, and join thousands of creators shaping the future of AI art!
# Compositor
Source: https://docs.imagine.art/compositor
Combine and layer multiple inputs directly on the Workflows canvas.
The **Compositor** node lets you merge and layer multiple inputs — images, videos, or both — into a single composited output, directly inside the Workflows canvas. No external tools needed.
## How to use the Compositor node
Open the node picker (`Space` or right-click on the canvas) and search for **Compositor**. Click to place it.
Connect the output handles of any Image or Video nodes to the input slots on the Compositor. You can connect multiple inputs — each one becomes a separate layer.
In the Eedit layers mode, reorder layers by dragging them up or down. Set blend mode, opacity, position, and scale for each layer independently.
Connect the Compositor's output handle to any downstream node — a Generate Video node, an Export node, or another Compositor for more complex compositions.
Click **Run** on the Compositor node or run the full workflow. The node merges all connected layers and produces a single composited output.
## Supported inputs
| Input type | Notes |
| ---------- | ------------------------------------------------------------- |
| **Image** | Static frames from any image node |
| **Video** | Clips from any video generation or edit node |
| **Mixed** | Image and video inputs can be combined in the same Compositor |
Use the Compositor to stitch together outputs from separate generation pipelines before passing them to a final export or post-processing node.
# Copyright Checker
Source: https://docs.imagine.art/copyright-checker
## Summary
The Copyright Checker Node scans a creative for copyright or trademark risk before you publish it. It's available in both the Image and Video node categories, so you can check a still or a clip directly inside a workflow instead of reviewing it manually afterward.
## How to Use
Click the Add (+) button and select **Copyright Checker** from either the Image or Video node category (it appears in both).
Upload an image (PNG, JPG, or WEBP, up to 30MB) or connect a video from another node.
Use the **Choose Creative Type** dropdown to specify what you're checking (for example, "Image (Static Asset)").
Click **Analyze Image** (or the video equivalent) to run the check.
This node flags potential copyright/trademark risk — it's a screening aid, not a legal clearance. Review anything it flags before publishing.
## Sample Use Cases
Route a finished image or video through Copyright Checker before it reaches an export or publish node, catching obvious risks early in the workflow.
Check an uploaded reference image or video before using it as the basis for further generation, to avoid building on something risky.
# Edit Image
Source: https://docs.imagine.art/edit-image
The Edit Image Node allows you to create new images inspired by reference images. Rather than simply modifying existing images, this node uses uploaded reference images as inspiration to generate entirely new visuals based on your prompt. Whether you want to create a variation of a concept, change the environment, or combine multiple design elements, this node gives you the freedom to generate fresh, creative outputs.
With models like Nano Banana Pro, Nano Banana 2, and Seedream v4.5, you can upload 10+ reference images to guide the AI in creating new content that fits your desired style, concept, or scene.

## How to Edit and Create with Reference Images
Upload multiple reference images that serve as the foundation for your new image. These images can represent different aspects of the final output, such as color schemes, styles, objects, or environments. Models like Nano Banana Pro, Nano Banana 2, and Seedream v4.5 support multiple references for a more refined and varied concept.
Provide a prompt that explains what you want the AI to create, using the reference images as inspiration. The prompt should describe what should be incorporated from the references into the new generated image (e.g., "Create a futuristic cityscape using the style from the uploaded images, with neon lights and flying cars").
Click Generate, and the AI will use your reference images and prompt to create a new, unique image that fits your concept, combining elements from the provided references into a fresh design.
### Working with Multiple Reference Images
For models like Nano Banana Pro, Nano Banana 2, and Seedream v4.5, you can upload 10+ reference images to provide greater inspiration and guide the AI in generating more complex or intricate scenes. This is particularly useful when:
* You need to combine multiple styles or visual elements.
* You're looking to generate a completely new concept or scene that combines aspects from several images.
* You want to maintain consistency while introducing variations or new ideas.
## Common Editing and Creation Tasks
Generate a fashion lookbook in a cyberpunk style, featuring futuristic clothing designs with glowing accessories, tech-integrated outfits, and bold streetwear.
**How to Achieve It:**
* **Prompt:** "Create a fashion lookbook with models wearing futuristic cyberpunk clothing in neon-lit city streets."
* **Multiple References:** Upload reference images of cyberpunk fashion and tech-inspired clothing, and direct the AI to combine these into dynamic fashion looks.
Generate fashion outfits and visualize them on models, perfect for virtual try-on apps.
**How to Achieve It:**
* **Prompt:** "Generate an image of a model wearing a futuristic jacket and glowing sneakers on a busy city street."
* **Multiple References:** Upload images of different clothing items and a model, prompting the AI to seamlessly combine them into a complete outfit.
Create mythical creatures (dragons, unicorns, etc.) in everyday real-world environments (e.g., parks, cities, beaches).
**How to Achieve It:**
* **Prompt:** "Generate a scene where a dragon flies over a busy city during the day."
* **Multiple References:** Upload images of real-world environments and mythical creatures, guiding the AI to blend fantasy elements into real settings.
Turn flat, 2D digital art or illustrations into 3D models with depth, shadowing, and realism.
**How to Achieve It:**
* **Prompt:** "Transform this 2D character design into a realistic 3D model with shading and depth."
* **Multiple References:** Upload a 2D character design and reference images of similar 3D models, asking the AI to turn the design into a 3D visual.
Create vintage-style posters with a modern twist, incorporating retro typography, bold colors, and vintage themes.
**How to Achieve It:**
* **Prompt:** "Generate a vintage travel poster for a futuristic city, using bright, retro colors and old-school typography."
* **Multiple References:** Upload reference images of classic vintage posters and modern cityscapes, asking the AI to combine both into a unique, retro-futuristic design.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Edit Video
Source: https://docs.imagine.art/edit-video
The Edit Video Node lets you transform existing videos by using reference clips and prompts to guide the AI in making changes. Rather than editing frame-by-frame manually, you describe the changes you want, style shifts, scene alterations, visual effects, or environment swaps and the AI generates a new version of the video based on your instructions and reference inputs.
## How to Use
1. **Add the Node:** Click the Add (+) button and select Edit Video from the Video node category.
2. **Provide a Reference Video:** Connect a video from another node (such as Generate Video or Import), or upload one directly. This serves as the base that the AI will transform.
3. **Write Your Prompt:** Describe the changes you want applied to the video (e.g., "Change the setting to a snowy mountain landscape" or "Apply a cyberpunk neon aesthetic to the entire scene").
4. **Generate:** Click Run, and the AI will produce a new version of the video with your requested changes applied.
## Sample Use Cases
Apply a completely different visual style to existing footage, turn a daytime street scene into a noir-style night sequence, or give realistic footage an anime or watercolor look.
Change the environment of a video without reshooting. Transform an indoor scene to an outdoor setting, swap a plain background for a futuristic cityscape, or shift the season from summer to winter.
Take raw video footage and apply consistent brand aesthetics color grading, visual tone, and stylistic elements across multiple clips to maintain a unified look for campaigns.
Add atmospheric effects like rain, fog, or dramatic lighting to an existing video, or shift the mood from bright and cheerful to dark and cinematic using a simple prompt.
## Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Extend Video
Source: https://docs.imagine.art/extend-video
The Extend Video Node generates additional footage that continues seamlessly from where your original video ends. Instead of looping or duplicating frames, the AI understands the scene's context, motion, and visual style—then creates new frames that naturally extend the narrative. It's ideal for turning short clips into longer sequences, adding breathing room to intros and outros, or building out scenes that need more runtime.
## How to Use
Click the Add (+) button and select Extend from the Video node category.
Link an existing video from another node (such as Generate Video or Import). This is the clip the AI will continue from.
Describe what should happen next in the extended footage (e.g., "The camera slowly pulls back to reveal the full cityscape"). If no prompt is provided, the AI continues the scene naturally based on the existing content.
Select your Model, Duration, Resolution, and Aspect Ratio from the Properties panel.
Click Run, and the AI generates new footage that picks up exactly where the original clip left off.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| -------------- | -------------------------------- | ----------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Model | Dropdown | Selects the AI model used for extension. Different models vary in how well they maintain continuity, motion quality, and visual consistency with the source clip. |
| Duration | Dropdown (e.g., 5s, 10s) | Sets how much additional footage to generate. |
| Resolution | Dropdown (e.g., 720p, 1080p) | Determines the output resolution. Match this to your source clip for seamless continuity. |
| Aspect Ratio | Dropdown (16:9, 9:16, 1:1, etc.) | Defines the frame dimensions. Should match the original video. |
| Generate Audio | Checkbox | When enabled, generates a matching audio track for the extended portion. |
| Seed | Number Input | A fixed number for reproducible results across generations. |
## Sample Use Cases
Turn a 5-second AI-generated clip into a 15–30 second video suitable for YouTube Shorts, TikTok, or Reels—without the content feeling repetitive or looped.
Extend a dramatic establishing shot or action sequence to give it more runtime for a trailer, showreel, or narrative project. Add a prompt to guide what happens next in the scene.
Add extra seconds to the beginning or end of a video for title cards, transitions, or fade-outs while keeping the visual style and motion perfectly consistent.
## Extend Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Extract Video Frame
Source: https://docs.imagine.art/extract-video-frame
## Summary
The Extract Video Frame Node captures a single frame from a video and outputs it as a still image. Specify the exact moment using a frame number or timecode, and the node extracts that frame for use in downstream image nodes. It's the bridge between your video and image workflows.
## How to Use
Click the Add (+) button and select Extract Frame from the Video Utilities sub-category (the palette label is "Extract Frame," though this page is titled "Extract Video Frame").
Link a video from another node (such as Generate Video, Import, or Extend).
Enter a Frame number or a Timecode (format: HH:MM:SS) to specify exactly which moment to extract.
The node outputs the extracted frame as a still image, ready to connect to any image node downstream.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| -------- | ------------------------------ | --------------------------------------------------------------------------------------------------------------------------- |
| Frame | Number Input (default: 1) | Specifies which frame to extract by its sequence number. Frame 1 is the first frame of the video. |
| Timecode | Time Input (default: 00:00:00) | Specifies which frame to extract by timestamp. Useful when you know the exact moment you want rather than the frame number. |
## Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Account Deletion Request
Source: https://docs.imagine.art/faq/account-deletion
How to permanently delete your personal ImagineArt account or a team workspace, and what data is removed.
Once permanent deletion is complete, your data cannot be recovered. Personal accounts have a 7-day cooling-off window — team workspace deletion is immediate and cannot be undone.
## What gets deleted
Deleting your personal account removes:
* Personal account details
* Generated images
* Collections
* All other data linked to your user account
Any active subscriptions are cancelled automatically. You have a **7-day cooling-off period** after initiating deletion during which you can reverse the process.
Deleting a team workspace removes:
* Team information
* Shared assets and directories
* Workspace access for all members
* All references tied to that workspace
**Team workspace deletion is immediate and cannot be undone.** Only the team owner can delete a team or transfer ownership. The default personal team cannot be deleted.
## Delete your personal account
Click your profile picture in the top-right corner and navigate to account settings.
From the sidebar, select the **Account** section.
Scroll down and expand the **Danger Zone** section.
Click **Delete Account** and confirm via the popup that appears.
Approve the deletion. This activates the 7-day cooling-off period. If you do not reverse the request within 7 days, your account is permanently deleted.
## Delete a team workspace
This action is immediate and irreversible. Only the team owner can perform this action.
Click the team selector at the top of your dashboard.
Select the specific team you want to delete.
Go to the **Team Overview** section in settings.
Scroll down and expand the **Danger Zone** section.
Click the **Delete Team** button.
Select either permanent deletion or transfer ownership to another member.
Confirm your selection. If transferring ownership, approve the transfer to complete the process.
## Need help?
For assistance with account access or deletion, contact [support@imagine.art](mailto:support@imagine.art).
# How to Cancel Your Subscription
Source: https://docs.imagine.art/faq/cancel-subscription
Step-by-step guide to cancelling your ImagineArt subscription on web or mobile, and what happens after cancellation.
## Cancelling on the web
Go to [imagine.art](https://www.imagine.art) and log in with your credentials.
Click your profile picture in the top-right corner and click on the settings icon from the dropdown menu.
Locate and select the \*\*Plan and Billing \*\*tab.
Follow each step carefully until you see a confirmation that your cancellation has been processed.
## Mobile and web subscriptions
Mobile and web subscriptions are linked. Cancelling on either platform automatically updates the other. To sync your accounts, sign in at [imagine.art](https://www.imagine.art) using the same email address associated with your Google Play Store or Apple App Store account.
If you subscribed through the iOS App Store or Google Play and are having trouble cancelling, contact the [ImagineArt support team](mailto:support@imagine.art) for assistance.
## What happens after cancellation
* Access to your subscription features continues through the end of your current billing period.
* No additional charges occur unless you choose to resubscribe.
* You can restart your subscription at any time from your account dashboard.
## Need help?
If you encounter any issues during cancellation, contact [support@imagine.art](mailto:support@imagine.art).
# Credit Assignment and Usage Policy
Source: https://docs.imagine.art/faq/credit-policy
How credits are assigned across subscription plans, the no-rollover policy, and your options when credits run out.
## How credits work
ImagineArt operates on a **monthly credit assignment cycle**. Every subscription tier receives credits at the beginning of each monthly period, regardless of whether you are on a monthly, quarterly, or annual plan.
## Credit distribution by plan type
| Plan | How credits are assigned |
| ------------- | ---------------------------------------------------------------------------- |
| **Monthly** | Full credit allotment assigned at the start of each monthly renewal |
| **Quarterly** | A set amount of credits assigned each month for three consecutive months |
| **Annual** | A set amount of credits assigned each month throughout the twelve-month term |
All plans receive credits on a monthly basis — the commitment period determines your billing cycle, not when credits arrive.
## No-rollover policy
Unused credits do not roll over to the next month. At the start of each new cycle, your balance resets and your new monthly allotment is applied.
To get the most value from your plan, select a tier that matches your typical monthly usage. You can view your current balance and next renewal date from your Billing and subscription tab under the Plan and Billing section.
## Options when credits run out
If you use all your credits before the monthly cycle resets, you have three options:
Your credits automatically replenish at the start of your next monthly cycle. You can check the exact renewal date in your account dashboard.
Buy a one-time credit bundle to continue generating without waiting for your cycle to reset. Top-ups are available from the **Manage Subscription** section of your dashboard.
If you consistently run out of credits before your cycle ends, upgrading to a higher tier gives you a larger monthly credit allotment. See [Upgrading or Downgrading Your Subscription](/faq/upgrade-downgrade) for steps.
## Need help?
For questions about your credit balance or subscription, visit your account dashboard or contact [support@imagine.art](mailto:support@imagine.art).
# Other FAQs
Source: https://docs.imagine.art/faq/other-faqs
Answers to common questions about commercial use, content filters, cross-platform accounts, and copyright of AI-generated art.
Yes — paid subscribers have the freedom to incorporate generated images into commercial projects.
**Benefits of commercial use:**
* Unique, customizable content aligned with your brand's needs
* Time and cost savings compared to traditional design workflows
* Flexible modification options — adjust colors, text, and elements as needed
* A source of creative inspiration that you can build on
**Required guidelines:**
* Respect copyright and intellectual property. Do not reproduce logos, trademarked materials, or other protected content without authorization.
* Follow applicable laws regarding advertising, privacy, and content usage in your jurisdiction.
* Review and customize generated images before commercial deployment to ensure they meet your standards.
ImagineArt implements content filters to prevent explicit, harmful, or inappropriate material from being generated. These filters exist to maintain platform safety and community standards.
While filters may occasionally limit certain creative directions, the development team actively refines these systems based on user feedback. If you believe a prompt was incorrectly blocked, contact [support@imagine.art](mailto:support@imagine.art).
Your account, credits, and gallery are shared across web and mobile — sign in with the same credentials on both platforms to access your content everywhere.
However, **web and mobile subscriptions are separate**. A subscription purchased on the web cannot be transferred to the mobile app, and vice versa. You will need to manage subscriptions independently on each platform.
The ImagineArt team is working toward fuller cross-platform compatibility in future updates.
Under current U.S. law, AI-generated art cannot be copyrighted because it lacks human authorship. As a result, purely AI-generated works enter the public domain and can be used freely by anyone.
**Strategies for stronger protection:**
* **Incorporate substantial human creative input** — The more you modify, direct, and build on generated images, the stronger your claim to ownership.
* **Trademark your designs** — Brand-specific designs can be protected through trademark registration.
* **Blend AI tools with human creativity** — Combining AI-generated elements with original human artwork produces a result that may qualify for copyright protection.
Copyright law around AI-generated content is evolving. Consult a legal professional for advice specific to your jurisdiction and use case.
# Password Recovery and Sign-in
Source: https://docs.imagine.art/faq/password-recovery
How to recover your password, set up email login from a social account, and sign in using any of ImagineArt's four authentication methods.
## Sign-in methods
ImagineArt supports four ways to sign in:
* **Google** authentication
* **Facebook** authentication
* **Discord** authentication
* **Email and password**
You can sign in with a different method than the one you originally used to sign up. For example, if you registered with Google, you can also log in using Facebook, Discord, or your email address.
## Recover your password
Navigate to [imagine.art](https://www.imagine.art) and click **Log In**.
Enter the email address associated with your account and click **Continue**.
On the next screen, select the **Forgot Password** option.
Open the password reset email sent to your inbox and click the reset link.
Follow the link to create and confirm your new password.
## Set up email login from a social account
If you originally registered with Google, Facebook, or Discord but want to add email-and-password access to your account:
On the sign-in page, enter the email address associated with your social account and click **Continue**.
A password reset email will be sent to your inbox automatically.
Open the email, click the link, and create a password. You can now sign in with email and password in addition to your social login.
## Still having trouble?
Contact [support@imagine.art](mailto:support@imagine.art) for help with account access.
# Troubleshooting Your ImagineArt Account
Source: https://docs.imagine.art/faq/troubleshooting
Eight steps to resolve common sign-in issues, loading problems, and browser compatibility issues with ImagineArt on the web.
If you're experiencing issues with your ImagineArt web account, work through the steps below in order. Most problems are resolved by steps 1–3.
Never share your password with anyone to ensure your account's safety.
## Step-by-step troubleshooting
Simply logging out and logging back in can refresh your session and fix minor issues. Navigate to [imagine.art](https://www.imagine.art), click logout, wait a moment, then re-enter your credentials.
Open your browser settings and clear browsing data — specifically **cached images and files** and **cookies**. Restart the browser before attempting to sign in again.
Switch to a different web browser to determine if the issue is browser-specific. Common alternatives include Chrome, Firefox, Safari, and Edge.
Confirm you have a stable internet connection. Refresh the page or reconnect to your network if needed.
Ensure you are using the latest version of your browser. Outdated browser software can cause compatibility issues with the ImagineArt website.
Some browser extensions can interfere with the ImagineArt website. Temporarily disable all extensions, then try again. If this resolves the issue, re-enable extensions one at a time to identify the cause.
If you see visual loading problems, force a full page reload that bypasses cached content:
* **Mac:** `Command + Shift + R`
* **Windows:** `Ctrl + F5`
If none of the above steps resolve the issue, email [support@imagine.art](mailto:support@imagine.art) with a description of the problem and any error messages you see.
# Updating Billing Information
Source: https://docs.imagine.art/faq/updating-billing
How to update your payment method, download invoices, and manage billing details through ImagineArt's billing portal.
## Update your billing information
Click your profile picture in the top-right corner of the dashboard and select **Billing and Subscription** from the dropdown.
Click the **Edit Billing** button on the billing page. You can also view your current plan and next charge date here.
You will be taken to Stripe's official payment page where you can update your payment method, billing address, and other details.
## Download invoices
From the same Stripe billing page, you can view all past invoices. Select a specific invoice to view its details, then download either the invoice or receipt for your records.
## Common issues
Billing information can only be updated when you have an active subscription. Purchase a plan first — billing updates will be available from your second invoice onward.
* **New customers** can add a billing address during checkout.
* **Returning customers** using the same email address cannot directly edit the first invoice's billing address. Contact [support@imagine.art](mailto:support@imagine.art) or purchase a new plan to apply address updates to future invoices.
Name changes apply to future invoices only. To update the name on past invoices, contact [support@imagine.art](mailto:support@imagine.art).
Taxes are calculated by Stripe, ImagineArt's payment processor, based on the country associated with your billing address.
## Need help?
For billing inquiries or adjustments that can't be made through the portal, contact [support@imagine.art](mailto:support@imagine.art).
# Upgrading or Downgrading Your Subscription
Source: https://docs.imagine.art/faq/upgrade-downgrade
How to upgrade, downgrade, cancel a scheduled change, or purchase credit top-ups from your ImagineArt subscription dashboard.
## Upgrade your subscription
From your profile, locate and open the **Billing and Subscription** tab.
Click the **Upgrade** button on the card of the plan you want to upgrade to
Select when you want the upgrade to take effect:
* **Immediate activation** — The new plan starts right away. A new billing cycle begins immediately with credits allocated to the new plan.
* **End-of-period activation** — The upgrade takes effect when your current billing cycle ends. You won't be charged until the new cycle starts.
Review the plan summary and any adjusted charges, then confirm your selection.
If you select end-of-period activation, cancellation options will be unavailable until after the new billing cycle begins.
## Downgrade your subscription
From your profile, open the **Billing and Subscription** tab.
Review the available plans and select the one that meets your needs.
Click **Downgrade** and confirm your selection. Your current plan will remain active until the end of your current subscription cycle. Upon the next renewal you will only be charged for the downgraded plan.
Downgrades take effect only after your current billing period ends. For an immediate downgrade, contact [support@imagine.art](mailto:support@imagine.art).
## Cancel a scheduled plan change
If you've scheduled an end-of-period change and want to cancel it before it takes effect:
Open the **Billing and Subscription** tab from your dashboard.
Navigate to the \*\*Billing and payment \*\*section.
Click the **Cancel** option to remove the scheduled change. Your current plan will continue as-is.
## Purchase credit top-ups
If you need more credits without changing your plan, you can buy a one-time credit bundle:
1. Open the **profile** tab.
2. Select **Buy Credits**.
3. Choose a bundle and confirm your purchase.
## Need help?
Contact [support@imagine.art](mailto:support@imagine.art) for assistance with subscription changes.
# Catalog shoots
Source: https://docs.imagine.art/fashion-studio/catalog-shoot
Generate clean, consistent studio shots for product listings.
A Catalog shoot is for clean, consistent studio shots for product listings — the mode to reach for when you need product-page-ready images or clips without an editorial "look."
Go to [imagine.art/fashion-studio](https://imagine.art/fashion-studio) and click **Catalog image** (or **Catalog video**). This opens the Catalogue Shoot editor — the same editor as an [Editorial shoot](/fashion-studio/editorial-shoot), with the same **[Model](/fashion-studio/models)**, **[Choose outfit](/fashion-studio/choosing-an-outfit)**, **Model styling**, and **[Shot settings](/fashion-studio/scenes-and-poses)** panels.
Pick a model, add your **Top** and **Bottom** images, and set a scene and pose — the same controls as Editorial. **Mode** also has **[Video](/fashion-studio/video-shoot)** and **[Edit](/fashion-studio/editing-a-creation)** options alongside Image. Click **Shoot** to generate.
Looking for the more styled, campaign-oriented mode instead? See [Editorial shoots](/fashion-studio/editorial-shoot).
# Choosing an outfit
Source: https://docs.imagine.art/fashion-studio/choosing-an-outfit
Dress your model with preset garments and footwear, or upload your own to generate a custom shot.
Every shot needs an outfit. Fashion Studio lets you dress the model with preset garments, footwear, and accessories, or upload your own for a fully custom look.
Under **Choose outfit**, click **Top** or **Bottom** to open a searchable catalog of preset garments. Everything here is ready to use as-is — no upload required.
Click **Create New** to add a garment that isn't in the preset catalog. Choose **Shirt** or **Dress**, then upload clear-background **front** and **back** photos of the item (an optional **fabric & print details** photo helps with texture). **Auto cutout** is on by default, so you don't need to pre-mask the photos yourself.
Under **Model styling**, click **Footwear & accessories** to open a similar picker — **Footwear** and **Accessories** each have their own preset catalog, plus **My uploads** for anything you've added yourself and **Create New** for uploading a new item. Both are optional; skip them if the shot doesn't need them.
# Editing a creation
Source: https://docs.imagine.art/fashion-studio/editing-a-creation
Regenerate an existing image with a new scene, angle, or description — a reference-driven workflow, not a brush or mask tool.
**Edit** is the third mode under **Mode**, alongside Image and Video. It's for reworking an image you already have rather than building a shot from scratch — but it works by regenerating from a reference image and a text description, not by painting masks or swapping specific regions by hand.
Click **Edit** under **Mode**. Like Video mode, this replaces the entire left panel — Model, Choose outfit, Model styling, and Shot settings from Image mode aren't part of it. You need a reference image before anything else appears.
Click **Select** to open **Select Assets**, and choose an image from **Recent uploads** or **My creations** — or upload a new one directly. You can add a second reference image afterward (for example, a model shot plus a separate garment photo) using the **+** tile next to it.
With a reference image in place, the panel adds **Shot settings** (**Set scene**, **Camera angle**), **Number of variations**, aspect ratio, and an optional **Describe your shot** field — this is where you say what should change (e.g. "outdoor afternoon light, model looking away, editorial style"). Click **Shoot** to generate.
There's no brush, mask, or region-select tool. If you want a specific change (a different background, a swapped garment), describe it in **Describe your shot** rather than looking for a dedicated tool icon.
# Editorial shoots
Source: https://docs.imagine.art/fashion-studio/editorial-shoot
Generate campaign-style fashion imagery and video with a styled, editorial edge.
An Editorial shoot is for campaign imagery and video with a styled, editorial edge — the mode to reach for when mood and styling matter as much as the product itself: brand campaigns, lookbooks, and social content.
Go to [imagine.art/fashion-studio](https://imagine.art/fashion-studio) and click **Editorial image** (or **Editorial video**). This opens the Editorial Shoot editor.
Configure the shot from the left panel:
* **Mode** — choose **Image**, **[Video](/fashion-studio/video-shoot)**, or **[Edit](/fashion-studio/editing-a-creation)**.
* **[Model](/fashion-studio/models)** — pick a preset AI model, or one of your own saved models.
* **[Choose outfit](/fashion-studio/choosing-an-outfit)** — add images for the **Top** and **Bottom**.
* **Model styling** — add footwear and accessories (see [Choosing an outfit](/fashion-studio/choosing-an-outfit)).
* **[Shot settings](/fashion-studio/scenes-and-poses)** — pick a **Set scene** and **Pose** from the built-in libraries, or describe the shot yourself in the optional text field.
Set the number of variations and aspect ratio, then click **Shoot** to generate.
Prefer to start from an existing look? Switch to the **Templates** tab to browse a gallery of ready-made looks, filterable by **All**, **Female**, or **Male**. Pick one to generate from directly instead of building a shot from scratch — see [Templates](/fashion-studio/templates) for the difference between this gallery and the one on the landing page.
Editorial and Catalog shoots share the same editor and the same scene/pose libraries — see [What is Fashion Studio?](/fashion-studio/what-is-fashion-studio) for how the two modes differ.
# Models
Source: https://docs.imagine.art/fashion-studio/models
Pick a preset AI model, reuse your own saved models, or create a brand-new one from a photo or a text description.
Every shoot starts with a model. Fashion Studio gives you three ways to get one: pick from around 85 built-in preset models, reuse a model you've created before, or create a new one from scratch.
From the Editorial or Catalog Shoot editor, click the **Model** card (it shows your current model, e.g. "Alina"). This opens **Select your model**, filterable by **All**, **Favorites**, **My models**, **Female**, or **Male**, with a search box for finding a specific model by name.
Click the **My models** filter to see only models you've previously generated or uploaded on this account, instead of the full preset library.
Click **Generate** on the create tile at the start of the grid to open **Create your own model**. You can either:
* **Upload** a reference photo of a real person to base the model on, or
* **Generate** one from a text description — describe the model's looks and face, then narrow it down with **Female/Male**, **Ethnicity**, and **Body size**. Turning on **Character sheet** produces multiple consistent angles of the same model instead of a single image.
A model you create this way is saved automatically and shows up under **My models** the next time you need it.
# Scenes & poses
Source: https://docs.imagine.art/fashion-studio/scenes-and-poses
Set where your model is shot and how they're standing, from a large preset library or your own reference photos.
**Shot settings** control the background and body position of a shot: **Set scene** for the location, and **Pose** for how the model is standing. Both work the same way — a large preset library, plus a way to bring your own reference.
Click **Set scene** (it shows the current background, e.g. "Beige studio") to open **Select background** — roughly 25 preset locations spanning studio backdrops, European architecture, city streets, and interiors, filterable by **All**, **Saved**, or **Personal**.
Click **Create new background** to upload a photo of your own location, or pick from images you've generated before. This is upload-only — there's no text-to-image option for backgrounds, so you need an actual photo or an existing creation to work from.
Click **Pose** (it shows the current pose, e.g. "Arms crossed stand") to open **Select Pose** — roughly 90 preset poses covering standing, sitting, walking, and back-view shots, filterable by **All**, **Saved**, **My poses**, **Female**, or **Male**.
Click **Create Custom Pose**, the first tile in the grid, to upload a photo of the pose you want the model to replicate. Like backgrounds, this is a bring-your-own-reference flow rather than a text description.
You can also skip both pickers and describe the shot yourself in the optional **Describe your shot** text field, right below Set scene and Pose.
# Templates
Source: https://docs.imagine.art/fashion-studio/templates
Start from a ready-made look instead of building a shot from scratch — two different galleries, depending on how you want to browse.
If you'd rather start from an existing look than build one field by field, Fashion Studio has two separate template galleries — they look similar but serve different purposes.
On the [Fashion Studio landing page](https://imagine.art/fashion-studio), scroll to **Generate across formats** and filter by **All**, **Brand**, **Outfit styling**, **E-commerce**, or **Influencers**. Each filter shows a curated set of named looks (e.g. "Urban Sport," "Parisian Rooftops") built around a mood or use case rather than a specific model.
Once you're in an Editorial or Catalog Shoot editor, click the **Templates** tab at the top. This opens a much larger gallery — around 380 looks — organized by individual model persona rather than use case, filterable only by **All**, **Female**, or **Male**.
This is a genuinely different set of templates from the landing page gallery, not the same templates re-filtered.
Click any template card. There's no preview or confirmation step — it immediately switches to the **Create** tab with the model, outfit, footwear, and scene all filled in to match, ready to click **Shoot** as-is or tweak first.
# Video mode
Source: https://docs.imagine.art/fashion-studio/video-shoot
Turn a still shot into a short video — a different, media-upload-driven panel from Image mode.
Switching to **Video** under **Mode** doesn't just add video-specific fields to the Image mode panel — it replaces the entire left panel with a different workflow. Model, Choose outfit, Model styling, and Shot settings from Image mode aren't part of it.
From the Editorial or Catalog Shoot editor, click **Video** under **Mode**. Configure the shot from:
* **Effects** — a preset style for the video (e.g. "Studio Editorial — E-commerce").
* **Upload media** — the starting image or video the clip is generated from.
* **Aspect ratio**, **Duration**, and **Resolution** pills (defaults: 3:4, 15s, 720p).
Turning on **Advance mode** adds a **Camera movement** setting (default: Auto) and an optional **Describe your shot** text field, for when you want more control than the Effects preset gives you.
Video mode works identically in Editorial and Catalog shoots — same Effects preset, same defaults, same Advance mode controls. Only the panel title changes.
# What is Fashion Studio?
Source: https://docs.imagine.art/fashion-studio/what-is-fashion-studio
Fashion Studio is ImagineArt's dedicated environment for generating fashion campaign and catalog content.
Fashion Studio is a dedicated AI environment purpose-built for fashion content. Instead of a general-purpose image or video prompt, it gives you a structured workflow built around fashion assets: a model, an outfit, styling details, and a scene.
## Editorial vs. Catalog
Both modes use the exact same editor — the same model picker, outfit slots, styling options, scene and pose pickers — so switching between them doesn't mean learning new controls. What changes is the kind of shot you're going for:
* **[Editorial](/fashion-studio/editorial-shoot)** — campaign imagery and video with a styled, editorial edge. Use this for brand campaigns, lookbooks, and social content where mood and styling matter.
* **[Catalog](/fashion-studio/catalog-shoot)** — clean, consistent studio shots for product listings. Use this when you need product-page-ready images or clips without an editorial "look."
Because the underlying controls are identical, you can build a look in one mode and know your way around the other immediately.
## Building a shot
Whichever mode you're in, the editor is built around the same pieces:
* **[Models](/fashion-studio/models)** — pick a preset model, reuse one you've created before, or generate a new one.
* **[Choosing an outfit](/fashion-studio/choosing-an-outfit)** — preset garments, footwear, and accessories, or upload your own.
* **[Scenes & poses](/fashion-studio/scenes-and-poses)** — a large preset library for backgrounds and poses, or bring your own reference photo.
* **[Templates](/fashion-studio/templates)** — skip building a shot entirely and start from a ready-made look.
Beyond still images, each mode also has:
* **[Video mode](/fashion-studio/video-shoot)** — turn a shot into a short video.
* **[Editing a creation](/fashion-studio/editing-a-creation)** — regenerate an existing image with a new scene, angle, or description.
# Adding Audio
Source: https://docs.imagine.art/film-studio/adding-audio
Generate a soundtrack or narration directly on your timeline — no separate audio tool required.
A film isn't just picture. The **Add Audio** button on the Create Video timeline gives you two real generation tools — an AI music composer and a text-to-speech narrator — plus the option to reuse existing audio.
## Four ways to add audio
In the **Create Video** tab, click **Add Audio** on the timeline (next to your scenes). This opens a menu with four options: **AI Music**, **Text to Speech**, **Select from Library**, and **Upload**.
Choose **AI Music** to open a dialog powered by **ElevenLabs Music**. Describe what you want to hear in **Music Description**, optionally turn on **Instrumental** to skip vocals, pick a **Style** preset (e.g. "Dark Cinematic"), and set a **Duration**. Click **Generate** to add the track to your timeline.
Choose **Text to Speech** to open a dialog powered by **ElevenLabs v3**. Type your narration into **Speech**, or click **AI Writer** to have it drafted for you. Pick a voice from **Select voice**, and optionally expand **Parameters** to adjust **Stability** (how consistent the voice sounds versus more expressive and variable). Click **Generate** to add the narration to your timeline.
**Select from Library** pulls in an audio reference you've already added to the project's References library. **Upload** lets you bring in your own audio file directly.
Audio you generate or add here sits on the timeline alongside your video scenes — it isn't the same as attaching an **Audio** reference in the [References](/film-studio/references) library, which is for reusing a voice, track, or sound effect as an influence across multiple scenes rather than placing actual audio on the timeline.
# The Assets Library
Source: https://docs.imagine.art/film-studio/assets-library
Every image and video you've generated in a project, with its exact prompt, one-click regeneration tools, and comments — all behind one button.
The **Assets** button (top right of the workspace) is more than a media browser. Every result you've generated in a project lives here, grouped by date, each with its full generation history and a set of one-click tools for reworking it.
Click **Assets** to switch the canvas into a reverse-chronological feed of every image and video generated in this project, grouped under date headers like "Today."
Click any asset to open its detail view. The **Details** tab shows the exact prompt that produced it and the model used (e.g. "Nano Banana Pro") — useful for understanding why a result looks the way it does, or for recreating a similar look later. A **Comments** tab sits alongside it for per-asset feedback and discussion, and a **Located** chip shows which folder the asset belongs to.
## Creation Actions
Every asset's detail view has a **Creation Actions** panel for regenerating or refining it without starting over:
| Action | What it does |
| ------------- | ------------------------------------------------------------------------------------------------- |
| **Upscale** | Increase resolution — **Subtle** or **Creative** (Creative adds more reinterpreted detail). |
| **Remove Bg** | Cut the subject out from its background. |
| **Variate** | Generate variations — **Subtle** (close to the original) or **Strong** (looser reinterpretation). |
| **Pan** | Extend the image outward in any of the four directions. |
A matching toolbar at the bottom of the viewer (**Relight**, **Camera Angles**, **Variate**, **Background Changer**, **Upscale**, **Edit**) offers the same kind of one-click regeneration for images already in the canvas view, not just from the Assets feed.
## Organization
Assets aren't a flat list — each one belongs to a folder (a project's **Default Folder** unless you've organized further). Use the folder icon next to **Located** to move an asset elsewhere.
Before you regenerate something from scratch, check its **Details** tab first — reusing the exact prompt (with one change) is often faster than reconstructing what worked.
# The Cinematographer's Toolkit
Source: https://docs.imagine.art/film-studio/camera-controls
Four dials separate Film Studio from a generic image generator: camera body, lens, focal length, aperture.
Four dials separate Film Studio from a generic image generator: **camera body, lens, focal length, aperture.** These choices, more than anything else, shape the personality of your image.
Sets the underlying **texture** — digital cleanliness, film grain, large-format scale.
Sets the **optical character** — anamorphic, swirly, surgical, vintage.
Sets the **field of view** — how much of the world you see, how compressed.
Sets the **depth of field** — how much of the image is in sharp focus.
## Camera bodies
Cameras fall into two families: **digital** (clean, modern, color-rich) and **film** (grainy, organic, classic). Pick the family first, then the specific model based on the feel you want.
### Digital cameras
| Camera | Character | Great for |
| ---------------- | --------------------------------------------------------------------------------------------- | ----------------------------------------------------------------------------------------------------- |
| **ARRI Alexa** | Industry standard for high-end cinema. Soft highlights, natural skin tones, restrained color. | Premium dramatic features, character-driven scenes, anywhere you want "this looks like a real movie." |
| **Sony FX6** | Compact cinema-style camera. Versatile, slightly cleaner and sharper than Alexa. | Documentary, indie films, run-and-gun shoots, modern naturalistic stories. |
| **Red V Raptor** | High-resolution 8K with vibrant, saturated color. Crisp detail in every frame. | Action, sci-fi, VFX-heavy scenes, music videos, anywhere maximum detail matters. |
### Film cameras
| Camera | Character | Great for |
| -------------------- | ------------------------------------------------------------------------------- | ------------------------------------------------------------------------------------------------------ |
| **ARRI Flex (35mm)** | Classic film texture. Visible grain, warm tonality, organic falloff in shadows. | Period pieces, nostalgic looks, music videos, anywhere you want a hand-made feel. |
| **IMAX** | Large-format film. Massive resolution, vast dynamic range, epic scope. | Landscapes, scale shots, sci-fi spectacle — Nolan films, blockbusters, anything that should feel huge. |
If you are unsure which camera to pick, **start with ARRI Alexa.** It is the safest choice for almost any narrative scene and tends to produce the most universally pleasing results.
## Lenses
Where the camera body sets the underlying texture, the lens sets the optical personality. Lenses fall into two groups — **anamorphic** (cinematic widescreen) and **spherical** (everything else).
### Anamorphic
| Lens | Character | Great for |
| -------------- | -------------------------------------------------------------------------------------------- | -------------------------------------------------------------------------------------------------------- |
| **Anamorphic** | Cinematic widescreen with characteristic horizontal lens flares, oval bokeh, slight stretch. | Hollywood-style epics, sci-fi, action — anywhere the audience should immediately feel "this is a movie." |
### Spherical
| Lens | Character | Great for |
| ------------- | ---------------------------------------------------------------------- | ----------------------------------------------------------------------------------------- |
| **Fish eye** | Extreme curved distortion, near-180° field of view. | POV action, skating, surreal sequences, dream logic. |
| **Macro** | Surgical detail on small subjects, very shallow focus. | Textures, eyes, food, mechanisms, jewelry — anything where detail is the subject. |
| **Petzval** | Vintage portrait lens with a famous swirly bokeh around the edges. | Romantic portraits, dream sequences, fashion editorials. |
| **Probe** | Long, thin lens that can reach into tight spaces and exaggerate depth. | Tabletop, hyper-detail shots, unique angles into small worlds — food, models, miniatures. |
| **Spherical** | The neutral, naturalistic cinema lens. | Most narrative work — the safe, transparent choice when you want the lens to disappear. |
| **Vintage** | Soft contrast, lower sharpness, character flaws. | Period pieces, music videos, dreamy retro aesthetics, mood-first scenes. |
## Focal length
Measured in millimeters; controls two things at once — **how wide your field of view is**, and **how compressed your perspective looks.** Wide focal lengths make space feel deep. Long focal lengths make space feel flat and stacked.
| Category | Focal | What it does | Great for |
| -------------- | ----- | ----------------------------------------------------------- | ------------------------------------------------------------ |
| **Fish eye** | 8mm | Extreme wide with strong curved distortion. | POV action, skate videos, surrealism. |
| **Ultra wide** | 14mm | Very wide, mild distortion at the edges. | Tight interiors, big landscapes, immersive scenes. |
| **Ultra wide** | 18mm | Slightly more controlled ultra-wide. | Landscapes, architecture, environmental establishing shots. |
| **Wide** | 35mm | Documentary-style wide. Sees more than the eye does. | Environmental portraits, dialogue, anywhere context matters. |
| **Standard** | 50mm | Matches the natural perspective of the human eye. | Naturalistic narrative — the most versatile choice. |
| **Portrait** | 85mm | Slight background compression, beautiful subject isolation. | Portraits, close-ups, emotional reactions. |
| **Portrait** | 100mm | More compression, tighter framing, dreamier separation. | Beauty shots, intimate close-ups, hero moments. |
Focal length is the **single biggest lever** for how a scene "feels." Wide lenses (8–35mm) make spaces feel large, energetic, immersive. Long lenses (85–100mm) make spaces feel intimate, compressed, emotional. **When in doubt, start with 35mm or 50mm** — the workhorse focal lengths of cinema.
## Aperture
Controls **depth of field** — how much of your image is in sharp focus. Wide-open (small f-number) gives creamy background blur and razor-thin focus. Stopped-down (large f-number) keeps everything sharp from foreground to horizon.
| Category | Aperture | What it does | Great for |
| ---------------- | -------- | ------------------------------------------------------------- | ------------------------------------------------------------------- |
| **Wide open** | F/1.4 | Maximum background blur, very shallow focus. | Dreamy portraits, romantic scenes, isolating a subject from chaos. |
| **Near wide** | F/2.0 | Soft, blurred background with slightly more in-focus subject. | Standard portraits, interviews, intimate dialogue scenes. |
| **Middle range** | F/2.8 | Balanced look — subject sharp, background pleasantly soft. | Most cinematic scenes; the go-to default. |
| **Stopped down** | F/5.6 | More of the scene in focus, less background blur. | Group shots, documentary-style work, scenes with multiple subjects. |
| **Deep focus** | F/16 | Everything from foreground to horizon in sharp focus. | Landscapes, epic establishing shots, classic deep-focus cinema. |
## Camera Presets — a shortcut past the four dials
If you'd rather start from a recognizable look than tune camera, lens, focal, and aperture yourself, the camera picker has a second tab: **Camera Presets**, alongside **All Cameras**. Each preset is a named director- or film-style look that bundles all four dials into one click.
| Preset | Bundle |
| ------------------------- | ----------------------------------- |
| **A24 Prestige** | ARRI Alexa · Spherical Lens · 35mm |
| **Neon Fever** | ARRI Alexa · Anamorphic Lens · 85mm |
| **Denis Villeneuve Epic** | IMAX · Spherical Lens · 18mm |
| **Wong Kar-Wai Romantic** | ARRIFLEX · Vintage Lens · 85mm |
More presets are available in the gallery — including **Christopher Nolan** and **Terrence Malick Poetic** — each following the same pattern of a named look bundled from the same camera/lens/focal/aperture dials above.
Camera Presets are a fast starting point, not a replacement for the four dials — apply one, then still adjust individual settings from **All Cameras** if you want to nudge the look further.
## Putting it together
Reading tables one row at a time is useful — but the magic comes from combining them. Seven starter combos you can copy directly into a project.
**Camera:** ARRI Alexa · **Lens:** Spherical · **Focal:** 50mm · **Aperture:** f/2.0
The safe, classic cinema look. If you take one combo from this page, take this one.
**Camera:** IMAX · **Lens:** Anamorphic · **Focal:** 35mm · **Aperture:** f/2.8
Big scale, widescreen flares. Trailer energy, the big-screen feel.
**Camera:** ARRI Flex · **Lens:** Vintage · **Focal:** 85mm · **Aperture:** f/1.4
Grainy, soft, romantic. Reach for this when you want texture and warmth.
**Camera:** Red V Raptor · **Lens:** Anamorphic · **Focal:** 35mm · **Aperture:** f/2.0
Crisp, vivid, widescreen. Neon-on-wet-asphalt territory.
**Camera:** Sony FX6 · **Lens:** Spherical · **Focal:** 35mm · **Aperture:** f/5.6
Naturalistic, lots in focus. Honest, observational.
**Camera:** ARRI Alexa · **Lens:** Petzval · **Focal:** 85mm · **Aperture:** f/1.4
Swirly bokeh, soft glow. Fashion, beauty, hero portraits.
**Camera:** Red V Raptor · **Lens:** Fish eye · **Focal:** 8mm · **Aperture:** f/2.8
Wide, curved, immersive. POV action, dream sequences, anything that should feel "inside the moment."
# Creating Images
Source: https://docs.imagine.art/film-studio/create-image
Images are the foundation of everything you will do in Film Studio.
Images are the foundation of everything you will do in Film Studio. Even when your goal is a video, you will often start by generating images for individual frames, then bring them into the video tab as references.
## Two ways to start: Upload or Generate
When you open the Image tab, you have two starting points:
* **Upload media.** Use this when you already have an image you want to work with — a photograph, concept art, a frame from another video. Click *Upload media* in the optional media box and pick the file. You can also click *Select* to pull from media you have generated earlier in this project.
* **Generate from a prompt.** Skip the upload box and just write a description in the Prompt field. The system will create the image for you. Add reference media via *Add media* inside the prompt area if you want to influence style or composition.
These two are not mutually exclusive. The most common workflow is to **upload a reference image and then write a prompt** that tells the system what to change or what to use the reference for.
## Output controls — the three chips below the prompt
| Chip | What it controls | When to change it |
| -------- | ------------------------------------------------------------------------------------------------ | -------------------------------------------------------------------------------------------------- |
| **1/4** | Number of variations generated per click. Increase to compare options; decrease to save credits. | Increase when exploring; decrease when iterating on something you already like. |
| **16:9** | Aspect ratio. 16:9 is cinematic widescreen; portrait and square ratios also available. | Pick the ratio that matches the final delivery — social vertical, web horizontal, or print square. |
| **4K** | Output resolution. Higher means more detail and a larger file. | Use 4K for final deliverables. Drop it lower for fast iteration. |
## Storyboard mode
A storyboard is a sequence of frames that map out a film before you shoot it. Film Studio gives you a **Storyboard toggle** (just below the upload area on the Image tab) that lets you generate a full sequence of frames in one go, instead of generating each frame separately. When Storyboard is on, the system reads your prompt as a sequence of shots rather than a single frame, and produces a series of images that share a consistent style.
Use Storyboard mode **early.** Even if you only plan to deliver a single hero image at the end, generating a storyboard first helps you decide on framing, color palette, and pacing before you commit to a final look.
### Writing a storyboard prompt
There's no separate multi-shot UI — just describe each shot inline in the one prompt box, e.g. *"Shot 1: a lighthouse at dawn. Shot 2: a boat approaching the shore. Shot 3: a fisherman waving from the deck."* The system decides how many frames to actually produce, and can expand well beyond what you literally described — three described shots can come back as nine frames.
### The result is one grid image, not separate files
Storyboard output isn't delivered as individual images — it's a single contact-sheet-style grid image containing every frame.
To work with an individual frame — download it, reuse it, regenerate it — hover the grid and click **Split grid**. This produces independent, full-resolution images for every frame.
There's no in-place way to regenerate or reorder a single frame while it's still part of the grid. **Split grid first**, then treat each resulting image like any normal Create Image result.
### Sending a frame into Create Video
Once you've split the grid, any frame is available as a **Start Frame** or **End Frame** for a video. Open the **Create Video** tab, click **Select** under Start Frame (or End Frame), and switch to the **My Creations** tab in the picker — your split storyboard frames show up there alongside everything else you've generated.
## Saving & reusing presets
Consistency is the difference between a collection of nice images and a film. A preset captures camera, lens, focal length, and aperture under one name — apply it with a single click.
Once you find a camera/lens/focal/aperture combination that works for your project, **save it as a preset.** A preset captures all four settings under a name you choose, so you can apply it again with one click on a future image or video — without having to remember what you picked.
Saving presets matters because **consistency is what makes a series of images feel like a film.** If every frame of your project uses the same preset, your work will feel like it came from the same hand. Without that, even strong individual images can feel disconnected.
### How to save a preset
In the Camera panel, set the four values (camera, lens, focal length, aperture) the way you want them.
Click the **save preset button** under the camera values selection area
Apply it later from the same panel by selecting it from your saved presets list.
**Try this** — For your first project, build **exactly one preset** and use it for every image and every video. This single discipline will do more for the consistency of your work than any other technique in this guide.
# Creating Videos
Source: https://docs.imagine.art/film-studio/create-video
From frames to motion. Single-shot and multi-shot, up to 15 seconds.
From frames to motion. Same cameras, same lenses — plus how the camera moves, how time bends, which genre the result feels like, and how long each piece of the shot lasts. Two flavors: single-shot and multi-shot, up to 15 seconds.
## Single-shot videos
A single-shot video is **one continuous clip generated from one prompt.** This is the right starting point when you want a quick result, when your idea is one moment rather than a sequence, or when you are testing how a particular look behaves in motion.
### How to make a single-shot video
Open the **Create Video** tab in your project.
Optionally upload a **start frame** and an **end frame**. The system will treat these as the bookends of your shot.
Optionally add reference images using the *Add media* button. References influence style without dictating exact content.
Write your prompt. Describe subject, action, setting, and mood.
Set the cinematic controls (genre, camera movement, speed ramp, camera/lens/focal/aperture).
Hit **Generate.**
Start and end frames are powerful, but they constrain the result. Use them when you specifically want to control where the shot begins or ends. **If you just want the system to be creative, leave them empty.**
### Output settings
Aspect ratio and resolution live alongside the cinematic controls. Create Video's options are narrower than Create Image's — don't assume the two tabs match:
| Setting | Options |
| ---------------- | ------------------------------------ |
| **Aspect ratio** | 16:9 · 9:16 · 1:1 · 4:3 · 3:4 · 21:9 |
| **Resolution** | 720p · 1080p |
4K is available for still images in the **Image** tab, but not for video generation here.
## Multi-shot videos
A multi-shot video is a **single piece of footage made up of several distinct shots** that the system will stitch together into one continuous film. This is how you build a sequence — a scene followed by a closeup followed by a wide reveal — without combining clips manually.
**5** shots per video
**15 seconds** total
If you use all 5 shots, each averages 3 seconds. If your story needs a long establishing shot, you might use 1 shot for 8 seconds and 3 short shots of 2–3 seconds each. The choice is yours — but the total never exceeds 15 seconds and you never have more than 5 shots.
### How to build a multi-shot video
Open the **Create Video** tab and switch to multi-shot mode.
For each shot, write its own **prompt**. Treat each as a complete brief for that moment.
For each shot, set the **duration**. The interface will show how many seconds are left in your 15-second budget.
For each shot, you can also add reference media and adjust cinematic controls (camera movement, speed ramp, etc.).
Click the **"+"** icon in the scene timeline to add another shot. Drag shots to reorder.
When all shots are configured, hit **Generate.**
**Duration is the only per-shot setting.** Aspect ratio, resolution, genre, and speed ramp apply to the whole scene, not to individual shots within it — if you need different genres or ramps for different moments, put them in separate scenes rather than separate shots of the same scene.
## Adding audio
The timeline also has an **Add Audio** button for generating a soundtrack (AI Music) or narration (Text to Speech) directly onto your video — not just attaching a passive audio reference. See [Adding Audio](/film-studio/adding-audio) for the full walkthrough.
Write each prompt as if briefing a film crew on each scene of a one-page script. Keep your **subject and style consistent** across shots, but vary the framing and the action. **Wide → medium → close-up** is a classic three-shot rhythm that almost always works.
## Genres
Selecting a genre tells the system what **emotional and visual vocabulary** to draw from. Genres are stylistic shortcuts — they change lighting, color palette, pacing cues, and reference material the system uses. Pick one that matches the story you are telling.
*Fast, kinetic, high-energy.*
**Visuals:** Sharp contrast, bold color, hard light, motion blur. **Use for:** chases, fights, sports.
*Larger than life, journey-driven.*
**Visuals:** Wide vistas, golden hour, rich earthy palette. **Use for:** exploration, hero's journey moments.
*Tense, investigative, shadowed.*
**Visuals:** Low-key lighting, muted colors, hard edges. **Use for:** detective scenes, noir.
*Neon-soaked, urban dystopian future.*
**Visuals:** Magenta & cyan, rain, holograms, dense cities. **Use for:** sci-fi cities, hacker scenes.
*Grounded, period-accurate, classical.*
**Visuals:** Restrained color, soft natural light, era-correct detail. **Use for:** period pieces, biographical drama.
*Dread, unease, claustrophobic.*
**Visuals:** Deep shadows, desaturated palette, off-balance framing. **Use for:** scary scenes, suspense.
*Warm, intimate, soft.*
**Visuals:** Warm light, soft focus, pastel and golden palette. **Use for:** love scenes, tender moments, dreamy memories.
*Imaginative, otherworldly, "what if."*
**Visuals:** Surreal color, unusual physics, dreamlike composition. **Use for:** fantasy, sci-fi, magical realism.
*Psychological tension, simmering dread.*
**Visuals:** Cool tones, tight framing, restrained motion. **Use for:** suspense, slow-burn dread, paranoia.
*Dusty, sun-baked, mythic.*
**Visuals:** Warm yellows and oranges, harsh sun, wide vistas. **Use for:** westerns, frontier scenes, desert showdowns.
## Camera movements
How the camera moves is part of how your scene speaks. A pan reveals. A dolly emphasizes. A handheld engages. A static shot pauses. **Pick the movement that matches the feeling of the moment, not just the geography of it.**
| Movement | What it does | Use it when you want to… |
| ---------------------- | ----------------------------------------------------------------- | --------------------------------------------------------------------------------------------- |
| **Static** | No camera movement at all. | Hold a quiet moment, signal control or stillness, let action move within the frame. |
| **Camera follows** | The camera tracks alongside a moving subject. | Show someone moving through space — running, walking, driving — and stay with them. |
| **Dolly in** | The camera moves forward toward the subject. | Build tension, focus the audience, signal that something important is happening. |
| **Dolly out** | The camera moves backward, away from the subject. | Reveal context, signal isolation, end a scene or a thought. |
| **Dolly left / right** | The camera slides sideways while staying parallel to the subject. | Show parallel motion, reveal something hidden offscreen, add gentle motion to a static scene. |
| **Drone shot** | A high aerial view from above, often moving. | Establish scale, open a film, geographic reveal. |
| **Handheld** | Camera operated by hand — slightly shaky, organic. | Add intimacy, urgency, or chaos. Documentary and action. |
| **Jib up** | The camera rises smoothly on a crane. | Build hope, reveal what is above, end on a sweeping note. |
| **Jib down** | The camera descends smoothly on a crane. | Reveal what is below, descend into a scene, ground the audience. |
| **Orbit around** | The camera circles the subject. | Examine a subject from all sides, emphasize importance, add motion to a static figure. |
| **Pan left / right** | The camera rotates horizontally on a fixed axis. | Follow action across a space, reveal something to the side, scan an environment. |
| **Tilt up** | The camera angles upward. | Convey grandeur, scale, intimidation, or hope. |
| **Tilt down** | The camera angles downward. | Convey power dynamics, reveal something below, end on a grounded note. |
| **Zoom in** | The lens zooms in toward the subject. | Snap audience attention, mark an emotional beat, hint at obsession. |
| **Zoom out** | The lens zooms out away from the subject. | Provide context, reveal scale, isolate a subject within a larger world. |
## Speed ramps
A speed ramp is a **controlled change in playback speed** over the course of a shot. Speed ramps are a hallmark of modern cinema — used to emphasize an impact, stretch out a beautiful moment, or accelerate through a transition.
| Speed ramp | What it does | Use it when… |
| ------------ | --------------------------------------------------------------------------------------- | -------------------------------------------------------------------------------------------- |
| **Linear** | Constant speed for the entire shot. No ramping. | You want a natural, even pace — the default for most scenes. |
| **Slo Mo** | The whole shot plays in slow motion. | You want a beautiful, dramatic moment — the lead-up, the reveal, the kiss, the punch. |
| **Speed Up** | The whole shot plays faster than real time. | You want energy, urgency, compressed time — montages, transitions, action. |
| **Impact** | The shot ramps sharply at a key moment — typically slow on either side of a fast burst. | You want to emphasize one specific frame inside the shot — the hit, the explosion, the snap. |
| **Custom** | You define your own speed curve across the shot. | You have a specific rhythm in mind that the presets do not match. |
Speed ramps are best used **sparingly.** If every shot uses Slo Mo or Impact, none of them will feel special. Use Linear as your default, and bring in a ramp on the one or two shots where you really want to draw attention.
## Scene & shot timing
At the bottom of the workspace you will see your timeline — one card per scene, with each shot laid out inside the scene. The timeline is where you control pacing.
From the timeline you can:
* **Adjust the duration** of any individual shot by dragging its edges.
* **Reorder shots** by dragging the cards.
* **Add new shots** with the "+" button.
* **Delete** shots you no longer want.
* **Rename scenes** for your own organization.
Remember the **15-second total ceiling.** The timeline shows how much of your budget each shot is consuming. If you try to extend a shot past your remaining budget, you will need to shorten another shot first.
# Drama Studio
Source: https://docs.imagine.art/film-studio/drama-studio
A second project type inside Film Studio, built for vertical, multi-episode series — persistent characters, episode structure, platform-ready output.
**Drama Studio** is a distinct project type inside Film Studio, purpose-built for short-form vertical drama series — the kind of content that runs on ReelShort, TikTok, and DramaBox. It shares Film Studio's editor shell, but the project structure underneath is genuinely different: persistent characters that carry across episodes, an episode/scene hierarchy, and a fixed vertical format.
Click **Create Project** from the Film Studio home page. The **Create Project** dialog asks what you want to create:
* **Film** — a cinematic short with full director-level controls.
* **Drama** — a vertical, multi-episode series for ReelShort, TikTok & DramaBox, with **persistent cast** and **platform-ready** output.
Selecting Drama opens an editor with **Create Image**, **Create Video**, and **Edit Video** tabs — the same shell as a Film project, but everything inside is oriented around building a series: an **Episodes** panel on the right (starting with Episode 1), and a **Scene Builder** built around **Characters**, **Props**, and **Locations**.
The **Episodes** panel lists your episodes and an **Add Episode** row. Adding one simply appends the next sequentially-numbered episode — there's no separate naming or synopsis prompt at creation time. Each episode has its own **Scene 1 / Shot** timeline underneath, but the character/prop/location library itself is shared across every episode (and every project), not scoped per-episode.
Click **Characters** (in Scene Builder or the References panel) to see your character library — a real preset roster (Nova, Atlas, Pixel, Luna, and more) alongside **All / Saved / My Characters** filters. Click the **+** tile to create your own: upload reference photos (up to 10 images), give it a **Name**, pick **Male / Female / Neutral**, and optionally add instructions describing the character further.
A character created this way is reusable by name across every episode and scene — this is the "persistent cast" the project type is built around. **Props** and **Locations** work the same way (reference photo + name + optional description), just without the gender field.
**Create Video** inside Drama Studio looks like Film Studio's, with real differences: aspect ratio is fixed **vertical (9:16)** with no ratio picker, duration defaults to **3s** per shot, resolution defaults to **720p**, and **Style / Genre / Movement all default to Auto** rather than exposing Film Studio's full genre/movement library. Camera controls (camera body, lens, focal length, aperture) are still present underneath, so you aren't fully locked out of cinematographic control. **Add Audio** works the same way as regular Film Studio, directly on the scene timeline.
The video model Drama Studio generates with isn't shown anywhere in the interface — Style, Genre, and Movement all default to Auto, and the model is selected automatically behind the scenes.
## Browsing a real example
The Film Studio home page's project list has **All / Drama / Film** filter tabs. The **Drama** filter surfaces example series like "Don't Tell My Husband" — a set of individually playable episode clips (S01E01, S01E02...), each with its own synopsis, forming a continuing storyline. These are playable showcase examples, not projects you can open and inspect scene-by-scene.
Watch a couple of episodes from the showcase examples before building your own series — the pacing (roughly 1-1.5 minutes per episode) and per-episode cliffhanger structure are the pattern platforms like ReelShort and DramaBox actually expect.
# Editing an Existing Video
Source: https://docs.imagine.art/film-studio/edit-video
For when you already have footage and want to redirect it. Provide the video, write a prompt describing what should change — the system produces a new version of the clip that follows your instruction. The original footage shapes the result.
## When to use Edit Video
* You have a generated clip that is close but not quite right (wrong color, wrong mood, wrong genre).
* You have a real video and want to apply a stylized look to it.
* You want to keep the composition and motion of an existing shot but change its subject or setting.
## How to edit a video
Open the **Edit Video** tab.
Provide the source video. Upload a file or pick one of your existing generations from the project library.
Write a prompt that describes what you want to change. **Be specific about what should stay the same and what should change.**
Optionally add reference media to influence the new style.
Set the cinematic controls — apply a new genre and new camera movements just as you would for a brand-new video.
Hit **Generate.** You will get back an edited version of your clip.
Edit Video works best when you give it **one clear instruction at a time.** *"Change the time of day to sunset"* produces a cleaner result than *"Change the time of day to sunset, add a horse, make it look like a Western, and slow it down."* If you have multiple changes to make — edit, review, then edit again.
## Tips for consistent results
* Use the **same camera preset** you used for the original — this keeps the visual style consistent across the edit.
* Keep your prompts focused on **the change**, not on re-describing the entire scene.
* Use **reference images** when language alone isn't enough — a mood board is faster than a long description.
# Extending a Video
Source: https://docs.imagine.art/film-studio/extend-video
For when you have a video you love and want more of it. Provide the source clip, write a prompt describing what should happen next — the system generates new frames that continue naturally from where the original left off.
## When to use Extend
* You generated a clip that ends too soon and you want to see more of the same.
* You have a 5-second moment and want it to play out for 10 seconds.
* You want a sequel-style continuation — **same subject, same world, what happens next.**
## How to extend a video
Open the **Extend** tab.
Provide the source video — upload a file or pick from your project library.
Write a prompt that describes what should happen in the extension. The system will use the end of your original clip as the starting point, and your prompt as the direction forward.
Optionally add reference media and tune cinematic controls (genre, movements) for the extended portion.
Hit **Generate.** You will receive a new clip that continues from the original.
Extend is **forward-only.** It continues from the end of your source, not before the start. If you want a lead-in to a clip, generate that as a separate shot and assemble the two in a multi-shot video.
## Tips for natural-feeling extensions
* Describe an action that naturally follows the end of your source. *"She turns and walks away"* is easier to extend from a clip of someone standing than *"she suddenly flies into space."*
* **Reuse your camera preset** so the optical character matches across the seam.
* Keep the lighting consistent in your prompt — sudden changes in lighting are the most common reason extensions feel disjointed.
# Your First Project
Source: https://docs.imagine.art/film-studio/getting-started
Follow this section once and the rest of the guide will make a lot more sense.
Follow this section once and the rest of the guide will make a lot more sense.
## Step 1 — Open Film Studio
Open your browser and go to **imagine.art / film-studio**.
Sign in if you are not already logged in.
You will land on the Film Studio home view, which lists your existing projects on the left side.
## Step 2 — Start a new project
Click the **"+"** button next to *Projects* in the top-left corner.
Give your project a name. A descriptive name (e.g. *"Cyberpunk Trailer V1"*) is easier to find later than the default *"Untitled Project"*.
Press **Enter** or click outside the field to save.
Every piece of work you create — images, videos, storyboards, presets — lives inside a project. **Make a habit of starting a new project for each idea**, the same way a film team would treat each production as its own folder.
## Step 3 — Open the project
Click the project name in the left sidebar. The main workspace opens. You are now ready to create — the next section walks through what each part of the workspace does.
# Glossary
Source: https://docs.imagine.art/film-studio/glossary
A quick reference to terms used throughout this guide and inside Film Studio.
A quick reference to terms used throughout this guide and inside Film Studio.
| Term | Meaning |
| -------------------- | ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| **AI Music** | Generates a soundtrack directly onto your video timeline from a text description, powered by ElevenLabs Music. See [Adding Audio](/film-studio/adding-audio). |
| **Anamorphic** | A lens type that produces widescreen cinematic images with characteristic horizontal flares and oval-shaped bokeh. |
| **Aperture** | The opening in a lens that controls how much of the image is in focus. Smaller f-numbers = more background blur. |
| **Aspect ratio** | The shape of the image, expressed as width:height. 16:9 is widescreen; 9:16 is vertical / social. |
| **Bokeh** | The visual quality of out-of-focus areas in an image. |
| **Camera Preset** | A named, director-style look (e.g. "A24 Prestige") that bundles a camera/lens/focal/aperture combination for one-click apply. See [Camera Presets](/film-studio/camera-controls#camera-presets-a-shortcut-past-the-four-dials). |
| **Creation Actions** | The one-click regeneration tools (Upscale, Remove Bg, Variate, Pan) available on any asset in the [Assets Library](/film-studio/assets-library). |
| **Depth of field** | How much of the image, from front to back, is in sharp focus. |
| **Dolly** | A camera move along the ground — forward, backward, left, or right. |
| **Focal length** | Measured in millimeters; controls field of view and perspective compression. |
| **Genre** | A stylistic category that informs lighting, color, and reference material the system uses. |
| **Handheld** | A camera operated by hand, with characteristic small movements. |
| **IMAX** | A large-format film camera known for massive resolution and epic scale. |
| **Jib** | A camera mounted on a crane that moves vertically (up or down). |
| **Multi-shot video** | A video made of multiple distinct shots stitched into one clip — up to 5 shots and 15 seconds total in Film Studio. |
| **Pan** | A horizontal camera rotation on a fixed axis. |
| **Preset** | A saved combination of camera body, lens, focal length, and aperture — reusable across images and videos. |
| **Reference media** | Images or clips you attach to a prompt to influence the style or content of the result. |
| **Scene** | An organizational unit in the timeline. A scene contains one or more shots. |
| **Shot** | A single continuous piece of footage within a scene. |
| **Speed ramp** | A controlled change in playback speed across a shot. |
| **Spherical lens** | The standard, naturalistic cinema lens family — the default optical look. |
| **Split grid** | The action that breaks a Storyboard's single grid image into independent, full-resolution frames. See [Storyboard mode](/film-studio/create-image#storyboard-mode). |
| **Storyboard** | A sequence of frames that plan out a film before it is produced. |
| **Text to Speech** | Generates narration directly onto your video timeline from typed or AI-written text, powered by ElevenLabs. See [Adding Audio](/film-studio/adding-audio). |
| **Tilt** | A vertical camera rotation on a fixed axis (up or down). |
| **Zoom** | A change in focal length within a shot — optical rather than physical motion. |
# Pro Tips & Best Practices
Source: https://docs.imagine.art/film-studio/pro-tips
Habits that lift your output from 'good generation' to 'good film.'
Once you are comfortable with the basics, these habits will lift your output from "good generation" to "good film."
## Prompting
* Write prompts the way a director briefs a crew: **subject, action, setting, time of day, mood.**
* Mention only what matters. Long prompts with conflicting instructions confuse the result.
* Lean on references for things hard to describe in words — a shade of teal, a face, a silhouette.
## Cinematography
* Pick a camera/lens/focal/aperture preset for your project and **stick to it.**
* Default to **35mm or 50mm.** Reach for ultra-wide or long lenses when the story specifically calls for it.
* Aperture **f/2.0–f/2.8** covers most narrative scenes. Use f/1.4 for romance and dreams; f/16 for landscapes and epic reveals.
## Video structure
* **Build storyboards before videos.** Cheap to iterate, easy to share.
* Use the **wide → medium → close-up** rhythm for multi-shot scenes. Classic for a reason.
* Save your strongest budget for the moment that matters most. Not every shot has to be the same length.
## Production rhythm
* Work in iterations. **Four variations → pick one → refine → move on.** Trying to nail it in one click is slower.
* Name projects, scenes, and presets descriptively. *"Hero Shot V3"* beats *"Untitled."*
* Use the **Assets library.** Anything you generate is available to reuse.
## Common mistakes to avoid
Loading **every cinematic control** at once on your first try. Start with prompt + camera preset, then add movements and ramps.
Asking for too many things in a single edit. **One change per pass.**
Ignoring the **15-second ceiling** on multi-shot videos until the timeline forces you to. Plan your time budget before generating.
# Quick Reference
Source: https://docs.imagine.art/film-studio/quick-reference
One-page summary — all the key dials, limits, and disciplines in one place.
One-page summary you can return to whenever you forget which dial does what.
## Starting fresh
1. Open **imagine.art / film-studio** → click **"+"** next to *Projects* → name it → click in.
2. Pick a tab: **Image · Create Video · Edit Video · Extend.**
## The four image dials
| Dial | Options |
| ------------ | ------------------------------------------------------------------------------- |
| **Body** | Texture · Alexa, FX6, Raptor, ARRI Flex, IMAX |
| **Lens** | Optical · anamorphic / fish eye / macro / petzval / probe / spherical / vintage |
| **Focal** | FoV · 8mm fish eye → 100mm portrait |
| **Aperture** | Depth · f/1.4 dreamy → f/16 deep focus |
Or skip the dials — the **Camera Presets** tab bundles a look in one click (A24 Prestige, Neon Fever, Denis Villeneuve Epic, and more).
## Video limits
| | |
| ---------------- | ----------------------------------------------------- |
| **Max shots** | 5 per video |
| **Max length** | 15 seconds total |
| **Mode** | Single-shot for one moment; Multi-shot for a sequence |
| **Aspect ratio** | 16:9 · 9:16 · 1:1 · 4:3 · 3:4 · 21:9 |
| **Resolution** | 720p · 1080p (4K is Image-tab only) |
## Video controls
| Control | Options |
| -------------- | ---------------------------------------------------------------------------------------------------------------- |
| **Genre** | Picks the visual vocabulary |
| **Movement** | Pan · tilt · dolly · jib · orbit · drone · handheld · static · zoom |
| **Speed ramp** | Linear · Slo Mo · Speed Up · Impact · Custom |
| **Audio** | AI Music or Text to Speech, generated straight onto the timeline — see [Adding Audio](/film-studio/adding-audio) |
Multi-shot note: only **duration** is per-shot — aspect ratio, resolution, genre, and speed ramp apply to the whole scene.
## Editing & extending
| Tool | How it works |
| -------------- | --------------------------------------------------- |
| **Edit Video** | Existing video + prompt → modified version |
| **Extend** | Existing video + prompt → continues forward in time |
## Storyboard & Assets
| | |
| --------------------- | ----------------------------------------------------------------------------------------------------------------------------------------------- |
| **Storyboard output** | One grid image — use **Split grid** to get individual frames |
| **Assets Library** | Full generation history, prompts, comments, and one-click Upscale/Variate/Remove Bg/Pan — see [The Assets Library](/film-studio/assets-library) |
## Discipline that pays off
| | |
| -------------- | ------------------------------------------------------- |
| **One preset** | Build one per project. Use it everywhere. |
| **Storyboard** | Before you video. |
| **One change** | Per edit pass. |
| **Budget** | Save 15-second budget for the moment that matters most. |
**Safe default:** ARRI Alexa · Spherical · 50mm · F/2.0 — the classic cinema look. When in doubt, start here.
# References
Source: https://docs.imagine.art/film-studio/references
Build a persistent library of characters, products, images, videos, and audio for consistent visual continuity across your film.
The **References** library in Film Studio is a persistent collection of assets — characters, products, images, videos, and audio — that you can attach to any scene to maintain visual and auditory continuity throughout your project.
References carry over across the entire film. Add a character once and use them in every scene without re-uploading or re-describing.
## Reference types
| Type | Use case |
| ------------- | ---------------------------------------------------------------------------------- |
| **Character** | A person, creature, or avatar — keep their appearance consistent across all scenes |
| **Product** | A physical product — maintain accurate shape, color, and texture throughout |
| **Image** | Any static visual reference — props, locations, mood boards |
| **Video** | A motion reference — use for style, movement, or continuity reference |
| **Audio** | A voice, music track, or sound effect — attach to scenes for audio continuity |
## Adding a reference
Click **References** in the Film Studio sidebar or top toolbar.
Click **+ Add Reference** and select the type (Character, Product, Image, Video, or Audio).
Upload the relevant files. For characters and products, uploading multiple angles improves consistency — front, side, three-quarter, and back.
Give the reference a clear name and an optional description. The name is how you'll reference it when attaching it to scenes.
Click **Save**. The reference is now available in your library for the entire project.
## Attaching references to scenes
Once saved, you can attach a reference to any scene from the **Camera Controls** or **Create Video** panel — look for the **References** slot in the scene settings. You can attach multiple references to a single scene.
## Managing your library
References persist for the lifetime of the project. You can edit, replace, or delete any reference from the References panel at any time. Changes take effect on all future generations that use that reference — previously generated scenes are not affected.
For character references, upload a neutral expression frontal shot as the primary image. Add smiling, profile, and three-quarter shots as additional references. The more angles you provide, the more reliably the model maintains your character's look across different scene contexts.
# What is Film Studio
Source: https://docs.imagine.art/film-studio/what-is-film-studio
A complete production environment for visual storytellers — the controls of a real cinematographer, paired with generative AI.
ImagineArt Film Studio is a complete production environment for visual storytellers. Whether you are a solo creator with a single idea or a team building a multi-shot trailer, Film Studio gives you the tools that a traditional cinematographer would use — **cameras, lenses, focal lengths, apertures, movements, genres** — and pairs them with generative AI so you can produce production-ready images and films from a prompt.
## What you will be able to do
Seven concrete capabilities — the ones that take you from blank workspace to multi-shot film with a consistent look.
1. Start a new project on **imagine.art / film-studio** and find your way around the workspace.
2. Generate a single image from a prompt and control its look using **camera, lens, focal length, and aperture**.
3. Build a **multi-scene storyboard** for any production.
4. Generate a single-shot video, then graduate to **multi-shot films up to 15 seconds** long.
5. Apply **genres, camera movements, and speed ramps** to give your footage a deliberate, cinematic feel.
6. Edit an existing video with a prompt, and extend an existing video into a longer sequence.
7. Save your favorite camera setups as **presets**, so your work has a consistent look across projects.
**Building a vertical drama series instead?** See [Drama Studio](/film-studio/drama-studio) — a second project type inside Film Studio built for multi-episode, platform-ready series with a persistent cast, rather than a single cinematic short.
**A note on philosophy** — Film Studio is a director's tool, not a slot machine. The more deliberate your direction — your camera choice, your aperture, your shot rhythm — the more deliberate your result. This guide is structured to help you make those choices with intention.
## How this guide is organized
The guide follows a simple principle: **learn the smallest useful thing first, then build on it.** We start with the workspace and a basic image, then add cinematographic controls, then move into video, and finally cover editing and extending your work.
* **Getting Started & Workspace** cover orientation: what Film Studio is, what you need, and how the workspace is laid out.
* **Creating Images** covers image creation — the foundation for everything else, because a great frame is the seed of a great shot.
* **Creating Videos** covers video creation — single-shot and multi-shot workflows.
* **Edit & Extend** cover tools for refining and lengthening existing video.
* **Pro Tips, Glossary, and Quick Reference** are reference material you can return to at any time.
**Read this first** — If you only have ten minutes, read [Getting Started](/film-studio/getting-started) and [Understanding the Workspace](/film-studio/workspace). That alone is enough to start producing. Come back to the other sections when you want more control.
## Before you begin
Film Studio runs entirely in your browser. There is nothing to install — but a few minutes of preparation will save you frustration later.
A modern browser — **Chrome, Edge, Safari, or Firefox**, kept reasonably up to date.
A **stable connection** — generation runs on our servers, so a steady link makes for a smoother experience.
An **ImagineArt account** — sign up or log in at imagine.art before opening Film Studio.
Any **photos, mood-board images, or video clips** you want to influence the look. Keep them on your computer so you can upload quickly.
## A useful mindset
Film Studio rewards specificity. A prompt like *"a person walking"* will get you something generic. A prompt like *"a tired detective in a long beige coat walking down a wet alley at 2 a.m., neon reflections in puddles"* will get you something cinematic.
As you go through this guide, you will learn how the camera, lens, focal length, aperture, genre, and movement controls let you be even more specific — but the prompt is always the foundation.
Treat each prompt the way a director would brief a crew. Tell the system **who is in the shot, what they are doing, where they are, what time of day it is, and what mood you are after.** The richer the brief, the closer the result.
# Understanding the Workspace
Source: https://docs.imagine.art/film-studio/workspace
Three areas, four core actions, one top bar. Learn this once and every later section gets faster.
Three areas, four core actions, one top bar. Learn this once and every later section gets faster.
## The four core actions
Across the top of the left panel you will see four tabs. These are the four things you can do in Film Studio, and the rest of this guide is organized around them:
Generate or upload still images, with full cinematographic controls.
**When:** Designing a frame, building a storyboard, or producing key art.
Generate a video from prompts — single shot, or multi-shot up to 15 seconds.
**When:** You want to produce new motion footage from scratch.
Take an existing video and re-direct it with a prompt — change the look, the genre, or the motion.
**When:** You have footage that is close but not quite right.
Take an existing video and continue it forward — new frames that follow on naturally.
**When:** You have a clip and want a longer version of it.
## The three workspace areas
### Control panel (left)
Set up your shot — prompt, references, camera, lens, aspect, resolution. Hit Generate.
### Canvas (center)
Your viewer. Generated images & videos appear here. Zoom, compare versions, inspect detail.
### Scene timeline (bottom)
Where multi-shot work lives. Add, rename, reorder, and time your shots.
## The top bar
The top bar gives you account-level controls — **project navigation, search, contact sales, plan upgrades, notifications, and your profile**. The *Assets* button on the right opens a library of all the media you have generated or uploaded in this project, so you can reuse anything you have already made — see [The Assets Library](/film-studio/assets-library) for the full depth of what's in there (generation prompts, comments, one-click regeneration tools).
## Reading the left panel
The **left panel changes based on the tab you are in**. In *Image*, you see prompt, references, output chips (variations, aspect, resolution), and the camera panel. In *Create Video*, you also see motion controls — genre, movement, speed ramp — plus optional start and end frames. *Edit* and *Extend* require a source video before the prompt becomes active.
## Reading the timeline
The bottom strip is your **scene timeline**. When you work on a single image, you will only see one scene. When you build a multi-shot video, each shot appears as a card you can drag to reorder, resize to retime, or click to focus.
**Try this** — Before you read further, open Film Studio, create a project called *"Practice Project,"* and click through each of the four tabs (Image, Create Video, Edit Video, Extend). You do not need to generate anything yet — just notice how the left panel changes based on the tab. That five-minute tour will make the rest of this guide much easier to follow.
**Keep in mind** — Anything you make in any tab is automatically available in the others. Generate an image, then drop it as a start frame for a video. Render a video, then use Extend to continue it. The four actions are designed to compose.
**Where things live** — All output is filed under the current project. Use the **Assets** button (top right) to browse everything you have made, grouped by date. See [The Assets Library](/film-studio/assets-library) for the full detail view, prompt history, and one-click regeneration tools available from there.
**Team folders vs. personal folders** — The **All team creations** dropdown (top left) isn't just a project switcher. It also holds **Team folders** and **Private folders** you can create to organize projects — every project belongs to one of these folders, shown as a **Located** chip in that project's asset details.
# Generate Audio
Source: https://docs.imagine.art/generate-audio
Audio Nodes in ImagineArt Workflows enable you to generate professional spoken voice, original music, and custom sound effects entirely from text. Whether you need a voiceover for video, a background track, or sound effects for animation, these AI-powered nodes produce broadcast-quality audio that integrates seamlessly into your creative pipeline.
## How to Add Audio Nodes
1. Click the Add (+) button on the left toolbar in the workflow canvas.
2. Select Audio under node categories.
3. Choose from the available nodes listed below.
You can also double-click anywhere on the canvas and search for any audio node by name.
## Generation Nodes
These nodes create audio content from text prompts.
Convert text into natural-sounding speech. Write what you want spoken, select a voice character, and generate realistic audio. Supports multiple voices and models including ElevenLabs TTS and Minimax 2.8 HD.
Generate original music tracks from text descriptions or lyrics. Describe the mood, genre, and instruments or write full lyrics and the AI composes a complete track. Powered by ElevenLabs Music.
Generate custom sound effects from text descriptions. Describe any sound like rain, footsteps, explosions, sci-fi ambiance and the AI creates a matching audio clip. Powered by ElevenLabs Sound Effects.
## Combining Nodes
Audio nodes are designed to feed into the rest of your workflow. Here are some common combinations:
* **Voice → Combine Audio & Video**: Generate a voiceover and layer it onto an AI-generated video.
* **Music → Combine Audio & Video**: Create a custom soundtrack and pair it with a product video or brand reel.
* **Sound Effects → Combine Audio & Video**: Add ambient sounds or action effects to animated scenes.
* **Voice → Lipsync**: Generate speech, then use it as the audio input for a Lipsync node to create a talking-head video.
* **Text Iterator → Voice**: Enter multiple scripts and generate a separate voiceover for each — perfect for multilingual content or personalized outreach.
* **Prompt → Music + Prompt → Generate Video → Combine Audio & Video**: Create both a video and its soundtrack from text, then merge them into a finished piece.
# Generate Image
Source: https://docs.imagine.art/generate-image
## Overview
The Generate Image Node leverages advanced AI models to create high-quality images from text descriptions or by transforming existing images. With support for multiple models and customizable parameters, you can generate everything from photorealistic landscapes to stylized artwork tailored to your exact specifications.
Each model offers distinct capabilities and settings, so experimenting with different configurations helps you achieve your creative vision. You maintain full control over resolution, style, strength, and other parameters to fine-tune the output.
## Summary
The Generate Image Node is a powerful tool for generating high-quality images using AI models. It allows you to select from a variety of models, each with its own unique set of parameters and properties. Whether you're creating realistic landscapes, abstract art, or product visuals, this node gives you the flexibility to fine-tune the output according to your needs.
Each AI model has different variables and settings, so it's important to explore and experiment with the options available. You have full control over the properties, ensuring that the generated image aligns perfectly with your creative vision.
## How to Use
Select an AI model that suits your creative vision. Each model has unique properties that influence the style and detail of your generated image.
Adjust settings like resolution, image size, and style to match the exact look you want. Some models offer extra settings like strength (how much the input influences the image) or seed (for consistent results).
**Text to Image:** Write a prompt describing the image you want (e.g., "A vibrant cityscape at sunset").
**Image to Image:** Upload an existing image and describe how you want to transform it (e.g., "Change the background to a futuristic city skyline").
Click Generate, and the AI model will process your input and create the image based on your settings.

### Choosing the Right Settings
| Setting | Type | Impact on Output |
| ---------------- | ------------------------------------------------ | ----------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Prompt | Text | The text prompt defines the visual elements and context, guiding the AI model to generate the corresponding image. |
| Style | Preset | Select a stylistic preset (e.g., 3D, Cyberpunk, Watercolor, etc.) to influence the rendering style of the image, such as textures, color schemes, and artistic techniques. |
| Strength | 0-100% | Controls the influence of the text prompt on the image. A higher strength increases the effect of the prompt over the generated image. |
| Image Size | Landscape 16:9, Portrait 4:3, 1:1 (square), etc. | Defines the output resolution. |
| Seed | Seed | A fixed number that ensures reproducible results. Set a seed if you want the same output across different generations, while keeping all parameters the same. |
| Guidance Scale | 0%-20% | Defines how strongly the text prompt influences the output. A higher guidance scale ensures that the text has more influence over the final image while balancing randomness. |
| Resolution | High (4K)/Medium (2K)/Low (1K) | Determines the output resolution of the image. Higher resolution provides more detail but may take longer to generate. |
| Prompt Optimizer | Button (Action) | Enhances and refines your existing prompt. It can either generate a new prompt from scratch or improve the current one by analyzing its effectiveness for better results. |
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Generate Music
Source: https://docs.imagine.art/generate-music
### Summary
The Music Node generates original music tracks from a text prompt or lyrics. Describe the mood, genre, tempo, or instruments you want, or write full lyrics, and the AI composes a complete audio track. Use it to create background music for videos, jingles for ads, or full songs for creative projects, all within your workflow.
### How to Use
Click the Add (+) button and select Generate Music from the Audio node category.
Describe the music you want (e.g., "Upbeat lo-fi hip hop beat with soft piano and vinyl crackle") or write full lyrics for a vocal track. You can also use Copilot Node or connect the Prompt/Text handle.
Click Run, and the AI generates an original music track based on your input.
### Sample Use Cases
Generate a custom soundtrack and connect it to a Combine Audio & Video node to add music to any video in your workflow, no licensing headaches.
Describe your brand's tone and energy, and generate a short jingle or audio signature for ads, intros, or social media content.
Write full lyrics as your prompt and generate a complete song. Experiment with different genres and moods by adjusting your description.
# Generate Video
Source: https://docs.imagine.art/generate-video
## Summary
The Generate Video Node is the primary tool for creating videos using AI models in Workflows. It supports both text-to-video and image-to-video generation, giving you the flexibility to produce video content from a written description, a reference image, or a combination of both. You can choose from a wide range of AI models, each with its own strengths in style, motion quality, and output fidelity, and fine-tune settings like duration, resolution, aspect ratio, and more.
Each AI model has different available parameters and supported settings, so it's worth experimenting with the options to find the combination that best fits your creative vision.
## How to Use
Select an AI video model from the Model dropdown. Each model offers different strengths—some excel at cinematic motion, others at realistic physics or fast iteration. See the Video Models page for a full comparison.
**Text-to-Video:** Connect a Prompt or AI Copilot node and describe the scene you want (e.g., "A drone shot flying over a misty mountain range at sunrise").
**Image-to-Video:** Connect an image from an Import, Generate Image, or Edit Image node to use as the first frame or visual reference, and add a prompt to describe the desired motion.
Adjust Duration, Resolution, Aspect Ratio, and other parameters to match your output requirements.
Click Run, and the AI model will process your inputs and produce a video based on your settings.
## Multi-Shot (Kling models)
Multi-Shot lets you chain multiple independently authored shots inside a single Kling generation, producing one continuous video sequence without manual stitching.
Multi-Shot is available on Kling models only. Look for the **Multi-Shot** toggle inside the node settings when a Kling model is selected.
**How it works:**
In the Generate Video node, choose any supported Kling model from the model dropdown.
Toggle **Multi-Shot** on in the node settings panel. Shot slots will appear — Shot 1, Shot 2, Shot 3, and so on.
Write a separate prompt for each shot. Every shot has its own creative direction and is independently wirable — you can connect different upstream nodes to different shot inputs.
Each shot supports a duration between **3 and 9 seconds**. Adjust each shot individually to control pacing across the sequence.
Use the shared **Start Frame** and **Last Frame** inputs to lock visual continuity at the beginning and end of the full sequence. These apply across all shots in the generation.
Click **Run**. All shots are processed as a single Kling job and rendered as one continuous video timeline.
Use the **Start Frame** and **Last Frame** controls to anchor your sequence visually — especially useful when chaining multi-shot clips into a longer narrative.
## Elements in Workflows
Elements are a persistent reference library for characters, objects, or any subject you want to keep visually consistent across shots. Once created, an Element can be referenced in any workflow via `@mention` in your prompt.
See [Elements](/video-tools/elements) for full setup instructions, including how to create an Element and manage your library.
**Using an Element in a Generate Video node:**
Type `@` in any prompt field and pick the Element from your library. It appears as a chip in the prompt and tells the model to lock that subject's appearance for the generation.
## Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Generate Voice
Source: https://docs.imagine.art/generate-voice
## Summary
The Voice Node converts text into natural-sounding speech using AI text-to-speech models. Write or connect a prompt, select a voice, and the node generates an audio output that sounds like a real person speaking. Use it for voiceovers, narration, dialogue, or any workflow that needs spoken audio from text.
## How to Use
Click the Add (+) button and select Voice from the Audio node category.
Type what you want spoken into the prompt field, or connect a Prompt or AI Copilot node.
Choose a voice from the Voice dropdown (e.g., Roger). Each voice has a distinct tone, pitch, and character.
Click Run, and the AI generates an audio file of the selected voice speaking your text.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| ---------------- | ---------------------- | ----------------------------------------------------------------------------------------------------------------- |
| Voice | Dropdown (e.g., Roger) | Selects the voice character used for speech generation. Different voices vary in tone, gender, accent, and style. |
| Prompt | Text Input | The text content that will be converted into speech. |
| Stability | Slider (0–100%) | Controls how consistent the voice sounds across the output. |
| Similarity Boost | Slider (0–100%) | Controls how closely the output matches the selected voice. |
| Speed | Slider | Adjusts the speaking pace of the generated audio. |
| Timestamps | Checkbox | When enabled, returns word-level or sentence-level timestamps alongside the audio output. |
## Sample Use Cases
Generate a voiceover and connect it to a Combine Audio & Video node to add narration to any video in your workflow.
Write the same script in multiple languages and generate a voice for each — perfect for localizing video content without hiring voice actors.
Quickly generate audio previews of scripts, blog posts, or ad copy to hear how they sound before recording with a real voice.
# Generative Fill (Node)
Source: https://docs.imagine.art/generative-fill-node
## Summary
The Generative Fill Node fills, replaces, or extends part of an image using AI, based on a masked region and a prompt. It's the workflow-canvas equivalent of the Generative Fill tool available elsewhere in the app's image editing surfaces (for example, in the Photoshop and Figma plugin panels) — brush a mask, describe what should appear there, and the node generates content that blends into the rest of the image.
## How to Use
Click the Add (+) button and select **Generative Fill** from the Image node category.
Link an image from another node or upload one, then mark the region you want to fill, replace, or extend.
Write a prompt describing what should appear in the masked area.
Click Run. The node outputs the image with the masked region filled according to your prompt.
This node's full settings panel wasn't fully captured during this pass — if you hit a setting not covered here, treat this page as a starting point and flag it for a follow-up pass.
## Sample Use Cases
Mask an unwanted object in a generated or uploaded image and describe what should replace it.
Mask the edge of an image and describe what should continue into the newly added space.
# Getting Started
Source: https://docs.imagine.art/getting-started
Create your ImagineArt account, choose a plan, and generate your first image or video in minutes.
## Create your first Image
Go to [imagine.art](https://www.imagine.art) and click **Sign In** in the top-right corner. You can sign up using any of the following:
* **Google** — sign in with your Google account
* **Facebook** — sign in with your Facebook account
* **Discord** — sign in with your Discord account
* **Email** — enter an email address and create a password
Once you complete sign-up, you land on your ImagineArt dashboard and your account is ready to use.
All new accounts receive **100 free credits per day** — no subscription required. Free credits reset every 24 hours and are a great way to explore the platform before committing to a plan.
When you're ready to unlock more credits, Pro models, and advanced features, visit the [Pricing page](https://www.imagine.art/subscription) to compare subscription plans (Standard, Ultimate, and Creator).
Pro models (labeled "Pro" in the interface) require an active subscription. Free daily credits cannot be used to access them.
From your dashboard, navigate to the **Image** tab at [imagine.art/image](https://www.imagine.art/image) to generate still images, or the **Video** tab at [imagine.art/video](https://www.imagine.art/video) to produce short video clips. Each tab is a self-contained creative studio.
At the bottom of the screen you'll find the **prompt box** — this is where you describe what you want to create. You can be brief or highly detailed; the AI handles both styles.
A few prompts to get you started:
* `a cozy coffee shop on a rainy evening`
* `futuristic cityscape at sunset, neon lights, cyberpunk style`
* `playful golden retriever puppy in a field of flowers`
Prompt crafting is a skill that develops over time. Short prompts produce quick results; detailed prompts with style, lighting, and composition cues give you greater control over the output.
Click the **Create** button. ImagineArt sends your prompt to the selected AI model and returns results shortly:
* **Images** — four variations, typically ready in 4–15 seconds depending on the model and settings.
* **Videos** — one clip, typically ready in 30–60 seconds depending on the model and settings.
You've created your first ImagineArt image or video.
## Use images to guide your creations
Beyond text prompts, you can use images as creative input. Inside the prompt box, click the **+** icon to open the **Add Image** tray. Four modes are available:
| Mode | What it does |
| ------------------------- | ----------------------------------------------------------------------------- |
| **Create (Image Prompt)** | Defines content, composition, style, and colors based on your reference image |
| **Edit Mode** | Customizes and tweaks an existing image with AI assistance |
| **Style** | Matches the visual look and feel of a reference image across your generations |
| **First or Last Frame** | Controls the start or end frame of a video for dynamic sequencing |
## Add animation to images
You can turn any still image — including images you generated with ImagineArt — into a video. Head to the Video tab and use the **Image to Video** feature to upload a static image and animate it with fluid, realistic motion.
## Explore the creative suite
Generate high-fidelity images from text prompts, edit with advanced models, apply style references, and use image prompts to guide your compositions.
Produce cinematic video clips from text or images using models like Google Veo 3, Kling 2.5, and Hailuo 02 Pro. Add VFX and animate images.
# Image Iterator
Source: https://docs.imagine.art/image-iterator
The Image Iterator Node lets you feed multiple images into a workflow and process each one individually through the connected downstream nodes. Instead of running your workflow manually for every image, the iterator handles the batch automatically, passing each image one by one to the next node in the chain.
## How to Use
Click the Add (+) button and select Image Iterator from the Image node category.
Connect multiple image inputs from Import, generated images, or click + to add another Image within the node directly.
Link the iterator's output to any node you want to apply to each image such as Generate Image, Edit Image, Upscale, or AI Resize.
When the workflow runs, the iterator feeds each image one by one through the connected nodes, generating a separate output for every input image.
## Sample Use Cases
Connect multiple product photos to the iterator and link it to a Generate Image node with a style prompt. Each photo gets the same style treatment automatically.
Import a set of reference images—different models, backgrounds, or scenes—and iterate through them with a single prompt to produce consistent branded visuals for each.
Chain the iterator to an Upscale Image node to enhance an entire batch of images in one run, rather than processing each one individually.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Camera Angles
Source: https://docs.imagine.art/image-tools/camera-angles
**Camera Angles** lets you match the perspective, framing, and compositional point of view of a reference in your generation. Instead of describing a shot in text — "low angle," "bird's eye view," "extreme close-up" — you show the AI a reference that has exactly the angle you want, and it replicates that vantage point for your generated scene.
## What Camera Angles does
When you choose a reference angle from our presets, ImagineArt reads the spatial perspective encoded in it — the camera height, angle relative to the subject, distance, and compositional framing. It then uses that perspective as a constraint on how your generated scene is composed, so the output mirrors the shot angle of your reference while depicting the subject and setting from your prompt.
This is particularly useful for maintaining shot consistency across a sequence — for example, holding a fixed eye-level perspective across a storyboard panel, or replicating the dramatic low angle of a specific photograph across multiple character renders.
## How to use Camera Angles
In the image generation panel, click **Add References** to open the references modal.
Choose the **Camera Angles** tab from the six available reference types.
Describe the scene and subject you want to generate. You can also reference the selected camera angle directly in your prompt using @midshot.
Click **Create**. Your generated image will frame your described subject from the same vantage point as your reference.
## Tips for better results
* **Match the reference angle to your subject** — a dramatic low-angle shot of a building will transfer differently to a portrait than to an architectural render. Consider what your subject is and pick a reference angle that makes spatial sense for it.
* **Useful for storyboarding** — locking in a specific shot angle across multiple generations keeps your visual continuity consistent without needing to re-describe camera position in every prompt.
* **Combine with Style or Effects** — Camera Angles handles perspective; the other reference types handle look and treatment. Using them together gives you full compositional and aesthetic control.
# Characters
Source: https://docs.imagine.art/image-tools/characters
Create custom characters and products from reference images and use them consistently across your generations.
**Characters** lets you train a custom character or product from your own reference images and then use it directly in your prompts. Instead of describing a character from scratch every time, you build a reusable identity once—and the AI references that identity every time you generate.
You can access Characters from the **Add References** option in the image generation panel.
## How Characters works
When you create a character, ImagineArt analyzes your uploaded reference images and builds an internal representation of that subject's visual identity. You can then call that identity into any prompt, either by selecting it from the image panel or by tagging it with `@CharacterName` directly in your prompt text.
## Setting up a Character
In the image generation panel, click **Add References** and select the **Characters** tab.
Click **New Characters** (+) to create your own, or select one of our presets. Upload up to 10 reference images for your character — including different angles improves consistency. Give your character a name, select their gender, and optionally add a description.
After clicking **Create**, your character will be saved under the **My Characters** tab. You can also reference it in any prompt using `@CharacterName`.
## Tips for better results
* **Use 4–10 reference images** for characters when possible. A wider variety of reference angles significantly improves consistency.
* **Name your characters descriptively** — names like `ElenaStudio` make it easier to remember which reference to call.
* **Combine with other references** — You can use Characters alongside Style and Image Prompt to further control the output's composition and aesthetic.
# Color Palettes
Source: https://docs.imagine.art/image-tools/color-palettes
Extract the color scheme from any reference image and apply it to your generations.
**Color Palettes** lets you pull the exact color scheme from a reference image and carry it into your next generation. Instead of manually specifying hex codes or describing colors in your prompt, you show the AI an image whose colors you want to replicate — and it matches that palette across your output, helping maintain consistency across brand assets and logo generation.
## What Color Palettes does
When you upload a reference image, ImagineArt analyzes its dominant colors, tonal balance, and color relationships, and uses that analysis to guide the color output of your generation. Alternatively, you can choose a palette from the provided presets. The content of your image is defined by your prompt; the colors are anchored to your reference color palette.
This is useful when you want chromatic consistency across a project — for example, keeping a brand's color language intact across multiple generated visuals.
## How to use Color Palettes
In the image generation panel, click **Add References** to open the references modal.
Choose the **Color Palettes** tab from the six available reference types.
This opens a modal as shown below. You can either manually select and edit individual colors, or upload an image to extract colors from it — e.g., a logo.
Once you're happy with the colors and their intensity, give your palette a name and click **Create**.
Once created, find your palette under the **My Color Palettes** tab, or reference it in your prompt using @paletteName.
## Tips for better results
* **Use images with a clear, intentional palette** — mood boards, film stills, or artwork tend to produce cleaner color extraction than cluttered photos.
* **Pair with Style** — Color Palettes focuses on hue and tone; Style handles the full aesthetic. Using both together gives you fine-grained visual control.
* **Try desaturated or monochromatic references** to produce cohesive, restrained outputs — useful for editorial and brand work.
* **The subject of your reference doesn't need to match your prompt** — a reference image of a sunset can apply its orange-gold palette to an architectural render or a portrait.
# Create Image
Source: https://docs.imagine.art/image-tools/create-image
Experience the next frontier of AI artistry. Choose from Imagine 1.0, Flux, Seedream 4.0, and more to create stunning visuals from text prompts with unmatched details.
### **What is Text to Image?**
**ImagineArt's Text-to-Image** is the bridge between your vocabulary and a visual masterpiece. By translating natural language into high-fidelity pixels, this feature allows you to bypass complex software and manual rendering.
With ImagineArt’s advanced style controls, you can fine-tune the "soul" of your image; perfect for high-impact social media assets, professional marketing campaigns, or personal passion projects. If you can describe it, ImagineArt can build it.
ImagineArt features the latest AI models alongside our proprietary options. See the model comparison at the end of this page for more details.
## Generating your first image
Go to [imagine.art/image](https://www.imagine.art/image). This opens the Image Studio, where all generation controls are available in the left-hand settings panel.
Choose the AI model you want to generate with. Each model has different strengths. See the [model comparison](#available-models) below.
Hover over any model name for about 3 seconds to see its credit cost and a short description before committing to it.
Type your description in the prompt box. The more specific and detailed your prompt, the more accurately the AI will match your intent.
**Example prompts to try:**
* `A cinematic portrait of a woman in a neon-lit Tokyo alley, rain-soaked streets, shallow depth of field, 35mm film`
* `Architectural visualization of a minimalist beach house at golden hour, aerial view, photorealistic`
* `A whimsical children's book illustration of a fox wearing a red scarf, soft watercolor style`
* `Product shot of a matte black coffee mug on a marble surface, studio lighting, commercial photography`
Include descriptive details about lighting, mood, style, camera angle, and subject matter. Prompts like "a dog" will produce generic results; "a golden retriever puppy sitting in autumn leaves, soft natural light, bokeh background, portrait photography" will produce something specific and polished.
Configure your generation settings in the panel:
* **Aspect ratio** — Choose from standard ratios like 1:1, 16:9, 9:16, 3:2, and more depending on the model. Select the ratio that matches your intended output format (e.g., 9:16 for mobile, 16:9 for widescreen, 1:1 for social media).
* **Number of images** — Generate between 1 and 4 images per run. Generating more images per batch lets you compare results and pick the best one.
* **Resolution** — Some models support higher resolution outputs (2K or 4K). Higher resolutions cost more credits.
* **Prompt enhancer** — When enabled, ImagineArt automatically expands and refines your prompt before sending it to the model.
Your total credit cost updates dynamically as you change settings. Hover over the **Create** button to see the exact cost before generating.
Click the **Create** button to start generation. Most models complete within a few seconds to a minute depending on the model, resolution, and number of images requested. Your results appear directly in the canvas area.
Review your generated images. From here you can:
* Use **Quick Actions** (Animate, Variate, Edit, Upscale) directly from the gallery
* Adjust your prompt and regenerate
* Download your favorite result
## Available models
Different models suit different creative goals. Here is a summary of what each model is designed for:
| Model | Best for |
| ------------------------ | ------------------------------------------------------------------------------------------- |
| **Flux Dev** | Fast, high-detail generation with bold artistic diversity. Good default for most prompts. |
| **Flux 2 Pro** | High-quality artistic visuals with flexible creative control. |
| **Flux 2 Max** | High-resolution photorealistic rendering for detailed outputs. |
| **Flux 1.1 Ultra** | Fine-grained texture fidelity, ideal for product visualizations and architectural mock-ups. |
| **ImagineArt 1.0 / 1.5** | Versatile generation for creative and professional projects. |
| **ImagineArt 1.5 Pro** | Higher resolution version of ImagineArt 1.5, up to 4K. |
| **Nano Banana** | High-fidelity textures, complex illustrations, and stylized designs. |
| **Nano Banana 2** | Upgraded version supporting up to 4K output. |
| **ChatGPT Image** | Prompt-accurate visuals, especially strong at incorporating readable text into images. |
| **Imagen 3** | Photorealistic images with accurate lighting, textures, and object relationships. |
| **Imagen 4** | High-resolution photorealistic images with fine detail. |
| **Ideogram v3** | Advanced typography—generates images with crisp, correctly spelled text in complex layouts. |
| **Ideogram Character** | One-shot character consistency model. |
| **Minimax Image** | Realistic portraits, product images, and architectural structures. |
| **Seedream v3** | Cinematic, dream-like scenes with rich lighting and atmospheric depth. |
| **Seedream v4 / v4.5** | Captures cinematic visuals with lush lighting and creative depth, up to 4K. |
| **Seedream v5 Lite** | Balanced quality and speed, up to 3K output. |
| **Dreamina 3.1** | Mood, design, worldbuilding, and storytelling visuals. |
| **Qwen Image** | Highly detailed and stylized images with enhanced lighting and texture control. |
| **Recraft v4** | Clean, high-quality stylized outputs. |
| **Recraft v4 Pro** | Professional-grade stylized rendering. |
| **xAI Grok Imagine** | Fast generation at an accessible credit cost. |
| **Z Image Turbo** | Budget-friendly rapid generation. |
The model list evolves as new models are added. For the latest details, visit [Image mode](https://www.imagine.art/image) and hover any model name to see its current description and credit cost.
# Edit Image
Source: https://docs.imagine.art/image-tools/edit-image
With ImagineArt’s Image Edit feature, you can guide the AI to make specific changes to your images using text-based prompts. Powered by the latest AI models.
## What is Image Edit?
**Image Edit** is an AI-powered tool designed to help you make detailed edits to your images through simple text descriptions. Whether you need to change lighting, adjust colors, remove unwanted objects, or even add new elements, Image Edit allows you to give specific instructions to the AI, which then generates high-quality results based on your prompts.
## How to use Image Edit
1. Open [ImagineArt](https://www.imagine.art/image) and select **Edit Mode**.
2. \*\*Add an image to be used as a reference: \*\*Upload the image you want to edit or simply drag it into the "+" box. This will serve as the reference for the AI to make edits based on your description.Choose an editing model from the options below.
3. \*\*Choose an image editing model: \*\*Select from various available models (e.g., **Nano Banana**, **Seedream v4**, **ChatGPT Edit**) depending on the level of detail and style you need.
4. \*\*Write the editing prompt: \*\*Now, write a detailed description of the changes you want to make. Be specific about what you want altered—whether it’s adjusting brightness, adding objects, or changing the composition.
5. \*\*Choose aspect ratio and generation settings: \*\*Select the aspect ratio (e.g., square, landscape, portrait) and any other settings that match your project’s needs.
6. \*\*Click Create: \*\*Once everything is set, click **Generate** to let the AI process your image and apply the changes based on your prompt.
Be specific in your edit prompt. Instead of "change the background," try "replace the background with a sun-drenched Tuscan vineyard at golden hour." Precision helps the AI understand exactly what you want.
## Editing models
**Best for:** Complex natural language instructions, scene adjustments, and perspective changes.
**Nano Banana**, developed by Google, understands complex editing instructions similar to how a conversational AI processes language. It excels at interpreting multi-part instructions and making nuanced adjustments to scenes, angles, and object placement.
**Key capabilities:**
* Understands complex, multi-step editing instructions in plain language
* Adjusts scene angles, perspectives, and spatial relationships
* Adds, removes, or repositions objects based on descriptions
* Produces high-fidelity textures and stylized outputs
**How to use Nano Banana:**
1. Select **Edit Mode** and choose **Nano Banana** as your model.
2. Upload **reference images**.
3. Write your editing prompt, describing what you want to change or add.
4. Click **Create**.
**Example prompt:** `"Shift the camera angle to a low-angle shot looking up at the subject. Add dramatic storm clouds to the sky."`
**Best for:** Multi-element compositions, detailed scene generation, and images with embedded text.
**ChatGPT Image** by OpenAI excels at prompt accuracy and complex scene construction. It is particularly strong when you need multiple distinct elements to coexist in a single image or when your image needs to include readable, correctly spelled text.
**Key capabilities:**
* Highly prompt-accurate: follows detailed, multi-part instructions closely
* Strong at generating complex scenes with many interacting elements
* Best-in-class text rendering inside images
* Character and object design from detailed descriptions
**How to use ChatGPT Image:**
1. Select **Edit Mode** and choose **ChatGPT** as your model.
2. Upload **reference images**.
3. Write your detailed editing or generation prompt.
4. Click **Create**.
**Example prompt:** `"Add a chalk-board sign in the background that reads 'Grand Opening' in bold script. Keep the storefront and lighting unchanged."`
**Best for:** Character consistency across multiple scenes, poses, and styles.
**Imagine You** is a specialized character consistency model. Upload a single reference photo of a person or character, and it generates variations of that same character in different poses, outfits, environments, and artistic styles—while preserving their core visual identity.
**Key capabilities:**
* One-shot character consistency from a single reference photo
* Generates variations across styles (photorealistic, cartoon, fantasy, etc.)
* Explore different poses, clothing, and settings while maintaining likeness
* Select presets or guide output through the prompt box
**How to use Imagine You:**
1. Select **Edit Mode** and choose **Imagine You** as your model.
2. Upload a clear reference photo of your character.
3. Select a style preset or write a custom prompt to describe the variation you want.
4. Set aspect ratio and image count, then click **Create**.
**Example prompt:** `"Same character, wearing a futuristic silver spacesuit, standing on a Martian landscape, dramatic cinematic lighting."`
**Best for:** High-accuracy editing with real-time knowledge, multi-character consistency, and sharp text rendering.
**Nano Banana 2**, powered by Google's Gemini 3.1 Flash Image model, is an upgraded successor to Nano Banana with significantly improved accuracy, richer visual quality, and real-time web search integration. It can handle up to 5 characters and 14 objects simultaneously while maintaining consistency, and supports resolutions from 512px up to 4K. Its real-time web grounding means it can generate visually accurate representations of real-world subjects, brands, and locations.
**Key capabilities:**
* **Real-time web search grounding** — Pulls live online information so generated visuals are factually and visually accurate (e.g., real product appearances, landmarks, logos).
* **Multi-character and multi-object consistency** — Maintain up to 5 characters and 14 objects consistently across a single image or across multiple outputs.
* **Precise text rendering** — Generate legible, accurate text within images for marketing mockups, greeting cards, posters, and product labels.
* **High-resolution output** — Supports resolutions up to 4K with vibrant lighting, richer textures, and sharper details.
* **Fast Flash-speed generation** — Delivers high quality at the speed expected from Gemini Flash architecture.
**How to use Nano Banana 2:**
1. Select **Edit Mode** and choose **Nano Banana 2** as your model.
2. Upload up to **14 reference images** (optional) or start with a text prompt.
3. Describe your desired output or edits with as much detail as needed.
4. Choose your resolution and aspect ratio.
5. Click **Create**.
**Example prompt:** `"Generate a product advertisement for a luxury watch on a marble surface. Include the text 'Timeless Precision' in elegant serif font at the bottom."`
**Best for:** Professional-grade, photorealistic image creation with maximum prompt fidelity and multilingual text support.
**Nano Banana Pro**, powered by Google's Gemini 3 Pro Image model, is Google's highest-quality image generation and editing model. Designed for creative professionals and advanced workflows, it produces ultra-high-resolution 4K images with industry-leading text rendering, localized editing, and search-grounded accuracy. It's the best choice when image quality, detail, and precision are non-negotiable.
**Key capabilities:**
* **Ultra-high-resolution 4K output** — Generates images at 3840×2160 or higher with exceptional sharpness and detail.
* **Industry-leading text rendering** — Create infographics, slides, diagrams, and layouts with multi-language text rendered clearly, from short taglines to full paragraphs.
* **Localized editing and inpainting** — Fine-tune specific areas of an image—adjusting focus, lighting, color grading, or camera angles—without affecting the rest.
* **Search-grounded generation** — Google Search grounding allows the model to research topics and generate factually accurate, context-aware visuals for maps, diagrams, or real-world subjects.
* **Multi-identity consistency** — Preserve facial features, clothing, and posture for up to 5 people across multi-image sequences.
* **Multilingual support** — Accurately render text across different writing systems and languages.
**How to use Nano Banana Pro:**
1. Select **Edit Mode** and choose **Nano Banana Pro** as your model.
2. Upload **reference images** or start from a detailed text prompt.
3. Describe your desired output, specifying style, lighting, text elements, and any localized edits.
4. Select your resolution (up to 4K) and aspect ratio.
5. Click **Create**.
**Example prompt:** `"Create a magazine cover layout featuring a woman in a tailored red suit against a white studio backdrop. Include the headline 'The Future of Design' in bold black font at the top, and a subheading in French below."`
**Best for:** 4K image generation with deep reasoning, real-time web search, and exceptional text accuracy.
**Seedream v5 Lite**, ByteDance's most advanced image generation model, combines chain-of-thought visual reasoning with real-time web search to produce native 4K images in 2–3 seconds. It's the first Seedream model with live internet connectivity, enabling it to retrieve up-to-date visual references during generation. With 99%+ text spelling accuracy in both English and Chinese, it excels at any content that requires precise typography — from product labels to marketing posters.
**Key capabilities:**
* **Real-time web search integration** — The first Seedream model to retrieve live information during generation, producing more accurate and up-to-date visuals.
* **Native 4K generation** — Produces high-resolution images natively, without upscaling artifacts.
* **Chain-of-thought visual reasoning** — The model thinks through complex prompts step by step, resulting in better-composed and more coherent images.
* **99%+ text accuracy** — Industry-leading spelling accuracy for rendered text in both English and Chinese across diverse font styles, rotations, and layouts.
* **Multi-reference compositing** — Combine elements from up to 14 reference images in a single generation.
**How to use Seedream v5 Lite:**
1. Select **Edit Mode** and choose **Seedream v5 Lite** as your model.
2. Upload **reference images** or start from a text prompt.
3. Write a detailed prompt describing the scene, style, and any text elements you need.
4. Set your resolution and aspect ratio.
5. Click **Create**.
**Example prompt:** `"A high-resolution product mockup of a skincare bottle on a pale pink marble surface, with soft natural lighting. Include the brand name 'Lumière' in elegant gold cursive on the label."`
**Best for:** Versatile, high-volume image generation across photorealistic, artistic, and illustrative styles.
Grok Imagine is xAI's dedicated image generation model, built for creative flexibility and production-scale output. It supports a wide range of visual styles — from photorealistic portraits to anime, oil painting, and pencil sketches — and accepts up to three reference images per prompt. With generation modes optimized for either quality or speed, and support for up to 300 requests per minute, it's well-suited for both individual creators and high-volume workflows.
**Key capabilities:**
* Multi-style generation — Produce images ranging from ultra-realistic photography to anime, watercolor, oil painting, digital illustration, and beyond.
* Image-to-image generation — Upload up to 3 reference images and describe the transformation you want.
* Quality and Speed modes — Choose between Speed Mode for fast output and Quality Mode for finer detail and higher fidelity.
* Precise text and logo rendering — Generate legible text, accurate logos, and detailed realistic elements within images.
* High-volume capability — Handles up to 300 API requests per minute for production use cases.
**How to use xAI Grok Imagine:**
1. Select Edit Mode and choose xAI Grok Imagine as your model.
2. Upload up to **reference images** to guide the visual style or subject.
3. Write a prompt describing the desired image, style, and any elements to change or preserve.
4. Select your generation mode (Quality or Speed) and aspect ratio.
5. Click Create.
**Example prompt:** `"A realistic portrait of a woman with curly red hair standing in a forest at golden hour. Painterly style, warm tones, soft background bokeh."`
**Best for:** Precise image editing, inpainting, transparent background generation, and real-time streaming previews.
ChatGPT 1.5 (GPT Image 1.5) is OpenAI's most advanced image generation and editing model, built as the successor to DALL-E 3. It delivers superior instruction-following accuracy with highly targeted edits that change only what you ask for, leaving the rest of the image intact. Key additions include streaming generation (see partial previews as the image builds), transparent background support, and prompt-guided inpainting — making it a powerful tool for production-ready image workflows.
**Key capabilities:**
* Precise targeted editing — Change specific elements (shirt color, background, object position) using natural language while keeping everything else consistent.
* Prompt-guided inpainting — Edit a defined region of an image using a mask, with the model blending changes naturally into the surrounding content.
* Transparent background generation — Produce images with no background, ready for compositing into designs or placing on custom surfaces.
* Consistent facial and detail fidelity — Maintains original composition, lighting, style, and fine-grained detail across edits and iterations.
**How to use ChatGPT 1.5:**
1. Select Edit Mode and choose ChatGPT 1.5 as your model.
2. Upload your image (required for editing workflows) or start from a text prompt.
3. For inpainting, indicate the region you want to edit; for full edits, describe the changes in natural language.
4. Enable streaming if you want partial preview images during generation.
5. Click Create.
**Example prompt:** `"Remove the background entirely and replace it with a transparent layer. Keep the product and all its shadows exactly as they are."`
**Best for:** Unified text-to-image generation and editing with exceptional typography and multi-image consistency.
Seedream v4.5 is ByteDance's professional-grade image model that unifies text-to-image generation and image editing in a single architecture. Released in December 2025, it delivers up to 4-megapixel (2048×2048) output with outstanding text rendering — supporting complex typography, multiple font styles, curved and rotated text, and multilingual content. Its multi-image support makes it especially effective for projects requiring consistent visual output across many assets.
**Key capabilities:**
* Unified generation and editing — Generate new images from text and edit existing images within the same model and workflow.
* Exceptional text rendering — Accurately renders complex words, multiple text elements, diverse font styles, rotated or curved text, and multilingual content.
* Multi-reference support — Process up to 4 reference images simultaneously, maintaining consistent subjects, lighting, and style across complex compositions.
* Multi-image consistency — Ideal for product catalogs, brand campaigns, and visual storytelling where multiple images must share the same character, style, or theme.
**How to use Seedream v4.5:**
1. Select Edit Mode and choose Seedream v4.5 as your model.
2. Upload up to 4 **reference images** for consistency or editing guidance.
3. Write a prompt describing your scene, style, and any text elements to include.
4. Choose your output resolution and aspect ratio.
5. Click Create.
**Example prompt:** `"A series of product packaging designs for a tea brand in a minimalist Japanese style. Each image should feature the same logo, clean typography with the flavor name, and a botanical illustration in muted earth tones."`
**Best for:** Top-tier photorealism, character consistency, and professional-grade image editing with maximum prompt fidelity.
FLUX.2 Max is Black Forest Labs' highest-quality image generation and editing model in the FLUX.2 family. Built on a hybrid Mistral-3 24B vision-language model with Rectified Flow Transformer architecture, it delivers unmatched photorealism, the strongest prompt following in its class, and real-time web grounding for factually accurate visuals. It's the premium choice for commercial campaigns, product imagery, and any creative work where quality is the top priority.
**Key capabilities:**
* Unmatched photorealism — Professional-grade output with exceptional sharpness, lighting, and material detail.
* Real-time web grounding — Performs live web searches to incorporate accurate, current information into generated images.
* Multi-reference editing — Supports up to 4 reference images with reliable character consistency, product styling, and brand identity preservation.
* Precise hex color control — Specify exact color values for pixel-perfect brand and design consistency.
* Strongest prompt adherence — Maximum faithfulness to complex, multi-element prompts across styles and subjects.
**How to use Flux.2 Max:**
1. Select Edit Mode and choose Flux.2 Max as your model.
2. Upload up to 4 reference images (optional) to guide character, style, or product consistency.
3. Write a detailed prompt specifying style, lighting, colors, and any elements to preserve or change.
4. Set your resolution (up to 4MP) and inference steps.
5. Click Create.
**Example prompt:** `"A luxury fashion editorial photograph of a model in a tailored ivory trench coat, standing on a rain-slicked Paris street at night. Cinematic lighting, reflections in puddles, sharp focus on the fabric texture."`
**Best for:** High-resolution generation and editing with reliable text rendering and consistent character identity across multiple outputs.
Released in November 2025, it delivers accurate text rendering, precise color matching, and consistent character identity across multiple generations — making it a strong choice for creators who need dependable, high-quality output at scale without the maximum compute of FLUX.2 Max.
**Key capabilities:**
* Accurate text rendering — Reliably places legible, well-styled text within images.
* Precise color matching — Consistent color reproduction for brand and product imagery.
* Character identity consistency — Maintain the same character's appearance across multiple generated outputs.
* Multi-reference editing — Combine reference images for compositing, style matching, and character-guided generation.
**How to use Flux.2 Pro:**
1. Select Edit Mode and choose Flux.2 Pro as your model.
2. Upload 4 reference images for style or character guidance.
3. Write a prompt describing the desired image, specifying details like lighting, composition, and text.
4. Set your resolution and aspect ratio.
5. Click Create.
**Example prompt:** `"A high-resolution product shot of a premium chocolate box on a dark wood surface, with soft studio lighting. Include 'Maison Noël' in gold serif text embossed on the lid."`
**Best for:** Multi-image compositing, precise character consistency across scenes, and context-aware editing using multiple source images.
Flux Kontext Max Multi is Black Forest Labs' highest-quality context-aware image editing model, optimized specifically for multi-image workflows. While the standard Flux Kontext Max excels at single-image editing, the Multi variant is designed to combine, blend, and maintain consistency across up to 3input images simultaneously — ideal for complex compositions, multi-scene campaigns, and character-driven visual storytelling where elements from different sources must seamlessly coexist.
**Key capabilities:**
* Multi-image input support — Accepts up to 3reference images, enabling compositing and element fusion from diverse sources.
* Strongest instruction following — Maximum accuracy in understanding and executing complex, multi-element editing instructions.
* Character and identity consistency — Preserves facial features, clothing, proportions, and visual identity across all inputs and outputs.
* Detail preservation — Fine-grained preservation of textures, lighting, and compositional elements from source images.
* Context-aware editing — Processes all input images in context to maintain coherent spatial relationships, lighting, and style.
* High-resolution output — Generates at up to 4MP across any aspect ratio.
**How to use Flux Kontext Max Multi:**
1. Select Edit Mode and choose Flux Kontext Max Multi as your model.
2. Upload up to 3 reference images from which you want to draw elements.
3. Write a prompt specifying what to combine, preserve, or change across the input images.
4. Set your aspect ratio and output resolution.
5. Click Create.
**Example prompt:** `"Combine the character from image 1 with the outdoor environment in image 2. Keep the character's clothing and facial features identical, and match the lighting to the scene in image 2."`
**Best for:** High-fidelity, subject-preserving image edits — outfit swaps, product restyling, and interior design changes.
Seedream v4 Edit is ByteDance's dedicated image-to-image editing model, optimized for making precise, targeted changes to existing images while maintaining subject identity, lighting, and overall composition. It's specifically tuned for tasks like swapping outfits, changing hair or makeup, restyling products, and transforming interior finishes — all without disrupting what you want to keep. Its context-aware approach handles depth, perspective, and lighting consistency automatically.
**Key capabilities:**
* Subject-preserving editing — Accurately identifies and preserves the main subject (person, product, or object) while applying targeted changes.
* High-fidelity output — Reliable skin tones, fabric and material detail, logo reproduction, and fine edge clarity on people and products.
* Natural language spatial instructions — Describe edits in plain language (e.g., "change the wall color to sage green") for intuitive control.
* Context-aware transformations — Automatically maintains depth, perspective, and lighting consistency when integrating new elements — no manual blending required.
* Multi-reference support — Use multiple reference images to guide editing style, subject identity, or material appearance.
**How to use Seedream v4 Edit:**
1. Select Edit Mode and choose Seedream v4 Edit as your model.
2. Upload the image you want to edit as your base image.
3. Upload 4 reference images to guide style or subject identity.
4. Write a prompt specifying what you want to change and what should stay the same.
5. Click Create.
**Example prompt:** `"Change the model's jacket to a cream-colored oversized blazer. Keep her face, hair, pose, and all background elements exactly as they are."`
## Supported models summary
| Model | Primary strength |
| ---------------- | --------------------------------------------------- |
| Flux Kontext Max | Precise local edits, character consistency |
| Nano Banana | Complex natural language instructions |
| ChatGPT Image | Multi-element scenes, embedded text |
| Imagine You | Character variations from a single photo |
| Runway Reference | Cross-image style and visual consistency |
| Seedream v4 Edit | Dream-like, artistic edits with creative flair |
| Qwen Edit | Intuitive edits with style transfer and fine-tuning |
# Effects
Source: https://docs.imagine.art/image-tools/effects
Replicate the visual effects from a reference image — lighting, atmosphere, texture overlays, and more.
**Effects** lets you capture and transfer the visual treatment of a reference image into your generation. This goes beyond overall style or color — it isolates the specific effects layered onto an image, such as light leaks, film grain, lens blur, atmospheric haze, glow, or other post-processing treatments, and applies them to your output.
## What Effects does
When you upload a reference image, ImagineArt analyzes the visual effects and treatments applied to it — the way light interacts with the scene, the texture overlays, the depth-of-field characteristics, and any stylistic post-processing. These qualities are then applied to the image your prompt generates, giving your output a treated, finished look without any manual editing.
This is valuable when you have a specific "finished" look in mind and want to apply it consistently — for example, matching the cinematic grade of a film still, replicating the grain and fading of an analog photograph, or carrying a specific glow or atmospheric haze across a set of images.
## How to use Effects
In the image generation panel, click **Add References** to open the references modal.
Choose the **Effects** tab from the six available reference types.
Browse and select from the available effects.
Describe the scene or subject you want to generate. Click **Create**. The AI will generate your described scene with the visual effects from your reference applied.
## Tips for better results
* **Combine Effects with Style** — Effects handles the treatment layer; Style handles the overall aesthetic. Together they give you comprehensive visual control.
# Elements
Source: https://docs.imagine.art/image-tools/elements
Pull specific objects, motifs, or visual elements from a reference image into your generations.
**Elements** lets you extract a specific object or motif from a reference image and inject it into your generation. Rather than describing an object in text and hoping the AI interprets it correctly, you show it exactly what you want — and the AI places that element into your generated scene.
## What Elements does
When you upload a reference image, ImagineArt identifies and isolates the visual elements within it. You can then specify which element you want to carry into your generation, and the AI incorporates it into the scene described by your prompt — adapting its lighting, perspective, and style to match the rest of the output.
This is especially useful for product integration, prop consistency, or adding a specific motif (a logo shape, a piece of furniture, a vehicle) into multiple different scenes without re-describing it each time.
## How to use Elements
In the image generation panel, click **Add References** to open the references modal.
Choose the **Elements** tab from the six available reference types.
You can either choose from our existing presets or click **New Elements** to create a new one. Upload up to 10 reference images, name your element, and optionally add a description.
Describe the scene you want to generate, and reference the element you're bringing in. For example: *"a living room interior with the chair from the reference image, afternoon light, architectural photography."*
You can reference your element in your prompt using **@elementName**.
## Tips for better results
* **Use clean, uncluttered reference images** — the clearer the element, the more accurately the AI can extract and transfer it.
* **Reference the element in your prompt** — explicitly mentioning it helps the AI understand how to integrate it into the scene.
* **Combine with Characters** — use Characters for human subjects and Elements for props, products, or objects to get both consistent people and consistent items in the same generation.
* **Use high-contrast reference images** when the element has fine detail — this helps preserve edge quality and surface texture in the output.
# Image Credits
Source: https://docs.imagine.art/image-tools/image-credits
Understand how credits are consumed for image generation, editing, personalization, and upscaling.
Credit consumption in Image Studio depends on the **model you choose** and your **generation settings**. Different models have different base costs, and certain settings—like resolution, aspect ratio, and the number of images—adjust the total cost before you generate.
For a full explanation of how credits work, how they are allocated, and how to manage your balance, see the [Understanding Credits](/overview/understanding-credits) page.
## Check your cost before generating
You never need to guess what a generation will cost. ImagineArt shows you the exact credit cost in real time.
In [Image mode](https://www.imagine.art/image), hover over any model name for about 3 seconds to see its base credit cost and a short description.
As you change your settings, the credit cost updates automatically. Settings that affect cost include:
* **Image size and aspect ratio** — Larger output dimensions cost more.
* **Resolution** — 2K and 4K outputs cost more than 1K.
* **Prompt enhancer** — Enabling the prompt enhancer adds a small credit cost.
* **Number of images** — Generating 4 images costs 4× the per-image rate.
Before clicking **Create**, hover over the button to see the **total credit cost** for your exact current settings. This is the definitive cost—use it to confirm before generating.
Credit costs are updated regularly. The tooltip in [Image mode](https://www.imagine.art/image) is always the source of truth. The tables below reflect current costs but may not reflect future changes.
## Create / Edit image (base cost per image)
Credits shown are the base cost for generating **1 image**. Generating multiple images multiplies this cost.
| Model | Credits (per image) | Max resolution | Supported aspect ratios |
| ---------------------- | ----------------------------- | -------------- | ----------------------------------------------------------------------- |
| **ChatGPT Image 2** | 6 | Up to 4k | 1:1, 3:4, 4:3, 9:16, 16:9, 3:2, 21:9 |
| **Nano Banana 2** | 49 (1K) / 73.5 (2K) / 98 (4K) | Up to 4K | 1:1, 4:3, 3:4, 3:2, 2:3, 5:4, 4:5, 9:16, 16:9, 21:9 |
| ImagineArt 2.0 | 25 | See tooltip | 1:1, 9:16, 16:9, 3:2, 2:3, 4:3, 3:4, 3:1, 1:3, 21:9, 4:5, 5:4, 4:1, 1:4 |
| **ImagineArt 1.5** | 15 | Up to 2K | 1:1, 4:3, 3:4, 1:3, 3:1, 2:3, 3:2, 9:16, 16:9 |
| **ImagineArt 1.5 Pro** | 25 | Up to 4K | 1:1, 4:3, 3:4, 1:3, 3:1, 2:3, 3:2, 9:16, 16:9 |
| Recraft v4.1 | 150 | See tooltip | 1:1, 4:3, 3:4, 9:16, 16:9 |
| Ideogram v4 | 36 | See tooltip | 1:1, 4:3, 3:4, 9:16, 16:9 |
| Krea flux 1 | 10 | See tooltip | 1:1, 3:4, 4:3, 9:16, 16:9 |
| **Midjourney V7** | 60 | See tooltip | 1:1, 4:3, 3:4, 2:3, 3:2 9:16, 16:9, 21:9, 9:21 |
| **Nano Banana Pro** | 80 (1K/2K) / 160 (4K) | Up to 4K | 1:1, 4:3, 3:4, 3:2, 2:3, 5:4, 4:5, 9:16, 16:9, 21:9 |
| **Seedream v5 Lite** | 20 | Up to 3K | 1:1, 4:3, 3:4, 2:3, 3:2, 9:16, 16:9 |
| **Recraft v4** | 24 | See tooltip | 1:1, 4:3, 3:4, 16:9, 9:16 |
| **Recraft v4 Pro** | 150 | See tooltip | 1:1, 3:4, 4:3, 9:16, 16:9 |
| **ChatGPT 1.5** | 35 | Up to 1K | 1:1, 2:3, 3:2 |
| **xAI Grok Imagine** | 12 | 1K | 1:1, 3:4, 4:3, 2:3, 3:2, 9:16, 16:9 |
| **Flux 2 Max** | 50 | See tooltip | 1:1, 3:4, 4:3, 9:16, 16:9 |
| **Flux 2 Pro** | 28 | Up to 2K | 1:1, 3:2, 2:3, 5:4, 4:5, 9:16, 16:9 |
| **Z Image Turbo** | 5 | See tooltip | 1:1, 4:3, 3:4, 1:3, 3:1, 2:3, 3:2, 9:16, 16:9 |
| **Seedream v4.5** | 24 | 4K | 1:1, 4:3, 3:4, 2:3, 3:2, 9:16, 16:9 |
| **Flux 1.1 Ultra** | 36 | 2K | 1:1, 4:3, 3:4, 3:2, 2:3, 9:16, 16:9 |
| **ChatGPT** | 45 | Up to 1K | 1:1, 2:3, 3:2 |
| **Nano Banana** | 24 | Up to 1K | 1:1, 4:3, 3:4, 3:2, 2:3, 5:4, 4:5, 9:16, 16:9 |
| **Dreamina 3.1** | 18 | Up to 2K | 1:1, 4:3, 3:4, 9:16, 16:9 |
| **Ideogram v3** | 25 | Up to 1K | 1:1, 4:3, 3:4, 9:16, 16:9 |
| **Flux Dev** | 5 | Up to 1K | 1:1, 3:4, 2:3, 3:2, 5:4, 9:16, 16:9 |
| **Qwen Image** | 24 | Up to 1K | 1:1, 4:3, 3:4, 9:16, 16:9 |
| **Minimax Image** | 6 | Up to 1K | 1:1, 4:3, 3:4, 3:2, 2:3, 9:16, 16:9 |
| **Seedream v4** | 18 | Up to 4K | 1:1, 4:3, 3:4, 2:3, 3:2, 9:16, 16:9 |
## Upscale image
Upscale costs vary depending on the input image size and the selected output scale factor.
| Engine | Supported scales | Credits |
| ------------------------- | ---------------- | ----------------------------------------------------------------------------------------- |
| **Topaz** | 2×, 4× | - 2× = base cost × megapixels × 1
- 4× = base cost × megapixels × 3
|
| **Magnific Precision v2** | 2×, 4×, 8×, 16× | - 2×: 65
- 4×: 260
- 8×: 520
- 16×: 1040
|
> *Example: A 1920×1080 image at 2× costs approximately 41 credits with the Topaz model*
Note: Credit consumption for each generation may differ depending on the number of images generated. The breakdown provided above is for single image generation.
# Image Prompt
Source: https://docs.imagine.art/image-tools/image-prompt
Use an existing image as a visual reference to guide AI image generation alongside your text prompt.
### **What is Image Prompt?**
Think of an **Image Prompt** as the "visual shorthand" for your creative ideas. Instead of relying solely on words, you provide a starting image; a photo, a sketch, or even a previous generation to act as a structural anchor.
You simply upload or drag your image into the Image Prompt field, and the AI uses it as a foundational layer of inspiration. It’s like a conversation where your image sets the scene, and your text prompt provides the direction.
## How to use Image Prompt
You can upload any image from your device by clicking the "+" button to open the file browser, or simply drag and drop an image directly into the field. Supported formats include JPEG, PNG, and WEBP. For best results, use a clear, well-lit image that reflects the composition or style you want the AI to draw inspiration from.
Once uploaded, the AI will analyse your image and use it as a visual foundation alongside your text prompt. The more relevant your image is to your desired outcome, the more accurately the AI can capture the mood, structure, and style you're going for. You'll see a thumbnail preview of your uploaded image in the field confirming it's been applied.
**Example combinations:**
* Reference image of a misty forest + prompt: `"A lone cabin with warm glowing windows, winter evening, oil painting style"`
* Reference image of a sports car + prompt: `"Same car on a coastal highway at sunset, cinematic photography, motion blur"`
* Reference image of a character sketch + prompt: `"Full color illustration, fantasy armor, dramatic lighting"`
Not happy with the first result? That's completely normal — and part of the fun! You can adjust your text prompt, swap out your reference image, or tweak your settings and generate again. Each iteration builds on your feedback, helping you get closer to your ideal image with every attempt.
When you're happy with your image prompt and text, hit the Create button to generate your image. The AI will combine your visual reference and text instructions to produce a unique result. Depending on the selected model and settings, generation typically takes just a few seconds.Supported models
## Support
## Support
Image prompt is supported by all available image generation models.
| Model | Description |
| ------------------ | --------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Flux dev | A pioneering experimental engine engineered for rapid, high-detail image synthesis with a bold artistic diversity. |
| Realistic | A high-fidelity generator that produces true-to-life visuals with precise textures, authentic lighting, and natural depth for portraits and scenes. |
| ChatGPT Image | Delivers prompt-accurate visuals, excels at incorporating readable, precise text into images. |
| Imagen 3 | Specializes in producing photorealistic, detailed images with accurate lighting, textures, and object relationships. |
| Flux Ultra 1.1 | Excels at high-resolution photorealistic rendering with fine-grained texture fidelity, making it ideal for product visualizations and architectural mock-ups. |
| Ideogram v3 | Noted for its advanced typography control—it can generate images that incorporate crisp, correctly spelled text in complex layouts. |
| Minimax Image | Great for realistic portraits, product images, and structures. |
| Seedream v3 | Specializes in dream-like cinematic scenes; it synthesizes rich lighting, atmospheric depth, and vibrant color palettes reminiscent of concept-art mood boards. |
| Ideogram Character | Best in class one-shot character consistency model. |
| Dreamina Image 3.1 | Great for mood, design, worldbuilding, and storytelling. Not ideal for prompt-as-spec use cases like typography-heavy layouts. |
The model you choose affects how strongly the reference image influences the output. Experiment with different models to find the balance between reference fidelity and creative freedom that works best for your project.
# Quick Actions
Source: https://docs.imagine.art/image-tools/quick-actions
.png?alt=media\&token=3c33ed74-ac07-404a-a6b9-d33618a0bd21)
## What are Quick Actions?
**Quick Actions** are one-click tools that let you modify images instantly from your gallery without opening a separate editor. Hover over any image to reveal a shortcut menu with four key actions:
* **Animate** – Add motion effects to bring your image to life
* **Variate** – Generate multiple variations to explore different creative directions
* **Edit** – Use text prompts to modify specific elements
* **Upscale** – Enhance resolution and clarity for sharper, higher-quality output

## How to Use Quick Actions
Navigate to your home view where generated images are displayed. Hover over any image to reveal the Quick Actions options: **Animate**, **Variate**, **Edit**, and **Upscale**.
Transform your image into a dynamic visual with motion effects.
1. Click **Animate** – your image automatically becomes the animation reference
2. Add a text prompt to customize the motion level
3. Click **Create** to generate your animated result

Generate 4 instant variations of your image to explore different creative options.
1. Click **Variate** – processing begins immediately
2. Review the generated variations
3. Select your preferred version or use it as a new starting point


Use text prompts to make specific edits like changing colors, adding elements, or adjusting composition.
1. Click **Edit** – your image becomes the reference
2. Enter a text prompt describing your desired changes
3. Click **Create** to apply the edits


Increase image resolution and clarity for sharper details and finer textures. Cost depends on input image size.
.png?alt=media\&token=bde6dde8-efd0-4fa9-8339-ec9d820e1991)
## Best Practices
* **Use Variate to explore ideas** – Instantly generate multiple directions without starting from scratch
* **Use Animate for visual impact** – Create engaging motion effects for social media and visual projects
* **Use Edit for precision** – Make targeted adjustments to specific elements, backgrounds, or compositions
# Style
Source: https://docs.imagine.art/image-tools/styles
Use any existing image as a visual blueprint to guide the aesthetic of your generations.
## **What is Style?**
**Style** is your shortcut to visual consistency. It allows you to use any existing image as a "blueprint" for the aesthetic of your next creation. Instead of trying to describe a complex art style with hundreds of words, you simply show the AI what you want, and it maps that "DNA" onto your new prompt.
Whether you want your work to align with a specific aesthetic or want to keep the style consistent across multiple designs, **Style** makes it easy to apply and replicate that look to new images.
## **How to use Style?**
Style is supported by SDXL models only.
Select **Style**: Begin by selecting the **Style** option under **Add References** in the image generation panel.
**Add a reference image**: Upload an image that embodies the style you want to replicate — this guides the AI in applying that aesthetic to your new creation. Alternatively, you can select one of our presets.
Click the **New Styles** button to create your own. Clicking it opens a modal as shown below:
Start by uploading an image, or select one from your previous creations. Make sure to give your style a name.
Once you're satisfied with your image, name, and optional description, click **Create**. Your style will be saved to the **My Styles** tab inside the Add References modal.
You can apply a style by clicking on it in the panel, or by referencing it in your prompt using the @ symbol.
#### **Why Use Style?**
* **Unshakeable Aesthetic Consistency:** Stop guessing and start creating. Style ensures that the lighting, mood, and visual DNA remain identical across every image you generate. This is the ultimate tool for storyboarding, brand identity, and cohesive concept design.
* **Total Creative Command:** Take the driver’s seat. Instead of hoping the AI understands your description of a "mood," you provide the blueprint. This gives you pixel-perfect control over the final artistic direction, ensuring every output aligns with your exact vision.
* **Rapid Style Replication:** Found a look you love? Don’t let it go. Once you’ve captured a specific aesthetic with a reference image, you can instantly deploy it across multiple projects. Save hours of prompting time while maintaining professional-grade visual coherence.
# Upscale
Source: https://docs.imagine.art/image-tools/upscale
Upscaler is an AI-powered tool designed to enhance your image resolution, making it sharper and more detailed without compromising on quality. Whether you need a subtle increase in resolution or a creative boost to bring more depth and richness, the Upscaler service has you covered.
### **How to use Upscaler?**
Video Studio offers a suite of AI-powered tools to enhance your creative workflow:
**Select Your Image**\
Upload your image that you want to upscale.
**Choose the Upscale Option**
* **Subtle**: This option simply increases the resolution of your image while retaining the original details and characteristics.
* **Creative**: This option not only upscales the image but also adds creative details to enhance the image, giving it a fresh look and new elements.
You can view the credit consumption for upscaling [here](/image-tools/credit-consumption#upscale-image)
# What is Add References
Source: https://docs.imagine.art/image-tools/what-is-add-references
Use reference images to guide your generations — control style, color, characters, objects, effects, and camera perspective.
**Add References** gives you six ways to anchor your generations to real visual inputs rather than relying on description alone. Each reference type targets a different dimension of the image — from the overall look to the specific angle of the shot.
Use any image as a visual blueprint. The AI extracts the aesthetic — lighting, mood, color, texture — and applies it to your new prompt.
Extract the color scheme from a reference image and carry it into your generation, independent of subject or style.
Train a reusable identity from reference photos — a person, character, or product — and call it into any prompt by name.
Pull a specific object or motif from a reference image and inject it into your generated scene.
Replicate the visual treatment of a reference — film grain, light leaks, atmospheric haze, glow — and apply it to your output.
Match the perspective and framing of a reference shot — bird's eye, low angle, close-up — and apply it to your generated composition.
## When to use each
| | What it controls | Best for |
| ------------------ | ------------------------------------------------- | --------------------------------------------------------------- |
| **Styles** | Overall aesthetic — lighting, mood, texture | Mood boards, consistent campaign aesthetics, art direction |
| **Color Palettes** | Dominant colors and tonal balance | Brand color consistency, themed series, editorial work |
| **Characters** | Subject identity — who or what appears | Character consistency, product shoots, brand mascots |
| **Elements** | Specific objects or motifs | Product integration, prop consistency, recurring visual details |
| **Effects** | Visual treatment — grain, glow, grade, atmosphere | Cinematic looks, analog aesthetics, post-processing consistency |
| **Camera Angles** | Perspective, framing, and shot angle | Storyboarding, shot-matching, compositional consistency |
## Combining reference types
You can use multiple reference types together in a single generation. For example:
* Use **Characters** to anchor the subject and **Styles** to set the overall aesthetic.
* Use **Color Palettes** alongside **Effects** for tight control over both hue and visual treatment.
* Use **Camera Angles** with any other reference type to lock in both the composition and the look.
Each reference type operates on a different layer of the image, so combining them gives you more precise control without one overriding another.
# What is Image Studio?
Source: https://docs.imagine.art/image-tools/what-is-image
[**Image Studio**](https://www.imagine.art/image) is the intersection of cutting-edge AI research and intuitive design. It is your digital darkroom, allowing you to synthesise high-fidelity visuals from scratch, evolve existing work, and experiment with limitless artistic styles. Whether you are building a brand or an imaginary world, **Image Studio** provides the precision tools to make it a reality.
### **Powerful Creative Tools at Your Fingertips**
### **How it works?**
Image Studio offers a suite of AI-powered tools to enhance your creative workflow:
1. **Create** a new image using your prompt. Breathe life into your concepts using our most advanced AI models.
2. **Edit** your image editing through text prompts and image references and take total control of your compositions.
3. Use **Style Reference** to map the aesthetic, color palette, or brushwork of one image onto another. This allows you to explore different artistic movements while keeping your core content perfectly intact.
## Image Models
To view all available image models, [go here](https://docs.imagine.art/ai-models/image/imagineart-2-0).
# Connectors
Source: https://docs.imagine.art/imagine-computer/connectors
Give the agent real access to your actual tools — Slack, Notion, Gmail, GitHub, and 30 more — with per-action permission control.
Connectors let the agent read and act on your real accounts instead of working from a description alone. Find them under **Customize → Connectors**.
A visual hub diagram sits above a searchable **My connectors** grid — 33 real integrations at last count, including Slack, Notion, Gmail, GitHub, Google Workspace (Docs/Drive/Sheets/Slides/Meet/Calendar/Classroom), HubSpot, Jira, Linear, QuickBooks, Sentry, Stripe, and more. There's no category grouping — just one searchable list.
Click **Connect** on any connector. A confirmation modal shows the connector's name and description first — clicking **Connect** there redirects you to that provider's real OAuth sign-in page (confirmed for Slack: a genuine `slack.com` sign-in requesting a real, specific scope list). Nothing is faked or simulated — you're authorizing a real third-party connection.
Click into any connector (its name or the chevron on its row) to see **Tools permission** — every individual action the connector could take (for Notion: adding content blocks, appending code/media/table blocks, creating databases, and more), each with its own allow/deny toggle and a search box to filter the list. This is real per-action scoping, not a single on/off switch.
## Using a connector in a chat
There's no `@`-mention syntax for connectors. Instead, the prompt bar has a dedicated icon that opens a **Search connectors** picker — a compact version of the same list, with inline **Connect** for anything not yet linked, and a **Manage Connectors** link back to the full settings page.
Connect only what a given task actually needs. The per-action Tools permission list means you can, for example, let the agent read from Slack without giving it permission to post messages — set that up once per connector rather than re-deciding it in every chat.
# Customize
Source: https://docs.imagine.art/imagine-computer/customize
Give it your own files, let it remember things, and build a custom personality — three tabs alongside Connectors.
**Customize** in the sidebar has four tabs: [Connectors](/imagine-computer/connectors), **Sources**, **Memory**, and **Personalization**.
## Sources
Sources let you upload your own files so the agent can reference them in any chat.
Click **Create new Source**, give it a name and description, choose whether it's **org-shared** or **private**, then upload files. Accepted formats are documents (PDF, DOCX, DOC, PPTX, PPT, XLSX, CSV, MD, TXT) and a wide range of code/config files (JS, TS, PY, GO, JAVA, RS, SQL, JSON, YAML, XML, and more) — no image formats. Each source has a **Use in chat** action, plus a menu for **Download All** or **Delete**.
## Memory
A single **Reference saved memories** toggle (on by default) controls whether the agent saves and uses memories when responding. The **Memory from your chats** panel is meant to list facts it's automatically extracted from past conversations — there's no manual "add a fact" option. Memory can also personalize the queries it sends to search providers like Bing.
## Personalization
This isn't a settings panel — it's a custom-persona builder. Click **Create new personality** and fill in: **Your name**, **Personality name**, **Your profession**, **Personality traits** (free text, with suggested tags like Witty, Poetic, Encouraging, Gen Z), and **What Personality should know about you**. The result is a persona that "thinks and speaks exactly like you," per the tab's own description, rather than a simple tone-of-voice setting.
# Getting Started
Source: https://docs.imagine.art/imagine-computer/getting-started
The interface, the mode pills, and the two controls that shape every task: model tier and run mode.
## The sidebar
Go to **imagine.art/computer**. You'll land on the main chat view.
* **New Chat** — start a fresh conversation.
* **Sites** *(Beta)* — the separately-gated Imagine Build website tool.
* **Search Chats** — find a past conversation.
* **AI Tools** — shortcuts into ImagineArt's other generation tools.
* **[Customize](/imagine-computer/customize)** — Sources, Memory, and Personalization.
* **[Scheduled](/imagine-computer/scheduled-tasks)** *(Beta)* — recurring agent jobs.
* **Projects** and **Recents** — your saved work and chat history.
## Modes: presets on the same chat
The pills below the prompt box — **Chat, Sites, Images, AI Slides, AI Sheets**, and under **More**: **AI Videos, AI Docs, Graphs, AI Music, Deep Research, Plan Trips, Motion Design, AI Resumes, AI Contract** — mostly aren't separate tools. Each one (other than Sites) is the same chat framework on its own sub-route, with a mode-specific composer hint and starter prompts. **Deep Research**, for instance, adds its own quick-start chips and a Standard/Deep-research toggle; **AI Slides** adds a slide-count selector and a template gallery.
**Sites is the exception.** Clicking the Sites pill goes to the same invite-only Imagine Build gate mentioned in [What is Imagine Computer?](/imagine-computer/what-is-imagine-computer) — it isn't a more-open version of the sidebar's Sites item.
## Two controls that shape every task
Every prompt bar has two dropdowns worth understanding before you run anything.
### Model tier
* **Flash** — fast and efficient for everyday tasks, quick ideas, and lightweight problem-solving.
* **Pro** — more power for deeper research, bigger workflows, and multi-step tasks that need extra capacity.
* **Ultra** — maximum intelligence for your most complex, high-stakes, end-to-end work.
A separately promoted **Opus 5** banner sits above the prompt box — it isn't part of this tier dropdown, implying it's a distinct offering rather than a fourth tier.
### Run mode
* **Auto run without approval** — full access to create, edit, or run intensive actions without stopping.
* **Ask for approval** — meant to ask before creating or editing artifacts.
In testing, setting **Ask for approval** and running a real research task end-to-end did not actually pause for approval at any point, including at the report-writing step. Don't rely on this setting as a safety gate yet — treat every run as if it were Auto.
# Glossary
Source: https://docs.imagine.art/imagine-computer/glossary
A quick reference to terms used throughout this guide and inside Imagine Computer.
A quick reference to terms used throughout this guide and inside Imagine Computer.
| Term | Meaning |
| ----------------------------- | ------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------ |
| **Ask for approval** | A run mode meant to pause before creating or editing artifacts. In testing, it did not actually pause during a full run — see [Running a Task](/imagine-computer/running-a-task). |
| **Auto run without approval** | The run mode with full, uninterrupted access to create, edit, or run intensive actions. |
| **Connector** | A real, OAuth-authenticated link to a third-party app (Slack, Notion, Gmail, and 30+ more) that lets the agent read and act on your actual data. See [Connectors](/imagine-computer/connectors). |
| **Customize** | The sidebar section holding Connectors, Sources, Memory, and Personalization. |
| **Memory** | A toggle plus an auto-extracted list of facts the agent remembers across chats. |
| **Mode pill** | A preset on the same chat framework (Deep Research, Plan Trips, AI Slides, etc.) — not a separate tool, with one exception: Sites. |
| **Model tier** | Flash, Pro, or Ultra — speed vs. depth. A separately promoted "Opus 5" sits outside this tier ladder. |
| **OmniAgent** | Imagine Computer's internal product name. |
| **Personality** | A custom persona built in Personalization — name, profession, traits, and background — that shapes how the agent responds. |
| **Plan** | The numbered, editable step-by-step outline the agent proposes before running a task. |
| **Run mode** | The Auto/Ask-for-approval toggle controlling how much the agent does without stopping. |
| **Scheduled task** | A recurring agent job, configured with a trigger (repeat frequency + time), output format, and output destination. |
| **Sites** | Imagine Build, ImagineArt's website/app builder — currently Early Access, Invite Only, and the one part of Imagine Computer this guide couldn't fully explore. |
| **Source** | A named, uploaded collection of your own files (documents or code) available to reference in any chat. |
| **Tools permission** | The per-connector, per-action allow/deny list controlling exactly what a connector can do. |
# Quick Reference
Source: https://docs.imagine.art/imagine-computer/quick-reference
One-page summary — the core loop, the controls, and the one bug to watch for.
One page to return to whenever you need a fast reminder.
## Starting a task
1. Open **imagine.art/computer**.
2. Type a goal (or pick a mode pill: Deep Research, Plan Trips, AI Slides, etc.).
3. Review the **plan** → **Start Research** (or equivalent) → watch the live two-pane progress view → get a finished, editable document.
## The two controls on every prompt bar
| Control | Options |
| -------------- | ------------------------------------------------------------------------------------- |
| **Model tier** | Flash (fast) · Pro (deeper) · Ultra (max) — "Opus 5" is a separate, promoted offering |
| **Run mode** | Auto run without approval · Ask for approval (unreliable — didn't pause in testing) |
## Customize (sidebar)
| Tab | What it's for |
| ------------------- | -------------------------------------------------------------------- |
| **Connectors** | 33 real integrations, OAuth-based, per-action Tools permission |
| **Sources** | Upload your own files (docs + code, no images) to reference in chats |
| **Memory** | Toggle + auto-extracted facts remembered across chats |
| **Personalization** | Build a custom persona (name, profession, traits, background) |
## Scheduled tasks
* 24+ templates, filterable by Analysis/Automation/Communication/Information, or start from scratch.
* Trigger = Repeat (Once/Daily/Weekdays/Weekly/Monthly/Quarterly/Annually/Custom) + time.
* Output format: Text / Artifacts / Audio. Output destination: in-chat / Slack channel / Slack DM / Gmail.
## What's gated
**Sites** (the "Imagine Build" website builder) is Early Access, Invite Only — everything else in this guide is fully usable.
**The one bug to remember:** a task can report "Done" while its actual document is missing the real write-up (only a Sources/citation list). Always open and check the finished document yourself before trusting a completed status.
# Running a Task
Source: https://docs.imagine.art/imagine-computer/running-a-task
From a goal to a plan to a finished document — how an agent task actually runs, and one real failure mode to watch for.
Submitting a prompt doesn't generate an answer directly — it produces a plan you review first, then a live-running task you can watch.
After you describe what you want (e.g. in **Deep Research** mode), the agent proposes a numbered plan — its research angles, an analysis step, and a final write-up step — with an ETA.
There's no drag-and-drop plan editor. Click **Edit Plan** and describe what you want changed in plain language (different ranking criteria, a dropped angle, an added axis) — the agent replies conversationally and regenerates the whole plan.
Click **Start Research**. The view splits into two panes: your chat collapses to a short "Working…" status on the left, while the right-hand canvas streams a timestamped log — search angles, result counts, which sources it's reading, and interim takeaways.
Once research wraps up, the right pane switches into a Google-Docs-style editor titled after your topic, with **Contents**, **Download**, and **Share** controls and a live word count. It's a fully editable rich-text document, not a read-only report — you can keep working in it directly.
## A real failure mode to watch for
In testing, a task reported completion ("Done — the report is in your chat") while the actual document contained only a **Sources** heading — the synthesized write-up itself was silently missing, despite the chat log showing sources had been read and a report was being written. Separately, the stated "4-5 minutes" estimate isn't a reliable ceiling — a research run can still be mid-flight well past that.
Always open and check the actual document before trusting a "Done" status or relying on it for anything time-sensitive. A completed status is not proof the content is actually there.
## Run modes
The mode pill next to the model dropdown controls how much the agent asks before acting:
* **Auto run without approval** — full access to create, edit, or run intensive actions without stopping.
* **Ask for approval** — meant to pause before creating or editing artifacts.
In testing, **Ask for approval** did not actually pause at any point during a full research run, including at the report-writing step. Treat every run as if it were Auto until this is confirmed to work as described.
# Scheduled Tasks
Source: https://docs.imagine.art/imagine-computer/scheduled-tasks
Set an agent task to run on a repeat, instead of triggering it yourself every time.
**Scheduled** *(Beta)*, in the sidebar, turns a one-off agent task into a recurring job.
Click **Scheduled** in the sidebar. It tracks **Active**, **Pending**, and **Total** job counts, and lists your scheduled agents below.
Click **New schedule** to open a template gallery — over 24 real, named templates (Annual Vision Deck, Bedtime Wind-Down Audio, Daily Affirmation Audio, Daily Habit Tracker Prompt, and more), filterable by **Analysis**, **Automation**, **Communication**, and **Information**. Each template shows its own default schedule (e.g. "Daily at 7:00 AM"). Click **Start from scratch** if none fit.
Fill in a **Name** and **Task description**, then set the **Trigger** — click it to open **Repeat** options (Run once, Daily, Daily on weekdays, Weekly, Monthly, Quarterly, Annually, or Custom) and a time picker. No raw cron syntax is exposed.
Choose an **Output format** — **Text** (a simple plain-text message), **Artifacts** (attaches relevant files, e.g. a Document), or **Audio** (delivers the message as speech) — and an **Output destination**: the default in-Computer chat, a Slack channel, a Slack DM, or Gmail.
Templates are a faster starting point than they look — even if the exact use case doesn't match yours, copying a template's trigger/output setup and swapping the task description is quicker than configuring everything from scratch.
# What is Imagine Computer?
Source: https://docs.imagine.art/imagine-computer/what-is-imagine-computer
An agentic AI assistant that plans and runs multi-step tasks for you — research, documents, connected tools, and scheduled work.
Imagine Computer is ImagineArt's agentic AI product (internally called OmniAgent). Instead of a single prompt-in, result-out generation, you give it a goal — research something, write a document, work across your connected apps — and it plans the steps, runs them, and hands you back a real result.
## What you can do with it
* **[Run agent tasks](/imagine-computer/running-a-task)** — describe a goal, review or edit the plan, and let it research, browse, and write for you.
* **[Connect your tools](/imagine-computer/connectors)** — link real accounts (Slack, Notion, Gmail, GitHub, and 30 more) so the agent can act on your actual data.
* **[Customize](/imagine-computer/customize)** — give it your own files to reference, let it remember facts across chats, and build a custom personality.
* **[Schedule recurring work](/imagine-computer/scheduled-tasks)** — set a task to run on a repeating schedule instead of triggering it by hand every time.
## Entry point
Open **imagine.art/computer** — it redirects to the main chat view. Imagine Computer is reachable from a featured card on the ImagineArt homepage rather than the primary left-nav tool list.
**Sites is separately gated.** The **Sites** mode pill and sidebar item point to "Imagine Build," a website/app builder that's currently **Early Access · Invite Only** — a disabled "Start Creating" button and an "Apply for access" link are as far as this account (or any non-invited account) can go. Everything else described in this section is fully available.
Imagine Computer is genuinely autonomous, and that comes with a real, observed failure mode: a task can report "Done" while the actual written content is missing or incomplete (see [Running a task](/imagine-computer/running-a-task#a-real-failure-mode-to-watch-for)). Always check the actual output before trusting a "completed" status.
# ImagineArt Documentation Home
Source: https://docs.imagine.art/index
Welcome to the **ImagineArt Help Center**. You'll find guides, references, and tips here to help you create at your best.
Sign up, understand credits, and create your first image or video in minutes.
Generate, edit, and enhance images using state-of-the-art AI models.
Turn prompts and images into cinematic video clips with powerful AI video generation.
Build node-based creative pipelines that chain 50+ AI models on an infinite canvas.
## Quick Start
Visit [imagine.art](https://www.imagine.art/) and sign up with Google, Facebook, Discord, or email.
Start for free with 100 daily credits, or [subscribe](https://www.imagine.art/subscription) to unlock Pro models and more generations.
Head to the [Image tab](https://www.imagine.art/image), type a prompt, and click **Create**. Your images are ready in seconds.
Try [Video Tools](/video-tools/text-to-video) to animate your creations, or dive into [Workflows](/workflows/welcome) for advanced pipelines.
## Explore by Tool
Generate high-fidelity visuals from text prompts.
Refine and transform images with precision editing tools.
Use an image as a creative reference to guide generation.
Animate any static image into a fluid, AI-generated video.
Sync audio to video with realistic lip animations.
Share credits and collaborate with your team on one plan.
Need help? Email [support@imagine.art](mailto:support@imagine.art) or join the [Discord community](https://discord.gg/QdkX3tcu).
Welcome to the **ImagineArt Help Center**. You'll find guides, references, and tips here to help you create at your best.
Sign up, understand credits, and create your first image or video in minutes.
Generate, edit, and enhance images using state-of-the-art AI models.
Turn prompts and images into cinematic video clips with powerful AI video generation.
Build node-based creative pipelines that chain 50+ AI models on an infinite canvas.
## Quick Start
Visit [imagine.art](https://www.imagine.art/) and sign up with Google, Facebook, Discord, or email.
Start for free with 100 daily credits, or [subscribe](https://www.imagine.art/subscription) to unlock Pro models and more generations.
Head to the [Image tab](https://www.imagine.art/image), type a prompt, and click **Create**. Your images are ready in seconds.
Try [Video Tools](/video-tools/text-to-video) to animate your creations, or dive into [Workflows](/workflows/welcome) for advanced pipelines.
## Explore by Tool
Generate high-fidelity visuals from text prompts.
Refine and transform images with precision editing tools.
Use an image as a creative reference to guide generation.
Animate any static image into a fluid, AI-generated video.
Sync audio to video with realistic lip animations.
Share credits and collaborate with your team on one plan.
Need help? Email [support@imagine.art](mailto:support@imagine.art) or join the [Discord community](https://discord.gg/QdkX3tcu).
# Imagine MCP
Source: https://docs.imagine.art/integrations/imagine-mcp
Connect ImagineArt to Claude, Cursor, OpenClaw, Hermes, and any MCP client — no API key, billed through your existing imagine.art credits.
Imagine MCP connects ImagineArt to **Claude, Cursor, OpenClaw, Hermes**, and any client that speaks the Model Context Protocol. Use ImagineArt's full set of creative tools — image, video, music, fashion, ads, and more — with no API key, billed through your existing imagine.art credits.
By the end of setup, you can tell your agent "Generate an image of a red bicycle at sunset" and get the finished media back in the conversation.
In your client, add a custom MCP server pointing to `https://mcp.imagine.art`.
Sign in with your imagine.art account when prompted — no API key.
Ask your agent to create something. Generations draw from your existing credits.
## What is MCP?
**MCP (Model Context Protocol)** is an open standard that lets an AI agent connect to external tools. ImagineArt runs a hosted MCP server; your AI client (Claude, Cursor, and so on) acts as the host that connects to it.
When you register the ImagineArt server, its creative tools appear as native tools your agent can call. The connection is:
* **Hosted** — the server lives at `https://mcp.imagine.art`. There's nothing to run or install.
* **Authenticated by your account** — you sign in with your imagine.art login. There is no API key to generate, store, or rotate.
* **Billed through your credits** — no separate MCP pricing. The free tier includes 100 credits/day.
Two addresses, don't mix them up. `https://www.imagine.art/mcp` is the **information page** you read in a browser. `https://mcp.imagine.art` is the **server endpoint** that goes into your client's setup. Use the server endpoint for setup.
## Why Imagine MCP
Most creative MCP servers require API keys, separate billing, and credential management across platforms. Imagine MCP skips all of that.
* **No API key needed** — every request authenticates through your imagine.art account.
* **Every tool in one connection** — image, video, music, upscaling, background removal, fashion, and ad generation are all reachable through a single server, so you can chain them in one conversation.
* **Uses your current balance** — runs on the same credit system as the platform; your existing plan and balance carry over with no extra charges.
* **Zero data retention** — each request is processed independently. Prompts, outputs, and session data aren't retained.
## What you can build
Tell your agent what you're working on and it picks the right tools, chains them together, and delivers production-ready results without you leaving the conversation. Generate an image, upscale it, strip its background, then feed it into a video tool. Shoot a model in an outfit, then animate the still into a campaign clip. Turn a product photo into a finished ad — all in one agent session.
## Available tools
### Six core tools
| Tool | What it does |
| ------------------ | ------------------------------------------------------- |
| Text-to-image | Generate images from a prompt at multiple resolutions |
| Text-to-video | Generate short clips from a prompt or a reference image |
| Music generation | Produce original music or instrumentals from a prompt |
| Image upscaler | Enhance and increase the resolution of an image |
| Background remover | Cleanly strip the background from an image |
| Balance inquiry | Check your remaining credits and renewal date |
### Guided studio workflows
Beyond the six core tools, two full multi-step workflows are available — each walks your agent through a short setup, then generates a polished, production-ready result:
* **Fashion Studio** — compose a model, wardrobe, scene, and pose into a finished photoshoot, or animate a still into a campaign video. See [Fashion Studio](#fashion-studio) below.
* **Ad Studio** — turn a product and an avatar into a finished ad, image or video, using a format, hook, and setting. See [Ad Studio](#ad-studio) below.
### Specialized creative tools
Purpose-built recipes that combine the core engines into polished, ready-to-use outputs:
* **Logo generation** — clean, scalable brand marks
* **3D logo animation** — turn a flat logo into a cinematic reveal
* **Cinematic product ad** — animate a product photo into a commercial clip
* **Giant product showcase** — surreal building-scale product hero shot
* **UGC lifestyle try-on** — authentic influencer-style product photos
* **Instagram post** — scroll-stopping hero image with caption and hashtags
* **YouTube thumbnail** — high-CTR 16:9 thumbnail with overlay guidance
* **Interior design** — redesign a room from a photo or a concept
* **Jewelry video** — luxury macro product commercial
* **Cooking video** — turn a person's photo into a tutorial clip
* **Drone/aerial video** — sweeping flyover, orbit, and top-down shots
### Workflow and account helpers
* **Upload images** — bring your own reference assets into a generation
* **List generations/uploaded assets** — browse what you've made
* **Select organization and folder** — choose the workspace and where outputs are saved
## Fashion Studio
Fashion Studio is a guided, multi-step workflow for AI fashion photoshoots and fashion videos — model, wardrobe, scene, and pose composed into finished editorial or catalogue shots, or animated into short campaign clips.
Ask your agent for a fashion shoot — for example, "Generate me a Fashion post."
Your agent walks through organization → shoot type → project, then model → wardrobe → background → pose.
It generates the shoot and shows you the results. From there you can animate any still into a video.
### How it works
Fashion Studio shares one setup, then splits into two branches:
* **Shared setup:** organization → shoot type (editorial or catalogue) → project
* **Photoshoot branch:** model → wardrobe (outfit + footwear/accessories, optional) → background (optional) → pose (optional) → generate
* **Video branch:** pick a finished still → template or free-text motion → camera movement (optional) → duration/aspect/resolution → generate
| Step | What it does |
| ------------------- | ------------------------------------------------------------------------------------------ |
| Fashion model | Select an existing AI model, or create one from your own reference photos |
| Wardrobe | Select an outfit (top/bottom or dress) and, optionally, footwear or accessories |
| Background and pose | Choose a scene and pose — or shoot at a plain root with no background selected |
| Generate photoshoot | Compose model + wardrobe + scene + pose into finished photos |
| Animate to video | Turn any finished still into a short clip using an editorial or catalog template |
| Composite shoot | Merge multiple stills (for example, the same model across shots) into one reconciled frame |
Generating a photoshoot or a composite requires an active (paid) subscription. Free-tier organizations can browse and set up a project, but generation needs an upgrade.
### Example workflow
> Create an editorial lookbook shot: put the navy blazer and the white sneakers on my "Studio Model 1," in a soft daylight loft setting, 4:3.
The agent walks org → shoot type → project (reusing your existing project if you have one), then model, wardrobe, and background selection, then generates the shoot and shows the results. Ask it to "turn that into a 6-second campaign clip" afterward and it animates the still using a matching video template.
## Ad Studio
Ad Studio turns a product photo (or an existing product) into a finished ad — image or video — by chaining together a product, an avatar, a format, and an optional hook and setting.
Ask your agent to create an ad — for example, "Make a UGC-style ad for my water bottle."
It walks you through picking (or adding) a product and an avatar, then a format, hook, and setting.
Confirm resolution, aspect ratio, and duration, and it generates the ad.
### The pipeline
Ads are generated through a sequential setup: organization → marketing project → product → avatar → format → hook → setting → generate.
| Step | What it does |
| ----------------- | ---------------------------------------------------------------------------------------------- |
| Marketing project | The campaign container the ad is linked to — pick an existing one or create a new one |
| Product | Pick an existing product, or add one by uploading a photo or pasting a product URL |
| Avatar | Pick an existing AI avatar, or create one from a text prompt or reference photos |
| Format | The ad type — for example, UGC, testimonial, unboxing. Skipped if you already named a format |
| Hook and setting | Optional — an opening hook and a scene/setting for the ad |
| Generate | Compose everything into the finished ad, at your chosen resolution, aspect ratio, and duration |
Resolution, aspect ratio, and duration are all required before generating — confirm each first so credits aren't spent on an unintended output.
### Example workflow
> Create a testimonial-style ad for my ceramic mug using my "Sarah" avatar, 9:16, 1080p, 10 seconds.
The agent resolves the mug as your product and Sarah as your avatar, matches "testimonial" to its format, confirms the hook/setting, then generates the ad at the specs you gave.
## Passing parameters and references
You control each generation through a few simple parameters. You don't pass these as raw code — just describe them to your agent (for example, "make it 16:9, 4K, using the veo model") and it maps them to the right tool.
### Common parameters
| Parameter | Applies to | Notes |
| ------------ | -------------- | -------------------------------------------------------------------------- |
| Prompt | All generators | The text description of what to create |
| Model | Image/video | Each tool has a default; you can request a specific model |
| Aspect ratio | Image/video | For example, 1:1, 16:9, 9:16 — invalid values fall back to a supported one |
| Resolution | Image/video | Images: 1K/2K/4K. Video: 480p–4K (model-dependent) |
| Duration | Video | In seconds; supported lengths vary by model (commonly 4–15s) |
### Using references
* **By URL** — pass the URL of an existing image as a reference.
* **By upload** — upload your own image(s) directly and use them as references.
* **Multiple references** — for video, you can supply more than one reference image (and on some models, reference video or audio) to guide the result.
## Example workflows
Each example is just what you'd type to your agent.
**A single image**
> Generate an image of a red bicycle at sunset, 16:9, 4K.
The agent calls text-to-image and returns the finished image in the conversation.
**Chain tools into a product hero shot**
> Generate a sleek matte-black water bottle on a marble surface, then upscale it and remove the background.
Image → upscale → background removal, all in one session, ending with a transparent PNG ready to drop into a design.
**Turn your own photo into an ad**
> Here's my product photo — turn it into a 6-second cinematic ad with a luxury mood.
Upload the reference, then the cinematic product-ad tool animates it into a commercial clip.
**Build a mini brand kit**
> Create a minimal wordmark logo for "Northwind Coffee", then animate it into a 3D reveal.
Logo generation → 3D logo animation, producing both a static mark and a motion intro.
**Shoot and animate a fashion look**
> Put this jacket on a model for an editorial shoot, then turn the best still into a 6-second campaign clip.
Fashion Studio composes the shoot, then animates the chosen still using a matching video template.
**Build a UGC-style ad from a product photo**
> Here's my product — create a UGC-style ad with an avatar, 9:16, 15 seconds.
Ad Studio resolves the product and avatar, applies the UGC format, and generates the finished vertical ad.
## How to connect
### Claude
Launch Claude, then go to **Settings → Connectors → Add custom connector**.
Name it `Imagine MCP` and paste the URL: `https://mcp.imagine.art`
Click **Add → Connect**, sign in with your imagine.art account, then ask Claude to generate an image.
### Cursor
In Cursor, go to **Settings → MCP → Add new MCP server**, or edit `~/.cursor/mcp.json` directly.
Paste this configuration (the token comes from your signed-in imagine.art account):
```json theme={null}
{
"mcpServers": {
"ImagineArt": {
"type": "http",
"url": "https://mcp.imagine.art",
"headers": {
"Authorization": "Bearer "
}
}
}
}
```
Save and restart Cursor, then ask the agent to generate a hero image — it picks the right tool.
### OpenClaw
Make sure OpenClaw is installed, updated (`openclaw update`), and running. Then:
```bash theme={null}
# Add the server
openclaw mcp add imagine --url https://mcp.imagine.art --transport streamable-http --auth oauth --timeout 180 --connect-timeout 60
# Sign in (opens an authorization URL)
openclaw mcp login imagine
openclaw mcp login imagine --code 'YOUR_CODE'
# Load the new tools into running agents
openclaw mcp reload
```
**Authorizing:** `login` prints an `https://imagine.art/mcp/authorize?...` URL. Open it, approve access, then copy only the value between `code=` and `&` from the redirect (the "site can't be reached" page is normal). Wrap the code in single quotes so a stray `&` can't break the command. Codes expire in about 1–2 minutes — re-run `login` for a fresh one if needed.
```bash theme={null}
# Inspect
openclaw mcp status --verbose # list saved servers + auth state (no network call)
openclaw mcp show imagine # show this server's raw config
openclaw mcp probe imagine # connect and list tools (makes the network call)
# Maintain / reset
openclaw mcp configure imagine --timeout 180 --connect-timeout 60
openclaw mcp logout imagine # clear stored credentials (keeps the server)
openclaw mcp unset imagine # remove the server entirely
```
A healthy `status` shows `authorized` and `tokens=yes`. A probe should report the full tool list.
### Hermes and other agents (device flow)
Open a terminal and run:
```bash theme={null}
hermes mcp add ImagineArt --url "https://mcp.imagine.art"
```
For `Does this server require authentication? [Y/n]:`, type `Y` and then paste your bearer token.
For `Enable all 26 tools? [Y/n/select]:`, type `y`.
Then type `hermes` and start generating.
## Security and privacy
* **Your account is the only key** — authentication runs through your existing imagine.art login. No shared API keys, no separate credentials to secure.
* **Tokens stay local** — your client stores the OAuth token on your machine and refreshes it automatically.
* **Zero data retention** — the server processes each request independently and doesn't retain prompts, outputs, or session data.
* **Trust only your own sign-in** — never approve an authorization link or paste a code that came from anything other than your own client's login command.
## Troubleshooting
| Symptom | Cause | Fix |
| ---------------------------------------------------------- | ----------------------------------------------------------------------- | ----------------------------------------------------------------------- |
| Browser shows "This site can't be reached" after approving | Normal — nothing serves the local callback page | Ignore it; copy the `code` from the address bar |
| Login never finishes / command suspends | You included `&state=...`; the `&` backgrounded the command | Copy only up to the `&`, wrap the code in single quotes |
| `code is expired / invalid` | Auth codes expire in about 1–2 min | Re-run login, re-approve, use the fresh code |
| `Request timed out` (-32001) | The streaming connection is being dropped or slowed | Use generous timeouts; switch off VPN/proxy or try a stable network |
| Agent receives messages but never replies | Its language model is out of credits/keys | Top up or switch the agent's model |
| Ad generation fails citing an invalid format | A format name (for example, "UGC") was passed instead of its numeric id | Look up the id via the format list, or use the format picker, and retry |
| New project silently reuses an old one | Duplicate project names in the same folder | Rename projects distinctly, or select by creation date in the picker |
| Generation stalls or fails on credits | Organization credit balance exhausted | Check balance and top up, or switch organizations |
| Newly created avatar/product "not found" | Its id wasn't picked up from the picker's response | Re-select from the picker rather than reusing an old id |
**Connection vs. model:** a working ImagineArt connection only means the tools are available. Your agent still needs a working language model (with credits) to drive the conversation and decide to call those tools.
## FAQ
It uses the Model Context Protocol, an open standard that gives AI agents access to external tools. Once connected, your agent can generate images, create videos, produce music, upscale assets, remove backgrounds, run fashion photoshoots, build ads, and check your balance — all within a single conversation.
Claude, Cursor, Hermes, and OpenClaw. Any agent or client that speaks MCP can connect, including custom setups running locally or on a server.
No. Add the server URL in your agent's settings and authenticate through your imagine.art account. No keys to generate, store, or rotate.
Imagine MCP uses the same credit system as the platform. Each generation costs credits based on the tool and model selected, drawn from your existing plan. Check your balance and renewal date anytime with the balance tool.
Images typically complete in a few seconds. Videos take longer depending on duration and model. Generation runs asynchronously — your agent polls for results and delivers them the moment they're ready.
Images at multiple resolutions, short videos, and original music — all from a single prompt, plus full fashion photoshoots and ad campaigns through the studio workflows. You can chain tools in sequence within one agent session.
Yes. Fashion Studio composes a model, wardrobe, scene, and pose into finished photoshoots or short campaign videos. Ad Studio turns a product and an avatar into a finished ad, image or video, using a format, hook, and setting. Fashion Studio generation is limited to paid organizations.
# ImagineArt for After Effects
Source: https://docs.imagine.art/integrations/plugins/after-effects
Generate images and video, animate a still into a clip, upscale and clean up footage, and drop the results straight onto your comp, without leaving After Effects.
The ImagineArt panel docks inside After Effects so you can generate, animate, and clean up footage right where you're compositing — no browser tab, no export, no round trip.
## Install
Download `ImagineArt-AfterEffects.pkg` from the [After Effects plugin page](https://imagine.art/plugins).
Double-click the `.pkg` and follow the prompts. Quit After Effects first if it's open.
If macOS says it "could not verify" the developer: click **Done**, open **System Settings → Privacy & Security**, scroll to the Security section, and click **Open Anyway** next to `ImagineArt-AfterEffects.pkg`. Run the installer again and choose **Open** — you only do this once.
In After Effects, go to **Window → Extensions → ImagineArt**. Dock the panel and sign in once.
Get a free ZXP installer — [ZXPInstaller](https://zxpinstaller.com) or Anastasiy's Extension Manager — then drag `ImagineArt-AfterEffects.zxp` onto it (or **File → Install**). Quit After Effects first.
Reopen After Effects and go to **Window → Extensions → ImagineArt**. Dock the panel and sign in once.
## How the panel works
Open the panel and choose what to make — image, video, animation, upscale, or remove background.
Type a prompt, then set the model, aspect ratio, resolution, and duration. The panel shows the credit cost before you run anything.
The panel generates right where you're docked — no browser tab, no export, no round trip.
Add the result to your composition as footage, then keep working — the clip lands ready to animate.
Animate a still, upscale a clip, or clean a plate — every action targets the layer you already have selected.
## What you can do
| Tool | What it does |
| ---------------------------- | ------------------------------------------------------------------------ |
| Image generation (txt2img) | Describe an idea, pick a model and aspect ratio, generate in the panel |
| Video generation (txt2video) | Text-to-video with resolution and duration controls |
| Image to video (i2v) | Turn a still into motion, or animate from a start frame |
| Typo animation | Animate a typography still into a moving title clip |
| Logo animation | Turn a static logo into an animated reveal |
| Poster animation | Bring a poster to life as a short animated loop |
| Remove object from video | Erase an object from a clip with a prompt — the background fills back in |
| Video object replace | Swap an object in a clip for something new, by description |
| Video reframe | Reframe a clip to a new aspect ratio while keeping the subject in shot |
| Video background changer | Replace a clip's background with a prompt — the subject stays put |
| AI video translator | Translate a clip into another language with natural voice and lip-sync |
| Text to speech | Turn a script into narration, dropped onto your timeline |
| Creative upscale | Boost any image or selection to crisp, hi-res |
| Remove background | Clean cutouts in one click |
| Camera angle | Re-angle a shot by spinning an orbit globe; the subject stays intact |
| Relight | Drag a light globe to relight a scene from any direction |
## FAQ
No — nothing comes from the Marketplace. On macOS you download the `.pkg` and run it; on Windows you install the small `.zxp` with a free ZXP installer.
Quit After Effects completely and reopen it — the extension is only picked up on launch. Then check **Window → Extensions → ImagineArt**. If it's still missing, reinstall with After Effects closed.
Every result has an **Add to comp** action — it adds the image or video as a footage layer, ready to animate. You can also select a footage layer already in your comp to feed the edit tools.
Return to After Effects and give it a moment — the panel picks up the session automatically. If it still doesn't connect, copy the code shown after sign-in and paste it into the panel; it's accepted as a manual fallback.
Generation, animation, upscaling, and background removal use the same ImagineArt credits as the web app, and the panel shows the estimated cost next to each model and tool before you run anything.
Free with your ImagineArt account. Works on macOS and Windows.
# ImagineArt for Figma
Source: https://docs.imagine.art/integrations/plugins/figma
Generate images and video, run upscales and background removal on any canvas selection, and drop the results straight into your file, without leaving Figma.
ImagineArt is live on Figma Community. Run it as a plugin inside any file — browser or desktop app — to generate images, video, and editable vectors, or edit what's already on your canvas.
## Install
Open the [Figma Community listing](https://imagine.art/plugins), or in Figma open **Plugins** and search "ImagineArt."
Click **Open in…** from the listing, or in any file open **Plugins → ImagineArt**. It runs everywhere Figma plugins run, browser included.
A browser tab opens — approve, and you're in. If it doesn't connect automatically, click **Copy sign-in code** and paste it into the panel. The session sticks across all your files.
## How the plugin works
Open the plugin in any file, browser or desktop app, and choose what to make — image, video, vector, upscale, or remove background.
Type a prompt and set the model and aspect ratio. The panel shows the credit cost before you run anything.
The plugin generates right inside the panel — no tab-hopping, no exporting out of Figma.
One click drops the result onto your canvas as a named, full-resolution layer, placed beside your selection.
Pick the Vector tool and the plugin drops your graphic onto the canvas as real, editable vector layers — recolor and reshape them like anything else in your file.
## What you can do
| Tool | What it does |
| ---------------------------- | ------------------------------------------------------------------------------------ |
| Image generation (txt2img) | Describe an idea, pick a model and aspect ratio, generate in the panel |
| Video generation (txt2video) | Text-to-video with resolution and duration controls |
| Image to video (i2v) | Turn a still into motion, or animate from a start frame |
| Creative upscale | Boost any image or selection to crisp, hi-res |
| Vector (Recraft → SVG) | Generate clean, editable vector graphics that drop in as real SVG layers |
| Remove background | Clean cutouts in one click |
| Camera angle | Re-angle a shot by spinning an orbit globe; the subject stays intact |
| Relight | Drag a light globe to relight a scene from any direction |
| Generative Fill | Brush a mask and fill, replace, or extend anything in the frame |
| Color grading | Cinematic presets, palettes, and color-correct sliders, applied in place — no tokens |
| Edit / reference (img2img) | Reimagine an image from a prompt or reference while keeping its structure |
| Outfit try-on | Dress a subject in any outfit from a reference image, fit and folds intact |
## FAQ
From Figma Community. It installs and auto-updates from there — no files to manage.
Both. It runs everywhere Figma plugins run, browser and desktop app alike.
Inserted on the canvas next to your selection, sized and named to match. They stay fully editable Figma layers.
Figma only allows video fills on paid plans, so placing generated video on the canvas needs one. Images work on any plan.
Return to Figma and give it a moment — the plugin picks up the session automatically. If it still doesn't connect, copy the code shown after sign-in and paste it into the plugin; it's accepted as a manual fallback.
Generation, upscaling, and background removal use the same ImagineArt credits as the web app, and the plugin shows the estimated cost next to each model and tool before you run anything.
Free with your ImagineArt account. Runs in the browser and the desktop app.
# ImagineArt for Framer
Source: https://docs.imagine.art/integrations/plugins/framer
Generate images and video, run upscales and background removal on any selection, and drop the results straight onto your Framer canvas, without leaving Framer.
ImagineArt is live on the Framer Marketplace. Run it as a panel inside any Framer project — browser or desktop app — to generate images and video or edit what's already on your canvas, without exporting or tab-hopping.
## Install
Open the [Framer Marketplace listing](https://imagine.art/plugins), or in Framer open the **Plugins** menu and search "ImagineArt."
Click **Install** from the listing, then open ImagineArt from Framer's Plugins menu. It runs everywhere Framer runs, browser included.
A browser tab opens — approve, and you're in. If it doesn't connect automatically, click **Copy sign-in code** and paste it into the panel. The session sticks across all your projects.
## How the plugin works
Run the plugin in Framer and choose what to make — image, video, upscale, or remove background.
Type a prompt and set the model and aspect ratio. The panel shows the credit cost before you run anything.
The plugin generates right inside the panel — no tab-hopping, no exporting out of Framer.
One click drops the result onto your canvas as a named, full-resolution layer, placed beside your selection.
Point the same panel at art you already have — upscale it to crisp hi-res or knock out the background in one click, no round-trips, no re-exporting.
## What you can do
| Tool | What it does |
| ---------------------------- | ------------------------------------------------------------------------- |
| Image generation (txt2img) | Describe an idea, pick a model and aspect ratio, generate in the panel |
| Video generation (txt2video) | Text-to-video with resolution and duration controls |
| Image to video (i2v) | Turn a still into motion, or animate from a start frame |
| Creative upscale | Boost any image or selection to crisp, hi-res |
| Remove background | Clean cutouts in one click |
| Camera angle | Re-angle a shot by spinning an orbit globe; the subject stays intact |
| Relight | Drag a light globe to relight a scene from any direction |
| Edit / reference (img2img) | Reimagine an image from a prompt or reference while keeping its structure |
## FAQ
From the Framer Marketplace. It installs and auto-updates from there — no files to manage.
Both. It runs everywhere Framer runs, browser and desktop app alike.
Inserted on the canvas next to your selection, sized and named to match. They stay fully editable Framer layers — move, mask, and compose them like anything else in your project.
Return to Framer and give it a moment — the plugin picks up the session automatically. If it still doesn't connect, copy the code shown after sign-in and paste it into the plugin; it's accepted as a manual fallback.
Generation, upscaling, and background removal use the same ImagineArt credits as the web app, and the plugin shows the estimated cost next to each model and tool before you run anything.
Free with your ImagineArt account. Runs in the browser and the desktop app.
# ImagineArt for Premiere Pro
Source: https://docs.imagine.art/integrations/plugins/premiere-pro
Generate images and video, upscale and clean up shots, and conform Film Studio projects into editable Premiere timelines, without leaving the edit.
The ImagineArt panel docks inside Premiere Pro so you can generate, upscale, and edit footage right where you're cutting — no browser tab, no export, no round trip. It also rebuilds an exported Film Studio project as an editable Premiere sequence, scenes mapped to clips with transitions already applied.
## Install
Download `ImagineArt-Premiere.pkg` from the [Premiere Pro plugin page](https://imagine.art/plugins).
Double-click the `.pkg` and follow the prompts. Quit Premiere first if it's open.
If macOS says it "could not verify" the developer: click **Done**, open **System Settings → Privacy & Security**, scroll to the Security section, and click **Open Anyway** next to `ImagineArt-Premiere.pkg`. Run the installer again and choose **Open** — you only do this once.
In Premiere, go to **Window → Extensions → ImagineArt Film Studio**. Dock the panel and sign in once.
Get a free ZXP installer — [ZXPInstaller](https://zxpinstaller.com) or Anastasiy's Extension Manager — then drag `ImagineArt-Premiere.zxp` onto it (or **File → Install**). Quit Premiere first.
A one-click `.exe` installer is also available from the plugin page. It isn't code-signed yet, so Windows shows a SmartScreen warning — click **More info → Run anyway** to continue. The `.zxp` route skips this screen, which is why it's the default.
Reopen Premiere and go to **Window → Extensions → ImagineArt Film Studio**. Dock the panel and sign in once.
## How the panel works
Open the panel and choose what to make — image, video, upscale, remove background, or a full Film Studio export.
Type a prompt, then set the model, aspect ratio, resolution, and duration. The panel shows the credit cost before you run anything.
The panel generates right where you're docked — no browser tab, no export, no round trip.
Export a project from Film Studio on the web and the panel rebuilds it as an editable Premiere sequence, scenes mapped to clips with transitions applied.
Animate a still, upscale a clip, or clean a plate — every action targets the selection you already have open.
## What you can do
| Tool | What it does |
| ---------------------------- | ------------------------------------------------------------------------- |
| Film Studio Export | Conform a Film Studio project into an editable, ready-to-cut timeline |
| Image generation (txt2img) | Describe an idea, pick a model and aspect ratio, generate in the panel |
| Video generation (txt2video) | Text-to-video with resolution and duration controls |
| Image to video (i2v) | Turn a still into motion, or animate from a start frame |
| Remove object from video | Erase an object from a clip with a prompt — the background fills back in |
| Video object replace | Swap an object in a clip for something new, by description |
| Video reframe | Reframe a clip to a new aspect ratio while keeping the subject in shot |
| Video background changer | Replace a clip's background with a prompt — the subject stays put |
| AI video translator | Translate a clip into another language with natural voice and lip-sync |
| Text to speech | Turn a script into narration, dropped onto your timeline |
| Creative upscale | Boost any image or selection to crisp, hi-res |
| Remove background | Clean cutouts in one click |
| Camera angle | Re-angle a shot by spinning an orbit globe; the subject stays intact |
| Relight | Drag a light globe to relight a scene from any direction |
| Edit / reference (img2img) | Reimagine an image from a prompt or reference while keeping its structure |
## FAQ
No — nothing comes from the Marketplace. On macOS you download the `.pkg` and run it; on Windows you install the small `.zxp` with a free ZXP installer. Either way, open the panel in Premiere afterward.
Quit Premiere completely and reopen it — the extension is only picked up on launch. Then check **Window → Extensions → ImagineArt Film Studio**. If it's still missing, run the installer again with Premiere closed.
Return to Premiere and give it a moment — the panel picks up the session automatically. If it still doesn't connect, copy the code shown after sign-in and paste it into the panel; it's accepted as a manual fallback.
Make sure you're signed in inside the panel, then click **Export to Premiere** again in Film Studio. You can always build a project manually: open Film Studio Export in the panel and browse workspace → folder → project.
Each missing shot is named explicitly in the build summary. Open Premiere's Media Browser and relink them — everything else is already placed at the right position on the timeline.
Generation, upscaling, and background removal use the same ImagineArt credits as the web app, and the panel shows the estimated cost next to each model and tool before you run anything. Deductions happen on your account exactly as they would on imagine.art.
Free with your ImagineArt account. Works on macOS and Windows.
# ImagineArt for Shopify
Source: https://docs.imagine.art/integrations/plugins/shopify
Generate product photos and videos, put apparel on a model, spin it 360°, then publish straight to your products, all inside your Shopify admin.
ImagineArt is live on the Shopify App Store as an embedded admin app. Shoot your whole catalog without a studio — generate product photography, UGC-style video, and fashion try-ons, then publish the results directly to a product.
## Install
Open the [Shopify App Store listing](https://imagine.art/plugins), or search "ImagineArt" in the Shopify App Store from your admin.
Click **Install** and approve. The app asks for one permission — write access to products — and nothing else.
Click **Connect** inside the app. Your browser opens, you approve, and you're back in the admin. You bring your own ImagineArt credits — the app never charges through Shopify.
Go to **Apps → ImagineArt** in your admin sidebar. The session sticks, so you go straight back to the studio.
## How the app works
ImagineArt opens inside your admin with its own workspace — Studio, Image, Video, and One Tap Apps across the top.
Seventeen tools, grouped the way you'd shop for them: Generate, Product FX videos, Edit & enhance, Fashion, and UGC. Every card previews what it does before you spend a credit.
Write the shot you want, then pick a model, an aspect ratio, and how many variations. Drag in a reference image or pull one straight from your catalog.
Results fill in as they finish, up to four at a time, each with its own status. Images take about 30 seconds, so you can keep working elsewhere in your admin.
Attach the result to an existing product, or create a new draft product with the media already in place. It lands in your catalog at full resolution.
## What you can do
Seventeen tools across five groups:
| Group | Tools |
| ----------------- | ----------------------------------------------------------------------------------------------------------------- |
| Generate | Image generation (txt2img), video generation (txt2video) |
| Product FX videos | Product ad cinematic (brief → hero ad), UGC try-on, UGC video factory, 360° rotation, animate image |
| Fashion | Outfit try-on, AI clothes changer |
| Edit & enhance | Change background, variate, generative fill, outpaint, camera angle, relight, creative upscale, remove background |
Some highlights:
* **Product ad cinematic** — turn a product photo and a short brand brief into a cinematic ad clip.
* **UGC video factory** — generate a UGC-style creator video talking about your product, ready for paid social.
* **360° rotation** — spin a single product photo into a smooth turntable video for the product page.
* **AI clothes changer** — dress a model in your apparel, with a realistic fit and matching lighting.
## FAQ
The Shopify App Store. It installs in a couple of clicks and updates itself from there — nothing to download, no theme code to touch.
The app itself is free to install, with no subscription and no charge through Shopify. Generations run on ImagineArt credits from your own account, so you only ever pay for the generating you actually do.
One: write access to products. That's what lets the app attach generated media to a product or create a draft product. It doesn't read orders, customers, or payment data.
Yes, deliberately — you connect your own ImagineArt account and generations run on your own credits, so nothing is billed through Shopify.
Not for everything. Image and video generation work from a prompt alone. Try-ons, 360° spins, ad cinematics, relight, upscale, and background/inpaint work run on a photo you already have — a plain supplier shot is enough.
It only ever adds. Publishing to an existing product appends the new media to that product's media, and new products are created as drafts — nothing goes live until you publish it yourself. Titles, prices, and inventory are never touched.
The app stores only what it needs to work — your Shopify session and your ImagineArt connection. Uninstalling purges that automatically.
Free to install. Bring your own ImagineArt credits — the app never charges through Shopify.
# LipSync
Source: https://docs.imagine.art/lip-sync
The LipSync Node (labeled "Lipsync" in the node palette) synchronizes audio with video by automatically adjusting lip movements to match speech or music. It analyzes the audio track and generates realistic mouth animations that align perfectly with the sound, making it ideal for dubbing, voice-over work, multilingual content, and creating expressive character animations.
The Model dropdown now goes well beyond classic audio-driven lip-sync — it also includes full **avatar-generation models** (confirmed live: Heygen Avatar v4, Veed Fabric, Creatify Aurora, Kling Avatar 2.0 Pro/SD, Omnihuman 1.5, Kling Avatar 1.0 Pro/Standard, Infinitalk Audio) alongside the original sync-focused models. Those generate a talking avatar from a photo and an audio track, rather than re-syncing an existing video's mouth movements — pick one of those if you want an avatar presenter, not just corrected lip movement on footage you already have.
### How to Use
1. Add the Node:
* Click the Add (+) button and select LipSync from the Video node category.
2. Connect Video and Audio:
* Link a video from another node (such as Generate Video, Import, or Extend).
* Link an audio track from an audio node or import.
3. Configure Settings:
* Select your preferred Model and adjust other parameters from the Properties panel.
4. Generate:
* Click Run, and the AI will produce a video with lip movements synchronized to the audio.
### Choosing the Right Settings
| Setting | Type | Impact on Output |
| ------- | ------------ | ---------------------------------------------------------------------------------------------------------------------------------------- |
| Model | Dropdown | Selects the AI model used for lip-sync generation. Different models balance between accuracy, speed, and naturalness of mouth movements. |
| Seed | Number Input | A fixed number for reproducible results across generations. |
### Sample Use Cases
Generate or import a video, then replace the original audio with a dubbed version in another language. Use LipSync to automatically adjust the lip movements to match the new dialogue perfectly.
Record or generate a voice-over track, then sync it to existing video footage. LipSync ensures the speaker's mouth movements align with the new audio for a polished, professional result.
Pair generated character videos with custom audio to create expressive animations. LipSync automatically generates realistic mouth movements that match the speech, dialogue, or singing.
### LipSync Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Localize Image
Source: https://docs.imagine.art/localize-image
## Summary
The Localize Image Node translates the text inside an image into another language, keeping the rest of the image intact. Use it to adapt a poster, ad, or packaging shot for a new market without rebuilding it from scratch.
## How to Use
Click the Add (+) button and select **Localize Image** from the Image node category.
Link an image from another node, or upload one directly. This input is required.
In the Properties panel, choose the target language from the flag-based **Language** dropdown (defaults to English).
Click Run, and the node outputs a version of the image with its text translated into the selected language.
## Sample Use Cases
Take a finished ad image and generate a version with its on-image copy translated, instead of recreating the design in a new language by hand.
Run the same base image through Localize Image multiple times with different target languages to produce a set of market-specific variants.
# Motion Transfer
Source: https://docs.imagine.art/motion-transfer
Choose a reference video with clear, well-defined movements. Avoid clips with heavy occlusion (objects blocking the person), rapid camera movement, or multiple people.
## Summary
The Motion Transfer Node captures motion from a reference video and applies it to a character image, generating a new video where your character performs the exact movements from the reference. Connect a still character image and a reference video of someone dancing, walking, gesturing, or performing any action and the AI transfers that motion onto your character with realistic body movement and natural physics.
This node requires two inputs:
* **Character Image** — the subject you want to animate.
* **Reference Video** — the motion source you want to transfer from.
## How to Use
Click the Add (+) button and select Motion Transfer from the Video node category.
Link a character image via the Character Image input handle (marked in orange). This is the subject that will be animated with the transferred motion. Use a clear, well-lit image where the character's full body or relevant body parts are visible.
Link a video via the Reference Video input handle (marked in green). This is the motion source—the AI will extract the movement from this video and map it onto your character.
Select your Model, adjust Guidance Scale, Inference Steps, and other parameters from the Properties panel (see settings table below).
Click Run, and the AI will produce a video of your character performing the motion from the reference video.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| --------------- | ---------------------------------- | -------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Model | Dropdown (e.g., Wan 2.2 Move) | Selects the AI model for motion transfer. Different models handle body types, motion complexity, and rendering quality differently. |
| Guidance Scale | Slider (default: 1) | Controls how closely the output follows the reference motion. Lower values allow more creative freedom; higher values produce a stricter match to the source motion. |
| Resolution | Dropdown (e.g., 480p, 720p) | Determines the output video resolution. Higher resolution captures finer detail but takes longer to generate. |
| Inference Steps | Slider (default: 20) | Controls how many processing passes the AI runs. More steps generally produce smoother, higher-quality results but increase generation time. |
| Video Quality | Dropdown (e.g., High, Medium, Low) | Sets the overall rendering quality of the output video. |
| Seed | Number Input | A fixed number for reproducible results across generations. |
## Sample Use Cases
Grab a trending dance video as your reference and transfer the choreography onto an AI-generated character, brand mascot, or illustrated figure perfect for jumping on social trends without filming yourself.
Bring concept art or illustrated characters to life by transferring real human performances onto them. Record a quick acting reference, and your character inherits every gesture, head tilt, and body movement.
Combine a fashion image with a walking or posing reference video to showcase clothing in motion giving e-commerce customers a dynamic view of how garments look and flow on a moving body.
## Motion Transfer Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Multiple Camera Angles
Source: https://docs.imagine.art/multiple-camera-angles
## Summary
The Multiple Camera Angles Node lets you generate new views of a subject from different perspectives using a single reference image. With an interactive 3D camera controller and precise sliders for rotation, movement, and zoom, you can create realistic new angles without reshooting or re-rendering.
## How to Use
Click the Add (+) button and select Multiple Camera Angles from the Image node category.
Connect an image from another node (such as Generate Image or Edit Image), or upload one directly.
Drag the point on the interactive 3D camera controller to orbit around your subject visually. Fine-tune with the Rotation, Move, and Zoom sliders.
Click Generate, and the AI will produce a new image from your configured camera angle.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| --------------------- | ------------------------------- | ---------------------------------------------------------------------------------------------------------------------- |
| Camera Controller | Interactive 3D Widget | Drag to visually orbit the camera around your subject before fine-tuning with sliders. |
| Rotation (Left/Right) | Slider | Controls horizontal rotation. Negative values rotate left; positive values rotate right. |
| Move (Up/Down) | Slider | Adjusts vertical camera position. Negative moves down; positive moves up. |
| Zoom | Slider (default: 5) | Controls camera distance. Higher values zoom in; lower values pull back. |
| Aspect Ratio | Dropdown (16:9, 4:3, 1:1, etc.) | Defines the output image dimensions. |
| Wide Angle Lens | Checkbox | Simulates a wide-angle lens, expanding the field of view for more immersive perspectives. |
| Guidance Scale | Slider (default: 4.5) | Controls how strongly the input and camera settings influence the output. Higher values stay closer to your reference. |
| Seed | Number Input | A fixed number for reproducible results across generations. |
## Sample Use Cases
Upload/Generate a single product shot and rotate the camera to generate front, side, and top-down views, perfect for e-commerce listings without a full photo shoot.
Turn a single scene into a full storyboard of camera angles. Use Move (Up/Down) for a dramatic low-angle hero shot, Rotation for an over-the-shoulder perspective, and Zoom for a tight close-up all from one base image.
Generate eye-level, aerial, and dramatic low-angle views of a building from a single render. Enable Wide Angle Lens for expansive interior or exterior perspectives.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Music
Source: https://docs.imagine.art/music
Generate original music tracks from text descriptions using AI.
Describe the type of music you want — mood, instruments, energy, or genre — and the AI will compose an original track. Generated tracks range from **1 to 5 minutes** in length.
## Model
ElevenLabs' music generation model. Produces original, full-length instrumental and vocal tracks from text descriptions across a wide range of genres and styles.
## Styles
Choose a style to guide the sonic direction of your track:
`Dark Cinematic` `Cinematic Ambient` `Relaxing Ambient` `Bass Techno` `Emotional Piano` `Percussive Rhythm` `Commercial Pop` `Upbeat Pop` `Peak Energy` `High-Energy House` `Afro House Beats` `Reggaeton` `Arabic Groove` `70s Cambodian Rock` `80s Nu-Disco Revival` `Golden Hour Indie` `Wooden Slit Drum` `Brazilian Funk` `Baile Beats` `Rock Français` `18th Century Symphony`
# Output
Source: https://docs.imagine.art/output-node
## Summary
The Output Node marks the final result of a workflow or app. It takes a single connection from wherever your workflow ends, and displays that result — with playback controls for video — in a dedicated results panel. It's part of the new **App** node category, alongside an **Input** node for defining an app's starting parameters (Input wasn't fully verified during this pass — not yet documented on its own page).
## How to Use
Click the Add (+) button and select **Output** from the App node category.
Link the Output node's single input to whatever node produces your workflow's finished result.
Once run, the result appears in the Output node's panel — with a play control if it's a video — instead of just wherever the generating node left it on the canvas.
Output is distinct from the [Export](/workflows/utilities) node under Essentials. Export is for saving/downloading a result; Output is for designating and previewing a workflow's designated final result, particularly relevant when building a published App.
## Sample Use Cases
In a workflow with several generation and edit steps, connect Output to the last node so anyone running the workflow sees a clear, single final result rather than having to trace the canvas.
Pair Output with Input when building a workflow into an App — Input defines what a user provides, Output defines what they get back. See [App Builder](/workflows/app-builder).
# ImagineArt Teams
Source: https://docs.imagine.art/overview/imagineart-teams
Share a single subscription and credit pool with your collaborators using ImagineArt Teams.
ImagineArt Teams lets you consolidate your creative workflow by sharing a **single subscription** and **credit pool** across multiple accounts. It's built for agencies, studios, classrooms, and creator groups who need centralized billing and shared resources.
Every team member logs in with their own individual ImagineArt account. Teams simply connects those accounts under one shared plan and credit balance.
## How Teams work
The **Team Owner** creates the team through their account. The owner controls billing, owns all invoices, and is the only person who can change the plan or purchase top-ups.
The owner sends invitations to collaborators. Each plan sets a maximum number of seats — including the owner's seat — so you can only add members up to that limit. Members join via an invite link and use their own ImagineArt login.
The team has a **single shared credit pool**. Any member — owner or otherwise — can spend credits from it. All plan credits and any top-ups purchased by the owner go into this same pool.
If the team runs low on credits, only the owner can buy top-ups or upgrade the plan. Members have no access to billing controls.
## Roles and permissions
| | Team Owner | Team Member |
| --------------------------------------- | ---------- | ----------- |
| Creates the team | ✓ | — |
| Invites and removes members | ✓ | — |
| Upgrades or downgrades the plan | ✓ | — |
| Purchases credit top-ups | ✓ | — |
| Owns invoices and billing details | ✓ | — |
| Uses shared credits to generate content | ✓ | ✓ |
| Accesses plan features | ✓ | ✓ |
## Seats (member limits) by plan
Your plan determines the maximum number of people who can be in your team, including the owner.
| Plan | Maximum seats (including owner) |
| -------- | ------------------------------- |
| Standard | 3 |
| Ultimate | 6 |
| Creator | 20 |
Seat counts and pricing can change. Always confirm the current limits on the [Subscription page](https://www.imagine.art/subscription) and the [Teams plan page](https://www.imagine.art/teams-plan) before purchasing.
If you need more seats, upgrade to a higher plan. If you downgrade, you'll need to remove members until your team size falls within the new plan's limit.
## Multiple teams per account
You can belong to more than one team at the same time. Each team operates as a completely isolated workspace with its own subscription and credit pool.
**Example:** You can be the owner of one team for your studio while also being a member of a client's separate team. Those two teams have entirely separate subscriptions and credit balances — activity in one does not affect the other.
## Frequently asked questions
No. Only the Team Owner manages billing. Members have no visibility into or control over payment methods, invoices, or plan changes.
Generation stops for all members until credits are replenished. The owner can buy a top-up to add credits to the shared pool immediately, or upgrade the plan for a larger monthly allocation.
Yes. If you belong to multiple teams, you can switch between them using the team dropdown in the top-right corner of the platform. Each team's credits and settings are completely separate.
Ownership is tied to the account that created the team and holds the billing relationship. If you need to transfer ownership or restructure your team, contact [support@imagine.art](mailto:support@imagine.art).
The team's shared subscription ends along with the owner's plan. Members lose access to plan features and the shared credit pool until the owner renews or the team is re-established under an active subscription.
Yes. You can create and own multiple teams from a single account. Each team has its own billing, seat limit, and shared credit pool.
Need help with Teams setup or billing? Email [support@imagine.art](mailto:support@imagine.art).
***
* [Creating & Joining Teams](/overview/teams/create-and-join)
* [Credits & Billing](/overview/teams/credits)
* [Folder Sharing](/overview/teams/sharing)
# Creating & Joining Teams
Source: https://docs.imagine.art/overview/teams/create-and-join
Set up a new team workspace, invite members, and collaborate with comments.
## Creating or joining a team
* To **change your team/workspace**, start by navigating to the top left corner of your screen.
* Click on the **team name dropdown** located beside the upgrade button. This action will open a dropdown menu where you can choose from the teams you are already a part of.
* If you are looking to create a new team, select the option to create a new one.
Provide the **name** and **description** for this team and it will create your dedicated workspace.
You can optionally also select or upload a team icon. This will be displayed in the team profile picture beside its name.
Once you have successfully **created a team**, you will be greeted with a confirmation pop-up. This will include an **invitation link** that you can copy and distribute to the relevant individuals you'd like to add to your team. Alternatively, **you can manually invite members by entering their email addresses into the invitation field**, streamlining the process.
To reopen the invitation popup, click the invite button next to the desired workspace name.
Once a user requests access using your **invitation link**, the workspace admin will receive a **notification**. The admin can then review the pending request by accessing the teams dropdown. Here, admins have the ability to **accept or decline** the request, ensuring that only the appropriate individuals gain access to the collaborative workspace.
All assets created in a shared workspace will be visible to all individuals in that workspace.
## Team Comments
Each team member can **leave comments on assets generated by other members in the workspace**. Simply click on the asset you want to leave a comment on and go to the comments tab. These comments are visible to all team members, allowing for collaborative feedback and discussion.
Once addressed, comments can be marked as resolved, ensuring the team maintains clear and organised communication around their shared projects. The comment can be resolved or deleted by the person who left the comment, or replied to by team members.
# Credits & Billing
Source: https://docs.imagine.art/overview/teams/credits
How shared credits work, auto top-ups, and setting per-member credit limits.
## Credits, refresh, and top-ups
All monthly plan credits go directly into the team's shared credit pool. Top-up credits purchased by the owner also go into the same pool.
Subscription credits refresh monthly and do not roll over. Top-up credits never expire. If the team runs out of credits mid-cycle, the owner can purchase a top-up to replenish the shared pool without waiting for the next renewal.
For more detail, see [Understanding Credits](/overview/understanding-credits).
## Auto Top-ups
Auto top-ups are available exclusively on enterprise plans and can only be configured by **Admins** and **Owners**.
Auto top-ups automatically replenish the team's shared credit pool when credits fall below a threshold you define. Once configured, the system monitors the balance and triggers a top-up purchase without any manual action from the owner.
Navigate to your workspace settings from the team dropdown in the top-left corner.
Select the **Plan & Billing** tab inside your workspace settings.
Toggle **Auto Top-up** on, then set your threshold (the credit level that triggers a replenishment) and the top-up amount to add each time it fires.
Confirm and save. The system will automatically purchase the configured top-up amount whenever the shared pool drops below the threshold.
Set your threshold high enough to avoid generation interruptions during peak usage. Credits purchased via auto top-up never expire.
## Setting Credit Limits per Member
Administrators have the **ability to set monthly credit limits for individual team members**, or they can choose to allow members to use unlimited credits. When monthly credit limits are set for a member, their usage is automatically reset at the beginning of each month, allowing for consistent and manageable tracking of credit consumption over time.
To edit the monthly credit limit of your team members, simply go to your workspace settings:
Here, you can monitor detailed metrics, including total credits remaining, available seats and are able to invite members for team space. Click on "Manage Access" beside the invite members button:
To edit individual credit usage limits, navigate to the member's section in the workspace settings. Each **member's usage limits are clearly displayed on this page**. By clicking on a specific member, you can edit the respective field to adjust their monthly credit limit.
# Folder Sharing
Source: https://docs.imagine.art/overview/teams/sharing
Grant specific team members access to individual private folders without exposing your entire workspace.
Folder-level sharing is available on enterprise plans and is managed by **Admins** and **Owners**.
By default, your private folders are only visible to you. Folder-level sharing lets you grant specific team members access to individual private folders, so you can collaborate on a project without exposing your entire workspace.
## Sharing a private folder
Navigate to the **Assets** section from the left sidebar.
Navigate to the folder you want to share from your assets library.
Click the **···** menu on the folder and select **Share**.
Search for and select the team members you want to grant access. You can add multiple members at once.
Click **Share**. The selected members will see the folder in their shared view and can access its contents.
Shared folder access does not affect the assets inside — members can view and use the folder's content, but generation history and settings remain tied to the original owner's account.
# Understanding Credits
Source: https://docs.imagine.art/overview/understanding-credits
Learn what credits are, how they're consumed, the three types available, and how to manage your balance.
Credits are how ImagineArt measures usage across Image, Video, and other tools. Every generation or edit you run deducts credits from your balance. Understanding how credits work helps you choose the right plan and avoid running out mid-project.
## What credits are (and what affects usage)
Credits are deducted each time you generate or edit content. The number of credits consumed depends on several factors:
* **Tool** — Image, Video, and other tools have different base costs.
* **Model** — some models (especially Pro-tier models) cost more credits per generation than standard models.
* **Settings** — options such as video duration, output quality, and resolution affect consumption.
* **Number of results** — generating more variants or re-running a generation multiplies the credit cost.
The "estimated generations" shown on plan pages is a baseline for standard model settings. Pro and premium models will consume more credits per generation than those estimates suggest.
## Types of credits
Your account can hold up to three separate credit balances at any time:
### Free credits (daily)
All accounts without an active subscription receive **100 free credits every day**. These credits reset each day and are a great way to explore the platform at no cost.
Pro models (labeled "Pro" in the interface) require an active subscription — free daily credits cannot be used to access them. If you've upgraded and still can't access Pro models, contact [support@imagine.art](mailto:support@imagine.art).
### Subscription credits (monthly)
When you subscribe to a paid plan, you receive a monthly credit allocation that refreshes on your billing cycle date.
Unused subscription credits do not roll over to the next month. Any remaining subscription credits expire when your cycle renews. Check your renewal date any time under [**Billing & Subscription**](https://www.imagine.art/subscription) in your account settings.
For detailed information about how credits are assigned across monthly, quarterly, and yearly plans, see [Subscription Plans](/account/subscription-plans).
### Top-up credits (one-time)
Top-up credits are additional credits you purchase on top of your subscription, sold in bundles.
Top-up credits never expire, making them a good safety net for heavy production periods. However, you must have an active subscription to purchase them. If your subscription lapses, any remaining top-up credits can still be spent — but only on features available under your current plan tier (Free plan restrictions apply).
## Which credits get used first?
To help you use expiring credits before non-expiring ones, ImagineArt always draws from your balances in the following order:
1. **Free credits** (expire daily)
2. **Subscription credits** (expire at monthly renewal)
3. **Top-up credits** (never expire)
This order ensures you never waste credits that would otherwise expire.
## How to check your credit balance
1. Click your **profile picture** in the top-right corner of the platform.
2. Look at **Credits Left** for a quick summary. Click the tooltip beside it to view your credits broken down by type.
3. Click on "Buy Credits" to replenish your credits or upgrade your plan.
## Estimating your credit needs
Use this formula to estimate how many credits you'll need per month before choosing a plan:
```text theme={null}
monthly credits = (images per month × image cost) + (videos per month × video cost) + buffer
```
Add a **20–30% buffer** to your estimate if you expect to: iterate through multiple variants, switch between different models, or use higher-quality output settings.
**Example calculation**
Suppose you generate:
* 30 images/day at 24 credits/image
* 10 videos/day at 210 credits/video
* Working 22 days/month
| Category | Calculation | Subtotal |
| --------- | ------------------------------------ | ------------------------ |
| Images | 30 × 22 days = 660 images; 660 × 24 | 15,840 credits |
| Videos | 10 × 22 days = 220 videos; 220 × 210 | 46,200 credits |
| **Total** | 15,840 + 46,200 | **62,040 credits/month** |
Add your 20–30% buffer on top of that total when selecting a plan.
For the latest credit costs per model and studio, see:
* [Image Credit Consumption](/image-tools/credit-consumption)
* [Video Credit Consumption](/video-tools/video-credits)
## What happens if you run out of credits?
You have three options when you exhaust your current balance:
1. **Wait for your next monthly refresh** — subscription credits renew automatically on your billing date.
2. **Buy a top-up** — available any time while you hold an active subscription.
3. **Upgrade your plan** — if you routinely run out, a higher-tier plan gives you a larger monthly allocation.
## Buying top-ups (additional credits)
Top-ups may appear labeled as **Tokens** inside the app.
Select **Buy Tokens** in the top-right area of the platform interface.
Review the available credit bundles and select the one that fits your needs.
Finish the checkout process. Top-up credits are added to your account immediately and never expire.
## Usage Analytics dashboard
The Usage Analytics dashboard gives you a detailed breakdown of where your credits are going — by model, feature, and team member.
**To access analytics for your current workspace:**
1. Click your **profile picture** in the top-right corner.
2. Click **Usage Analytics**.
You can also go directly to [imagine.art/settings/analytics](https://www.imagine.art/settings/analytics).
**To access analytics for a different team or workspace:**
1. Click the **team dropdown** in the top-right corner.
2. Hover over the team or workspace you want to inspect.
3. Click the **Settings** icon that appears next to it.
4. In the left navigation panel, select **Analytics**.
On the **Analytics** page you can view:
* Which **models** you've used most frequently
* How many **credits** each model consumed
* Usage broken down by **user**, **feature**, and **model**
* A **date range filter** to view usage for any specific period
**Direct access:**\
You can go directly to the Analytics page here:\
[https://www.imagine.art/settings/analytics](https://www.imagine.art/settings/analytics)
This link will open the analytics dashboard for the **team/workspace you are currently working in**.
## Need help?
Email [support@imagine.art](mailto:support@imagine.art) and include:
* Your account email
* Your current plan
* A screenshot of your credits modal (optional)
# What is ImagineArt?
Source: https://docs.imagine.art/overview/what-is-imagineart
An overview of ImagineArt, how to create an account, and the core tools available to you.
ImagineArt is an AI-powered creative platform designed to close the gap between imagination and execution. Whether you're a professional designer, a digital artist, or someone exploring AI for the first time, ImagineArt gives you the tools to generate stunning images, cinematic videos, and immersive audio with just a few keystrokes.
More than 6 million people use ImagineArt to bring creative ideas to life.
## Sign up for an account
Visit [imagine.art](https://www.imagine.art) and click the **Sign In** button in the top-right corner of the dashboard. A secure window guides you through creating an account or returning to your existing workspace.
You can sign up using any of the following methods:
* **Google**
* **Facebook**
* **Discord**
* **Email**
Once signed in, you have access to your ImagineArt profile and can start creating immediately.
## Sign in to your account
Select the same method you used when you first signed up.
If you can't remember which method you used, try each option one by one — Google, Facebook, Discord, and then Email — until you find the one that matches your original sign-up.
When signing in with email, enter the same address you registered with and use your password. If you've forgotten it, select **Forgot Password** to receive a reset link.
## Key features
ImagineArt is built around two core creative areas.
Transform ideas into visual masterpieces using a full suite of generation and editing tools — from text-to-image generation to advanced editing with models like Nano Banana and Seedream v4.
Bring visuals to life with cinematic movement and professional-grade effects, powered by the latest AI video models including Google Veo 3, Kling 2.5, and Hailuo 02 Pro.
### Image Tools
Image Tools cover everything from initial generation to precise editing:
* **Image Generation** — turn descriptive prompts into high-fidelity art instantly.
* **Image Prompt** — combine a reference image with your text prompt so ImagineArt analyzes the core elements and creates new, unique visuals based on it.
* **Image Edit** — refine your work using advanced models like Nano Banana and Seedream v4. Change lighting, add objects, or tweak details with surgical precision.
* **Style** — maintain consistent visual styles across your images by uploading reference images to guide your creative direction.
### Video Tools
Video Tools give you cinematic control over motion and effects:
* **Video Generation** — generate videos, short clips, or dub existing videos based on your text and image prompts. Access the latest AI video models including Google Veo 3, Kling 2.5, and Hailuo 02 Pro.
* **Image to Video** — upload any static image and watch the AI breathe life into it with fluid, realistic animation.
* **Video Effects** — add realistic VFX to your videos and animate images for a truly cinematic experience.
# Account Deletion Policy
Source: https://docs.imagine.art/policies/account-deletion-policy
When you initiate an account deletion, your account enters a 7-day pending state before permanent removal. This action affects not only your personal profile but also any teams you are associated with.
### Immediate Impacts upon Scheduling
The moment you hit "Delete Account," the following actions occur automatically:
* Plan Cancellation: All active subscriptions and paid plans in every organization where you are an owner will be canceled immediately.
* Service Access: You will lose access to premium features associated with those plans, though your account remains "active but scheduled for deletion" for the duration of the grace period.
* Billing: No further charges will be incurred. Please note that ImagineArt does not provide pro-rated refunds for the remaining period of a canceled plan upon account deletion.
### Roles and Organizations
Because you may belong to multiple organizations, your deletion affects them based on your permission level:
| **Role** | **Impact on Organization** |
| -------- | --------------------------------------------------------------------------------------------------------------------------------------------- |
| Member | You are removed from the organization roster after 7 days. Shared assets remain accessible to the team. |
| Admin | You lose administrative privileges immediately. The team persists, but you will be removed as a contact. |
| Owner | Action Required. You cannot delete your account while remaining the sole owner of an active organization without making a choice (see below). |
#### Transfer of Ownership
If you are the Owner of an organization with other members added, the system will give you an option to transfer the ownership during the 7 days period.
### The 7-Day Grace Period
We provide a 7-day window to protect you from accidental deletions or a change of heart.
* The Cooling-Off Period: Your data remains on our servers for 7 days.
* Account Recovery: If you log back into ImagineArt within these 7 days, you will see an option to Cancel Deletion.
* Reactivating Plans: Note that while you can recover your account, canceled plans will not automatically restart. You will need to manually re-subscribe to your previous plans if you choose to recover your account.
### Permanent Deletion
Once the 7-day period expires:
* Your account, personal data, and any non-transferred organizations are permanently purged from our active databases.
* This action is irreversible. We cannot recover images, prompts, or team projects once this final stage is reached.
* Some data may remain in encrypted backups for a limited time as required by law or for security audits, but this data will not be accessible for standard use
# Cookie Policy
Source: https://docs.imagine.art/policies/cookie-policy
**Last updated:** January 12th, 2026
### 1. Introduction
This Cookie Policy explains how ImagineArt ("we", "our", or "us") uses cookies and similar technologies when you visit or use our platform. It describes what cookies are, why we use them, and how you can manage your preferences.
Cookies are small text files stored on your device that help websites function efficiently, improve user experience, and provide insights into how the platform is used.
### 2. Why We Use Cookies
We use cookies and similar technologies for the following purposes:
* To enable essential platform functionality, including security and authentication
* To remember user preferences and settings
* To understand how users interact with our platform so we can improve performance and usability
* To measure the effectiveness of marketing and promotional efforts
Some cookies are strictly necessary for the platform to operate and cannot be disabled. Other cookies are optional and can be managed through your cookie preferences.
### 3. Managing Your Cookie Preferences
You can manage your cookie preferences at any time through the **cookie management settings** available on the platform.
* **Necessary cookies** are always enabled, as they are required for the platform to function properly.
* **Preferences, Statistics, and Marketing cookies** are optional and can be enabled or disabled based on your choices.
Please note that disabling certain optional cookies may affect some features or functionality of the platform.
### 4. Types of Cookies We Use
#### 4.1 Necessary Cookies (Always Enabled)
These cookies support essential platform functionality, including security, authentication, core features, and other functions required for the platform to operate reliably.
These cookies include, but are not limited to:
| Cookie Name | Purpose | Expiration |
| ----------------------------- | ----------------------------------------------------------- | ------------- |
| `token` (BEARER\_TOKEN) | Stores authentication token for secure API requests | Varies |
| `refreshToken` | Enables secure session renewal | 7 days |
| `cachedSessionV1` | Caches session data to reduce repeated authentication calls | Varies |
| `impersonateUserToken` | Supports admin user impersonation for support purposes | Session-based |
| `countryCode` | Detects user country for pricing and localization | Session-based |
| `currency` | Stores preferred currency for subscriptions | Persistent |
| `fair_usage_tracker` | Enforces fair usage limits and abuse prevention | Varies |
| `folder_sidebar` | Persists UI state for navigation | Persistent |
| `refundNotificationExpiresAt` | Controls display of refund notifications | 1 hour |
#### 4.2 Preferences Cookies (Optional)
These cookies store non-critical user choices that improve consistency and personalization but are not required for core functionality.
| Cookie Name | Purpose | Expiration |
| -------------------- | ------------------------------------------------ | ---------- |
| `ia_pref_theme` | Stores selected theme (e.g., light or dark mode) | 6 months |
| `ia_pref_language` | Remembers preferred language or locale | 6 months |
| `ia_pref_layout` | Saves UI layout preferences | 3 months |
| `ia_pref_dismissals` | Records dismissed tips or pop-ups | 30 days |
| `ia_pref_defaults` | Stores default tool or workflow selections | 6 months |
#### 4.3 Statistics Cookies (Optional)
These cookies help us understand how the platform is used so we can monitor performance, diagnose issues, and improve usability. Data collected through these cookies is aggregated and used for analytical purposes.
| Cookie Name | Purpose | Expiration |
| -------------------- | ----------------------------------------------- | ---------- |
| `ia_stats_session` | Tracks session-level usage for analytics | Session |
| `ia_stats_pageviews` | Counts page or feature interactions | 24 hours |
| `ia_stats_events` | Records anonymized interaction events | 30 days |
| `ia_stats_perf` | Monitors performance metrics such as load times | 7 days |
| `ia_stats_returning` | Identifies returning users for usage trends | 90 days |
#### 4.4 Marketing Cookies (Optional)
These cookies are used to measure the effectiveness of marketing efforts and understand how users arrive at and engage with the platform.
| Cookie Name | Purpose | Expiration |
| -------------------- | ------------------------------------------------------------- | ---------- |
| `ia_mkt_source` | Stores visit source or referral information | 30 days |
| `ia_mkt_campaign` | Tracks marketing campaign attribution | 30 days |
| `ia_mkt_conversion` | Measures whether marketing efforts led to sign-ups or actions | 90 days |
| `ia_mkt_retention` | Analyzes engagement following acquisition | 90 days |
| `ia_mkt_experiments` | Supports controlled marketing-related experiments | 30 days |
### 5. Cookie Security & Configuration
We take appropriate measures to protect cookies and the information stored within them:
* Most cookies are scoped to the application domain
* Secure and SameSite attributes are applied where appropriate
* Cookies are retained only for as long as necessary for their intended purpose
### 6. Updates to This Cookie Policy
We may update this Cookie Policy from time to time to reflect changes in technology, legal requirements, or our practices. Any updates will be posted on this page with an updated "Last updated" date.
### 7. Contact Us
If you have any questions about this Cookie Policy or how we use cookies, please contact us at: [support@imagine.art](mailto:support@imagine.art)
# Privacy Policy
Source: https://docs.imagine.art/policies/privacy-policy
**Last updated:** November 4th, 2025
Imagine.art ("Imagine Art," "we," or "us") is a creative technology company offering AI-powered image and video generation services. This Privacy Policy describes how we collect, use, and disclose personal information in connection with your use of our Service. By accessing or using Imagine Art (including signing up for any plan at Pricing), you agree to the terms of this Policy. If you do not agree with any part of this Policy, **please do not use our Service**.
## Scope and Acceptance
This Policy applies to all users of Imagine Art worldwide, including users in the U.S., EU/EEA, UK, California, and other jurisdictions. It covers information collected through our website, apps, APIs, and any other Imagine Art services (the "Service"). By creating an account or otherwise using the Service, you accept the collection, use, and disclosure of your information as described here. Your continued use of the Service after changes to this Policy are posted will constitute acceptance of those changes.
## Information We Collect
We collect several categories of information from and about you:
* **Contact and Account Information:** When you register for an account or contact support, we collect your name, email address, billing address, and other contact details. We also record your chosen username and details of your subscription plans and usage.
* **Payment Information:** We use **Stripe** as our payment processor. When you subscribe or make a purchase, Stripe collects your credit card or financial data. We do **not** store full credit card details on our servers; we only retain a transaction confirmation from Stripe.
* **User-Provided Content:** We collect any images, videos, text prompts, or other content you upload or generate through the Service. This includes any inputs (such as text or reference images) and the outputs (generated images/videos) you create.
* **Usage and Technical Data:** We automatically collect log and usage data about your interaction with the Service, including your IP address, device and browser information, page views, feature usage, timestamps, and error reports.
* **Tracking and Analytics Data:** We use cookies and similar technologies to track your activity. This may include session cookies, device identifiers, and analytics data (e.g. pages visited, features used).
* **Communications:** If you correspond with us (via email, support tickets, surveys, etc.), we collect the information you provide (e.g. message content, email address).
* **Third-Party Sources:** We may also receive information about you from third parties. This can include profile information or social login data you choose to share.
## How We Use Your Information
We use your information for the following purposes:
* **Operate and Improve the Service:** We use your personal data to provide, maintain, and enhance the Imagine Art platform. This includes processing your image/video generation requests, delivering the features you use, monitoring usage and performance, and fixing errors.
* **Account Management:** We use your data to create and manage your account and subscription. For example, we authenticate you, manage your plan and billing, and communicate with you about your account (including security notices).
* **Payment Processing:** Your payment and billing information is used to process transactions for subscriptions, credits, or other purchases you make. We confirm payments and may store transaction receipts, but actual credit card data is handled by Stripe.
* **Communications:** We may contact you by email or other means to provide updates, notices, technical support, and security alerts related to the Service. If you opt in, we may send you newsletters, promotions, or special offers about our services. You can opt out of promotional emails at any time by following the "unsubscribe" link in those messages.
* **Marketing and Personalisation:** With your consent where required, we use your information to personalise our communications and to improve the Service. For example, we may analyse usage data to tailor content, features, and advertisements. We do not sell your personal data to third parties for their marketing; any ad targeting by partners is based on anonymised or aggregated data.
* **Analytics:** We analyse usage patterns to better understand how the Service is used and to improve it.
* **Legal and Security:** We use your information to comply with legal obligations, to protect our rights and the rights of other users, and to detect or prevent fraud and abuse. This includes reviewing user-generated content for compliance with our policies.
* **Other Purposes:** We may use your data for other legitimate business purposes, such as debugging, data analysis, user support, and enforcing our Terms and policies.
## Legal Basis for Processing
If you are in the European Union, European Economic Area, or the UK, we process your personal data under one or more of the following legal bases:
* **Contract Performance:** Processing necessary to fulfil our contract with you, e.g. providing the Service you have subscribed to.
* **Consent:** Processing based on your consent (for example, marketing communications or any additional uses where we request your permission).
* **Legitimate Interests:** Processing for our legitimate interests (such as improving our Service, ensuring security, or managing our business) provided those interests do not override your privacy rights.
* **Legal Obligations:** Processing necessary to comply with laws or regulations (for example, tax and financial record-keeping).
You may withdraw consent at any time (e.g., unsubscribe from marketing emails) without affecting prior processing. For EU users, you also have rights described below under "User Rights and Controls."
## Sharing and Disclosure of Information
Imagine Art does not rent or sell your personal information to third parties for their marketing. We disclose your information only as described below:
* **Service Providers and Third Parties:** We share information with trusted third-party service providers who perform services on our behalf. This includes cloud hosting customer support platforms, email delivery services, and others. These parties are contractually obligated to protect your data and use it only to provide their service to us.
* **Third-Party AI Models:** The AI content generation is powered by third-party models (e.g. Stability AI, OpenAI). When you submit prompts or inputs, we send that data to these model providers to generate outputs. **Important:** because these are third-party services, do not include any personal or sensitive information in your prompts or uploads.
* **Public Areas and Community:** If you choose to publish or share your generated content in a public gallery or community forum, that content will be visible to other users. In that case, Imagine Art may allow others to view, remix, or use your shared content as permitted by the community features. We clearly mark what content is public; do not post anything you wish to keep private.
* **Advertising/Analytics Partners:** With your consent, we may share anonymized or aggregate usage data (via cookies or similar technologies) with analytics and advertising partners to deliver relevant ads and understand trends.
* **Legal Compliance and Safety:** We may disclose your information if required by law or in response to valid legal process (subpoena, court order, government request). We may also disclose information to protect our rights, to comply with a legal obligation, to enforce our policies, or to protect the safety of our users or the public.
## Data Retention
We retain personal information only as long as necessary to fulfil the purposes in this Policy. For example, we keep account and billing records for as long as your account is active and for legal purposes (such as tax requirements). Log and usage data are generally kept for a shorter period, except when needed for security or to comply with a legal obligation. When you delete your account or request deletion, we will erase your personal data promptly unless retention is required by law or for legitimate business purposes (e.g. fraud prevention).
## User Rights and Controls
Imagine Art respects your data protection rights. Subject to applicable law, you have the following rights regarding your personal data:
* **Access and Portability:** You can request a copy of the personal information we hold about you. We can provide it in a portable format.
* **Rectification:** You may correct or update inaccurate or incomplete information in your account.
* **Erasure ("Right to be Forgotten"):** You can request deletion of your personal data. We will comply unless we need to retain it for legal or legitimate purposes.
* **Restriction/Objection:** You may ask us to restrict or stop processing your personal data in certain circumstances (for example, if you contest accuracy or object to direct marketing).
To exercise any of these rights, please contact us at [privacy@imagine.art](mailto:privacy@imagine.art). We will verify your identity and respond within the timeframe required by law.
## Cookies and Similar Technologies
We and our partners use cookies, web beacons, and other tracking technologies to collect information about your use of the Service. These technologies help us remember your settings, enable core functionality, analyse traffic, and personalize your experience. You can control or disable cookies through your browser settings, but disabling cookies may prevent parts of the Service from working correctly.
## Security of Your Personal Data
At ImagineArt, we take the protection of your Personal Data seriously. We implement industry-standard security measures to safeguard your information from unauthorized access, disclosure, alteration, or destruction. Your Personal Data is stored and processed only within the jurisdictions of certain countries where we operate and where data protection laws provide adequate safeguards.
However, please note that no method of transmission over the Internet or method of electronic storage can be guaranteed to be completely secure. While we make every reasonable effort to protect your Personal Data, we cannot ensure or warrant its absolute security.
## Data Access, Portability, and Deletion
You can view, export, or delete much of your information via your account dashboard. To download your personal data or generated content, follow the instructions in your account settings. To delete your account and all associated data, please use the account deletion process in your dashboard or contact support. We will honor deletion requests promptly, except for data we must keep for legal or business reasons (e.g. we may retain anonymized logs or financial transaction records as required).
## Children's Privacy
Our Service is **not directed to children under 13**. We do not knowingly collect personal information from anyone under 13 years of age. If we become aware that a child under 13 has provided us with personal data, we will promptly delete that data. If you are a parent or guardian and believe your child under 13 has used the Service or submitted information, please contact us to have their data removed.
## Changes to This Privacy Policy
We may update this Privacy Policy from time to time to reflect changes in our practices, technology, or legal requirements. When we make significant changes, we will post the new version on our website with a revised "Last updated" date and, if required, notify you (for example, via email or in-product notification). Your continued use of the Service after any update indicates your acceptance of the revised Policy. We encourage you to review this Policy periodically.
## Contact Information
If you have any questions, concerns, or requests regarding this Privacy Policy or our privacy practices, please contact us:
* **Email:** [privacy@imagine.art](mailto:privacy@imagine.art)
* **Company Name:** Vyro Turkey Teknoloji Limited Sirketi
* **Company Address:** Idealtepe Mh. Dik Sk. No. 13 K2 Maltepe, Istanbul, Turkey
We will try to address and resolve your inquiries in a timely manner.
# Refund Policy
Source: https://docs.imagine.art/policies/refund-policy
### 1. General Policy Statement
This Refund Policy ("Policy") governs the terms under which refunds may be issued by Imagine Art ("Company", "we", "us", "our") to its users ("Customer", "User", "you") for services delivered through its software-as-a-service platform. This Policy is incorporated by reference into Imagine Art's Terms of Service and applies to all subscriptions, credit-based purchases, and digital goods offered by the Company.
### 2. Refund Eligibility
#### 2.1 Eligible refund cases
The Company offers refunds only under the following specific circumstances, subject to verification and compliance with the conditions set forth herein:
1. **Monthly and quarterly subscription plans**
* Customer may request a refund within **Three (3) calendar days** from the date of initial purchase of a monthly subscription.
* Refund eligibility is contingent upon Customer having used no more than **Three hundred (300) credits** during such period.
2. **Annual subscription plans**
* Customer may request a refund within **Five (5) calendar days** from the date of initial purchase of an annual subscription.
* Refund eligibility is contingent upon Customer having used no more than **Three hundred (300) credits** during such period.
3. **Verified technical malfunctions**
* If the system experiences a verified failure that prevents the Services from functioning as intended, due to technical issues fully attributable to the Company's platform, and such failure persists for more than 24 consecutive hours without resolution, refunds or credit restorations may be granted at the Company's discretion.
4. **Duplicate charges**
* Unintentional duplicate billing or erroneous multiple charges will be fully refunded upon verification.
5. **Unauthorized transactions**
* Verified fraudulent use or **Unauthorized** charges will be refunded in full upon investigation.
6. **Mandatory compliance with applicable law**
* Where required by applicable consumer protection statutes, including but not limited to the California Automatic Renewal Law, Federal Trade Commission Act Section 5, EU Consumer Rights Directive, refunds shall be issued in accordance with statutory obligations.
7. **Refund eligibility for recent payments (subscription updates)**
* Refund requests for subscription updates, meaning any payment occurring after the initial subscription charge, will only be accepted within **three (3) calendar days** from the date of the charge.
* In such cases, a **fifty percent (50%) deduction** will apply to the refundable amount.
8. **Promotional offers**
* Refunds for promotional offers will be assessed based on credit consumption of the model or tool involved.
* For models or tools with unlimited generations, refunds will be calculated accordingly, considering usage during the promotional period.
### 3. Non-Refundable Circumstances
#### 3.1 Refunds are not provided under the following conditions
1. Dissatisfaction with style, creative output, aesthetic choices, or subjective preferences once Services have been delivered.
2. User error, including but not limited to incorrect prompts, improper configuration, or misunderstanding of platform features.
3. Failure to cancel recurring subscriptions prior to the renewal billing date.
4. Substantial consumption of services beyond the defined three hundred (300) token usage threshold.
5. Abuse of promotional offers, free trials, or repeated refund requests deemed excessive, unreasonable, or fraudulent.
6. Accidental or mistaken transactions initiated by the Customer.
7. Changes in the appearance, theme, or other non-functional modifications to the platform.
8. A change of mind after purchasing the product.
9. Charges automatically applied upon expiration of a free trial period that the Customer initiated, where the Customer did not cancel prior to the trial's conclusion.
### 4. Administrative Charges
#### 4.1 Approved refund deductions
In the event a refund request is approved, the Company reserves the right to deduct a reasonable administrative fee from the refunded amount. The specific deduction shall be disclosed at the time of refund approval and may vary based on payment processing costs and internal handling.
### 5. Subscription Cancellation
#### 5.1 Auto-renewal
All subscription plans are subject to automatic renewal unless actively canceled by the Customer prior to the renewal date.
#### 5.2 How to cancel
Customers may cancel future renewals at any time via email to [support@imagine.art](mailto:support@imagine.art) or by following the steps in [How to Cancel Your Subscription](https://help.imagine.art/frequently-asked-questions/how-to-cancel-your-subscription). Cancellations shall apply prospectively to future billing periods. Retroactive refunds for unused portions of active billing cycles shall not be granted, unless otherwise provided herein.
### 6. Refund Request Procedure
#### 6.1 Submission channel
All refund requests must be submitted through the following official channel:
* By email to [support@imagine.art](mailto:support@imagine.art)
#### 6.2 Review timeline
The Company shall review and respond to refund requests within a commercially reasonable timeframe, generally not to exceed ten (10) business days.
#### 6.3 Refund method
Approved refunds shall be processed to the original form of payment unless otherwise agreed. Refunds may, at Company's sole discretion and subject to customer consent where applicable, be issued in the form of account credit.
### 7. Disputes and Chargebacks
#### 7.1 Pre-chargeback resolution
Customers are expected to engage with Company's support team to resolve any billing disputes prior to initiating any chargebacks or payment reversals.
#### 7.2 Fraudulent chargebacks
Illegitimate or fraudulent chargebacks may result in permanent suspension of customer accounts and recovery of associated costs.
### 8. Policy Modifications
#### 8.1 Amendments
Company reserves the right to amend this Policy at any time. Material modifications shall be communicated to active subscribers via email or in-app notification.
#### 8.2 Continued use
Continued use of Services following policy amendments shall constitute acceptance of revised terms.
# Terms and conditions
Source: https://docs.imagine.art/policies/terms-and-conditions
### 1. Acceptance of Terms
Welcome to **ImagineArt**, an AI-powered platform offering image generation, video creation, audio tools, model training, and API access (collectively, the "Service"). These Terms and Conditions ("Terms") govern your use of the Service and form a binding legal agreement between you and ImagineArt (referred to as "ImagineArt," "we," "us," or "our"). By creating an account or otherwise accessing or using the Service, you acknowledge that you have read, understood, and agree to be bound by these Terms. If you do not agree, you must not use the Service.
If you are using the Service on behalf of an organization or other entity, you represent that you have the authority to bind that entity to these Terms, and "you" as used in these Terms includes both you as an individual and that entity. You are responsible for ensuring that all persons who access the Service through your account are aware of and comply with these Terms.
Additional guidelines, policies, or terms may apply to certain features of the Service (e.g., API use, community forums, or specific tools). Such **Supplemental Terms** will be presented to you for acceptance when you use those features, or are otherwise referenced in these Terms. Supplemental Terms are incorporated by reference and form part of this Agreement. If there is any conflict between these Terms and Supplemental Terms, the Supplemental Terms will govern with respect to the applicable feature or Service.
### 2. Eligibility and Account Registration
You must be at least 13 years old (or the minimum digital age of consent in your country, if higher) to use ImagineArt. If you are under 18 (or under the age of majority in your jurisdiction), you may only use the Service under the supervision of a parent or legal guardian who agrees to be bound by these Terms. By using the Service, you affirm that you are old enough to enter into this agreement or have obtained proper consent from a parent/guardian. While Imagine Art strives to generate content that aligns with user expectations and is appropriate for all audiences, the assets are created by an artificial intelligence system based on user inputs. As a result, we cannot guarantee that the generated content will always meet specific suitability or appropriateness standards for every user.
To access most features, you must create an ImagineArt account. You agree to provide accurate, current, and complete information during registration and to keep it updated. You are responsible for maintaining the confidentiality of your account credentials and for all activities that occur under your account. You must promptly notify us of any unauthorized access to or use of your account. We recommend using strong, unique passwords and enabling available security features. We are not liable for any loss or damage arising from your failure to secure your account.
Each user may register and use only one account, and you may not share your account with others. You must not misrepresent your identity or affiliation with any person or entity. ImagineArt reserves the right to suspend or terminate any account that it suspects is being shared or used in violation of these Terms.
If you create an account on behalf of a company or other entity, you represent that you have the authority to do so and to bind the entity to these Terms. In such case, the term "you" will refer to both you and the entity. The entity shall be fully responsible for the account and compliance with these Terms.
### 3. Subscription Plans, Credits, and Fees
ImagineArt offers both free access (with limited features or usage quotas) and paid subscription plans that provide enhanced features, higher usage limits, or access to premium tools. Some aspects of the Service may use a **credit system** for consumption of resources (e.g., generating images or videos may cost tokens/credits). The specific details of available plans, including pricing, features, and credit allotments, are described on our Pricing page or in the Service interface. By selecting a subscription tier or purchasing credits, you agree to pay the applicable fees and abide by any additional terms for that plan.
You must provide a valid payment method when signing up for a paid plan. Subscription fees are billed in advance on a recurring basis (e.g., monthly, quarterly or annually) according to the plan you select, and will auto-renew at the end of each billing cycle unless canceled. By subscribing, you authorize ImagineArt (or its payment processor) to charge your provided payment method automatically **each renewal period** for the subscription fee, until you cancel. We will charge applicable taxes as required by law. No contract for services is formed until we confirm your subscription (for example, by email or by providing access to the paid features).
After your initial term, your subscription will automatically renew for successive periods of the same length at the then-current price for your plan. We may adjust the pricing and/or features of subscription plans and will provide advance notice of any material changes (for instance, by email or via the Service). If you do not agree to a change in fees or terms for a renewed term, you must cancel your subscription before the next billing cycle. Continued use of the Service after price or feature changes take effect constitutes your acceptance of the new terms. Promotional or discounted subscription offers, if any, are subject to the terms of the offer and may be available only for first-time or eligible users; after the promotional period ends, regular rates will apply.
If your plan includes usage credits , these credits may be deducted as you use certain tools or API calls. Credits may have an expiration date (as specified in the plan or at purchase) and are only usable for the ImagineArt Service. They are not refundable or redeemable for cash and may not be transferred to other users (except as explicitly allowed by a specific business or team plan). If your account is terminated or closed, unused credits are forfeited and will not be refunded, except where required by law or at our sole discretion.
We may offer free trials or free tiers for new users. Free trials are for a limited period as specified, and are intended to allow new users to try the Service. We may require you to provide a payment method to start a free trial; however, you will not be charged until the trial period ends. If you do not cancel before the trial ends, your trial may convert to a paid subscription and your provided payment method will be charged the applicable fees on the first day after the trial period. You can cancel a trial at any time before it ends to avoid incurring charges. We reserve the right to modify or discontinue free trial offers at any time.
You may cancel your subscription at any time by visiting your account settings (e.g., the "Plans & Billing" page) or through the platform from which you purchased the subscription. If you cancel, you will continue to have access to your paid features until the end of the current billing period, but your subscription will not renew thereafter. We do not provide prorated refunds for unused time in a billing cycle – the cancellation will take effect at the next renewal and you will not be charged going forward. If you downgrade your plan, the change will occur at the start of the next billing period; you will retain access to the higher-tier features until that time.
Subscription fees (and purchased credit packs) are generally **non-refundable** except as required by law or explicitly stated otherwise. ([Refund Policy](https://help.imagine.art/terms-and-policies/refund-policy))
If we cannot charge your provided payment method for any reason (e.g., card expiration or insufficient funds), we will attempt to notify you and may retry billing. If payment remains outstanding, ImagineArt may suspend or downgrade your account, or terminate your subscription, at our discretion. You remain responsible for any unpaid amounts, and you agree to reimburse us for any collection costs incurred.
### 4. Service Availability and Changes
ImagineArt is continually evolving. We are constantly improving and updating the Service's features and capabilities. You acknowledge that the Service (including any content, algorithms, models, or user interfaces) may change over time as we refine our products. We reserve the right to add, modify, or remove features or functionality of the Service at any time, with or without notice, provided that if you are a paid subscriber and a change materially reduces core functionality of your plan, we will endeavor to provide you notice or an adjustment.
From time to time, we may offer early access to new or experimental features (often labeled as "Beta" or preview). Such features are provided on an as-available basis for testing and feedback, without any warranties, and may be modified or discontinued at our discretion. Beta features may have separate or additional terms.
We strive to maintain the Service's availability but do not guarantee uninterrupted, error-free operation or any specific uptime. The Service may occasionally be unavailable for maintenance, updates, or network issues. ImagineArt is provided "as is" and "as available" without any warranty of quality, reliability, or availability. You agree that you will not rely on the continuous availability of the Service for any critical uses. We are not liable to you for any harm or losses arising from Service outages, disruptions, or changes to any features.
We may provide user support or documentation to help you use the Service. However, unless you have a separate support agreement, we do not guarantee any specific response times or resolutions for support inquiries.
### 5. Acceptable Use and Conduct
By using ImagineArt, you agree to use the Service only for lawful purposes and in compliance with these Terms and all applicable laws and regulations. You are responsible for your conduct and any data, text, images, audio, video, or other content that you input into or generate via the Service (collectively, "Content"). **Without limiting the generality of the foregoing, you agree not to:**
* **Illegal Activities:** Use the Service for any unlawful, fraudulent, or malicious activities, or in any way that violates any applicable law or regulation. You may not upload or submit any content that is illegal in the jurisdiction in which you reside or in which we operate, nor use the Service in furtherance of any illegal purpose.
* **Infringement of Rights:** Upload, submit, or generate any content that infringes or violates the intellectual property rights or other rights of any person or entity. This includes, for example, prompts or inputs that attempt to recreate copyrighted images or audio without authorization, as well as any output content that you use in a manner that violates someone else's copyright, trademark, patent, trade secret, right of publicity, or privacy rights. Do not attempt to use the Service to violate the IP rights of others.
* **Prohibited Content:** Use the Service to create or share content that is obscene, pornographic, sexually explicit, excessively violent or graphic (gore), or hateful/harassing toward any individual or group. Content that promotes self-harm, suicide, terrorism, or other extreme violence is strictly forbidden. We also do not allow content that is defamatory, libelous, threatening, or that advocates bigotry or discrimination against protected classes. ImagineArt aims to maintain a PG-13, family-friendly environment, and we reserve the right to block or filter prompts that produce disallowed content.
* **Personal Data and Privacy:** Do not input personal data about others without their consent, especially sensitive personal information. You must not upload images, audio, or any media depicting another person (especially minors) without proper rights or permission to do so. **No Doxing:** You may not use the Service to disclose another person's personal identifying information ("doxing") or invade another's privacy.
* **Deception and Misinformation:** You may not use the Service to create content with the intent to deceive, defraud, or mislead anyone. This includes generating content for phishing, impersonation of others (without parody/satire protections), deepfakes of private individuals or public figures for malicious purposes, or any fraudulent activity. You agree not to misrepresent the nature or origin of generated content. For example, you should not present AI-generated outputs as human-created when it is likely to cause confusion or harm.
* **Political and Election Use:** You may not use the Service to generate content for political campaigning or to attempt to influence the outcome of an election. Political figures and scenarios can be sensitive; any use of the Service in a political context must comply with all laws and platform guidelines and not be misleading or manipulative.
* **Spam and Advertising:** Do not use the Service to transmit any unsolicited or unauthorized advertising or promotional materials, junk mail, spam, pyramid schemes, or any other form of solicitation. The Service should not be used to harvest contact information or send mass communications without consent.
* **Interference and Misuse:** You will not interfere with or disrupt the integrity or performance of the Service or the data contained therein. This means you must not attempt to hack, DDOS, overload ("flood"), or disrupt the Service for others. Avoid any activity that could harm or place an unreasonable burden on our infrastructure. You are prohibited from using any automated means (such as scripts, bots, or scrapers) to access or use the Service, except as allowed through our official API and tools.
* **Reverse Engineering and Competitive Use:** You may not reverse engineer, decompile, or attempt to extract the source code or underlying models of any part of the Service. Also, you must not use the Service to develop, train, or improve any competing product or service. Access to the Service is provided for your direct use only – you may not resell, redistribute, or provide third parties with unauthorized access to our Service or content.
* **Content Sharing:** If the Service allows you to share or publish content (e.g., in a community gallery or forum), you agree to follow any additional Community Guidelines provided. Do not post others' content created on ImagineArt without their permission, and do not falsely claim authorship of output that is not yours. Respect other users and do not harass or abuse others in any community interaction.
We reserve the right (but do not assume the obligation) to monitor use of the Service and content to ensure compliance with these Terms and applicable law. ImagineArt, in its sole discretion, may refuse to process any input or may block or remove any content (including generated outputs) that we believe violates these Terms or our policies, or that is otherwise objectionable. We may use a combination of automated filters and human review to achieve this. You understand that using our AI generation tools may occasionally produce unintended or inappropriate results, and you agree to use your best judgment and to notify us of any serious issues.
If you become aware of any misuse of the Service or any content that you believe violates these Terms, please report it to us through the designated channels (by contacting support). All reports should be made in good faith and include as much detail as possible. We will review reports and take actions we deem appropriate, which may include content removal, user warnings, or account suspension.
Violation of this Acceptable Use Policy may result in suspension or termination of your account, removal of content, and/or other legal action where appropriate. **We have a zero-tolerance policy for egregious abuses.** ImagineArt reserves the right to immediately suspend or ban your access to the Service at any time, **with or without notice**, for any conduct that we determine, in our sole discretion, is objectionable or in violation of these Terms, or which exposes us or others to risk of harm or liability. Repeated violators or those committing severe offenses may be permanently banned from the Service. We may also cooperate with law enforcement and report unlawful conduct.
### 6. User Content and Intellectual Property Rights
#### 6.1 Your Inputs and Outputs
**User-Provided Content (Inputs):** In the course of using ImagineArt, you may provide text prompts, images, audio, video, or other materials as input ("Input") to the AI models. You might also upload content such as profile information, comments, or other media. You affirm that you either own all rights to any Input you provide, or that you have obtained all necessary permissions, licenses, and consents from any applicable rights holders (such as photographers, authors, performers, or data subjects) to use that Input with the Service. **Responsibility:** You remain solely responsible for any Input you provide and for any Output generated from your Input. This means you must ensure your Inputs (and the resulting Outputs) do not violate any laws or anyone's rights, and are not subject to any third-party confidentiality or contractual obligations. If you upload or provide content that includes third-party intellectual property (for example, uploading an image or audio clip you did not create), you must have the legal right to do so. By providing any Input, you represent and warrant that doing so, and the resulting generation, will not infringe or misappropriate the rights of any third party, and that you have complied with all relevant license requirements.
You acknowledge that ImagineArt owns all rights, title, and interest in the Services, including the platform, underlying technologies, and all content generated through the platform, except for the Outputs for which we grant you a license.
**Personal vs. Commercial Use:** If you are on a free or trial plan, your use of Outputs may be limited to non-commercial, personal purposes unless otherwise explicitly permitted. Paid subscribers (or those with a proper commercial license or subscription) are allowed to use the Outputs for commercial purposes, such as in business projects, products, or content monetization. Regardless of free or paid status, all use of Outputs must adhere to the Acceptable Use rules in Section 5 and any applicable law. (For example, even a paid user cannot use an Output in a way that infringes someone's copyright or violates content restrictions.) We do not restrict your right to monetize or otherwise exploit your own Outputs, so long as you comply with these Terms.
You are **solely responsible** for how you use and distribute Outputs. If an Output inadvertently resembles or incorporates someone else's content, you may need permission from that party to use it. We encourage you to review Outputs, especially for commercial projects, to ensure they can be safely used. ImagineArt will not be liable for claims arising from your use of the Outputs, and you agree to use them at your own risk.
#### 6.2 License to ImagineArt (for Operation and Improvement of the Service)
**Service Operation:** When you use the Service, the processing of your Inputs and generation of Outputs necessarily involves copies of and modifications to your content (e.g., our systems may multiply, transform, or analyze your Input to generate results). Subject to any applicable account settings you select, you grant ImagineArt a fully paid, royalty-free, perpetual, irrevocable, worldwide, non-exclusive, and fully sub licensable license (including any moral rights) to host, use, license, distribute, reproduce, modify, adapt, publicly perform, and publicly display, in whole or in part, Your Content for the purpose of operating and providing the Services to you and other registered users. This license includes the right to use Your Content to improve, promote, and enhance ImagineArt's platform and Services.
**Optional AI Training Usage:** ImagineArt gives you control over whether your Inputs and Outputs may be used to further train or improve our AI models. By default, **ImagineArt will not use your personal content for AI model training without your consent**.
**Public Content Sharing:** If you explicitly publish or share your content in a public area of the Service (for example, by posting in a public gallery or community forum within ImagineArt), you **allow other users to view and use your shared content**. In such cases, you also grant ImagineArt a license to make that content available to others as part of the community. For instance, if a community gallery allows remixing of images, you permit other ImagineArt users to remix or build upon the images you post there. We will clearly indicate which areas are public. If you prefer to keep your generations private, use the appropriate settings or refrain from posting them publicly. ImagineArt is not responsible for what other users do with content you choose to make public.
**Feedback:** If you provide feedback, suggestions, or ideas regarding ImagineArt ("Feedback"), you agree that we are free to use or not use such Feedback in any manner, and you grant us a perpetual, sublicensable, worldwide, royalty-free license to incorporate and use your Feedback for any purpose, without any obligation to compensate you. Providing Feedback is entirely voluntary and will not create any confidentiality obligation.
#### 6.3 ImagineArt's Intellectual Property
**Our Property:** Except for your own content, all right, title, and interest in and to the Service and its components are and will remain the exclusive property of ImagineArt and/or our licensors. This includes, but is not limited to, all software, algorithms and AI models, code, design, user interfaces, trademarks ("ImagineArt" and associated logos), content provided by us (such as sample images or text), and the compilation of all materials on the Service. The Service is protected by copyright, trade secret, trademark, and other intellectual property laws.
**License to Use the Service:** Subject to your compliance with these Terms and payment of any applicable fees, we grant you a limited, non-exclusive, non-transferable, non-sublicensable, revocable license to access and use the Service for your personal or internal business purposes. You may not use ImagineArt's name, logos, or trademarks without our prior written consent, except as necessary to attribute the source of generated content or as permitted under fair use or other applicable laws.
**Restrictions:** You agree not to copy, distribute, modify, create derivative works based on, publicly perform, publicly display, or exploit any portion of our Service except as expressly allowed by these Terms. You shall not remove or obscure any copyright, trademark, or other proprietary rights notices on the Service or on any content delivered to you. All rights not expressly granted to you are reserved by ImagineArt. We welcome you to use ImagineArt's tools, but **please respect our intellectual property** in how they are implemented.
**Third-Party Software and Models:** If the Service includes any third-party software or open source components, such components may be subject to separate license terms, which will be provided to you. We also use or provide access to certain third-party AI models or services; your use of those may be subject to additional terms or license requirements, which will be communicated when applicable.
### 7. Privacy and Data Protection
Your privacy is important to us. Please review our [Privacy Policy](https://help.imagine.art/terms-and-policies/privacy-policy) to understand how we collect, use, disclose, and safeguard your personal information. The Privacy Policy is incorporated into these Terms by reference. By using the Service, you consent to the collection and use of information as described in the Privacy Policy. In summary, and not to limit the details in that policy, note the following:
* **Personal Data:** We will process certain personal data from you (such as account information, usage data, and any personal data contained in your Inputs or Outputs) in order to provide and improve the Service, to handle billing, and for other legitimate business purposes (like preventing abuse and complying with law).
* **GDPR and EU Users:** If you are in the European Economic Area (EEA), United Kingdom, or a similar jurisdiction with comprehensive data protection laws, you have specific rights regarding your personal data (such as rights to access, correction, deletion, and objection). ImagineArt is committed to compliance with the General Data Protection Regulation (GDPR) and relevant laws. We will only process your personal data on a valid legal basis, such as your consent or our legitimate interests, and we have implemented measures to safeguard your data. You can learn more in our Privacy Policy and can contact us with any privacy-related questions or requests.
* **Data Security:** We employ administrative, technical, and physical security measures to protect your information. However, no system is perfectly secure; you acknowledge that you provide information at your own risk. Notify us immediately if you discover any security vulnerabilities or suspect any unauthorized access to your account.
* **Data Retention:** We retain personal data and content for as long as necessary to fulfill the purposes for which it was collected, or as required by law. If you delete your account or certain content, we will delete or de-identify the related data, except to the extent we are permitted or required to retain it (for example, for legal compliance, dispute resolution, or backups).
If there is a conflict between these Terms and the Privacy Policy with respect to personal data practices, the Privacy Policy will govern. Remember that any content you voluntarily make public (such as in a community forum) is **not private** and may be seen or used by others; only share what you are comfortable being public.
### 8. Copyright and Intellectual Property Complaints
ImagineArt respects the intellectual property rights of others and expects our users to do the same. If you believe that any content on the Service infringes your copyright (or other intellectual property rights), please notify us promptly in writing. We have implemented procedures for receiving and addressing such complaints in accordance with the Digital Millennium Copyright Act (DMCA) and other applicable laws.
**DMCA Takedown Notices:** A notification of claimed copyright infringement must be sent to our designated Support Email at:
```text theme={null}
Attn: Support – ImagineArt
[Mailing Address]
Email: support@imagine.art
```
Your notice must include the following information (please see 17 U.S.C. §512(c)(3) for further detail):
1. **Identification of the copyrighted work** you believe has been infringed (or a representative list of such works if the notice covers multiple).
2. **Identification of the material that is claimed to be infringing** and information reasonably sufficient to permit us to locate the material on the Service (e.g., a URL or screenshot).
3. **Your contact information** – including your name, mailing address, telephone number, and email – so we can reach you.
4. A statement that you have a **good-faith belief** that the use of the material in the manner complained of is not authorized by the copyright owner, its agent, or the law.
5. A statement that the information in your notice is **accurate**, and under penalty of perjury, that you are the **owner of the copyright** (or an agent authorized to act on the owner's behalf).
6. Your physical or electronic **signature** (typing your full name at the end of your notice will suffice as an electronic signature).
If your notice does not substantially comply with these requirements, it may not be processed. Upon receipt of a valid notice, we will expeditiously investigate and, if appropriate, remove or disable access to the allegedly infringing material. We may notify the user who posted the content (if applicable) and provide them an opportunity to submit a counter-notification.
**Counter-Notifications:** If you believe your content was removed or disabled by mistake or misidentification, you may send us a counter-notice. A valid counter-notification should include: (a) identification of the material that was removed and where it appeared, (b) a statement under penalty of perjury that you have a good-faith belief the material was removed due to mistake or misidentification, (c) your contact information.
**Repeat Infringers:** In accordance with the DMCA and other applicable law, ImagineArt has a policy of terminating, in appropriate circumstances, users who are deemed to be repeat infringers of others' copyrights or other intellectual property rights. If a user repeatedly uploads or posts infringing material, we may suspend or delete their account at our discretion. We also reserve the right to limit access to the Service or take other appropriate action against any user who infringes intellectual property rights, even for a first offense.
**Trademark Complaints:** If you believe someone's use of a trademark on our Service is infringing or violates your rights, you may report it to [support@imagine.art](mailto:support@imagine.art). Please provide details of the trademark, the offending use, and proof of your rights in the mark. We will review and take appropriate action under applicable trademark laws.
### 9. Disclaimers of Warranties
**As-Is Service:** **ImagineArt and all of its services, information, and content are provided on an "AS IS" and "AS AVAILABLE" basis**. To the maximum extent permitted by law, ImagineArt disclaims all warranties, representations, and conditions of any kind, whether express, implied, or statutory, including but not limited to any implied warranties of merchantability, fitness for a particular purpose, title, non-infringement, and any warranties that may arise from course of dealing or course of performance. We do not guarantee that the Service will meet your requirements or expectations, that it will be available on an uninterrupted, secure, or error-free basis, or that the results obtained from its use (including any Outputs) will be accurate, reliable, or suitable for your purposes.
**Outputs and Content:** You understand that using an AI generative service involves probabilistic outputs that may not always be appropriate or correct. **Use at Your Own Risk:** Any content generated, downloaded, or otherwise obtained through ImagineArt is used at your own discretion and risk. You are solely responsible for any harm to your computer system, loss of data, or other damage that results from your use of the Service or any content (Inputs or Outputs). ImagineArt makes no warranty that the AI-generated content will be free of objectionable material or errors. We also do not warrant that our filters or safeguards will catch all prohibited content or perfectly meet your sensitivity expectations.
**Third-Party Content & Services:** ImagineArt may allow you to access or use third-party services or content through the platform. We do not control or endorse third-party content and disclaim any liability for such content. Any third-party services or content are subject to the terms and privacy policies of those third parties, not these Terms.
**No Guarantee of Data Integrity:** While we try to preserve your content, we do not guarantee that your Inputs or Outputs will be stored or retrievable forever. You are encouraged to keep backups of important Outputs on your own. We are not responsible for any loss or corruption of data or content.
**No Advice or Information:** No advice or information, whether oral or written, obtained from ImagineArt or through the Service shall create any warranty not expressly stated in these Terms.
**Exceptions:** In certain jurisdictions, the law may not permit the disclaimer of certain warranties, so some of the above disclaimers may not apply to you. In such case, ImagineArt's warranties shall be limited to the minimum extent required by law.
### 10. Limitation of Liability
**Indirect Damages:** To the fullest extent permitted by law, in no event will ImagineArt or its affiliates, officers, employees, agents, partners, or licensors be liable to you for any indirect, incidental, special, consequential, or exemplary damages, or for any loss of profits, revenues, goodwill, data, or other intangible losses, arising out of or related to your access to or use of (or inability to use) the Service or any content, whether based on warranty, contract, tort (including negligence), statute, or any other legal theory, and even if ImagineArt has been advised of the possibility of such damages.
**Cap on Liability:** To the extent not prohibited by law, **ImagineArt's total cumulative liability to you for any claims arising out of or related to these Terms or the Service will not exceed the amount you have paid to ImagineArt for the Service in the twelve (12) months immediately preceding the event giving rise to liability** (or, if no fees have been paid, \$100 USD). This limitation applies to all causes of action in the aggregate.
**User Content and Use:** You are solely responsible for your use of the Service and any content you create or publish using the Service. You acknowledge that AI-generated outputs may be unpredictable, and you assume all risk for any actions you take based on the outputs. We will not be liable for any dispute between you and any third party (for example, if you believe someone else's output is similar to yours, or if someone claims your output infringes their rights). Any such dispute must be resolved between the parties involved, and you release ImagineArt from any claims or liability arising out of third-party actions or content.
**No Liability for Certain Types of Claims:** Some jurisdictions do not allow exclusion or limitation of certain damages or liability (such as in the case of intentional misconduct or gross negligence, or for death or personal injury caused by negligence, or breach of statutory duty). In those jurisdictions, ImagineArt's liability will be limited to the greatest extent permitted by law. Nothing in these Terms limits or excludes any liability that cannot be limited or excluded by law.
### 11. Dispute Resolution and Governing Law
If a dispute, controversy, or claim arises out of or relates to these Terms ("Dispute"), it will be resolved through binding arbitration rather than in court.
Both parties agree to first attempt, in good faith, to resolve any Dispute within thirty (30) days from when it arises. If the Dispute cannot be resolved within this period, it will be settled by binding arbitration before a single, neutral arbitrator mutually selected by the parties. The arbitration will be conducted in English, in Newcastle County, Delaware, United States.
By agreeing to arbitration, both you and Imagine Art knowingly and irrevocably waive the right to a trial by jury in any legal action, proceeding, or counterclaim, except that either party may seek injunctive or equitable relief in a court of competent jurisdiction to protect its rights while arbitration is pending. The arbitrator may also award equitable or injunctive relief consistent with these Terms.
The arbitrator's decision will be final and binding on both parties, and judgment on the award may be entered and enforced in any court with proper jurisdiction.
Each party will be responsible for its own attorneys' and experts' fees and expenses, regardless of the arbitrator's final decision regarding the Dispute.
These Terms are governed by and will be construed in accordance with the laws of the State of Delaware, USA, without regard to conflict of laws principles.
### 12. Indemnification
You agree to defend, indemnify, and hold harmless ImagineArt, its parent company, affiliates, and their respective officers, directors, employees, and agents (the "ImagineArt Parties") from and against any and all claims, liabilities, damages, judgments, awards, losses, costs, or expenses (including reasonable attorneys' fees) arising out of or related to: (a) **Your Content** (Inputs or Outputs) and your use of the Service; (b) your breach or alleged breach of these Terms; (c) your violation of any applicable law or regulation in connection with your use of the Service; or (d) your infringement or violation of any intellectual property, privacy, or other rights of any third party. This means you will pay all amounts another party claims from or against the ImagineArt Parties, as well as any expenses incurred by the ImagineArt Parties, resulting from your actions or content, to the extent such claims are caused by you.
ImagineArt reserves the right, at its own expense, to assume the exclusive defense and control of any matter otherwise subject to indemnification by you. In that case, you agree to cooperate with our defense of that claim. You must not settle any such claim without our prior written consent if the settlement requires you to admit any liability or to pay any money or otherwise affects our rights. We will use reasonable efforts to notify you of any such claim, action, or proceeding upon becoming aware of it.
**Exceptions:** You are not required to indemnify the ImagineArt Parties to the extent the claim arises from our own gross negligence, wilful misconduct, or fraud, or other liability that by law cannot be imposed on you. The indemnification obligations will survive any termination of your account or these Terms.
### 13. Termination and Suspension
**By You:** You may stop using the Service at any time. You may also delete your account at any time through your account settings or by contacting us (subject to account verification). Termination of your account will be effective once processed (it will take 7 days to process).
**By ImagineArt:** We may suspend or terminate your access to the Service (or certain features of the Service), or terminate these Terms as they apply to you, at any time **for any reason**, with or without notice. For example, we may do so if you violate these Terms, if we discontinue the Service (in whole or part), or if your use of the Service creates risk for us or for other users, or is unlawful. In most cases of minor violations, we will attempt to warn you and/or work with you to remedy the issue; however, we are not required to provide notice or an opportunity to cure before termination, especially for serious violations or repeat offenders. We also reserve the right to terminate accounts that have been inactive for an extended period or that were registered with throwaway email addresses, to free up resources.
**Effect of Termination:** Upon any termination of your account or these Terms: (a) the rights and licenses granted to you hereunder will immediately end; (b) you must stop using the Service and, if applicable, delete any software or applications provided to you as part of the Service; and (c) any user content or data associated with your account may no longer be accessible to you (we are not obligated to provide copies of your Outputs or data post-termination, so please ensure you save any important content beforehand). However, termination does not automatically delete content you have made public (for example, content you contributed to a public gallery or shared with others may remain accessible by those users). Sections of these Terms which by their nature should survive termination (such as indemnification, disclaimers, limitation of liability, content licenses granted to ImagineArt (to the extent necessary for ongoing rights), dispute resolution, etc.) **will survive**.
If your account was terminated by us due to a violation of these Terms or law, you are not entitled to create a new account to use the Service. We may prevent you from re-registering (for example, by blocking your email or IP address). If you believe your account was wrongfully suspended or terminated, you may contact us to appeal, though we make no guarantee of reinstatement.
**Data after Termination:** We will handle personal data following account deletion in accordance with our Privacy Policy. We may retain certain information as required by law or for legitimate business purposes (such as records of payments or communications). Any licenses you granted to ImagineArt to use your content (per Section 6.2) will survive as described therein (for example, content used for AI training with your opt-in may continue to be used to improve models, and public posts remain visible to others).
### 14. Governing Law
These Terms and any dispute arising out of or relating to these Terms or the Service ("Dispute") will be governed by and construed in accordance with the laws of the State of Delaware, United States **without regard to its conflict of law principles**, except as may be otherwise provided in the **Dispute Resolution** section below (for example, the Federal Arbitration Act governs the interpretation and enforcement of the arbitration agreement) or in supplemental terms for specific jurisdictions. If you reside outside of the United States, nothing in this governing law section has the effect of depriving you of the protections of consumer laws in your own country of residence which, by law, cannot be waived or overridden by contract. In such cases, you will retain the benefit of any mandatory provisions of the laws of your country of residence.
### 15. Updates to These Terms
ImagineArt may modify or update these Terms from time to time. If we make material changes, we will notify you by reasonable means, such as by posting the updated Terms on our website (with a new "Last updated" date at the top), and/or by sending a notice to the email address associated with your account. **Please review any changes carefully.** Unless we state otherwise, changes are effective immediately upon posting. By continuing to use the Service after updated Terms are in effect, you agree to be bound by the revised Terms. If you do not agree to any update, you must stop using the Service and, if applicable, cancel your subscription. We encourage you to periodically review the Terms to stay informed of any changes.
For changes that are necessary to comply with law or for any urgent security, legal, or regulatory reasons, we may not be able to provide advance notice, but will still post the updated Terms and indicate the changes.
### 16. Miscellaneous
* **Entire Agreement:** These Terms, together with any Supplemental Terms and our Privacy Policy, constitute the entire agreement between you and ImagineArt regarding the Service and supersede all prior agreements or understandings relating to the Service. Any additional or different terms proposed by you (for example, in a purchase order or email) are hereby rejected and will not apply unless expressly agreed to in writing by an authorized representative of ImagineArt.
* **No Waiver:** Our failure to enforce any provision of these Terms is not a waiver of our right to do so later. If we do expressly waive any provision of these Terms, such waiver is limited to the specific instance and context and does not constitute a continuing waiver.
* **Severability:** In the event that any provision of these Terms is held to be invalid, illegal, or unenforceable by a court or arbitrator of competent jurisdiction, that provision will be enforced to the maximum extent permissible and the remaining provisions of these Terms will remain in full force and effect.
* **No Agency:** You and ImagineArt are independent contractors, and these Terms do not create any partnership, joint venture, employment, franchise, or agency relationship between us. Neither party has the authority to bind the other or incur obligations on the other's behalf without prior written consent.
* **No Third-Party Beneficiaries:** These Terms are for the benefit of you and ImagineArt (and our successors and permitted assigns). Except as expressly provided, they are not intended to confer any rights or remedies on any third party. For example, no third party can enforce any provision of these Terms under any legal theory.
* **Assignment:** You may not assign or transfer these Terms or your rights or obligations under these Terms, by operation of law or otherwise, without our prior written consent. Any attempt by you to do so without consent is void. ImagineArt may freely assign or transfer these Terms (for example, in the event of a merger, acquisition, sale of assets, or by operation of law) without notice or consent. These Terms are binding on and will inure to the benefit of each party's permitted successors and assigns.
* **Force Majeure:** Neither ImagineArt nor you will be liable for any delay or failure to perform any obligation under these Terms (except payment obligations) if the delay or failure is due to unforeseen events beyond the reasonable control of the party, such as acts of God, natural disasters, war, terrorism, riots, embargoes, internet or telecommunications outages, government orders, or other force majeure event. The affected party shall use reasonable efforts to mitigate the impact of the force majeure event and resume performance as soon as practicable.
* **Notices:** ImagineArt may provide notices or communications to you via email, through your account, by posting on our website, or via other electronic means. You consent to receive electronic communications and you agree that any such notices satisfy any legal requirement that such communications be in writing. If you need to give notice to us, you must do so in writing via email to [support@imagine.art](mailto:support@imagine.art) or via registered mail to our mailing address (see Section 8 for DMCA notices, and our website for general contact information).
* **Headings and Interpretation:** The section titles in these Terms are for convenience only and have no legal or contractual effect. Words like "including" and "for example" are deemed to be followed by "without limitation." These Terms were drafted in English, and to the extent any translated version conflicts with the English version, the English version controls.
If you have any questions or concerns about these Terms or the Service, please contact us at [support@imagine.art](mailto:support@imagine.art). By using ImagineArt, you acknowledge that you have read and agree to these Terms. Thank you for reading, and we hope you enjoy creating with ImagineArt!
# Unlimited Generation Policy
Source: https://docs.imagine.art/policies/unlimited-generation-policy
This policy outlines how the Unlimited Generation feature works, including speed, fair usage, and system health considerations.
Unlimited models are not available in ImagineArt workflows. Usage will be charged at the standard credit rate.
## How Unlimited Generation Works
Unlimited Generations allows you to create images and videos without consuming credits. This feature is available only on specific plans and is tied to a designated model for a limited time.
* **Eligibility:** You must be on the required subscription plan.
### Model-Specific Access Windows
Unlimited access is granted on a per-model basis and is tied to your subscription billing cycle.
* **Activation:** Unlimited access for eligible models activates immediately upon subscription purchase.
* **Model-Level Durations:** Each model may have a different duration of unlimited availability (e.g., 7 days for specific Image models, 3 days for Video models). Check the specific model details on your manage subscription screen for its current window.
* **Expiration:** Once the designated window for a specific model expires, that model will return to credit-based usage.
## Tools with Unlimited Generations Usage
The following features do not consume any credits when utilizing the Unlimited Promo campaign:
* Image Generation
* Video Generation
These tools are part of limited-time promotional campaigns and will appear on your plan periodically.
## Speed and Queueing
With Unlimited Generations, you can generate content at no cost, but limits are in place to ensure a fast and fair experience for everyone.
### Simultaneous Generation Limit
To ensure consistent speed and stability for all users, the Unlimited feature supports one active generation at a time.
### Daily Fast-Generation Limit
All users have a daily limit for Fast-Generation speed.
* Once this limit is reached, you will be automatically blocked until 12:00 AM local time (a "cooling-off period").
* Alternatively, if the model supports it, your requests will automatically be moved to the Relaxed Generation Queue.
### Relaxed Generation Queue
The Relaxed Queue provides unlimited quantity of generations without a daily cap.
* Requests in this queue are processed more slowly (relaxed processing).
* This ensures you can keep generating without interruption while maintaining platform stability and fairness.
| Generation Type | Quantity Limit | Processing Speed | Applies When |
| -------------------------------------- | --------------------- | ------------------ | --------------------------------------------- |
| Fast | Per-day limit applies | Fastest processing | Within your daily limit |
| Relaxed — Available on specific models | No quantity limit | Slower processing | After hitting the daily Fast-Generation limit |
Speed and concurrency limits may be adjusted periodically to maintain optimal platform performance and fairness for all users.
## Fair Use Policy
Unlimited Generations is strictly intended for individual, human use only.
The following activities are prohibited and constitute a violation of our Fair Use Policy:
* Use of automation or scraping tools.
* Account or credential sharing with others.
* Reselling access to the Unlimited Generations feature.
If unusual activity is detected, your Unlimited Generations access may be temporarily paused while our Support Team reviews your account. Repeated violations may result in permanent suspension.
### Throughput and System Health
To maintain system stability and performance, we may occasionally:
* Limit concurrency (the number of simultaneous requests).
* Reduce speed after periods of sustained high activity.
If this occurs, user requests may temporarily enter a system queue. Specific thresholds and limits will evolve as we scale capacity.
## When Your Access is Paused or Disabled
To safeguard the service and ensure equitable access, we may pause or permanently disable the Unlimited Generations feature if we detect violations, including but not limited to account sharing.
### How to Appeal a Pause or Strike
If your access is paused, you may submit an appeal:
1. Email: `support@imagine.art`
2. Subject Line: "Request to appeal Unlimited access strike."
We will review your case promptly. If no violation is found, we will restore your Unlimited Generations access. Repeated suspicious activity, even after a restored appeal, may lead to the permanent suspension of your account.
## Need More Help?
If you are unsure how the Unlimited Generations feature applies to your specific subscription plan, please contact Customer Service at `support@imagine.art`. We are here to review your account and help you get back to creating as quickly as possible.
# Prompt
Source: https://docs.imagine.art/prompt
## Summary
The Prompt Node allows you to give clear instructions to the AI models to generate text, images, and videos. It's perfect for crafting custom prompts for text, images, or videos based on your specific needs.

### Key Features
* Write detailed prompts for specific outputs.
* Works for all; text, image, and video generation.
* Use it with other nodes for complex workflows.
## How to Use
1. **Add the Node:** Click the Add (+) button and select Prompt from the Text node category (it sits at the top level, not inside Text Utilities), or use the shortcut (P).
2. **Enter Your Prompt:** Write a detailed prompt to define the content you want to generate (e.g., "Generate a product description for an eco-friendly jacket").
3. **Connect with Nodes:** Connect this with other nodes for generating content, such as Image or Video.
### Sample Use Cases
"Write a catchy caption for a social media ad about our new eco-friendly sneakers."

"Generate an image of a beach sunset with a silhouette of a person practicing yoga."

# Quick Actions
Source: https://docs.imagine.art/quick-actions
**Available quick actions:**
* **Upscale** — Increase video resolution for better quality
* **Edit video** — Trim, crop, or add effects
* **Extend video** — Add additional content or length
* **Extract frame** — Capture a single frame as an image
**Available quick actions:**
* **Upscale** — Increase image resolution for improved quality
* **Edit image** — Apply color correction or saturation adjustments
* **Animate** — Create moving visuals from static images
* **Crop** — Remove unwanted areas and adjust composition
* **Remove background** — Isolate the subject by eliminating the backdrop
* **AI resize** — Smart resizing that maintains quality while adjusting dimensions.
**Video Node Quick Actions:**
* **Upscale** — Enhance video resolution and clarity by increasing pixel density. Useful for improving the quality of lower-resolution source footage or preparing videos for higher-definition displays and platforms.
* **Edit video** — Trim unwanted sections, crop the frame to focus on specific areas, adjust playback speed, and apply visual effects such as filters, transitions, and color grading to refine your video.
* **Extend video** — Lengthen your video by adding additional content, inserting supplementary clips, or repeating sections. Helpful for creating longer-form content or filling specific duration requirements.
* **Extract frame** — Capture a single frame from your video and export it as a high-quality image. Ideal for creating thumbnails, promotional graphics, or still images from video content.
**Image Node Quick Actions:**
* **Upscale** — Enhance image resolution and clarity by increasing pixel density. Useful for improving the quality of lower-resolution source images or preparing images for high-definition displays, printing, or larger canvas sizes.
* **Edit image** — Apply color correction to adjust brightness, contrast, and white balance. Use saturation adjustments to enhance or mute colors, fine-tune hue, and make targeted tonal refinements to achieve your desired visual aesthetic.
* **Animate** — Transform static images into moving visuals by applying motion effects, transitions, or frame-by-frame animations. Create dynamic content suitable for social media, presentations, or video backgrounds.
* **Crop** — Remove unwanted areas from your image and adjust composition to improve framing. Reposition subjects, change aspect ratios, or focus on specific details by precisely defining the visible area.
* **Remove background** — Isolate your subject by automatically detecting and eliminating the backdrop. Useful for creating transparent backgrounds, compositing images, or preparing product shots and portraits for further editing.
* **AI resize** — Intelligently resize images while maintaining visual quality and detail. The AI-powered algorithm preserves important content and sharpness while adjusting dimensions to fit specific requirements or canvas sizes.
# Relight
Source: https://docs.imagine.art/relight
## Summary
The Relight Node lets you change the lighting of any image by repositioning, recoloring, and adjusting the intensity of the light source. Using an interactive 3D light controller and quick-position presets (Top, Front, Back, Bottom, Left, Right), you can dramatically transform the mood and depth of an image without regenerating it.
## How to Use
Click the Add (+) button and select Relight from the Image node category.
Connect an image from another node (such as Generate Image or Edit Image), or upload one directly.
Drag the light point on the interactive 3D controller to place the light source around your subject, or use the quick-position buttons (Top, Front, Back, Bottom, Left, Right) to snap to a preset angle. Click Reset to return to the default position.
Set the Light Intensity, Color, and output settings to match the look you want.
Click Run Selected, and the AI will produce a relit version of your image based on your configured lighting.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| ---------------- | ----------------------------------------------- | ----------------------------------------------------------------------------------------------------------------- |
| Light Controller | Interactive 3D Widget | Drag to visually reposition the light source around your subject. |
| Quick Positions | Buttons (Top, Front, Back, Bottom, Left, Right) | Snap the light to a preset direction with one click. |
| Reset | Button | Returns the light position to its default. |
| Light Intensity | Slider (default: 7) | Controls brightness of the light source. Higher values produce stronger, more pronounced lighting. |
| Color | Color Picker (default: #ffffff) | Sets the color of the light source. Use warm tones for golden-hour effects or cool tones for moonlit atmospheres. |
| Aspect Ratio | Dropdown (1:1, 16:9, 4:3, etc.) | Defines the output image dimensions. |
| Resolution | Dropdown (1k, 2k, 4k) | Determines the output resolution. Higher values provide more detail but may take longer to generate. |
## Sample Use Cases
Relight a product image with clean, even front lighting for a professional listing, or add dramatic side lighting to emphasize texture and shape.
Transform a flat scene into a moody cinematic shot. Use a low-intensity warm light from the side for a golden-hour look, or a cool-toned top light for a thriller atmosphere.
Shift the light source on a portrait to simulate classic studio setups—front lighting for beauty shots, side lighting for dramatic contrast, or back lighting for silhouette effects.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Remove Background (Node)
Source: https://docs.imagine.art/remove-background-node
## Summary
The Remove Background Node strips the background out of an image, leaving just the subject. It's a simple, single-purpose utility node — one image in, one background-removed image out — for when you need a clean cutout inside a larger workflow rather than as a standalone tool.
## How to Use
Click the Add (+) button and select **Remove Background** from the Image Utilities sub-category.
Link an image from another node, or upload one directly.
Click Run. The node outputs the same image with its background removed.
This is the workflow-canvas node version of background removal. For the same capability outside the Workflow builder, see the standalone **Remove Background** tool available across the app's image editing surfaces.
## Sample Use Cases
Remove a product photo's background inside a workflow, then feed the cutout into a downstream node that composites it onto a new scene.
Clean up a reference image before passing it into a node like Carousel Maker or Ad Maker, so the subject isolates cleanly.
# Sound Effects
Source: https://docs.imagine.art/sound-effects
The Sound Effects Node generates audio effects from a text prompt. Describe any sound—footsteps on gravel, thunder rolling, a car engine starting, a sci-fi laser blast—and the AI creates a matching audio clip. Connect a prompt or type directly into the node to produce sound effects for videos, animations, games, or any creative project.
## How to Use
Click the Add (+) button and select Sound Effects from the Audio node category.
Connect a Prompt node to the Prompt input handle, or type a description directly (e.g., "Heavy rain on a tin roof with distant thunder").
Click Run, and the AI generates an audio clip matching your description.
## Sample Use Cases
Generate ambient sounds like rain, crowd noise, wind, traffic, and combine them with AI-generated videos using the Combine Audio & Video node.
Create specific action sounds like explosions, footsteps, door creaks, swooshes, etc., to layer onto animated scenes or motion graphics.
Quickly generate placeholder sound effects for game prototypes, UI clicks, power-ups, environmental audio, without sourcing from libraries.
# Split Image
Source: https://docs.imagine.art/split-image
## Summary
The Split Image Node takes a single image and splits it into a grid of smaller cropped sections. A visual preview shows exactly where the cuts will be made, so you can see how your image will be divided before running. This is useful when you need to isolate specific parts of an image for individual processing, create tiled compositions, or break down a large design into manageable pieces for downstream nodes.
## How to Use
Click the Add (+) button and select Split Image from the Image node category.
Connect an image from another node (such as Import or Generate Image), or upload one directly. A preview with dashed grid lines will appear on the image.
Choose how to divide the image from the Grid Size dropdown (e.g., 2x2, 3x3, 4x4). The preview updates in real time.
The node outputs each cropped section as a separate image that can be passed to downstream nodes individually.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| --------- | -------------------------------------------- | -------------------------------------------------------------------------------------------------------- |
| Grid Size | Dropdown (2x2, 3x2, 3x3, 4x2, 4x3, 4x4, 5x5) | Defines how many sections the image is split into. A 2x2 grid produces 4 images; a 5x5 grid produces 25. |
## Sample Use Cases
Break a storyboard sheet into separate frames so each scene can be processed independently—feed individual frames into Generate Image or Video nodes to bring each panel to life.
Split a full comic page into individual panels for translation, editing, or re-sequencing across different layout formats.
Divide a collage-style mood board into its individual reference images, then pass each one separately into an Edit Image or Generate Image node for inspiration-driven generation.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Split Text
Source: https://docs.imagine.art/split-text
## Summary
The Split Text Node allows you to divide a single text input into multiple separate outputs. It's ideal for projects that require breaking down large content into smaller, manageable pieces for further processing or analysis.
## How to Use
Click the Add (+) button and select Split Text from the Text Utilities node category.
Connect a text input (such as a document, paragraph, or content block) into the node.
Set your split parameters (by delimiter, line breaks, or character count) to divide the text as needed.
The node will automatically split the input into separate text outputs ready for downstream processing.
## Sample Use Cases
When generating video content from a script, break the script down into smaller sections or scenes. Each section can be processed to generate specific video content (e.g., background, animations, or characters) to later be stitched together.
For video or multimedia projects, break down the voiceover script into manageable sections. Each part can be converted into audio clips to be later synced with video content or used for interactive experiences.
## When to Use
1. **Content Segmentation for AI Models:** Split descriptive text into smaller sections for targeted image, video, and audio generation.
2. **Interactive Media:** Break scripts into manageable sections to generate interactive elements for AI-powered applications.
3. **Voiceover Creation:** Divide long voiceover scripts into parts for easy audio generation and syncing with visual content.
4. **Scene Structuring:** Organize complex scenes or stories into segments to simplify content generation for animation, effects, and video.
# Storyboard
Source: https://docs.imagine.art/storyboard
The Storyboard Node generates a complete storyboard image with multiple scenes arranged in a grid, all from a single text prompt or reference image. The AI interprets your description, breaks it into sequential scenes, and produces a numbered grid of visuals that tell the story. Choose from multiple grid sizes (2x2, 2x3, 3x3, and more) depending on how many scenes you need. The output can be used directly as a visual reference or split and animated into individual video clips downstream.
## How to Use
Click the Add (+) button and select Storyboard from the Image node category.
Link a Prompt or AI Copilot node with a description of the scene, story, or sequence you want to visualize (e.g., "A cinematic scene of Peaky Blinders").
Select how many panels the storyboard should contain from the grid size options (e.g., 2x2, 2x3, 3x3).
Click Run, and the AI generates a single storyboard image with numbered scenes arranged in your chosen grid layout.
## Sample Use Cases
Describe a scene or sequence and generate a quick storyboard to share with your team before committing to full video generation. Perfect for aligning on direction early.
Generate a storyboard, then connect it to a Split Image node to break it into individual frames. Feed each frame into a Generate Video node to animate the full sequence.
Create a series of visually connected scenes for Instagram Stories, carousel posts, or TikTok sequences — all from a single prompt.
## Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Text iterator
Source: https://docs.imagine.art/text-iterator
The Text Iterator Node lets you feed multiple text inputs into a workflow and process each one individually through the connected downstream nodes. Enter multiple prompts, descriptions, or any text content — and the iterator passes each one separately to the next node in the chain. It's essential for workflows where you need to generate multiple outputs from different text inputs in a single run.
## Summary
The Text Iterator Node lets you feed multiple text inputs into a workflow and process each one individually through the connected downstream nodes. Enter multiple prompts, descriptions, or any text content — and the iterator passes each one separately to the next node in the chain. It's essential for workflows where you need to generate multiple outputs from different text inputs in a single run.
## How to Use
Click the Add (+) button and select Text Iterator from the Text node category.
Type your content into the available input fields (Input 1, Input 2, Input 3, etc.). Click + Add Another Input to add more.
Link the iterator's output to any node you want to apply to each text input — such as Generate Image, Generate Video, or AI Copilot.
When the workflow runs, the iterator feeds each text input one by one through the connected nodes, generating a separate output for every input.
## Sample Use Cases
Enter a series of different scene descriptions and connect to a Generate Image node. Each prompt produces its own image — perfect for generating a full set of visuals in one run.
Write individual scene descriptions as separate inputs and connect to a Generate Video node. Each text input generates its own video clip, ready to be combined downstream.
Enter multiple ad copy variations and connect to an AI Copilot node to refine, expand, or translate each one — all in a single workflow run.
# Text-to-Speech
Source: https://docs.imagine.art/text-to-speech
Convert text into natural-sounding speech using AI voice models.
Add your script directly or use the AI writer to help craft it. Text-to-Speech supports multiple languages and lets you fine-tune the output with the following settings.
## Parameters
| Parameter | Description |
| ---------- | -------------------------------------------- |
| **Speed** | Controls how fast or slow the voice speaks. |
| **Volume** | Adjusts the loudness of the generated audio. |
| **Pitch** | Raises or lowers the tone of the voice. |
## Models
ElevenLabs' latest generation model with highly expressive, natural-sounding speech and advanced emotional range.
`Expressive` `High Quality`
MiniMax's high-definition speech model delivering studio-quality audio with nuanced intonation and clarity.
`HD Audio` `Studio Quality`
MiniMax's fastest speech model optimized for low-latency generation without compromising on voice naturalness.
`Fast` `Low Latency`
## Emotions
Choose an emotion to shape how the voice delivers your script:
`Happy` `Sad` `Angry` `Fearful` `Disgusted` `Surprised` `Neutral`
## Voices
All voices support multiple languages.
| Voice | Gender |
| --------------- | ---------- |
| Wise Woman | Female |
| Friendly Person | Non-binary |
| Deep Voice Man | Male |
| Calm Woman | Female |
| Casual Guy | Male |
| Lively Girl | Female |
| Patient Man | Male |
| Young Knight | Male |
| Determined Man | Male |
| Lovely Girl | Female |
| Decent Boy | Male |
| Imposing Manner | Male |
| Elegant Man | Male |
| Abbess | Female |
| Sweet Girl 2 | Female |
| Exuberant Girl | Female |
# Upscale Image
Source: https://docs.imagine.art/upscale-image
## Summary
The Upscale Image Node uses advanced AI models to enhance the resolution and quality of images. This node is ideal for increasing the detail, sharpness, and clarity of low-resolution images, making them suitable for high-quality prints, larger displays, or just improving overall visual appeal. Whether you're working with photos, artwork, or digital graphics, this node helps you refine your images with minimal effort, while preserving the original look and feel.
## How It Works
The Upscale Image Node leverages state-of-the-art AI models to intelligently upscale images by analyzing their content and enhancing pixel data.
* The AI works to retain fine details, improve sharpness, and fill in information that is missing due to low resolution.
* You can adjust the upscaling factor to control how much the image is enlarged, as well as apply additional refinement settings for the best result.
### Adjusting Settings for Optimal Results
| Settings | Type | Impact on Output |
| ---------------- | --------------- | ---------------------------------------------------------------------------------------------------------------------- |
| Upscaling Factor | 2x, 4x, 8x, 16x | Higher values increase the resolution and detail of the image. |
| Strength | 0-100% | Controls the intensity of the upscale effect. Higher strength increases the AI's impact on fine details and sharpness. |
| Noise Reduction | Slider | Higher values will smooth out the image but may remove some fine details. |
| Sharpness | 0-100% | Higher values increase the edge definition and clarity, but too high may lead to over-sharpening. |
| Detail Level | 0-100% | Specifies how much detail is added to the image during upscaling. |
| Seed | Seed | A fixed number to ensure reproducible results. |
## Sample Use Cases
Improve old or low-quality images that need to be enlarged for better clarity.
**How to Achieve It:**
* Upload the image: Choose a low-resolution image (e.g., old photo, low-quality digital artwork).
* Set the Upscaling Factor: Choose how much you want the image to be upscaled (e.g., 2x, 4x, or more).
* Generate: The AI will refine and upscale the image, improving its quality without losing essential details.
Enhance images from photoshoots, improving sharpness, resolution, and clarity for professional presentations.
**How to Achieve It:**
* Upload the photo: Choose the image from your photoshoot that needs refinement.
* Set Refinement Options: Customize sharpness, contrast, and other factors to improve the image's professional quality.
* Generate: The AI will upscale the photo, improving its overall visual quality for high-end projects.
## Upscale Image Models
Visit [Image Models](https://docs.imagine.art/ai-models/image/imagineart-2-0) to explore all available models and find the one that fits your needs for creating or transforming images.
# Upscale Video
Source: https://docs.imagine.art/upscale-video
## Summary
The Upscale Video Node increases the resolution and visual quality of an existing video using AI. It sharpens details, improves clarity, and enhances overall fidelity all while preserving the natural motion and smoothness of the original footage. Whether you're working with low-resolution AI-generated clips or older footage that needs a quality boost, this node prepares your videos for large displays, cinematic projects, or professional delivery.
## How to Use
Click the Add (+) button and select Upscale Video from the Video node category.
Link a video from another node (such as Generate Video, Import, or Extend).
Select your preferred Model and adjust the Upscaling Factor and other parameters from the Properties panel.
Click Run, and the AI will produce an upscaled version of the video with enhanced resolution and detail.
## Choosing the Right Settings
| Setting | Type | Impact on Output |
| ---------------- | ----------------------- | --------------------------------------------------------------------------------------------------------------------------- |
| Model | Dropdown | Selects the AI model used for upscaling. Different models balance between speed, detail preservation, and output quality. |
| Upscaling Factor | Dropdown (e.g., 2x, 4x) | Controls how much the resolution increases. Higher factors produce larger, more detailed output but take longer to process. |
| Seed | Number Input | A fixed number for reproducible results across generations. |
## Sample Use Cases
Most AI video models output at 720p or lower. Chain the Upscale Video node after Generate Video to bring clips up to 1080p or 4K for client delivery, showreels, or broadcast.
Enhance older or compressed video files from social media downloads, archived footage, or screen recordings by upscaling them to a higher resolution with improved sharpness and clarity.
Upscale video content for presentations on large screens, digital signage, or trade show displays where low-resolution artifacts become highly visible.
## Upscale Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# Voice Cloning
Source: https://docs.imagine.art/video-cloning
Generate speech in the voice of real people and original characters using AI.
## Parameters
| Parameter | Description |
| ------------- | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| **Stability** | Controls how consistent the voice sounds across generations. Range: **0–1**. Lower values allow more expressive variation; higher values produce a more stable, uniform delivery. |
## Voices
### Celebrity Voices
`Cait Blanchett` `Ema Watson` `Donuld Trumpp` `James Earl Johns` `Morgon Freeman` `Samual L. Jacksom` `Adelle` `Sir David Attenbro` `Tom Hiddleston` `Beyoncy` `Jasinda Ardern` `Grahem Norton` `Snoopy Dogg` `Tayler Swift` `Michela Obama` `Nelsan Mandela` `Jorden Peterson` `Elena DeGeneres` `Gorden Ramsay` `Jimmie Fallon` `Ophra Winfrey` `Ricki Gervais` `Elen Musk` `Barack Obama` `Queen Elizabet II` `Scarlett Johansson`
### Generic Voices
`Henry` `James` `Emilia` `William` `Alice` `Aiden` `Robert` `Brad` `Mason` `Quinn` `Jannet` `Lucy` `Thomas` `Rachel` `Monica` `Ben` `Sophia` `Knox` `Ember` `Luna` `Liam` `Jennie` `Anderson` `Aria` `Beckett` `Harlow` `Samantha` `Mogan` `Jaxon` `Maddox` `Ryder` `Asher`
# Video Iterator
Source: https://docs.imagine.art/video-iterator
## Summary
The Video Iterator Node lets you feed multiple videos into a workflow and process each one individually through the connected downstream nodes — the video equivalent of [Image Iterator](/image-iterator) and [Text Iterator](/text-iterator). Instead of running your workflow manually for every clip, the iterator handles the batch automatically, passing each video one by one to the next node in the chain.
## How to Use
Click the Add (+) button and select **Video Iterator** from the Video Utilities sub-category.
Upload video files directly (MP4, WEBM, or MOV, up to 30MB each) or click **Select from Assets** to pull in videos already in your library.
Connect the node's output to whatever should run on each video — an edit, upscale, or export node, for example.
Click Run. The workflow processes each video in the batch in turn.
## Sample Use Cases
Feed a folder's worth of raw clips through Video Iterator into a single edit or upscale node, instead of running the workflow once per clip by hand.
Use Video Iterator to apply the same downstream chain — for example Copyright Checker followed by Export — consistently across every video in a batch.
# Video production brief
Source: https://docs.imagine.art/video-production-brief
# ImagineArt Docs — Video Production Brief
**Prepared for:** Design Team\
**Date:** April 27, 2026\
**Purpose:** Identify documentation pages that require instructional demo videos\
**Total pages needing videos:** 132
***
## How to Use This Document
Each page listed below has written step-by-step instructions but no accompanying video. The goal is to create short screen-recorded demo videos for each page so users can follow along visually.
**For reference**, these pages already have videos and can be used as style/format templates:
* `generate-image.mdx` — Image generation tutorial
* `generate-video.mdx` — Video generation tutorial
* `getting-started.mdx` — Onboarding flow
* `workflows/Intro-to-workflows.mdx` — Workflow introduction
* `image-tools/edit-image.mdx` — Image editing walkthrough
* `video-tools/edit-video.mdx` — Video editing walkthrough
***
## Priority 1 — Workflows & App Builder
*Core user journey for power users. Highest impact.*
| File | Page Topic |
| ----------------------------------------- | -------------------------------------- |
| `workflows/run-your-first-workflow.mdx` | Running your first workflow end-to-end |
| `workflows/create-app.mdx` | Creating a workflow app |
| `workflows/app-builder.mdx` | Using the App Builder interface |
| `workflows/canvas-interface.mdx` | Navigating the canvas UI |
| `workflows/collaboration-and-sharing.mdx` | Sharing and collaborating on workflows |
| `workflows/organize-your-workflow.mdx` | Organizing and managing workflows |
| `workflows/core-concept.mdx` | Core concepts of the workflow system |
| `apps/imagine-apps.mdx` | Overview of Imagine Apps |
***
## Priority 2 — Image Tools
*Everyday image creation and editing features.*
| File | Page Topic |
| ------------------------------------ | ------------------------------------- |
| `image-tools/create-image.mdx` | Creating an image from scratch |
| `image-tools/characters.mdx` | Generating and customizing characters |
| `image-tools/camera-angles.mdx` | Controlling camera perspective |
| `image-tools/color-palettes.mdx` | Applying and managing color palettes |
| `image-tools/effects.mdx` | Adding image effects |
| `image-tools/elements.mdx` | Using image elements |
| `image-tools/image-prompt.mdx` | Writing effective image prompts |
| `image-tools/quick-actions.mdx` | Using quick action shortcuts |
| `image-tools/credit-consumption.mdx` | Understanding credit usage for images |
***
## Priority 3 — Video Tools
*Core video creation features.*
| File | Page Topic |
| -------------------------------- | ------------------------------------- |
| `video-tools/image-to-video.mdx` | Converting images to video |
| `video-tools/text-to-video.mdx` | Generating video from text prompts |
| `video-tools/elements.mdx` | Adding elements to video |
| `video-tools/video-effects.mdx` | Applying video effects |
| `video-tools/video-extend.mdx` | Extending video length |
| `video-tools/lipsync.mdx` | Animating lip-sync on characters |
| `video-tools/motion-control.mdx` | Controlling motion in generated video |
| `video-tools/video-credits.mdx` | Credit costs for video generation |
***
## Priority 4 — Editing & Transformation Tools
*Advanced tools for editing and transforming existing media.*
| File | Page Topic |
| ------------------------------- | --------------------------------- |
| `edit-video.mdx` | Full video editing workflow |
| `extend-video.mdx` | Extending an existing video |
| `extract-video-frame.mdx` | Extracting frames from video |
| `motion-transfer.mdx` | Transferring motion between clips |
| `multiple-camera-angles.mdx` | Generating multiple camera angles |
| `relight.mdx` | Relighting images and scenes |
| `split-image.mdx` | Splitting images |
| `storyboard.mdx` | Creating storyboards |
| `upscale-image.mdx` | Upscaling image resolution |
| `upscale-video.mdx` | Upscaling video resolution |
| `ai-resize.mdx` | AI-powered resizing |
| `combine-videos-and-audios.mdx` | Merging video and audio files |
***
## Priority 5 — AI Generation Features
*Core generation tools including audio, music, voice, and text utilities.*
| File | Page Topic |
| -------------------- | -------------------------------------- |
| `generate-audio.mdx` | Generating audio (speech, sound) |
| `generate-music.mdx` | Creating AI music tracks |
| `generate-voice.mdx` | Text-to-speech voice generation |
| `sound-effects.mdx` | Generating sound effects |
| `music-tools.mdx` | Overview of music tools |
| `image-iterator.mdx` | Batch image variation generation |
| `prompt.mdx` | Guide to writing effective prompts |
| `ai-copilot.mdx` | Using AI Copilot for text and analysis |
| `combine-text.mdx` | Merging text inputs |
| `split-text.mdx` | Splitting text inputs |
| `text-iterator.mdx` | Batch text processing |
***
## Priority 6 — Account & Onboarding
*Setup and billing flows — important for new user experience.*
| File | Page Topic |
| --------------------------------- | ---------------------------------- |
| `account/sign-up-sign-in.mdx` | Creating an account and signing in |
| `account/upgrade-downgrade.mdx` | Changing subscription tier |
| `account/updating-billing.mdx` | Updating payment method |
| `account/subscription-plans.mdx` | Overview of subscription plans |
| `account/cancel-subscription.mdx` | Cancelling a subscription |
| `account/password-recovery.mdx` | Recovering a forgotten password |
***
## Priority 7 — AI Model Reference Pages (44 pages)
*Individual model pages. Lower priority per page but high volume. Consider one demo video per model family rather than per page.*
### Image Models (20 pages)
| File | Model |
| ---------------------------------------- | ------------------ |
| `ai-models/image/dreamina-3-1.mdx` | Dreamina 3.1 |
| `ai-models/image/flux-2-max.mdx` | Flux 2 Max |
| `ai-models/image/flux-2-pro.mdx` | Flux 2 Pro |
| `ai-models/image/flux-dev.mdx` | Flux Dev |
| `ai-models/image/flux-ultra.mdx` | Flux Ultra |
| `ai-models/image/grok-imagine.mdx` | Grok Imagine |
| `ai-models/image/ideogram-v3.mdx` | Ideogram V3 |
| `ai-models/image/imagineart-1-5-pro.mdx` | ImagineArt 1.5 Pro |
| `ai-models/image/imagineart-1-5.mdx` | ImagineArt 1.5 |
| `ai-models/image/imagineart-2-0.mdx` | ImagineArt 2.0 |
| `ai-models/image/midjourney-v7.mdx` | Midjourney V7 |
| `ai-models/image/minimax-image.mdx` | Minimax Image |
| `ai-models/image/nano-banana-2.mdx` | Nano Banana 2 |
| `ai-models/image/nano-banana.mdx` | Nano Banana |
| `ai-models/image/qwen-image.mdx` | Qwen Image |
| `ai-models/image/recraft-v4-pro.mdx` | Recraft V4 Pro |
| `ai-models/image/seedream-v4-5.mdx` | Seedream V4.5 |
| `ai-models/image/seedream-v5-lite.mdx` | Seedream V5 Lite |
| `ai-models/image/z-image-turbo.mdx` | Z Image Turbo |
### Video Models (25 pages)
| File | Model |
| ----------------------------------------- | ------------------- |
| `ai-models/video/google-veo-3-1-fast.mdx` | Google Veo 3.1 Fast |
| `ai-models/video/google-veo-3-1-lite.mdx` | Google Veo 3.1 Lite |
| `ai-models/video/google-veo-3-1.mdx` | Google Veo 3.1 |
| `ai-models/video/grok-video.mdx` | Grok Video |
| `ai-models/video/hailuo-02-pro.mdx` | Hailuo 02 Pro |
| `ai-models/video/hailuo-02-sd.mdx` | Hailuo 02 SD |
| `ai-models/video/kling-2-1-pro.mdx` | Kling 2.1 Pro |
| `ai-models/video/kling-2-5-pro.mdx` | Kling 2.5 Pro |
| `ai-models/video/kling-2-6-pro.mdx` | Kling 2.6 Pro |
| `ai-models/video/kling-3-0-pro.mdx` | Kling 3.0 Pro |
| `ai-models/video/kling-o1.mdx` | Kling O1 |
| `ai-models/video/kling-o3.mdx` | Kling O3 |
| `ai-models/video/lucy.mdx` | Lucy |
| `ai-models/video/luma-ray-2.mdx` | Luma Ray 2 |
| `ai-models/video/pika-2-2.mdx` | Pika 2.2 |
| `ai-models/video/pixverse-v5-5.mdx` | Pixverse V5.5 |
| `ai-models/video/pixverse-v5.mdx` | Pixverse V5 |
| `ai-models/video/pixverse-v6.mdx` | Pixverse V6 |
| `ai-models/video/runway-4-5.mdx` | Runway 4.5 |
| `ai-models/video/runway-gen-4-turbo.mdx` | Runway Gen-4 Turbo |
| `ai-models/video/seedance-1-0-pro.mdx` | Seedance 1.0 Pro |
| `ai-models/video/seedance-1-5-pro.mdx` | Seedance 1.5 Pro |
| `ai-models/video/sora-2-pro.mdx` | Sora 2 Pro |
| `ai-models/video/sora-2.mdx` | Sora 2 |
| `ai-models/video/wan-2-2.mdx` | Wan 2.2 |
***
## Summary by Priority
| Priority | Category | Pages |
| --------- | ------------------------ | ------- |
| 1 | Workflows & App Builder | 8 |
| 2 | Image Tools | 9 |
| 3 | Video Tools | 8 |
| 4 | Editing & Transformation | 12 |
| 5 | AI Generation Features | 11 |
| 6 | Account & Onboarding | 6 |
| 7 | AI Model Reference Pages | 44 |
| — | Community & Other | 34 |
| **Total** | | **132** |
***
*Generated April 27, 2026 — ImagineArt Documentation Audit*
# Edit Video
Source: https://docs.imagine.art/video-tools/edit-video
Transform existing videos using natural language commands — remove objects, change environments, adjust lighting, and more without traditional editing software.
Edit Video lets you modify an existing video by describing what you want to change in plain text. Rather than using a timeline editor or compositing tools, you write a sentence describing the edit and the AI makes precise, pixel-level modifications to the clip.
This mode is powered by Kling and supports a wide range of edits — from removing unwanted elements and swapping backgrounds to changing the weather, time of day, and visual style.
## What you can edit
Add, remove, or replace characters, props, or items anywhere in the scene. Remove distracting elements cleanly without visible artifacts.
Swap the background or relocate the scene entirely — from a city street to a forest, from an office to an open field.
Shift the lighting mood (sunset, dawn, overcast), change the time of day, or adjust colour temperature and shadow intensity.
Add atmospheric effects such as falling snow, rain, fog, or heat haze. Create dramatic environmental changes from a single description.
Apply a new visual style to the entire video — convert footage to look like a painting, sketch, or cinematic grade.
Alter the camera's perspective, tilt, or zoom to change the viewpoint of a scene without reshooting.
## How to edit a video
Upload the video clip you want to edit.
**Input requirements:**
* Duration: 3–10 seconds
* Maximum file size: 200 MB
* Maximum resolution: 2K
Write a clear, specific natural language command describing the change you want to make.
**Example edit prompts:**
* `Remove all cars from the street and change the time of day to sunset`
* `Add snow falling in the background`
* `Replace the concrete wall with a dense forest`
* `Change the lighting to a warm golden hour glow`
* `Apply a vintage film grain style to the entire video`
Be direct and specific. Instead of "make it look more dramatic," describe the actual visual change: "add storm clouds overhead and reduce ambient light by half." Specific descriptions produce more predictable results.
You can upload up to 4 reference images or elements alongside the video to guide how new content should look. For example, if you want to replace a character's outfit, upload a reference image of the target clothing style.
When a video is present, you can upload up to **4 images or elements combined**. Without a video, you can upload up to **7 images or elements**.
Choose the video duration, resolution, and output format for the edited result.
Click **Generate**. The AI applies your described edits to the uploaded video while preserving the elements you haven't asked to change, maintaining subject consistency throughout the clip.
## Supported input media
| Media type | Specification |
| --------------------------- | ----------------------------------- |
| Video quantity | 1 video per generation |
| Video duration | 3–10 seconds |
| Video max file size | 200 MB |
| Video max resolution | 2K |
| Reference images | Up to 4 images (with video present) |
| Reference images (no video) | Up to 7 images or elements |
## What to do next
Add 5 seconds of new content to the end of your edited video.
Generate a new video from scratch using reference images to guide appearance.
Add talking-avatar lip movements to a video using audio or text.
See credit costs for Edit Video generations.
# Elements
Source: https://docs.imagine.art/video-tools/elements
Save a character or object as a reusable reference that keeps its appearance consistent across your videos and Workflows.
Elements are a persistent reference library — a set of images (or videos) of the same character or object that the model uses to keep its appearance consistent throughout a generation, or across multiple projects. Once an element is created, you can reference it in any supported tool without re-uploading your references each time.
Elements are available in both the **Video Tools** suite and directly inside **Workflows**, where you can `@mention` any Element inline in a prompt.
## Where to create an element
Elements can be created from multiple places:
**Video Tools**
* Reference Image to Video
* Edit Video
* Reference Video
In any of these, look for the **Elements** button at the bottom left of the prompt box.
**Workflows**
* Any Generate Video node with a supported Kling model — use the **Create New Element** option in the Elements panel, or type `@` in any prompt field to reference an existing Element.
## How to create a new element
Click the **Elements** button at the bottom left of the prompt box from either **Reference Image to video**, **Reference Video** or **Edit Video.**
Click **Create a new element** in the panel that appears.
Upload up to 4 reference images of your character or object. A frontal image is required adn you may enter 3 more images. Adding images from different angles — side, three-quarter, back — trains the model on all sides of the element and produces more consistent results.
Set an element name and an optional description to keep your library organized.
Save the element. It will now appear in your Elements library, ready to reference in any supported workflow.
## Using an existing element
Once an element is saved, you can reference it in two ways:
* **From the Elements panel** — click the Elements button and select an existing element from your library to attach it to the current generation.
* **Using @ in your prompt** — type `@` followed by the element name directly in any prompt field. The element renders as a chip in the prompt itself, and the model uses it as a visual reference for that subject.
Elements created in Video Tools are available in Workflows, and vice versa — your library is shared across the platform.
For best results, upload a clear frontal image as your primary reference and supplement it with side or three-quarter angle shots. The more angles the model has, the more reliably it maintains the element's appearance across different shots and camera angles.
## When to use elements
If your video features a recurring character, saving them as an element ensures their appearance stays consistent from shot to shot, regardless of changes in lighting, angle, or scene.
For product videos or branded content, saving a product as an element prevents visual drift between cuts and keeps details like color, shape, and texture accurate throughout.
Elements persist across sessions. You can reference the same character or object in separate videos without re-uploading your images each time.
# Starting Frame
Source: https://docs.imagine.art/video-tools/image-to-video
Animate a static image into a video by providing a start frame, or both a start and end frame, and letting the AI generate the motion in between.
Image to Video converts static images into dynamic video sequences. You provide one or two key frames — a start image and optionally an end image — and the AI generates the motion and intermediate frames that connect them into a smooth, continuous clip.
This mode is ideal when you already have a specific visual in mind and want to bring it to life, rather than generating imagery from scratch via a text prompt.
## How to use Starting Frame
Navigate to [Video Studio](https://www.imagine.art/video) from the left sidebar
Upload the image you want the video to begin from. This image defines the opening scene — the AI uses it as a precise anchor and builds outward from this frame.
For best results:
* Use a high-resolution image (minimum 300px on the shortest side)
* Prefer JPG, JPEG, PNG, or WEBP formats
* Keep the file under 10 MB
* Choose images with clear subjects and unambiguous composition
If you want to control where the video ends, upload a second image as the end frame. The model will generate the transition between your two images, interpolating motion, lighting, and perspective changes to create a coherent sequence. This option is available only with specific models and if you cannot find an option to upload your end frame, try changing the model selected.
When using both start and end frames, choose images that share some visual continuity — similar subjects, environments, or colour palettes — to help the model produce a plausible and visually coherent transition.
Select the **aspect ratio**, **duration**, and **resolution** that match your output requirements. The available options depend on the model you choose.
Not all models support end frames. Models that do are marked with **Start/End frame** in the [credit consumption table](/video-tools/video-credits).
Click **Generate**. The AI produces a video that transitions smoothly from your start frame, and to your end frame if provided. Generation typically completes within 30–60 seconds.
## Example use cases
Use a daytime landscape as the start frame and a night version of the same scene as the end frame. The model generates a natural time-lapse-style transition, shifting light, shadows, and atmosphere across the clip.
Upload a still of a character or vehicle and use a text prompt alongside the image to describe the intended motion — for example, "a racing car accelerating from a standing start." The AI animates the subject based on both the visual and descriptive inputs.
Use two images with different lighting or compositional states to create an abstract visual journey. Gradual shifts in colour, background detail, or perspective can produce compelling artistic animations with minimal effort.
## Supported input formats
| Input | Requirements |
| ------------------ | ----------------------------------- |
| Format | JPG, JPEG, PNG, WEBP |
| Minimum resolution | 300px (shortest side) |
| Maximum file size | 10 MB per image |
| Maximum images | Up to 4 images total per generation |
Image to Video uses the same video models as Text to Video. Each model has different support for start-only vs start-and-end frames. Check the tooltip in Video mode or the [credit consumption page](/video-tools/video-credits) to confirm what a specific model supports before uploading.
## What to do next
Use multiple reference images to guide the style and content of a new video, rather than defining start and end frames.
Add 5 more seconds to the end of your generated clip.
Apply cinematic VFX effects to your animated output.
See how credits are calculated for Image to Video generations.
# Lipsync
Source: https://docs.imagine.art/video-tools/lipsync
Transform a still image into a talking avatar by syncing facial movements and lip animations to audio or text-to-speech.
Lipsync turns a static portrait or avatar image into a speaking video. You provide an image and either an audio file or written text, and the AI generates realistic facial movements and lip animations that match the speech. The result is a professional-quality talking-head video produced in minutes.
## How to create a lipsync video
Select **Lipsync** from the modal that appears when you hover over **Video** from the left navbar
Select the model that best fits your project requirements. Each model offers different duration, resolution, and aspect ratio options.
| Model | Duration | Resolution | Takes Audio Input? |
| --------------------- | --------------- | ------------ | ------------------ |
| Kling 2.6 Pro | 5s, 10s | View Tooltip | No |
| Google Veo 3.1 Fast | 8s | 720p-1080p | No |
| Google Veo 3.1 | 8s | 720p-1080p | No |
| Wan 2.5 Speak | 5s, 10s | 480p-1080p | No |
| Kling Avatars 2.0 Pro | Audio dependent | View tooltip | Yes |
| Infini Talk | Audio dependent | 480p-720p | Yes |
| OmniHuman (Bytedance) | Audio dependent | View tooltip | Yes |
| Fabric 1.0 VEED | Audio dependent | 480p-720p | Yes |
| Wan 2.6 | 5-15s | 720p-1080p | Yes |
Models marked **Same as image** preserve the aspect ratio of your uploaded photo, which is useful when you want to avoid cropping or letterboxing.
Upload a clear photo or illustration of the face you want to animate. Select from your library of existing ImagineArt creations, or upload a new image.
For the best lip sync results:
* Use a front-facing or near-front-facing portrait (slight angles are acceptable)
* Ensure the face occupies a significant portion of the frame
* Avoid heavy occlusion of the mouth area (scarves, masks, hands)
* Use a well-lit image with the face clearly in focus
* A clean or simple background produces cleaner output
Depending on the model, you have two input options:
Type the script you want the avatar to speak. The model converts your text to speech using a built-in voice and syncs the facial animation to match. This option is supported by all lipsync models.
Upload a recorded audio file containing the speech you want to sync. This gives you full control over the voice, tone, pacing, and language. Check the individual model's settings panel to confirm audio upload support.
Ensure any audio you upload is speech you have the rights to use. Do not upload audio recordings of other individuals without their consent.
Some models accept a text prompt alongside the image and audio. Use this to describe the context or setting, for example: `avatar speaking in a classroom` or `presenter delivering a keynote on stage`. This can influence the generated background, lighting, and overall mood.
Click **Generate**. The AI processes the image and audio or text, then produces a video with the avatar's face animated to match the speech. Generation typically takes 30–60 seconds.
## Tips for best results
* **Use a high-quality source image.** Blurry or low-resolution portraits produce less accurate facial animations. A sharp, well-lit photo at 512px or above gives the model more detail to work with.
* **Keep audio clear and clean.** Audio with background noise, music, or multiple overlapping voices can confuse the sync algorithm. Use isolated speech recordings when possible.
* **Match duration to content.** Choose a model duration that fits your script length. If your script is 4 seconds of speech, selecting a 10-second model will result in silence or padding at the end.
* **Front-facing portraits perform best.** Profiles and severe three-quarter angles reduce the quality of lip movement mapping. The more of the front of the face that is visible, the more accurate the animation.
* **Simple backgrounds reduce visual artefacts.** Busy or complex backgrounds can sometimes show distortion around the face boundary. Solid or blurred backgrounds produce cleaner-looking output.
## What to do next
Generate video from text prompts to create the source clip you want to lipsync.
Change the background, lighting, or environment of an existing video.
Add 5 more seconds of content to the end of your video.
Understand how credits are consumed for lipsync generations.
# Motion Control
Source: https://docs.imagine.art/video-tools/motion-control
Animate a character image by transferring body motion from a reference video — the AI reads the movements in the clip and applies them to your character.
Motion Control transfers body movements from a reference video onto a character image. You supply a portrait or full-body image of a character and a short video clip containing the movements you want to replicate. The AI reads the body motion from the reference clip and animates your character to perform those same movements.
This is distinct from other video modes: you are not generating a scene from a prompt, and you are not animating an image with added effects. Motion Control is specifically about **motion transfer** — taking human movement data from one source and applying it to a different subject.
## How to use Motion Control
Navigate to **Video** in the left sidebar. Scroll through the available modes and select **Motion Control**. The input panel opens, accepting a character image and a reference video.
Upload the image of the character you want to animate. The quality of motion transfer depends significantly on the source image.
For best results:
* Show the full body or at least the upper body clearly in the frame
* Use a front-facing or near-front-facing pose (slight angles work; extreme side profiles reduce accuracy)
* Avoid heavily cropped images where limbs are cut off
* Use a clean, simple background — complex backgrounds can produce visual noise around the character boundary
* Ensure the image is well-lit and in focus
Upload the video containing the movements you want to transfer to your character. This is the **driving video** — the AI reads body position, joint angles, and motion timing from this clip and maps them to your character.
For best results:
* Use a video with a single person as the primary subject
* Ensure the subject in the reference video is clearly visible and well-lit
* Minimal background clutter improves motion tracking accuracy
* The subject should be performing the movement you want to transfer — avoid clips where the subject is partially obscured or moving out of frame
You can also choose from the built-in **motion reference presets** provided in the interface. These are curated clips covering a range of movements including walking, dancing, waving, and more — a useful starting point if you don't have a specific reference video.
Once both inputs are in place, click **Generate**. ImagineArt processes the motion transfer and produces an animated video of your character performing the movements from the reference clip.
Generation time varies based on the length of the reference video and the current processing queue.
## What affects output quality
| Factor | Impact |
| -------------------------- | ------------------------------------------------------------------------------------- |
| Character image pose | Front-facing images produce more accurate motion mapping than profiles |
| Character image clarity | Sharp, well-lit images with visible limbs give the model more to work with |
| Reference video quality | Clear, well-lit subjects with minimal occlusion improve motion tracking |
| Reference video complexity | Single-person clips with simple backgrounds outperform crowded scenes |
| Reference video length | Longer reference clips generate longer animations; processing time scales accordingly |
## Common use cases
Transfer natural human movement onto an illustrated character, game asset, or AI-generated portrait. This is useful for previewing how a character design moves before investing in full animation production.
Use a dance performance video as the reference to make a character image perform the same choreography. Works well for music content, social media videos, and promotional material.
Apply a presenter's gestures and body language from a reference video to an avatar or illustrated character, creating an animated spokesperson without additional filming.
Only use reference videos of people whose motion you have the right to use. Do not upload recordings of individuals without their consent for commercial or public-facing projects.
## What to do next
Add speech and lip animations to a character after applying motion.
Modify the background or environment of your animated output.
Add 5 more seconds of content to the generated animation.
Understand credit costs for Motion Control generations.
# Reference to Video
Source: https://docs.imagine.art/video-tools/reference-to-video
Generate entirely new video content by uploading up to four reference images of characters, props, or scenes, and describing the scenario you want to create.
Reference to Video generates new video content inspired by reference images you provide. Unlike [Image to Video](/video-tools/image-to-video) — which uses an image as a literal starting or ending frame — Reference to Video extracts visual features from your references (appearance, clothing, props, environment details) and recreates those elements in a brand-new scenario you describe via a prompt.
Use this mode when you want a character or object from your reference images to appear in a completely new scene, action, or environment that doesn't exist in any of the source images.
## How Reference to Video differs from Starting Frame
| | Starting Frame | Reference to Video |
| -------------------------------- | ------------------------------------------------------------------ | -------------------------------------------------------------------------- |
| **Input role** | Defines the literal first (and optionally last) frame of the video | Provides visual features for the AI to extract and recreate |
| **Output relationship to input** | Video begins from and stays visually close to the uploaded image | Video depicts a new scenario; references guide appearance, not composition |
| **Prompt role** | Optional guidance for motion and style | Required to describe the scenario, action, and environment |
| **Best for** | Animating an existing scene or visual | Placing characters or objects in entirely new contexts |
## How to use Reference to Video
Navigate to [Video mode](https://www.imagine.art/video) from the left sidebar.
Upload one to four images of the subject(s) you want to appear in the video, characters, props, costume details, or scene elements. The model extracts visual features from all provided images and uses them to maintain consistency in the output.
For best results:
* Use images that show your subject clearly from multiple angles when possible
* Avoid heavily cropped or obscured images
* Provide images with consistent clothing, accessories, or design details if character or object consistency is important
* Formats supported: JPG, JPEG, PNG, WEBP (min 300px, max 10 MB each)
Describe the new scene or action you want the video to depict. Be specific about the environment, the action, the mood, and the camera angle.
**Example prompts:**
* `The character walking through a futuristic city at night, neon lights reflecting on wet streets, cinematic tracking shot`
* `A woman in a red dress dancing in a grand ballroom, warm candlelight, slow zoom out`
* `The robot standing on a rocky cliff overlooking a stormy ocean, dramatic wide angle, overcast sky`
Describe the setting in detail. Because the model generates an entirely new video rather than animating your reference, a strong environmental description helps anchor the output and maintain visual coherence with your references.
Choose the video **duration**, **aspect ratio**, and **resolution** for your output.
Click **Generate**. The model creates a new video that maintains the visual identity of your reference subjects while placing them in the scenario you described.
## Input media specifications
* Up to **4 images** per generation
* Minimum resolution: **300px** (shortest side)
* Maximum file size: **10 MB** per image
* Supported formats: **JPG, JPEG, PNG, WEBP**
Reference to Video can also be used alongside a reference video clip.
* **With a video present:** up to 4 images/elements combined
* **Without a video:** up to 7 images/elements
* Video duration: 3–10 seconds
* Video max file size: 200 MB
* Video max resolution: 2K
## Key capabilities
* **Multi-reference subject creation:** Combine up to four images of the same subject to give the model more information about their appearance, helping it maintain consistency in clothing, accessories, and distinguishing features.
* **Subject consistency across the clip:** Characters, props, and scenes remain visually stable throughout the generated video, even as the action and environment change.
* **Creative flexibility:** The AI can place your subjects in any scenario you can describe — new environments, action sequences, different camera angles, or lighting conditions entirely distinct from the source images.
## When to use Reference to Video vs other modes
* You have an existing character design, illustration, or photo and want to see it in a new scene
* You want to create multiple videos featuring the same character in different situations
* You need subject consistency across generated clips without being constrained by a specific starting frame
* You want the video to begin exactly from a specific image
* You want to animate a scene that already exists rather than create a new one
* The precise composition of your source image should be preserved in the output
* You don't have a reference image and want to generate everything from a text description
* You're exploring ideas and don't need visual consistency with existing assets
## What to do next
Animate a specific image as the literal start frame of your video.
Modify an existing video using natural language commands.
Transfer body motion from a reference video onto a character image.
Understand credit costs for Reference to Video generations.
# Reference Video
Source: https://docs.imagine.art/video-tools/reference-video
**Reference Video** mode allows you to create new video content using an existing reference video combined with additional reference images. Whether you want to replicate the style, scene elements, or specific features from your reference video, the mode enables you to generate a new video while maintaining key content and style consistency.
This mode is perfect for creators looking to create videos that match the look and feel of existing footage, but with new scenes, actions, or characters.
### How to Use?
Provide a reference video (3–10 seconds long) that will serve as the base for your new video.
Add up to four images or elements that guide the generation process, such as specific characters, props, or scene details.\\
Choose video duration, resolution, and output format according to your needs.
Click **Generate**, and the model will create a new video that aligns with your reference video and elements.
#### Input Media Supported
Reference video supports various input media to help guide video generation and editing, ensuring flexibility in your creative process.
**Videos:**
* **1 video** can be uploaded for use in **Reference Video**.
* **Duration**: 3 to 10 seconds.
* **Max File Size**: 200 MB.
* **Max Resolution**: 2K.
**Elements:**
* **Multiple images** (up to 4) can be uploaded or generated from different angles to form an element, providing more reference information for the model.
* **Note**:
* When a video is present, you can upload up to **4 images/elements combined**.
* Without a video, you can upload **up to 7 images/elements**.
#### Example Scenarios You Can Create
* **Style Matching**: Create a new video that matches the style, lighting, and camera work of your reference video, but with different characters or props.
* **Scene Replication**: Replicate the action or scene of your reference video in a new setting or with different characters, while keeping the same mood or tone.
* **Enhanced Elements**: Introduce new objects, characters, or backgrounds into an existing scene while maintaining visual coherence with the original reference video.
# Text to Video
Source: https://docs.imagine.art/video-tools/text-to-video

## What is Text to Video?
**Text to Video** is a tool designed to generate high-quality videos from textual descriptions. Simply describe a scene, action, or environment, and the AI model will generate a video that matches your vision. From dramatic action sequences to peaceful nature scenes, this feature allows creators to quickly generate engaging visual content based on nothing more than a few sentences.
## Make your first Video
Go to [Video Studio](https://imagine.art/video) to begin creating your video. You can choose from multiple options, choose "Text to Video"
In the prompt field, type a detailed description of the video you want to create. The more specific you are, the better the AI will understand your needs and generate a video closer to your vision. Describe actions, scenes, and visual elements clearly.
Select the AI Video model you want to use from the available options, including WAN, Kling, and Hailuo AI, based on the video style you need (e.g., cinematic, realistic, etc.). Then, choose the aspect ratio and other generation settings such as video length and resolution.
After finalizing your settings and prompt, click the **Generate** button. The AI will process your request and create a video based on the input you've provided.
# Video Credits
Source: https://docs.imagine.art/video-tools/video-credits
Understand how credits are consumed for video generation, including what factors affect cost and how to check your exact credit total before generating.
Credit consumption in Video Studio depends on the **model** you select and your **generation settings**. This page explains how costs are calculated and lists base credit costs per model.
Credit costs are updated regularly as new models are added and pricing changes. The **tooltip in [Video mode](https://www.imagine.art/video)** is always the authoritative source — check it before generating if you want to confirm the exact cost for your current settings.
## How to check credits before generating
In any [Video mode](https://www.imagine.art/video), click on the model dropdown and all video models display base credit cost
Select your desired options. The following settings can change the total credit cost depending on the model:
* **Duration** — longer videos cost more credits
* **Resolution** — higher resolutions (720p → 1080p → 4K) increase cost
* **Generate Preview** — creates preview frames before the final video, adding to the total
* **Prompt enhancer** — AI-assisted prompt improvement may consume additional credits
Hover over the **Create** button (or **Preview**, depending on the mode) to see the **total credits** for your exact current settings before you commit.
## Base cost vs total cost
**Base cost** is the starting price for a model at its minimum duration and default settings. **Total cost** increases when you select longer durations, higher resolutions, or optional features like preview generation.
Use the base cost table below to compare models, then use the in-app tooltip to confirm the exact cost for your specific configuration.
## Image to Video — base credit costs
Credits shown are the base cost for generating **1 video** at the model's minimum duration.
| Model | Base credits | Duration | Resolution | Aspect ratios | Options |
| ------------------- | ------------ | ---------------- | ----------------- | ----------------------------------- | ---------------------- |
| Happy Horse | 252 | 3s–15s | 720p, 1080p | 16:9, 9:16, 1:1, 4:3, 3:4 | Start frame, Audio |
| Kling 3.0 4K | 755 | 3s–15s | 4K | 16:9, 9:16, 1:1 | Start/End frame, Audio |
| Seedance 2 Fast | 515 | 4s–15s | 720p | 16:9, 9:16, 1:1, 3:4, 4:3, 21:9 | Start/End frame, Audio |
| Seedance 2 | 645 | 4s–15s | 720p, 1080p | 16:9, 9:16, 1:1, 3:4, 4:3, 21:9 | Start/End frame, Audio |
| Kling 3.0 Pro | 300 | 3s–15s | 1080p | 16:9, 9:16, 1:1 | Start/End frame, Audio |
| Kling O3 | 220 | 5s–10s | 1080p | 16:9, 9:16, 1:1 | Start/End frame |
| Runway 4.5 | 360 | 5s–10s | 720p | 16:9, 9:16, 1:1, 4:3, 3:4 | Start frame |
| Seedance 1.5 Pro | 72 | 4s–12s | 480p, 720p | 16:9, 9:16, 3:4, 4:3, 1:1 | Start/End frame, Audio |
| Pika 2.2 | 120 | 5s–10s | 720p, 1080p | 16:9, 9:16, 1:1, 2:3, 3:2, 5:4, 4:5 | Start frame, Audio |
| Luma Ray 2 | 300 | 5s–9s | 540p, 720p | 16:9, 9:16, 3:4, 4:3 | Start frame |
| Hailuo 02 SD | 160 | 6s–10s | 768p | 16:9 | Start/End frame |
| Hailuo 02 Pro | 290 | 6s | 1080p | 16:9 | Start/End frame |
| Hailuo 2.3 SD | 170 | 6s–10s | 768p | 16:9 | Start frame |
| Hailuo 2.3 Pro | 300 | 6s | 1080p | 16:9 | Start frame |
| Kling 2.1 Pro | 270 | 5s–10s | 1080p | 16:9, 9:16, 1:1 | Start/End frame |
| Kling 2.5 Pro | 210 | 5s–10s | 1080p | 16:9, 9:16, 1:1 | Start/End frame |
| Kling 2.6 Pro | 420 | 5s–10s | 1080p | 16:9, 9:16, 1:1 | Start frame, Audio |
| Kling O1 | 210 | 5s–10s | 1080p | 16:9, 9:16, 2:3, 3:2, 3:4, 4:3, 1:1 | Start/End frame |
| Seedance 1.0 Pro | 370 | 3s–12s | 480p, 720p, 1080p | 16:9, 9:16, 3:4, 4:3, 1:1 | Start/End frame |
| Seedance Pro Fast | 150 | 3s–12s | 480p, 720p, 1080p | 16:9, 4:3, 1:1, 3:4, 9:16, 21:9 | Start frame |
| Wan 2.2 | 30 | 5s | 720p | 9:16, 16:9, 1:1, 3:4, 4:3, 4:5, 5:4 | Start frame |
| Wan 2.5 | 300 | 5s–10s | 480p, 720p, 1080p | 16:9, 9:16 | Start frame, Audio |
| Wan 2.6 | 300 | 5s–15s | 720p, 1080p | 16:9, 9:16, 1:1, 4:3, 3:4 | Start frame, Audio |
| PixVerse v5 | 240 | 5s–8s | 540p, 720p | 16:9, 9:16, 3:4, 4:3, 1:1 | Start frame |
| PixVerse v5.5 | 120 | 5s–8s | 540p, 720p, 1080p | 16:9, 9:16, 3:4, 4:3, 1:1 | Start frame, Audio |
| PixVerse v6 | 135 | 5s–10s | 540p, 720p, 1080p | 16:9, 9:16, 1:1, 4:3, 3:4 | Start frame |
| Lucy | 240 | Shown in tooltip | 720p | 16:9, 9:16 | Start frame |
| Runway Gen 4 Turbo | 150 | 5s–10s | 720p | 16:9, 9:16, 1:1 | Start frame |
| Sora 2 | 240 | 4s–20s | 720p | 16:9, 9:16 | Start frame, Audio |
| Sora 2 Pro | 720 | 4s–20s | 720p, 1080p | 16:9, 9:16 | Start frame, Audio |
| Google Veo 3.1 Lite | 120 | 8s | 720p, 1080p | 16:9, 9:16 | Start/End frame |
| Google Veo 3.1 Fast | 350 | 4s–8s | 720p, 1080p, 4K | 16:9, 9:16 | Start/End frame |
| Google Veo 3.1 | 900 | 4s–8s | 720p, 1080p, 4K | 16:9, 9:16 | Start/End frame, Audio |
| xAI Grok 1.5 | 240 | 5s–15s | 480p, 720p | 16:9, 9:16, 1:1, 4:3, 3:4 | Start frame, Audio |
| xAI Grok Video | 180 | 6s–15s | 480p, 720p | 16:9, 9:16, 1:1, 2:3, 3:2, 4:3, 3:4 | Start frame, Audio |
| LTX 2.3 | 215 | 6s–10s | 1080p, 2160p | 16:9, 9:16 | Start frame, Audio |
## Other video modes
These modes use different inputs and have their own credit structures.
| Mode | Model | Base credits | Duration | Resolution | Notes |
| ---------------------------- | ---------------- | ---------------- | ------------------ | ----------------- | -------------------------------------------------------- |
| **Extend** | Same as original | Same as original | Based on extension | Based on settings | Uses the same model as the original video |
| **Reference Image to Video** | Kling O1 | 210 | 5s–10s | 1080p | Upload multiple reference images for subject consistency |
| **Edit Video** | Kling O1 | Shown in tooltip | Shown in tooltip | Shown in tooltip | Upload a video, then edit via prompt and references |
| **Reference Video** | Kling O1 | 1,000 | 5s–10s | 1080p | Upload a 3–10s reference clip to guide the next shot |
Credit consumption can vary based on your settings and how many preview frames or variations you generate. The figures above reflect a single generation at the minimum duration with default settings.
## Factors that affect credit cost
| Factor | Effect on credits |
| -------------------- | ------------------------------------------------------------------------------------------------------- |
| **Model choice** | Premium models (Sora 2 Pro, Google Veo 3.1, Kling 3.0 Pro) cost significantly more than standard models |
| **Duration** | Longer clips consume more credits; cost scales with duration |
| **Resolution** | Higher resolutions (1080p, 4K) cost more than 480p or 720p |
| **Generate Preview** | Enabling preview frames adds credits to the total before generation begins |
| **Prompt enhancer** | Using the AI prompt enhancer may add a small credit cost |
For a full explanation of how credits work, see [Understanding Credits](/overview/understanding-credits).
# Video Effects
Source: https://docs.imagine.art/video-tools/video-effects
Video Effects transforms static images into dynamic videos by applying realistic visual effects. Rather than generating a scene from a text prompt, you start from an image and select the type of effect you want to add. The result is a short video clip that brings your image to life with motion, lighting, and cinematic transitions.
## Available effect types
Video Effects supports a range of visual enhancements that you can apply to your input image:
Add natural movement to elements in the scene — camera pans, zooms, parallax depth, flowing water, swaying foliage, or other organic motion.
Introduce dynamic lighting changes such as flickering light sources, beams of light, reflections on surfaces, or shifting time-of-day illumination.
Apply signature cinematic looks — dramatic push-ins, slow reveals, depth-of-field shifts, or stylised camera movements that add production value.
Overlay weather and atmospheric elements such as rain, snow, fog, fire sparks, or particles to add mood and environmental character.
## How to apply video effects
Navigate to [Video Studio](https://www.imagine.art/video) and select **Text to Video** to begin.
At the top of the Video Generation panel, you have the option to select your effect.
Try multiple effects on the same image to compare results before committing credits. Each effect produces a distinct visual style, and the best choice depends on the mood and content of your source image.
Then you can select from any one of our pre-existing styles to help you achieve your vision:
Upload the image you want to apply the effect to. Use a clear, well-composed image for the best output.
Choose the **aspect ratio** and **generation settings** (duration and resolution) for your output video.
| Aspect ratio | Best for |
| ----------------- | ----------------------------------- |
| **16:9** | Landscape, YouTube, presentations |
| **9:16** | Portrait, TikTok, Instagram Stories |
| **1:1** | Square, Instagram feed |
| **4:3** / **3:4** | Tablet, poster formats |
Click **Generate**. The AI applies your selected effect to the image and produces a dynamic video clip. Generation typically completes within 30–60 seconds.
## What to do next
Add 5 more seconds of content to the end of your effects video.
Use natural language to change specific elements in your generated video.
Generate a new video from scratch with a text prompt.
See credit costs for Video Effects generations.
# Video Extend
Source: https://docs.imagine.art/video-tools/video-extend
Video Extend adds 5 seconds of new content to the end of an existing video. You can let the AI continue the clip naturally in Auto mode, or write a prompt in Manual mode to direct exactly what should happen in the extension.
## How to extend a video
Navigate to [Video Studio](https://www.imagine.art/video) and select **Video Extend** to begin.
You can select an existing asset from your uploads or upload one from your computer
The model analyses the last frames of your clip and generates a natural continuation that maintains the same scene, motion, and style.
Best when:
* The existing video has clear, ongoing motion that should continue (e.g. a camera pan, a walking character, flowing water)
* You want a quick extension without writing a prompt
* The visual content at the end of the clip provides enough context for the AI to continue coherently
Write a custom prompt describing what should happen in the additional video. You can specify new actions, introduce new scene elements, or direct a change in camera angle or subject focus.
**Example prompts:**
* `The character continues walking and turns a corner, revealing a busy marketplace`
* `The camera slowly pulls back to reveal the full landscape`
* `Rain begins to fall, intensifying with each second`
This is best when:
* You want to direct a specific continuation rather than letting the AI decide
* The clip ends at a neutral or static moment that doesn't naturally suggest what comes next
* You're building a multi-segment narrative and need each extension to advance the story
Click **Generate**. The AI produces the extension and appends it to your original clip, delivering a combined video as the output.
## What to do next
Change specific elements in your video using a text description.
Apply VFX effects to a generated or extended video.
Add talking-avatar lip animations to your clip.
Understand how credits are calculated for Video Extend.
# What is Video Studio?
Source: https://docs.imagine.art/video-tools/what-is-video
[**Video Studio**](https://www.imagine.art/video) is ImagineArt's end-to-end environment for AI video creation. Generate videos from text or images, edit existing footage with natural language, animate characters, apply cinematic effects, and extend clips — all without traditional editing software. Whether you're producing social content, concept animations, or production-ready footage, **Video Studio** gives you the tools to move from idea to finished video.
### **Powerful Creative Tools at Your Fingertips**
### Aspect Ratio Guide
Select the aspect ratio that fits your project's needs:
| Aspect Ratio | Usage |
| ------------ | --------------------------------------------------------------------- |
| **9:16** | Best for mobile-first content, Instagram Stories, and TikTok |
| **16:9** | Traditional video format, ideal for YouTube and presentations |
| **5:4** | Common for photography prints and framed artwork |
| **4:5** | Optimized for Instagram posts with more vertical space |
| **2:3** | Standard for photography, especially portraits |
| **3:2** | Classic for DSLRs and digital photography |
| **1:1** | Perfect for Instagram posts, profile pictures, and social media feeds |
| **3:4** | Frequently used for posters, digital flyers, and tablets |
| **4:3** | Traditional TV and PowerPoint presentations, used in older monitors |
### **How it works?**
Video Studio offers a suite of AI-powered tools to take your video from concept to completion:
1. **Create** a new video from a text prompt using our most advanced video generation models.
2. **Animate** a still image into a video with Image to Video, or build a scene from multiple reference images with Reference to Video.
3. **Edit** existing footage using natural language — remove objects, swap backgrounds, adjust lighting, and more.
4. **Apply effects** with Video Effects to add cinematic VFX, motion blur, or lighting treatments to any clip.
5. **Extend** any generated video by 5 seconds and guide the continuation with a custom prompt.
6. **Animate characters** with Motion Control by transferring movement from a reference video onto your subject, or bring portraits to life with Lipsync.
7. **Video Reframe** allows you to change the angle of an existing video e.g changing from a frontal shot to a high angle shot
8. **Video Recolor** allows you to change the color saturation and overall hue of the video
9. **LipSync Studio** allows you to
## Video Models
To view all available video models, [go here](https://docs.imagine.art/ai-models/video/seedance-2-fast).
# Video Trimmer
Source: https://docs.imagine.art/video-trimmer
## Summary
The Video Trimmer Node lets you cut a video down to a specific portion using a visual timeline editor. Preview the clip directly inside the node, drag the trim handles to set your start and end points, and output only the segment you want. It's the fastest way to remove unwanted sections before passing a clip to downstream nodes.
## How to Use
Click the Add (+) button and select Trim Video from the Video Utilities sub-category (the palette label is "Trim Video," though this page is titled "Video Trimmer").
Link a video from another node (such as Generate Video, Import, or Extend). The clip will appear in the built-in preview player.
Use the timeline strip at the bottom to drag the start and end handles to your desired segment. The preview player and time-codes update in real time so you can see exactly what you're keeping.
Click Run, and the node outputs only the trimmed portion of the video.
## Node Controls
| Control | Description |
| -------------- | -------------------------------------------------------------------------------------------------------------------------- |
| Preview Player | Play back the video directly inside the node to find the right moment. Includes a play/pause button and audio mute toggle. |
| Timeline Strip | A frame-by-frame filmstrip of the full video. Drag the start and end handles to define the trim range. |
| Timecodes | Displays the start time, current playhead position, and end time of the selected range. |
| Trim Presets | Quick-action buttons for common trims trim start, trim both ends, or trim end. |
| Undo / Redo | Step back or forward through your trim adjustments. |
## Video Models
Visit [Video Models](https://docs.imagine.art/ai-models/video/seedance-2-fast) to explore all available models and find the one that fits your needs for creating or transforming videos.
# What are Text Nodes
Source: https://docs.imagine.art/what-are-text-nodes
The Text category in ImagineArt Workflows offers a powerful suite of nodes designed to create, manipulate, and refine text-based content. These nodes are built to assist in everything from generating original text to combining, splitting, and enhancing it for complex workflows.
Text nodes are fully resizable — drag any edge or corner to make them as large or small as your canvas layout needs. Resizing is smooth and responsive, so even large multi-node workflows stay easy to navigate.
## Explore our Text Nodes
Write custom instructions to guide AI models in generating text, images, or videos.
Generate and analyze text using LLMs, with support for text, image, and video inputs.
Merge multiple text inputs into a single, unified output.
Break down long or complex text into smaller, manageable sections.
# Intro to Workflows
Source: https://docs.imagine.art/workflows/Intro-to-workflows
Build powerful AI pipelines on an infinite visual canvas. Connect 50+ models, organize your workspace, and collaborate with your team.
Workflows (also known as **Imagine Flow**) is ImagineArt's visual node-based canvas where you connect multiple AI models and tools into automated creative pipelines. Instead of switching between platforms or doing repetitive tasks manually, you build a graph of connected nodes—each performing one step—and run the whole sequence in one click.
## Why use Workflows?
Traditional AI design tools force you to choose between power and simplicity. You end up juggling multiple platforms or accepting limitations. Workflows removes these obstacles:
* **50+ models in one place.** Image, video, and text AI models are unified on a single canvas, eliminating the need to switch between tools.
* **Smart organization.** Use sections, pages, and folders to keep your workspace structured and clutter-free as your projects grow.
* **Collaboration.** Share your workflows with your team using a project link. Use built-in collaboration tools including shapes and sticky notes to communicate visually.
* **Node-based control.** A professional, node-based interface gives you precise control over every step of your pipeline—without requiring code.
## How it works
Every workflow is built from **nodes**—self-contained units that each perform one task. You connect nodes by dragging from an output handle on one node to an input handle on another. Data flows left to right across the canvas:
* **Input nodes** provide data (text prompts, uploaded images, audio files) to the next step.
* **Output nodes** generate or transform that data (a generated image, an upscaled video, a rewritten prompt).
* **Connections** are type-safe: image outputs connect to image inputs, text to text, video to video.
Once your nodes are wired together, click **Run** (or press `Ctrl/Cmd + Enter`) to execute the workflow. Each node processes in sequence, producing the final output from your input data.
Workflows are iterative by design. Adjust prompts, swap models, or rewire nodes between runs until the output matches your vision.
You can share any workflow with your team by sharing the project link from the Collaboration toolbar on the canvas.
## What you can build
* Generate an image from a prompt, upscale it to 4K, then resize it for every ad format in one run.
* Animate a still image into a video, extend it to a longer clip, add a voiceover with Lipsync, and export a finished asset.
* Feed a script through AI Copilot to expand it, split it into scenes, generate a video per scene, and stitch everything together.
* Batch-process dozens of product photos through the same editing prompt using Image Iterator.
## Explore the docs
Learn how to navigate the canvas, add and connect nodes, and organize your workspace.
Reference guide for every node type—image, video, text, and utility nodes.
Browse all 50+ image, video, and text models available on the canvas.
Open Workflows and start building.
# AI Models in Workflows
Source: https://docs.imagine.art/workflows/ai-models
Browse all image, video, and text AI models available on the Workflows canvas.
ImagineArt Workflows gives you access to 50+ AI models across three categories — image, video, and text — all available from the same canvas. Each node type (Image, Video, Audio) has a built-in **model selector**: drop a single node onto the canvas, pick any model from the dropdown, and the node's available parameters update dynamically to match that model's capabilities. You no longer need separate nodes for each model — one node handles them all.
Model availability may change as new models are added. The lists below reflect the current model catalog.
Image models are available in the [Generate Image](/workflows/understanding-nodes#generate-image), [Edit Image](/workflows/understanding-nodes#edit-image), and [Upscale Image](/workflows/understanding-nodes#upscale-image) nodes.
### Generation and editing models
| Model | Modes | Description |
| ------------------------------ | -------------- | ---------------------------------------------------------------------------- |
| **Nano Banana Pro** | Generate, Edit | Google's state-of-the-art image generation and editing model |
| **Seedream V4.5** | Generate, Edit | ByteDance Seedream 4.5 for text-to-image and image editing |
| **Nano Banana** | Generate, Edit | Fast-paced, dynamic rendering model for generation and editing |
| **Seedream V4** | Generate, Edit | ByteDance Seedream 4 for generating high-quality images |
| **Flux 2** | Generate, Edit | Great aesthetics and prompt adherence for text-to-image generation |
| **Ideogram Character** | Generate, Edit | Character-focused Ideogram generation and editing |
| **ChatGPT Image** | Generate, Edit | OpenAI's most advanced image model for realistic images |
| **Flux Pro Kontext** | Generate, Edit | Flux 1.0 Pro Kontext with text and reference image inputs for editing |
| **Flux Pro Kontext Max** | Generate, Edit | Premium editing with stronger prompt adherence and typography |
| **Flux Pro Kontext Max Multi** | Generate, Edit | Experimental multi-reference image editing for complex compositions |
| **Qwen Image** | Generate, Edit | Text-to-image with parallel CFG, LoRA, and turbo acceleration |
| **Flux Pro 1.1 Ultra** | Generate, Edit | Advanced version of Flux Pro 1.1 for detailed, realistic scenes |
| **Ideogram V3** | Generate, Edit | Advanced typography handling for intricate text elements and design visuals |
| **ImagineArt 1.5** | Generate | Great aesthetics and prompt adherence for high-quality images |
| **Z Image Turbo** | Generate | Latest image generation model from Alibaba with high-speed results |
| **Seedream V3** | Generate | ByteDance Seedream 3.0 for text-to-image generation |
| **Dreamina 3.1** | Generate | ByteDance Dreamina 3.1 with rich styles and enhanced prompt features |
| **Flux Pro 1.1** | Generate | High-quality text-to-image offering realistic and detailed outputs |
| **Flux Dev** | Generate | Flexible and customizable model ideal for experimental and conceptual design |
| **Recaraft V3** | Generate | Text-to-image designed for vector and typography-friendly outputs |
| **Minimax Image 01** | Generate | Fast, efficient text-to-image for quick, high-quality generation |
| **Ideogram V2 Turbo** | Generate | Fast, high-efficiency model for typography-heavy content |
### Upscale models
Used in the [Upscale Image](/workflows/understanding-nodes#upscale-image) node.
| Model | Description |
| ----------------------------- | ---------------------------------------------------------------------------- |
| **ImagineArt Subtle** | Enhances resolution with subtle refinement and natural detail |
| **Topaz Upscale** | Industry-leading upscaling technology enhancing texture while reducing noise |
| **ImagineArt Creative** | Creative upscaling with artistic detail and visual style enhancement |
| **FreePik Upscaler Creative** | AI-driven upscaling introducing stylistic detail and artistic enhancement |
| **FreePik Precision v1** | Focuses on image fidelity, sharpness, and natural color preservation |
| **FreePik Precision v2** | Advanced precision upscaling with adaptive content-aware algorithms |
For maximum multi-reference editing capability, use **Nano Banana Pro**, **Nano Banana 2**, or **Seedream V4.5**—these models support uploading 10+ reference images in the Edit Image node.
Video models power the [Generate Video](/workflows/understanding-nodes#generate-video), [Edit Video](/workflows/understanding-nodes#edit-video), [Extend Video](/workflows/understanding-nodes#extend-video), [Lipsync](/workflows/understanding-nodes#lipsync), [Motion Transfer](/workflows/understanding-nodes#motion-transfer), and [Upscale Video](/workflows/understanding-nodes#upscale-video) nodes.
### Generation models
| Model | Description |
| ----------------------- | ------------------------------------------------------------------------- |
| **Seedance 1.5 Pro** | Stable, fluid animations for high-end video production |
| **Seedance Pro Fast** | High-speed ByteDance model with flexible camera controls |
| **Wan 2.6** | Alibaba's cinematic multi-shot video AI with audio |
| **Kling 2.6 Pro** | Advanced AI video generation with detailed text/image-to-video creation |
| **Kling Omni** | Versatile video generator with multi-frame reference control |
| **Veo 3.1** | Google's high-quality cinematic video with audio generation |
| **Veo 3.1 Fast** | Rapid, cost-effective Google video with high-fidelity output |
| **Sora 2 Pro** | OpenAI's advanced physics-based cinematic video generation |
| **Pixverse v5.5** | Pro-level cinematic video with advanced motion brush control |
| **Hailuo 2.3 Pro** | Premium 1080p cinematic video with ultra-realistic physical dynamics |
| **Veo 3 Fast** | Rapid generation for quick creative iterations |
| **Sora 2** | Versatile world simulation with high-quality physics |
| **Wan 2.2 Turbo** | High-speed generation with advanced storytelling controls |
| **Wan 2.5** | Alibaba's realistic video model with native audio |
| **Kling 2.5 Pro Turbo** | Ultra-fluid motion with precise prompt control |
| **Hailuo 2.3 SD** | Balanced high-speed video with enhanced physics and micro-expressions |
| **Kling 2.1 Pro** | Professional motion quality for stable visuals |
| **Seedance 1.0 Pro** | ByteDance's high-quality text/image to video with flexible camera control |
| **Pixverse v5** | Hyper-realistic textures and improved consistency across frames |
| **Veo 2** | Cinematic videos with professional camera controls |
| **Kling 2.1 SD** | Balanced performance for high-speed video generation |
| **Pixverse v4.5** | Viral social media video with cinematic motion |
| **Pixverse v4** | Creative video generation with versatile cinematic camera controls |
| **Kling 1.6 SD** | Optimized for cinematic motion and complex narrative consistency |
| **Kling 1.6 Pro** | Maximum visual fidelity and longer narrative sequences |
| **Hailuo 02 Pro** | Handles complex physical interactions—object collisions, fluid dynamics |
| **Hailuo 02 Standard** | Optimized for rapid iteration and creative experimentation |
| **Seedance 1.0 Lite** | Fast and efficient ByteDance generation with flexible aspect ratios |
| **Decart Lucy 14B** | Lightning-fast image-to-video generation with cinematic motion |
### Upscale models
| Model | Description |
| ---------------------- | --------------------------------------------------------------- |
| **ImagineArt Upscale** | Upscales videos while maintaining quality and enhancing details |
| **Topaz Upscale** | Enhances video resolution with advanced algorithms for clarity |
### Lipsync models
| Model | Description |
| -------------------------------- | ------------------------------------------------------------------ |
| **Kling AI Avatar Pro i2v** | High-fidelity talking avatar with expressive emotion |
| **Kling AI Avatar Standard i2v** | Efficient, stable talking avatar for everyday use |
| **Omnihuman v1.5** | ByteDance's realistic full-body motion and lip-sync model |
| **Kling AI Avatar Pro** | Professional-grade lip-sync with detailed facial micro-expressions |
| **Kling AI Avatar Standard** | Reliable audio-driven lip-sync for static portraits |
| **Infinitalk Audio** | Video dubbing with precise lip synchronization |
### Motion transfer models
| Model | Description |
| ------------------- | ------------------------------------------------------------ |
| **Wan 2.2 Replace** | Transfers video motion to static character images |
| **Wan 2.2 Move** | Animates static characters using video motion reference |
| **Runway Act Two** | Advanced character animation with video performance transfer |
| **MoonValley** | Creative video transformation with style and motion control |
### Extend models
| Model | Description |
| ------------------------ | ------------------------------------------------------------- |
| **Pixverse Extend** | High-quality video extensions with seamless visual continuity |
| **Pixverse Extend Fast** | Rapid video lengthening with optimized processing speed |
Text models power the [AI Copilot](/workflows/understanding-nodes#ai-copilot) node. These are large language models (LLMs) from leading AI providers, available for generating, analyzing, and transforming text within your workflows.
| Model | Use cases |
| -------------------------- | --------------------------------------------------------------------------------------- |
| **Gemini 3.0 Pro Preview** | Multimodal powerhouse for complex reasoning, creative text, image, and video generation |
| **Gemini 2.5 Pro** | Google's model for deep text analysis, multi-step reasoning, and complex tasks |
| **GPT-5.1** | OpenAI's top-tier model for high-performance problem-solving and creative text |
| **GPT-5 Mini** | Optimized for fast response times; ideal for high-volume text generation |
| **GPT-4o** | Omni model with speed and multimodal capabilities for text, voice, and video tasks |
| **Claude Sonnet 4.5** | Balanced workhorse for content generation, analysis, and general-purpose workflows |
| **Grok 4** | Distinctive personality, real-time knowledge integration, and dialogue |
| **GPT-5 Nano** | Small, efficient model for basic text processing and simple tasks |
| **GPT-4o Mini** | Fast, cost-effective multimodal model for quick text generation |
| **Gemini 2.5 Flash** | Speed-optimized for high-quality multimodal tasks; ideal for large text content |
| **Gemini 2.0 Flash** | Efficient model for common tasks like blog posts, social media, and scripts |
| **Gemini 2.5 Flash Lite** | Compact model for rapid, low-latency text responses in real-time applications |
| **Claude Opus 4.1** | Ideal for complex reasoning, creative writing, and detailed content generation |
| **Claude Haiku 4.5** | Fast, cost-effective model for quick tasks: dialogue, captions, and summaries |
For workflows that require analyzing images or videos as inputs to the AI Copilot node, choose a multimodal model such as **Gemini 3.0 Pro Preview**, **GPT-4o**, or **Gemini 2.5 Flash**.
# What is App Builder
Source: https://docs.imagine.art/workflows/app-builder
### Summary
App Builder lets you package any workflow into a standalone app that anyone can use - without ever seeing the canvas. Define what goes in, define what comes out, and everything in between runs behind the scenes. Publish it privately, share it with your team, or launch it to the Marketplace and earn credits every time someone runs it.
### Overview
If you've built a workflow you want to reuse or share with others - App Builder is how you turn it into a one-click experience.
What you can do:
* **Hide the complexity:** Collapse a multi-node workflow into a clean app with just inputs and outputs.
* **Give users only what they need:** Choose exactly which parameters users can control (prompts, uploads, settings) and lock down everything else.
* **Add presets for instant results:** Provide pre-filled prompts and reference images so users can generate without writing anything from scratch.
* **Publish anywhere:** Keep it private, share it with your team, or launch it to the Community Marketplace.
* **Earn credits:** Set your own pricing tier and earn credits every time someone uses or clones your app.
### How It Works
Every app is built from a standard workflow with two special nodes:
| Node | What It Does |
| ----------- | -------------------------------------------------------------------------------------------------------------------------------------------- |
| Input Node | Collects everything the user controls - text prompts, image uploads, settings - into a single panel. This becomes the left side of your app. |
| Output Node | Displays the final result - the generated image, video, or audio. This becomes the center of your app. |
The workflow between them runs exactly as it would on the canvas. The user never sees the nodes, connections, or intermediate steps - just a clean interface with inputs on the left and results in the center.
> One Input Node + One Output Node = one app per canvas.
### Getting Started
1. Open any workflow on the canvas.
2. Click the Builder tab in the left panel.
3. Add an Input Node and an Output Node to your workflow.
4. Switch between Editor (canvas view) and App (user-facing view) using the toggle at the top center.
### The Builder Flow
App Builder walks you through four stages:
#### Build
These pages cover how to create and configure your app.
Add Input and Output nodes to your workflow, expose parameters with "Set As Input," add categorized presets, and preview the app view. One Input Node + One Output Node = one app per canvas.
Configure your app's name, description, visibility (Public, Team, or Personal), remix access, monetization tier, and thumbnail, then publish. Community apps go through moderation review before going live.
#### Discover & Earn
These pages cover the marketplace and how credits work.
Publish apps to the Community Marketplace and earn credits every time someone uses or clones your app. Set your earning tier (None, Low, Medium, High) and track your income from the My Apps dashboard.
Browse and manage all your apps across four tabs; Explore, Community, Team Apps, and My Apps. Track metrics (credits earned, runs, clones, likes), edit workflows, manage versions, and unpublish from one place.
### Key Concepts
| Concept | What It Means |
| ------------ | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Input Node | Collects all user-facing parameters (prompts, uploads, settings) into a single panel. Users interact with this when they use your app. |
| Output Node | Displays the final generated result (image, video, audio). Connects to the last generation node in your workflow. |
| Set As Input | A shortcut button on any node parameter that automatically connects its input handle to the Input Node — exposing it to app users. You can also connect handles manually by dragging on the canvas. |
| Presets | Pre-filled input options (prompt templates, reference images) organized by category. Makes apps easier to use out of the box. |
| Remix Access | When enabled, users can clone your app and modify the workflow behind it. You earn credits from every clone. |
| Monetization | Set an earning tier (0–30 credits) on top of the base generation cost. Earned credits are transferred to your account per use. |
| Versioning | Each republish creates a new version. Last 5 versions are archived. Users are notified when updates are available. |
### Common App Patterns
Here are some popular ways creators use App Builder. Each pattern follows the same idea: the user provides simple inputs, and the workflow handles the rest.
| Pattern | What Users Do | What the Workflow Does |
| ---------------------- | ------------------------ | ---------------------------------------------------------------------------------- |
| Simple image generator | Type a prompt, click Run | Generates an image. Add presets for one-click generation. |
| Photo enhancer | Upload a photo | Edits, upscales, and outputs a polished version. |
| Talking-head video | Enter a script | Generates video, adds lipsync, combines audio — complete video with synced speech. |
| Batch ad creatives | Enter multiple prompts | Generates and resizes assets for every platform using Text Iterator. |
| Product photography | Upload one product shot | Produces multiple angle variations automatically. |
| Storyboard generator | Describe a scene | Creates numbered frames ready for animation using Split Image. |
### What's Next
Ready to build? Start with [Create an App →](/workflows/create-app) to set up your Input and Output nodes and configure your app's interface.
# App Marketplace
Source: https://docs.imagine.art/workflows/app-marketplace
Publish apps to the community and earn credits when others use them
## App Marketplace
The Marketplace is where community-published apps live. Users discover, search, use, and remix apps built by other creators. As a publisher, you earn credits every time someone uses or clones your app - turning your workflows into a source of ongoing income.
### How the Marketplace Works
When you publish an app with Community visibility, it goes through moderation review. Once approved, it appears in the Community tab of Imagine Apps - searchable, categorized, and available to every ImagineArt user.
Community users can:
* **Use your app:** Enter inputs and generate assets. Credits are deducted from their account.
* **Like your app:** Help surface popular apps in the Marketplace.
* **Remix your app:** Clone it to create their own independent version (if you enabled Remix Access).
### How You Earn Credits
You earn credits in two ways:
#### From app usage
Every time someone runs your app, the credit cost breaks down like this:
**Total cost to user = Base cost + Your earning tier**
| Component | How It Works |
| ----------------- | -------------------------------------------------------------------------------------------------------------------------------------------- |
| Base cost | The sum of all generation nodes in your workflow, calculated automatically (e.g., 73.5 credits). This covers the AI model usage. |
| Your earning tier | The additional amount you set during publishing — None (0), Low (10), Medium (20), or High (30) credits. This goes directly to your account. |
**Example:** Your app has a base cost of 100 credits, and you set the earning tier to High (30). When a user runs it, they pay 130 credits total - 100 covers generation, and 30 are transferred to you.
#### From app cloning
When someone clones your app, a fixed credit amount is deducted from them and transferred to your account. This applies every time, regardless of the earning tier you've set.
### Tracking Your Earnings
All credit activity is tracked in the My Apps view. Each app card shows:
| Metric | What It Tracks |
| -------------- | --------------------------------------------- |
| Credits Earned | Total credits from usage and cloning combined |
| Runs | Number of times the app has been used |
| Clones | Number of times the app has been remixed |
| Likes | Total likes from community users |
Your full credit history including individual transactions is accessible from your account.
### Maximizing Your App's Reach
Your app competes with every other app in the Marketplace. Here's what separates the apps that get used from the ones that don't:
* **Choose a clear, specific name.** Users search by keywords. "Product Photo Background Remover" will outperform "My Cool App" every time.
* **Upload a strong thumbnail.** The thumbnail is the first thing users see when browsing. Use a compelling output image or a custom cover that shows what your app creates.
* **Add presets.** Apps with presets convert better. A new user can generate a result in one click — no prompt writing required.
* **Set a fair earning tier.** Higher tiers earn more per use but increase the cost for users. Start with Low or Medium and adjust based on demand.
* **Enable Remix Access.** Allowing clones increases your app's visibility and earns you credits from every clone, while building your reputation as a creator.
* **Write a clear description.** Tell users what your app does, what inputs it needs, and what kind of results it produces. Don't make them guess.
### Moderation
All Community apps go through a review before going live. The moderation team checks for quality, safety, and adherence to community guidelines.
Private and Team apps are published instantly without review.
# Canvas Interface
Source: https://docs.imagine.art/workflows/canvas-interface
Navigate the Workflows canvas, add and connect nodes, organize your workspace, and run your first workflow.
The **Canvas Interface** is the central space where you create, manage, and fine-tune your workflows. It provides a visual layout of connected nodes so you can see exactly how your tasks relate to each other—and modify, run, or monitor them at any point.
## Building your first workflow
Go to [imagine.art/flow](https://www.imagine.art/flow) and create a new workflow, or open an existing one from your dashboard.
Click the **+** button (or **Toolbox**) in the left toolbar to open the node picker. You can also right-click anywhere on the canvas, or press `Space`, to open the search panel.
Browse by category or search for a node by name, then click it to place it on the canvas.
Click the node to select it. The **right sidebar** displays the node's settings and the available AI models for that node type. Set your model, parameters (resolution, aspect ratio, seed, etc.), and any required inputs.
Add additional nodes as needed for your pipeline. To connect two nodes, click and drag from an **output handle** (right side of a node) to a compatible **input handle** (left side of another node), then release to form the connection.
Connections are type-safe: image handles connect only to image handles, text to text, and video to video.
Click the **Run** button inside a node or in the right sidebar, or press `Ctrl/Cmd + Enter` to execute. The workflow processes each node in sequence. You can also select specific nodes and click **Run Selected Nodes** to execute only part of your pipeline.
Review the output, adjust prompts or settings, and re-run. Workflows are designed for iteration—tweak any node and re-run without rebuilding from scratch.
## Canvas navigation
**Panning**
* Click and drag on empty canvas space to pan.
* Use the **middle mouse button** or a **two-finger drag** on a trackpad.
**Zooming**
* Scroll the mouse wheel up/down, or use a pinch gesture on a trackpad.
* Use the zoom controls in the bottom-right corner of the canvas.
**Minimap**
The minimap in the bottom-right gives you an overview of your full canvas layout, useful for large workflows with many nodes.
## Adding and managing nodes
| Action | How to do it |
| ----------------- | ------------------------------------------------------------------------- |
| Add a node | Click **+** in the left toolbar, right-click the canvas, or press `Space` |
| Search for a node | Press `Cmd/Ctrl + K` or use the Search button in the toolbar |
| Move a node | Click and drag it to the desired position |
| Duplicate a node | Select it and press `Ctrl/Cmd + D` |
| Delete a node | Select it and press `Delete` or `Backspace` |
## Connecting and removing handles
**To connect a handle:** Drag from an output slot on the right side of a node to a compatible input slot on the left side of another node. Release to form the connection.
**To remove a handle:** Hover over the connection handle—an **×** button appears. Click it to disconnect, or select the connection and press `Delete` or `Backspace`.
If a connection snaps back when you release, the two handles are incompatible types. Check that you're connecting an image output to an image input, text to text, or video to video.
## Left toolbar
The toolbar on the left side of the canvas gives you access to all major tools:
| Button | Function |
| ----------------- | ---------------------------------------------------------------------------- |
| **Core Nodes** | Essential generation nodes: Generate Image, Generate Video, Prompt, and more |
| **Utility Nodes** | Supporting nodes: Upscale, Crop, Resize, Extract Frame, Import, Export, etc. |
| **Search** | Find any node instantly (`Cmd/Ctrl + K`) |
| **Assets** | Access all previous generations from your account |
| **Move/Pan** | Switch to the pan tool for navigating the canvas |
| **Sections** | Add organizational sections to group related nodes together |
| **Collaboration** | Add shapes and sticky notes, and share the project link with your team |
## Right sidebar
When you select a node (or multiple nodes), the right sidebar shows settings specific to that selection:
* **Node properties and model settings** — Configure parameters for the selected node.
* **Model selection** — Choose from available AI models. Settings like aspect ratio, resolution, and other parameters update dynamically based on the model you select.
* **Number of runs** — Specify how many times the selected node(s) should execute per run.
* **Run selected nodes** — Execute only the selected nodes with their current settings.
## Version History
Every change you make to a workflow is automatically tracked. Version History lets you browse past states of your canvas and roll back to any previous version — so you can experiment freely without worrying about losing work.
**To access Version History:**
Click the **Version History** button in the top toolbar (clock icon), or open it from the workflow's context menu.
The panel lists all saved states with timestamps. Click any entry to preview that version of the canvas.
Select the version you want and click **Restore**. The canvas reverts to that state. Your current version is preserved as an entry in the history so you can always go forward again.
## Draw on Canvas
You can sketch, annotate, and ideate directly on the workflow canvas without leaving the platform. Drawing is free-form — use it to mark up connections, diagram ideas, or leave visual notes for collaborators.
**To start drawing:**
Click the **Draw** icon in the left toolbar (pencil icon) to activate free-draw mode.
Click and drag anywhere on the canvas to draw freely. Drawings float above the node layer so they don't interfere with connections.
Switch to the **Eraser** tool to remove specific strokes, or use **Clear Drawing** to wipe all annotations at once.
## Organizing your workspace
Use **Sections** (from the left toolbar) to group related nodes into labeled containers. This keeps large workflows readable and makes it easy to identify different stages of a pipeline—for example, separating a "Generation" section from a "Post-processing" section.
You can also use the **Collaboration** tools (shapes and sticky notes) to annotate your canvas with explanations or design decisions, which is useful when sharing workflows with teammates.
## Keyboard shortcuts
| Shortcut | Action |
| ---------------------- | ---------------------------------- |
| `Ctrl/Cmd + Enter` | Run the entire workflow |
| `Space` or right-click | Open node picker on canvas |
| `Cmd/Ctrl + K` | Search for a node |
| `Ctrl/Cmd + D` | Duplicate selected node |
| `Delete` / `Backspace` | Delete selected node or connection |
| `P` | Add a Prompt node |
| `T` | Add an AI Copilot node |
# Canvas Navigation
Source: https://docs.imagine.art/workflows/canvas-navigation
Learn how to navigate, add, and manage nodes on the canvas, connect handles, and use the toolbar and sidebar features.
## Adding and Managing Nodes
**Add a new node:**
* Click on `"+"` or `"Toolbox"` from the left toolbar

* Right-click anywhere on the canvas (or press `Space`)
* Search or browse by category

* Click to add the node
**Move a node:**
* Click and drag it to the desired location
**Remove a node:**
* Select it (Highlighted) and Press `Delete` or `Backspace`
**Duplicate a node:**
* Select it and Press `Ctrl/Cmd + D`
## Connect and Manage Handles
**Connect a Handle:**
* Click and drag from an output slot (right side of a node) and connect to a compatible input slot (left side of another node)
* Release to connect the handle

**Remove a Handle:**
* On hovering the handle, a `"x"` button appears to remove/delete it
* Click to select it (Highlighted) and Press `Delete` or `Backspace`
## Cut Connections with Scissors
The **Scissors** tool lets you sever node connections by drawing a cut line across them — useful for quickly cleaning up complex graphs without having to click each handle individually.
**To use the Scissors tool:**
Hold `Alt` (Windows) or `Option` (Mac) and click-drag across the canvas, or select the Scissors tool from the left toolbar.
Drag the cut line across any connections you want to remove. Every handle the line crosses will be severed.
Use the Scissors tool when reorganizing a large workflow — it's much faster than deleting handles one by one.
## Canvas Navigation
**Pan the canvas:**
* Click and drag on empty space
* Use the **middle mouse button** or **two-finger drag** on a trackpad
**Zoom:**
* Use the **scroll wheel** (up/down) or the **pinch gesture** on a trackpad
* Use zoom controls in the bottom-right corner
## Left Toolbar
| Button | Function |
| ----------------- | ------------------------------------------------------------------------------------------------------------------------------------------ |
| **Core Nodes** | Essential nodes for content generation such as Image, Video, Text, and Prompt Node |
| **Utility Nodes** | All other functional or helping nodes to assist you in editing or refining part, such as Upscale, Reframe, Crop, Extract Frame, Import etc |
| **Search** | Find any node instantly with `Cmd + K` |
| **Assets** | Access to all previous generations |
| **Move/Pan** | Navigate and move within the canvas |
| **Sections** | Organizational iterations for structuring and grouping canvas nodes |
| **Collaboration** | Intuitive collaboration tools featuring shapes, sticky notes, and share the project link with your team |

## Right Sidebar
**Nodes Properties and Model Settings:**
When you select a node or multiple nodes, the right panel displays settings specific to your selection and the chosen model.

| Feature | Description |
| -------------------------------------- | ----------------------------------------------------------------------------------------------------------------------------------------------- |
| **Node Properties and Model Settings** | When a node or multiple nodes are selected, the right panel shows specific settings and the model configurations |
| **Model Selection** | Choose from available models; settings and parameters (such as aspect ratios, resolution, etc.) will dynamically adjust based on your selection |
| **Number of Runs** | Specify how many times the selected node(s) should execute |
| **Run Selected Nodes** | Execute the selected node(s) with your configured settings |
# Collaboration and Sharing
Source: https://docs.imagine.art/workflows/collaboration-and-sharing
## Sharing Workflows
Share your workflows with your team or community:
Located in the top-right navigation.
Share with anyone on your team or community.

## Collaboration Tools
Easily organize and collaborate with your team:
* **Sticky Notes** — Add annotations and reminders to your workflows
* **Shapes** — Use geometric shapes with built-in text functionality for visual organization

# Core Concept
Source: https://docs.imagine.art/workflows/core-concept
Understanding the core concepts of workflows in ImagineArt will empower you to build and customize your creative processes efficiently. Here's a breakdown of the fundamental elements that make up a workflow:
## 1. Nodes: The Building Blocks of Workflows
Think of nodes as the building blocks of your workflow. Each node represents a task or action, and can either provide input for another node or generate output that serves as input for the next task.
There are two primary roles for nodes in a workflow:
* **Input Nodes**: These nodes provide data to another node. For example, a Prompt Node may provide the text for an AI model.
* **Output Nodes**: These nodes generate data, which can then be used as input for another node. For example, an Image Node generates an image based on a prompt.
Nodes can be connected to form a sequence that creates a fluid, step-by-step process in your workflow.

## 2. Connecting Nodes
Once you have nodes, you need to connect them. Nodes are connected through output handles (on the right side) and input handles (on the left side). Data flows from left to right across your canvas, just like reading a sentence.
* **Image** connects only to image handles.
* **Text** connects only to text handles.
* **Video** connects only to video handles.

### How Connections Work
* Data flows from Input Nodes (e.g., a text prompt) to Output Nodes (e.g., an image generation model).
* Proper connections ensure your workflow operates smoothly from start to finish.
You can easily visualize this data flow on your canvas, making it simple to make adjustments as needed.
**Key Takeaways**
* **Input Nodes** provide data that feeds into other nodes.
* **Output Nodes** generate data that becomes the input for the next task.
* You can link nodes together to create creative workflows, where the output from one task becomes the input for the next.
## 3. AI Models: Powering Your Workflow
AI models are where the actual AI-driven tasks happen within your workflow. By selecting different models, you can perform a wide range of actions, from generating images to creating videos. Models give you flexibility and power in your creative process.
**Examples**:
* An Image Model takes a Prompt Node (input) and generates a custom image (output).
* A Video Model takes an Image Node and a Prompt Node (input) and turns them into a video (output).

## 4. Execution and Running the Nodes
Once your nodes are set up and connected, it's time to run your workflow. Running the workflow processes each node in sequence, generating the final output based on the data you've provided.
**How to Run a Workflow**:
1. Click the Run button in the node or the node settings panel.
2. Alternatively, press Ctrl/Cmd + Enter to execute the workflow.
Running a workflow transforms your input data into tangible content, bringing your creative ideas to life.

## 5. Iteration: Refining Your Workflow
Workflows are iterative by nature, meaning you can adjust, test, and refine your workflows as you go. With each run, you can improve your workflow by tweaking nodes, updating prompts, or experimenting with different models.
**Why Iteration is Important**:
* Experiment with different node connections or models to refine your output.
* Tweak prompts for more precise results.
* Adjust settings to make sure the workflow evolves with your creative needs.
Iteration is about fine-tuning your workflow to achieve the best possible outcome. With each iteration, you move closer to your ideal result.
By mastering these core concepts, you'll be equipped to create powerful, dynamic workflows that bring your creative ideas to life. Whether you're starting from a featured workflow or building from scratch, understanding how nodes, connections, models, and iterations work together will enable you to customize and perfect your workflow.
# Create an App
Source: https://docs.imagine.art/workflows/create-app
Turn your workflow into a simple, shareable app
## Summary
App Builder turns any workflow into a standalone app. Add an Input Node to define what users can control, add an Output Node to define what they see, and everything in between runs behind the scenes. The result is a clean, focused tool — inputs on the left, results in the center.
### Before You Start
Make sure your workflow is complete and generating the results you want. App Builder packages what you've already built — so the better your workflow runs on canvas, the better it will run as an app.
Run your workflow at least once before entering the app builder. This ensures all your nodes are properly connected and producing output.
### Step 1: Build Your Workflow
Create your workflow as usual on the canvas — connect prompts, generation nodes, editing nodes, and any other steps you need. This is your pipeline. The app will run it exactly as-is, but users will only see the parts you choose to expose.
### Step 2: Open App Builder
Click the App Builder tab in the left panel. Your Input and Output nodes will directly be added in the canvas. You'll see options to add an Input Node and an Output Node if you mistakenly delete them from canvas.
### Step 3: "What should users control?"
This is where you define the app's interface. Input Node will collect all user-facing parameters into one place.
#### How the Input Node works
The Input Node collects parameters from your workflow through handle connections. Every node in your workflow has input handles - small connection points for parameters like prompts, images, or settings. To expose a parameter to app users, you connect that handle to the Input Node.
There are two ways to do this:
**Option 1: Connect handles directly** — Drag a connection from any node's input handle to the Input Node on the canvas. That parameter now appears in the Input Node and becomes part of the app's interface.
**Option 2: Use the "Set As Input" shortcut** — Go to any node's settings panel and click "Set As Input" next to the parameter you want to expose. This automatically creates the handle connection to the Input Node for you.
Not everything needs to be user-facing - that's the whole point. You choose which handles to connect. For example, you might expose the main prompt and an image upload, but keep the model selection, resolution, and step count locked to your preferred settings. Whatever isn't connected to the Input Node stays hidden and runs with your defaults.
#### Naming your inputs
Input field names are what users see when they open your app. Rename them to be clear and specific:
| Instead of... | Use something like... |
| ------------- | --------------------------- |
| Input 1 | "Describe your scene" |
| Image | "Upload your product photo" |
| Text | "Enter your script" |
| Parameter | "Choose a style" |
Good names tell users exactly what to provide — no guessing required.
### Step 4: Add Presets (Optional but Recommended)
Presets are pre-filled input options that let users generate results immediately — without writing a prompt from scratch.
#### How to add presets
1. Click "Add presets" next to any input field in the Input Node.
2. For text inputs: add preset prompts (e.g., a set of scene descriptions or style directions).
3. For image inputs: add a set of predefined reference images.
4. Click "+ Add another preset" to add more options.
5. Click "+ Add new category" to organize presets into groups (e.g., "Female", "Male", "Product Shots").
Apps with presets are dramatically easier to use. A new user can open your app, tap a preset, and get a result in one click — no prompt engineering required.
### Step 5: "What should users see?"
Drag the Output Node onto the canvas and connect it to the final generation node in your workflow. This defines what result appears when the app finishes running.
The Output Node supports multiple output types — images, videos, or audio — connected to a single node. Whatever your workflow produces at the end is what users see.
### Step 6: Preview Your App
Switch from Editor to App using the toggle at the top center of the screen. This shows you exactly what users will see:
| Panel | What It Shows |
| ------ | --------------------------------------------------------------------------------------------------------------------------- |
| Left | All input fields from the Input Node — prompts, file uploads, dropdowns, presets. Plus "Run App" and "Publish App" buttons. |
| Center | Generated output — images, videos, or audio displayed as a preview. |
| Right | Generation details — the prompt used, settings applied, and a thumbnail of the output. |
This is the real app interface. If something looks unclear or confusing to a first-time user, now is the time to rename inputs, adjust presets, or simplify.
### Step 7: Test Your App
Click "Run App" at the bottom of the left panel to test the full flow. Try different inputs, check that presets work correctly, and verify the output matches what you expect.
Test like a first-time user. Open the app view, ignore the canvas, and ask yourself: "If I had never seen this workflow, would I know what to do?" If the answer is no, simplify your input names or add more presets.
### Key Rules
* One Input Node and one Output Node per workflow — meaning one app per canvas.
* Both nodes must be connected for publishing to be available.
* Any node parameter can be exposed by connecting its input handle to the Input Node, or by using the "Set As Input" shortcut in that node's settings.
* Input field names are editable — always rename them to be user-friendly.
* Presets are optional but recommended — they lower the barrier for new users.
### Tips & Best Practices
* **Expose only what matters.** The fewer inputs users have to think about, the easier your app is to use. Lock down technical settings and only expose creative choices.
* **Use clear, specific names.** "Upload a front-facing product shot" is better than "Image Input." Users should know exactly what to provide.
* **Organize presets into categories.** If you have many presets, group them (e.g., "Portraits", "Landscapes", "Abstract") so users can browse without scrolling through a flat list.
* **Test with someone unfamiliar.** Your workflow makes sense to you because you built it. Have someone else try the app view to catch confusing labels or missing presets.
* **Keep the output clean.** Connect the Output Node to your final, polished result — not an intermediate step. Users expect the output to be the finished product.
### What's Next
Once your app is working the way you want, you're ready to publish it. See [Publish Your App →](/workflows/publish-app) to choose your visibility, set pricing, and go live.
# Feedback and Support
Source: https://docs.imagine.art/workflows/feedbaack-and-support
If you have suggestions, questions, or run into any issues, provide detailed feedback:
### On Canvas Feedback
Click the **Feedback** icon in the top-right corner to send us your thoughts directly.

### Community
* **Discord Community** - Join hundreds of Workflow creators at [discord.gg/F9XeAkH7](https://discord.gg/F9XeAkH7)
* **Support Email** - [support@imagineart.com](mailto:support@imagineart.com)
# The Brand Hub
Source: https://docs.imagine.art/workflows/imaginebusiness-brand-hub
Store your logo, colors, fonts, and visual direction once — apply them to every generation automatically.
**Brand Hub** is ImagineBusiness's brand-governance layer: two tools, **Brand Kits** and **Moodboards**, that let a team set its visual identity once and have it carried into every generation, rather than re-specifying brand details in every prompt.
## Brand Kits
Click **Brand Kits** to store your logo, colors, and fonts in one place.
Three ways to build one:
* **Enter your website URL** — the fastest path if your brand guidelines already live on your site.
* **Upload a brand file** — PDF or DOCX, up to 50MB, if you have an existing brand guidelines document.
* **Create from Scratch** — build a kit manually if you don't have either.
Sample brand kits (Blumenschon, Woofers, Quipli, Terracota) are shown as real examples of what a finished kit looks like — each with a logo, font pairing, and color palette.
## Moodboards
Click **Moodboard** to set a visual direction — mood, style, and look — that every generation then follows, separately from the literal brand assets in a Brand Kit.
Brand Kits and Moodboards solve different problems. A **Brand Kit** locks in *what's literally yours* — your actual logo, your actual brand colors and fonts. A **Moodboard** locks in *a feel* — the mood and visual style a generation should have, useful even before a formal brand kit exists.
Set up a Brand Kit before onboarding a new team member to ImagineBusiness. Once it exists, everyone generating content for that brand inherits the same visual identity automatically, instead of everyone individually re-describing the brand in their prompts.
# Keyboard Shortcuts
Source: https://docs.imagine.art/workflows/keyboard-shortcuts
| Shortcut | Action |
| ------------------------- | ---------------------------------------- |
| `Ctrl/Cmd + Enter` | Run Node (Selected Ones) |
| `P` | Prompt Node |
| `I` | Image Node |
| `V` | Video Node |
| `U` | Upload |
| `Double-click` | Anywhere on Canvas, will open Nodes List |
| `Cmd + K` | Open powerful search |
| `Delete` or `Backspace` | Remove selected nodes/elements |
| `Ctrl/Cmd + D` | Duplicate selected nodes or media |
| `Ctrl/Cmd + C` | Copy selected nodes or media |
| `Ctrl/Cmd + V` | Paste selected nodes or media |
| `Ctrl/Cmd + Z` | Undo |
| `Ctrl/Cmd + Y` | Redo |
| `Ctrl/Cmd + Tab` | Navigate among sections |
| `Enter` | Select children in the Section |
| `Shift + Enter` | Select Nodes again |
| `Ctrl/Cmd + Plus` | Zoom In |
| `Ctrl/Cmd + Minus` | Zoom Out |
| `Ctrl/Cmd + Mouse Scroll` | Zoom In / Zoom Out |
| `Shift + 1` | Zoom to Fit |
| `Ctrl/Cmd + A` | Select All |
| `Shift + S` | Add Section |
Workflow is built to keep you in flow. No friction. Just unlimited creative power.
**Happy creating!**
# Manage Your Apps
Source: https://docs.imagine.art/workflows/manage-apps
Browse, manage, and discover apps
## Summary
All your apps - published, in progress, or under review - live in the Imagine Apps hub. Browse what others have built, use apps shared by your team, or manage everything you've published from one place.
### Navigating the Tabs
When you open Imagine Apps, you'll see four tabs:
| Tab | What You'll Find |
| --------- | -------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Explore | Featured, trending, and recently published apps across the platform. This is the default landing page, a curated entry point for discovering what's available. |
| Community | All publicly published apps from the global Marketplace. Browse by category or search for specific use cases. |
| Team Apps | Apps shared by your team members. Visible to everyone in your team, regardless of which folder or workspace they were created in. |
| My Apps | Every app you've published — your personal dashboard with metrics, edit access, and version history. |
### My Apps: Your Dashboard
My Apps is where you manage everything. Each app card shows at-a-glance metrics so you can track performance without digging into details.
| Metric | What It Tracks |
| -------------- | ---------------------------------------------------------- |
| Status | Draft, Published, or Under Review |
| Credits Earned | Total credits earned from usage and cloning |
| Runs | Number of times the app has been used to generate assets |
| Clones | Number of times the app has been duplicated by other users |
| Updated | When the app was last modified |
#### What you can do from My Apps
* **Edit:** Jump directly into the workflow behind the app to make changes. Adjust nodes, update prompts, swap models - everything is accessible.
* **Republish:** After editing, republish to create a new version. Users on the current version will be notified that an update is available.
* **Unpublish / Delete:** Remove the app from its published scope (Community, Team, or Private) or delete it entirely.
* **View Version History:** Track which versions were published and when.
### What Users Can Do With Your App
Access levels depend on how you've configured the app:
| Access Level | What Users Can Do |
| ------------ | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| View & Use | Open the app, enter inputs, and generate results. They cannot see or modify the underlying workflow. |
| Clone | Duplicate the app to create an independent copy. The cloned version can be modified freely without affecting your original. *(Only available if you enabled Cloning Permission.)* |
| Edit | Only available to you - the app creator - from the My Apps view. Opens the workflow behind the app for direct editing. |
### Community Apps
Apps published with Community visibility appear in the global Marketplace after passing moderation. Community users can:
* **Explore and search:** Browse apps by category or search for specific use cases.
* **Use the app:** Input data and generate assets directly. Credits are deducted based on the base cost plus your earning tier.
* **Like the app:** Show appreciation and help surface popular apps in the Marketplace.
* **Clone the app:** Duplicate it to customize for their own needs (if cloning is enabled). A fixed credit amount is transferred to you.
### Team Apps
Apps published with Team visibility appear for all team members. Team members can:
* View and use the app to generate assets.
* Like the app.
* Clone the app (if cloning is enabled) to create their own independent version.
### What's Next
Want to maximize your app's reach and earnings in the Marketplace? See [App Marketplace →](/workflows/app-marketplace) for details on credit earning, pricing strategy, and discoverability tips.
# Monitor Status
Source: https://docs.imagine.art/workflows/monitor-status
Click the **Active Runs** tab in the left sidebar to monitor your workflow status:
Currently generating
Finished generations
Errors with details

# Organize your Workflow
Source: https://docs.imagine.art/workflows/organize-your-workflow
## Sections
Sections help you organize iterations and variations within a single workflow.
### Create a Section
For example, "Color Variations" or "Character Poses". Customize the color or tidy up the layout.
Drag existing nodes or create new ones within the section.
Use `Ctrl/Cmd + Tab` to navigate between sections and nodes.

### Use Cases
Test different prompt variations and compare results side by side.
Evaluate multiple model responses to find the best output.
Explore creative elements and refine your designs progressively.
Test alternative layouts and design compositions.
# Publish Your App
Source: https://docs.imagine.art/workflows/publish-app
Share your app with your team, the community, or keep it private
## Summary
Once your app is built and tested, publishing makes it available to others. Keep it for yourself, share it with your team, or launch it to the Community Marketplace where anyone can use it - and you earn credits every time they do.
### How to Publish
#### 1. Confirm your app is ready
Switch to App mode using the toggle at the top of the screen. Walk through the experience as a user would - check that inputs are clearly named, presets work, and the output looks right.
#### 2. Click "Publish App"
Hit the "Publish App" button at the bottom of the left panel. This opens the publishing configuration.
#### 3. "Who should have access?"
Choose where your app should be available:
| Visibility | Who Can See It | Where It Appears | Goes Live |
| ---------- | ---------------------- | ------------------------------- | ----------------------- |
| Private | Only you | My Apps | Instantly |
| Team | All your team members | Team Apps + My Apps | Instantly |
| Community | Everyone on ImagineArt | Community Marketplace + My Apps | After moderation review |
Not sure? Start with Private or Team. You can always change visibility later by republishing.
#### 4. Configure your listing
Fill in the details that help others discover, understand, and use your app:
* **App Name:** A clear, descriptive name. This is the first thing people see in the Marketplace.
* **Description:** What your app does and what kind of results it produces. Be specific - users search by keywords.
* **Cover Image:** Your app's thumbnail. Auto-generated from the output, or upload a custom one. A strong thumbnail is the single biggest driver of clicks in the Marketplace.
* **Cloning Permission:** Choose whether other users can clone (duplicate) your app. Cloning creates an independent copy they can modify without affecting your original. You earn credits from every clone.
* **Credit Pricing** *(Community only):* Set an additional credit charge on top of the base generation cost. This is your earning per use.
| Earning Tier | Credits You Earn Per Use |
| ------------ | ------------------------ |
| None | 0 |
| Low | 10 |
| Medium | 20 |
| High | 30 |
#### 5. Submit
* **Private & Team** apps go live instantly.
* **Community** apps are submitted for moderation review. The team reviews for quality and safety, this usually takes a few hours. Once approved, your app appears in the Community Marketplace.
### Moderation (Community Only)
Apps published to the Community go through a review before going live. The moderation team checks for quality, safety, and adherence to community guidelines.
Private and Team apps skip moderation entirely and are published instantly.
### What's Next
After publishing, head to [Manage Your Apps →](/workflows/manage-apps) to track metrics, edit workflows, manage versions, and monitor your credit earnings.
# Run your first Workflow
Source: https://docs.imagine.art/workflows/run-your-first-workflow
Using Workflows is intuitive and powerful. If you're new to node-based AI design, here are some quick tips to get started.
## Getting Started with Workflows
Using Workflows is intuitive and powerful. If you're new to node-based AI design, here are three ways to get started:
Use pre-configured presets for common tasks without manual setup.
Browse curated workflows designed for specific use cases.
Create custom workflows with complete creative control.
## Option 1: Start with Presets on Canvas
Presets are the quickest way to get started without manually adding each node. These pre-configured presets cover common tasks and streamline your workflow.
You can use them as-is or modify them to match your unique use case. Presets help you leverage the power of ImagineArt Workflow capabilities without needing to know the technical details. Here's how you can use them:
#### **Step 1: Select a Preset**
On your blank canvas, you'll see placeholder options like **Get Prompt Ideas**, **Animate Image**, **Merge Styles**, **Edit Image**, etc. These presets are designed to handle specific use cases with minimal effort.
* **Click on a Preset**: When you click on any preset, it automatically adds the required nodes to the canvas, pre-configured and ready to go.

#### **Step 2: Customize the Preset**
Once the preset is loaded into the canvas, you can customize it to fit your needs:
* **Update the Prompt Node**: In case the preset requires a specific prompt (e.g., for animating an image or generating art), modify the prompt text within the **Prompt Node**.
* **Replace Input Image**: For image-related presets (like **Animate Image** or **Edit Image**), click on the **Input Image Node** and upload your image.
#### **Step 3: Run the Workflow**
Once you customize preset inputs:
* **Click the Run →** button in the node, or in the right panel under node settings.
* Alternatively, **select all nodes** and press **Ctrl/Cmd + Enter** to run the entire workflow.
Your content will automatically be generated in the corresponding image or video node, based on the preset you've chosen.
## Option 2: Start from Featured Workflows
The second fastest way to get started is using our curated workflows.
#### **Step 1: Choose a Workflow**
Browse featured workflows on the dashboard designed for common use cases:

**Featured Workflows:**
* **Ad Placement**
* **Product Photoshoot**
* **Photo Composer**
* **Camera Angle Ideation**
* **Virtual Try on**
#### **Step 2: Duplicate Workflow**
Click on any workflow to load and duplicate it. All nodes and models are pre-configured and ready to use.

You only need to:
1. **Update the prompt** in the `Prompt` node
2. **Or upload/replace an image** in the `Import` node (if the workflow requires input)
#### **Step 3: Run the Workflow**
If everything is loaded correctly, click the **Run →** button in the node or in the right panel of node settings, or select all and press `Ctrl/Cmd + Enter`.
Your generated content will appear in the image/video node on the canvas, depending on the workflow you're running.
## Option 3: Build From Scratch
For complete creative control, start with a blank canvas.
#### **Step 1: Open Blank Canvas**
* Click on **Create new workflow** in the dashboard
#### **Step 2: Add Your First Node**
* Press `double-click` anywhere on the canvas.
.png?alt=media\&token=d92f6657-563f-4e1f-8410-3e1e6024f2a1)
* Alternatively, click the **+** button in the left panel to add the required nodes (Text, Prompt, Image, or Video).
.png?alt=media\&token=86f970be-78ba-47ba-9dcc-fcc4884e4212)
* For a quick search, press **Cmd + K** and search for any node (e.g., "Image", "Video", "Prompt"), or search by AI model name (e.g., "Veo 3", "Nano Banana").
#### **Step 3: Add Required Input Nodes**
1. Click on the visible node to add it to the canvas (e.g., **Text** or **Image** node). Write your prompt in it.
2. Connect the Text or Prompt node to the Image or Video node, depending on your setup.
#### **Step 4: Generate**
* Click **Run →** and watch your creation appear in real time.

# Understanding Nodes
Source: https://docs.imagine.art/workflows/understanding-nodes
Complete reference for every node type in Workflows—image, video, text, and utility nodes.
In the ImagineArt Workflows canvas, **nodes** are the fundamental building blocks of every pipeline. Every operation—importing a file, running an AI model, trimming a video, combining prompts—lives inside a node. Connect nodes together to form a visual data pipeline that automates your creative process.
## What is a node?
A node is a self-contained unit designed to perform one specific task. Each node has:
* **Inputs (left side):** Where the node receives data—text, images, video, or audio.
* **Outputs (right side):** Where the node sends its processed result to the next step.
* **Parameters:** Settings inside the node that let you fine-tune how it performs its task.
Data flows left to right across the canvas. Nodes are type-safe: a video output can only connect to a video input. When you trigger a workflow, data travels through the chain you've built—a text prompt might become an image, which flows into a background remover, and finally into an export node.
Nodes fall into four categories: **Generational** (AI-powered creators), **Transformation** (specialists that modify content), **Utility** (organizers that route and adjust data), and **Essential** (import/export gates).
Image nodes cover the full image creation pipeline, from generating visuals from scratch to fine-tuning lighting, resizing for multiple platforms, and batch-processing asset sets.
**To add an image node:** Click **+** in the left toolbar, select **Image**, then choose the node. You can also double-click the canvas and search by name.
### Generate Image
**Purpose:** Create high-quality images from text prompts (text-to-image) or transform existing images using a reference and a description (image-to-image).
The Generate Image node supports a wide range of AI models, each with its own style characteristics. Key parameters:
| Setting | Type | Effect |
| ---------------- | -------- | ------------------------------------------------------------ |
| Prompt | Text | Describes the visual elements the AI generates |
| Style | Preset | Applies a stylistic preset (3D, Cyberpunk, Watercolor, etc.) |
| Strength | 0–100% | Controls how strongly the prompt influences the output |
| Image Size | Dropdown | Sets the output aspect ratio (16:9, 4:3, 1:1, etc.) |
| Seed | Number | Locks the result for reproducible outputs |
| Guidance Scale | 0–20% | Balances prompt adherence against creative randomness |
| Resolution | Dropdown | Sets output resolution: High (4K), Medium (2K), or Low (1K) |
| Prompt Optimizer | Button | Enhances or rewrites your prompt for better results |
**Inputs:** Text prompt (from a Prompt or AI Copilot node), optional reference image. **Output:** Generated image.
### Edit Image
**Purpose:** Generate new images inspired by one or more reference images. Rather than modifying an image directly, this node uses your uploaded references as creative inspiration—combining styles, concepts, and visual elements into a fresh output.
Models like **Nano Banana Pro**, **Nano Banana 2**, and **Seedream V4.5** support uploading 10+ reference images, making this node useful for complex compositions that blend multiple visual ideas.
**Common use cases:** Fashion lookbooks, virtual try-ons, 3D art from flat designs, vintage poster variations, mythical creatures in real-world settings.
**Inputs:** One or more reference images, text prompt describing the desired output. **Output:** Newly generated image based on the references and prompt.
### Upscale Image
**Purpose:** Increase the resolution and sharpness of any image using AI. Ideal for preparing low-resolution assets for print, large displays, or professional presentations.
| Setting | Type | Effect |
| ---------------- | --------------- | ------------------------------------------------------ |
| Upscaling Factor | 2×, 4×, 8×, 16× | Higher values increase resolution and detail |
| Strength | 0–100% | Controls the AI's impact on fine details and sharpness |
| Noise Reduction | Slider | Smooths the image; high values may remove fine details |
| Sharpness | 0–100% | Increases edge definition; too high may over-sharpen |
| Detail Level | 0–100% | Controls how much detail is added during upscaling |
| Seed | Number | Ensures reproducible results |
**Input:** Image. **Output:** Upscaled image at the specified factor.
### Multiple Camera Angles
**Purpose:** Generate new perspective views of a subject from a single reference image. Rotate, tilt, and zoom a virtual camera using an interactive 3D controller to produce angles that weren't in the original shot.
| Setting | Type | Effect |
| --------------------- | --------------------- | ----------------------------------------------------------- |
| Camera Controller | 3D widget | Drag to orbit the camera around your subject |
| Rotation (Left/Right) | Slider | Horizontal rotation; negative = left, positive = right |
| Move (Up/Down) | Slider | Vertical camera position |
| Zoom | Slider (default: 5) | Camera distance; higher = closer |
| Wide Angle Lens | Checkbox | Simulates a wide-angle lens for more immersive perspectives |
| Guidance Scale | Slider (default: 4.5) | Controls adherence to input and camera settings |
| Seed | Number | Ensures reproducible results |
**Input:** Reference image. **Output:** New image from the configured camera angle.
### Relight
**Purpose:** Reposition, recolor, and adjust the intensity of the light source on any image. Use the interactive 3D light controller or quick-position presets (Top, Front, Back, Bottom, Left, Right) to transform the mood and depth of a scene without regenerating it.
| Setting | Type | Effect |
| ---------------- | ------------------------------- | ------------------------------------------------------------------------ |
| Light Controller | 3D widget | Drag to reposition the light source |
| Quick Positions | Buttons | Snap to a preset direction in one click |
| Reset | Button | Returns the light to its default position |
| Light Intensity | Slider (default: 7) | Controls brightness; higher = stronger, more pronounced lighting |
| Color | Color picker (default: #ffffff) | Sets the light's color; use warm tones for golden-hour, cool for moonlit |
| Aspect Ratio | Dropdown | Defines the output image dimensions |
| Resolution | Dropdown (1k, 2k, 4k) | Determines output resolution |
**Input:** Reference image. **Output:** Relit version of the image.
### AI Resize
**Purpose:** Intelligently resize images to different aspect ratios without awkward cropping or stretching. The AI adapts the composition—extending backgrounds and repositioning elements—to fit the new ratio.
Unlike the standard Resize utility node (which scales dimensions directly), AI Resize understands the scene and fills in new areas coherently.
| Setting | Type | Effect |
| ------------ | ------------------------------------- | --------------------------------------- |
| Aspect Ratio | Dropdown (1:1, 16:9, 4:3, 9:16, etc.) | Target dimensions for the resized image |
**Common use cases:** Multi-platform ad campaigns (1:1 for Instagram, 9:16 for Stories, 16:9 for YouTube), poster format variations, e-commerce banner adaptation.
**Input:** Image. **Output:** Image resized to the target aspect ratio.
### Split Image
**Purpose:** Divide a single image into a grid of smaller cropped sections. A visual preview shows exactly where the cuts will be made before you run the node.
| Setting | Type | Effect |
| --------- | -------------------------------------------- | ---------------------------------------------------- |
| Grid Size | Dropdown (2×2, 3×2, 3×3, 4×2, 4×3, 4×4, 5×5) | Number of output sections (2×2 = 4 images; 5×5 = 25) |
**Common use cases:** Breaking storyboards into individual frames, extracting panels from comic pages, dividing mood boards into separate reference images.
**Input:** Image. **Output:** Multiple cropped section images, each passed individually downstream.
### Image Iterator
**Purpose:** Feed multiple images into a workflow and process each one individually through the connected downstream nodes. Instead of running your workflow manually for every image, the iterator handles the batch automatically.
**Common use cases:** Batch style transfer across product photos, bulk ad variations with consistent branding, multi-image upscaling in a single run.
**Inputs:** Multiple images (from Import, generated images, or added directly in the node). **Output:** Each image passed one by one to the next connected node, generating a separate output per input.
### Common image node combinations
| Combination | Use case |
| ------------------------------------ | --------------------------------------------------------------------- |
| Generate → Upscale | Create an image, then enhance resolution for print |
| Generate → Multiple Camera Angles | Produce a base image, then generate product or character perspectives |
| Import → Image Iterator → Edit Image | Batch-process reference photos through the same style prompt |
| Generate → AI Resize | Create a hero visual, then resize for every ad platform |
| Generate → Relight | Produce a scene, then experiment with different lighting setups |
| Split Image → Upscale | Break a storyboard into frames, then upscale each one |
Video nodes cover the full video production pipeline—generating clips from text or images, enhancing quality, synchronizing audio, transferring motion, and assembling final sequences.
**To add a video node:** Click **+** in the left toolbar, select **Video**, then choose the node. You can also double-click the canvas and search by name.
### Generate Video
**Purpose:** The primary node for creating videos. Supports both **text-to-video** (describe a scene) and **image-to-video** (provide a still image and describe the motion). Supports a wide range of AI models.
| Setting | Type | Effect |
| -------------- | ------------------------------- | ---------------------------------------------------------------------------------------- |
| Model | Dropdown | Selects the AI model; each has unique motion quality, realism, and speed characteristics |
| Duration | Dropdown (e.g., 5s, 10s) | Length of the generated video |
| Resolution | Dropdown (720p, 1080p) | Output video resolution |
| Aspect Ratio | Dropdown (16:9, 9:16, 1:1, 4:3) | Frame dimensions |
| Generate Audio | Checkbox | When enabled, generates a matching audio track (select models) |
| Camera Fixed | Checkbox | Locks the camera; no camera movement in the output |
| Seed | Number | Produces reproducible results with the same seed and settings |
**Input modes:**
* **Text-to-video:** Connect a Prompt node and describe the scene. The more specific your description, the closer the output matches your vision.
* **Image-to-video:** Connect an image as the first frame and describe the desired motion.
**Output:** Generated video clip.
### Edit Video
**Purpose:** Transform existing videos using reference clips and prompts. Describe the changes you want—style shifts, scene alterations, environment swaps, visual effects—and the AI generates a new version of the video.
**Common use cases:** Style transfer (daytime scene → noir night sequence), environment swaps (indoor → futuristic cityscape), brand consistency across clips, adding atmospheric effects like rain or fog.
**Inputs:** Reference video, text prompt describing the changes. **Output:** Transformed video.
### Extend Video
**Purpose:** Generate additional footage that continues seamlessly from where your original video ends. The AI understands the scene's context, motion, and visual style—then creates new frames that naturally extend the narrative.
| Setting | Type | Effect |
| -------------- | -------- | ----------------------------------------------------------------- |
| Model | Dropdown | Affects continuity quality and visual consistency with the source |
| Duration | Dropdown | Amount of additional footage to generate |
| Resolution | Dropdown | Match to your source clip for seamless continuity |
| Aspect Ratio | Dropdown | Should match the original video |
| Generate Audio | Checkbox | Generates matching audio for the extended portion |
| Seed | Number | Ensures reproducible results |
You can optionally add a prompt to guide what happens next. Without a prompt, the AI continues the scene naturally.
**Input:** Video clip to extend. **Output:** Extended video.
### Lipsync
**Purpose:** Generate a talking video from a single character image. Provide a character image and either write a prompt (the AI generates speech and lip movements) or connect your own audio track (the AI syncs lip movements to that audio).
| Setting | Type | Effect |
| -------------- | -------- | --------------------------------------------------------------------------------- |
| Model | Dropdown | Controls speech realism, facial expression quality, and native audio capabilities |
| Duration | Dropdown | Length of the generated video |
| Resolution | Dropdown | Output resolution; higher values produce sharper facial details |
| Aspect Ratio | Dropdown | Frame dimensions |
| Generate Audio | Checkbox | When enabled, generates native spoken audio from your prompt |
| Seed | Number | Reproducible results |
**Common use cases:** AI influencer content, personalized video outreach at scale, UGC-style ads without on-camera talent.
**Inputs:** Character image (portrait or headshot), text prompt or audio file. **Output:** Video of the character speaking with synchronized lip movements.
### Motion Transfer
**Purpose:** Capture motion from a reference video and apply it to a character image. Upload a reference video of someone dancing, walking, or performing any action—the node transfers that exact movement onto your target character.
This node requires two inputs:
* **Character Image** — the subject you want to animate (clearly lit, full body or relevant body parts visible).
* **Reference Video** — the motion source (use clips with clear, well-defined movements; avoid heavy occlusion, rapid camera movement, or multiple people).
| Setting | Type | Effect |
| --------------- | ----------------------------- | --------------------------------------------------------------------------------- |
| Model | Dropdown (e.g., Wan 2.2 Move) | Affects how well body types, motion complexity, and rendering quality are handled |
| Guidance Scale | Slider (default: 1) | Lower = more creative freedom; higher = stricter match to source motion |
| Resolution | Dropdown | Higher captures finer detail but takes longer |
| Inference Steps | Slider (default: 20) | More steps = smoother, higher-quality output at the cost of generation time |
| Video Quality | Dropdown (High/Medium/Low) | Overall rendering quality |
| Seed | Number | Reproducible results |
**Common use cases:** Viral dance videos, animated character performances, virtual try-on with movement.
**Inputs:** Character image, reference video. **Output:** Video of the character performing the motion from the reference.
### Upscale Video
**Purpose:** Increase the resolution and visual quality of an existing video using AI. Sharpens details, improves clarity, and enhances fidelity while preserving natural motion and smoothness.
| Setting | Type | Effect |
| ---------------- | ----------------- | ------------------------------------------------------------- |
| Model | Dropdown | Balances speed, detail preservation, and output quality |
| Upscaling Factor | Dropdown (2×, 4×) | Higher = larger, more detailed output; longer processing time |
| Seed | Number | Reproducible results |
**Common use cases:** Polishing AI-generated clips from 720p to 1080p/4K, restoring low-resolution footage, preparing content for large displays.
**Input:** Video clip. **Output:** Upscaled video.
### Video Trimmer
**Purpose:** Cut a video down to a specific portion using a visual timeline editor. Preview the clip directly inside the node, drag the trim handles to set start and end points, and output only the segment you need.
| Control | Description |
| -------------- | ----------------------------------------------------------------------------- |
| Preview Player | Play back the video inside the node; includes play/pause and audio mute |
| Timeline Strip | Frame-by-frame filmstrip; drag start and end handles to define the trim range |
| Timecodes | Displays start time, current playhead, and end time of the selected range |
| Trim Presets | Quick-action buttons: trim start, trim both ends, or trim end |
| Undo / Redo | Step back or forward through trim adjustments |
**Input:** Video clip. **Output:** Trimmed video segment.
### Combine Videos & Audios
This node category covers two distinct operations:
**Combine Audio & Video** — Merges a video input and an audio input into a single video file with the audio track embedded. Both inputs are required. Useful for pairing a generated video with a voiceover, music track, or sound effects.
**Combine Videos** — Joins multiple video clips sequentially into a single continuous video. Clips play in the order they are connected.
**Inputs for Combine Audio & Video:** Video (green handle), Audio (pink handle). **Inputs for Combine Videos:** Multiple video clips. **Output:** Single merged video file.
### Extract Video Frame
**Purpose:** Capture a single still frame from a video and output it as an image. Specify the exact moment using a frame number or timecode. This is the bridge between your video and image workflows.
| Setting | Type | Effect |
| -------- | ------------------------ | -------------------------------------------------------- |
| Frame | Number (default: 1) | Specifies which frame to extract by sequence number |
| Timecode | Time (default: 00:00:00) | Specifies which frame to extract by timestamp (HH:MM:SS) |
**Common use case:** Pull a key frame from an existing video, edit it with Generate Image or Edit Image, then regenerate a new video from the updated frame.
**Input:** Video clip. **Output:** Still image extracted from the specified frame.
### Common video node combinations
| Combination | Use case |
| ----------------------------------------------------- | --------------------------------------------------------------------- |
| Generate Video → Upscale Video | Create a clip, then enhance for professional delivery |
| Generate Video → Extend | Produce a short clip, then extend to the desired length |
| Image → Generate Video → Lipsync | Animate a portrait, then sync lip movements to a voiceover |
| Generate Video → Motion Transfer | Generate a base character video, then apply motion from a reference |
| Generate Video → Video Trimmer → Combine Videos | Create multiple clips, trim each, then assemble into a final sequence |
| Extract Video Frame → Generate Image → Generate Video | Pull a frame, enhance the still, then regenerate a new video |
Text nodes create, refine, and manipulate text—from writing prompts that drive AI generation to combining or splitting content for complex workflows.
**To add a text node:** Click **+** in the left toolbar, select the node from the Text Utilities category. You can also use keyboard shortcuts on the canvas.
### Prompt
**Purpose:** Write custom instructions to guide AI models in generating text, images, or videos. This is the most common input node in any workflow—it holds the text description that downstream nodes act on.
**Key features:**
* Works with all generation types: text, image, and video.
* Combine with multiple generation nodes to reuse the same prompt across different models.
* Chain with AI Copilot to first generate or refine text, then pass it downstream.
**Keyboard shortcut:** `P`
**Input:** None (you type the prompt directly into the node). **Output:** Text string.
**Example prompts:**
* *Text:* "Write a catchy caption for a social media ad about our new eco-friendly sneakers."
* *Image:* "A beach sunset with a silhouette of a person practicing yoga, golden hour light."
* *Video:* "A slow-motion close-up of coffee being poured into a ceramic mug, steam rising."
### AI Copilot
**Purpose:** Generate and analyze text using large language models (LLMs). The AI Copilot node can process text, image, and video inputs to produce high-quality written outputs—scripts, descriptions, analyses, creative pieces, ad copy, and more.
**Key features:**
* **Advanced text processing:** Leverage top LLMs to process and refine any text input.
* **Contextual understanding:** Extracts context, personas, and identifiers from inputs, including visual content.
* **Multi-input integration:** Combine text, images, and videos as inputs for more informed outputs.
* **High-quality output:** Rich, contextually accurate text for scripts, stories, campaigns, and workflows.
**Keyboard shortcut:** `T`
**Inputs:** Text, images, or videos (any combination). **Output:** Generated or analyzed text.
**Example use cases:**
* Write a 5-minute YouTube script from a topic prompt.
* Analyze a video ad and generate consistent marketing copy that matches its tone.
* Create 5 ad copy variations for a Google campaign.
* Write a brand story for a product launch.
### Combine Text
**Purpose:** Merge multiple text inputs into a single, unified output. Ideal for assembling a complete document from separate parts—sections of a script, product features, research notes—before passing the combined text to a generation node.
**How to use:**
1. Connect multiple text inputs (from Prompt nodes, AI Copilot outputs, or other text sources).
2. The node automatically merges them into one cohesive block.
**Common use cases:**
* Assembling a full blog post from separate research, quotes, and original writing.
* Merging intro, body, and call-to-action sections into a complete video script.
* Combining product feature bullets into a single e-commerce description.
**Inputs:** Two or more text streams. **Output:** Single merged text string.
### Split Text
**Purpose:** Break down long or complex text into smaller, manageable sections. Useful for workflows that need to process or generate content scene-by-scene or section-by-section.
**Split criteria options:** By sentence, by paragraph, or by a custom delimiter (commas, keywords, etc.).
**Common use cases:**
* Breaking a video script into individual scenes, then generating a video clip per scene.
* Dividing a voiceover script into sections for audio generation and video syncing.
* Segmenting descriptive text for targeted image generation per section.
**Input:** Text string. **Output:** Multiple text segments, each passed individually to downstream nodes.
Utility nodes handle the connective work of any workflow—importing assets, exporting results, previewing outputs, routing data, and applying non-AI visual adjustments like cropping, resizing, and color correction.
**To add a utility node:** Click **+** in the left toolbar, select **Utilities**, then choose the node. You can also double-click the canvas and search by name.
### Essentials
**Import** Upload images, videos, or audio files directly into your workflow. This is the starting point for any workflow that uses your own assets rather than AI-generated content.
* **Input:** None (you upload a file directly).
* **Output:** Image, video, or audio file ready to connect to downstream nodes.
**Export** Save your workflow's final output to a specific format or destination. Use this as the last node in a workflow to download or publish your finished content.
* **Input:** Image, video, or audio.
* **Output:** Downloaded/exported file.
**Preview** View the output of any node at any point in your workflow without exporting. Useful for checking intermediate results, debugging, or comparing outputs between pipeline stages.
* **Input:** Image, video, or text.
* **Output:** Visual display inside the node (no data passed downstream).
**Router** Direct content to different paths within your workflow. Use the Router to split a workflow into multiple branches—sending the same input to different nodes for parallel processing, or selecting which path to follow based on your setup.
* **Input:** Image, video, or text.
* **Output:** Same data routed to one or more downstream branches.
### Adjustment nodes
These nodes make quick, non-AI visual modifications to images and videos.
**Crop** Trim an image or video frame to focus on a specific area. Remove unwanted edges, center a subject, or reframe a composition before passing it downstream.
**Resize** Change the dimensions of an image or video by a specific pixel size or scale factor. Unlike [AI Resize](/workflows/understanding-nodes#ai-resize) (which intelligently adapts the composition), this is a straightforward scale operation—useful for meeting specific pixel requirements or optimizing file sizes.
**Blur** Apply a blur effect to an image or video, uniformly or to specific areas. Great for softening backgrounds, creating depth-of-field effects, or obscuring sensitive information.
**Levels** Adjust brightness, contrast, and tonal range. Fine-tune shadows, midtones, and highlights to correct exposure, enhance detail, or give content a specific look.
**Filters** Apply visual filters that change color tone, texture, or overall aesthetic. Options include vintage, high-contrast, warm, cool, and artistic effects for quickly setting the mood of your content.
### Common utility node combinations
| Combination | Use case |
| ------------------------------------------ | ----------------------------------------------------------------------- |
| Import → Crop → Generate Image | Upload a reference photo, crop to the relevant area, use it as AI input |
| Generate Video → Levels → Filters → Export | Create a video, correct exposure, apply color grade, export |
| Import → Router → (Multiple Branches) | Route a single asset to different processing paths simultaneously |
| Generate Image → Resize → Preview | Create an image, resize to platform specs, preview before exporting |
| Import → Blur → Crop → Export | Upload a photo, blur the background, reframe, and export |
Audio Nodes in ImagineArt Workflows enable you to generate professional spoken voice, original music, and custom sound effects entirely from text. Whether you need a voiceover for video, a background track, or sound effects for animation, these AI-powered nodes produce broadcast-quality audio that integrates seamlessly into your creative pipeline.
**To add an Audio node:** Click **+** in the left toolbar, select **Audio**, then choose the node. You can also double-click the canvas and search by name.
# Utilities
Source: https://docs.imagine.art/workflows/utilities
## Summary
The Utility category provides the foundational nodes that handle importing, exporting, previewing, routing, and making quick adjustments to your content. These nodes don't generate or transform content with AI; instead, they manage the flow of data through your workflow and apply straightforward visual adjustments like cropping, resizing, blurring, and color correction.
Think of them as the connective tissue of any workflow—they keep everything organized, properly sized, and moving in the right direction.
## How to Add Utility Nodes
1. Click the Add (+) button on the left toolbar in the workflow canvas.
2. Select Utilities (Text/Image/Video) under node categories.
3. Choose from the available nodes listed below.
You can also double-click anywhere on the canvas and search for any utility node by name.
## Essentials
These nodes bring content into your workflow and send it out.
Upload images, videos, or audio files directly into your workflow. This is the starting point for any workflow that uses your own assets rather than AI-generated content.
Save your workflow's final output to a specific format or destination. Use this as the last node in a workflow to download or publish your finished content.
View the output of any node at any point in your workflow without exporting. Useful for checking intermediate results, debugging, or comparing outputs between different stages of a pipeline.
Direct content to different paths within your workflow based on your setup. Use the Router to split a workflow into multiple branches—sending the same input to different nodes for parallel processing, or selecting which path to follow.
## Adjustment Nodes
These nodes make quick, non-AI visual modifications to images and videos.
Trim an image or video frame to focus on a specific area. Remove unwanted edges, center a subject, or reframe a composition before passing it downstream.
Change the dimensions of an image or video. Unlike AI Resize (which intelligently adapts the composition), this is a straightforward scale—useful for meeting specific pixel requirements or optimizing file sizes.
Apply a blur effect to an image or video, either uniformly or to specific areas. Great for softening backgrounds, creating depth-of-field effects, or obscuring sensitive information.
Adjust brightness, contrast, and tonal range. Fine-tune shadows, midtones, and highlights to correct exposure, enhance detail, or give your content a specific look.
Apply visual filters that change the color tone, texture, or overall aesthetic. Includes options like vintage, high-contrast, warm, cool, and artistic effects to quickly set the mood of your content.
## Combining Utility Nodes
Utility nodes are designed to slot into any workflow. Here are some common patterns:
* **Import → Crop → Generate Image** – Upload a reference photo, crop to the relevant area, then use it as input for AI generation.
* **Generate Video → Levels → Filters → Export** – Create a video, correct the exposure, apply a color grade, and export the final file.
* **Import → Router → (Multiple Branches)** – Bring in a single asset and route it to different processing paths simultaneously—one for image generation, another for video, another for upscaling.
* **Generate Image → Resize → Preview** – Create an image, resize it to platform specs, and preview the result before exporting.
* **Import → Blur → Crop → Export** – Upload a photo, blur the background, crop to frame, and export the final version.
# What is ImagineBusiness?
Source: https://docs.imagine.art/workflows/what-is-imaginebusiness
Workflows, built for teams — brand governance, shared assets, and pre-built templates layered on top of the same node-based canvas.
**ImagineBusiness** is a team workspace built on top of Workflows' node-based canvas — same underlying builder, with a governance and collaboration layer for teams working together on brand-consistent content.
## Switching into it
ImagineBusiness isn't a separate URL you navigate to directly — it's a **workspace** you switch into from the workspace switcher (top-left, where your personal workspace normally shows). Once switched, every page lives under a distinct `/enterprise/*` area of the app.
The sidebar adds several sections not present in personal Workflows:
* **Home** — a dashboard with an AI workflow generator and your team's recent projects. See below.
* **Projects** — the same node-based workflow canvas you'd use in regular Workflows.
* **Templates** — a searchable gallery of pre-built workflows to start from instead of building a canvas from scratch.
* **Assets** — a shared team asset library.
* **Apps** — Team Apps, My Apps, and Community Apps: ready-to-run apps powered by your workflows.
* **Brand Hub** — Brand Kits and Moodboards, covered on their own page: [The Brand Hub](/workflows/imaginebusiness-brand-hub).
* **Configure → Integrations** — connect third-party tools.
ImagineBusiness is a paid tier — "Contact Sales" and "Upgrade" prompts appear throughout, and clicking **Upgrade** opens a real pricing modal (a Teams Scale tier at \$350/mo alongside a custom Enterprise plan). Core functionality (Projects, Templates, Assets, Brand Hub) is fully usable once you're on it; a handful of integrations are still "Coming Soon."
## Home
The **Home** page is the landing screen when you switch into ImagineBusiness — a prompt box ("Turn ideas into AI workflows... Describe what you want to automate...") with suggestion chips (for example, "On-brand social carousel," "Product mockup for Etsy," "UGC-style ad video") for generating a starter workflow from a plain-language description, plus a **Recent Projects** grid below it.
## The top bar
Every ImagineBusiness page shares the same top bar, none of it specific to any one section:
| Element | What it does |
| ------------------------------------- | ----------------------------------------------------------------------------------------------------------------------------------------- |
| Scope dropdown ("All team creations") | Filters what you see across the workspace |
| Team switcher pill | Shows member roles and seat counts, with **Invite** and any pending **New joiner requests** — real team/role management, not just a label |
| Grid icon | Opens a **Featured tools** quick-launch panel (Image, Video, Audio, Workflow, Edit, Upscale, Assist, AI Docs, Slides) |
| Account menu | Total Credits balance, Buy Credits, Usage Analytics, Appearance, Billing & Subscription, Settings, Logout |
## Templates
Click **Templates** for a searchable gallery of pre-built workflows. Each template is a real, ready-to-run workflow — not just an example.
The category filter chips are **not fixed** — confirmed live across two sessions on the same team, the set changed from Campaigns/Cinematic/Advertising/Fashion/Branding/Editing to All/Cinematic/Advertising/**Fashion & Apparel**/Branding/**FMCG**/**Fast Food**/Editing. Treat any specific category list (including this page's) as illustrative, not exhaustive — expect it to drift.
## Assets
Click **Assets** for the team's shared library. Beyond a simple grid, assets carry real status tracking — **All / To Do / In Progress / Under Review / Done** — and are organized into **Team folders** (shared, e.g. "All public creations") and **Private folders** you create yourself. Saved views for **Liked**, **Uploads**, and **Recently Deleted** sit alongside the status pipeline, plus a multi-select **Filters** accordion, an Upload button, a grid/list toggle, and a zoom slider.
Use the status filters as a lightweight review pipeline — move an asset from **To Do** to **Under Review** once a draft is ready, rather than relying on a separate project-management tool for creative approvals.
## Connecting your tools
Under **Configure → Integrations**, connect third-party tools your team already uses.
**Google Drive** is connectable today. Slack, Facebook, Instagram, LinkedIn, and Google Sheets are all listed as **Coming Soon** — plan around Google Drive for now if you need a working integration.