# Ideart > Generate AI images and AI videos from one prompt box. Browse tens of thousands of free AI prompts, remix a proven example, and create your own in seconds. ## Pages - [Home](https://ideart.ai/): AI image and video creation agent - [AI Tools](https://ideart.ai/tools): Focused AI creation tools - [AI Video Effects](https://ideart.ai/ai-video-effects): Preset video effect templates - [Pricing](https://ideart.ai/pricing): Pricing plans - [Privacy Policy](https://ideart.ai/privacy-policy): How personal information is handled - [Refund Policy](https://ideart.ai/refund-policy): Cancellations, refunds, and billing errors - [Terms of Service](https://ideart.ai/terms-of-service): Terms for using the service - [Contact](https://ideart.ai/contact): Support, billing, and partnership enquiries - [AI Image Generator](https://ideart.ai/ai-image-generator): Generate polished AI images from prompts and continue organizing outputs in the Ideart workspace. - [Image to Video](https://ideart.ai/image-to-video): Turn static images into cinematic AI videos with motion, camera direction, and optional prompts. - [Text to Video](https://ideart.ai/text-to-video): Turn written prompts into short AI videos with scene direction, motion, and cinematic pacing. - [Image to Image](https://ideart.ai/image-to-image): Transform reference images into new styles, compositions, and production-ready visual variations. - [Consistent Character](https://ideart.ai/consistent-character-video): Create AI videos where a character keeps a recognizable look across motion, shots, and prompt variations. - [Video to Video](https://ideart.ai/video-to-video): Restyle, transform, and iterate on existing videos with AI-powered video-to-video workflows. - [Video Creator](https://ideart.ai/video-creator): Create short AI videos from prompts, images, and creative direction in an Ideart workflow. - [Sexy Me](https://ideart.ai/ai-video-effects/sexy_me): Create a stylized short character video from one portrait reference with the Sexy Me template on Ideart. - [French Kiss](https://ideart.ai/ai-video-effects/french_kiss_8s): Create a short romantic AI video concept from a two-person reference with the French Kiss template on Ideart. - [Image to Prompt Generator](https://ideart.ai/image-to-prompt): Read the prompt out of any picture, then run it in the Ideart image and video tools. ## Blog Posts ### What Is an AI Video Agent and How Does It Work? URL: https://ideart.ai/blog/ai-image-video-agent Description: An AI video agent turns a plain-language brief into a finished clip: what it decides for you, what it will not do, and how to brief one so take one lands. An AI video agent is a system you brief in plain language that then makes the production decisions itself: whether the job is an image or a video, whether it is a generation from text or an animation of a picture you attached, and how long, how large and what shape the output should be. You describe the result; it works out the settings. That is the difference from a generator, and it changes what a good request looks like. This article covers what an AI video agent decides for you, what it will not do, how it handles a follow-up edit, and how to brief one so the first take is close. ## What is an AI video agent? An AI video agent sits between your description and the generation models. A plain generator gives you a form — prompt box, model picker, duration, resolution — and executes exactly what you set. An agent reads the brief and fills the form itself. In practice that means three jobs: 1. **Interpretation.** Turning "a fifteen-second launch clip for a skincare bottle, clean and clinical, for a website hero" into a concrete set of parameters. 2. **Routing.** Deciding which operation the request actually is — a generation from text, an animation of an attached still, or a pass that works from reference material — and, when the composer is left on auto, whether the answer is an image or a video at all. 3. **Continuity.** Keeping the brief, the attachments, the settings and the results in one thread, so the next request is understood as a change to the last one rather than a fresh start. The third job is the one that most changes how it feels to use. A generator treats every render as an isolated submission. An agent treats a session as one piece of work. ## AI video agent vs AI video generator | | AI video generator | AI video agent | | ------------------------- | ------------------------------------ | ---------------------------------------- | | **You supply** | Prompt plus every setting | A brief, plus the model | | **Duration, size, shape** | Yours to set | Read from the brief, inside model limits | | **Which operation runs** | Whichever form you opened | Read from the prompt and attachments | | **Follow-up edit** | Fill the form again | "Same shot, slower camera" | | **Best when** | You know the exact settings you want | You know the outcome, not the parameters | Neither is strictly better. If you already know you want six seconds at 720p, setting it yourself is faster than describing it. If you know you want a product reveal that looks expensive and have no view on duration or framing, the agent gets there in fewer steps. ## What an AI video agent decides for you ### The limits of the model you picked You still pick the model — in the composer, or by opening a model page that locks the workspace to one. What the agent does is stay inside that model's limits, and they differ enough to matter: | Model | Clip length | Resolutions | Audio | | ------------ | ----------- | ----------------- | ----- | | MiniMax H3 | 5–15s | 768P, 2K | No | | Seedance 2.5 | 4–30s | 480p, 720p | Yes | | Seedance 2.0 | 4–15s | 480p, 720p, 1080p | Yes | Ask for a twenty-second clip with sound and only one of those models can serve it. An AI video agent clamps a request to the selected model's bounds instead of failing at the provider, which is why the model list is published rather than hidden — you can see the full specification for each on the [AI generation models](/models) pages, and pick accordingly before you brief anything. ### The output settings Duration, resolution and aspect ratio all cost money in different ways. Video is billed per second of output, so length is the biggest lever, and resolution multiplies the rate on top. Aspect ratio is the one with an auto setting; duration and resolution always carry a concrete value. An AI video agent can override any of the three from the destination you named — a 9:16 story, a 16:9 hero banner — instead of leaving you to guess and crop afterwards. ### The operation Attaching a picture changes the job. With no attachment the request is a text-to-video generation; with a still attached it becomes an animation of that still, which the [AI image to video generator](/image-to-video) also exposes directly if you would rather drive it yourself. The same goes for the [AI text to video generator](/text-to-video) when there is nothing to attach. ## How an AI video agent handles a follow-up edit This is where the conversation earns its keep. After the first take you do not restate the brief — you name the one thing to change: > Keep the product and the lighting. Halve the camera push-in, calm the background, and end on a centred pack shot. The previous prompt, references and settings are still in scope, so the request is understood as a delta. That is both easier to write and easier to reason about than editing a form, because you can see exactly which change produced which result. The discipline that makes it work: **one or two changes per turn.** Three simultaneous adjustments and you cannot tell which one fixed the shot and which one broke it. ## What an AI video agent will not do Being clear about the ceiling saves credits: - **It cannot exceed a model's limits.** No twenty-second clip from a model capped at fifteen seconds, and no 4K video from a catalog that tops out at 2K. - **It never swaps your model.** If the configured provider cannot serve the model you selected, the request is refused rather than quietly rerouted to a different one. - **It does not bypass billing.** Every generation is priced from the validated settings and charged to the same credit balance before the render starts. - **It does not supply taste.** It will execute "cinematic and premium" as literally as it can. Naming a concrete reference, a camera move and a colour direction still beats mood words. - **It cannot rescue a vague brief.** Garbage in is still the governing rule; the agent just fills in fewer blanks badly than a form does. ## How to brief an AI video agent well A good brief answers five questions, in roughly this order: 1. **What is the subject?** The specific product, person or scene. 2. **What happens?** One clear action over the length of the clip. 3. **How does the camera behave?** One move, not three. 4. **What must not change?** Identity, logo, packaging, wardrobe. 5. **Where is it going?** Feed, story, hero banner, ad — this decides aspect ratio and length. Attach a reference only when it carries information the words cannot: a face, a product shape, an approved colourway. Then say what to take from it. A file attached without explanation is an instruction the agent has to guess at. Every signed-in account gets 100 free credits a day. The cheapest way to learn how an agent reads a brief is to run two versions of the same request and compare what it chose. ## AI video agent FAQ ### Is an AI video agent the same as an AI video generator? No. A generator executes the settings you supply. An AI video agent reads a brief, works out the operation and the output settings itself, and keeps the session's context so follow-up edits are understood as changes rather than new jobs. ### Do I still choose the model myself? Yes — the model is yours to pick in the composer, and a model page locks the workspace to one. What the agent decides is the operation and the output settings, and those stay visible and adjustable. ### Does using an agent cost more credits? No. Pricing comes from the model, resolution and clip length of the render that actually runs, not from how the request was written. The charge is taken when the render starts, and a failed render is refunded. ### What should I attach as a reference? Only what the words cannot carry: a face, a product shape, an approved layout or colourway. Then say what should be borrowed from it — identity, composition, or pacing. ## Ready to brief an AI video agent? Bring the outcome, not the parameters: what the shot is, what moves, where it will be published. The settings get chosen for you, the charge follows the model, resolution and length of the clip that actually runs, and the next turn is an edit rather than a restart. [Try the AI video creator →](/video-creator) ## Keep reading More on briefing an AI video agent and writing the prompts behind it: - [How to write image to video prompts](/blog/image-to-video-prompt-guide) - [Text to video vs image to video](/blog/text-to-video-vs-image-to-video) - [What is Ideart AI](/blog/what-is-ideart) --- ### How to Write Image to Video Prompts That Work URL: https://ideart.ai/blog/image-to-video-prompt-guide Description: A five-part structure for writing an image to video prompt that keeps the subject stable: what to lock, one action, one camera move, and how to iterate. An image to video prompt is not a second description of your picture. The image already fixes the subject, the composition, the colours and most of the light. What the model still needs from you is movement: what moves, how the camera behaves, and what has to stay exactly where it is. Get that split right and the same source frame stops drifting between takes. This guide gives you a five-part structure for an image to video prompt, a worked before-and-after example, the limits each model in the [AI image to video generator](/image-to-video) imposes on what you can ask for, and the five mistakes that waste the most credits. ## What an image to video prompt actually controls Your source image is the strongest instruction in the whole request. It carries the subject, the framing, the palette and the lighting, and the model treats it as the opening state of the clip. Re-describing all of that in the prompt does not reinforce it — it competes with it, because the words and the pixels are rarely a perfect match. An image to video prompt controls the parts the image cannot express: - **Motion** — what the subject does over the next few seconds. - **Camera** — whether the frame pushes in, pulls back, tracks or holds. - **Secondary movement** — hair, fabric, steam, water, background traffic. - **Pacing** — slow and deliberate, or quick and energetic. - **The final frame** — where the shot needs to land. Everything else is already decided. A prompt that spends its first two sentences on "a woman in a red coat standing on a wet street at night" is spending its budget on facts the model can already see. ## The five-part image to video prompt structure Write every image to video prompt in the same order. The order matters: models weight the front of a prompt more heavily, so what must not change goes first. ### 1. Lock what must not change Name the elements whose identity is load-bearing before you ask for any motion. This is the single highest-leverage sentence in an image to video prompt. Things worth locking: - facial identity, hairstyle, skin tone - product geometry, label text, logo placement - wardrobe and accessories - background architecture and signage - packaging layout Write it plainly: _Preserve the bottle shape, the label text and the studio lighting._ ### 2. Name one subject action One action. Not three. A clip of four to six seconds cannot show a person turn, walk, sit and smile without rushing all four, and rushed motion is where faces and hands break down. Use verbs the model can act on — turning, walking, blinking, pouring, rotating, stepping forward — rather than mood words like _dynamic_ or _lively_, which describe how the result should feel rather than what should happen. ### 3. Choose one camera move Pick a single camera intention and let it run for the whole clip: - slow push-in - gentle pull-back - left-to-right tracking shot - subtle handheld drift - locked camera, subject motion only If your source image is already well composed, a locked camera or a small push-in protects that composition. A full orbit throws it away and asks the model to invent three-quarters of a scene it has never seen. ### 4. Add one layer of background motion Secondary movement is what separates a video from a slowly zooming photo. Keep it subordinate to the main action: - hair or fabric moving in a light breeze - steam, mist, dust, rain or falling leaves - a travelling highlight across glass or metal - out-of-focus pedestrians or traffic - water, curtains, or drifting particles One layer is usually enough. Two competing layers of atmosphere pull attention away from the subject. ### 5. Say how the shot ends Most image to video prompts stop describing the shot halfway through, and the last second is where a clip either lands or falls apart. Say where it should settle: _End on a centred hero frame with the label fully readable_, or _Hold the final pose for the last half second._ ## A worked example: weak prompt to strong Here is the same job written twice, against a studio product photo. **Weak:** > A beautiful cinematic video of a luxury serum bottle on a reflective surface, elegant lighting, premium feel, high quality, 4k, amazing detail. Nothing in that prompt is actionable. Every adjective describes the still image the model already has, and not one word says what should move. **Strong:** > Preserve the bottle shape, the label text and the studio lighting. The bottle rotates slowly by about fifteen degrees while a narrow highlight travels across the glass. The camera makes a gentle push-in. Fine mist drifts across the background. End on a centred hero frame with the label fully readable. The second version is barely longer. It is better because every clause is a decision the model has to execute, and because a reviewer can check each one against the result. ## Match the image to video prompt to the model Asking for motion the selected model cannot produce is the most common reason a good prompt returns a bad clip. Ideart publishes three video models, and each of them animates a still image under different limits: | Model | Clip length | Resolutions | Audio | Images accepted | | ------------ | ----------- | ----------------- | ----- | --------------- | | MiniMax H3 | 5–15s | 768P, 2K | No | Exactly 1 | | Seedance 2.5 | 4–30s | 480p, 720p | Yes | Up to 2 | | Seedance 2.0 | 4–15s | 480p, 720p, 1080p | Yes | Up to 2 | Two practical consequences: 1. **Length is a constraint, not a quality setting.** A five-second clip that holds identity beats a fifteen-second clip that loses the face at second eight. Ask for the shortest duration that fits the action you described. 2. **Longer costs proportionally more.** Video is billed per second, so a ten-second clip costs twice a five-second one at the same settings. The price is fixed by those settings, and a render that fails is refunded automatically. Aspect ratio is worth setting deliberately too. Every published video model offers 16:9, 9:16, 1:1, 4:3, 3:4 and 21:9, and cropping a finished clip afterwards throws away resolution you already paid for. Decide where the video is going — feed, story, hero banner — before you generate, not after. If you want the specification for a single model — supported operations, clip lengths and reference limits — open its page, for example [Seedance 2.5](/models/seedance-2-5), or browse all the [AI generation models](/models). ## Five mistakes that ruin an image to video prompt 1. **Re-describing the image.** The picture is already in the request. Words spent on it are words not spent on motion. 2. **Stacking transformations.** A dramatic subject action, a fast orbit, a lighting change and a wardrobe swap in one four-second clip give the model four goals that fight each other. 3. **Mood words instead of verbs.** _Cinematic_, _stunning_ and _epic_ do not specify movement. _Push in slowly while she turns towards the window_ does. 4. **Asking for readable text to change.** Logos, packaging copy and signage are the first things to smear once they move. Lock them, keep the camera calm, and keep them large in frame. 5. **Rewriting the whole prompt after every take.** You lose the parts that worked and cannot tell which change caused which result. ## How to iterate without starting over After the first take, change one thing. An image to video prompt is easier to debug when each generation differs from the last in a single, named way: - _Keep the motion, but halve the camera push-in._ - _The product shape is correct. Slow the rotation and remove the background mist._ - _Same shot, but hold the final frame for longer._ Two or three targeted passes usually beat ten rewrites, and they cost less. Every signed-in account gets 100 free credits a day, which covers still renders rather than clips — so the cheap place to practise the structure is on images, before you spend video credits on it. If you would rather start from something proven, the [video prompt library](/prompts/video) is free to read and remix — browsing it costs no credits. ## Image to video prompt FAQ ### How long should an image to video prompt be? Two to four sentences is usually enough. Length is not the goal; a short prompt where every clause is an instruction outperforms a long one padded with adjectives about the source image. ### Should I describe the image in the prompt? Only the parts that must stay fixed. Naming the subject's identity, a product's shape or a logo's position helps the model protect them. Describing colours, framing and lighting the image already shows adds nothing. ### Why does my subject's face change during the clip? Usually too much motion for the duration. Shorten the clip, pick a calmer camera move, and name the facial identity explicitly in the preserve clause before asking for any action. ### Can one image to video prompt produce a longer scene? Not in one clip. Generate several short clips that each hold up, keeping the same source image and preserve clause, then cut them together. That is more reliable than asking one generation to carry thirty seconds of story. ### Can an image to video clip have sound? That depends on the model, not the prompt. Seedance 2.5 and Seedance 2.0 produce audio-capable output; MiniMax H3 does not. Pick the model before you write the brief if the clip needs sound. ### How much does an image to video generation cost? It depends on the model, the resolution and the length. A default five-second MiniMax H3 clip costs 275 credits, and credits are priced at 100 to the US dollar. The same settings always cost the same, so the figure is settled before you send anything. ## Ready to run your image to video prompt? Bring one sharp source frame and a five-part motion brief. Pick the model, set the duration and resolution — those three decide the price — and iterate on the take instead of rewriting it. [Try the AI image to video generator →](/image-to-video) ## Keep reading More on writing an image to video prompt and choosing between the generation modes: - [Text to video vs image to video: which to use](/blog/text-to-video-vs-image-to-video) - [What is an AI video agent](/blog/ai-image-video-agent) - [What is Ideart AI](/blog/what-is-ideart) --- ### Text to Video vs Image to Video: Which to Use URL: https://ideart.ai/blog/text-to-video-vs-image-to-video Description: Text to video vs image to video: how each handles control, consistency and cost, when to pick which, and the workflow that uses both in the right order. Every AI video starts one of two ways: from words alone, or from a picture you already have. That is the whole of text to video vs image to video, and picking the wrong one is the most common reason a generation comes back looking nothing like the thing in your head. This comparison covers how each mode works, what each one gives you control over, where each one fails, what they cost, and a decision table you can apply to your next shot. ## Text to video vs image to video in one paragraph **Text to video** builds the whole frame from your description — subject, setting, framing, lighting, motion. You get range and speed, and you give up precise control of what things look like. **Image to video** starts from a still you supply and animates it, so composition, colours and identity are already settled and the prompt only directs movement. Text to video is for exploring an idea; image to video is for moving something specific that has to stay recognisable. ## How text to video works You write a prompt, choose a model, set the duration, resolution and aspect ratio, and the model generates every frame. Nothing about the picture is fixed in advance, so two runs of the same prompt produce two different-looking clips. That variance is the feature. When you do not yet know what the shot should look like, [AI text to video generation](/text-to-video) is the cheapest way to see five interpretations of an idea and find out which one you actually wanted. It is also the limitation. If the clip has to show _your_ product, _your_ character or _your_ brand colours, a text prompt is a lossy way to specify them, and each regeneration re-rolls the details you were trying to keep. ## How image to video works You upload a still image and describe the motion. The image supplies the subject, framing, palette and most of the lighting; the prompt supplies what moves, how the camera behaves and what has to stay put. Because the first frame is fixed, [AI image to video generation](/image-to-video) is repeatable in a way text to video is not. Run it three times and you get three versions of the same look, which is what makes it usable for product shots, campaign assets and anything with a face or a logo in it. The trade is that the clip can only ever be about what is in the picture. The model cannot show you the other side of a building it has never seen, and it will invent badly if you ask. ## Text to video vs image to video: the differences that matter | | Text to video | Image to video | | ---------------------------- | ----------------------------------------- | --------------------------------------------- | | **What you supply** | A written prompt | A still image plus a motion prompt | | **Who decides the look** | The model | You, in the source image | | **Consistency across takes** | Low — every run differs | High — the first frame is fixed | | **Prompt should describe** | Subject, setting, style, motion | Motion, camera, what must not change | | **Typical failure** | The right idea, the wrong-looking subject | Identity drift once the motion gets ambitious | | **Best for** | Concepts, moods, backgrounds, b-roll | Products, characters, brand assets, ads | The row that decides most projects is consistency. If you need the same object to appear the same way in six clips, text to video will make you fight for it and image to video gives it to you for free. ## When to pick text to video Choose text to video when the picture does not exist yet and nothing in the shot has to match something real: - **Concept exploration.** Five takes on "a lone rider crossing a dune field at sunrise" tells you what the sequence wants to be. - **Backgrounds and b-roll.** Establishing shots, textures, weather, ambience — footage nobody will scrutinise for identity. - **Moodboards and pitches.** You need the feel of a campaign before anyone approves the product photography. - **Anything you would otherwise licence as stock.** The practical tell: if you cannot name a specific real object that must appear correctly, text to video is the faster route. ## When to pick image to video Choose image to video when something in the frame has to survive intact: - **Product video.** The bottle, the shoe, the packaging, the label copy. - **Faces and characters.** A person the audience is meant to recognise across shots. - **Brand assets.** Approved photography that has already been signed off. - **Reworking a still you like.** You already generated the perfect image; now it needs to move. If you need a character to hold a look across several different shots rather than one, that is a third case — the [consistent character video generator](/consistent-character-video) exists for exactly that job. ## The workflow that uses both Treating text to video vs image to video as an either/or decision misses the most reliable production route, which uses both in order: 1. **Generate a still first.** Make the frame in an image model, where a bad result costs a fraction of a video and you can iterate quickly. 2. **Approve the look.** Composition, colour, product accuracy, the face — settle all of it while it is still a picture. 3. **Animate the approved still.** Now the video model only has to solve motion, which is the part it is good at. This is cheaper as well as more controllable. An image render costs a fraction of a video clip at the same quality bar, so every problem you catch at step two is a problem you did not pay video prices to discover. There is a middle option too. The Seedance models take reference inputs alongside the prompt — images, and on some operations video or audio — so you can hand the model a subject to keep without giving it the entire opening frame. That sits between the two modes: more direction than text alone, more freedom than animating one fixed still. ## Cost: does text to video vs image to video change the price? The mode itself does not change the rate. Video is billed per second of output, so the model, the resolution and the clip length set the price, not whether you started from text or from a picture. A default five-second MiniMax H3 clip costs 275 credits, and credits are priced at 100 to the US dollar, so a ten-second version of the same clip costs twice as much. Where the two modes differ is in **wasted** spend. Text to video takes more attempts to hit a specific look, and every attempt is a full video render. Image to video front-loads the iteration into cheap image renders and usually reaches an approved clip in fewer video generations. The price follows from the model, resolution and clip length, so the same settings always cost the same; a render that fails is refunded automatically, and every signed-in account gets 100 free credits a day to test image work with. Plans are listed on the [credit-based pricing](/pricing) page. ## Text to video vs image to video FAQ ### Which produces better quality? Neither, at the same model and resolution. They differ in control, not in fidelity. Image to video looks better more often only because you approved the first frame before spending video credits on it. ### Can I use the same prompt for both? No, and this is the most common mistake. A text to video prompt has to describe the whole scene. An image to video prompt should describe almost nothing except the motion, because the picture already carries the rest. ### Do the same models handle both modes? All three published video models — MiniMax H3, Seedance 2.5 and Seedance 2.0 — support generating from text and animating a still. They differ in clip length, resolution and whether they produce audio, which is listed on each model page. ### Can I turn a text to video result into an image to video source? Yes, and it is a good habit. Pull a frame you like out of a text to video clip, or regenerate that look as a still, then animate the still. You keep the composition the model found and stop re-rolling it on every take. ### Which one should a beginner start with? Image to video. The result is easier to judge because you can compare it to the source frame, and a single fixed input removes most of the variables while you learn how motion prompts behave. ## Ready to test text to video vs image to video on your own shot? Run the same idea both ways: once from a written prompt, once from a still you approve first. The difference in control is obvious within two takes, and each one is priced from the model, resolution and length you picked. [Try the AI video creator →](/video-creator) ## Keep reading More on text to video vs image to video and the prompts behind each mode: - [How to write image to video prompts](/blog/image-to-video-prompt-guide) - [What is an AI video agent](/blog/ai-image-video-agent) - [What is Ideart AI](/blog/what-is-ideart) --- ### What Is Ideart AI? Models, Credits and Limits Explained URL: https://ideart.ai/blog/what-is-ideart Description: Ideart AI is a credit-based AI image and video workspace. What it generates, which models it runs, what one render costs, and the limits worth knowing. Ideart AI is a credit-based workspace for generating AI images and AI videos from a prompt. One balance pays for every render, each job has its own workbench, and the models you can pick from are published with their limits in the open. This page covers what Ideart AI makes, which models it runs, what a single render costs, and where the limits sit — the four things worth knowing before you sign up for anything. ## What is Ideart AI? Ideart AI is a generation workspace rather than a single prompt box. You describe the result you want, attach a reference when identity matters, choose the model and the output settings, and the request goes to a video or image model at a price those settings decide on their own. Three things shape how it works day to day: - **One currency.** There is no per-tool or per-model billing. Every account holds a credit balance, and each generation draws from it. - **A workbench per job.** Text to video, image to video, image generation, image editing, video-to-video and consistent-character shots each open on their own page, with the controls that job actually needs. - **A prompt library that costs nothing to read.** Tens of thousands of prompts are free to browse and remix, whether or not you have credits left. Every signed-in account gets 100 free credits a day, which is enough to run image work before deciding whether to subscribe. ## What Ideart AI can generate ### AI images The [AI image generator](/ai-image-generator) turns a written prompt into an image, and accepts up to 16 reference images when you want the result to follow an existing subject, style or layout. You choose the aspect ratio and resolution before generating, and downloads are full-resolution. ### AI video from a text prompt The [AI text to video generator](/text-to-video) builds a clip from a description alone: subject, setting, motion, camera and pacing. Nothing is fixed in advance, which makes it the fast way to explore an idea you have not visualised yet. ### AI video from a still image The [AI image to video generator](/image-to-video) animates a picture you supply. The still fixes the composition, colour and identity; the prompt only directs the motion. This is the repeatable option, and the one to use when a product or a face has to stay recognisable. ### Editing, restyling and character work Image to image transforms an existing picture with a prompt. Video to video restyles a clip or transfers motion from one to another. Consistent character video keeps a person recognisable across separate shots rather than within a single clip. Each of these opens on its own workbench with the controls that job needs, so an image request is not buried under video settings and vice versa. Whichever one you start in, the output lands in the same place and is paid for out of the same balance. ## Which models Ideart AI runs The catalog is small and published on purpose — each model's clip lengths, resolutions and supported operations are listed rather than hidden behind an "AI" label. | Model | Type | Clip length | Resolutions | Audio | | ------------ | ----- | ----------- | ----------------- | ----- | | GPT Image 2 | Image | — | 1K, 2K, 4K | — | | MiniMax H3 | Video | 5–15s | 768P, 2K | No | | Seedance 2.5 | Video | 4–30s | 480p, 720p | Yes | | Seedance 2.0 | Video | 4–15s | 480p, 720p, 1080p | Yes | Seedance 2.5 is the broadest of the three: as well as generating from text and animating a still, it can work from reference inputs, edit an existing clip and extend one. Every video model has its own page with the full specification — start from the [AI generation models](/models) directory. ## What a render costs on Ideart AI Credits are priced at 100 to the US dollar, so one credit is one US cent. What a generation costs depends on three things: the model, the resolution and — for video — the length. - A default five-second MiniMax H3 clip costs **275 credits**, about $2.75. - A ten-second clip at the same settings costs twice that, because video is billed per second of output. - An image render is far cheaper. A 1K image at medium quality on GPT Image 2 costs **24 credits**, so the same monthly balance goes much further on image work. The same model, resolution and length always cost the same, so the figure is settled before you send anything, and a render that fails is refunded automatically. Plans start at $9.90 a month for 1,000 credits and are listed on the [credit-based pricing](/pricing) page; every plan includes free downloads and a commercial licence for what you generate. ## Where Ideart AI's limits are The honest constraints, so nothing is a surprise: - **Clips are short.** Four to thirty seconds depending on the model. Longer sequences are several clips cut together, not one generation. - **Video resolution tops out at 2K.** There is no 4K video tier. Images do go to 4K. - **Audio is model-dependent.** Seedance 2.5 and 2.0 produce audio-capable output; MiniMax H3 does not. - **Plan credits expire monthly.** They do not roll over, and the balance always spends the soonest-expiring credits first. Top-up credits last a year but require an active plan. - **Free credits do not stack.** Each day's 100-credit grant replaces the previous one rather than accumulating. ## Who Ideart AI is for Ideart AI suits people who need finished visual assets rather than a research playground: marketers producing ad variants, ecommerce sellers animating product photography, founders making launch clips, and creators who want the same character or product to look the same across a set of outputs. If you only ever need one image occasionally, the free daily credits will cover you without a plan. ## Ideart AI FAQ ### Is Ideart AI free? There is no free tier, but every signed-in account gets 100 free credits a day, which covers image work. Paid plans start at $9.90 a month for 1,000 credits. ### How much does one video cost? It depends on the model, resolution and duration. A default five-second MiniMax H3 clip is 275 credits, and the same settings always cost the same, so the price is predictable before you send anything. ### Which AI models does Ideart AI run? Four are published: GPT Image 2 for images, and MiniMax H3, Seedance 2.5 and Seedance 2.0 for video. Each video model has its own page listing its clip lengths, resolutions and supported operations. ### Can I use the results commercially? Yes. Every plan includes a commercial licence for the images and videos you generate, and downloading a finished file costs no extra credits. ### What happens if a generation fails? The credits are refunded automatically. You are not charged for a render that did not produce a result. ### Are my uploads used to train models? No. Uploads and results are encrypted in transit and at rest, and any conversation can be deleted from your history. ## Ready to try Ideart AI? Start with the job you actually have — an image, a clip from a prompt, or a photo that needs to move. Each tool opens on its own workbench, the price follows from the model, resolution and length you pick, and the first 100 credits arrive the day you sign in. [Browse the AI image and video tools →](/tools) ## Keep reading More on how Ideart AI generates images and video: - [How to write image to video prompts](/blog/image-to-video-prompt-guide) - [Text to video vs image to video](/blog/text-to-video-vs-image-to-video) - [What is an AI video agent](/blog/ai-image-video-agent) ---