← Home
PRODUCT ANALYSIS

By Valmera Editorial · Published · Updated

Vizard Review 2026: Pricing, Agent, API & Limits

Vizard now has two materially different products behind one brand. AI Studio is the established long-to-short system with one-source-minute clipping credits, transcript and timeline editing, automatic framing, captions, templates, API jobs, scheduling, and social publishing. Vizard Agent accepts an outcome instead of a sequence of tools. Its public page offers Flash, Pro, Max, and Ultra reasoning tiers, 22 starting templates, and 24 public showcases; first-party guides document work from a URL, idea, article, reference ad, photos, songs, raw footage, synchronized cameras, or finished video. The critical August update is evidence about Agent economics and behavior. Vizard still withholds unsigned plan prices and allowances, but its own guide pages now publish median credit use by tier, measured time and credit distributions for some categories, revision-billing descriptions, and probed output-file examples. This review separates those vendor measurements from promises, calculates the mature Studio plans independently, and documents when an agentic result creates a factual, rights, consent, or provenance obligation.

Is Vizard worth it in 2026?

Vizard is worth testing for teams that repeatedly turn podcasts, webinars, interviews, and events into social clips and want transcript correction, caption styling, brand templates, an API, scheduling, and direct publishing in one workflow. Its newer Vizard Agent is much broader: the current public surface exposes 22 templates and 24 showcases spanning product ads, explainers, multicam interviews, vlogs, trailers, localization, AI films, documentaries, and more. Agent pricing is still not published for signed-out buyers, but Vizard's own August use-case pages now reveal useful metering evidence: Flash is described as a free quick-edit tier, and the vendor reports cross-project median use of 47 Flash, 55 Pro, 242 Max, and 263 Ultra credits. Comparable jobs are described in tens of minutes rather than instantly; one measured advertising/promotional category reports a 28-minute median and 82-credit median. Those are vendor benchmarks, not guarantees or dollar prices. Agent keeps project assets for revisions and says only the new change is billed, but long jobs can lose coherence and rights, factual claims, likeness edits, source access, and realistic synthetic footage still require human control. Use AI Studio for predictable long-to-short operations; pilot Agent on a real job and inspect its signed-in credit balance and checkout before committing.

How much does Vizard cost in 2026?

These are the live US-dollar AI Studio prices and allowances checked August 26, 2026. Creator and Business can be configured with larger credit bundles. Annual plans are prepaid as one charge and begin at 7,200 annual credits. Vizard Agent is a separate public beta: its guides call Flash free but do not disclose the signed-out credit allowance, dollar price, or the prices of Pro, Max, and Ultra. Vendor-published per-job credit medians are workload evidence, not plan entitlements.

PlanMonthly priceIncluded creditsWhat it means
Free$060 credits/monthOne ordinary source minute per credit, 60-minute uploads, 1GB files, 720p watermarked exports up to ten minutes, one TikTok account, and three-day project retention
Creator — monthly$29/moFrom 600 credits/monthUp to 600-minute and 30GB uploads, 4K input and export, any-length export, no watermark, 100GB permanent storage, six social accounts, scheduling, custom templates, API, SRT and TXT
Creator — annual$174/year ($14.50/mo equivalent)From 7,200 credits/yearCreator features with annual prepayment; saves $174 versus 12 current monthly payments, and each annual credit release is documented as valid for 13 months
Business — monthly$39/moFrom 600 credits/monthShared workspace, Brand Kit, custom fonts, project sharing, faster AI, 20 social accounts, unlimited storage while subscribed, and higher API limits
Business — annual$234/year ($19.50/mo equivalent)From 7,200 credits/yearBusiness features with annual prepayment; saves $234 versus 12 current monthly payments before extra seats
Business seat — monthly$10/seat/moNo added credits publicly specifiedAdditional Business workspace access; current public material does not say that a seat adds to the workspace credit allowance
Business seat — annual$60/seat/year ($5/mo equivalent)No added credits publicly specifiedAnnual extra seat; the help center says seats added mid-term are prorated for remaining months
Vizard Agent — Flash$0 tier; allowance not publishedVendor reports 47-credit median across all projectsLabeled Junior and described as the free tier for quick, simple edits; not a guarantee that a 47-credit job fits the free allowance
Vizard Agent — ProNot publicly disclosedVendor reports 55-credit median across all projectsLabeled Skilled; inspect signed-in price, included balance, expiry, model access, and revision meter
Vizard Agent — MaxNot publicly disclosedVendor reports 242-credit median across all projectsLabeled Senior and repeatedly recommended by Vizard's own guides for publishable or customer-facing work
Vizard Agent — UltraNot publicly disclosedVendor reports 263-credit median across all projectsLabeled Master; public page does not explain when its higher meter is justified over Max

For ordinary AI Studio v1 clipping or transcription, one credit covers one source minute and a project using both is not charged twice. API v2 costs 1.25× v1. AI Enhance uses about 1 credit per output second at 1080p, 1.5 at 2K, and 2 at 4K. Agent uses a different workload meter. Vizard-authored August guides report cross-project medians of 47 Flash, 55 Pro, 242 Max, and 263 Ultra credits; advertising and promotional work is reported at 82 credits median with a 48–251 middle-half range and 28 minutes median with a 15–47-minute middle-half range. Many other categories were not measured separately and are compared with 28–38-minute medians from similar work. These figures cannot yield a dollar cost until Agent publishes or displays plan price, included credits, top-up price, expiry, and exact billing events.

QUICK VERDICT

Choose Vizard AI Studio when the recurring unit is a long speech-led source that should become selected, captioned, branded, reframed, scheduled clips. Creator's $29 monthly or $174 annual entry point is straightforward at one credit per ordinary v1 source minute, and the editor plus API provide useful correction paths. Choose Vizard Agent for a broader outcome such as a product ad from a URL, a synced multicam interview, a narrated explainer, a footage-led vlog, a trailer, a localized version, or a generated short film. Start with Flash only for a real low-risk test; Vizard itself describes it as free and suited to quick simple edits, while its guides recommend Max for many customer-facing jobs. Do not translate Agent credits into dollars until the signed-in account reveals price, included balance, expiry, and revision rules. Keep source files, prompts, approvals, rights records, factual checks, and delivered masters outside the project. Valmera is the stronger comparison when real uploaded footage and open-ended post-production are the center; Vizard retains the stronger mature clip-to-publish stack.

Method: this review analyzes Vizard's official product, help, and pricing material, checked August 26, 2026. Valmera publishes it and makes a competing video editor; there are no affiliate links, and Vizard's advantages are included because sometimes it is the right choice.

Vizard at a Glance

CategoryLong-to-short AI clipping, transcript and timeline editing, social publishing, API automation, and a separate general video agent
Best forPodcasts, webinars, interviews, events, social teams, agencies, and outcome-level brief-to-video experiments through Vizard Agent
Free plan60 credits every month, one ordinary source minute per credit, 720p watermarked output, ten-minute export maximum, and three-day retention
Creator$29 monthly or $174 yearly, starting at 600 monthly or 7,200 annual credits, 100GB, six social accounts, 4K, scheduling, templates, and API
Business$39 monthly or $234 yearly before seats, with a shared workspace, brand controls, 20 social accounts, unlimited storage while subscribed, faster AI, and higher API limits
Credit modelv1 clipping or transcription: 1/source minute; API v2 clipping: 1.25×; AI Enhance: 1–2 credits/output second depending on resolution
LanguagesThe current help center explicitly lists 40 transcription languages and 128 subtitle-translation languages; those are different entitlements
APIPaid key required according to the quickstart; remote files and 13 named platform-source types; polling or webhook completion; four output ratios and up to 100 returned clips
SocialTikTok, YouTube, LinkedIn, X, Instagram, and Facebook Pages; one TikTok connection on Free, six accounts on Creator, and 20 on Business
Legal postureSelf-serve terms grant a broad service-operation content license and permit machine-learning use; clear output-ownership and commercial-use promises were not found
Agent tiersFlash Junior, Pro Skilled, Max Senior, and Ultra Master; Flash is described by Vizard as the free quick/simple-edit tier
Agent public catalog22 starting templates and 24 showcases checked on the live signed-out page August 26, 2026
Agent credit benchmarkVizard-authored guides report cross-project medians of 47 Flash, 55 Pro, 242 Max, and 263 Ultra credits
Agent time benchmarkComparable jobs are described with 28–38-minute medians; measured advertising/promotional work reports 28 minutes median and 15–47 minutes for the middle half
Agent revision meterUse-case guides say project assets are retained and a revision is billed for the new change rather than another source upload; exact rates remain unpublished
Agent output examplesFirst-party guides probe delivered H.264/AAC files from 720×1280 at 24/30 fps through 1080×1920 at 30/60 fps; examples are not a format guarantee
Multicam requirementAgent's multicam guide depends on shared audio reference across camera files for sync and still requires review of speaker cuts
Synthetic speechAgent can replace a few spoken words with cloned voice and lip sync; exact replacement length, clean audio, permission, on-screen text, and regulated claims matter
Main trade-offMature Studio economics plus ambitious Agent breadth, but Agent dollar pricing is hidden and vendor benchmarks, rights, truth, consent, and provenance need active interpretation

What Vizard Actually Is

Vizard's established product is a browser-based repurposing system for speech-led media. A user uploads a file or imports a supported online source, selects the spoken language, and asks AI to find candidate highlights. The result is not merely a folder of extracted segments. Vizard connects clip selection to transcript editing, automatic reframing, captions, headline and keyword treatments, emojis, b-roll, layouts, brand templates, download, share links, scheduling, and direct publishing. That end-to-end operational loop is the product's clearest advantage over a tool that returns timestamps and leaves every finishing and distribution step elsewhere.

The clipping model works best when language carries the video's value. Vizard's own help says the input should be speech-based and recommends a source longer than 15 minutes. The number of clips is variable rather than guaranteed. Asking for shorter outputs can produce more candidates, while a keyword or prompt search narrows the result. A podcast interview, conference talk, webinar, tutorial, coaching call, or talking-head recording is therefore a natural fit. A silent montage, music video, sports sequence, or visually driven narrative gives the selection model less verbal evidence and should be tested rather than assumed to work equally well.

Clip preferences make the first pass more useful. The user can request one or several duration bands, choose vertical 9:16 or square 1:1 output, include emojis or highlighted words, and use a default or custom template. The API exposes a wider ratio set—9:16, 1:1, 4:5, and 16:9—and duration bands from under 30 seconds through 90 seconds to three minutes. Preferences are instructions, not a deterministic render specification. Vizard says AI can adapt a custom layout when the chosen template does not match the source composition, so a brand-critical team still needs to review framing and placement.

The transcript editor is a real editing surface. Deleting transcript text cuts the corresponding video, while editing a word corrects the subtitle rather than replacing recorded speech. Users can restore struck-through text, correct transcription, and work against both transcript and timeline. Original-video editing supports manual clip creation when AI does not find the required passage. Paid users can download the complete transcript as SRT or TXT, while the original editor also permits copying text. This is more controllable than accepting AI highlights unchanged and is especially efficient for dialogue-heavy corrections.

AI-generated clip boundaries are not final editorial truth. Vizard's own extension guide acknowledges that a result can start or stop awkwardly or begin in the middle of a sentence. The user can open the clip in the editor, restore removed transcript text, choose a longer duration preference, or return to the original-video editor and make a manual selection. That workflow is important: a virality or relevance model optimizes a candidate segment, but a human must protect setup, context, claim accuracy, speaker intent, and a natural ending.

Scenes expand the edit beyond a single continuous excerpt. From the editor, a user can insert material from the original transcript, upload local files, or use Brand Kit assets to create an intro, outro, insert, or b-roll scene. Scenes can be added from the timeline or transcript. This makes Vizard more than an automatic cropper, although it remains a streamlined web editor rather than a deep compositing or audio-post environment. A production that needs keyframes across many parameters, multicamera source management, advanced color, detailed sound mixing, or interchange with a mature NLE should validate those requirements separately.

Speaker framing has continued to evolve. Vizard's July 20, 2026 update says speaker identification now uses both audio and video context. Active Speaker is the default, and a user can select specific speakers who should stay visible. This helps interviews and podcasts in which a reaction matters even when that person is not currently talking. It is still an inferred layout: crossings, off-camera speakers, fast turns, occlusion, unusual framing, or visually important non-speakers can produce the wrong crop. Every automated vertical reframing pass deserves a full watch on the actual target aspect ratio.

Auto Censorship is another production-oriented addition. Vizard says it detects profanity and other sensitive words at the audio level, then lets the user silence them or replace them with a beep and customize the word list. That can accelerate brand-safety versions for social platforms, schools, or sponsors. It should not be treated as a compliance guarantee. Context, slurs, names, product claims, visual text, gestures, and platform-specific rules can fall outside a word detector, while an aggressive list can mute legitimate language. The final audio and captions need review together.

Captions and localization require careful vocabulary. Vizard's homepage uses a broad 'over 100 languages' translation claim, while its transcript-editing page says more than 180. The current help article is the more concrete source: it explicitly enumerates 40 transcription languages and 128 subtitle-translation languages. Translation coverage is therefore much broader than speech recognition coverage, and neither number proves equal accuracy, typography, line breaking, right-to-left rendering, speaker labeling, or cultural adaptation. Test the exact source and destination language pair with names, jargon, numbers, and branded terms.

Export resolution also needs a precise distinction. Creator and Business advertise output up to 4K, but selecting a 4K canvas does not invent lost source detail. Vizard's export guidance notes that recent YouTube imports may not yet expose the highest-quality version, vertical reframing can enlarge a small crop, and some high-quality or HDR sources may be adapted in the processing pipeline. If a sharp master matters, upload an appropriate source, inspect the downloaded file rather than the preview alone, and compare text, faces, motion, gradients, and audio synchronization.

AI Enhance is Vizard's paid super-resolution pass. It is applied to the final export, not to the original media or editor timeline, and it must be selected again after later edits or resolution changes. The feature can rebuild apparent edges, textures, faces, hair, and text after compression, cropping, or reframing. Results depend on the source: severe blur, noise, fast motion, camera shake, defocus, or an extreme jump such as 240p to 4K leaves little reliable information to reconstruct. Vizard recommends testing a short section first, a sensible safeguard for both quality and credits.

AI Enhance changes the economics dramatically. A 60-second output costs about 60 credits at 1080p, 90 at 2K, or 120 at 4K. A five-minute 4K enhancement would therefore consume about 600 credits—the entire starting monthly Creator or Business allowance—before any other credit-bearing work. Vizard shows an estimate before submission, requires at least one credit, and says failed enhancements are automatically refunded. While enhancement runs, downloading, publishing, trimming, and opening the clip in the editor may be temporarily unavailable; files above 500MB may fail.

Ordinary clipping is easier to budget. Vizard documents one credit per minute of source for v1 clipping or transcription, with no double charge when both happen inside the same project. A 60-minute episode therefore costs 60 credits, so the Free plan can process roughly one such episode in a month and a starting 600-credit paid plan can process roughly ten. API model v2 is documented at 1.25 times the v1 rate: that same hour costs 75 credits. Vizard describes v2 as more curated and coherent, while v1 remains the default for higher volume and lower cost.

The headline price creates useful but incomplete unit economics. Creator's starting monthly configuration is $29 for 600 credits, about 4.83 cents per ordinary source minute before rollover. Its $174 annual configuration supplies 7,200 starting credits, about 2.42 cents each, and saves $174 against 12 current monthly payments. Business is 6.5 cents per starting monthly credit or 3.25 cents annually and saves $234 against 12 monthly payments. Those figures do not include taxes, larger interactive bundles, Agent work, revision burn, AI Enhance, staff review, storage migration, or publishing operations.

Credit rollover is time-limited. Vizard's plan-change help says each monthly release remains valid for two months and each yearly release for 13 months, with rollover credits spent first. That permits uneven production but not indefinite banking. The same page says moving from Creator to Business starts a new cycle immediately and unused Creator credits do not carry. Increasing a usage tier adds credits immediately and charges the price difference without changing the existing billing date; an annual-plan upgrade starts a new cycle. Downgrades take effect at the end of the current cycle.

Vizard's own discount copy is not fully synchronized. The live pricing page and exact plan feed show Creator at $174 annually versus $348 across 12 monthly payments, and Business at $234 versus $468—a mathematical 50% saving. An upgrade help article still says annual billing saves five months, approximately 42%. The exact totals support the current 50% result, but the contradiction demonstrates why a buyer should preserve the checkout screen and invoice rather than rely on one marketing or help sentence.

Free is a meaningful product trial but a fragile workspace. It provides 60 monthly credits, a 60-minute maximum upload, a 1GB file limit, input up to 1080p, 720p export, a ten-minute export ceiling, the full editor, one connected social account, and a watermark. Project retention is only three days. The social-account help narrows the Free connection to TikTok. Anyone evaluating quality should schedule time to upload, process, correct, export, and archive the result before the project disappears rather than casually returning the following week.

Creator raises the practical production ceiling: sources can be as long as 600 minutes and 30GB, input and export can reach 4K, final duration is not capped by the plan table, the watermark disappears, storage becomes 100GB, and the account receives custom templates, social scheduling, six connected accounts, a private workspace, API access, and SRT/TXT downloads. It is the logical self-serve tier for an individual or small operation that needs ongoing output. The 600-credit entry allowance can still be exhausted by a single long enhancement or several large recordings.

Business changes collaboration and distribution rather than the starting credit pool. Its base selector also begins at 600 monthly or 7,200 annual credits, but it adds a shared workspace, Brand Kit, custom fonts, project sharing, faster processing, 20 social accounts, unlimited storage while subscribed, and higher published API limits. Extra access costs $10 per seat monthly or $60 per seat annually. Vizard's public material does not say that an added seat brings additional credits, so budget seats as access until a current quote or checkout states otherwise.

Business seat billing has edge cases. The help center says an annual extra seat costs $5 per month, charged as $60 for a full year, and a seat added during the term is prorated for remaining months with a separate invoice. On monthly billing, a new member's charge appears on the next bill, with a reminder three days before billing. The article's example uses an outdated base-plan amount, so it is appropriate evidence for seat mechanics but not for today's $234 Business base price. Current base amounts should come from the live subscription feed.

Workspace storage is not just a marketing number. Free includes 1GB of uploaded assets and three-day projects. Creator includes 100GB of permanent storage; once the quota is exceeded, later projects can become temporary and normally expire after 28 days. Business is described as unlimited while subscribed. Original media counts after Vizard's transcoding, and uploaded assets plus Brand Kit material also count. Removing a project from permanent storage turns it into a temporary project, usually with a 28-day expiry, so storage cleanup can become data deletion if misunderstood.

A downgrade changes team state. When a Business workspace ends and becomes Free, Vizard says non-admin members are removed and projects remain but begin a 28-day expiry period. Scheduled social posts are similarly time-sensitive: the help center says projects and videos remain for 28 days after subscription expiry, posts scheduled inside that window can publish, and posts scheduled later fail. A team leaving Vizard should export masters, transcripts, brand assets, schedules, account mappings, and API job records before canceling rather than relying on cloud retention.

Publishing is unusually complete for a clipping product. Vizard documents connections to TikTok, YouTube, LinkedIn, X, Instagram, and Facebook Pages. The calendar can target multiple accounts, generate descriptions and hashtags, and regenerate copy with a selected tone or voice. Auto Schedule can use a minimum virality score, daily cap, pattern, active days, date range, and time window; AI chooses clips and copy, and the team can revise them later. The current help says up to ten clips per day can be scheduled from a project.

Direct publishing reduces repetitive work but increases review risk. An AI-selected clip, generated headline, hashtags, caption, and post description can all be individually plausible while collectively misrepresenting the source or brand. Account permissions can expire, platform limits change, and a scheduled asset can outlive the intended campaign context. The calendar should be treated as a controlled queue: name an approver, preview the final crop with captions, verify destination and time zone, inspect claims and links, and retain the exported master independently of a social connection.

The API turns recurring repurposing into a pipeline. Vizard's quickstart says API keys are available on paid plans under Workspace settings and should be kept secret. A job can use a remote file or sources such as YouTube, Google Drive, Vimeo, StreamYard, TikTok, X, Twitch, Loom, Facebook, LinkedIn, Dropbox, and Instagram. The caller submits the source, spoken language, desired durations and ratio, and optional template, silence removal, keywords, subtitles, headline, emoji, highlights, or automatic b-roll, then polls or receives a webhook.

API results are asynchronous. Vizard recommends polling about every 30 seconds and says processing can take minutes to tens of minutes, with 4K sources taking longer. Its example says a 30-minute talking-head video can yield more than 20 clips in about ten minutes; that is an illustration, not a service-level promise. The submit endpoint accepts mp4, 3gp, avi, and mov for remote files, four aspect ratios, a maximum return setting from one to 100 clips, and duration preferences through three minutes. The short-video editing API separately requires input under three minutes according to the official machine-readable docs index.

Public API entitlement needs a small qualification. The plan comparison visually displays a Free rate-limit row of one request per minute and ten per hour, but Vizard's API quickstart says a paid subscription is required to obtain an API key. Creator is shown at three requests per minute and 20 per hour; Business at ten per minute and 60 per hour. The safe reading is that supported API access is paid-only unless the actual Free workspace exposes a key. Do not design a production integration around the visual Free limit without confirming authentication.

Vizard Agent is strategically broader than the AI Studio. The August 10 launch says a user can start from raw footage, an existing video, script, article, product page, images, creative brief, document, website, or simply an idea. Agent may research, develop an angle, write or restructure a script, create visuals, edit existing footage, work with audio and voice, generate missing material, localize, adapt to a platform or audience, and carry the project through revision. The intended unit is an outcome such as a launch video or ad, not a sequence of feature-level commands.

The public Agent surface calls itself 'The First Video AGI' and currently exposes Flash, Pro, Max, and Ultra model choices. It offers more than 22 starting templates, including product and software ads, AI motion, talking head, vlog, tutorial, UGC ad, explainer, podcast clips, localization, listicle, avatar, customer story, photo story, event recap, trailer, multicam, music video, movie recap, AI film, mini-documentary, and documentary. Those templates demonstrate breadth, but public model names are not a price card, quality benchmark, speed guarantee, or rights statement.

Agent also changes the interface comparison. Vizard's conventional AI Studio has transcript and timeline controls. A separate Vizard article describes Agent as having no timeline to drag because the user directs the outcome conversationally. Both can be true: the agentic surface abstracts the sequence, while the established editor exposes direct controls. A buyer should decide whether the desired relationship is delegation, manual correction, or both—and confirm how an Agent result moves into editable Studio state when precision changes are required.

Vizard is unusually candid about the current Agent ceiling. Its launch post says a long or complicated task can lose coherence; an early decision may not carry through; visual style and tone can drift; the system may make a reasonable but unwanted decision; and one prompt will not always produce a finished result. Human direction remains necessary. That is the correct procurement model: define the audience, purpose, facts, prohibited claims, source hierarchy, visual references, delivery specifications, and approval checkpoints, then price multiple revision rounds rather than a single generation.

The self-serve contract deserves equal attention. Vizard's Terms were last updated October 31, 2022, years before the public Agent launch. They say the user keeps responsibility for content and grants Vizard a nonexclusive, perpetual, irrevocable, royalty-free, worldwide, transferable, sublicensable license to access, use, host, cache, reproduce, transmit, and display it, stated solely to operate and maintain the service. They also permit collection of Usage Data and use of machine learning produced while processing content and usage to test, tune, optimize, validate, and enhance analytics, models, and algorithms.

Those Terms do not provide a simple no-training promise. They also do not clearly state that a customer owns generated output or may use every output commercially. That absence is not proof that commercial use is forbidden; it means the public self-serve contract does not resolve the question clearly. Inputs can contain music, trademarks, faces, confidential information, client material, and third-party clips, while Agent may add generated or researched assets. For valuable campaigns, regulated work, minors, unreleased products, or confidential footage, obtain written rights, training, deletion, subprocessor, and indemnity terms before upload.

The Privacy Policy is also dated October 31, 2022. It describes purpose-based retention rather than a universal fixed period, says aggregated non-identifying information may remain after account deletion, allows service providers to process data for designated functions, and notes international transfers. A user can request deletion by contacting Vizard, but copies may be retained for legal obligations. The policy mentions SSL, digital signatures, and PCI-compliant payment processors; this review did not find a current first-party SOC 2 or ISO certification claim. Enterprise procurement should request present evidence rather than infer it.

Refund protection is narrower than a casual seven-day guarantee. Vizard's current help says a refund may be available for a double or incorrect charge, or when the request arrives within seven days and the service has not been used. Uploading data or using tools counts as use. Processing fees may be deducted unless Vizard caused the billing error, support makes the final determination, and an approved refund can take five to ten business days. The Terms similarly describe a first-seven-days unused condition. Test meaningful workflow fit before paying, but understand that the test itself can end refund eligibility.

Cancellation ends renewal rather than immediately deleting a paid term. The pricing FAQ and help center direct users to Workspace Subscription and say access continues until the billing cycle ends. Plans auto-renew unless terminated. Vizard accepts cards, PayPal, Stripe, and bank transfers according to its pricing FAQ, with Stripe handling payment details. Because storage, members, projects, scheduled posts, rollover credits, and API access can change at expiry, cancellation should be treated as a migration project, not merely a billing toggle.

The fairest conclusion is that Vizard has a mature specialist core and an ambitious agentic expansion. The specialist system has clear value when the same organization repeatedly processes spoken long-form media and must publish across many channels. Agent may become the more important product, but its public pricing and contract coverage lag its capability claims. A disciplined buyer should run two evaluations: one measuring clips accepted per source hour and minutes of human correction, and another measuring Agent's complete brief-to-approved-video cost, factual accuracy, visual consistency, revision count, rights, and export fitness.

The live signed-out Agent page is more specific than the original launch announcement. It exposes 22 templates: Product ad, Software ad, AI motion, Talking head, Vlog story, Software tutorial, UGC ad, Explainer, Podcast clips, Localize, Listicle, AI avatar, Customer story, Photo story, Event recap, Trailer, Multicam, Music video, Movie recap, AI film, Mini-doc, and Documentary. It separately links 24 public showcase cases such as a product ad from a screenshot, article-to-explainer, keynote trailer, Figma-led editing style, and raw-clips-to-vlog. Template count is breadth, not proof of equal reliability.

The tier labels communicate intended capability rather than pricing transparency. Flash is labeled Junior, Pro Skilled, Max Senior, and Ultra Master. Vizard's own guides call Flash free and suited to quick, simple edits, while repeatedly recommending Max for customer-facing explainers, launch trailers, channel uploads, releases, event recaps, and other work intended to publish. That is vendor guidance and an upsell signal, not an independent quality threshold. A buyer needs side-by-side outputs from the same brief before deciding whether Max's much higher median credit use is justified.

Agent credit evidence now exists even though public plan economics do not. Across its August use-case pages, Vizard reports all-project median consumption of 47 credits on Flash, 55 on Pro, 242 on Max, and 263 on Ultra. The jump is not linear: Max's median is about 4.4 times Pro's and Ultra's only about 8.7 percent above Max's. Those numbers may reflect different job complexity selected for each tier rather than the same task run four ways. They do not reveal included balance, dollar price, top-up price, expiry, failure refunds, reservation timing, or what happens when an agent changes tiers mid-project.

Vizard also publishes a measured distribution for advertising and promotional work. Its trailer guide reports a 28-minute end-to-end median, a 15–47-minute middle-half range, an 82-credit median, and a 48–251-credit middle-half range. Other pages often say their category was not separately measured and compare it with jobs whose medians run from 28 to 38 minutes, with upper quartiles as high as 47 or 61. These are valuable vendor benchmarks because they reject instant-output expectations, but they are neither a service-level promise nor an independent benchmark methodology.

The fair way to use those statistics is as a test design. Record prompt time, tier, source bytes and duration, start time, first playable result, completion, credits before and after, number of generated or searched assets, revisions, export specs, and human correction time. Report the median and middle half of approved jobs, not the fastest demo. Compare the same brief on Flash, Pro, and Max when the account permits. A 242-credit Max result can still be cheaper if it eliminates several revisions; a 47-credit Flash result is expensive if it never becomes publishable.

Revision language on the Agent guides changes cost planning. Vizard repeatedly says the source footage and built assets remain inside the project, so requesting a shorter platform version or changed audience does not require another upload and the user is billed only for the new change. That is better than restarting, but ‘only the change’ is not a rate card. The system may need to regenerate narration, visuals, music, captions, or many scenes to satisfy one sentence. Check the projected or posted credit change for a micro-edit, structural revision, new aspect ratio, and new language before promising marginal cost.

The current output examples prove variability rather than one spec. A first-person synthetic clip was probed at 720×1280 H.264, 30 fps, 19.93 seconds with AAC. A generated ASMR piece was 720×1280 at 24 fps for 40.31 seconds. A meme-format example was 1080×1920 at 60 fps for 57.58 seconds. A three-minute B-roll edit was 1080×1920 at 30 fps with preserved AAC audio. These are Vizard-authored examples from individual runs, not a public promise that every tier supports every dimension, frame rate, codec, duration, bitrate, or audio configuration.

Multicam is one of Agent's most consequential footage-first claims. The current guide says to upload every camera angle and the audio, after which Agent aligns the files, cuts to the person speaking, and removes dead air. The shared audio reference is the synchronization anchor and must exist in the recordings; it cannot be manufactured after the shoot. The result still needs a full watch for the wrong face, reaction shots, interruptions, off-camera speech, waveform drift, camera gaps, and whether dead-air removal damages rhythm.

Short spoken corrections combine transcription, synthetic voice, and lip synchronization. Vizard says Agent locates the phrase in a word-level transcript, clones the speaker for the replacement, syncs the visible mouth for those seconds, and seeks a quiet waveform point for the join. Exact written wording matters because ‘the seventeenth’ and ‘August seventeenth’ occupy different time. A longer replacement can force timing changes, noisy or reverberant audio clones poorly, close faces expose errors, and any matching date or price visible on screen must be corrected separately.

The deeper limit is consent and truth. Replacing someone's words changes the audiovisual record. Vizard's own guide says the video must be yours or used with permission and warns that a changed regulated claim remains a regulated claim. A responsible workflow stores the original, identifies the synthetic span, obtains speaker and rights approval, verifies every affected caption and graphic, preserves the reason for the correction, and discloses the change where audience context or law requires it. Technical plausibility does not make fabricated testimony legitimate.

URL-to-ad work shows Agent acting as a researcher and data extractor. One official example says Agent reads a shop, identifies branding and images, inspects source code, discovers the underlying product feed, collects catalog items and prices, and builds an ad around a chosen product, category, sale, audience, length, and destination. That can remove tedious asset collection. It can also reproduce stale prices, select discontinued inventory, misread variants, encounter an automated-access block, or make a claim the landing page does not support. The merchant's current approved feed and offer terms must beat the generated video.

Product-feed access is not permission to publish every field or bypass a site's rules. Use a domain and catalog the customer controls, identify products that are off limits, recheck price, stock, currency, tax, shipping, sale dates, and required disclaimers immediately before launch, and preserve the extracted source. If automated access is blocked, Vizard advises supplying assets directly. A video ad should not depend on probing private endpoints, hidden customer data, or a third-party site the user has no right to reuse.

Explainers demonstrate Agent's no-footage path. A topic, audience, question, length, and any company-specific definitions can be enough for Agent to write narration, generate matching visuals, record voiceover, add captions and music, and return a finished cut. The audience often shapes vocabulary and pacing more than a style adjective. The dangerous point is that the first close reading of the script may occur after the video is rendered. Review the narration, source every factual claim, and correct confident errors before evaluating visual polish.

Vlogs, trailers, and event recaps demonstrate footage culling and story construction. Agent can watch unsorted material, find usable moments, choose a spine, build escalation or an arc, cut to music, add title cards or captions, and make platform variants. The prompt must still name what the piece is about, its intended feeling, the ending, anything that must or must not appear, and the target duration. An agent can find a story that exists in the footage; it can also manufacture a misleading story when the brief rewards drama over source truth.

Reference-ad editing has a clear rights boundary. Vizard says Agent can study a commercial's pacing, shot order, camera moves, and music beat structure, then apply that structural lesson to the user's footage and a separately licensable track. It also warns not to use the reference footage or music and notes that trade dress can be protected. ‘Inspired by’ becomes risky when a viewer could mistake the output for the source brand or campaign. Preserve originality in footage, product, graphics, copy, sound, and overall impression.

Agent's generated AI-film and realistic-footage templates intensify provenance duties. Official guides say it can build a short film from a premise with generated shots, characters, voices, music, and grading, or generate first-person phone footage with deliberately imperfect camera and sound. The realistic-footage guide explicitly warns that mistakable synthetic footage should be labeled and should not depict real people or actual incidents. Keep prompt, generation log, asset provenance, disclosure copy, and approval beside the master, and never let a plausible render become false evidence.

The movie-recap guide is unusually candid about copyright. Agent can upload a film, write a narration, voice it, and cut the source footage to match, but Vizard says the unresolved question is what the user may lawfully publish and monetize. A transformation workflow is not a license. Rights, jurisdiction, amount used, market effect, commentary purpose, platform claims, and local legal advice can matter. The correct production gate occurs before upload and credit spend, not after a finished recap has already been created.

Sourcing has the same problem even when Agent searches for material marked reusable. In a meme-video example, Vizard says Agent searches stock and then narrows web searches toward no-copyright vertical gameplay. The guide still says copyright follows the footage, music is a frequent claim trigger, and platforms may demote mass-produced formats. ‘No copyright’ in a title or search result is not chain-of-title evidence. Prefer customer-owned or clearly licensed assets and preserve the actual license, author, URL, retrieval date, and permitted use.

The Agent catalog therefore creates a new procurement checklist beyond the old AI Studio plans. Establish where uploaded files and generated assets are stored, whether the 2022 Terms govern Agent, whether Agent credits are separate from Studio credits, how prices and top-ups work, when credits are reserved and restored, how long projects persist, which models and search providers process inputs, what commercial output rights apply, how deletion reaches subprocessors, and which actions are recorded for audit. Public creative examples cannot answer those contractual questions.

A rigorous trial should run four jobs. First, one predictable AI Studio hour to confirm ordinary credit and export behavior. Second, the same modest Agent brief on Flash and Pro. Third, one Max job with mixed real and generated media. Fourth, a revision sequence: change one word, one scene, aspect ratio, audience, and language. Record all credits and elapsed time. Evaluate factual accuracy, source fidelity, identity consistency, visual style drift, caption errors, audio joins, legal clearance, correction time, and whether the result can be reopened or exported without the platform.

Where Vizard Wins

The workflow reaches all the way to publishing

Vizard combines long-video ingestion, highlight selection, transcript correction, reframing, captions, branding, post copy, a calendar, and direct social publishing. That removes substantial handoff work for teams whose deliverable is a steady stream of platform-ready clips rather than one master timeline.

Ordinary clipping cost is easy to forecast

One v1 credit per source minute is unusually legible, and transcription plus clipping inside the same project is not double-charged. A team can translate its podcast, event, or webinar calendar into a baseline allowance before accounting for v2 and enhancement.

The editor can correct AI rather than merely accept it

Transcript deletion, subtitle correction, restored text, manual clipping, timeline work, added scenes, layouts, templates, and b-roll give an operator several ways to repair awkward boundaries, missing context, bad crops, or an unsuitable first pass.

Spoken-media features are purposeful

Active-speaker detection, selected visible speakers, silence and filler control, subtitles, headline hooks, keyword highlights, emojis, ratios, and Auto Censorship map directly to the needs of podcasts, interviews, educational recordings, and social excerpts.

Creator is a complete individual production tier

At the published starting price, Creator removes the watermark, supports 4K and any-length exports, raises uploads to 600 minutes and 30GB, adds 100GB permanent storage, scheduling, six social accounts, templates, transcript downloads, and paid API access.

The API exposes meaningful creative controls

Automation is not limited to a file upload. Callers can specify source, language, output ratio, duration bands, template, silence removal, keywords, subtitles, headline, emoji, highlights, b-roll, and a clip-count ceiling, then use polling or webhooks.

Source integrations cover common media systems

The submission reference names remote files plus YouTube, Drive, Vimeo, StreamYard, TikTok, X, Twitch, Loom, Facebook, LinkedIn, Dropbox, and Instagram. That breadth can remove download-and-reupload steps in an authorized pipeline.

Publishing controls fit multi-channel operations

Six named social networks, multiple connected accounts, generated descriptions and hashtags, adjustable tone, a calendar, virality thresholds, day and time rules, and up to ten scheduled clips per project per day make Vizard operationally useful after editing.

Business adds practical team structure

A shared workspace, projects and export sharing, Brand Kit, custom fonts, 20 social accounts, faster processing, unlimited storage while subscribed, and higher API limits address agency and in-house social teams better than a single-user clipper.

Agent is genuinely broader than repurposing

Vizard Agent can begin from a goal and mixed context, decide which operations are needed, generate missing material, and revise the finished result. Ads, launches, tutorials, localization, documentaries, and footage-led jobs can share one conversational surface.

Vizard states Agent's weaknesses plainly

The launch post openly says long tasks can lose coherence, style and tone can drift, prior decisions can be lost, and one prompt may not be enough. That candor gives evaluators a realistic basis for checkpoints and revision budgets.

Official documentation is unusually detailed

The help center publishes exact storage behavior, rollover windows, social-account limits, credit rates, enhancement failure refunds, language lists, API parameters, expiry behavior, and Agent limitations. Some pages conflict, but the evidence needed to spot those conflicts is public.

Agent publishes real workload benchmarks

Vizard's own guides disclose median credits by tier plus measured time and credit ranges for some categories. The figures are not independent or complete, but they provide much more planning evidence than a generic faster-video promise.

Flash provides a genuine zero-price test surface

Vizard describes Flash as the free tier for quick, simple edits. A buyer can test brief interpretation and output behavior before paying, even though the public allowance is not disclosed.

Agent spans real footage and generated media

The public catalog covers multicam interviews, vlogs, trailers, event recaps, localization, ads, explainers, avatars, AI films, documentaries, and mixed-media work rather than only one prompt-to-video format.

Project assets can support incremental revisions

First-party guides say Agent retains footage, scripts, and generated assets, then bills the requested change rather than another source upload. Platform and audience variants can reuse prior work.

Multicam sync attacks a high-friction edit

Using a shared audio reference to align angles, cut by speaker, and tighten dead air can remove hours from an interview workflow when the source was recorded correctly.

Agent guides acknowledge rights and truth limits

Vizard's own pages warn about consent for synthetic speech, factual review for explainers, film and music rights, reference-ad trade dress, and disclosure for realistic generated footage. That candor helps teams build review gates.

Limits and Trade-offs

Agent pricing is not public on the unsigned page

Flash, Pro, Max, and Ultra are visible, but their prices, included units, generation rates, and revision economics are not. The $29 and $39 plans belong to AI Studio and cannot safely stand in for Agent cost.

The credit unit changes by feature

One credit equals one source minute only for ordinary v1 clipping or transcription. API v2 costs 1.25×, while AI Enhance consumes 1–2 credits per output second. A generic balance can hide very different production quantities.

AI Enhance can consume a whole starting allowance

Enhancing five minutes at 4K costs about 600 credits. That equals the starting paid monthly pool before clipping, transcription, or other work and makes high-resolution finishing a major budget line rather than a small toggle.

Annual billing is prepaid

The $14.50 and $19.50 figures are monthly equivalents, not monthly invoices. Creator charges $174 and Business $234 for the year, before tax, larger allowance selections, extra seats, or other purchases.

Rollover expires

Monthly releases remain usable for two months and annual releases for 13 months according to current help. This smooths a production spike but does not create a permanent bank of unused credits.

Creator credits disappear on a Business upgrade

Vizard explicitly says a Creator-to-Business change starts a new cycle and unused Creator credits do not carry. Upgrade timing can therefore erase paid capacity even though rollover is advertised elsewhere.

Vizard's annual-discount documents conflict

The current prices mathematically save 50% versus 12 monthly payments, matching the live page. An upgrade help article still describes about 42%, so marketing and support copy are not fully synchronized.

Free retention is only three days

A Free project can disappear before a relaxed evaluation is complete. The tier requires a planned upload-edit-export session and immediate local archiving, especially if stakeholders need to approve the result.

Free output is constrained

Free adds a watermark, exports only at 720p, caps an export at ten minutes, limits a file to 1GB and an upload to 60 minutes, and connects only one TikTok account according to current help.

Creator storage is finite and overflow becomes temporary

Creator includes 100GB, but projects created beyond that permanent quota can receive a roughly 28-day expiry. Uploaded assets, Brand Kit files, and transcoded original media count toward workspace use.

Removing permanent storage status starts an expiry clock

A project taken out of permanent storage becomes temporary and is normally retained for 28 days. A cleanup action can therefore become deletion unless the team has independent masters and a clear archive policy.

Business storage depends on an active subscription

The unlimited promise applies while subscribed. At downgrade, projects move toward a 28-day expiry state and non-admin workspace members are removed, so cancellation requires an orderly export and access migration.

Scheduled posts can fail after expiry

Vizard says posts inside the 28-day post-subscription window can still publish, but later ones fail. A future social calendar is not a durable archive or guaranteed queue after cancellation.

AI chooses candidates, not editorial truth

A clip can be catchy and still omit qualification, reverse the speaker's intent, expose confidential context, or repeat a factual error. Virality and coherence scores cannot replace source-aware approval.

Clip boundaries may be awkward

Vizard's own help acknowledges starts or endings that land mid-sentence. Repair can require restored transcript text, a longer duration preference, or manual clipping from the original-video editor.

Speech-led sources are the natural fit

The clip generator is built around spoken content and Vizard recommends inputs over 15 minutes. Music, cinematic montage, sports, demonstrations, or other visually led stories may not offer enough transcript evidence for strong selection.

Clip count is not guaranteed

Duration preferences, source structure, prompt search, and model judgment affect the number returned. The API's maxClipNumber is only a ceiling from one to 100, not a commitment to supply that many usable clips.

Automated reframing can choose the wrong subject

Audio-video speaker identification improves the crop, but reactions, products, slides, demonstrations, off-camera voices, and rapid exchanges can make the active speaker an incomplete visual priority. Every ratio needs review.

A 4K setting does not restore real detail

A low-resolution or heavily cropped source can be exported in a 4K container without becoming truly sharp. AI Enhance may reconstruct plausible texture, but blur, compression, noise, and missing information constrain the result.

Enhancement can block actions and fail on large exports

Downloading, publishing, trimming, or reopening a clip may be unavailable during processing, and Vizard warns that files larger than 500MB may fail. High resolution also takes longer than standard export.

Transcription and translation coverage differ

Vizard lists 40 recognition languages and 128 subtitle-translation languages. A broad translation headline does not mean the system can transcribe an arbitrary source language or preserve every name, dialect, glyph, or line break.

The public API entitlement is ambiguous at Free

The pricing comparison displays a Free rate limit, but the quickstart says a paid plan is required for an API key. Production planning should follow the paid-only requirement unless a real Free workspace proves otherwise.

API processing time is not an SLA

Vizard describes minutes to tens of minutes and offers an illustrative ten-minute result for a 30-minute talking-head source. Network import, source resolution, queue, model, job settings, and failures can change the result.

API secrets and external-source rights remain the customer's job

Keys must be protected, source URLs must remain reachable, and a platform URL does not confer permission to download or republish its content. Automation magnifies both access and rights mistakes.

Agent can lose coherence

Vizard says decisions made early in a long job may not carry through, while visual style or tone can drift. Complex work needs explicit checkpoints and a final consistency pass.

Agent can make a reasonable but unwanted choice

Outcome-level delegation necessarily gives the system creative discretion. A result can be technically sound yet wrong for the audience, offer, platform, brand, or factual position, requiring directed revisions.

Agent has no public one-prompt guarantee

Vizard explicitly rejects the magic-button framing. Teams should estimate review labor and multiple revisions, and the undisclosed public Agent meter makes that full-job cost difficult to compare before sign-in.

The Terms predate the Agent launch

The self-serve contract was last updated in October 2022, while Agent launched in August 2026. It does not clearly address the modern product's output ownership, generation inputs, or commercial-use expectations.

There is no simple self-serve no-training promise

The Terms permit machine-learning use arising from processing user content and usage data to tune, validate, optimize, and enhance analytics, models, and algorithms. Sensitive buyers should obtain a contrary written commitment if needed.

The content license is broad

Although stated as solely for operating and maintaining the service, the license is described as perpetual, irrevocable, transferable, and sublicensable. Counsel should evaluate it against client, confidential, biometric, minor, and unreleased-product material.

Output ownership and commercial rights are not clearly promised

The public Terms do not plainly allocate generated-output ownership or grant an express commercial-use right. That is a contract gap, not a conclusion that all use is prohibited; obtain written clarification for valuable work.

The Privacy Policy is also dated 2022

The public policy predates Agent, provides purpose-based rather than universal fixed retention, and allows some aggregated non-identifying data to persist. Current subprocessors, transfer terms, and controls should be requested for procurement.

No current first-party SOC 2 or ISO claim was found

The policy discusses SSL, digital signatures, and PCI-compliant payment processors, but this review did not find a current certification claim. An enterprise buyer should request reports and scope instead of inferring certification.

The seven-day refund is effectively unused-only

Uploading data or using tools counts as use under current help. A normal product test can therefore remove refund eligibility, processing fees may be deducted, and approval remains with support.

Agent credit medians are not dollar prices

47, 55, 242, and 263 credits cannot be converted to cost without public plan price, included balance, top-up rate, expiry, failure rules, and exact billing events. Signed-in checkout remains necessary.

The tier medians may reflect different workloads

Max jobs may be more complex than Flash jobs, so a 4.4× median difference does not prove the same project costs 4.4× more. Vizard does not publish a controlled same-brief methodology.

Time statistics are vendor benchmarks, not an SLA

A 28-minute promotional median and 15–47-minute middle-half range do not guarantee one project. Source volume, generation retries, tier, queue, search, export, and revision scope can create a long tail.

Incremental revision pricing remains undefined

Guides say only the requested change is billed, but one sentence can trigger many regenerated assets. No public rate table maps a text correction, recut, new ratio, new audience, or localization to credits.

Agent outputs do not have one published format contract

Official examples vary from 720×1280 at 24 or 30 fps to 1080×1920 at 30 or 60 fps. Individual probed files do not establish supported codecs, bitrates, frame rates, resolutions, durations, HDR, or audio across tiers.

Multicam depends on shoot-time reference audio

Shared audio is the synchronization anchor. Missing scratch audio, drift, dropped frames, isolated recorders, or wrong sample clocks can defeat an automated sync that cannot reconstruct the shoot after the fact.

Synthetic word replacement can falsify the record

Voice cloning and lip sync make a changed statement look authentic. Permission, exact wording, claim review, disclosure, original preservation, and on-screen-text correction are mandatory safeguards.

URL research can ingest stale or unauthorized data

A live product feed can contain wrong prices, variants, unavailable items, or fields not intended for an ad. Automated access may also be blocked or prohibited. The approved merchant record must control.

Reference commercials create imitation risk

Learning pacing is different from copying footage, music, trade dress, or a distinctive campaign impression. Agentic execution does not erase copyright, trademark, or unfair-competition boundaries.

Realistic synthetic footage needs disclosure

First-person generation is intentionally capable of looking accidental and real. Depicting an actual person or incident can create false evidence, reputational harm, policy violations, and legal exposure.

Search-found media is not automatically licensed

A no-copyright label or search result is weak provenance. Stock, gameplay, music, archival footage, and social clips need a retrievable license and permitted-use record.

The 2022 contract predates these capabilities

Agent now researches sites, clones voices, generates realistic footage, and invokes multiple creative operations, while the public Terms and Privacy Policy predate launch. Critical rights and data promises need current written terms.

Who Should Choose Vizard?

  • Podcast teams that need multiple captioned, reframed clips from every long episode.
  • Webinar, interview, conference, and event marketers searching speech-led recordings for campaign moments.
  • Social teams that want editing, captions, post copy, a multi-account calendar, and direct publishing in one system.
  • Agencies using templates, Brand Kit, shared workspaces, many social accounts, and repeated API submissions.
  • Developers automating authorized long-to-short workflows with polling or webhooks and structured output preferences.
  • Creators who prefer editing dialogue by deleting and restoring transcript text instead of navigating a deep NLE.
  • Teams willing to supervise Vizard Agent for ads, launches, explainers, tutorials, localization, or mixed-media stories.
  • Operators who can measure accepted clips, correction time, revision cost, and distribution throughput on representative media.
  • Buyers whose footage and contractual requirements are compatible with Vizard's current self-serve data terms.
  • Teams willing to benchmark Flash, Pro, and Max on the same outcome and record credit, time, and correction evidence.
  • Multicam interview producers whose cameras share usable reference audio and who can review every automated speaker cut.
  • Marketers building product ads from controlled first-party URLs, catalogs, prices, footage, and approved claims.
  • Creators turning large unsorted footage collections into vlogs, trailers, event recaps, or short social versions through iterative briefs.
  • Organizations prepared to preserve provenance, consent, rights, disclosure, and factual review around generated or synthetic media.

Vizard vs Valmera

Vizard and Valmera now overlap more than older comparisons suggest. Vizard's established AI Studio is the more mature specialist for long-to-short repurposing and distribution. It offers purpose-built clip preferences, transcript and timeline correction, social-account connections, automatic scheduling, team brand controls, and a public API. If the recurring job is turning each podcast or webinar into a measured flow of social excerpts, Vizard has the clearer operational fit.

Vizard Agent is the closer direct comparison. It can accept an outcome rather than a chain of editing commands, begin from an idea or mixed context, create missing assets, and revise a broader finished video. Valmera is centered on existing footage and a versioned edit decision list: the agent interprets the brief, rewrites the EDL, uses media tools, renders the result, and accepts further direction against the same project. That makes Valmera legible when source footage is irreplaceable and the exact cut is the core asset.

The products meter work differently. Vizard's classic v1 workflow maps cleanly to source minutes, v2 uses a 1.25 multiplier, and enhancement uses output seconds. Vizard Agent's public unsigned price is incomplete. Valmera exposes one server-reported credit balance while charging agent turns according to model work rather than a public per-source-minute rule. A fair cost comparison must measure a completed approved deliverable—including discarded clips, revisions, enhancement, human correction, and publishing—not compare nominal credits.

Data posture may be decisive. Vizard's dated public Terms permit meaningful machine-learning use and do not clearly promise output ownership or commercial use. Valmera publishes its own terms and privacy policy at valmera.io/legal; review both with the same rigor. For confidential footage, obtain written commitments from whichever provider is selected and keep independent source and export archives.

Use the same evaluation pack: one long interview, one visually led recording, a written brief, target aspect ratios, caption style, prohibited claims, brand assets, required final duration, and an approval rubric. Track usable first-pass minutes, boundary repairs, factual errors, framing mistakes, transcript corrections, revision rounds, elapsed time, credit spend, and export quality. Choose Vizard when its clip-to-publish throughput wins; choose Valmera when an open-ended footage edit directed through conversation is the better working model.

Vizard Agent now looks much closer to Valmera's outcome-level interaction than its original clipping product did. Both accept a brief and revise a rendered result. Vizard's public templates extend much further into research, generation, product feeds, synthetic voice, AI films, and mixed-media creation. Valmera remains more explicitly centered on indexing real uploaded footage, rewriting a versioned edit decision list, and executing deterministic media tools around that source.

Metering should be normalized by accepted outcome. Vizard Agent publishes median credits but not signed-out dollar economics; Valmera exposes a unified credit balance while agent turns vary with model cost. Run the same owned-footage brief, record all credits and elapsed time, and compare approved seconds, source-faithfulness, number of correction turns, manual rescue, export specs, and total paid cost. Do not compare 242 Agent credits with a Valmera balance as if the units were equivalent.

Vizard AI Studio remains the stronger mature option for scheduled high-volume clipping because its one-minute v1 meter, social calendar, Auto Schedule, six platform families, API, templates, and Business workspace already form an operational system. Valmera is stronger when the desired edit is not reducible to candidate clips and social distribution. Vizard Agent is the broader experimental competitor when the brief includes missing-asset generation or research as well as editing.

For the full feature-by-feature comparison, read Valmera vs Vizard.

Primary Sources Checked

Frequently Asked Questions

Vizard is an AI video platform with two major surfaces. AI Studio turns long speech-led recordings into edited social clips and publishes them. Vizard Agent accepts broader outcome-level briefs and can research, script, generate, edit, localize, and revise a finished video.
It is worth testing for repeatable podcast, webinar, interview, and event repurposing when captions, branding, scheduling, publishing, and API automation matter. Agent is promising for broader jobs, but its public price is incomplete and its output still needs human direction.
Yes. Free includes 60 credits every month, the editor, one social connection, and 720p export. It also adds a watermark, limits uploads to 60 minutes and 1GB, caps exports at ten minutes, and stores projects for only three days.
The public pricing page says Free receives 60 credits per month. Free project retention is separate and much shorter: three days. Do not interpret a monthly credit refresh as a month of cloud storage.
Yes. Free output carries Vizard branding. Creator and Business remove the watermark. Evaluate on the tier that matches the intended resolution, retention, templates, social connections, and collaboration rather than assuming only the logo changes.
The current starting month-to-month price is $29 for 600 credits. Larger credit configurations are available interactively. Creator also includes 4K, no watermark, 100GB, six social accounts, scheduling, templates, transcripts, and API access.
Creator currently charges $174 for the year, equivalent to $14.50 per month, with a starting 7,200-credit annual configuration. That saves $174 versus 12 current $29 payments, or 50%, before tax and optional larger allowances.
Business currently starts at $39 month-to-month for 600 credits. It adds a shared workspace, Brand Kit, custom fonts, project sharing, faster AI, 20 social accounts, unlimited storage while subscribed, and higher API limits.
Business currently charges $234 for the year, equivalent to $19.50 per month, and starts at 7,200 annual credits. That saves $234 versus 12 current $39 payments, excluding taxes and extra seats.
An extra Business seat is currently $10 on monthly billing or $60 per year, equivalent to $5 per month. Annual seats added mid-term can be prorated. Public material does not say that a seat adds credits.
No. The displayed $14.50 and $19.50 figures are monthly equivalents. The actual starting annual charges are $174 for Creator and $234 for Business, paid for the year.
Yes. The pricing page says Creator and Business renew automatically at the end of the billing cycle unless terminated. Cancellation takes effect at cycle end, so preserve the confirmation and plan any data migration before expiry.
Vizard's pricing FAQ lists credit cards, PayPal, Stripe, and bank transfers. It says Stripe processes payment details and Vizard does not access the customer's card data directly.
For ordinary v1 clipping or transcription, one credit covers one processed source minute, and using both in the same project is not charged twice. v2 clipping and AI Enhance use different rates.
A 60-minute source uses 60 credits with ordinary v1 clipping or transcription. API clipping model v2 uses 1.25 times the v1 rate, so the same source costs 75 credits.
Not when both happen in the same project. Current credit guidance says v1 clipping or transcription costs one credit per source minute and a project using both is not charged twice for that processing.
v2 is an API-selectable clipping model that Vizard describes as more curated and coherent, with fewer higher-quality candidates. It costs 1.25 times v1. v1 remains the default and favors volume and lower cost.
Yes, temporarily. Each monthly release is documented as valid for two months, and each annual release for 13 months. Rollover credits are spent first, but they do not remain forever.
Increasing a usage tier adds credits immediately and charges the difference. An annual upgrade starts a new cycle. Moving from Creator to Business also starts a new cycle, and Vizard says unused Creator credits do not carry over.
The current live prices are exactly 50% below 12 monthly payments, matching the pricing page. An older upgrade article says about 42%. Use the exact checkout total and preserve the invoice rather than relying on the stale percentage.
Vizard Agent is an outcome-directed video system launched August 10, 2026. It can start from an idea, brief, article, product page, document, image, existing video, or raw footage; perform the needed work; and revise from feedback.
The unsigned public Agent page shows Flash, Pro, Max, and Ultra tiers but does not publish their prices, allowances, or revision rates. Inspect the signed-in meter and checkout; do not infer Agent cost from Creator or Business clipping credits.
Public material checked for this review does not establish that the Creator or Business upload-credit pool pays for Agent work. Treat Agent as a separate buying surface until the signed-in account states the exact entitlement and meter.
Vizard lists ads, product launches, educational content, social videos, YouTube work, repurposing, localization, and projects mixing real and generated media. Its public templates also include tutorials, vlogs, trailers, multicam, music videos, AI films, and documentaries.
Yes. The launch material says an idea or creative brief can be enough. Agent may research, create an angle and script, generate missing visuals, assemble the video, and revise it based on the requested outcome.
Yes. Vizard says Agent can work from raw footage or an existing video, restructure content, add or create missing scenes, adapt it for a platform or audience, and carry the project through multiple revisions.
No. Vizard explicitly says one prompt will not always produce the final result. Long jobs can lose coherence, style and tone can drift, and the agent can make a reasonable choice that is not the choice the user wanted.
Yes. Deleting transcript text cuts the linked video, editing text corrects subtitles, and struck-through content can be restored to extend a clip. Users can also work in the timeline or manually select from the original video.
Yes. Editing a transcript word corrects the subtitle. Deleting the word removes the corresponding video segment. That distinction lets users repair captions without accidentally changing the spoken recording.
Yes. When AI does not find the desired moment, the original-video editor provides a manual path. This is also useful when an AI clip begins mid-sentence, lacks context, or ends awkwardly.
There is no fixed guarantee. Count depends on source structure, duration preferences, prompt or keyword filtering, and model judgment. Shorter requested durations often produce more candidates, while narrow searches produce fewer.
Speech-led videos such as podcasts, interviews, webinars, tutorials, talks, and coaching recordings are the natural fit. Vizard recommends sources longer than 15 minutes for AI clipping. Visually led media should be tested separately.
Yes. Add Scenes can insert material from the original transcript, local uploads, or Brand Kit into the editor as intros, outros, inserts, or b-roll. Scenes can be added from the timeline or transcript.
Yes. Its current speaker identification uses audio and video context, defaults to Active Speaker, and can keep selected speakers visible. The crop remains an inference, so reactions, products, slides, and fast exchanges need review.
Yes. Auto Censorship detects selected sensitive words in audio and can silence them or add a beep. Users can customize the word list. It does not replace broader legal, platform, visual, or contextual review.
The public API supports vertical 9:16, square 1:1, portrait 4:5, and horizontal 16:9. Clip preferences in the web workflow emphasize 9:16 and 1:1. Confirm the exact editor export needed for each channel.
Creator and Business advertise export up to 4K. A 4K setting increases output dimensions but does not recreate missing source detail. Cropped, compressed, low-resolution, or not-yet-fully-processed online sources may still look soft.
AI Enhance is a paid super-resolution pass applied to the final 1080p, 2K, or 4K export. It attempts to reconstruct detail after low resolution, compression, cropping, or reframing without changing the original media or timeline.
The current rates are about one credit per output second at 1080p, 1.5 at 2K, and two at 4K, with a one-credit minimum. A 60-second result therefore costs about 60, 90, or 120 credits.
Yes. The help center says credits used for an enhancement that fails to process are refunded automatically. A technically completed output that merely looks disappointing is not described as an automatic refund.
It can output 4K and reconstruct plausible details, but Vizard warns that extreme upscaling, blur, noise, compression, defocus, and fast motion limit the result. Test a short representative clip before spending credits on a long export.
Three days. The limit is short enough that evaluation should be scheduled: process, inspect, correct, export, and archive promptly rather than assuming the project will still exist next week.
Creator includes 100GB of permanent workspace storage. New work beyond the quota can become temporary and normally expires after 28 days. Original media after transcoding, uploaded assets, and Brand Kit files count.
Business is advertised as unlimited while the subscription remains active. When the workspace drops to Free, projects enter a 28-day expiry window and non-admin members are removed, so unlimited does not mean permanent after cancellation.
The workspace downgrades, non-admin members are removed, and projects are retained temporarily with a 28-day expiry period. Export media, transcripts, brand assets, schedules, and records before the paid term ends.
Vizard supports anyone-with-link and invited-only access for exported videos. Business adds project sharing. Current help also says workspace members can override link settings, so access governance should include workspace roles and link review.
Vizard currently documents TikTok, YouTube, LinkedIn, X, Instagram, and Facebook Pages. Free connects one TikTok account, Creator up to six social accounts, and Business up to 20.
Yes. Creator and Business include scheduling. The calendar supports multiple accounts, generated descriptions and hashtags, tone or voice changes, and later editing or deletion.
Auto Schedule can select clips above a virality threshold and publish them using a daily cap, schedule pattern, active days, date range, and time window. AI also drafts post copy. Users can revise or reschedule afterward.
The current help center says up to ten clips per day can be scheduled from a project. Platform rules, connected-account state, approval policy, and subscription expiry can still affect delivery.
Vizard says posts scheduled within the 28-day post-expiry project-retention window can publish, while posts beyond that window fail. Migrate the future calendar and archive exports before canceling.
Yes. Vizard documents direct Google Drive export. Connection permissions, destination organization, file naming, storage quota, and independent retention still need to be managed by the customer.
Yes. Current guidance says a paid plan is required for an API key. Jobs consume the workspace's upload-credit pool, can be tracked through polling or webhooks, and appear in Vizard for editing and publishing.
The quickstart says paid is required to obtain a key, even though the pricing matrix visually shows a Free rate-limit row. Treat supported API access as paid-only unless an actual Free workspace exposes and authorizes a key.
The public plan matrix lists Creator at three requests per minute and 20 per hour and Business at ten per minute and 60 per hour. It also displays Free at one per minute and ten per hour despite the paid-key requirement.
The submit reference lists remote files, YouTube, Google Drive, Vimeo, StreamYard, TikTok, X, Twitch, Loom, Facebook, LinkedIn, Dropbox, and Instagram. The customer still needs authorization to access and republish the media.
The current submit endpoint lists MP4, 3GP, AVI, and MOV for remote-file submissions. Reachability, codec behavior, resolution, duration, file size, and URL expiry should be tested with the production pipeline.
Vizard says minutes to tens of minutes and notes that 4K can take longer. Its example of a 30-minute talking-head video returning 20-plus clips in about ten minutes is illustrative, not a guaranteed SLA.
Yes. Vizard documents both polling and webhook completion. For polling, the quickstart recommends an interval around 30 seconds. Production integrations should handle retries, duplicate notifications, terminal failures, and expired source URLs.
Yes. The submit reference exposes removeSilenceSwitch for silent gaps and fillers such as 'um' and 'uh.' It also exposes subtitles, headlines, emoji, highlights, keywords, templates, ratios, durations, and automatic b-roll options.
The current help article explicitly lists 40 transcription languages, including Arabic, English, Cantonese, simplified and traditional Chinese, Hindi, Japanese, Korean, Spanish, Turkish, Ukrainian, and Vietnamese.
The same current help article enumerates 128 translation languages. That is separate from its 40 transcription languages. Verify the exact pair, names, terminology, script, line breaking, and cultural meaning.
Yes. Paid plans can download the original video's full transcript as SRT or TXT, and users can copy text from the editor. Free's pricing row shows TXT while paid tiers show TXT and SRT.
Yes. Business provides a shared workspace, member invitations, Brand Kit, custom fonts, project sharing, 20 social accounts, unlimited storage while subscribed, and higher API limits. Invitations are documented as desktop-only.
The current public sources specify a seat price but do not state that each seat adds credits. Plan on the selected workspace allowance being the relevant pool unless checkout or a written quote says otherwise.
Possibly for a double or incorrect charge, or within seven days when the service remains unused. Uploading data or using tools counts as use, support decides eligibility, fees may be deducted, and approved refunds can take five to ten business days.
Open the workspace subscription settings and complete cancellation. The plan remains active through the paid cycle and then stops renewing. Export assets and account records first because storage, members, schedules, credits, and API access change at expiry.
The self-serve Terms do not offer a simple no-training promise. They permit machine-learning use created while processing user content and usage data to test, tune, optimize, validate, and enhance analytics, models, and algorithms.
The public Terms do not clearly say that Vizard owns generated output, but they also do not clearly promise customer ownership. They grant a broad service-operation license over user content. Obtain written clarification for commercially important or sensitive work.
A clear general commercial-use grant was not found in the public Terms checked for this review. That does not prove commercial use is forbidden. It means the contract should be clarified, and all input and generated-asset rights still need review.
This review did not find a current first-party SOC 2 claim in the public sources checked. The Privacy Policy mentions SSL, digital signatures, and PCI-compliant payment processors. Procurement should request current reports, scope, DPA, and subprocessors.
Both specialize in long-to-short repurposing. Vizard emphasizes a connected editor, social scheduling, six named networks, API controls, Business workspace features, and now a general Agent. Compare accepted clips, correction time, publishing, pricing, rights, and export needs.
Vizard automates transcript-led selection, reframing, captions, b-roll, formatting, and distribution. It is faster for repeatable spoken-media clips but offers less direct depth than a mature NLE for complex timelines, keyframes, color, audio mixing, and interchange.
Vizard has a mature clip-to-publish platform and a separate broad Agent. Valmera is centered on directing an agent over uploaded footage through a versioned edit decision list and rendered revisions. Choose based on repurposing throughput versus footage-first open-ended editing.
The main disadvantages are unpublished Agent pricing, multiple credit rates, expiring rollover, lost Creator credits on Business upgrade, three-day Free retention, storage and schedule clocks after downgrade, probabilistic edits, public document conflicts, dated legal policies, and unclear output rights.
Yes according to Vizard's current use-case guides, which call Flash the free tier for quick, simple edits. The signed-out site does not publish its included credit allowance, expiry, queue, or other entitlement details.
The live page labels them Flash Junior, Pro Skilled, Max Senior, and Ultra Master. Those names express intended capability. Public dollar prices and detailed allowances for Pro, Max, and Ultra are not shown.
Vizard-authored guides report cross-project median use of 47 credits on Flash, 55 on Pro, 242 on Max, and 263 on Ultra. A specific job can differ substantially.
The published medians may combine different job types and complexity, not controlled repeats of the same brief. Max's 242-credit median versus Pro's 55 cannot prove a 4.4× same-job multiplier.
It cannot be calculated from the signed-out public material. Vizard does not publish Agent plan prices, included balances, top-up price, or expiry there. Inspect the actual account and checkout.
Vizard describes comparable jobs with medians from 28 to 38 minutes. Its measured advertising and promotional category reports a 28-minute median and 15–47-minute middle-half range. These are vendor benchmarks, not an SLA.
The trailer guide reports 82 credits median for advertising and promotional work, with the middle half from 48 to 251. That is a vendor category statistic, not a guaranteed trailer quote.
Its guides say the project keeps footage and built assets, so a revision is charged for the new change rather than another source upload. The exact credit cost of each type of revision is not publicly tabulated.
Likely because it is a new requested change, but no public rate maps aspect-ratio variants to credits. Check the posted estimate or ledger before authorizing a batch of platform versions.
The live signed-out page showed 22 starting templates and 24 separate showcase cases on August 26, 2026. Templates range from ads and tutorials to multicam, trailers, music videos, AI films, mini-docs, and documentaries.
Vizard describes Flash as free for quick simple edits and often recommends Max in its own guides for publishable or customer-facing work. Run the same real brief across tiers; vendor recommendations are not independent proof.
Vizard says yes: upload every angle and the audio, and Agent aligns them, cuts to the speaker, and tightens dead air. The cameras need shared reference audio for sync.
A common audio reference across every camera or recorder. That correlation is created during the shoot. Missing scratch audio, drift, dropped frames, or isolated files can prevent reliable automatic alignment.
Yes. Vizard's own guide tells users to watch back and correct cuts that land on the wrong face. Reactions, interruptions, off-camera speakers, and overlapping dialogue need human editorial judgment.
Yes. Its guide describes finding the phrase by word-level transcript, cloning the speaker's replacement audio, lip-syncing the visible mouth, and joining at a quiet waveform point.
A much longer phrase may not fit, noisy or reverberant audio clones poorly, a close face exposes lip-sync errors, on-screen text must be corrected separately, and the user needs permission and claim approval.
It should. Vizard's guide says the source must be your video or used with permission. Preserve the original, document the synthetic span, obtain speaker approval, and disclose where context or law requires it.
Yes. One official workflow says it reads the site, extracts brand assets, inspects the underlying product feed, collects catalog items and prices, and builds an ad from the selected goal.
No guarantee. It reads what the shop publishes, which can be stale, regional, variant-specific, unavailable, or changed after extraction. Recheck price, stock, currency, tax, shipping, dates, and disclaimers before launch.
Vizard says Agent tries several public routes and advises supplying products and images directly when a hard block remains. Do not bypass access controls or use a catalog you lack permission to republish.
Yes. It can write narration, generate matching visuals, add voiceover, captions, and music from a topic, audience, question, length, and any definitions supplied by the user.
Read the complete narration, map each material claim to an authoritative source, check company-specific definitions, numbers, dates, captions, and visuals, then render again. Polished delivery can conceal confident errors.
Yes. Its guide says Agent can watch the footage, find usable moments, build an arc, tighten pauses, caption, and reframe. Give it the story, ending, must-use or forbidden material, length, and destination.
Yes. Vizard says it selects shots, builds escalation, cuts to music, and adds title cards from uploaded footage. Specify the mood, purpose, length, ending card, and anything that must remain unseen.
Yes. Upload the clips and photos, name the event, intended feeling, length, people or quotes that matter, and final call to action. Review every face, name, quote, consent, and brand asset.
It can study pacing, shot order, camera moves, and beat structure, then use the customer's footage and separately licensed music. Do not copy the reference footage, track, protected trade dress, or distinctive campaign identity.
Yes. Vizard says it can turn a premise into generated shots, characters, dialogue voices, music, grading, and a finished cut. Story, identity consistency, rights, disclosure, and audience suitability still need review.
It can deliberately generate handheld first-person footage with imperfect framing and sound. Vizard itself warns to label mistakable synthetic footage and not depict real people or actual incidents as fabricated evidence.
There is no single public contract. Vizard-authored examples include H.264/AAC files at 720×1280 or 1080×1920 and 24, 30, or 60 fps. Those individual runs do not guarantee every tier or job.
At least one official meme example was probed at 1080×1920 and 60 fps. Treat that as evidence from one delivered file, not a universal export entitlement. Test the exact template and tier.
Yes. Its guide describes transcribing with word timings, examining framing, searching per subject, selecting candidates, and inserting cutaways while preserving the original talk. Licensed owned footage should be preferred.
It may search for stock and material described as reusable, but Vizard warns that copyright follows the footage. Preserve an actual license and use customer-owned or clearly licensed media whenever possible.
The editing capability does not answer the rights question. Vizard itself says lawful publication and monetization must be decided before spending credits. Obtain jurisdiction-specific advice for important uses.
No. They are useful first-party statistics published inside Vizard's own guide pages. The company does not disclose a complete sample, methodology, queue conditions, approval criteria, or controlled same-brief tier comparison.
Capture tier, prompt, source duration and size, start and completion time, credits before and after, generated assets, revisions, output specs, factual errors, rights issues, human correction time, and approved result.
Ask for plan price, included and top-up credits, expiry, reservation timing, failure refunds, revision billing, project retention, model and search providers, commercial rights, training use, deletion, export specs, and support terms.
It is broader, not universally better. Agent handles outcome-level research, generation, mixed media, and long workflows. AI Studio has clearer clipping economics, transcript correction, templates, API, and social operations for repeatable repurposing.
Vizard Agent is broader for research, generation, URL ads, synthetic media, and many templates. Valmera is more explicitly footage-first, with indexed uploads, versioned edit decisions, media-tool execution, rendered previews, and conversational revision.
Agent now exposes useful credit benchmarks but still hides dollar pricing and allowances. Other risks are tier-selection ambiguity, undefined revision rates, variable output specs, synthetic-speech consent, source rights, realistic-media disclosure, and 2022 legal terms.

Want the AI to Perform the Edit?

Upload your footage to Valmera, describe the outcome, and refine the rendered edit through conversation.

Edit with Valmera →
See pricing →

Related Articles

Valmera vs Vizard
Compare footage-first agentic editing with Vizard AI Studio, Agent, API, and publishing.
Turn Long Videos into Clips
A practical workflow for selection, context, reframing, captions, approval, and publishing.
Best AI Clipping Tools
Compare long-to-short specialists, agentic editors, and general video systems.
AI Video Editor vs Generator
Choose based on whether footage, information, or a new scene is the source of value.