Skip to content

Video AI apps

AI video generation and editing — text-to-video models, avatar studios, and automated post-production.

41 apps · researched & kept current by Claude Code

Filter & search these 41 apps
  • View Vozo details
    TranslationFREEMIUM

    Vozo

    Vozo

    AI video translation with dubbing, lip sync, and voice cloning.

    Vozo is an AI video platform that translates and dubs videos into other languages while preserving the original speaker's voice. It combines automatic transcription, translation, voice cloning, lip-sync, subtitle generation, and on-screen text translation in a single editor, plus an API for higher-volume workflows. It supports 160+ languages and targets creators, marketers, and localization teams.

    160+ languages supported
    Lip sync adds to credit usage
    • video-translation
    • dubbing
    • voice-cloning
    • lip-sync
  • View Rask AI details
    TranslationFREEMIUM

    Rask AI

    Brask Inc.

    AI video localization and dubbing with voice cloning.

    Rask AI is a video localization and dubbing platform that translates and re-voices video and audio into 130+ languages. It generates synthetic voiceovers with voice cloning in 32 languages, adds automatic lip-sync, detects multiple speakers, and produces translated captions and subtitles. Aimed at creators and marketing teams adapting content for global audiences, it runs in the browser and via API.

    Voice cloning in 32 languages
    Cloud-only, minute/credit-based pricing
    • dubbing
    • video-localization
    • voice-cloning
    • subtitles
  • View Wan details
    VideoFREEMIUMOpen core

    Wan

    Alibaba (Tongyi Lab)

    Open-source text- and image-to-video generation from Alibaba.

    Wan is an open-source family of video generation models from Alibaba's Tongyi Lab, covering text-to-video, image-to-video, and image generation and editing. The open weights can be self-hosted or used free on the wan.video site. Successive versions, including Wan 2.2's mixture-of-experts architecture, have topped open-video benchmarks.

    Public weights, training and inference code
    Self-hosting needs strong GPUs
    • video-generation
    • open-source
    • text-to-video
    • image-to-video
  • View NightCafe details
    ImageFREEMIUM

    NightCafe

    NightCafe Studio

    Multi-model AI art generator wrapped in a creative community.

    NightCafe is a browser-based AI art platform that generates and edits images from text prompts across many engines in one place — FLUX, Stable Diffusion, DALL·E, Google Imagen, Ideogram and more — plus image-to-video and style-transfer tools. It runs on daily free credits and a credit-pack/PRO model, and is built around an active community with shared galleries, chat rooms, and a daily AI Art Challenge.

    Many models accessible in one place
    Web/PWA only — no native desktop app
    • image-generation
    • ai-art
    • community
    • style-transfer
    • +1
  • View Perso AI details
    TranslationFREEMIUM

    Perso AI

    ESTsoft

    AI video dubbing and localization with voice cloning and lip-sync.

    A video localization platform that translates and dubs footage across 99+ languages with voice cloning and AI lip-sync. The pipeline covers audio separation, speech-to-text, subtitle editing, and per-speaker voices, and the suite extends to interactive AI avatars and an AI human studio.

    99+ languages supported
    Credit-based plans cap monthly output
    • dubbing
    • localization
    • voice-cloning
    • lip-sync
  • View Magic Hour details
    VideoFREEMIUM

    Magic Hour

    Magic Hour AI, Inc.

    AI video generator and creative suite with a developer API.

    Magic Hour is an AI creative platform for video, image and audio generation aimed at creators, marketers and developers. It bundles 100+ tools — text-to-video, image-to-video, face swap, lip sync, talking photo, image generation and upscaling, plus voice cloning and music generation — behind one shared credit balance. Beyond the browser app it exposes a REST API with official Python, Node.js, Go and Rust SDKs.

    One credit balance across all tools
    Breadth over best-in-class depth per tool
    • ai-video
    • image-generation
    • video-editing
    • api
    • +1
  • View Odyssey details
    VideoFREE

    Odyssey

    Odyssey

    Real-time world model that generates interactive, explorable AI video.

    Odyssey is a world-model startup whose AI generates interactive video you can steer in real time — explorable, 3D-like scenes streamed frame by frame without a game engine. Its current model, Odyssey-2, turns a text prompt or image into a multi-minute simulation you navigate as it is generated. It is positioned for film, gaming, and interactive media, and is available as a free research-preview experience.

    Real-time interactive generation
    Research preview, not production-ready
    • world-model
    • interactive-video
    • real-time
    • simulation
  • View Genmo details
    VideoFREEMIUMOpen core

    Genmo

    Genmo

    Open-source text-to-video generation, with a free hosted playground.

    AI video startup behind Mochi 1, an open-source text-to-video model built on an Asymmetric Diffusion Transformer (AsymmDiT) architecture. The weights and code are on GitHub and Hugging Face for local or self-hosted use, while genmo.ai hosts a free playground to generate clips from prompts in the browser.

    Downloadable open weights
    Local runs need heavy GPU memory
    • text-to-video
    • open-source
    • video-generation
    • diffusion
  • View invideo AI details
    VideoFREEMIUM

    invideo AI

    invideo

    Describe a video in plain text and get a finished, edited cut.

    An AI video creation platform that turns a text prompt into a publish-ready video — selecting stock footage, generating a voiceover, adding subtitles and music, and assembling the edit. Its newer Agent One flow routes across models like Sora, Veo, and Kling and keeps project context so you revise the cut by chatting rather than editing a timeline. Built for marketers, social creators, and faceless-channel video.

    Prompt-to-finished-video in one pass
    Credit limits on higher-quality output
    • text-to-video
    • video-editing
    • voiceover
    • social-video
  • View Vizard details
    VideoFREEMIUM

    Vizard

    Vizard

    Turns long videos into short, captioned clips for social in one click.

    Vizard is an AI video tool that automatically slices long-form videos — podcasts, webinars, interviews, livestreams — into short, social-ready vertical clips. It transcribes the source, finds the highlight moments, adds auto-captions, and reframes to keep the speaker centered for TikTok, Reels, and YouTube Shorts. A free plan lets you try clipping without signing up, with paid Creator and Business tiers for more exports and longer videos.

    Auto-finds highlight moments
    Cloud-only, no offline editing
    • video-clips
    • short-form
    • captions
    • repurposing
    • +1
  • View OpenArt details
    ImageFREEMIUM

    OpenArt

    OpenArt

    All-in-one AI studio for images, video, and characters.

    OpenArt is a consumer AI creation platform for generating and editing images and video, building consistent AI characters, and producing multi-scene visual stories. Rather than one proprietary model, it routes across many engines — Stable Diffusion, FLUX, DALL-E, and thousands of community-trained models — and adds LoRA training, canvas editing, and a model-and-prompt marketplace. It began as one of the largest community hubs for Stable Diffusion.

    Access to many models in one place
    Web-only
    • image-generation
    • video-generation
    • ai-characters
    • storytelling
    • +1
  • View Viggle details
    VideoFREEMIUM

    Viggle

    Viggle AI

    Animate any character by mapping a motion video onto a still image.

    Viggle is an AI video tool that brings still images to life — drop in a character photo and a motion reference clip, and it generates a video of that character performing the movement. It powers viral memes, VTubing, and previz, with web and mobile apps and a near real-time webcam-driven Live mode.

    Maps reference motion onto any character
    Narrow focus on character motion, not general video
    • video-generation
    • character-animation
    • motion-transfer
    • memes
  • View Argil details
    VideoPAID

    Argil

    Argil

    Generate avatar videos from a script — clone yourself or pick a ready-made avatar.

    Argil is an AI video platform that turns scripts and articles into talking-avatar videos. You can clone your own likeness from a single photo and a short voice recording, or use one of 100+ ready-made avatars, then let Argil auto-add captions, B-roll, backgrounds, and transitions. It targets creators and marketers producing short-form and faceless video at scale, with output rendered in roughly a couple of minutes.

    Self-clone from one photo + short audio
    No free tier (5-day trial only)
    • ai-avatar
    • video-generation
    • ugc
    • faceless-video
    • +1
  • View Dreamina details
    ImageFREEMIUM

    Dreamina

    ByteDance

    All-in-one AI suite for image and video generation.

    Dreamina is ByteDance's creative platform for turning text and reference images into images and video, powered by its Seedream image models and Seedance video models. It offers text-to-image, image-to-image, and a multi-layer canvas with inpainting, expansion, and removal, plus text- and image-to-video — and integrates with the CapCut video editor. Use cases span character design, product photography, marketing assets, and social content.

    Seedream 4.0 ranks among top image models
    Video rollout limited by region
    • image-generation
    • video-generation
    • creative-suite
    • bytedance
    • +1
  • View MiniMax details
    InferenceFREEMIUM

    MiniMax

    MiniMax

    Multimodal foundation models and developer API for text, code, video, speech, and music.

    MiniMax is a Shanghai foundation-model lab whose platform serves its own model family through a developer API and agent app: the M-series LLMs (M2/M3) built for coding and agentic workflows with up to a 1M-token context, the Hailuo video models, and MiniMax Speech and Music. Developers get chat completions, text-to-speech, and text-to-video on token-based pricing, with a free agent tier for getting started.

    Coding- and agent-tuned M-series models
    China-based; data-residency considerations
    • foundation-models
    • llm
    • api
    • video
    • +1
  • View LTX Studio details
    VideoFREEMIUM

    LTX Studio

    Lightricks

    An AI platform for end-to-end video production.

    LTX Studio is Lightricks' AI filmmaking suite that turns a script or prompt into characters, storyboards, and edited video sequences. It pairs generation (text-to-video, image-to-video, video-to-video) with director-style controls — camera framing, keyframe animation, character consistency, and a timeline editor — so creators can shape a full production, not just one clip. It runs on Lightricks' own LTX models and also integrates third-party models for generation and audio.

    End-to-end: script → storyboard → edited film
    The hosted platform itself isn't open source
    • video
    • text-to-video
    • filmmaking
    • storyboard
    • +1
  • View Decart details
    GamingPAID

    Decart

    Decart

    Real-time generative world and video models (Oasis, Lucy).

    Decart is an AI lab building real-time generative world and video models, including Oasis (a fully AI-generated, playable Minecraft-style open world) and Lucy (real-time video transformation used in gaming, ads, and virtual try-on). Its Oasis 3 is positioned as an API-accessible world model for Physical AI, with a low-latency inference stack. Products are accessed via the platform API and interactive demos.

    Real-time world-model generation
    Sales-led, API-first, no free tier
    • gaming
    • world-models
    • real-time-video
    • generative
  • View VEED.IO details
    VideoFREEMIUM

    VEED.IO

    VEED

    Browser-based AI video editor with subtitles, dubbing, and avatars.

    An all-in-one online video creation platform: a real timeline editor paired with AI tools for auto-subtitles, translation and dubbing, text-to-speech, AI avatars, background and noise removal, eye-contact correction, and text-to-video generation. Runs entirely in the browser, aimed at marketers, creators, and teams making social and brand video.

    Full editor plus AI tools in the browser
    Free tier watermarks and caps exports
    • video-editing
    • subtitles
    • dubbing
    • avatars
    • +1
  • View Topview details
    MarketingFREEMIUM

    Topview

    Topview

    Collaborative AI workspace for video ads, UGC, and avatar content.

    Topview is an AI video workspace for marketers and ecommerce brands: paste an Amazon, Shopify, or TikTok product link and it assembles complete video ads with AI avatars that can hold or wear the product. A database of 10M+ high-performing ads powers its recreator tool, and Topview 4.0 turns the editor into a collaborative, multiplayer content workspace.

    URL-to-video from product links
    Free tier watermarked, non-commercial
    • video-ads
    • ugc
    • ai-avatars
    • ecommerce
  • View FLORA details
    DesignFREEMIUM

    FLORA

    FLORA

    Node-based creative canvas chaining AI models into team workflows.

    An infinite-canvas creative environment where designers chain image, video, and text models into reusable node-based workflows. One subscription covers 50+ models — Veo, FLUX, Kling, Nano Banana, GPT, and more — and the built-in FAUNA agent helps assemble and iterate on pipelines. Pitched at professional teams who want repeatable, taste-driven output instead of one-off prompting.

    Many image/video/text models, one plan
    Node workflow has a learning curve
    • creative-canvas
    • node-workflow
    • image-gen
    • video-gen
  • View Moonvalley details
    VideoPAID

    Moonvalley

    Moonvalley AI

    Cinematic AI video generation from a fully licensed model.

    An AI research company whose Marey model generates text-to-video and image-to-video with cinematography-grade controls — pose transfer, camera moves, motion transfer, and trajectory control. Trained exclusively on owned and licensed high-resolution footage, it targets filmmakers and studios that need commercially safe output. Available as a web app, with API access via Fal and ComfyUI integrations.

    Trained only on licensed footage
    No free plan or trial
    • video-generation
    • filmmaking
    • licensed-data
    • image-to-video
  • View Vidu details
    VideoFREEMIUM

    Vidu

    Shengshu Technology

    Text-, image-, and reference-to-video generation at up to 1080p.

    An AI video generator from Tsinghua-spinout Shengshu Technology. Turns text prompts, still images, or reference images into short high-resolution clips, with text-to-video, image-to-video, and reference-to-video modes. Known for fast, low-cost generation and keeping multiple uploaded subjects consistent across a shot.

    Text-, image- and reference-to-video
    Short clip lengths
    • video-generation
    • text-to-video
    • image-to-video
    • reference-to-video
  • View PixVerse details
    VideoFREEMIUM

    PixVerse

    PixVerse

    Fast AI video from text or photos, tuned for stylized and social content.

    An AI video creation studio that turns text prompts or images into short clips in seconds. PixVerse is known for speed and for handling stylized, anime, and effects-heavy output rather than only photorealism, with a web app, mobile apps, and a developer API platform.

    Fast generation
    Less photoreal than frontier rivals
    • video-generation
    • text-to-video
    • image-to-video
    • anime
  • View AdCreative.ai details
    MarketingPAID

    AdCreative.ai

    Appier

    Generates conversion-focused ad creatives, video, and product shots with AI scoring.

    AdCreative.ai generates ad banners, copy, product photography, and video ads for platforms like Google and Meta without design skills. Its Creative Scoring AI predicts which creatives will perform, and it pulls competitor and audience insights to inform generation.

    Generates banners, copy, product shots, video
    No free tier — paid only
    • ad-creative
    • advertising
    • design
    • video-ads
    • +1
  • View Hedra details
    VideoFREEMIUM

    Hedra

    Hedra

    Turn a photo and voice into talking, expressive characters.

    Hedra generates lip-synced, expressive talking-character video from a single image plus audio or a script. Its Character-3 model handles facial performance and emotion, and a Live Avatars tier streams those characters in real time for conversational AI agents.

    Phoneme-accurate lip-sync from one image
    Maxes out at 720p
    • talking-avatar
    • lip-sync
    • character-video
  • View OpusClip details
    VideoFREEMIUM

    OpusClip

    OpusClip

    Turns long videos into viral short clips with AI captions and auto-reframing.

    Repurposes long-form videos and podcasts into short, vertical clips ready for TikTok, Reels, and Shorts. The AI finds the most engaging moments, adds animated captions, reframes to keep speakers centered, and scores each clip's virality. Credits are billed per minute of source video, not per clip produced.

    Strong auto-highlight detection
    Billed per source-minute, not per clip
    • video
    • short-form
    • clipping
    • repurposing
  • View Tavus details
    VideoFREEMIUM

    Tavus

    Tavus

    Real-time conversational video AI and digital human replicas.

    A developer platform for building face-to-face AI agents that see, listen, and respond in live video through its Conversational Video Interface (CVI). It also generates personalized videos at scale from digital replicas of a real person. Built on Tavus's own models — Phoenix for rendering, Raven for perception, and Sparrow for conversational timing — with the ability to plug in custom LLMs and text-to-speech.

    Live face-to-face AI video
    Developer-first, not no-code
    • video
    • avatars
    • digital-twin
    • conversational
  • View Adobe Firefly details
    ImageFREEMIUM

    Adobe Firefly

    Adobe

    Commercially-safe generative AI for images, video, and design.

    Adobe's generative AI for creators — text-to-image, generative fill, and increasingly video, available as a standalone web app, mobile apps, and inside Creative Cloud tools like Photoshop. Built around Adobe's own Firefly models trained on licensed and public-domain content, with select third-party models now integrated. A free tier ships monthly credits; paid plans add more credits and IP indemnification.

    Built into Photoshop and Creative Cloud
    Raw image quality trails Midjourney/FLUX
    • image-gen
    • video
    • commercially-safe
    • design
  • View Captions details
    VideoFREEMIUM

    Captions

    Mirage

    AI video editor and avatar creator for short-form, talking-head content.

    An AI video app for creators that auto-edits talking-head footage — generating captions, inserting B-roll, correcting eye contact, and dubbing into other languages. Its AI Creator mode renders a talking video from a script using AI personas. Built by Mirage on its own generative-video foundation model.

    In-house Mirage video model
    Focused on talking-head/short-form only
    • video
    • short-form
    • creators
    • avatars
  • View Submagic details
    VideoFREEMIUM

    Submagic

    Submagic

    AI editor that turns long videos into short-form clips.

    AI video editor for short-form content that auto-generates captions in dozens of languages, removes silences, inserts B-roll, and extracts the highest-engagement clips from long videos. Upload footage or a YouTube link and get TikTok/Reels/Shorts-ready edits. Pricing is per finished video rather than per credit.

    Removes silences automatically
    Web-only, no mobile editor
    • video
    • short-form
    • captions
    • clipping
  • View Hailuo AI details
    VideoFREEMIUM

    Hailuo AI

    MiniMax

    Text- and image-to-video generation from MiniMax.

    MiniMax's consumer video generator, turning text prompts and reference images into short cinematic clips with subject-reference for consistent characters. Available on the web and as iOS and Android apps. A free tier offers limited credits; subscriptions add HD output, faster generation, and commercial use.

    Realistic motion and physics
    Short clip-length cap
    • video-gen
    • text-to-video
    • image-to-video
    • minimax
  • View ComfyUI details
    ImageFREEMIUMOpen core

    ComfyUI

    Comfy Org

    Node-based visual AI — wire up image, video, and audio diffusion pipelines.

    An open-source, node-graph interface for diffusion models — build precise, reproducible image/video/audio pipelines on an infinite canvas where every model and parameter is visible. Runs locally on your own GPU or as a desktop app, with thousands of community nodes.

    Total control over the pipeline
    Steep learning curve
    • image-gen
    • video-gen
    • node-based
    • open-source
  • View Google Flow details
    VideoFREEMIUM

    Google Flow

    Google

    Google's AI filmmaking studio — Veo video + Imagen, in one canvas.

    Google Labs' creative studio for filmmakers — generate and stitch shots with Veo, craft keyframes with Imagen, and direct camera + scene with a Gemini-powered agent. Now folds in Whisk and ImageFX.

    Veo 3.1 quality with native audio/lip-sync
    ~8-second clip limit per generation
    • video-gen
    • filmmaking
    • veo
    • google-labs
  • View Kling AI details
    VideoFREEMIUM

    Kling AI

    Kuaishou

    AI video and image generation with strong motion and multishot sequences.

    Kuaishou's creative studio — text- and image-to-video with convincing motion, lip-sync, and multishot sequences up to ~15s, plus image generation. A leading Runway/Sora rival.

    Convincing complex motion and physics
    Credit-based pricing adds up
    • video-gen
    • image-to-video
    • motion
    • multishot
  • View Higgsfield details
    VideoFREEMIUM

    Higgsfield

    Higgsfield AI

    Cinematic AI video with camera controls — many models, one subscription.

    An AI video + image studio built around cinematic camera motion and presets. Aggregates 15+ third-party models (Sora, Veo, Kling, and more) so you switch engines without switching tools.

    Deep cinematic camera-motion controls
    Learning curve on advanced motion
    • video-gen
    • cinematic
    • camera-control
    • aggregator
  • View Synthesia details
    VideoFREEMIUM

    Synthesia

    Synthesia

    AI avatar video for training, marketing, and comms. Enterprise default.

    Studio for AI presenter videos — pick or clone an avatar, type a script in 140+ languages, and render a talking-head video. The go-to for L&D, onboarding, and corporate comms.

    230+ avatars, custom avatar cloning
    Avatars limited for high-emotion content
    • avatar-video
    • training
    • localization
    • enterprise
  • View Pika details
    VideoFREEMIUM

    Pika

    Pika Labs

    Playful AI video with Pikaffects, ingredients, and quick edits.

    Pika Labs' video generator — text- and image-to-video with signature effects (Pikaffects), character ingredients, and fast iteration. Popular for social-native, fun clips.

    Signature Pikaffects for fun clips
    Less photoreal than Veo/Sora/Kling
    • video-gen
    • effects
    • social
    • image-to-video
  • View HeyGen details
    VideoFREEMIUM

    HeyGen

    HeyGen

    Avatar video at scale. Talking-head clips from a script.

    AI avatar platform for B2B content — generate a presenter from a photo, give them a script, get a finished video with lip-sync, voiceover, and translation. Used heavily in marketing and corporate training.

    Highly realistic Avatar IV presenters
    Avatar IV burns credits fast
    • avatars
    • talking-head
    • translation
    • b2b-content
  • View Luma Dream Machine details
    VideoFREEMIUM

    Luma Dream Machine

    Luma AI

    Mobile-first AI video generation from text and images.

    Luma Labs' video model surfaced as an iOS app and web tool. Strong text-to-video, image-to-video, and keyframe controls — the indie creator's go-to before reaching for Runway's heavier toolkit.

    Fast generations
    Less granular control than Runway
    • video-gen
    • image-to-video
    • keyframes
    • mobile
  • View Descript details
    AudioFREEMIUM

    Descript

    Descript

    Podcast + audio editing where the transcript is the timeline.

    Audio and video editing built around an editable transcript — cut words, get cut audio. Add AI cleanup, overdub voice clones, and screen-recording for podcasts and tutorials in one tool.

    Transcript-as-timeline editing
    Transcription accuracy varies by audio
    • editing
    • podcasts
    • transcript
    • voice-clone
  • View Runway details
    VideoFREEMIUM

    Runway

    Runway

    Video AI with Gen-series models and a full timeline editor.

    End-to-end video AI platform — text-to-video, image-to-video, in-painting, motion brush, and a timeline editor. Longest-running studio in the space; default choice when the output ships to clients.

    Full editor, not just a generator
    Credits deplete fast on long clips
    • video-gen
    • editing
    • motion
    • film