This is now the BEST AI video generator! Free & open-source

AI SearchAbout 6 min readJul 31, 2025Watch original
THE SUMMARYAI-generated

Key Concepts

  • Juan 2.2: An open-source AI video generator by Alibaba, successor to Juan 2.1.
  • Text-to-Video: Generating videos from text prompts.
  • Image-to-Video: Animating a still image into a short video.
  • ComfyUI: A popular open-source platform for running image and video generators offline, offering customization and low VRAM support.
  • VRAM: Video RAM, the memory on a graphics card.
  • Quantization: A technique to reduce the size of AI models, allowing them to run on hardware with less VRAM.
  • Workflows: Pre-built configurations in ComfyUI that define the steps for generating images or videos.
  • Nodes: Individual processing units within a ComfyUI workflow.
  • K Sampler: A node in ComfyUI that uses sampling algorithms to generate images or videos.
  • CFG Scale: Classifier-Free Guidance scale, a parameter that controls how closely the AI follows the prompt.
  • Loras: Lightweight, fine-tuned models that can be added to existing models to enhance specific aspects of generation or speed up the process.
  • Self-Forcing Lora: A specific type of Lora that allows for video generation with fewer steps.
  • GGUF: A file format for quantized models.

Juan 2.2 Overview

Juan 2.2 is presented as the best free and open-source AI video generator currently available, emphasizing its uncensored nature. It excels at generating cinematic and realistic-looking videos, demonstrating strong capabilities in lighting, wide-angle shots, tracking shots, anatomy understanding, and handling high-action scenes.

  • Key Features:
    • Realistic lighting and cinematic visuals.
    • Good anatomy understanding and coherent action scenes.
    • Camera control: ability to specify camera movements like tracking shots.
    • Can generate fight scenes and high-action scenes well.
    • Image-to-video and text-to-video capabilities.

Online Platform Demonstration

The video showcases Juan 2.2's capabilities through examples on its online platform (one-api.alibaba.com).

  • Text-to-Video Examples:

    • Complex Prompt (Victorian Lady): Successfully generates a video based on a detailed prompt with multiple elements, including a Victorian lady, lavish bedroom, antique bottles, a parrot, a dinosaur, and a butler in a superhero cape.
    • Ballerina Example: Accurately depicts a ballerina spinning in a studio with scattered shoes and sheet music, a rabbit on a piano, and an elephant balancing on a circus ball outside.
    • Snowboarder Tracking Shot: Generates a tracking shot of a snowboarder launching off a cliff with accurate physics and anatomical correctness.
    • Gymnast on Balance Beam: Correctly generates a gymnast performing a flip on a balance beam, maintaining anatomical accuracy.
    • Cat Figure Skating: Generates a cat figure skating on an ice rink with graceful leaps and spins, without anatomical errors.
    • Fight Scene (Tuxedos): Generates a fight scene between two men in white tuxedos on a rooftop with rain and lightning, although the motion is somewhat slow.
    • Fight Scene (Animal Heads): Generates an intense fight between a man with a cat's head and another with a dog's head, with fast movements and anatomical correctness.
    • Will Smith Eating Spaghetti: Fails to generate the likeness of Will Smith using text-to-video.
  • Image-to-Video Examples:

    • Will Smith Eating Spaghetti: Successfully generates a video of Will Smith eating spaghetti by using an initial image of Will Smith.
    • Anime Scene: Animates an anime image into a scene of friends talking in a cafe, capturing the characteristic style of anime.
    • 3D Pixar Scene: Animate a 3D Pixar image with multiple characters moving and talking.
    • Chinese Watercolor Painting: Animates a Chinese watercolor painting, specifically animating the fish while keeping the background still.
    • K-Pop Dance Scene: Animates a photo of a K-pop group into a coherent dance scene with consistent character appearances.
    • Minecraft Scene: Animate a Minecraft scene, moving certain characters but not the buildings.
    • Genshin Impact Gameplay: Animate a Genshin Impact gameplay scene, keeping the interface elements consistent.
    • Image with Instructions: Follows instructions written on an image, such as adding a dog walking and a balloon floating.

Local Installation with ComfyUI

The video provides a step-by-step guide to installing and running Juan 2.2 locally using ComfyUI.

  • Prerequisites:

    • ComfyUI installed (refer to a previous tutorial for initial setup).
    • Update ComfyUI to the latest version.
  • 5 Billion Parameter Model (Hybrid - Text & Image to Video):

    1. Download Models:
      • one_2_2_base_5b.safetensors (Diffusion Model) - Place in ComfyUI/models/diffusion_models.
      • one_2_2_vae.safetensors (VAE File) - Place in ComfyUI/models/VAE.
      • UMT5-XXL-FP8.scaled.safetensors (Text Encoder) - Place in ComfyUI/models/text_encoders.
    2. Download Workflow: Download the ComfyUI workflow file.
    3. Load Workflow: Drag and drop the workflow file onto the ComfyUI interface.
    4. Select Models: In the ComfyUI workflow, select the downloaded models in the corresponding dropdown menus for each node (diffusion model, clip loader, VAE).
    5. Text-to-Video: Enter a text prompt, specify width, height, and length, and adjust K Sampler settings (steps, CFG).
    6. Image-to-Video: Upload an image, enter a prompt, and run the workflow.
  • 14 Billion Parameter Model (Text-to-Video):

    1. Download Models:
      • one_2_2_text_high_noise.safetensors (High Noise Model) - Place in ComfyUI/models/diffusion_models.
      • one_2_2_text_low_noise.safetensors (Low Noise Model) - Place in ComfyUI/models/diffusion_models.
      • one_2_1_vae.safetensors (VAE File) - Place in ComfyUI/models/VAE.
      • UMT5-XXL-FP8.scaled.safetensors (Text Encoder) - Place in ComfyUI/models/text_encoders.
    2. Download Workflow: Download the ComfyUI workflow file.
    3. Load Workflow: Drag and drop the workflow file onto the ComfyUI interface.
    4. Select Models: Select the downloaded models in the corresponding dropdown menus.
    5. Enter Prompts: Enter positive and negative prompts, specify video dimensions and length, and adjust K Sampler settings.
  • 14 Billion Parameter Model (Image-to-Video):

    1. Download Models:
      • one_2_2_image_high_noise.safetensors (High Noise Model) - Place in ComfyUI/models/diffusion_models.
      • one_2_2_image_low_noise.safetensors (Low Noise Model) - Place in ComfyUI/models/diffusion_models.
      • one_2_1_vae.safetensors (VAE File) - Place in ComfyUI/models/VAE.
      • UMT5-XXL-FP8.scaled.safetensors (Text Encoder) - Place in ComfyUI/models/text_encoders.
    2. Download Workflow: Download the ComfyUI workflow file.
    3. Load Workflow: Drag and drop the workflow file onto the ComfyUI interface.
    4. Select Models: Select the downloaded models in the corresponding dropdown menus.
    5. Enter Prompts: Enter positive and negative prompts, upload an image, and adjust K Sampler settings.

Speeding Up Generation and Reducing VRAM Usage

The video provides several techniques to speed up video generation and reduce VRAM usage.

  • Quantized Models (GGUF):
    • Download quantized versions of the models (e.g., Q8 version) from the linked page.
    • Replace the "Load Diffusion Model" node in the ComfyUI workflow with a "GGUF Loader" node.
    • Select the downloaded GGUF model in the GGUF Loader.
  • Lower Resolution: Reduce the width and height of the video.
  • Self-Forcing Lora:
    • Download the self-forcing Lora (either the I2V or T2V version, depending on the workflow).
    • Add a "Lora Loader Model Only" node between the high-noise/low-noise models and the K Sampler nodes.
    • Connect the nodes appropriately.
    • Set the K Sampler steps to 4-8, CFG to 1, and use an LCM sampler.
    • Adjust the start and end steps of the K Samplers accordingly.

Uncensored Nature and Lora Compatibility

The video mentions that Juan 2.2 is uncensored and compatible with Loras from Juan 2.1, including those that enable the generation of explicit content.

Monica AI Assistant

The video includes a sponsored segment for Monica, an AI assistant that provides access to various AI tools in one platform.

  • Key Features:
    • Access to top AI models (GPT, DeepSeek, Gemini).
    • Access to top image and video generators (Flux, Stable Diffusion, Cling, High Law).
    • Browser extension with context-aware capabilities.
    • Summarization of web pages and YouTube videos.
    • Mind map generation.

Conclusion

Juan 2.2 is presented as a powerful and versatile open-source AI video generator, offering impressive capabilities in both text-to-video and image-to-video generation. The video provides a comprehensive guide to installing and using Juan 2.2 locally with ComfyUI, along with techniques to optimize performance and reduce VRAM usage. The uncensored nature and Lora compatibility further enhance its appeal.

AI summaries can miss context or contain errors. Check important details against the original video.

Go a little deeper.

Have a question about this video? Load its transcript to open the video chat.