How To Make AI Videos with Google Al Ultra (made with veo 3)

Corbin BrownAbout 5 min readMay 27, 2025Watch original
THE SUMMARYAI-generated

V3 AI Video Generation Software: A Deep Dive

Key Concepts:

  • V3: Google's AI video generation software that creates videos with both visual and audio elements from text prompts and image inputs.
  • Text-to-Video: The core functionality of V3, converting textual descriptions into video content.
  • Image Ingredients: The ability to incorporate user-uploaded images into the generated videos.
  • Uncanny Valley: The feeling of unease or revulsion experienced when encountering realistic but not fully human representations.
  • AI Mimicry: The potential for AI to replicate a person's appearance, voice, and behavior based on available data.

Initial Impressions and Capabilities

The video explores V3, a $250/month AI video generation software, focusing on its ability to create videos with both visual and audio elements. The presenter emphasizes that V3 is the first AI video generation software to incorporate audio. An example is shown of a fake video featuring someone playing Fortnite, complete with keyboard clicks and spoken commentary, all generated by AI.

Experiment 1: Image Integration and Uncanny Results

The presenter begins by testing V3's ability to integrate a user-uploaded image into a generated video. He uploads a photo of himself and uses the prompt: "a man in the middle of the image have him jump up and down happy and clapping hands."

  • Process:
    1. Upload an image.
    2. Write a text prompt describing the desired action.
    3. Generate the video.
  • Results: The generated videos are described as "weird" and "uncanny." One video shows the presenter's likeness doing jumping jacks. The presenter notes that the physics of the jumping motion are surprisingly realistic. He expresses discomfort with the realism, highlighting the "uncanny valley" effect.

Experiment 2: Simulating a Protest

The presenter attempts to generate a video of a protest using the prompt: "have a protest of people yelling they are not made from a prompt."

  • Prompt: "have a protest of people yelling they are not made from a prompt"
  • Results: The generated video depicts a crowd of people yelling, "We are not made from a prompt! We are real!" The presenter finds this result "crazy" and notes the implications of Google training its AI on diverse datasets, including protest footage. He acknowledges that the faces in the video are blurry, making it obvious that it's AI-generated, but speculates on the potential realism of future versions of the software (V10). He jokingly questions whether the entire video he is making is AI-generated. The audio didn't come through on one of the generated videos.

Experiment 3: Mimicking a YouTuber

The presenter tests V3's ability to mimic a YouTuber with the prompt: "intro by a YouTuber with a bucket hat who has a green screen behind him and is saying 'Today we will make a fake AI video with Google.'"

  • Prompt: "intro by a YouTuber with a bucket hat who has a green screen behind him and is saying 'Today we will make a fake AI video with Google.'"
  • Results: The generated video features a figure with a bucket hat in front of a green screen, saying the prompted phrase. The presenter is surprised that the AI generated a generic green screen rather than attempting to replicate his own studio setup. He also notes a strange "cough" sound at the end of one of the generated videos. He jokingly wonders if he has sold his face data to Google through uploading so many videos.

Experiment 4: Coding Tutorial Simulation

The presenter explores V3's ability to simulate a coding tutorial with the prompt: "make a video of a YouTuber who is in the corner of a video and it's a coding tutorial showing them how to code a web app in an IDE and the YouTuber is actually coding in it."

  • Prompt: "make a video of a YouTuber who is in the corner of a video and it's a coding tutorial showing them how to code a web app in an IDE and the YouTuber is actually coding in it."
  • Results: The generated video shows a YouTuber in the corner of the screen with code displayed. The presenter notes that the code is gibberish but acknowledges that the AI accurately placed the YouTuber's face in the corner. He believes that providing actual code in the prompt might yield more realistic results.

Conclusion and Future Implications

The presenter concludes that V3 is a significant step forward in AI video generation and plans to create more videos on the topic. He encourages viewers to subscribe to his channel for future updates. He acknowledges that the technology is rapidly evolving and speculates on the potential of future versions of the software.

Notable Quotes:

  • "This is the first AI video generation software that also puts audio on it."
  • "This just got a little weird."
  • "What happens in V10 at this rate?"
  • "Did I just sell my face data to Google?"

Key Takeaways:

  • V3 is a powerful AI video generation tool capable of creating videos with both visual and audio elements.
  • The software can integrate user-uploaded images into generated videos.
  • The results can be uncanny and raise ethical questions about AI mimicry and data privacy.
  • The technology is rapidly evolving, with the potential for future versions to be even more realistic.

AI summaries can miss context or contain errors. Check important details against the original video.

Go a little deeper.

Have a question about this video? Load its transcript to open the video chat.