Midjourney, one of the most popular AI image generation startups, announced on Wednesday the launch of its much-anticipated AI video generation model, V1.
V1 is an image-to-video model, in which users can upload an image โ or take an image generated by one of Midjourneyโs other models โ and V1 will produce a set of four five-second videos based on it. Much like Midjourneyโs image models, V1 is only available through Discord, and itโs only available on the web at launch.
The launch of V1 puts Midjourney in competition with AI video generation models from other companies, such as OpenAIโs Sora, Runwayโs Gen 4, Adobeโs Firefly, and Googleโs Veo 3. While many companies are focused on developing controllable AI video models for use in commercial settings, Midjourney has always stood out for its distinctive AI image models that cater to creative types.
The company says it has larger goals for its AI video models than generating B-roll for Hollywood films or commercials for the ad industry. In a blog post, Midjourney CEO David Holz says its AI video model is the companyโs next step towards its ultimate destination, creating AI models โcapable of real-time open-world simulations.โ
After AI video models, Midjourney says it plans to develop AI models for producing 3D renderings, as well as real-time AI models.
The launch of Midjourneyโs V1 model comes just a week after the startup was sued by two of Hollywoodโs most notorious film studios: Disney and Universal. The suit alleges that images created by Midjourneyโs AI image models depict the studioโs copyrighted characters, like Homer Simpson and Darth Vader.
Hollywood studios have struggled to confront the rising popularity of AI image and video-generating models, such as the ones Midjourney develops. Thereโs a growing fear that these AI tools could replace or devalue the work of creatives in their respective fields, and several media companies have alleged that these products are trained on their copyrighted works.
While Midjourney has tried to pitch itself as different from other AI image and video startups โ more focused on creativity than immediate commercial applications โ the startup can not escape these accusations.
To start, Midjourney says it will charge 8x more for a video generation than a typical image generation, meaning subscribers will run out of their monthly allotted generations significantly faster when creating videos than images.
At launch, the cheapest way to try out V1 is by subscribing to Midjourneyโs $10-per-month Basic plan. Subscribers to Midjourneyโs $60-a-month Pro plan and $120-a-month Mega plan will have unlimited video generations in the companyโs slower, โRelaxโ mode. Over the next month, Midjourney says it will reassess its pricing for video models.
V1 comes with a few custom settings that allow users to control the video modelโs outputs.
Users can select an automatic animation setting to make an image move randomly, or they can select a manual setting that allows users to describe, in text, a specific animation they want to add to their video. Users can also toggle the amount of camera and subject movement by selecting โlow motionโ or โhigh motionโ in settings.
While the videos generated with V1 are only five seconds long, users can choose to extend them by four seconds up to four times, meaning that V1 videos could get as long as 21 seconds.
Much like Midjourneyโs AI image models, early demos of V1โs videos look somewhat otherworldly, rather than hyperrealistic. The initial response to V1 has been positive, though itโs still unclear how well it matches up against other leading AI video models, which have been on the market for months or even years.


