Midjourney Launches First AI Video Model V1
Midjourney debuts V1, an image-to-video model that turns static visuals into short animations, marking a leap in generative AI creativity.
Midjourney has officially entered the AI video generation arena with the launch of V1, its first model capable of transforming static images into short, dynamic video sequences. Announced on June 18, 2025, this move marks a pivotal expansion for the company, which has built a reputation over the past three years as a leader in AI-generated imagery favored by artists, designers, and creative professionals. Unlike previous offerings limited to stills, V1 enables users to generate four unique five-second video clips from a single image—either uploaded or previously created within Midjourney’s ecosystem—with the option to extend each clip up to 21 seconds through incremental additions. The release signals a broader industry shift: generative AI is no longer confined to producing snapshots but is now animating them.
The introduction of video capabilities matters not only for creative workflows but also for the competitive and legal landscape of generative AI. As tools evolve from generating individual assets to crafting moving narratives, questions around authorship, training data rights, and commercial viability grow more urgent. Midjourney’s entry intensifies pressure on rivals like OpenAI and Google, even as it faces mounting legal scrutiny. Simultaneously, the company’s continued independence—operating without external funding—positions it as a rare counterweight to well-financed tech conglomerates shaping the future of AI creativity.
How V1 Works and Who Can Access It
V1 operates as an image-to-video model, meaning it requires a visual input to begin animation. Users submit one image and receive four distinct five-second video outputs, each interpreting motion differently. These can be extended in four-second increments, up to a maximum of 21 seconds, allowing for modest narrative development or looping animations. Control options include selecting between automatic motion settings or specifying animation style via text prompts, giving users influence over how elements move within the frame. Additionally, sliders for camera and subject motion intensity—set to either low or high—offer further refinement.
At launch, access is restricted to Midjourney’s existing platforms: Discord and the web interface. There are no standalone apps or API integrations available yet. This maintains the company’s established user experience, where commands are issued through chat-based interactions rather than complex graphical interfaces. However, the computational demands of video generation have led to changes in resource allocation across subscription tiers. The $10/month Basic plan includes video functionality, but each video generation consumes eight times the credit of a standard image. For heavy users, this makes experimentation costly under the base tier. In contrast, Pro ($60/month) and Mega ($120/month) subscribers enjoy unlimited video generations when using the slower “Relax” mode, which prioritizes queue efficiency over speed.
David Holz, Midjourney’s CEO and founder, emphasized that pricing is provisional. “We’ll reassess within the first month based on actual usage patterns,” he stated, acknowledging the uncertainty around demand and infrastructure strain. This flexibility suggests Midjourney may adjust its model as user behavior becomes clearer—a necessary approach given the exponential increase in processing power required for video versus image generation.
A Strategic Step Toward Simulated Worlds
Holz described V1 not as a final product but as a “stepping stone” toward a more ambitious vision: real-time open-world simulations. This long-term goal follows what he frames as a logical progression—first mastering static images, then adding time through video, followed by full 3D environments, and ultimately interactive, navigable worlds. While such a future remains speculative, V1 represents a tangible advance along that path. By introducing temporal dynamics, Midjourney begins to explore continuity, causality, and motion logic—foundational components for any simulated environment.
This trajectory differentiates Midjourney from competitors focused primarily on commercial content creation. OpenAI’s Sora and Google’s Veo, for instance, emphasize photorealism and long-form video generation aimed at filmmakers and media studios. Midjourney, by contrast, continues to prioritize artistic exploration and emergent aesthetics. Its user base—largely composed of illustrators, concept artists, and digital creators—values unpredictability and stylistic diversity over pixel-perfect realism. V1 inherits this ethos, producing results that lean expressive rather than documentary.
Yet the line between creative tool and production engine is blurring. With video now part of its suite, Midjourney expands its utility beyond mood boards and character design into storyboarding, animated prototypes, and short-form social content. Executives in entertainment, advertising, and game development may find new value in rapid ideation cycles enabled by quick video iterations. The ability to go from sketch to motion in minutes could compress early-stage creative pipelines significantly.
Legal Risks Loom Amid Expansion
The launch arrives amid escalating legal challenges. On the same day V1 was unveiled, Disney and Universal filed a joint lawsuit alleging that Midjourney’s underlying image models were trained on copyrighted material without authorization and routinely produce outputs that closely resemble protected characters and works. The complaint asserts that the company’s technology facilitates infringement at scale, undermining intellectual property rights across film, television, and publishing.
While the case centers on Midjourney’s image generator, the implications extend directly to V1. Since the video model builds upon prior image-generation capabilities—both technically and legally—the lawsuit threatens the foundation of its new offering. If courts rule that training on copyrighted data constitutes infringement, all derivative applications, including video, could face restrictions or require licensing agreements. Midjourney has not commented specifically on how V1 might be affected, but the timing underscores the precarious position of many generative AI firms operating in regulatory gray zones.
This legal pressure contrasts sharply with Midjourney’s self-positioning as a tool for individual expression. The company has long promoted its use among independent creators, often highlighting non-commercial art and personal projects. But as its technology gains functionality—and potential commercial appeal—it becomes harder to distance itself from industrial applications where copyright compliance is paramount. For enterprise clients considering integration, the unresolved litigation introduces risk that could delay adoption or prompt caution in regulated sectors.
Market Position and Future Trajectory
Midjourney’s unfunded status sets it apart in a field increasingly dominated by deep-pocketed players. Without venture capital or corporate backing, the company retains operational autonomy and avoids shareholder-driven timelines. This independence allows for deliberate, iterative development—but also limits scalability and infrastructure investment compared to rivals backed by Microsoft, Google, or Amazon.
Nonetheless, V1 demonstrates that innovation in generative AI is not solely the domain of large tech firms. By focusing on niche usability and community-driven feedback, Midjourney continues to punch above its weight. Its emphasis on creative utility over benchmark supremacy resonates with professionals who value control and aesthetic range. Early reactions from beta testers suggest enthusiasm for the model’s dreamlike, impressionistic output—qualities less prioritized by systems aiming for cinematic fidelity.
Looking ahead, the success of V1 will depend on both technical refinement and legal resolution. Can Midjourney improve coherence and duration while managing compute costs? Will it develop tools for attribution or opt-in training data to mitigate legal exposure? And how quickly will it iterate toward multimodal capabilities, such as text-to-video or audio synchronization?
Sources
- Midjourney launches its first AI video generation model, V1
- Midjourney launches its first AI video generation model, V1 | TechCrunch
- Midjourney V1 Launch: New AI Video Era Begins
- TechCrunch
- Midjourney V1 AI video model: Price and how to try it | Mashable
Written by an AI editorial process from the sources above. Errors may occur.
Newsletter
Get the AI news that matters
One short brief with the day's most important AI stories — written for professionals.
We send a confirmation link. No spam. Unsubscribe anytime.
Read next
Gemini's New AI Model Transforms Photo Editing with Multi-Turn Control
Google's Gemini 2.5 Flash Image enables precise, context-aware edits across multiple steps, advancing creative workflows for professionals.
27 Sep 2026
Autonomous AI Agents: The Rise of Digital Workforce
How self-reasoning AI systems are transforming business workflows and redefining automation across industries.
26 Sep 2026
Adobe Launches AI Video Generation in Creative Cloud
Adobe unveils Firefly Video Model and faster image generation, embedding AI deeply into professional creative workflows.
26 Sep 2026