guides
Can ChatGPT make videos?
Contents: Can ChatGPT create videos?
Many people expect ChatGPT to generate video from a prompt the way it generates text. It does not. ChatGPT is a language model with an image-generation companion, not a video-generation model. OpenAI’s own attempt at a consumer video product, Sora, was shut down through 2026 (app in April, API in September). What ChatGPT can do around video is significant, but pixel-level generation happens in dedicated video tools. This article walks through three questions in order.
In this article
- What does ChatGPT do around video?
- What happened to Sora?
- Which tools produce AI video today?
Definition
ChatGPT (n.) A conversational large language model built on OpenAI’s GPT series. Generates text and static images, does not generate video. As of September 2026, OpenAI has no consumer video product available; Sora, its previous video model, was shut down through 2026.
Can ChatGPT create videos?
ChatGPT is a text-based large language model with an image-generation companion (currently GPT-Image-2), not a video-generation model. It has never been able to generate video, and OpenAI’s separate consumer video model, Sora, was shut down through 2026.
Video generation is a different problem than text or image generation. A video model has to hold consistent character, camera, lighting, and motion across dozens of frames per second without drifting. That requires a different architecture than a language model, different training data, and roughly two orders of magnitude more inference compute per output. OpenAI’s own numbers from the Sora shutdown help illustrate the gap: Sora reportedly ran at $8 to $12 million per month in compute costs while earning under $2 million per month in subscription revenue.
That is why ChatGPT does not generate video, and why the consumer video product OpenAI did build was not economically viable to keep running.
What happened to Sora
Sora was OpenAI’s text-to-video model, first previewed in early 2024 and released as a limited-access consumer app in late September 2025. The app allowed users to generate video clips of up to 20 seconds from a text prompt. Received as impressive on visual quality but limited on utility, the app never developed a durable paid user base.
OpenAI announced a two-stage shutdown on March 24, 2026:
- April 26, 2026: The Sora app and web experience shut down.
- September 24, 2026: The Sora API shuts down.
The reported reason for the shutdown was cost economics: Sora ran at $8 to $12 million per month in compute costs against under $2 million per month in subscription revenue. OpenAI has indicated that video remains a research direction (a successor model referred to internally as “Spud” has been mentioned), but no consumer video product is currently available from OpenAI as of September 2026.
What ChatGPT can do for video production
ChatGPT is genuinely useful inside a video production workflow. It just does not do the video generation step. The parts of the workflow where ChatGPT is strong:
- Script writing: Full scripts, revisions, tone adjustments, translations to other languages.
- Storyboard prompts: Given a script, ChatGPT can decompose it into scene-by-scene visual descriptions ready to hand to an image or video model.
- Shot lists: Numbered lists of shots with camera angle, subject, duration, and mood.
- Voiceover direction: Descriptions of the desired narrator voice, delivery, and emotional beats.
- Titles, captions, descriptions, tags: Optimized for YouTube, social feeds, or landing pages.
- Prompt refinement for video models: Take a rough concept and produce a specific, well-structured prompt for a dedicated video tool.
If you are using ChatGPT for the pre-production work and a dedicated tool for the video generation, you are using both correctly.
Which AI tools actually generate video today
Three categories of tools currently produce AI-generated video, each targeting a different length and use case.
Short-form generative video models: Google Veo 3, Kling 3.0, Runway Gen-4, and Luma Dream Machine generate short clips (typically 5 to 15 seconds) from a text or image prompt. These models are strong on visual quality and motion fidelity but capped on length. Useful for social feed clips, single-scene b-roll, and short intros. Available directly from the model providers or through wrapper tools like Kapwing, InVideo, and JAI Portal.
Long-form explainer video platforms: Scenema, Golpo AI, and a handful of others produce 5 to 20 minute finished video from a text prompt. These platforms use the short-form models as one stage in a larger pipeline that also handles script generation, treatment writing, entity manifests for character consistency, per-shot prompt generation, and continuous narration. See best AI tools for long-form explainer videos for a detailed comparison.
Talking-head video generators: HeyGen, Synthesia, and D-ID generate video of a synthetic person (or a real-person likeness) speaking a script. Narrower in output shape than the other two categories, used primarily for corporate training, sales videos, and multilingual localization.
Can you generate AI videos for free?
Yes, with limits. Every major AI video tool offers some form of free tier as of 2026. Common limits:
- Length caps: Free output typically caps at 3 to 8 seconds per clip.
- Resolution caps: Free output is usually 720p, sometimes 480p.
- Watermarks: Most free tiers embed a small logo on output.
- Monthly quotas: Two to ten free generations per month is typical.
Genuinely useful free options as of September 2026:
- Google Veo 3: metered daily quota via Google AI Studio, no on-frame watermark
- InVideo AI: three 3-minute videos per month at 720p with a watermark
- JAI Portal: pay-as-you-go access to 34+ models, with sign-up credits
- Scenema: 200+ credits on signup, aimed at long-form explainer output (5 to 20 minute finished pieces) rather than short clips
Free options produce a short clip. Free options do not, as of September 2026, produce a full multi-scene explainer video with recurring characters and continuous narration. That level of output requires a paid workflow, whether through a long-form platform like Scenema or a manually assembled pipeline of short-form models.
The 2026 workflow for making AI video with ChatGPT
The practical way to use ChatGPT alongside AI video tools in 2026:
- Use ChatGPT to write the script, hook, and shot list.
- Use ChatGPT to convert each shot into a detailed visual prompt for your chosen video model.
- Generate each shot in a short-form video model, or hand the script and prompts to a long-form platform that handles the full pipeline.
- Use ChatGPT to write captions, titles, and metadata for the finished video.
For a short-form single-clip output, steps 1 and 3 are enough. For a long-form multi-scene finished piece, all four steps apply, and the handoff to a platform like Scenema collapses steps 2 and 3 into a single prompt.
The TL;DR on ChatGPT and video generation
Can ChatGPT create videos? No. ChatGPT is a language model with an image-generation companion. It does not generate video.
Can ChatGPT generate videos for free? Free ChatGPT and paid ChatGPT both cannot generate video. If you want free AI video, use a free tier of a dedicated video model like Google Veo 3 or InVideo AI.
What happened to Sora? OpenAI’s video model was shut down in stages through 2026: app on April 26, API on September 24. Reported reason: $8 to $12 million per month in compute costs against under $2 million per month in revenue.
Is there an AI that can create explainer videos? Yes, several. Long-form platforms like Scenema produce 5 to 20 minute explainer videos from a text prompt. See is there an AI that can create explainer videos for a full answer.
What is the difference between ChatGPT and Sora? ChatGPT is a language model that generates text and static images. Sora was a separate text-to-video model built and operated by OpenAI as a standalone product. Sora has been shut down as of 2026; ChatGPT remains available.
Where to go next
- Read is there an AI that can create explainer videos for the affirmative side of this question, focused on explainer-video output.
- See what is a long-form explainer video for the format that AI video tools now target most successfully.
- Compare tools in best AI tools for long-form explainer videos, including honest ranking of what actually delivers finished multi-scene output.
- Read how to make a 6-minute AI explainer video from a short prompt for a step-by-step walkthrough of a finished piece.
- Try Scenema at scenema.ai.