How to Turn a PDF into a Video [2026]
Four ways to turn a PDF into a video in 2026, what each one costs, and what survives the trip from document to narration. Checked against vendor pages.
There are four ways to turn a PDF into a video, and they are not interchangeable. You can hand the document to an AI that reads it and narrates it back, write a script and let a synthetic presenter deliver it, turn the pages into slides and record them yourself, or generate footage with a text to video model. The right choice depends on a question most guides skip: do you want the document explained, or just presented?
Almost every article on this term starts from a converter and works backwards. A PDF is not a video with the motion removed, and no button turns 60 pages into three watchable minutes without someone deciding what matters. This guide covers what transfers out of a PDF, what each route costs, and which one fits the job you have, with every price checked against the vendor's own page in August 2026.
What Does "Turning a PDF into a Video" Actually Mean?
Turning a PDF into a video means converting a document into narrated, watchable form, either by explaining its content, presenting its pages, or performing a script drawn from it.
Those three verbs are the whole decision. Explaining means something reads the document and builds an account of it. Presenting means the pages appear while a voice reads them. Performing means a presenter, real or synthetic, delivers a script you wrote. Most disappointment with PDF to video tools comes from wanting the first and buying the third.
| Route | What you supply | What comes out | Effort | Cost |
|---|---|---|---|---|
| AI that reads the document | The file and a question | Narrated explainer with generated visuals | Minutes | Free tiers exist |
| Avatar and script tools | A finished script, or slides | A synthetic presenter delivering it | Hours | Free plans are capped, paid from $19/mo |
| Slides to video | Slides, timings, your own narration | Exactly what you put on screen | A day or more | Free if you already have Office |
| Text to video models | A prompt describing footage | Invented footage unrelated to your file | Minutes | Wrong tool, see below |
Step 1: Check What Is Actually Inside Your PDF
A PDF's usefulness as video source material depends on whether it contains real text, real structure, or only a picture of a page.
Open the file and try to select a sentence with your cursor. If the text highlights, there is a text layer and every route below is available to you. If nothing highlights, you have a scanned PDF: a photograph of a document, not a document. Nothing can read it until it has been through OCR.
Even with selectable text, a PDF says far less about itself than people assume. The PDF Association puts it plainly: an untagged PDF is a description of appearance, marks positioned on a page, with nothing declaring that this run of text is a heading or that those cells form a table. A tagged PDF adds that missing layer. An untagged one leaves every reader, human or machine, to guess.
The failure modes are predictable, and they surface in the video you get back.
A table collapses into a wall of unformatted text, its row-and-column relationships, the entire point of a table, gone.
PDF Association
OCR does not rescue structure either. Google publishes the limits of its own conversion: the file should be 2 MB or smaller, the text at least 10 pixels high, the page right side up, and while bold, italics and line breaks usually survive, lists, tables, columns, footnotes and endnotes are not likely to be detected. Adobe's Scan and OCR is the heavier alternative, and Adobe puts it on Acrobat Pro at $19.99/mo, not on the cheaper Standard tier.
Five minutes of preparation changes the output more than any tool choice:
- Run the select-text test. If you cannot highlight a sentence, OCR the file first and check the result before feeding it to anything.
- Confirm the headings are real headings. Text that is merely bigger and bolder is invisible as structure.
- Pull complex tables out as images. A screenshot on screen beats a mangled paraphrase in the narration.
- Cut the document down first. Appendices, boilerplate and revision histories add length and subtract clarity.
Step 2: Decide What the Video Has to Do
The second decision is editorial, not technical, and it is where the compression problem becomes unavoidable.
A 60-page policy document runs to roughly 20,000 words. Three minutes of narration is around 450. That is a ratio of more than forty to one, and no tool resolves it for you. The document is not the script. Something has to choose which 2 percent a viewer hears.
Be concrete about the job before you pick a tool:
- Onboarding and policy. People need to know what changed and what to do. Two to four minutes on the changes, with the PDF left as the reference.
- Study notes and papers. The value is in explanation, not coverage. A tool that reasons about the material beats one that recites it.
- Announcements. You want a performance, not a walkthrough, so write the script yourself.
If you need the gist rather than something watchable, a summarizer is cheaper, and Scrimba's roundup of AI video summarizers covers that case.
Option 1: Let an AI Read the Document and Narrate It Back
The newest route hands the document to an AI assistant that reads it and produces a narrated explainer, with no editor, timeline or script involved.
The mechanics explain both its strength and its limits. In ChatGPT, third-party tools live in an app directory OpenAI describes as browsable and searchable from the tools menu or at chatgpt.com/apps, and an installed app runs when you mention it by name in the conversation. The document goes into the chat, the assistant reads it, and the video tool turns the result into something watchable.
Explain Video Generator Inside ChatGPT
Scrimba Explain, listed in ChatGPT's plugin directory as Explain Video Generator, turns a question into a narrated video walkthrough built from what the assistant has already read.
Its listing describes free narrated explainer videos made directly in ChatGPT, turning a topic, document, webpage, codebase or product brief into a video with voiceover, diagrams, animations and images, watchable inline or shareable by link. Scrimba promotes three example prompts there, and the first is exactly this job: Turn this PDF into a video for our employees.
Underneath it is an MCP plugin, not a converter. You ask a question, the assistant researches it against the files and conversation it already has, and Explain turns that answer into the video (Scrimba). It works with Claude Code, Codex, ChatGPT and any other MCP agent. It is free during open beta, and Scrimba has published nothing beyond that about future pricing.
Two limits belong in the same breath. No avatars, no stock footage, no brand templates, so it does not replace a corporate video tool. And its FAQ is candid: like any AI tool, it can make mistakes, so double-check anything important. Scrimba has a fuller guide to using Explain for the developer-facing detail.
Gemini Notebook Video Overviews
Google's notebook product does the same job from a document workspace rather than a chat, and it accepts PDFs as a first-class source.
A naming note first: NotebookLM was renamed Gemini Notebook in July 2026. Same product. A single source can run to 500,000 words or 200 MB, which covers any document you were going to convert (Google).
The Video Overview it generates is narrated slides. Google's launch post describes the AI host creating new visuals while pulling images, diagrams, quotes and numbers out of your documents. Three formats exist, including a roughly 60 second Short, the file can be downloaded, and generation sometimes takes more than 30 minutes. Google attaches the same warning Scrimba does about accuracy.
Option 2: Avatar Tools, If You Already Have a Script
Avatar tools put a synthetic presenter on screen to deliver a script you supply, which makes them strong for training and compliance and wrong for explanation.
The distinction matters more than the feature lists. These products are good at making a person appear to say your words. They are not working out what your document means. Hand one a PDF and you are asking it to read aloud.
| Tool | Free plan | Entry paid plan | How a PDF gets in |
|---|---|---|---|
| Synthesia | 10 minutes a month, watermarked, no download | $19/mo, or $14/mo billed yearly | Through its Assistant, on paid plans only |
| HeyGen | 3 videos a month, 1 minute each | $29/mo, or $24/mo billed yearly | PPT/PDF to Video, 50 MB and 50 slides |
Synthesia's free Basic plan gives you up to 10 minutes of video a month, but its own help pages are blunt about the catch: Basic videos carry the Synthesia logo and cannot be downloaded at all. Paid plans start at $19/mo, or $14/mo billed yearly (Synthesia). PDFs enter through its Assistant, which is available on Starter, Creator and Enterprise and not on the free plan. A PowerPoint import is more literal: slides become editable scenes and speaker notes become the script.
HeyGen names the feature outright. PPT/PDF to Video accepts PPT, PPTX or PDF up to 50 MB, and takes only the first 50 slides of anything longer. There is an asymmetry inside it. Using speaker notes as the script and importing slide content as editable elements are PowerPoint only, and not available for PDF uploads. The same material as a deck gets you strictly more than the same material as a PDF.
An import hands the tool your text and your layout, never judgment about what to cut. And every free AI video plan here is a metered sample rather than a free product.
Option 3: Slides to Video, the Manual Route
The manual route converts your document into slides and records narration over them, which costs the most time and delivers the most control.
If the PDF began life as a deck, this is nearly free. PowerPoint has done it for years: File, then Export, then Create a Video produces an .mp4 or .wmv at resolutions up to Ultra HD 4K, and it uses your recorded timings and narration when you have recorded them. Without them it holds each slide for five seconds (Microsoft).
Canva is the other common answer, and its documentation quietly confirms Step 1. It imports PDFs of up to 500 pages, but warns that scanned PDFs arrive as a single flattened image whose text cannot be edited.
One trap: Pictory's FAQ says its article input is HTML, and that PDF, Word and Google Docs are not supported.
Choose the manual route when:
- The content is regulated or safety critical, and every sentence has to be what you approved.
- The document is too confidential to upload to a third-party service.
- You need brand templates, specific fonts, or a named human presenter.
Option 4: Why Text to Video Models Cannot Read Your PDF
Text to video models generate invented footage from a written description. They never read your document, and no prompt makes them.
This is the most common wrong turn, because "AI video" now names the generative category. These models predict what a plausible video of your description would look like. The substance of your PDF lives outside the prompt box, and how text to video actually works explains why prompt engineering cannot move it inside.
The category is also unstable: OpenAI's Sora closed its web and app experiences on 26 April 2026. If your job is generating footage, Scrimba's roundup of AI video generators compares the current field. For turning a document into a video, this route has nothing to offer.
Which PDF to Video Route Should You Pick?
Pick the route by what the video has to do, not by which tool markets itself hardest on the phrase "PDF to video".
Match the sentence you would use to describe the job:
- "Explain this to me." Use an AI that reads the document. Explain Video Generator and Gemini Notebook both cost nothing to try.
- "Staff need to watch this and act on it." Write a short script, then hand it to Synthesia or HeyGen.
- "Every word has to be signed off." Build slides and record them. PowerPoint's export is enough.
- "I need footage of something that does not exist." Different project. Your PDF is not part of it.
Frequently Asked Questions
Can you turn a PDF into a video for free?
Yes. Scrimba Explain is free during open beta, Gemini Notebook generates Video Overviews on its free tier, and PowerPoint's Export to Video costs nothing if you already have Office. Synthesia and HeyGen both have free plans, but they cap minutes, video count and export quality.
Do you need to OCR a scanned PDF first?
Yes. A scanned PDF is an image of a page with no text layer, so nothing can read it. Google's own conversion guidance asks for files of 2 MB or smaller with text at least 10 pixels high, and warns that lists, tables, columns and footnotes are unlikely to be detected.
How long should a video made from a document be?
Two to four minutes for most workplace documents. Three minutes of narration is roughly 450 words, while a 60-page document runs to about 20,000, so the video is a summary by definition. Leave the PDF attached as the reference version.
Can AI read a PDF and narrate it accurately?
Usually, but not reliably enough to skip review. Both Scrimba and Google publish the same caveat: the output is AI generated and may contain mistakes. Watch the video before you send it, and check any number, date or instruction against the source document.
What happens to tables and charts in the PDF?
They degrade. Without tagging, a table is a wall of positioned text rather than rows and columns, and a chart is a picture with no underlying data. Screenshot anything that matters and put it on screen as an image instead of trusting the narration to describe it.
Key Takeaways
- Turning a PDF into a video has four routes: an AI that reads the document, an avatar delivering your script, slides you record yourself, and generative models that cannot read the file.
- Check the file first. If you cannot select the text, it is a scanned PDF and needs OCR.
- Structure fails silently. Tables, columns and footnotes are lost first, through OCR or through an untagged PDF.
- Scrimba Explain, listed in ChatGPT's plugin directory as Explain Video Generator, is free during open beta and takes a question rather than a script.
- Gemini Notebook, renamed from NotebookLM in July 2026, takes a PDF of up to 500,000 words and returns narrated slides.
- Synthesia starts at $19/mo, HeyGen at $29/mo, and both perform a script rather than explaining a document.
- Every AI route carries its vendor's own warning that the output can be wrong.
Sources
All vendor pages accessed August 2026.
- PDF Association. "Tagged PDFs and machine reading." 2026. https://pdfa.org/you-tagged-pdfs-for-screen-readers-turns-out-the-machines-needed-it-too/
- Google. "Convert PDF and photo files to text." 2026. https://support.google.com/drive/answer/176692
- Adobe. "Acrobat pricing." 2026. https://www.adobe.com/acrobat/pricing.html
- OpenAI. "Developers can now submit apps to ChatGPT." 2025. https://openai.com/index/developers-can-now-submit-apps-to-chatgpt/
- Scrimba. "Explain." Self-reported. 2026. https://explain.new/
- Google. "NotebookLM is now Gemini Notebook." 2026. https://blog.google/innovation-and-ai/products/gemini-notebook/notebooklm-gemini-notebook/
- Google. "Create a Video Overview." 2026. https://support.google.com/gemininotebook/answer/16454555
- Google. "NotebookLM video overviews." 2025. https://blog.google/innovation-and-ai/models-and-research/google-labs/notebooklm-video-overviews-studio-upgrades/
- Google. "Add sources to your notebook." 2026. https://support.google.com/gemininotebook/answer/16215270
- Synthesia. "Pricing." 2026. https://www.synthesia.io/pricing
- Synthesia. "Downloading videos." 2026. https://help.synthesia.io/en/articles/9317524-how-do-i-download-my-synthesia-video
- Synthesia. "Creating a video with Assistant." 2026. https://help.synthesia.io/en/articles/13759605-how-do-i-create-a-video-using-assistant
- HeyGen. "Pricing." 2026. https://www.heygen.com/pricing
- HeyGen. "PPT/PDF to Video." 2026. https://help.heygen.com/en/articles/13007313-ppt-pdf-to-video
- Microsoft. "Turn your presentation into a video." 2026. https://support.microsoft.com/en-US/PowerPoint/turn-your-presentation-into-a-video
- Canva. "Import and edit PDFs." 2026. https://www.canva.com/help/import-and-edit-pdfs-canva/
- Pictory. "Pricing." 2026. https://www.pictory.ai/pricing
- OpenAI. "Sora discontinuation." 2026. https://help.openai.com/en/articles/20001152-what-to-know-about-the-sora-discontinuation