Skip to main content
v2026.11,610 entries · CC-BY 4.0
LAC HealthLaboratory & ResearchLab & research supplies.Reagents, consumables, PPE & instruments — documented, fast, chain-of-custody shipping.Shop lac.us lac.us

How to Write for JoVE: Manuscript Text and Video-Narration Script Conventions

JoVE pairs a written protocol manuscript with a filmed, narrated video as the primary article format. This guide covers JoVE’s title/abstract limits, manuscript structure, and how to write and script protocol text for filming and narration.

JoVE (the Journal of Visualized Experiments) is a peer-reviewed publication whose primary article format is not a text-only paper but a pairing of a written protocol manuscript with a filmed procedural video, produced with recorded narration walking through the method step by step. This is a structurally different requirement from a video abstract, which is a short promotional summary of a paper that has already been published elsewhere, or a graphical abstract, a single static image. In JoVE’s model, the video is not a summary of the article — the video and the text manuscript together are the article, and the written manuscript is prepared alongside a video-narration script from the outset. This guide covers JoVE’s stated format requirements for the written manuscript, the structure and conventions of the video-narration script, and how writing protocol text meant to be filmed and narrated differs in practice from writing standard journal prose.

How the JoVE format differs from a video abstract or graphical abstract

Researchers preparing a JoVE submission sometimes conflate it with the video-summary formats several other journals now accept. The distinction matters because it changes what you are actually writing:

JoVE video-article format Video abstract (other journals) Graphical abstract
The video and written protocol manuscript are co-primary; together they constitute the peer-reviewed article itself A short (typically a few minutes) promotional/explanatory video accompanying an already-complete text article A single static image, not a video at all
Written for a film crew or the author’s own camera setup: shot descriptions plus a word-for-word narration script Written as a spoken summary of findings already reported in the paper Not narrated; a self-explanatory visual only
Structured around demonstrating a reproducible procedure (Introduction, Protocol, Representative Results, Discussion) Structured around motivation, method (briefly), and key finding Structured as one panel: inputs, process, outcome

See CASRAI’s video abstract guide and graphical abstract guide for those adjacent, but distinct, formats. If your paper is already published elsewhere and you want a short video to promote it, one of those is the correct format — not JoVE. JoVE is the right venue specifically when the contribution itself is a laboratory or clinical protocol, and demonstrating the procedure on camera is part of what makes the method reproducible.

JoVE’s written manuscript: format requirements

JoVE publishes its formatting requirements directly for authors (jove.com/authors). The core constraints an author should plan around before drafting:

  • Title: maximum 150 characters, in standard sentence/title capitalization — JoVE’s own guidance specifically notes that titles should not be written in all caps.
  • Abstract: maximum 300 words. The abstract text must also appear within the body of the manuscript itself, not only as a separate submission field.
  • Manuscript sections built around the procedure, not around a hypothesis-testing narrative: a JoVE manuscript is organized to support step-by-step demonstration rather than the IMRaD (Introduction, Methods, Results, Discussion) structure of a conventional research article. The Protocol section is written as an explicit, numbered, reproducible sequence of steps — closer in register to a laboratory SOP than to conventional Methods prose — because it is the section a reader (and the videographer) will follow directly.
  • Representative Results: unlike a conventional Results section reporting a study’s full findings, this section shows results that demonstrate the protocol worked correctly (and, where useful, what it looks like when it does not), oriented toward helping another lab confirm they’ve executed the method correctly.

Because the written manuscript and the video are prepared together, format decisions in the text (how a step is broken up, what order steps appear in, what counts as one “step” versus several) directly shape what the video has to show. Draft the Protocol section with the video in mind from the first pass, not as a separate adaptation step afterward.

The video-narration script: structure and conventions

JoVE’s guidance for author-produced videos (published as the Author Submitted Video / Author Produced criteria on jove.com/authors) specifies that the video itself follows a chaptered structure mirroring the manuscript, typically:

  1. Introduction — brief on-camera or narrated framing of why the protocol matters and what problem it solves.
  2. Protocol — the demonstrated procedure itself, step by step. JoVE’s guidance is explicit that this section should make up the majority of the video’s runtime, since demonstrating the method is the point of the format.
  3. Representative Results — a narrated walkthrough of what successful execution produces.
  4. Discussion / Conclusion — brief closing narration on significance, pitfalls, or variations.

JoVE’s guidance calls for each of these sections to be visually separated with a clear chapter title card, and for the script to be written out in full before filming — not filmed from an outline or improvised on set. A JoVE-style script is written in two parallel columns or clearly labeled blocks: what the camera shows (the shot list — camera angle, what’s in frame, any on-screen text or graphic) and the exact words of the voice-over for that shot. Authors are expected to script the Introduction and Discussion narration in full, not only the Protocol steps, since those framing sections are easy to under-prepare and then film loosely. JoVE’s guidance also specifies that the narration track should be a real human voice, not a synthetic/computer-generated voice-over.

Writing protocol text to be narrated and filmed, not just read

The single biggest craft adjustment for researchers used to writing conventional Methods prose is that JoVE protocol text has two audiences at once: a reader following along on the page, and a narrator (often the author) reading it aloud on camera while a demonstration happens in sync. That changes several habits that work fine in ordinary manuscript prose:

  • Write in short, sequential, imperative steps. “Pipette 200 µL of the reagent into the microcentrifuge tube” narrates and times naturally against a matching shot. A dense compound sentence describing three sub-steps and a rationale in one breath does not — it either forces the narrator to speed-talk over footage that hasn’t caught up, or forces an edit that fights the pacing of the action.
  • One step, one action, one shot. If a sentence in your draft protocol contains “and then” joining two physically distinct actions, that is usually a sign it should be two numbered steps, each with its own shot, not one.
  • Say quantities, instruments, and durations the way they’ll be spoken, not just written. “37 °C” reads fine on the page; write the accompanying narration line the way it will actually be said aloud (“thirty-seven degrees Celsius”) so the voice-over doesn’t stumble reading symbols cold in the studio.
  • Minimize spoken abbreviations and acronyms on first use in narration even where the written manuscript can use standard scientific shorthand after definition — a viewer hearing an unfamiliar acronym once, with no way to reread it, needs more context than a reader does.
  • Match narration length to the real duration of the action being filmed. A step that takes ninety seconds of actual bench time needs narration that fills roughly that time without padding or rushing; timing the read-aloud pace of a draft script against a rough estimate of filming time before shooting catches mismatches early, rather than in the edit.
  • Keep the camera’s-eye and the narrator’s voice in the same script document, aligned line by line, so a script reviewer (a co-author, or JoVE’s own editorial team if using JoVE’s production model) can check that what’s being said matches what’s being shown at every point, not just that each exists somewhere in the document.

Two production paths, and what each means for the script

JoVE offers more than one production route, and which one an author uses changes how much of the filming/scripting work falls to the research team versus JoVE’s own production staff. Under JoVE’s own videography model, a JoVE producer typically works from the author’s manuscript and Protocol text to help develop the shot list and film on-site or remotely with JoVE’s crew. Under the Author Submitted Video (ASV) route, the author’s own team scripts, films, and edits the video themselves, and that video must still meet JoVE’s published technical and content criteria for author-produced videos before it will be accepted — this is the model where having a complete, filming-ready script (not just a protocol summary) matters most, since there is no JoVE producer translating manuscript text into a shot list on the author’s behalf. Confirm current production-route availability, technical specifications (resolution, audio, file format), and the applicable author-produced-video criteria directly on jove.com/authors before committing to a shooting schedule, since these production-side details are updated independently of the manuscript text requirements.

Practical preparation checklist

  • Draft the Protocol section as a numbered, single-action-per-step sequence from the start — don’t write conventional Methods prose and adapt it afterward.
  • Write the full narration script (Introduction, Protocol, Representative Results, Discussion) before filming, not an outline to improvise from.
  • Keep title to 150 characters or fewer and abstract to 300 words or fewer, and make sure the abstract text also appears in the manuscript body.
  • Pair every narration line with its shot description in the same document, so mismatches between what’s said and what’s shown are caught on paper, not in editing.
  • Plan for a real human narrator; don’t budget for a synthetic voice-over.
  • Confirm which production route (JoVE videography vs. author-submitted video) applies to your submission before scripting, since it changes who needs the filming-ready script in hand.

Frequently asked questions

Is a JoVE video the same thing as a video abstract?

No. A video abstract summarizes a paper that is already complete and published (or being published) elsewhere in text form; see CASRAI’s video abstract guide. A JoVE video is not a summary of a separate text article — the filmed protocol and the written manuscript are prepared together as one peer-reviewed article.

How long should the video-narration script be?

JoVE does not publish a fixed word count for the script itself, since length depends on the protocol’s complexity; the working discipline is to time the narration against the real duration of each filmed step, and to keep the Protocol section as the majority of total runtime, per JoVE’s own guidance for author-produced videos.

Can I use text-to-speech or an AI-generated voice for the narration?

JoVE’s author-produced video guidance calls for a human narration track, not a synthetic or computer-generated voice-over — confirm current policy directly on jove.com/authors before production, since guidance in this area can be updated.

Does the written manuscript need to match the video script word for word?

They serve different purposes and are not identical documents: the manuscript’s Protocol section is the numbered, reproducible procedure as it would appear in print, while the narration script adds shot descriptions and the exact spoken phrasing for the video. They should describe the same procedure consistently, but the narration script is written to be heard, not simply read aloud from the manuscript text.

For related CASRAI reference material, see the Operational Definition entry (relevant to writing an unambiguous, reproducible Protocol section), the Supplementary Materials guide, and the ANSI/NISO Z39.14 entry on abstract-writing standards more broadly.

Referenced across the research world

University of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logoUniversity of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logo
  • University of Cambridge logo
  • Columbia University logo
  • Crossref logo
  • University of Edinburgh logo
  • Harvard University logo
  • University of Oxford logo
  • Princeton University logo
  • Stanford School of Medicine logo
  • University College London logo
  • ORCID logo

View CASRAI adoption →