Interactive Video Interaction Types: A Practical Guide
Choose the right interactive video interaction for recall, reasoning, reflection, context, action, spatial practice, or branching across supported sources.
Interaction types are tools for collecting different kinds of evidence. A multiple-choice question can reveal a misconception. An ordering task can reveal whether someone understands a process. A poll can surface a prediction. A hotspot can show whether a learner can locate evidence on an uploaded video frame.
This guide helps you choose across those families without treating the interaction picker as a list of interchangeable effects.

Start with the evidence you need
Before choosing a type, write what a useful response would prove. “The viewer interacted” is not enough. A useful statement is specific:
- The viewer can distinguish a safe action from a tempting unsafe one.
- The viewer can put the five setup steps in the correct order.
- The viewer can point to the damaged component in the frame.
- The viewer can explain which evidence supports the claim.
- The viewer has chosen the most relevant next resource.
Format-first
We should add a poll because the video needs more variety.
Evidence-first
We need to capture each viewer's prediction before the result is revealed, so an ungraded poll fits.
The five interaction families
Knowledge checks
Recognize, retrieve, calculate, sequence, or connect
- Typical types
- Multiple choice, true or false, numeric input, fill blank, ordering, matching
- What it produces
- A scoreable response to a focused question
Constructed responses
Write, speak, select evidence, or manipulate text
- Typical types
- Open response, audio response, evidence highlight, mark words, drag words
- What it produces
- The viewer's own answer or reasoning
Audience input
Predict, rate, choose a preference, or supply structured details
- Typical types
- Poll, rating scale, embedded form, timestamped comment
- What it produces
- Opinion, confidence, discussion, or submitted information
Context and next steps
Orient, read, reveal, navigate, or follow a relevant action
- Typical types
- Chapter, info card, timed reveal, call to action, Choice Moment, leaderboard
- What it produces
- Viewing structure, supporting context, or a chosen next step
Spatial and scenario interactions
Locate, label, draw, practise in a workspace, or choose a consequence path
- Typical types
- Hotspot, image label, annotation, drawing, workspace, complete branching scenario
- What it produces
- A response tied to the media coordinates or segment graph
Knowledge checks
Knowledge checks are useful when the response can be evaluated against a defined answer. Choose based on the thinking, not convenience.
- Multiple choice works well when each distractor represents a plausible misunderstanding.
- True or false is best for one precise claim whose conditions are clear.
- Fill in the blank asks viewers to retrieve a short term or phrase.
- Numeric input checks a calculation, quantity, or measurement with a stated unit and tolerance.
- Ordering checks whether a sequence or ranking is understood.
- Matching checks several relationships between concepts and examples.
Constructed responses
Constructed responses ask viewers to produce something rather than select a prepared conclusion. They are useful when reasoning, expression, or performance matters.
- Open response captures an explanation, recommendation, or reflection.
- Audio response captures pronunciation, rehearsal, or spoken reasoning.
- Evidence highlight records the selected source evidence and the viewer's explanation.
- Mark the words asks viewers to identify exact language in a short passage.
- Drag the words checks contextual vocabulary through a bounded word bank.
Be explicit when a response awaits human review. A provisional score should not be framed as a final result while written, spoken, or submitted work is still pending.
Audience input
Not every valuable response has a correct answer. Polls and rating scales can make prior beliefs or confidence visible. Timestamped comments can anchor discussion to the exact evidence. A form can collect several related fields at one meaningful moment.
Context and next steps
Passive and navigational interactions can improve a video without asking another question. Chapters organize a long recording. Info cards explain a term. Timed reveals disclose an answer. Calls to action connect the current moment with a relevant resource. A leaderboard can make available standings visible where comparison is appropriate.
These tools still need a purpose. A card that repeats the narration or a call to action unrelated to the current moment adds visual load without helping the viewer.
Spatial interactions and branching
Spatial interactions treat the video frame as a coordinate system. A hotspot asks the viewer to select a region. Image labels place targets on the visual. Annotation and drawing add marks to the frame. A workspace lets the learner practise in a contained activity. Interakly must own the media surface to save and reproduce that geometry, so these interactions require uploaded video.
Complete branching is a different structural tool. The learner makes a decision, moves into another uploaded segment, and can reach a different consequence or ending. Plan it as a route, not as a single card.
YouTube and uploaded-video compatibility
YouTube playback runs inside YouTube's embedded player. Interakly can place compatible non-spatial questions and cards over that player, but it cannot use YouTube's pixels as its own authored coordinate system.
Verified product inventory
Current YouTube-safe interaction set
These groups are sourced from the same human-curated inventory that is tested against Interakly's client and server allowlists.
Knowledge checks and text activities
Use these when the viewer should demonstrate recall, comprehension, sequencing, or evidence-based reading.
Responses and audience input
Use these to collect explanations, opinions, confidence, structured information, or spoken answers.
Context, navigation, and participation
Use these to structure the experience, reveal information, present a next step, or make participation visible.
Uploaded-video boundary
Tools that need Interakly to own the media surface
Hotspot
Tap the right spot on a frame
Image Label
Place labels on an image
Annotation
Draw on the video frame
Drawing Submit
Submit a whiteboard drawing
Workspace Check
Validate workspace state
Embed
Embed external content
Advanced logic: Variable is an uploaded-video-only set or check control. It changes journey state behind the scenes rather than presenting another viewer-facing card. Complete branching scenarios are also uploaded-video experiences. They connect distinct video segments into a route and are not equivalent to a single navigation menu overlay.
Choose whether playback pauses
- Pause when the viewer must answer, read carefully, inspect a stable frame, or make a route decision.
- Keep playing when a brief contextual card or structure cue remains understandable without stopping.
- Preview the transition whenever a passive card could cover important evidence or a blocking prompt could interrupt a sentence.
The interaction type does not decide this alone. The same information card may be safely passive over a slow product shot and distracting over a dense diagram.
A five-question selection method
What must the viewer do?
Choose one observable action such as retrieve, calculate, organize, explain, predict, locate, or continue.
What response would count as evidence?
Define whether you need a scoreable answer, reasoning, opinion, spatial selection, or next-step choice.
Which source supports it?
Check whether the interaction is YouTube-safe or requires an uploaded video surface or segment graph.
Should playback stop?
Pause only when responding or reading requires the video to hold its place.
How will the loop close?
Plan feedback, explanation, results, review status, or a relevant next action before publishing.
Accessibility and mobile checks
- Write a complete text prompt that does not depend on color or position alone.
- Add meaningful alternative text for question images and visual evidence.
- Provide captions or an equivalent transcript for spoken content.
- Do not require a precise spatial click when the target can be made larger and clearer.
- Test keyboard navigation, focus order, feedback, and any recording permission flow.
- Preview the complete interaction inside a phone-sized player, not only the editor.
- Check that concurrent elements remain understandable together and do not compete for the same space.
Selection rule: choose the least complex interaction that can produce the evidence you need. Complexity is worthwhile only when it changes what the viewer can practise or what you can learn from the response.
FAQ
What are interactive video interaction types?
Interactive video interaction types are structured ways for viewers to respond or receive context during playback. Examples include multiple choice, polls, open responses, information cards, timed calls to action, hotspots, and branching decisions.
Which interaction type is best for checking understanding?
Use the lightest format that produces the evidence you need. Multiple choice is useful for diagnosing a misconception, fill in the blank checks retrieval of a short term, ordering checks a process, and open response reveals reasoning that prepared options may hide.
Which interactive video types work with YouTube?
Interakly currently exposes 21 non-spatial interaction types in its verified YouTube editor profile. The YouTube source must still permit embedding. Spatial tools such as hotspots, image labels, annotation, drawing, workspaces, and embedded iframes require uploaded video. Variable-based advanced logic also requires uploaded video.
Is a Choice Moment the same as a branching video?
No. A YouTube-compatible Choice Moment can present navigation options, but Interakly's complete branching scenario workflow connects uploaded video segments into different paths. Full segment-based branching requires uploaded video.
Should every interaction pause the video?
No. Pause when a response is required or when continuing playback would remove the context needed to respond. Chapters, brief information cards, comments, and some calls to action may remain passive when the viewer can understand them without stopping.
Can several interactions appear at the same time?
Interakly can render simultaneous active interactions inside the same scaled video interaction layer. Use that capability sparingly. Several prompts competing for the same attention can be technically valid and still create a confusing viewing experience.
What Is Interactive Video?
Start with the broader format, its mechanics, and the difference between interaction and ordinary playback.
What Interactions Can You Add to a YouTube Video?
Use the source-specific, tested inventory when your project starts with YouTube.
How to Add Quizzes to Video
Turn a question family into a planned sequence with feedback, settings, and review.
20 Interactive Video Examples
See these interaction families applied to practical teaching, training, and communication moments.
Interactive Video Best Practices
Plan timing, questions, feedback, accessibility, and quality assurance across the complete experience.
Choose the response before the component
Define the evidence you need, confirm the source supports it, and use the lightest interaction that can collect it clearly.
Get started free