Short answer: fixed scrolling advances on a timer; voice following advances from recognized reading position. Choose fixed speed for rehearsed, time-coded delivery. Try voice following when pauses, rewrites, and repeated lines are part of the take.
Fixed scrolling vs. voice following
How voice following works
- The microphone provides live audio to the system speech-recognition service.
- The app compares recognition fragments with text near the current script position.
- The highlight and reading area advance only when the match meets the progression rules.
- During a temporary miss, the last trusted position remains; matching resumes after enough context is found.
CueBuddy for iPhone and iPad uses Apple's Speech framework. The current implementation does not require on-device recognition, so network needs can vary by device, language, and system support. Voice following should not be described as universally offline.
How to choose and test
- Record the same 60–90 second script once with each mode.
- Note manual corrections during pauses, repeated lines, noise, and mixed-language passages.
- Retest on the exact device, OS, language, and network conditions used for the real shoot.
- If recognition is unavailable or the room is noisy, switch to fixed scrolling and keep a manual pause control ready.
Limits and evidence
This page explains behavior and selection criteria; it does not promise fixed latency, offline recognition for every language, or accuracy in every environment. The implementation boundary was reviewed against the iOS source on 2026-08-11; production use still requires an on-device test.
Sources: Apple Speech audio-buffer requests; Apple on-device recognition support conditions.
Start your next recording with CueBuddy
Try the free teleprompter workflow first, then open CueBuddy if you need stronger recording and camera workflows.