Intuify

Intuify

Documentation

Question types4 min readUpdated 2026-08-19

Guided Story With Image Capture

Guided camera capture with an optional voice narrative.

On this page (7 sections)

Guided Story With Image Capture (plugin type unified_image) guides the respondent through capturing up to 4 photos, each with its own instructions, then optionally records a spoken narrative after the photos. It's wizard-driven.


What respondents see#

A greeting screen introduces the exercise. The respondent is guided through capturing a fixed number of images (1–4), one at a time, each with its own on-screen instruction text and — if configured — a semi-transparent overlay template to help them frame the shot and an orientation requirement (flat, portrait, or landscape). After the photos, an optional voice follow-up narrative records their commentary.

When to use it#

  • Visual research that also wants qualitative context — "show us your pantry/desk/closet and tell us about it" — without splitting it into two separate questions.
  • In-home audits, unboxing studies, or "show me your setup" research where a fixed, known number of photos is expected.
  • Anywhere Pantry Checker's variable photo count isn't needed — use this when you know exactly how many images you need.

Configuring it in the builder#

Drag Guided Story With Image Capture from the StorySync AI Qual group onto your canvas — the wizard opens with Questions, Display Settings, and Text for Alerts. The Questions step's Image Section is the distinctive part:

  • Greetings Text and Instructions — the welcome screen(s), splittable across multiple slides.
  • Number of Images — how many photos the respondent needs to capture (1–4).
  • Ask voice follow-up after image capture — toggle the narrative step on or off.
  • Per image (1–4): Instruction Text, Overlay Template (URL of a semi-transparent framing guide), Select Orientation (Any / Flat / Portrait / Landscape), Custom Rules, and Image Tags (for AI-powered categorization used in conditional checks).

Settings reference#

SettingWhat it controlsDefault
Number of ImagesFixed photo count the respondent must capture2
Ask voice follow-up after image captureWhether a spoken narrative follows the photosOn
Overlay Template (per image)Framing guide overlay shown during captureShared default overlay
Select Orientation (per image)Required device orientation for that shotAny (no requirement)
Is Loaded from IframeMarks the question as embedded in an iframe inside an external surveyOff
Auto CaptureAutomatically take the photo once framing guidance is satisfied, instead of a manual shutter tapOff
Real Time Image GuidanceLive on-screen feedback while framing the shot (only shown when Auto Capture is on)On
Allow Image UploadLet respondents upload an existing photo instead of using the camera liveOff
AI SourceWhich AI engine analyzes captured images (for Image Tags / Custom Rules)Gemini (or ChatGPT)
Voice MandatoryVoice input required for the follow-up narrativeOff
Honor voice-permission globalsFollow survey-level voice/mic-permission setting, auto-skip if unavailableOn (recommended)
Default to VoiceVoice recording is the default input method for the follow-upOn
Enable TranscriptionLive transcript while the respondent talksOn
Enable Voice Attribute DetectionDetect speaker gender/age/language/accent/human-vs-AIOn
Tap to Zoom ImageLet respondents tap the concept image to zoom inOff
Enable Image ScrollingLet respondents scroll through multiple imagesOff
Enable Auto ProbingLet the AI generate its own follow-up questions on the narrativeOff
Use Whole ConversationFeed the entire prior conversation into AI-generated probesOff
Request Voice or Text PreferenceAsk respondents whether they'd rather record or type (only when voice isn't mandatory)Off
Concept Image HeightHeight of the stimulus image as % of screen50% (range 30–70)
Question Font SizeFont size for question/instruction text17px (range 10–40)
Maximum Recording DurationCap on the follow-up narrative5 minutes (range 3–10)
Select LanguageSpeech-recognition languageEnglish (+ 6 others)
Select Text DirectionLTR or RTL layoutLeft to Right
Button ColorPrimary action-button color#34ac8b

Logic support#

Stores captured images plus an optional open-ended narrative — condition operators are limited to Answered/Not answered. Image Tags can drive conditional checks within the structured story itself. See Logic, Flow & Conditions for survey-level condition operators.

Good to know#

  • Custom Rules and Image Tags exist specifically to support AI-powered analysis and conditional logic inside the structured story flow — set these deliberately if downstream steps depend on what was photographed.
  • Unlike Pantry Checker, the image count here is fixed at build time, not a respondent-chosen range.

Was this page helpful?