Capture Guide

What to record. How much. How often.

Exact specifications for the voice, photos, and video our AI needs — so future generations get the most lifelike legacy possible.

The four output goals — what your captures unlock

Each Heirloom capture produces multiple outputs. The more you capture, the more outputs become available. Here's exactly how much you need for each one.

A

Audio archive (the foundation)Basic

Minimum: 5 minutes. Sweet spot: 20 sessions × 15 minutes = 5 hours. Maximum useful: 50 hours.

Every memoir starts with audio. Each session can be a single 15-minute call answering 3–5 guided questions. Quality requirements:

  • Recorded in a quiet room (no fan, no TV in the background)
  • Phone microphone is fine — no studio needed
  • WAV or MP3, 44.1 kHz mono is enough
  • Single speaker (you can both speak — we identify and separate)
  • WhatsApp voice notes work perfectly for daily capture
V

Voice clone (recognizable likeness)Recognizable

Minimum: 30 minutes of clean voice. Sweet spot: 60–90 minutes. Maximum: 6 hours for a truly natural clone.

The voice clone needs clean audio without background music, multiple speakers, or phone-line distortion. From any 30-minute clean session you'll get a recognizable clone. The richer the audio, the more natural the clone — including emotion, laughter, and natural pauses.

  • 30 min = recognizable but slightly synthetic
  • 60 min = clearly them, mostly natural
  • 2 hr = indistinguishable from a real recording
  • 6 hr+ = full emotional range; whispers, laughter, tears
P

Photo archiveVisual memoir

Minimum: 10 photos. Sweet spot: 30–50 photos. Maximum useful: 200 photos.

Photos serve two jobs: the print-ready Memory Book and the family archive. Quality requirements:

  • Minimum 10: 4 childhood, 3 young adult, 3 elder. Faces visible, well-lit, front-facing.
  • For the Memory Book: 20–30 photos your family curates and places in the chapters they choose
  • Phone photos OK; old paper photos scanned at 300 DPI minimum
  • Avoid: tiny resolution thumbnails, group photos where face isn't visible, sunglasses or shadow on face
F

Live video (most powerful, optional)Best output

Minimum: 30 seconds. Sweet spot: 5–10 minutes per major life chapter.

Live video of the subject answering questions is far more powerful than animated still photos. If they're alive and willing, capture short video clips of them telling key stories. Use cases:

  • 30-second clips for each life chapter (childhood, marriage, career, parenthood)
  • 1-minute clip per major story (how they met spouse, day they emigrated, war story)
  • Phone camera in landscape, eye-level, 4K if possible (1080p is fine)
  • Good lighting (window light is enough)
  • Kept in your family archive alongside the Memory Book and the interactive AI experience
D

Documents + ephemeraBonus

Optional but transformative for the Memory Book.

Letters, marriage certificates, immigration papers, military discharge papers, recipe cards in their handwriting, tickets from key events. Scan or photograph. Your family can weave these into the Memory Book chapters as you review and edit them.

The auto-processing pipeline (what Heirloom does behind the scenes)

  1. Ingest: WhatsApp/phone-call/app-recorded audio uploads to encrypted storage. EXIF data stripped. Each file tagged with memoir_id + session_id.
  2. Transcribe: Deepgram processes audio → text with speaker diarization (separating you from the subject). Transcript stored per session.
  3. Translate (if needed): Source language preserved; English translation generated for descendants who don't speak the original.
  4. Theme extract: Anthropic Claude reads the transcript, identifies which question(s) it answered, tags themes (childhood, marriage, immigration, etc.).
  5. Voice model train: Once 30+ minutes of clean audio accumulates, ElevenLabs voice-clone build is queued (consent gate must be PASSED).
  6. Photo enrich: Photos are face-detected, age-estimated, OCR'd for text in the image. Auto-tagged for chapter alignment.
  7. Memoir compile: When the family compiles the Memory Book, AI assembles draft chapters using only source-linked content (no hallucination). Your family reviews and edits every chapter, then downloads the print-ready PDF.
  8. AI ancestor: Voice clone + transcript bank + photo gallery → conversational interface. Strictly source-linked. Says "I don't know" when asked something the subject never spoke about.

Every step logged for audit. Every transcript reviewable. Every model deletable on request.

For best results — the 90-minute starter recipe

If you want to give your loved one the best possible Heirloom experience, give us this in week one:

That's a complete, production-ready foundation. Everything after week one builds on this.

What to avoid (so we don't have to ask twice)

Ready to capture?

Sign up, and we send the first questions tonight.

Start a memoir →