yaadein
We set out to save stories. People told us it changed their conversations.
Then two people called us back the next day.
This is a case study for what happens after the recording stops. yaadein turns the craft of asking deep questions into four rungs, writes each next question from the last answer, and the proof we care about is not what people recorded — it is what they did an hour later, and a week later.
Every bar is one of the 139 answers recorded in Wave 2. Its height is how long the person talked.YAAD
What yaadein is, right now
Three ways in, same engine behind them all.
A photo or video, a question, and the voice of why it matters. The on-ramp from an album to a memory.
See the design →Answer the L1-L4 ladder about yourself. A solo warm-up before you send a question to someone else.
How it works →Ask someone you love, their answer in their voice, every next question built from the last one. The hero mode.
How it works →Build on their own words. In Wave 1, the deepest question worked because it handed the person's own phrase back to them. Every rule in our engine protects that.
What we saw, before we built anything
Both of Shefali's parents are in memory care with advanced dementia. She grew up on the story of how they met, heard at the dinner table so often she could recite the beats. She has her version of that story. She will never have theirs.
That looked like a problem about dementia, or about running out of time. It isn't. Shefali is here and can speak, and nobody has ever asked her what it was actually like to grow up as an Indian migrant in Sydney and start again in America. Not because anyone is careless. Because when someone does ask, they ask "so what was it like?", and that question has no way in.
The bottleneck isn't time, storage, or the phone in your pocket. It's the question.
Restating the problem that way changed three things. The people we could test with stopped being only the elderly and became everybody. The barrier stopped being time and became a skill. And that skill turned out to belong to people who ask for a living, who can do it but can't tell you how.
What we believed, and what would prove us wrong
Our hypothesis changed twice, and the change is part of the story (see what changed). Where it stands now:
If an Asker uses questions built on a four-rung ladder, each written from the answer before it, the person answering will tell them something they have never heard before, and will talk longer the deeper the ladder goes.
What would prove it wrong
- The ladder's questions produce no more new disclosure than the Asker's own questions on the same topic.
- Answers don't lengthen as the rungs go deeper. If people talk just as long at the fact rung as the disclosure rung, the ladder isn't doing the work.
- People who ask genuinely good questions still get one-sentence answers. Then the barrier is the subject, not the question.
The baseline
In our first manual run, the fourth and hardest answer ran 1 minute 25 seconds in total across two unaided conversations. With framework-built questions on the same topics, it ran 4:16 (questions written by Claude) and 3:22 (an expert interviewer).
Wave 2 ran the app-guided conversation only. We did not capture an unaided comparison in Wave 2, so our baseline remains Wave 1's hand-timed result. When we tried running the unaided conversation a few times, participants got confused between the two conversations and recorded on their phone's voice memo app instead of through ours, so nothing comparable was captured. Running the unaided conversation through the app is our next step.
Our approach
We treated yaadein as an experiment with software attached, not software with a survey attached.
Eight steps, in order. The two marked in solid pink are where real people used it without us steering.
A lived problem, stated before any product.
A claim, and what would prove it wrong.
Wave 1: six conversations, no screen, no app.
Keep what broke, change what it showed.
The Mom Test app, to run more sessions.
Wave 2: Askers who aren't on the team.
Surveys for both people, evaluators on the recordings.
Including what failed.
Six rules we held ourselves to
Measure first, build second
The software is how we get to the measurement. We built only what the next test needed.
By hand before code
Wave 1 ran with no app at all, so we'd know what the app had to do before anyone wrote it.
Two people, not one
Both the Asker and the respondent are surveyed after every conversation, and the Asker's emotion counts as much as the respondent's.
Nobody performs for the score
Participants never see how we score. People who know emotion earns points may perform it.
The Asker stays in charge
A person always decides what gets asked. The model suggests; it never decides.
The diff is the deliverable
Every design change is traced to the evidence that caused it, including our own mistakes.
The ladder: four rungs, YAAD
Yaad is the root of yaadein, Hindi for memories. Each rung goes a little deeper, and costs the person answering a little more, than the one before. The first two are about the memory. The last two are about the person.
The app writes each next question from the last recorded answer. The Asker stays in charge: they can accept a question, reject it, or write their own. Off-limits topics can be set before the conversation starts.
What we ran
Wave 1, by hand. Six recorded conversations across two threads, three rounds each: the Asker's own questions, Claude's, then an expert interviewer's. Nobody answering ever saw a screen.
Redesign from the evidence. Expert round retired, two conversations on the same topic, nine measures defined with scales, surveys for both people after each conversation.
Mentor working session. Live demo of the Mom Test app with our mentor as the respondent, and a steer to prioritize new disclosure, stay close to users, and not let the technology become the bottleneck.
Wave 2, through the app. Askers and respondents recruited by the team and by referral, running the ladder through the Mom Test app, with every answer recorded, transcribed and timed.
How a session runs
Only two of the eight steps involve a model. Everything else is either a fixed rule or a person, and a person always decides what actually gets asked.
The ladder is why conversations keep going.








