How a rehearsal works
You bring one real talk: a defense, a conference talk, a pitch, a class. The coach works on recorded takes of it, one at a time, and every note it makes points at the exact moment it describes.
Say what you’re practising
A short guided setup asks what the talk is, what you want to say, and what you want it to achieve. Three answers plus consent is enough to start; everything else, audience, setting, style to preserve, specific concerns, accommodations, is optional and skippable.
Formats range from investor pitch to thesis defense to a situation you describe in your own words. Formats without a dedicated rubric yet are analyzed with the general speaking rubric and say so; interview practice is general practice, not answer-structure scoring.
Record a take, or upload one
Record right in the browser, with preview, countdown, and retake, or upload a recording you already have. Takes can be up to 20 minutes and 50 MB (mp4, mov, webm, or audio files).
If recording in your browser fails for any reason, upload always works.
Consent, plainly
Before anything runs you see exactly what will happen: only the audio of your recording is analyzed, which processors touch it, and how long it is kept (30 or 90 days, your choice). Declining means no analysis, full stop.
The take is transcribed and measured
The audio becomes a word-by-word transcript with timings, plus measured delivery signals: pace, pauses, fillers, emphasis, energy. Specialist readers then work through the take in parallel: content and structure (message, hook, through-line, signposting), vocal delivery, and engagement (from the language alone). When you say who the talk is for, a fourth reader checks audience fit; if you don’t, that section honestly says it was skipped rather than guessing an audience. It’s built for talks in English first.
Nothing visual in your recording is processed, even when you record video. Anything about your audience is judged only against what you stated, never inferred.
Measured, not guessed · sample values from a synthetic take
Pace156wpmreference 120-170Fillers1.9/minevery one timestampedPauses0.6s medianlongest ones locatedPitch rangeMeasurednever a personality scoreReference ranges, never targets. Numbers come from the recording’s audio; anything inferred is labeled as inferred, and nothing scores you as a person.
Every claim passes an evidence check
Before a note reaches you it must survive a validation gate: the timestamp has to exist, the quote has to match the transcript at that moment, the metric has to come from the measured signals, and uncertain inferences have to be labeled as such. Notes that fail are dropped and recorded, not shown.
You read the report next to your take
Strengths come first: two to four specific things worth keeping, at least one of them about your style. Then at most three priorities, each with why it matters for your stated aim, the evidence, a concrete suggestion, and an exercise. Measured numbers come with reference ranges, not targets, and nothing anywhere scores you as a person.
See a sample report built from synthetic fixture data. It’s labeled as such on the page; no real voice, no real person.
From a sample report · synthetic take · defense rehearsal, 12:40
5:12KeepYour aside “this surprised us too” is doing real work: it resets attention right before the method section. It’s yours; don’t polish it away.
9:47ReviewThe one-sentence version of your key result first appears here, nine minutes in.
“…so what this actually means is that the effect survives replication…”
Why it matters: your stated aim was for the committee to repeat this sentence. Give it the first two minutes, not minute ten.
Next take: focusOpen with the one-sentence result, then earn it.
KeepThe “surprised us too” aside and your unhurried closing minute.
How you’ll knowKey result stated before 2:00; filler rate steady or lower.
The next take is the point
The report ends with a three-line plan: one focus, what to keep, and how you’ll know it worked. Progress compares your next take to your last one, on measured metrics only. It never compares you to anyone else.
What it deliberately doesn’t do
- No person-scores. There is no overall grade of you as a speaker, anywhere, by design.
- No accent judgment. Accent is never labeled, inferred, or coached, and pacing rubrics account for speaking in a second language.
- No visual analysis of you. Your appearance, gaze, gesture, and expression are not processed in this version.
- No live overlay. This is a rehearsal room for recorded takes, not a meeting assistant listening in real time.
- No spoken opinions. The report is written. A walkthrough can read it aloud in the report’s own words, nothing added: a generated coach voice when your session agreed to it and the month’s voice budget allows, your browser’s voice otherwise.
Where it stands today
The full loop is live: a real recorded pitch has gone from upload through transcription, analysis, and the evidence check to a finished report. The pilot is free and invite-only while the product earns trust with a small group. Numbers about usefulness or improvement will appear here only once they’re measured on real, consenting users.
Start a rehearsal from the home page, or read how consent, data, and fairness work.