Add recording
Choose a video or audio recording. Analysis starts automatically once the recording can be read.
See how you speak. Improve & Iterate.
Choose a video or audio recording. Analysis starts automatically once the recording can be read.
No. Your recording is not sent to ChatGPT, Claude, Gemini or another cloud chatbot. Speech Coach uses small speech and vision models directly on your device, then applies fixed coaching rules to the results. This keeps your recording private and avoids sending your speech to a generative-AI service.
No recording upload is needed for analysis. For the strongest privacy check, finish Prepare first, then switch off Wi‑Fi or use Airplane Mode before choosing the recording.
Speech Coach was created by the DeepMirror.me team as part of our work on privacy-first tools that help people reflect, learn and communicate better.
We want practical self-development tools to be more accessible without asking people to pay monthly subscriptions or give up sensitive personal data. Speech Coach is being developed as a public-benefit DeepMirror.me tool focused on useful, private feedback rather than advertising or data monetization.
You can choose video or audio from iPhone, Android, Mac or PC. Common formats such as MOV, MP4, M4A, MP3 and WAV are accepted. Analysis starts automatically when the recording can be read on your device. If a particular codec cannot be read, Speech Coach tells you rather than uploading or converting it elsewhere.
You can enable “Notify me when done”. Speech Coach also uses sound and vibration where available. You can use the FAQ and other controls while analysis runs. Some phones still slow or pause browser work if the whole tab is sent to the background, so keeping this tab open gives the most reliable result.
Speech Coach loads its speech, video and audio tools onto your device before you choose a recording. It also checks that the local media decoder is actually working before showing Ready.
Speech recognition and body/face analysis run in your browser using local machine-learning models. For transcription, Speech Coach uses longer overlapping listening windows so words near a boundary keep their context. The coaching report combines transcript, voice, pose, hand and facial signals using deterministic rules. No cloud generative language model is used for the analysis or report.