- Home
- Speech to Text
Dashboard
How do you want to transcribe?
Free minutes are included. Upload a file or record audio to start.
Turn a voice memo to text on this page: import a memo exported from your phone, or record a new note with the browser microphone, then start a real transcription task in the workspace below. When the draft is ready, review the wording and timestamps, fix names and numbers, copy the passages you need, and download the edited transcript as TXT or another format. Verified new accounts start with the current transcription-minute allowance, shown before you begin.
Dashboard
How do you want to transcribe?
Free minutes are included. Upload a file or record audio to start.
Verified accounts use the current Whisper AI transcription allowance. Audio is processed by the active cloud transcription provider, and finished memo tasks are saved in your recordings with the voice-memo-to-text source tag.

This is the reviewed text for a 30-second English note, locally synthesized for this example: a quick hand-off with a date, a headcount, and a deadline. The timestamps show the editor view; a plain TXT export contains the text without them.
Example: 30-second English note, locally synthesized
Reviewed transcript
This is the reviewed text for a 40-second English memo, locally synthesized and imported as an M4A file for this example. It captures two site observations and a follow-up the speaker promised a colleague.
Example: 40-second English memo, imported as M4A, locally synthesized
Reviewed transcript
Start from the file your phone already produced and let the workspace create a timestamped draft you can correct.
On a phone, share or export the voice memo to Files or your device, then upload that file here. iPhone voice memos arrive as M4A and Android recorders often produce M4A, MP3, or WAV; the workspace checks the actual media rather than trusting the file name. When you know the spoken language, set it before starting so a short memo is not guessed from a few seconds of audio.
The workspace states the current minute and file-size limits before the task runs and shows the credit cost against your balance. Nothing is generated until you start the task, so opening this page never spends transcription minutes.
When the draft is ready, replay the timestamps and scan for names, dates, numbers, and product terms, which are the details a quick spoken note carries and a machine pass can still mishear. The first result is a draft to correct, not a finished document.
When there is no file yet, record the note directly on this page and keep the recording, the draft, and the download together.
Switch the workspace to recording, allow microphone access when the browser asks, and speak the note. The timer shows the length while you record, and a recording tip keeps the audio clear without a separate app. If you prefer not to grant access, the file upload path stays available and no recording is started.
Stop the recording, play it back to check that the first and last sentences are complete, and re-record if the note was cut off or contains a long silence. Recording continues only while the tab is open, and the microphone track is released when you stop, cancel, switch back to upload, or leave the page.
Send the confirmed recording to the same transcription task. The browser recording is handled exactly like an uploaded file, so the finished memo lands in your recordings with the voice-memo-to-text source tag and can be reopened instead of recorded again.
The editor keeps your corrections with the task, so the file you download matches the version on screen instead of an older machine draft. Whether you transcribe voice memo audio from an imported file or a browser recording, the review step is the same.
Correct names, dates, numbers, and specialist vocabulary directly in the transcript text. Replay the timestamp for any uncertain passage before changing it, and keep the speaker's meaning rather than smoothing the note into different words.
Select and copy the lines you want for a task list, a message, or a document without leaving the page. Keep the timestamp beside any line you plan to check later so the original moment is easy to find.
Download TXT for a plain, portable memo, or use the workspace exports for DOCX, JSON, SRT, and VTT when the words need to become a document, structured data, or captions. Download after saving so the exported file matches the current version.
Comparison describes the checked pages at a high level. It does not repeat competitor pricing, ratings, certifications, or accuracy claims, and naming a product is not an endorsement.
| Decision factor | Whisper AI | Otter | Notta |
|---|---|---|---|
| Starting from a phone voice memo | Import the exported memo (M4A or another supported format) on this page, record a new note in the browser, or import a supported public media URL | Confirm the current upload, recording, and import options on its live tool | Confirm the current upload, recording, and import options on its live tool |
| When the microphone is blocked | A denied microphone stops only the recording step; the file upload stays available, so a saved memo can still be transcribed | Confirm how the live product behaves when mic access is refused | Confirm how the live product behaves when mic access is refused |
| Editing before export | Fix wording, replay timestamps, copy passages, and keep edits saved with the task before downloading | An editor is offered; confirm which edits the current plan allows | An editor is offered; confirm which edits the current plan allows |
| Download formats | TXT download, plus DOCX, JSON, SRT, and VTT from the same transcript | Confirm the live export list and any plan gating | Confirm the live export list and any plan gating |
| Where the memo lives | The finished task is saved in your recordings and tagged with its voice-memo-to-text source | Confirm how the live product stores and reopens items | Confirm how the live product stores and reopens items |
| Trying the workflow | Verified new accounts start with the current transcription-minute allowance, shown before the task runs | Check the current free and paid plan terms | Check the current free and paid plan terms |
Whisper AI is the strongest fit when the input is a phone voice memo or a quick browser note and you want recording, upload, editing, and download to stay on one page, with the finished task traceable in your recordings. Check each product's live terms before processing a long or confidential recording.
Yes. Export or share the voice memo from your phone to a file, then upload it here. iPhone memos arrive as M4A and Android recorders often produce M4A, MP3, or WAV, so keep the original export and let the workspace check the media instead of renaming an unsupported file. A single upload starts one voice memo transcription task, which keeps the timestamps and the transcript easy to check.
Yes. Switch the workspace to recording, allow the microphone when the browser asks, and record the note on the page. The timer shows the length, and playback lets you confirm the start and end before you submit. The recording is transcribed like any uploaded file, and the finished memo is saved in your recordings instead of living only in the browser.
Only the recording step is affected. The browser reports that audio permission is required, the recording does not start, and the file upload path stays available, so you can still import a saved voice memo and get the same transcript. Nothing is charged for the blocked attempt because no transcription task is created without audio.
Upload the media your recorder produced, such as M4A, MP3, WAV, or another supported audio format, and use the dedicated M4A or MP3 pages if you want a format-specific walkthrough. The workspace checks the actual audio stream, so a file that was renamed to a supported extension but contains unreadable audio fails with a clear message, and your voice recording to text result never comes from a file the workspace could not read.
Yes. When the draft is ready, open it in the workspace, replay any timestamp, and correct names, dates, and wording directly. Your edits are saved with the task, and the TXT or other export is built from the version you saved, so the download matches what you reviewed on screen.
Yes, once a task has been created. Signing in, returning from a recharge, or refreshing restores the task and its status instead of starting a second transcription, so you are not charged twice. A refresh that happens before the task is created and after the browser has dropped the local file means choosing the memo again to start.
A recording can continue while you switch to another tab, but the microphone is released as soon as you stop or cancel the recording, switch the workspace back to upload, or leave the page, so there is no hidden recording afterward. Closing the tab ends the recording; keep the page open until the note is submitted if you want the capture to finish.
This is a cloud transcription workflow, so the audio you submit is handled by the active transcription provider and stored so the task can be processed and reopened. Review the privacy policy before uploading confidential, medical, or client material, and avoid uploading recordings you are not authorized to process.
Written and checked by the Whisper AI Editorial Team · Last reviewed 2026-09-23 · Editorial policy