No. The models run on your own device, so a client call, an interview or a lecture recording never leaves it. That also means it works with the network off entirely.
Yes - audio and video both, covering m4a, mp3, wav, aiff, caf, mp4 and mov. Importing is often the point: the recording you most need a transcript of is usually not the one you made.
Yes. Speaker identification separates the voices and you can name them, which is what turns a wall of text into something you can quote from accurately.
Yes - a keyboard shortcut (Control-Option-D) starts dictation wherever the cursor already is, so the text lands in the document you were writing rather than in a separate window to copy from.
Plain text and Markdown for notes and documents, SRT and WebVTT for subtitling the video the audio came from, and JSON when the transcript is feeding another tool rather than being read.
Well over ninety depending on the engine, including German, Spanish, French, Italian, Portuguese, Dutch, Polish, Czech, Ukrainian, Arabic, Hebrew, Hindi, Chinese, Japanese and Korean.
iPhone (iOS 17), iPad (iPadOS 17) and Mac (macOS 14) and later.