BETA

The service is in beta. You may see slight delays — thank you for your patience. Hit a problem? Write to us

Explainer

Hebrew speech to text

Updated 20 September 2026
A speech model built for spoken Hebrew

Speech to text turns spoken Hebrew into written Hebrew. It runs on a model trained on recorded speech, and the quality of the result depends far more on the recording and on how well the model knows the language than on anything you set in the interface. IvreetMeet runs its own Hebrew model, adds speaker separation and a summary, and gives you 120 free minutes a month to test it on your own audio.

What happens between the sound and the text

StageWhat it does
Audio preparationThe recording is read and the speech is located inside it
RecognitionA Hebrew speech model writes what it hears, word by word, with timings
Speaker separationThe recording is cut where the voice changes, and each passage is attributed
AssemblyPassages become a readable transcript you can search and quote
SummaryA language model reads the transcript text and writes key points, decisions and action items

Hebrew is not English with different letters

Three things make Hebrew speech recognition its own problem. Vowels are not written, so the same letters can be several different words and only context decides. Prefixes and suffixes attach directly to words, so a model that has learned whole-word patterns from English has to learn a different shape. And spoken Hebrew at work is full of English: product names, job titles, whole phrases. A model trained on Hebrew speech expects all three; a general model treats them as noise.

What you can control

  • Distance to the microphone. The single biggest factor, every time.
  • One microphone per speaker when possible, which is what Zoom, Meet and Teams recordings give you.
  • The original file. Re-compressing to save space throws away what the model needs.
  • Context for the summary. Naming the participants and the purpose changes what the summary emphasises.

Accuracy claims, generally

Treat any single accuracy number you read — ours included — as a measurement on one test set under one set of conditions. What matters for you is how a tool behaves on your own recordings, which is exactly what a free tier is for.

Frequently asked questions

Is there free Hebrew speech to text online?

Yes. Up to 60 seconds with no account at all, and 120 minutes per organization per month once you sign up. Details on the free page.

Does dictation work, not just recordings?

The service works on recordings you upload. Live dictation inside the product writes what you say as you speak; the guides here are about recorded files.

Is there an API or a model I can call?

There is a REST API with a personal key, and an MCP server that plugs your meetings into AI assistants. See API and MCP. The model itself is not published for download.

Why do general speech tools do worse in Hebrew?

Because they are trained mostly on English. Hebrew writes no vowels, glues prefixes onto words, and mixes English terms into ordinary speech; a model that has not heard much of it guesses.

What about speaker names?

Separating speakers inside one recording is automatic. Recognising the same person in later meetings uses a voiceprint, which is biometric data: it is off by default and only ever applies to speakers whose consent was recorded.

Try it on your own recording

IvreetMeet transcribes and summarizes Hebrew meetings. The free plan covers 7 meetings and 120 minutes per organization each month.

Start free