English text to speech

Two people either side of a desk with a studio microphone, headphones, a speaking head, an ear and a musical note โ€” paste your own English and listen to how it sounds.

Paste your own English โ€” a paragraph you have written, an email you are about to send, a sentence you cannot get your mouth around โ€” and hear it read back. Slow it down to catch the endings, replay one sentence as many times as you need, and compare it with the way you say it. The voice is the one already built into your device; nothing is uploaded and nothing is recorded.

Your own textSpeed 0.5ร—โ€“1.5ร—Sentence by sentenceNo sign-upNothing uploaded

Hear your own English read aloud

Paste anything you have written. It is split into sentences, and every sentence gets its own play button.

๐Ÿ”’ No sign-up  Your text never leaves this tab โ€” the voice is your own deviceโ€™s.

0.9ร—

5 sentences

  • Last summer I decided to write to my old teacher.

  • I had not spoken to her for eleven years, and I was not sure she would remember me.

  • She wrote back in three days.

  • Her letter was two pages long and she remembered everything, including the essay I had been so proud of.

  • I have read it four times since then.

How to actually improve with this, not just listen to it

Playing a paragraph back once and nodding is not practice. Here is the loop that works, and it takes about four minutes.

1. Read it aloud yourself first. Before you press play. This matters more than anything else on the page: if you listen first, you will copy what you just heard and learn nothing about what you would have done. Say it your way, then check.

2. Play one sentence, slowed down. Use the speed control at 0.7 or 0.75. At full speed English word endings vanish into the next word and you will not hear whether the -ed was there. Slow speech is not natural speech, but it is diagnostic speech.

3. Listen for three things specifically. Which syllable is loudest in each long word; whether the sentence goes up or down at the end; and where the speaker does not pause. Learners almost always pause too often, at commas, which is where English speakers frequently do not.

4. Say it again with the audio. At the same time, out loud, over the top of the voice. This feels ridiculous and it is the single most effective thing on this list โ€” it forces your timing onto someone else's rhythm instead of your own.

5. Then speed it back up to 1.0 and check you can still keep up. If you can, that sentence is done.

Why hearing your own writing is different from listening to a podcast

There is no shortage of English listening material, and this tool is not competing with any of it. Listening to real speakers is how you learn to understand English. It is not, on its own, how you learn to produce it โ€” because when you listen to a native speaker you are hearing sentences you did not write, containing words you might never choose.

The gap most intermediate learners are stuck in is different: they can read and write far better than they can say. They construct a good sentence on paper and then cannot get it out of their mouth, because they have never heard it, and because English spelling does not tell you what a word sounds like. Thought, through and thorough are not a joke; they are a real problem you cannot solve by reading.

Putting your own sentence through a voice closes that specific gap. The sentence is already yours โ€” you chose the words, you know what it means, and there is no comprehension work in the way. All that is left is the sound, which is the thing you were actually stuck on.

This is also the fastest way to find out that a word you have used for years is a word you cannot say. Almost everybody has several. Our pronunciation pages cover the sounds themselves, with minimal pairs and the British IPA for 5,500 words.

What a synthetic voice is honestly good for

Two columns: what a synthetic voice models reliably โ€” word stress, endings, intonation and natural speed โ€” and what it cannot give a learner, including real emphasis, a genuine accent and any feedback on their own speech.
The left column is what to correct yourself against. The right column is what still needs a person.

The voice reading your text is your browser's speech synthesiser, not a recording of a person. It is very good at some things and cannot do others at all, and knowing which is which is the difference between a useful tool and a misleading one.

Trust it for: word stress, the shape of a word ending, the rise and fall of a statement or a question, and the speed real speech actually runs at โ€” which is faster than most learners expect. On all four it is reliable enough to correct yourself against.

Do not trust it for: emphasis and emotion, which it flattens; regional accent, which it has one arbitrary version of; unusual names and rare words, which it will guess at from spelling and sometimes get wrong; and โ€” most importantly โ€” any judgement about your own speech, because it cannot hear you.

It is a model of an accent, not a teacher. Use it to check the shape of a sentence, then use a human being to check yourself. If you have nobody to practise with, our guide to practising English speaking alone is about exactly that problem.

Why it reads one sentence at a time

A pasted paragraph on one side and the same text split into one playable row per sentence on the other, each row with its own play button.
The split happens before a word is spoken. That is what makes replaying one line possible โ€” and what stops a long paragraph being cut off by the browserโ€™s own speech queue.

The reader splits your text into sentences before it speaks a word, and gives each one its own play button. That is a deliberate design decision with two reasons behind it.

The first is pedagogical. Repetition is what fixes pronunciation, and repetition means going back to this sentence twenty times without listening to the four before it again. A single play button on a whole paragraph makes that impossible, which is why almost nobody practises properly with one.

The second is technical, and it is the sort of thing worth knowing if you use text-to-speech elsewhere. Browser speech engines truncate very long utterances โ€” pass a whole essay in one go and it will simply stop somewhere in the middle, usually without telling you. Splitting at sentence boundaries first means every chunk is comfortably short, so nothing is silently cut off.

A worked example

This is the example text loaded by the button above the box, and how the reader breaks it up. Five sentences, five buttons, each replayable on its own.

The example text

Last summer I decided to write to my old teacher. I had not spoken to her for eleven years, and I was not sure she would remember me. She wrote back in three days. Her letter was two pages long and she remembered everything, including the essay I had been so proud of. I have read it four times since then.

Split into 5 playable sentences:

1. Last summer I decided to write to my old teacher.
2. I had not spoken to her for eleven years, and I was not sure she would remember me.
3. She wrote back in three days.
4. Her letter was two pages long and she remembered everything, including the essay I had been so proud of.
5. I have read it four times since then.

Notice that the split is on sentence-ending punctuation, not on commas or line breaks, and that abbreviations such as Dr. do not start a new sentence. A sentence longer than about four hundred characters is capped rather than truncated mid-word โ€” if yours is longer than that, it is worth breaking up for a reader anyway.

That paragraph is also worth a second look for what it does to a learner's mouth. It contains wrote and written, thought and three, and an -ed ending said three different ways โ€” decided /ษชd/, remembered /d/, proud of with the linking that makes it one word. Slow it to 0.7 and every one of those is audible.

Take it further

The reader shows you a shape. These pages teach the sounds inside it.

Frequently asked questions

Is this English text-to-speech tool free?

Yes, completely, with no account and no limit on how much you read. It uses the speech voice already installed in your browser, so there is no service behind it to charge for.

Does my text get uploaded anywhere?

No. The text stays in the page and is passed to your own device's speech engine. Nothing is sent to a server, stored or recorded, and closing the tab removes it.

Which accent does it use?

It asks your browser for a British English voice and falls back to any English voice if there is not one installed. Which specific voice you hear depends on your device, so the same text can sound different on a phone and a laptop.

Why does it sound robotic on some words?

Because it is guessing the pronunciation from spelling, which English makes unusually hard. Names, borrowed words and rare technical terms are where it slips most. If a word sounds wrong, check it against the IPA on our vocabulary pages rather than trusting the voice.

Can I download the audio as an MP3?

No. The speech is generated live by your browser and never exists as a file, which is the same reason nothing has to be uploaded to produce it. If you need a recording, record your own voice reading it โ€” for a learner that is the more useful artefact anyway.

Can I use it to practise for a speaking exam?

It is useful for the pronunciation half of the job: hearing a model of the sentences you plan to use, and checking your own stress and intonation against it. It cannot assess you. For IELTS specifically, our IELTS speaking pages have real part 1, 2 and 3 topics to practise on.