How Do They Say That? Using Text-to-Speech for Foreign Language DX Identification

By Don Moore

More of Don’s traveling DX stories can be found in his book Tales of a Vagabond DXer [SWLing Post affiliate link]. If you’ve already read his book and enjoyed it, do Don a favor and leave a review on Amazon.

When I first discovered shortwave radio way back in 1971, right from the start I was both a listener (looking for interesting content) and a DXer (trying to hear as many different places as possible). But my DXing had a problem in that I only spoke English. Sure, lots of countries had international broadcasters with programs in English. But how was I going to log the many that didn’t?

In July 1972, I ordered a sample copy of FRENDX, the bulletin of the North American Shortwave Association. It was the first time I had seen a DX bulletin, and I learned a lot. The most eye-opening thing I found out was that other North American DXers were routinely identifying and logging radio stations using languages that they didn’t themselves speak. I was genuinely surprised by this. But if they could do it, then I could, too.

I began by going after Spanish-language broadcasters. I had started studying Spanish in school, so I knew the pronunciation rules even if I didn’t yet understand much. Some station names, like Barquisimeto and Bucaramanga, were even fun to say. I soon branched out to DXing stations in Portuguese, Arabic, and French and eventually to other languages like Russian and Indonesian. One thing that helped was that in those days the World Radio TV Handbook included transcriptions of standard ID announcements from many stations, such as in the image below of Radiodiffusion Television Algerienne in the 1972 WRTH. But whatever the language, the difficulty was always the same. What my English-centered brain thought the words in a book should sound like didn’t always match how the native speakers of the language actually said it.

The anecdote, of course, was experience. Over time I developed an ear for the various languages used by lots of DX stations. Even so, long words could still trip me up. Sure, my knowledge of Spanish made Barquisimeto and Bucaramanga easy. But words like Tanjungkarang and Manokwari in Indonesian and Itatiaia and Goiânia in Portuguese still gave me headaches.

The Reasons Why

In the late 1980s I went back to school to get a master’s degree in Applied Linguistics with the goal of teaching English as a Second Language. One of the first classes I had to take was phonetics – the study of how our mouths make the sounds that turn into words. It’s more complex than you might think. The International Phonetic Alphabet has 107 symbols for distinct vowel and consonant sounds. I had to learn to audibly recognize each of them. (Just don’t ask me to do it now.) Each of those sounds can be further modified in various ways, and it’s those sounds with modifications that produce the unique phonemes that make up each language. Standards of measurement vary, but the best estimates put the total number of phonemes in all the world’s languages at around eight hundred. Most languages only use around forty. English has 44 phonemes.

So that station name spelled out on a printed list may not only sound different than you expect, but it likely also contains sounds that you don’t even know. And when the sounds get chained into words, there’s the issue of which syllable gets the stress (consider photograph versus photography). Many languages help by using accent marks when stress differs from the standard. As words get chained into sentences, the connections also affect pronunciation, such as how “got to” becomes “gotta” in colloquial English. Other languages do the same thing. And I’m not even going to go into the use of tones in languages like Mandarin, Vietnamese, and Thai.

If you think English is an easy language, you are wrong. In terms of how the pronunciation of a word matches the spelling, English is one of the worst. There are historical reasons for the inconsistencies, but about 20% of English words are pronounced significantly differently from how they are spelled. And we don’t use accents or other diacritic marks to indicate when the sound differs from what’s standard. That makes English a difficult language to learn to speak and to understand. My linguistic DX heroes are any non-native English speakers DXing American medium wave stations. Local people do not pronounce places like Louisville (Kentucky) or DuBois (Pennsylvania) the way you would expect.

So there are all types of reasons why that station identification may not sound like we think it should sound.

But There’s a Solution …

I didn’t lead you all the way here just to complain about how complicated it all is. We now have a solution.

Text-to-speech is the process of presenting a computer program with a word or phrase and having it generate the sounds. For most of us, our first experience with this probably happened two or three decades ago when calling an automated phone system that used a mechanical voice to read back unique information that couldn’t be prerecorded by a human. It was pretty bad, and text-to-speech of the era deservedly got a bad reputation.

But the good news is that, like so much technology, text-to-speech has gotten a lot better. Some of it is so good that it’s become difficult to distinguish computer-generated speech from real human speech. Of course, the ways that can be misused also make it scary, but that’s a topic for other forums. Dozens of companies provide text-to-speech software in dozens of languages. And most of them have websites where you can try out their software for free.

I do a lot of DXing of stations in languages that I don’t understand or know well. And sometimes I hear what seems to be an ID but doesn’t match what I would expect from the spelling of any listed station. Then I go to one of those text-to-speech websites and plug in the station name to create a sample audio file. I compare that file with what is being said in my DX recording to see if they match.

I’ve been using this method for several years with Brazilian medium wave stations, Russian airports, and Indonesian marine stations, among others. I’ve confirmed numerous IDs this way that otherwise I wouldn’t have figured out, or at least not have been as comfortable that I had gotten the logging right. A couple of times it’s even saved me from making mis-logs on similarly spelled station names on the same frequency.

Most text-to-speech generators offer multiple voices, including both male and female. I always pick the gender that matches that of the ID I’m comparing to. For the most widely spoken languages, some companies offer multiple dialects, like Nigerian English or Argentine Spanish. So if you’re trying to ID a Colombian station, pick Colombian Spanish, not Argentine.

And, very important, be sure to use the correct spelling when pasting in the station name to check. That means including all the accent marks, tildes, umlauts, cedillas, etc. They are there to indicate how the word is pronounced, and if you omit them you will not get the right result. To make sure I have the correct spelling with all the proper diacritic marks, I usually copy/paste from Google Maps, a Wikipedia article, or some other webpage.

Finally, be aware that some local pronunciations are so unusual that the speech engines still may not get it right. I grew up near DuBois, Pennsylvania, and not one of the text-to-speech sites I tested said “DuBois, Pennsylvania” the way the people who live there do. They call their town “DO-boys.” To any native French speakers reading this, I sincerely apologize for what we Pennsylvanians have done to your language.

Links

A quick way to find a TTS engine is to do a web search for “<language name> text to speech”. Here are a few of the better ones.

Carlos’ Illustrated Radio Listening Report and Recording of NHK (August 29, 2026)

Many thanks to SWLing Post contributor and noted political cartoonist Carlos Latuff, who shares the following illustrated radio listening report.


Carlos notes:

“…Due to a weather front and humid air, experiencing record-breaking heavy rain. A Level 5 Special Heavy Rain Warning has been issued for Fukui City, Ono City and Katsuyama City…”

Click here to view on YouTube.

Something new from Sangean? Is an ATS-909X2+/ATS-909X2SE on the Horizon?

Many thanks to SWLing Post contributor Bob, who spotted something interesting in Sangean’s documentation database.

Sangean has posted what appears to be a product brochure for an ATS-909XSE, dated August 27, 2026 (PDF). There’s also a corresponding manual (PDF) in the company’s documentation database.

Interestingly, another Sangean document (PDF) appears to indicate that the ATS-909X2+ may be the European designation, while the U.S. version is designated the ATS-909X2SE.

Both models appear to have the same feature set:

At this point, however, I haven’t found a formal announcement from Sangean about the ATS-909X2+ or ATS-909X2SE, so I’m not going to speculate about the radio or its availability.

It’s certainly an intriguing find, though, and I’ll be watching for an official announcement. I’m super pleased that Sangean continues to innovate!

Thanks again to Bob for the tip!

Carlos’ Illustrated Radio Listening Report and Recording of the BBC and CGTN (August 27-28, 2026)

Many thanks to SWLing Post contributor and noted political cartoonist Carlos Latuff, who shares the following illustrated radio listening report.


Carlos notes:

Rescue operations are ongoing in Nepal and Tibet after a devastating flash flood, via BBC and CGTN

Click here to view on YouTube.

A 1979 Paul Harvey Broadcast May Hold the Key to a Family Mystery

Many thanks to SWLing Post reader Jessica, who writes:

I’m trying to locate an off-air recording of Paul Harvey from approximately November 7–9, 1979, possibly heard via AFRTS shortwave. The story concerned William “Bill” C. Haynes Sr., age 34, of Dallas, Texas, who died on November 5, 1979. An Associated Press story published November 7 carried the headline “‘Nobody loves me,’ man shouts as car hits him” and described Haynes as “a very lonely man.” His daughter distinctly remembers Paul Harvey discussing her father.

We believe the broadcast was most likely Paul Harvey News and Comment, though the syndicated television feature Paul Harvey Comments is also being investigated. We are looking for any AFRTS recording from November 7, 8, or 9, even if the cassette/reel is labeled only with AFRTS, a frequency, UTC time, a general news description, or a full block of programming rather than Paul Harvey’s name.

A February 1979 NASWA/FRENDX report, describing AFRTS schedule changes effective January 1, says “Paul Harvey 1711 & 2111” GMT. An August 1979 FRENDX guide lists “Paul Harvey, News and Commentary” at about 2100 GMT on 17765, 15430, 15330, and 11790 kHz. A February 1980 program guide then shows the later feed more precisely as 2111–2122 GMT on 17765, 15430, and 15330.

I have not yet found a surviving November 1979 schedule that proves the exact minute/frequency configuration for November 7–9, so I would phrase this as a strong target rather than certainty. Anyone checking old AFRTS tape boxes should especially look around 1711 GMT and 2100–2122 GMT on those frequencies.

If anyone taped AFRTS around those dates, kept DX cassettes or reels, has old AFRTS transcription discs, or knows a collector who did, I would be grateful for any lead. A recording of even one of the relevant Harvey feeds could help solve a nearly 47-year-old family mystery.

Readers, I know this is a very long shot, but if you’re able to assist Jessica, please leave a comment in this post with any information you have. I can then help Jessica get in touch with you. 

Waterfall images from pirate radio

Many thanks to SWLing Post contributor Dan Greenall, who writes:

Hi Thomas

The concept of transmitting images through SDR waterfalls is relatively new to me. Last weekend, while listening to pirate Radio 48 (via Mix Radio International) remotely on SDRs in Rochester, NY, and Cape Cod, MA, I snapped these photos of their 10 kHz-wide USB signal on 6935, along with some of the pictures and text that scrolled down beside it.

Less than 24 hours later, I came across another stream of graphics, this time on 6960 kHz. It displayed a web page address for Radio Free Spirit. These were followed by a brief stream of audio on USB of a man and woman having a conversation, which ended abruptly without any announcement. I searched out an email address for Radio Free Spirit and sent off an inquiry. Within 15 minutes, a response was received confirming it had been them, simply running a test and had not expected anyone to hear them or reply.

Photos taken of both stations are attached here along with some related audio clips.

Radio 48:

Radio Free Spirit:

I am wondering if others have had experiences with this type of reception and if any international broadcasters have ever dabbled in it.

73

Dan Greenall, Ontario, Canada

Radio Waves: Voice of Nigeria, Capturing ISS Images, SRG SSR FM Radio Expansion, and Conrad Remembers September 11

Radio Waves:  Stories Making Waves in the World of Radio

Welcome to the SWLing Post’s Radio Waves, a collection of links to interesting stories making waves in the world of radio. Enjoy!

Many thanks to SWLing Post contributors David Iurescia and Richard Cuff for the following tips:


Inside VON, Where Technology, Diplomacy and Storytelling Converge By Muhammad Goni (PR Nigeria)

There is something quietly powerful about walking through the soundproof doors of a broadcasting facility whose signals are designed to travel far beyond Nigeria’s borders. Behind the microphones, studios, transmitters and antennas of the Voice of Nigeria (VON) lies a responsibility far larger than daily news bulletins. VON is one of the institutions through which Nigeria speaks to the world—an instrument of national identity, diplomacy, culture and global relevance.

My tour of VON’s facilities in Abuja revealed this clearly. It was more than a visit; it was an education in how international broadcasting works, why it matters and what it means for a country seeking to shape how it is understood. [Continue reading…]

Looking up: French radio enthusiast captures ISS signals from his garage (Radio France international)

Astronaut Sophie Adenot made history this month as the first French woman to perform a spacewalk. Back on Earth, space enthusiasts across France are finding their own ways to reach for the stars. In the second episode of a five-part series, RFI hears from Vincent Plousey, a radio ham from Tavaux in eastern France whose hobby is picking up signals from the ISS using homemade antennas built in his garage. [Continue reading…]

Switzerland’s SRG SSR to Resurrect Analog FM Broadcasts (Radio World)

Switzerland’s public radio network, SRG SSR, abandoned FM broadcasting last year in favor of the DAB+ digital mode.

But after the Swiss Parliament voted to postpone the nationwide switch-off of FM transmitters, and amidst reported listener declines, SRG SSR is bringing back analog FM transmissions, reportedly targeting the middle part of this fall for resumption.

And, according to one report, the public broadcaster desires further FM expansion. [Continue reading…]

Remembering 9/11: Conrad Trautmann (Radio World)

The 25th anniversary of the Sept. 11 terror attacks is approaching. Radio World is presenting a series in which people share their memories.

Conrad Trautmann was the senior VP of engineering at Westwood One and Metro Traffic. His office and the main satellite uplink site were at the CBS Broadcast Center at 524 West 57th Street in New York City. [Click here to read the full article…]


Do you enjoy the SWLing Post?

Please consider supporting us via Patreon or our Coffee Fund!

Your support makes articles like this one possible. Thank you!