Mooneer’s FreeDV Update – February 2025

This month was focused primarily on preparation for and attending Orlando HamCation (more information about our experience at that event here). However, there were some improvements made to the FreeDV application in the meantime:

  • The official 80/160 meter calling frequencies were tweaked based on user input to better fit with the Japanese band plan.
  • Fixed an issue preventing suppression of the User Message tooltip for shorter messages in the FreeDV Reporter window.
  • Fixed an issue causing the User Message column to randomly change sizes on user connection and disconnection.
  • Tweaked the SNR display in the main window to only show whole numbered SNRs (i.e. no decimals). This was done to make it more obvious that slow/fast SNR calculations were occurring.
  • Updated Hamlib in the macOS and Windows builds to 4.6.2. As part of that, loading the Hamlib supported radios was also changed to prevent issues with the names changing during runtime.
  • The “short” timeout for RX display in FreeDV Reporter was lengthened to five seconds to reduce flickering in the FreeDV Reporter window.
  • The ezDV implementation of FreeDV Reporter support was adapted for freedv-gui to eliminate a dependency on the sioclient-cpp library. This also has the effect of fixing issues with users being unable to reconnect if the FreeDV Reporter server goes down (or if there are network issues causing the connection to drop).

More information can be found in the commit history below:

(Note that all commit logs above were generated with the following command line:)

git log --author="member@email" --after "Month 1, 2025" --before "Month 31, 2025" --all > commit.log

David Feb 2025, Papers, ML EQ prototyping

RADE Documentation

One of the aims of this project is to document our work to a professional standard. This month I spent some time working on two research papers, one on the HF RADE work, and one on the baseband FM (BBFM) work. I’m also working on a presentation of RADE technology aimed at Hams, that I will present at my local AREG club in March.

The RADE work has been moving pretty fast over the past 12 months, so I’ve found writing up the work beneficial to help me collect my thoughts and prepare for further development. It’s very new technology, and a lot of people are curious about how RADE technology works. This all means it takes some time and effort to explain. Another good reason to document this work is to get it out of my head and into a form that others can work with in future (another one of our grant aims).

We hope to publish the papers later this year. By writing the papers we also hope to promote the project and help communicate our work at a professional level to commercial companies who may be interested in integrating RADE technology into their products.

ML Equalisation

RADE is a mixture of classical DSP and ML signal processing. One interesting design choice is how to partition the design – which chunks of signal processing should use old school DSP, and which ML?

For RADE V1 we use ML to generate QAM symbols, but classical DSP to “equalise” these symbols at the receiver. You can think of equalisation as removing any phase and frequency offsets in the received signal – a bit like fine tuning a SSB receiver. In regular data modems equalisation stops the received modem constellation from rotating or spinning, so the correct bits can be demodulated.

For RADE V2 I am prototyping the use of ML for the equalisation. The curves below show the performance of various schemes I have tested so far, with RADE V1 (blue) as the control. The “loss” is a measure of distortion, lower is better. You can see the loss decrease as SNR increases, just as we would expect.

Now in practice RADE V1 actual requires about 3dB more SNR once we add the classical DSP equalisation so for a fair comparison it should be shifted 3dB to the right, making all of the curves within a few dB of each other. So the ML network is indeed performing the equalisation function, but too early to say if we have something that can outperform the classical DSP approach used in RADE V1.

The yellow curve is intriguing – its suggests that with the right network we can get better speech quality than RADE V1 at high SNRs.

More work required to work through the equalisation question and it’s very much R&D rather than Engineering, which makes the timeline for RADE V2 hard to predict. More next month…..

FreeDV Experience at 2025 Orlando HamCation

The FreeDV project recently came back from Orlando, FL, where it had a booth in the North Hall at Orlando HamCation. At the booth, we had a demo of the RADE work we’ve been working on, including two headsets (one playing “analog” audio as what would be transmitted through a microphone and the other playing receive audio after being transmitted and received using RADE). We additionally had internet access in the booth so we were able to show a live view of FreeDV Reporter:

RADE and FreeDV Reporter display at the FreeDV booth.

Over the course of the weekend, we had significant interest in the FreeDV booth. On top of that, nearly everyone who listened to the RADE demo were impressed by the audio quality, especially when compared to the existing state of the art in digitial voice.

Additionally, Mooneer K6AQ gave a talk about RADE on Friday afternoon of the hamfest. For those who missed the talk or want to refer back to the slides, they’re available in PDF format below:

FreeDV at Orlando HamCation – February 7-9, 2025

FreeDV will have a booth at the North Hall of Orlando HamCation this weekend (near the HamCation prize booth). If you happen to be in the area, come check us out! More information about HamCation can be found at https://www.hamcation.com/.

Additionally, Mooneer K6AQ will be giving a talk about/demoing FreeDV’s new RADE mode tomorrow (February 7) at 3:30pm in CS II (the set of popup tents in the back of the fairgrounds). Hope to see you there too!

Mooneer’s FreeDV Update – January 2025

This month was spent firming up the preview release that just came out. This involved fixing various bugs discovered by the RADE test team, such as various crashes that began occurring during the implementation of SNR and RADE text support (required for FreeDV Reporter to fully work with RADE). Audio quality was also improved and the transmission of the End Of Over (EOO) block was also made more reliable, ensuring that users reported received signals to FreeDV Reporter and PSK Reporter more often.

Speaking of PSK Reporter, this is the worldwide map of activity over the last 24 hours as a result of this work, indicating heavy interest in RADE and FreeDV more generally:

For February, the focus is going to be on promoting the work we’ve done as a project. We’ll be at Orlando HamCation in just a few days (North Hall, booth 119, as well as a talk given by me on Friday February 7th at 3:30pm)–hope to see you all there!

More information can be found in the commit history below:

(Note that all commit logs above were generated with the following command line:)

git log --author="member@email" --after "Month 1, 2025" --before "Month 31, 2025" --all > commit.log

David Jan 2025 – SNR estimation, Bandwidth, EQ, 2024 in review

At the start of this month I did battle with the problem of SNR estimation on the RADE V1 signal. As I have mentioned previously, this had some challenges due to the lack of structure in the RADE constellation. After a few false starts I managed to get something viable running using the properties of the pilot symbols. The plot below shows the estimated against actual SNR for a range of channels. In the -5 to 10dB range (of most interest to us) it’s within 1dB for all but the MPP (fast fading) channel where the reported estimate reads a few dB lower than the actual (Note Es/No roughly the same as SNR for this example).

I’ve started work on RADE V2, where we hope to use lessons learned from RADE V1 to make some improvements and develop a “stable” waveform for general Ham use. This month I have made some progress in jointly optimising the PAPR and bandwidth of the RADE signals. For regulatory purposes, the bandwidth of signals like OFDM are often specified in terms of the “occupied bandwidth” (OBW) that contains 99% of the power. The figure below shows the spectrum of a 1000 symbols/s signal with a 1235 Hz 99% occupied bandwidth OBW in red.


Machine Learning Equalisation

Also for RADE V2, I have been prototyping ML based equalisation, and have obtained good results for some examples using the BER of QPSK symbols as a metric. The plot below shows the BER against Eb/No for the classical DSP (blue), and two candidate ML equalisers (red and green, distinguished by different loss functions). The channel had random phase offsets for every frame, which the equaliser had to correct. The three equalisers have more or less the same performance.

These results show the equalisation function can be performed ML networks, with equivalent performance to classical DSP.

Project Management

Quite a bit of admin this month, including time spent recruiting prospective new PLT members, updating budgets, and our annual report. Not as much fun as playing with machine learning, but necessary to keep the project running smoothly.

It was time to write our annual report for the ARDC who have kindly funded this project for the last two years. Writing this report underlined what a good year we had in 2024, some highlights:

  • The development and Beta release of the Radio Autoencoder RADE V1 which is well on the way to meeting our goals of being competitive with SSB at high and low SNRs. Special thanks to Jean-Marc Valin for your mentoring and vision on this project!
  • The BBFM project, paving the way for high quality speech on VHF/UHF land module radio (LMR) applications, in collaboration with Tibor Bece and George Karan.
  • New data modes to support FreeDATA, in collaboration with Simon DJ2LS.
  • The release of ezDV and continued maintenance of freedv-gui largely by Mooneer’s efforts.
  • Peter Marks joining our Project Leadership Team. He’s already making a big impact – thanks Peter!

FreeDV 2.0.0-20250130 released

This is the second preview release of FreeDV containing the new RADE mode. For more information about RADE’s development, check out the blog posts on the FreeDV website:

https://freedv.org/davids-freedv-update-feb-2024/
https://freedv.org/davids-freedv-update-march-2024/
https://freedv.org/davids-freedv-update-april-2024/
https://freedv.org/davids-freedv-update-may-2024/
https://freedv.org/davids-freedv-update-june-2024/
https://freedv.org/davids-freedv-update-july-2024/
https://freedv.org/davids-freedv-update-august-2024/
https://freedv.org/mooneers-freedv-update-august-2024/
https://freedv.org/mooneers-freedv-update-september-2024/
https://freedv.org/davids-freedv-update-september-2024/

Changes versus the first preview release:

* Signal to noise ratio (SNR) is now displayed while receiving RADE signals.
* Received signals are now reported to FreeDV Reporter (without callsigns) once per second. Once a callsign is received (at the end of the transmission), the callsign is reported to both FreeDV Reporter and PSK Reporter.
* Fixed bug preventing sync indicator from turning green with RADE.
* Visual Studio Redistributable is now installed if your PC does not already have it. (This is required for the Python packages FreeDV uses.)
* Fixed bug preventing Request QSY button from being enabled in RADE mode.
* RADE has been renamed to RADEV1 in the UI and FreeDV Reporter.
* macOS binaries are now signed and notarized, avoiding the need for the workaround in the previous build.
* Fixed issue causing FreeDV to segfault on exit when RADE is running.
* Python files are now precompiled to improve startup time.
* Core RADE code is now in C (versus Python).
* Uninstaller now fully cleans up after Python.
* Audio chain is cleaned up to improve audio quality.
* README has been updated to clarify Linux instructions and to provide a link to a script to auto-build with RADE support. (Thanks @barjac!)
* Maximum SNR displayed in the main window is now 40 dB to reflect real-world testing.
* “devel” in the version string is shortened to “dev” and incremented to “dev2” to reflect the second preview build.

Limitations:

* Multiple RX mode is not supported. If you choose RADE and push Start, that’s the only mode you can work; you’ll need to stop, choose another mode and start again to work FreeDV with the existing modes.
* Squelch cannot currently be disabled with RADE. It’s unknown at this time whether disabling squelch is possible.
* Due to compilation problems, 2020/2020B modes are disabled.
* There is currently no Windows ARM build; this will hopefully be included in a future preview build. You may be able to use the 64-bit Intel/AMD Windows build in the meantime.
* Minimum hardware requirements haven’t been fully outlined, so your system currently may not be able to use RADE. Future planned optimizations may improve this.

Other notes:

* The below builds are significantly bigger than previous releases. This is due to needing to include Python and the modules that RADE requires. Planned porting to C/C++ will eventually negate the need for Python.
* The Windows build includes Python but not the modules that RADE requires. As part of the install process, the version of Python built into FreeDV will go out to the internet to download the needed modules.
* As development is expected to happen quickly, these preview builds have a six month expiry date (currently July 30, 2025).
* 32-bit Windows is no longer supported due to its likely inability to work with RADE.

More information and download links can be found at https://github.com/drowe67/freedv-gui/releases/tag/v2.0.0-20250130.

David Dec 2024 – Testing RADE with Automatic Speech Recognition

An important goal of our project is improved speech quality over SSB and both low and high SNRs. We have anecdotal reports of good performance of RADE compared to SSB, but need an objective, controlled way of comparing performance. For speech systems this generally means ITU-T P.800 or P.808 standards based subjective testing. However this is complex and requires skills, experience and resources not available to our team.

A few months ago Simon, DJ2LS suggested the use of Automatic Speech Recognition (ASR). More recently, when discussing the issue of subjective testing, Jean Marc Valin also suggested ASR and provided suggestions for a practical test system. So I spent much of December building up a framework for ASR tests.

The general idea is to take a dataset of speech samples, pass them through simulations of SSB and RADE over HF radio channels, then use a ASR engine to detect the words in the received speech. A post processing system then compares the detected words to the original words and determines the Word Error rate (WER) as a performance metric. Our work uses the Librispeech dataset, and the Whisper ASR system.

These sentences are complex English sentences, spoken quickly with no contextual cues. I have trouble understanding many of them on the first listen. This is a much tougher test than the typical low SNR Amateur Radio contact where someone shouts their callsign 5 times then reports “5 by 9”. For example, here is one sample from the Librispeech dataset processed with SSB/RADE/original (listen to the original last); SSB and RADE were at about 6dB SNR on a MPP (fading) channel.

The plot below show some initial results over 500 sentences. The x-axis is receiver SNR measured in a 3kHz noise bandwidth. The y-axis is the word error rate WER). Green is RADE, and blue SSB. The solid lines are for a AWGN channel, the dashed lines the multi-path poor (MPP) fading channel. The dots (placed arbitrarily on the x-axis) in the lower right are controls, e.g. the FARGAN synthesizer used by RADE with no encoding, 4kHz band limited speech, and the original, clean speech.


A low word error rate (WER), say 5%, would correspond to an effortless “armchair copy”; a 30% WER could be the limits of practical voice communication (1 in 3 words can’t be understood). The distance between the RADE and SSB curves shows the benefits of RADE, at least using this test.

For example, if you draw a line across the 10% WER level, RADE achieves this (dashed MPP curves) at 3dB, SSB at 12dB. The x-axis doesn’t include the PAPR advantage of RADE, which is roughly an additional 5dB when using a transmitter with the same peak power output (depending on how hard the SSB is compressed).

Also this month I have been working on SNR measurement of received RADE signals. This is quite challenging, due to the lack of structure in the ML-generated RADE constellation. At present I’m attempting to use a classical DSP approach using the pilots symbols. This will be the last feature we will add to RADE V1, as we’d like to use the lessons learned to start designing RADE V2.

Mooneer’s FreeDV Update – December 2024

This month involved more improvements to the FreeDV GUI application. One improvement involved the unit test framework; it’s now possible to capture the features decoded by RADE (prior to being fed into the FARGAN codec). This is useful for quantifying changes in the receive pipeline and ensuring that what’s encoded by RADE is also mostly returned by the decoder on a clean channel.

The biggest improvement, however, is the implementation of the same LDPC based callsign encoding and decoding system that’s used in the legacy FreeDV modes. This data is placed in what’s known as the End Of Over (EOO) block at the end of the RADE transmission and allows the application to report received callsigns to FreeDV Reporter and PSK Reporter, albeit only at the end of the transmission. FreeDV Reporter specific logic was added to mitigate this by reporting that a RADE signal is being received once a second while still in sync (just with no callsign), hopefully still allowing people to see that someone’s possibly decoding them in real time.

Since we’re touching the FreeDV Reporter logic, it was also a good opportunity to make some significant changes to the FreeDV Reporter service and website. First, the “left the chat”/”entered the chat” messages were removed by PLT request in order to make it easier to see actual chat messages. Next, the separate popup window for viewing who’s in the chat was removed in favor of an always-visible bar at the bottom of the chat tab containing the callsigns of the users that are logged into chat. The message backlog was also extended to 30 days (from 7 days) and preserved into a database so that the chat messages aren’t lost in the event that the FreeDV Reporter server needs to be restarted.

Besides the above, there were some other minor fixes with the Windows installer/uninstaller along with logic added to detect whether microphone permissions have been granted. RADE is also now called RADEV1 in the FreeDV application to differentiate it versus a future version 2 of RADE. Some infrastructure was also added to be able to sign macOS builds (required to avoid errors involving “damaged” applications in newer versions of macOS).

In any case, we’re now going to focus on additional testing prior to releasing a new preview build of FreeDV for general usage. Hopefully we’ll have additional updates on that soon.

More information can be found in the commit history below:

(Note that all commit logs above were generated with the following command line:)

git log --author="member@email" --after "Month 1, 2024" --before "Month 31, 2024" --all > commit.log

DC Coupled Baseband FM Testing

Tibor Bece and George Karan are collaborating with me on the baseband FM (BBFM) project. Tibor and George are veterans of the land mobile radio (LMR) industry, having worked together for many years and helped develop commercial VHF and UHF radio hardware with over 2 million units manufactured. They are pretty excited about the Radio Autoencoder work and what it could mean for LMR.

George has managed to build the RADE V1 stack, and run the ctests on a variety of embedded platforms, including AM625 – this is a high end embedded processor with enough power to run RADE (including the FARGAN stack); and a Librem 5 phone!

Tibor has been interfacing the BBFM ML stack to a COTs LMR radio, using a modified conventional digital voice frame structure to carry the “analog” BBFM symbols. Unlike my passband demo, this implementation has direct access to the FM modulator and discriminator so it’s a “DC coupled” arrangement – closer to what a real world, commercial implementation would look like.

Like me, Tibor was initially thinking the speech quality and low SNR performance of this technology was in the “too good to be true” category. However he has now performed controlled experiments on his (very well equipped) RF work bench, as was quite surprised to be getting high quality speech at RX signals levels down to -125dBm, several dB lower than analog FM or digital LMR systems like P25 would allow. At this low RF level the cut off is due to framing of the RADE symbols (not BBFM), as he never dreamed it would be necessary to operate at such a low SNR.

Tibor writes:

The 11dB SINAD point (around -121dBm) is where the squelch would normally fail to open, and a P25 frame would start dropping out. The RADE decoder munches through this with great ease, there is some barely perceptible degradation.

All I can say – WOW!

Here are samples (over the same radios) of analog FM and BBFM at various RF input levels from Tibor’s workbench:

FM at -124dBm
BBFM at -124dBm
FM at -121dBm
BBFM at -121dBm
FM at -117dBm
BBFM at -117dBm
FM at -110dBm
BBFM at -110dBm