Hello, I would like to present the results of my last test. I cleared a speech sound from the noise. In the image below, you can see what the spectra of these speech sounds look like. Some might say that noise removal is nothing fancy, but in my case, I used an algorithm that I intend to use to reconstruct audio speech from conferencing applications such as MS Teams, Zoom, or Webex to recover what was lost during the compression and transmission process.. And I'm pleased with myself because I did it in less than the duration of the speech, so it's very likely that I'll be able to do this operation in a real-time application. Follow me, I will post some examples for you to listen to in my next post.
ArtC
poniedziałek, 5 czerwca 2023
piątek, 21 kwietnia 2023
Signal alignment
After a long break, I came back to this project. I figured the first thing I needed to do was objectively compare the original audio with the audio recorded by the microphone and the audio from MS Teams.
So I equalized the volume of these audios, and I equalized them together in the time domain.
I found a correlation between the original audio and the microphone audio to align them together. You can see the correlation and perfectly aligned signals
I did the same for MS Teams audio
I'm particularly interested in the frequency domain, so I've made an effort to display the spectrograms of these signals. My goal was to visualize the details of the signals and do the FFT manually to fully understand how the signals were processed.
sobota, 17 grudnia 2022
Test stand
Today assembled my test stand. There are two low-quality microphones and quite nice speakers. I'm sure that some of you would claim that the speakers are garbage as well but in my opinion, it's good enough. I can listen to speeches with some comfort, but I can't say that about the output of that mics. So, it seems to be working as I desired :D
You should see its heart as well <3
Blogging
Hi, today I started this blog because I wanted to share with Indie Hackers community some of my progress. Unfortunately can't share images there :(, but here I can :D
https://www.indiehackers.com/product/artc
wtorek, 6 grudnia 2022
Meat and no potatoes
I compared referential speech with MS Teams sound. The difference is great. You can hear it, you can see it. It turns out that MS adheres to the principle of "eat the meat, leave the potatoes", They don't care about user experience, the transfer usage is the only thing that matters.
Subskrybuj:
Posty (Atom)
Real-time application is possible!
Hello, I would like to present the results of my last test. I cleared a speech sound from the noise. In the image below, you can see what th...
-
I compared referential speech with MS Teams sound. The difference is great. You can hear it, you can see it. It turns out that MS adheres t...
-
Hello, I would like to present the results of my last test. I cleared a speech sound from the noise. In the image below, you can see what th...
-
Today assembled my test stand. There are two low-quality microphones and quite nice speakers. I'm sure that some of you would claim that...
