This invention describes an electronic device, such as a smartphone or smart speaker, designed to improve its ability to detect human speech. It works by first analyzing an initial audio signal using a convolutional neural network (CNN) to extract features. Then, it uses a recurrent neural network (RNN) and potentially another CNN to process a subsequent audio signal, comparing it to the learned features to more accurately identify speech sections and filter out background noise.
Why it matters: Since 2022, the efficiency of deploying complex AI models like CNNs and RNNs on consumer electronic devices has significantly improved. This makes the power-efficient implementation of advanced voice activity detection, as described, more practical today.
AI gives you a few directions you could take this. Pick one, and we check whether your version is different enough to patent, then write the filing.
Reinvent this with AI