Can a local LLM answer as fast as the cloud?
Benchmarking full voice pipelines on consumer GPUs to see whether a private, local assistant can match the responsiveness of cloud services.
Notebook 2 entries
Benchmarks, experiments, and engineering notes: what we tried, what we measured, and what we learned, including when the answer was no.
Benchmarking full voice pipelines on consumer GPUs to see whether a private, local assistant can match the responsiveness of cloud services.
Testing whether wake-word loudness can pick the nearest device in a multi-room setup built from commodity speakerphones.