Hi HN! I'm Bek, founder of Speko, a platform that finds an optimal combination of speech-to-text, LLM, and text-to-speech models, given your constraints, among all our public benchmarked options, and tells you why.Demo: https://youtu.be/no2LY2gRh-c Link: https://speko.ai/?utm_source=hnTypical production voice agent is an ensemble of three models: STT, an LLM, and TTS.Each of those layers offers a dozen credible vendors, and each month there are new models on the market. Almost everyone evaluates once, picks a stack of their choice, and never rechecks because switching from a vendor to another involves yet another integration and arguments about the numbers.The result is that you use voice agents running last quarter's models while better and cheaper options are available.Before founding Speko, I spent four years as cofounder and CTO building voice agents for enterprises across Asia in 10+ languages. Each time a new speech model would arrive, we repeated the same ritual: hire native-spe...
Want to discover more AI signals like this?
Explore Steek