LLM Tools|Index 04
Speko automates voice agent model optimization
Speko offers an API to continuously benchmark and switch between STT, LLM, and TTS models for optimal voice agent performance.
- Via
- AITECH TOKYO Editors
- Dateline
- Tokyo, August 17, 2026
- Date
- August 17, 2026
- Time
- 6 min read
Source
Hacker News TopTagline
An API to optimize voice AI models
Who & Why
For Tokyo-based product managers building enterprise voice agents, Speko automates the continuous selection and optimization of STT, LLM, and TTS models, ensuring optimal performance and cost efficiency across languages.
vs. Existing
Speko competes with the costly, manual process of benchmarking and integrating various STT, LLM, and TTS vendors, offering an automated API-driven approach that continuously optimizes performance without re-integration.
Tokyo Take
Speko offers a compelling solution for optimizing multi-language voice agents, crucial for Tokyo's global businesses. Its API-first approach is accessible, but true value for Japanese enterprises hinges on its specialized Japanese STT/TTS model benchmarking and localized pricing.
Speko is an API platform designed to continuously optimize the combination of speech-to-text (STT), large language model (LLM), and text-to-speech (TTS) components within production voice agents.
Founded by Bek, who previously built enterprise voice agents across Asia, Speko addresses the challenge of constantly evolving AI models. Companies typically evaluate models once and rarely update their stack due to the integration overhead and internal friction.
The platform functions as a router and gateway. Users send requests with optimization criteria—accuracy, latency, cost, or a balanced approach—along with language and region. Speko then selects the best-performing models from its public benchmarks.
This process replaces manual, labor-intensive benchmarking. Before Speko, teams would hire native speakers to rate new models, a ritual that Speko now automates via an API.
"we can literally go to this dashboard, switch the model, and it will do it for us."
The system offers failover during connection setup and pre-fetches session plans to minimize latency.
Speko maintains impartiality by not training or selling models itself. It transparently publishes benchmark results, even when its selections underperform alternatives. It tests spontaneous speech, long takes, and specific vocabulary like money and dates, recognizing that production demands differ from short demo clips.
An open-source gateway is also available, allowing teams to avoid extra network hops and keep API keys local. This BYOK (Bring Your Own Key) setup is free, with charges applying only to the hosted router and managed keys. Launched in late June, the platform has seen consistent weekly growth.
For a Tokyo-based professional managing a voice agent project, Speko offers a way to ensure their service uses the most current and cost-effective voice AI components without the continuous R&D burden. This could lead to better customer experience and operational efficiency, especially for multilingual deployments.
Adjacent Tools
LLM Tools
Amazon's AI Data Strategy Raises Concerns Over Rare Text Preservation
The e-commerce giant reportedly converts unique physical texts into digital training data, sparking debate on ethical data acquisition and cultural heritage.
LLM Tools
Grok's Image Generation Raises Content Safety Concerns
An alleged misuse of X's Grok AI to generate explicit imagery from a childhood photo underscores the critical challenges in AI content moderation and platform responsibility.
LLM Tools
AI's Mathematical Ceiling: Beyond Pattern Matching
A recent analysis highlights that even advanced AI models struggle with genuine mathematical reasoning, suggesting a fundamental gap in their ability to "think" like mathematicians.