August 17, 2026

LLM Tools|Index 04

Speko automates voice agent model optimization

Speko offers an API to continuously benchmark and switch between STT, LLM, and TTS models for optimal voice agent performance.

Via
AITECH TOKYO Editors
Dateline
Tokyo, August 17, 2026
Date
August 17, 2026
Time
6 min read
Speko automates voice agent model optimization

Tagline

An API to optimize voice AI models

Who & Why

For Tokyo-based product managers building enterprise voice agents, Speko automates the continuous selection and optimization of STT, LLM, and TTS models, ensuring optimal performance and cost efficiency across languages.

vs. Existing

Speko competes with the costly, manual process of benchmarking and integrating various STT, LLM, and TTS vendors, offering an automated API-driven approach that continuously optimizes performance without re-integration.

Tokyo Take

Speko offers a compelling solution for optimizing multi-language voice agents, crucial for Tokyo's global businesses. Its API-first approach is accessible, but true value for Japanese enterprises hinges on its specialized Japanese STT/TTS model benchmarking and localized pricing.

Speko is an API platform designed to continuously optimize the combination of speech-to-text (STT), large language model (LLM), and text-to-speech (TTS) components within production voice agents.

Founded by Bek, who previously built enterprise voice agents across Asia, Speko addresses the challenge of constantly evolving AI models. Companies typically evaluate models once and rarely update their stack due to the integration overhead and internal friction.

The platform functions as a router and gateway. Users send requests with optimization criteria—accuracy, latency, cost, or a balanced approach—along with language and region. Speko then selects the best-performing models from its public benchmarks.

This process replaces manual, labor-intensive benchmarking. Before Speko, teams would hire native speakers to rate new models, a ritual that Speko now automates via an API.

"we can literally go to this dashboard, switch the model, and it will do it for us."

The system offers failover during connection setup and pre-fetches session plans to minimize latency.

Speko maintains impartiality by not training or selling models itself. It transparently publishes benchmark results, even when its selections underperform alternatives. It tests spontaneous speech, long takes, and specific vocabulary like money and dates, recognizing that production demands differ from short demo clips.

An open-source gateway is also available, allowing teams to avoid extra network hops and keep API keys local. This BYOK (Bring Your Own Key) setup is free, with charges applying only to the hosted router and managed keys. Launched in late June, the platform has seen consistent weekly growth.

For a Tokyo-based professional managing a voice agent project, Speko offers a way to ensure their service uses the most current and cost-effective voice AI components without the continuous R&D burden. This could lead to better customer experience and operational efficiency, especially for multilingual deployments.

The Briefing

World AI tech, read from Tokyo. Once a week, in Japanese.

Each Friday: the five global AI tech stories Japanese business professionals should know about this week, translated and read through a Tokyo lens — what it means for Japan, what to act on, what to keep watching.

We respect your inbox. Unsubscribe anytime.