Kalpa LabsTalk to KalpaIt Talks Back
Conversational speech models · Open beta

Towards generalist
audio models.

Kalpa Labs is an audio research lab working towards generalist audio models: models that follow instructions and learn new tasks in context. Our first conversational speech models are in beta today. Direct a scene in Studio, talk to a model live, or build on the API.

Talk to KalpaIt Talks BackOpen StudioStart DirectingRead the DocsBuild With It
Backed by Y Combinator
Live strand · grab the dot
In beta today
Studio

Direct the conversation.

Write the lines, cast the voices, set the mood. Studio performs the whole scene: multi-speaker dialogue with pacing, emotion, and voices that stay themselves.

Open Studio
Realtime

Talk to it, live.

Open a call in your browser and just talk. Ask it anything and hear it answer in its own voice. The beta model, live and unedited.

Say Hello
API

Build on it.

The same models over a clean REST API. Docs written for humans and for models, where every page has a markdown twin, plus a playground that runs in the browser.

Read the Docs
Blind human preference

State-of-the-art performance.

kalpa-beta-v0.3vsElevenLabs eleven-flash
59.3%win rate
kalpa-beta-v0.3vsElevenLabs eleven-turbo
54.0%win rate
Full evaluation in the launch report
From the launch report

Capabilities

Most voice cloning products only preserve speaker identity, and fail to mimic speaker delivery nuances. To stress test our voice cloning capabilities, we intentionally clone "meme voices" that say things in an exaggerated or distorted manner for a memetic effect.

The Bible, like an episode of Love Islandcloned from
0:00 / 0:00
John Kiriakou memecloned from
0:00 / 0:00
Read the launch report
Research direction

Conversation first. The rest of audio, next.

Speech models today are roughly where language models were a few years ago: they can speak, but they can't really listen. They don't follow instructions about sound, learn a voice or a style from a few examples in context, or notice how something was said and not just what. Generalist audio models that close that gap is our whole mission. Read more

Until then, the beta is open. Go talk to it
Careers

If you want to build the next frontier of audio models, we are hiring.