A script, a brief, or a sound you can't find
Three apps cover speech, music and effects. Each returns a file that lands in your library.
Speech, music and effects — and nothing invented in between
Every run shows its credit figure on the button before you commit, and a failed run is refunded.
Ninety-six voices, and punctuation as direction
These models read punctuation as direction. A comma is a breath and a full stop is a beat — write the pauses in rather than asking for them. Ninety-six named voices across four families, up to 5000 characters a run.
- 96 voices
- Ten models
- 5000 characters
A voice built from a recording
Hand Text to Speech five to thirty seconds of clean speech and it reads your script in that voice. It is a mode inside the app rather than a separate surface, so the same speed and emotion controls apply.
- Voice cloning
- Five to thirty seconds
- A mode, not an app
Scoring, not searching
Two models, and the choice between them is the choice between songs and beds: one takes a description and the lyrics separately, the other writes instrumental textures from a plain description. Both are flat priced, so the length of the track does not change what a run costs.
- Two models
- Flat priced
- Lyrics or instrumental
The sound of the thing
Effects are short. Under two seconds for an impact, four to six for an atmosphere — asking for twenty gets you silence with a sound in it. A 'Follow the description' control decides how literally the words are read: lower lets the model invent, which is what an atmosphere wants and an impact does not.
- Sound FX
- Chosen durations
- Seamless loop
Who it's for
Whatever you're building, it fits.
Different teams, different needs, one workspace.
Made alongside the audio apps
Stills from work the audio apps were scored for. Every output lands in the Library and reopens from History.






Three apps, not a studio pretending to be twelve
There is no audio studio page in the product — audio is three apps in the dashboard nav, and this is all of them.
-
Text to Speech
Ten models, 96 voices, and cloning as a mode.
-
Music
Songs with lyrics, or instrumental beds.
-
Sound FX
One effect, at a length you choose from a handful.
Built for Growth at Every Stage
Every plan includes every model and every feature. Plans only change how many credits you get and how many generations run at once.
- Credits refresh monthly
- Top-up additional credits anytime
- Unused credits don't roll over
Not ready for the commitment?
FAQs
Questions, answered.
The questions buyers ask before their first run. Each app's own page carries the specifics for that app.
What kinds of audio can I generate?
Speech, music and sound effects. Text to Speech carries ten models and ninety-six named voices, and clones a voice from a recording in the same request. Speech is priced per thousand characters — 91 to 301 credits depending on the model — while both music models are flat-priced whatever the length of the track. The figure is on the button before you commit.
Can I use my own images and footage as a starting point?
Yes. Uploading your own assets is included in every plan — you can start a generation from your own image or clip, edit footage you already have, and pin character references and brand kits in your Library so later generations stay consistent with them.
What can I actually do with what I generate, including commercial use?
You own the outputs you generate, to the extent they can be owned, and you may use them commercially. We claim no licence over them beyond what is needed to store them in your library and deliver them to you. Note that the legal status of machine-generated output is unsettled in some jurisdictions — we can grant everything we hold, but not rights that nobody holds.
What happens to credits I don't use?
Unused subscription credits don't roll over for as long as your subscription stays active — they expire monthly. Pay-as-you-go credit packages never expire and stay in your workspace balance until you use them. If a generation fails because of a fault on our side, those credits are returned automatically.