natural text to speech for audiobooks

make long passages easier to hear and return to with narration that keeps the chapter moving

narration built for chapters

narration built for chapters

narration built for chapters

shape a narrator voice that can hold a chapter with 12 voice presets and pacing control

type ‹ to insert emotion tags

built for real production work

built for real production work

built for real production work

self published books
turn your finished manuscript into an audiobook with silk's text to speech.
fiction
novels, short stories and serials voiced chapter by chapter with an ai narrator.
nonfiction
business books, memoirs and how-to guides turned into audio through the silk api.
kids stories
bedtime stories, picture books and read along tales for young listeners.

silk keeps the voice intact even through language changes

silk keeps the voice intact even through language changes

silk keeps the voice intact even through language changes

narrator
description: a male 40s british voice, low pitch, gravelly timbre, slow pacing, neutral, formal register, like a dramatic narrator. text: the door creaked open. nobody was there. and yet, something watched.
0:00 / 0:00
podcast host
description: a female 30s hindi voice, normal pitch, smooth timbre, conversational pacing, energetic, casual register, like a podcast host. text: आज का episode थोड़ा अलग है। एक minute के लिए सीधा बैठ जाओ।
0:00 / 0:00
support
description: a female 30s indian voice, normal pitch, warm timbre, conversational pacing, neutral, neutral register, like a customer support agent. text: मैं आपकी help के लिए यहाँ हूँ। एक minute, मैं check करती हूँ।
0:00 / 0:00
streamer
description: a male 20s american voice, high pitch, smooth timbre, very fast pacing, excited, casual register, like a streamer reacting live. text: oh my god, did you see that play? that was insane!
0:00 / 0:00

generate production-ready audiobook narration in three steps

generate production-ready audiobook narration in three steps

generate production-ready audiobook narration in three steps

call silk from your product workflow, or open the playground to hear the voice before shipping

silk api

silk api

silk api

  1. Create your key

  1. Create your key

  1. Create your key

grab a key from api keys in the playground. name it whatever you like.

  1. choose a voice model

  1. choose a voice model

  1. choose a model

mulberry-1.6 covers all 22 languages. muga adds tone tags for hindi-english lines.

  1. call the api

  1. call the api

send the model and text, get a wav back. add audio_format for other formats.

from rumikai import Rumik

client = Rumik()  # reads RUMIK_API_KEY
audio = client.speech.create(
    text="Hello, what can I do for you?",
    model="mulberry-1.6",
    description="professional, Indian English accent, steady pace",
)
audio.save("speech.wav")
import requests

url = "[https://silk-api.rumik.ai/v1/tts](https://silk-api.rumik.ai/v1/tts)"

payload = {
    "model": "mulberry-1.6",
    "text": "Hello, what can I do for you?", 
    "description": "professional, British accent, steady pace", 
    "speaker": "aisha",
}
headers = {"Authorization": "Bearer <token>"}

response = requests.post(url, json=payload, headers=headers)

with open("speech.wav", "wb") as f:
    f.write(response.content)
import requests

url = "[https://silk-api.rumik.ai/v1/tts](https://silk-api.rumik.ai/v1/tts)"

payload = {
    "model": "mulberry-1.6",
    "text": "Hello, what can I do for you?", 
    "description": "professional, British accent, steady pace", 
    "speaker": "aisha",
}
headers = {"Authorization": "Bearer <token>"}

response = requests.post(url, json=payload, headers=headers)

with open("speech.wav", "wb") as f:
    f.write(response.content)

silk playground

silk playground

silk playground

  1. paste your text

    enter the english sentence you want to hear. keep punctuation, numbers, and mixed-language phrases exactly as you want them spoken.

  2. choose a model

    select the silk model for english, then adjust the available settings for the voice and delivery you need.

  3. generate the audio

    listen to the result, refine the text or settings when needed, and download the audio when the line is ready.

audiobooks

audiobooks across languages and accents

audiobooks across languages and accents

audiobooks

build speech for 22 indian languages, english accents, and regional delivery from the same silk stack

audiobooks across languages and accents

build speech for 22 indian languages, english accents, and regional delivery from the same silk stack

where ai audiobook narration falls short

where ai audiobook narration falls short

where ai audiobook narration falls short

a chapter needs more than a clear voice. it needs a narrator that can hold the page

problemwhat goes wronghow silk helps
sounds monotonethe voice sounds fine for a paragraph, then gets tiring across a full chapter.set the narrator style before generation, with enough pace control to keep long passages moving.
characters blurdialogue and narration start to sound too close to each other.delivery cues give character lines a different feel from the surrounding narration.
poor pacingquiet moments and important turns get the same speed as everything else.punctuation and paragraph breaks act like narrator notes, giving the story room where it needs it.

meet the teams already speaking through silk

curvet put mulberry and muga directly inside its ai workflow canvas. within two days, teams generated voices across education, design, crm and enterprise workflows.

100 + hours in 2 days

curvet ai

snaptv uses silk to give a voice to bite-sized lessons made for how india learns, quickly, on mobile, in simple hindi and easy english.

snap tv

Image (9) (no background)

jee concepts are difficult enough. monk learning uses silk to turn dense explanations into clear, natural voice for aspirants preparing every day.

monk learning

meet the teams already speaking through silk

meet the teams already speaking through silk

curvet put mulberry and muga directly inside its ai workflow canvas. within two days, teams generated voices across education, design, crm and enterprise workflows.

curvet put mulberry and muga directly inside its ai workflow canvas. within two days, teams generated voices across education, design, crm and enterprise workflows.

100 + hours in 2 days

100 + hours in 2 days

curvet ai

snaptv uses silk to give a voice to bite-sized lessons made for how india learns, quickly, on mobile, in simple hindi and easy english.

snaptv uses silk to give a voice to bite-sized lessons made for how india learns, quickly, on mobile, in simple hindi and easy english.

snap tv

Image (9) (no background)

jee concepts are difficult enough. monk learning uses silk to turn dense explanations into clear, natural voice for aspirants preparing every day.

jee concepts are difficult enough. monk learning uses silk to turn dense explanations into clear, natural voice for aspirants preparing every day.

monk learning

frequently asked questions

frequently asked questions

frequently asked questions

yes, silk text to speech can turn a manuscript into audiobook narration. it gives you a narrator-style voice for chapters, samples, kids stories, fiction, and nonfiction.
silk is a good choice for book narration because you can test a passage in the playground with ₹10 free credit, then generate chapters as wav or mp3 when you are ready to produce the book.
monotone audiobook narration happens when every paragraph gets the same weight. silk's text to speech models follow punctuation, paragraph breaks, and delivery cues, so the narration has more room around quiet moments and important turns.
yes, silk can handle fiction dialogue when the lines include clear speaker cues and delivery direction. silk can give dialogue a different feel from the surrounding narration, so character lines do not blur into the narrator voice.
yes, you can turn book text from a manuscript or extracted pdf text into narration with silk's text to speech models. clean the text first, remove page headers or footers, then generate a sample passage in silk playground before producing longer sections.
yes. use the silk playground to generate a short passage first and hear how the narrator handles the tone of the book. once the sample feels right, you can use the silk api to generate the rest in a repeatable workflow.

explore more voice use cases

where teams use english voice

explore more voice use cases

start building with rumik's tts api

start building with rumik's tts api

join us

we are a close-knit group of researchers, engineers, and designers working on the hardest problems in ai.