Pricing for Noty, your AI producer

Plans differ only by volume, products, and rights — everything else is included.

Monthly
Yearly
Save 16%
Noty

“I write, cast, narrate, paint, build story books and sing. On every plan.”

Starter

Try it out
$8/month

$96 billed yearly — save $12 (11%)

~25 hours of audio / year · 1.2M credits, deposited upfront

or ~50 story books · ~380 pictures · ~200 songs

What you’ll say to Noty

“Turn my PDF into a podcast”“Draw the dragon from chapter 2”“Voice my slides in Spanish”

2,600+ voices, 130+ languages

The full voice library, every accent

Everything Noty does, in small doses

Pictures, songs and story books included

Most Popular

Personal

For creators & families
$15.83/month

$190 billed yearly — save $38 (16%)

~120 hours of audio / year · 6M credits, deposited upfront

or ~250 story books · ~1,890 pictures · ~1,000 songs

What you’ll say to Noty

“A bedtime story about Mia and her cat, with pictures”“A lullaby for Leo about the moon”“Make this a two-host podcast”

2,600+ voices in 130+ languages

Podcasts, dialogs, songs and covers in any voice

Your own voice

Noty narrates stories, podcasts and dialogs in your voice

Your own cast of characters

Named heroes with their own voice and face, in every story

Pro

Publish & sell
$40.83/month

$490 billed yearly — save $98 (16%)

~360 hours of audio / year · 18M credits, deposited upfront

or ~740 story books · ~5,670 pictures · ~3,000 songs

What you’ll say to Noty

“Narrate my whole book, one voice per character”“Cut a 30-second Spotify ad”“A cover and a voiceover for my YouTube series”

Unlimited cloned voices

Clients, brands, a whole cast — every voice you need, with commercial rights

Full commercial rights

Sell what you make — audiobooks, ads, YouTube

2,600+ voices & 5 seats

Emotion tags, character builder, priority support

Cancel anytime · Secure checkout by Paddle · Prices in USD

Estimates based on Natural voices at 1.2 credits per character (introductory rate through Dec 31, 2026) and typical project sizes across Notevibes Studio. Classic voices use fewer credits and go further.

Don’t want a subscription?

One-time pack — ~20 hours of audio · 1M credits — never expires, no recurring charges. Same price as a month of Pro, no commitment.

$49 one-time

Creators produced 200+ hours of audio with Notevibes in the last 30 days.

Compare plans

Every plan includes 2,600+ voices in 130+ languages, Noty the AI producer, emotion tags, multi-voice dialogs, podcasts, audiobooks, music, pictures and stories, and MP3/WAV export.

StarterPersonalPro
What’s included
Voice cloning (your own voice)—1 voiceUnlimited
Commercial usage rights——Full
API access——
Team members——Up to 5
SupportPriority

What you can make per month

Hours of audio~2 hours~10 hours~30 hours
Audiobook chapters~7~35~100
Podcast episodes~25~120~360
Hours of transcription~2 hours~10 hours~35 hours
Hours of audio translation~1 hour~6 hours~15 hours
Narrated PowerPoint decks~70~350~1,040
Pictures~30~160~470
Songs & lullabies~15~85~250
Credits100K500K1.5M

Running a large team?

Organization plans: pooled credits, unlimited seats.

Frequently Asked Questions

Get answers to common questions about our Text-to-Speech service

How do credits work?

+

Every generation draws from your credit balance based on the length of the text and the voice you pick. Your balance refreshes to your plan’s full amount at the start of each billing cycle. Yearly plans deposit the entire year of credits upfront, so you can front-load big projects like an audiobook.

What are Natural voices, and why do they go further?

+

Natural is the default voice engine in the Studio: 2,600+ voices across 130+ languages. Natural voices use about half the credits of Expressive voices — 1.2 credits per character instead of 2.3 — so every plan makes about twice the audio for the same price. That is an introductory rate through Dec 31, 2026; from January 1 Natural voices use 2.0 credits per character, still less than Expressive.

Which languages are supported?

+

Voices: 148 languages and regional variants — Afrikaans, Albanian, Amharic, Arabic, Arabic (Egypt), Armenian, Assamese, Australian English, Azerbaijani, Balinese, Bangla (BD), Banjar (Arabic script), Banjar (Latin script), Basque, Belarusian, Bengali, Bhojpuri, Bosnian, Brazilian Portuguese, British English, Buginese, Bulgarian, Burmese, Canadian English, Cantonese, Catalan, Cebuano, Chhattisgarhi, Chichewa, Chinese, Crimean Tatar, Croatian, Czech, Danish, Dutch, Dyula, Dzongkha, Esperanto, Estonian, Filipino, Finnish, French, Fulani (Fulfulde), Galician, Georgian, German, Greek, Guarani, Gujarati, Haitian Creole, Hausa, Hebrew, Hindi, Hungarian, Icelandic, Igbo, Ilocano, Indian English, Indonesian, Irish English, Italian, Japanese, Javanese, Kabyle, Kamba, Kannada, Kashmiri (Arabic script), Kashmiri (Devanagari), Kazakh, Khmer, Kikuyu, Kinyarwanda, Kongo, Konkani, Korean, Kurdish (Sorani), Kyrgyz, Lao, Latgalian, Latin, Latvian, Lingala, Lithuanian, Luganda, Luxembourgish, Macedonian, Magahi, Maithili, Malagasy, Malay, Malayalam, Maltese, Marathi, Minangkabau, Mizo, Mongolian, Nepali, New Zealand English, Norwegian, Norwegian Nynorsk, Occitan, Odia, Oromo, Pangasinan, Pashto, Persian, Polish, Portuguese, Punjabi, Romanian, Russian, Santali, Scottish Gaelic, Sepedi (Northern Sotho), Serbian, Sesotho, Setswana, Shona, Sindhi, Sinhala, Slovak, Slovenian, Somali, South African English, Spanish, Spanish (LatAm), Spanish (Mexico), Sundanese, Swahili, Swati, Swedish, Tajik, Tamil, Tatar, Telugu, Thai, Tigrinya, Turkish, Twi (Akan), Ukrainian, Urdu, Uyghur, Uzbek, Vietnamese, Welsh, Wolof, Xhosa, Zulu. Audio transcription and translation: 46 languages — English, Spanish, French, German, Italian, Portuguese, Japanese, Korean, Chinese, Arabic, Hindi, Russian, Dutch, Turkish, Persian, Polish, Vietnamese, Indonesian, Thai, Swedish, Ukrainian, Greek, Czech, Romanian, Hebrew, Hungarian, Filipino, Danish, Norwegian, Finnish, Malay, Bengali, Urdu, Tamil, Telugu, Marathi, Gujarati, Kannada, Malayalam, Punjabi, Nepali, Sinhala, Burmese, Khmer, Lao, Mongolian; the spoken language is detected automatically.

Can I clone my own voice?

+

Yes, on the Personal plan and up: Personal includes one cloned voice, Pro as many as you need. Open the voice picker, choose Clone your voice, then record or upload a short sample and read one consent sentence. You get a private voice that only you can use, or you can share it with your workspace. Speech in your voice uses credits like any Natural voice.

What happens if I run out of credits mid-month?

+

You never lose work — the Studio simply pauses before generating anything you can’t afford and offers a top-up right in the chat. You can grab the $49 one-time pack (1M credits, never expires), upgrade your plan, or wait for your next cycle.

How does the credit system work?

+

Each plan includes credits that renew monthly or yearly. Credits are consumed based on the length of text and voice type selected. Premium HD voices may use more credits but deliver superior quality. Transcription and audio translation draw from the same credit balance, metered by the length of the audio — so every feature shares one simple, credit-based limit.

What makes NoteVibes Pro different from other TTS services?

+

NoteVibes Pro offers 2,600+ AI voices with emotional expression, voice cloning, AI-powered content extraction from any file type (PDFs, videos, images), podcast mode with 80+ emotion tags, and advanced controls like SSML support — all in one platform.

Can I upload documents and videos for conversion?

+

Yes! Upload PowerPoint (PPTX) presentations, PDFs, Word docs, images, videos, or audio files. Our AI automatically extracts text using OCR and transcription, then converts it to natural-sounding audio. You can also paste URLs to convert web articles.

What is Podcast Mode and how does it work?

+

Podcast Mode creates engaging multi-voice conversations with emotional expression. Choose from 80+ emotion tags (happy, sad, excited, calm, angry, etc.) to make dialogues sound natural and human-like - perfect for podcasts, audiobooks, and storytelling.

What audio formats and quality options are available?

+

Export in MP3, WAV, or ULAW formats. Adjust voice parameters including pitch (-20 to +20 semitones), speaking rate (0.25x to 4.0x), and volume gain (-96dB to +16dB) for professional audio production.

How many languages and voices are supported?

+

Access 2,600+ AI voices across 130+ languages — or clone your own voice from a short recording. Each language features native-speaking voices with multiple regional accents, authentic pronunciation, and 80+ emotion tags for expressive delivery.

Can I use this for commercial projects?

+

The Personal plan is designed for personal, non-commercial use only. For commercial projects — including videos, podcasts, Spotify ads, audiobooks, e-learning courses, and marketing content — choose the Pro plan, which includes full commercial usage rights.

What happens if I run out of credits?

+

If you run out of credits, you can upgrade your plan or purchase additional credit packs. Your existing audio files, notes, and projects remain accessible. We'll notify you before credits run low.

What audiobook features are available?

+

Upload PDF, EPUB, DOCX, MOBI, AZW3, TXT, or paste a URL — AI automatically splits content into chapters. Smart character detection identifies all speaking characters and assigns unique voices for multi-cast narration. Scene prompts guide delivery with mood and pacing directions for each scene. AI-generated character portraits bring your cast to life. Read Along mode highlights text word-by-word in sync with narration. Edit per-chapter, export complete audiobook or individual files.

Can I cancel or switch plans anytime?

+

Yes. Upgrade, downgrade, or cancel anytime from billing settings. No fees, no contracts. Changes at next billing cycle. Content stays accessible.

How does character detection and voice assignment work?

+

AI scans your entire book to identify all speaking characters, profiles their personality and role, and automatically assigns a unique voice to each one for multi-cast narration. Each character gets a custom vocal persona with tone, pacing, and emotional delivery directions. Scene prompts provide mood and pacing guidance for each scene to shape the performance.

What is Read Along mode?

+

Read Along highlights text word-by-word in real time as the audiobook plays, so you can follow along with the narration. The view auto-scrolls to keep the active text centered. If you scroll manually, a Follow button lets you jump back into sync instantly.

What are AI character portraits?

+

AI generates unique portrait illustrations for every speaking character in your audiobook. Portrait styles are automatically matched to your book's genre — painterly for fantasy, noir for mystery, cinematic for drama, and more. Portraits appear alongside character profiles in your published library listing.

Can I transcribe audio and video to text?

+

Yes. Upload an audio or video file — or dictate live in your browser — and AI returns clean, timestamped text in 70+ languages, with word-level timing you can edit just by chatting. Transcription is available on every plan and draws from your credits, metered by the length of the audio.

Can I translate audio between languages?

+

Yes. Audio translation converts speech from one language to another while preserving the original speaker’s tone, pacing, and intonation — not a flat robot voice. You get the translated audio plus both the source and translated transcripts, across 70+ language pairs. It’s included on every plan and uses credits based on the length of the audio.

What is AI music generation?

+

Generate original background music and soundtracks directly from a text prompt — no licensing or royalties needed. Pick a mood, genre, or scene description and the AI composes a track ready to use under your podcast, voiceover, or audiobook. Available on Personal and Pro.

Can Noty draw pictures?

+

Yes. Ask in the chat — “draw the dragon from chapter 2”, “a cover for this book”, “a portrait of Mia” — and Noty paints it right there. Name a character and their portrait is used as the reference, so the same face carries across every picture. Ask for bigger or sharper and it switches to the full image model. Available on every plan; each picture shows its price on the card and waits for your tap.

Can Noty make a story book or a song for my kid?

+

Yes. Drop a photo of your child or pet into the chat and Noty asks only for a name — from then on that character has a storybook portrait that stars in stories, songs and pictures. The photo itself is not kept, only the drawn portrait. “A bedtime story about Mia and her cat, with pictures” makes a whole picture book: story, cast, narration, cover and a picture on every page. “A lullaby for Leo about the moon” gets written and sung — fun, learning, adventure, birthday and dance songs work the same way. Available on every plan, metered by credits.

Still have questions?

Our team is here to help you get started with text-to-speech.