Murf Falcon enables you to build voice agents that are ultra-fast, expressive, scalable, and highly cost-efficient. It has 55 ms model latency and 130 ms end-to-end latency. It supports 35+ languages and pricing starting at 1 cent per minute.
Murf API delivers enterprise-grade AI voice generation with industry-leading text-to-speech models, transforming multimedia and conversational experiences.
With a library of 150+ natural-sounding voices across 35 languages and 20+ speaking styles, Murf enables high-quality voiceovers for diverse applications. Its proprietary Multilingual technology allows a single voice to seamlessly switch between multiple languages while preserving authentic pronunciation patterns specific to each language.
Murf Falcon model, is an ultra-fast streaming model, optimized for conversational AI and real-time voice agents. It delivers natural speech with time-to-first-audio under 130ms. It supports data residency in 10+ regions including USA, Canada, India, EU, UK, Japan, Australia and more.
Highlights
Falcon API offers 55ms model latency and 130ms end-to-end latency.
It supports data residency in 10+ regions including USA, Canada, India, EU, UK, Japan, Australia and more.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
You pay based on usage, measured by the number of characters you convert into speech. There is one pricing dimension: characters. Your cost scales directly with how much text you generate. There are no tiers or fixed commitments. As you send more characters through the text-to-speech API, your charges grow in proportion. This flat, usage-based structure lets you predict spending based on your actual volume, whether you generate a little or a lot.
Top-of-mind questions for buyers
What counts as one character for billing?
One character is a single character in the text you submit to the API, including letters, digits, spaces, and punctuation. Your charge scales with the total characters you send for speech generation. Longer scripts convert more characters and cost more; shorter scripts cost less.
Am I charged if I generate no speech in a given period?
You pay only for characters you convert into speech. There are no fixed commitments or minimum fees tied to this dimension. If you send no text through the API, no character charges accrue for that period. Your cost follows your actual usage directly.
How does character-based billing map to audio output like minutes of speech?
Billing meters the input text by character count, not by resulting audio length. The number of characters you submit determines your charge, regardless of how long the generated speech runs. Faster or slower speech playback does not change the character count you are billed for.
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Murf API offers enterprise-grade AI voice generation capabilities with its industry-leading text-to-speech models, transforming multimedia and conversational experiences. Our API achieves over 99% pronunciation accuracy and offers unparalleled speech customization through styles, pauses, duration matching and variation controls.
The Deepdub API integrates our groundbreaking emotive-based Text-to-Speech technology, providing businesses with an efficient tool to create lifelike, emotionally resonant speech for a variety of applications. Designed for enterprise-scale use, this API supports extensive customization options, including accent control and advanced voice modification, ensuring that each audio output is perfectly tailored to meet specific content needs.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.