Genesys Cloud - Main

 View Only

Sign Up

  • 1.  How are you testing and comparing TTS voices in Genesys Cloud?

    Posted 3 hours ago

    Hi everyone,

    We're currently building a Voice Bot POC in Genesys Cloud to understand the caller's reason for contacting us and route them to the appropriate queue based on the detected intent.

    As part of the POC, we'd also like to compare the different TTS voices available and understand what others are using in production.

    How are you testing and comparing different TTS voices? Have you built a separate test flow to listen to the same prompts across different voices, or found an easier way to compare them without continually changing the voice and republishing the flow?

    I'm thinking of creating a simple test flow using a Data Table with a set of sample phrases, so we can work through them and listen to the different voices one by one. Has anyone done something similar, or is there an easier way to compare voices in Genesys Cloud?

    We're currently using Genesys Enhanced TTS with Amazon Polly Olivia NTTS, but we're also interested in hearing whether others are using Enhanced TTS, Generative TTS, or third-party TTS providers and what influenced your choice.

    Also, for any Australian customers here, which voice have you found sounds the most natural for an Australian audience?

    Would love to hear what others are doing and any lessons learned.

    Thanks!


    #ConversationalAI(Bots,VirtualAgent,etc.)

    ------------------------------
    Phaneendra
    Technical Solutions Consultant
    ------------------------------


  • 2.  RE: How are you testing and comparing TTS voices in Genesys Cloud?
    Best Answer

    Posted 2 hours ago

    What I normally do during the implementation and ramp-up phase is something very similar.

    When I need to present and compare TTS options with a customer, I usually create one separate test flow for each of the main voices available within what the customer has contracted/licensed.

    I replicate the same basic flow and configure a different TTS voice in each copy. All of them use the same Data Table with the same sample phrases, so the comparison is consistent.

    For the phrases, I normally include scenarios where TTS differences or limitations are easier to notice, such as:

    • Dates and times

    • Currency values

    • Long numbers

    • Phone numbers

    • Acronyms

    • Proper names

    • Decimal values

    • Longer sentences or prompts

    If there are spare DIDs available in the organization, I assign one DID to each test flow. This makes the customer session very easy because they can simply call each number and compare the voices directly.

    If there are no numbers available for testing, I use the option to call the flow directly from within Genesys Cloud. In that case, I provide the customer with the exact flow name to test. Since I normally have one flow per voice, I usually include the TTS/voice name in the flow name as well, which makes it easier to identify which one they are listening to.

    The main reason I started doing it this way was to prepare everything before the customer testing session.

    Instead of changing the TTS voice during the meeting, republishing the flow, testing it, changing it again, and repeating the process, all the flows are already published with their respective voices.

    It was the easiest way I found to anticipate that work and make the actual testing session with the customer much smoother.

    So I think your Data Table approach makes a lot of sense. The only thing I would add, especially if you are planning to compare several voices live with the customer, is having one pre-published test flow per voice, all using exactly the same set of phrases.

    That has worked well for me during ramp-up and customer validation sessions.



    ------------------------------
    Raphael Poliesi
    ------------------------------



  • 3.  RE: How are you testing and comparing TTS voices in Genesys Cloud?

    Posted an hour ago

    Hi Raphael,

    Thanks! This is really helpful. I like the idea of having one pre-published flow per voice and using the same Data Table and sample phrases across all of them. That should make the comparison much easier.

    One other question if you happen to know: when using TTS and speech recognition/STT within a Dialog Engine Bot Flow, is there an additional usage cost associated with either of these?

    I've been looking through the Genesys documentation, but some of the pricing/included usage information seems to refer specifically to Virtual Agent/Agentic Virtual Agent or AI Guides, so I'm not completely clear on how TTS/STT usage is charged for a standard Bot Flow.

    Would be interested to know how this works in practice.

    Thanks again!



    ------------------------------
    Phaneendra
    Technical Solutions Consultant
    ------------------------------