微软Azure文本转音频，保存成MP3文件【代码python3】

achirandliu

已于 2023-10-29 23:18:15 修改

阅读量978

点赞数

文章标签： microsoft python azure 微软Azure 文本转音频保存mp3文件

于 2023-10-29 23:13:34 首次发布

本文链接：https://blog.csdn.net/achirandliu/article/details/134109593

版权

本文介绍了如何使用微软Azure的PythonSDK将文本转换为音频文件（MP3格式），并提供了代码示例，包括设置语音合成配置、指定语言和输出格式，以及处理结果和错误。

摘要由CSDN通过智能技术生成

标签：文本转音频并保存mp3文件；微软Azure；

微软Azure可以将文本转音频，并保存mp3文件，直接上代码
代码格式：python 3

import os
import azure.cognitiveservices.speech as speechsdk

# This example requires environment variables named "SPEECH_KEY" and "SPEECH_REGION"
speech_config = speechsdk.SpeechConfig(subscription=os.environ.get('SPEECH_KEY'), region=os.environ.get('SPEECH_REGION'))

# The language of the voice that speaks.
speech_config.speech_synthesis_voice_name='zh-CN-YunjianNeural'   # 这个男声 有 磁性
text = "讲一个笑话：和朋友去饭店吃饭，要了一盘红烧肉，结果发现怎么咬都咬不动，我顿时就火了，把服务员叫过来喊道：你们这肉怎么咬都咬不动，把你们经理叫来。服务员说：叫我们经理干啥啊，你都咬不动，他能咬得动啊！"
speech_config.set_speech_synthesis_output_format(speechsdk.SpeechSynthesisOutputFormat.Audio24Khz160KBitRateMonoMp3) # 这里配置文件为mp3格式，要保存其它文件格式，修改这里参数
speech_synthesizer = speechsdk.SpeechSynthesizer(speech_config=speech_config, audio_config=None)

result = speech_synthesizer.speak_text_async(text).get()
stream = speechsdk.AudioDataStream(result)
stream.save_to_wav_file("D:/file.mp3")  # mp3文件保存路径


if result.reason == speechsdk.ResultReason.SynthesizingAudioCompleted:
    print("Speech synthesized Completed,  for text [{}]".format(text))
elif result.reason == speechsdk.ResultReason.Canceled:
    cancellation_details = result.cancellation_details
    print("Speech synthesis canceled: {}".format(cancellation_details.reason))
    if cancellation_details.reason == speechsdk.CancellationReason.Error:
        if cancellation_details.error_details:
            print("Error details: {}".format(cancellation_details.error_details))
            print("Did you set the speech resource key and region values?")