文字合成语音的方法 -两个离线包，百度接口，一个开源项目

最新推荐文章于 2025-03-31 16:17:39 发布

卖香油的少掌柜

最新推荐文章于 2025-03-31 16:17:39 发布

阅读量3.3k

点赞数 1

文章标签： python 语音识别人工智能神经网络

本文链接：https://blog.csdn.net/qq_58832911/article/details/120868695

版权

这篇博客介绍了三种将文本转换为语音的方法：使用Python库pyttsx和SpeechLib，以及调用百度语音API。pyttsx和SpeechLib是Python实现的本地解决方案，而百度API提供了高质量的语音合成服务，但可能涉及费用。文中提供了详细的代码示例和使用流程。

摘要生成于 C知道，由 DeepSeek-R1 满血版支持，前往体验 >

目标是将我们输入的文本文字转换为语音。

使用 pyttsx 将文本转化为语音

使用名为 pyttsx 的 python 包，你可以将文本转换为语音。直接使用 pip 就可以进行安装，

命令：pip install pyttsx3

使用实例：

import pyttsx3 as pyttsx

engine = pyttsx.init()
engine.say('你好吗')
engine.runAndWait()

使用 SpeechLib将文本转化为语音

使用 SpeechLib，可以从文本文件中获取输入，再将其转换为语音。先使用 pip 安装。

命令：pip install comtypes

使用实例：

from comtypes.client import CreateObject
from comtypes.gen import SpeechLib

#实例化拼读对象
engine=CreateObject('SAPI.SpVoice')
#实例化文件流对象
stream=CreateObject('SAPI.SpFileStream')

infile='demo.txt'
outfile='demo_audio.wav'

# 打开流文件，以读文本录入
stream.open(outfile,SpeechLib.SSFMCreateForWrite)
#输出方式流媒体
engine.AudioOutputStream=stream
#读取文本内容
f=open(infile,'r',encoding='utf-8')
theText=f.read()
f.close()
engine.speak(theText)
stream.close()

使用百度接口api

使用百度接口，可以从文本文件中获取输入，再将其转换为语音。不过超过免费额度之后会收费。

API_KEY ，SECRET_KEY 码需要去百度ai平台注册，领取。

使用实例：

import sys
import json

# 保证兼容python2以及python3
IS_PY3 = sys.version_info.major == 3
if IS_PY3:
    from urllib.request import urlopen
    from urllib.request import Request
    from urllib.error import URLError
    from urllib.parse import urlencode
    from urllib.parse import quote_plus
else:
    import urllib2
    from urllib import quote_plus
    from urllib2 import urlopen
    from urllib2 import Request
    from urllib2 import URLError
    from urllib import urlencode

API_KEY = '******************************'

SECRET_KEY = '******************************'

TEXT = "三分钟前，由北京市顺义区二经路与二纬路交汇处北侧，北京首都国际机场T3航站楼 去往 东城区北三环东路36号喜来登大酒店(北京金隅店)"



TTS_URL = 'http://tsn.baidu.com/text2audio'

"""  TOKEN start """

TOKEN_URL = 'http://openapi.baidu.com/oauth/2.0/token'


"""
    获取token
"""
def fetch_token():
    params = {'grant_type': 'client_credentials',
              'client_id': API_KEY,
              'client_secret': SECRET_KEY}
    post_data = urlencode(params)
    if (IS_PY3):
        post_data = post_data.encode('utf-8')
    req = Request(TOKEN_URL, post_data)
    try:
        f = urlopen(req, timeout=5)
        result_str = f.read()
    except URLError as err:
        print('token http response http code : ' + str(err.code))
        result_str = err.read()
    if (IS_PY3):
        result_str = result_str.decode()


    result = json.loads(result_str)

    if ('access_token' in result.keys() and 'scope' in result.keys()):
        if not 'audio_tts_post' in result['scope'].split(' '):
            print ('please ensure has check the tts ability')
            exit()
        return result['access_token']
    else:
        print ('please overwrite the correct API_KEY and SECRET_KEY')
        exit()


"""  TOKEN end """

if __name__ == '__main__':

    token = fetch_token()

    tex = quote_plus(TEXT)  # 此处TEXT需要两次urlencode

    params = {'tok': token, 'tex': tex, 'cuid': "quickstart",
              'lan': 'zh', 'ctp': 1}  # lan ctp 固定参数

    data = urlencode(params)

    req = Request(TTS_URL, data.encode('utf-8'))
    has_error = False
    try:
        f = urlopen(req)
        result_str = f.read()

        headers = dict((name.lower(), value) for name, value in f.headers.items())

        has_error = ('content-type' not in headers.keys() or headers['content-type'].find('audio/') < 0)
    except  URLError as err:
        print('http response http code : ' + str(err.code))
        result_str = err.read()
        has_error = True

    save_file = "error.txt" if has_error else u'大姚的订单信息.mp3'

    with open(save_file, 'wb') as of:
        of.write(result_str)

    if has_error:
        if (IS_PY3):
            result_str = str(result_str, 'utf-8')
        print("tts api  error:" + result_str)

    print("file saved as : " + save_file)