这是一个使用Windows内置语音服务提供文本转语音(TTS)和语音转文本(STT)功能的模型上下文协议(MCP)服务器。该服务器通过PowerShell命令利用Windows语音API(SAPI),无需依赖外部API或服务。
git clone https://github.com/ExpressionsBot/MS-Lucidia-Voice-Gateway-MCP.git
cd MS-Lucidia-Voice-Gateway-MCP
npm install
npm run build
npm run test
http://localhost:3000使用Windows SAPI将文本转换为语音。
参数:
text(必需):要转换为语音的文本voice(可选):使用的语音(例如:"Microsoft David Desktop")speed(可选):语速从0.5到2.0(默认值:1.0)示例:
fetch('http://localhost:3000/tts', {
method: 'POST',
headers: { 'Content-Type': 'application/json' },
body: JSON.stringify({
text: "Hello, this is a test",
voice: "Microsoft David Desktop",
speed: 1.0
})
});
使用Windows语音识别记录音频并将其转换为文本。
参数:
duration(可选):录音时长(秒,默认值:5,最大值:60)示例:
fetch('http://localhost:3000/stt', {
method: 'POST',
headers: { 'Content-Type': 'application/json' },
body: JSON.stringify({
duration: 5
})
}).then(response => response.json())
.then(data => console.log(data.text));
确保启用了Windows语音识别:
检查可用语音:
Add-Type -AssemblyName System.Speech
(New-Object System.Speech.Synthesis.SpeechSynthesizer).GetInstalledVoices().VoiceInfo.Name
测试语音识别:
MIT