VoiceVox を介してテキスト読み上げ機能を提供する Model Context Protocol サーバーです。このサーバーにより、Claude は VoiceVox エンジンが提供する様々な音声を使用してテキストから音声を生成することができます。
🔗 リンク
- GitHub: https://github.com/Sunwood-ai-labs/mcp-voicevox
- PyPI: https://pypi.org/project/mcp-server-voicevox/
✨ 機能
- テキスト読み上げ: 指定したテキストを VoiceVox の音声で読み上げます。
- 話者選択: 多数の個性的な話者から音声を選択できます。
- 音声の自動再生: 生成した音声をその場で自動的に再生します。
- 音声ファイル保存: 生成した音声は
soundフォルダに.wavファイルとして保存されます。
🚀 前提条件
- VoiceVox エンジンが動作していること(ローカルまたはリモートで)
- Python 3.10 以上
📦 インストール
uv の使用(推奨)
uv を使用する場合は特別なインストールは必要ありません。直接 uvx を使用して mcp-server-voicevox を実行します。
⚙️ 設定
VoiceVox エンジン
このサーバーは動作するために VoiceVox エンジンが必要です。エンジンの起動は手動で行う必要があります。
デフォルトでは http://localhost:50021 への接続を試みます。--voicevox-url 引数で別の URL を指定することができます。
VoiceVox エンジンは 公式 VoiceVox リポジトリ からダウンロードしてインストールできます。
Claude Desktop 用の設定
Claude Desktop の設定に追加:
uvx を使用する場合
{
"mcpServers": {
"voicevox": {
"command": "uvx",
"args": ["mcp-server-voicevox", "--voicevox-url=http://localhost:50021"]
}
}
}
🛠️ 利用可能なツール
-
get_voices- VoiceVox から利用可能な音声のリストを取得- 引数は必要ありません
-
text_to_speech- VoiceVox を使用してテキストを音声に変換- 必須引数:
text(文字列): 音声に変換するテキスト
- オプション引数:
speaker_id(整数、デフォルト: 1): 使用する音声の IDspeed(数値、デフォルト: 1.3): 再生速度の倍率
- 必須引数:
🎵 特別な機能
- 生成後の音声は、プラットフォーム固有の方法で自動的に再生されます:
- Windows: デフォルトのシステムプレーヤーを使用
- macOS: 内蔵の
afplayユーティリティを使用 - Linux: まず
aplayを試し、失敗した場合はxdg-openにフォールバック
📁 プロジェクト構造
📄 ライセンス
mcp-server-voicevox は MIT ライセンスの下で提供されています。これは、MIT ライセンスの条件に従い、自由に使用、修正、配布することができることを意味します。
🔗 リンク
Release files for mcp-server-voicevox 0.2.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| mcp_server_voicevox-0.2.0.tar.gz | 34.2 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| mcp_server_voicevox-0.2.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 42.3 kB
Release files / mcp_server_voicevox-0.2.0.tar.gz
| Download URL | mcp_server_voicevox-0.2.0.tar.gz |
|---|---|
| Size | 34.2 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
32a90993767508c1c459d70acb92a5728ce005ae55c4611a2f36a4e53126e77a
|
|
BLAKE2b-256 checksum How to use checksums |
f094284909c94afe008945201671e4d9b46e0057afbb14af2eb131da76afb53c
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.12.9
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Jun 28, 2025.
Transparency logRelease files / mcp_server_voicevox-0.2.0-py3-none-any.whl
| Download URL | mcp_server_voicevox-0.2.0-py3-none-any.whl |
|---|---|
| Size | 8.1 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
aa00d2687a50220520c0185e47bcc4a83eab1d80fe0fb6b0b868f60df468a288
|
|
BLAKE2b-256 checksum How to use checksums |
59bfc81ce242abdd7e8a769371a627697f6d892ebcbc77e2fd8b1c2fea857f04
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.12.9
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Jun 28, 2025.
Transparency log