← ClaudeAtlas

potluck-ttslisted

Text-to-speech via Potluck /v1/audio/speech using OpenAI / ElevenLabs / Deepgram / Edge TTS / Google TTS / Hyperbolic / Inworld voices. Use when the user wants to convert text to speech, generate audio, voiceover, narrate, or read text aloud.
Ezero23/potluck · ★ 1 · AI & Automation · score 65
Install: claude install-skill Ezero23/potluck
# Potluck — Text-to-Speech Requires `POTLUCK_URL` (and `POTLUCK_KEY` if auth enabled). See https://raw.githubusercontent.com/Ezero23/potluck/refs/heads/main/skills/potluck/SKILL.md for setup. ## Discover ```bash # 1) List models curl $POTLUCK_URL/v1/models/tts | jq '.data[].id' # 2) Per-model metadata (params, voicesUrl if voice-by-id) curl "$POTLUCK_URL/v1/models/info?id=el/eleven_multilingual_v2" # 3) List voices (elevenlabs, edge-tts, deepgram, inworld, local-device). Optional ?lang=vi curl "$POTLUCK_URL/v1/audio/voices?provider=edge-tts&lang=vi" | jq '.data[].model' ``` `model` field in `/v1/audio/speech` = voice ID directly (e.g. `edge-tts/vi-VN-HoaiMyNeural`, `el/<voice_id>`, or `openai/tts-1` model+default voice). ## Endpoint `POST $POTLUCK_URL/v1/audio/speech` | Field | Required | Notes | |---|---|---| | `model` | yes | voice ID from `/v1/models/tts` | | `input` | yes | text to speak | Query `?response_format=mp3` (default, raw bytes) or `?response_format=json` (`{audio: base64, format}`). ## Examples Save MP3: ```bash curl -X POST "$POTLUCK_URL/v1/audio/speech" \ -H "Authorization: Bearer $POTLUCK_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"openai/tts-1","input":"Hello world"}' \ --output speech.mp3 ``` JS (save file): ```js import { writeFile } from "node:fs/promises"; const r = await fetch(`${process.env.POTLUCK_URL}/v1/audio/speech`, { method: "POST", headers: { "Authorization": `Bearer ${process.env.POTLUCK_KEY}`, "Content-