OATDA Generate Speech
Generate speech or audio from text using OATDA's unified audio API. Triggers when the user wants to convert text to speech, create narration, voiceovers, acc... Skill: OATDA Generate Speech Owner: devcsde Summary: Generate speech or audio from text using OATDA's unified audio API. Triggers when the user wants to convert text to speech, create narration, voiceovers, acc... Tags: latest:1.1.0 Version history: v1.1.0 | 2026-07-17T23:20:42.873Z | user Remove deprecated tts-1/tts-1-hd model references. Replace static model table with list_models guidance. Broaden description to a
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
1.1.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 1.1.0release · observed Jul 17, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s1715k9c0vyqryjet8h0ek3xc58447z6:oatda-generate-speech- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-generate-speech/snapshot"
Documentation
CLAWHUB
37,652 characters of source documentation, loaded on request.
Extracted files
3 files captured from the source.
SKILL.md
---
name: oatda-generate-speech
description: Generate speech or audio from text using OATDA's unified audio API. Triggers when the user wants to convert text to speech, create narration, voiceovers, accessibility audio, or use TTS models such as OpenAI tts-1 through OATDA.
homepage: https://oatda.com
metadata:
{
"openclaw":
{
"emoji": "🔊",
"requires": { "bins": ["curl", "jq"], "env": ["OATDA_API_KEY"], "config": ["~/.oatda/credentials.json"] },
"primaryEnv": "OATDA_API_KEY",
},
}
---
# OATDA Speech Generation
Generate spoken audio from text through OATDA's unified audio API.
## API Key Resolution
All commands need the OATDA API key. Resolve it inline for each `exec` call:
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}"
```
If the key is empty or `null`, tell the user to get one at https://oatda.com and configure it.
**Security**: Never print the full API key. Only verify existence or show first 8 chars.
## Model Mapping
| User says | Provider | Model |
|-----------|----------|-------|
| tts, tts-1, openai tts (default) | openai | tts-1 |
| tts hd, tts-1-hd | openai | tts-1-hd |
| gpt tts, gpt-4o mini tts | openai | gpt-4o-mini-tts |
**Default**: `openai` / `tts-1` if no model specified.
If the user provides `provider/model` format directly (for example `openai/tts-1`), split on `/`.
Common OpenAI voices include `alloy`, `ash`, `ballad`, `coral`, `echo`, `fable`, `nova`, `onyx`, `sage`, and `shimmer`. Use `alloy` if the user does not specify a voice.
> ⚠️ Models change over time. If a model ID fails, query `oatda-list-models` with `?type=audio` first.
## Discovering Audio Model Parameters
Query available audio models and inspect `supported_params` before sending optional fields:
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}" && \
curl -s -X GET "https://oatda.com/api/v1/llm/models?type=audio" \
-H "Authorization: Bearer $OATDA_API_KEY" | jq '.audio_models[] | {id, supported_params}'
```
Look for:
- `audio_modes` containing `tts`
- supported `voice` values
- allowed `response_format` values
- optional fields like `instructions` or `language`
## API Call
The speech endpoint returns **binary audio**, not JSON. Always save the response to a file.
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}" && \
curl -s -X POST "https://oatda.com/api/v1/llm/speech" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $OATDA_API_KEY" \
-d '{
"provider": "<PROVIDER>",
"model": "<MODEL>",
"input": "<TEXT_TO_SPEAK>",
"voice": "alloy",
"response_format": "mp3",
"speed": 1.0
}' \
--output speech.mp3
```
### Common Parameters
- `input`: Text to convert to s_meta.json
{
"ownerId": "kn7anz9axpc20rkkkfarxeh589845ssw",
"slug": "oatda-generate-speech",
"version": "1.1.0",
"publishedAt": 1784330442873
}skill-card.md
## Description: Generate speech or audio from text using OATDA's unified audio API. This skill is ready for commercial/non-commercial use. ## Publisher: [devcsde](https://clawhub.ai/user/devcsde) ### License/Terms of Use: MIT-0 ## Use Case: Developers, creators, and accessibility-focused users use this skill through an agent to generate spoken audio, narration, voiceovers, and other text-to-speech outputs from text with OATDA-supported providers. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: The command template can mishandle arbitrary user text and may allow command injection if text is pasted directly into JSON payloads. Mitigation: Build request JSON with a proper encoder such as jq --arg, validate output paths, and avoid reusing shell snippets with unsanitized user input. Risk: The skill reads an OATDA API key and sends provided text to OATDA for speech generation. Mitigation: Use the skill only when that data flow is acceptable, verify credential presence without printing the full key, and avoid sending sensitive text unless authorized. ## Reference(s): - [OATDA Homepage](https://oatda.com) - [OATDA Audio Models Endpoint](https://oatda.com/api/v1/llm/models?type=audio) - [OATDA Speech Endpoint](https://oatda.com/api/v1/llm/speech) - [ClawHub Skill Page](https://clawhub.ai/devcsde/skills/oatda-generate-speech) ## Skill Output: **Output Type(s):** [Shell commands, Configuration, Guidance, Files] **Output Format:** [Markdown with bash commands and generated audio file paths] **Output Parameters:** [1D] **Other Properties Related to Output:** [Requires curl, jq, OATDA_API_KEY, and safe JSON construction for user-provided text.] ## Skill Version(s): 1.1.0 (source: evidence.release.version and target metadata; artifact _meta.json reports 1.0.1) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/devcsde/skills/oatda-generate-speech",
"sourceUrl": "https://clawhub.ai/devcsde/skills/oatda-generate-speech",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T06:42:21.711Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-generate-speech/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-generate-speech/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T06:42:21.711Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T06:42:21.711Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.1.0",
"href": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-17T23:20:42.873Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-generate-speech/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-generate-speech/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.1.0",
"description": "Remove deprecated tts-1/tts-1-hd model references. Replace static model table with list_models guidance. Broaden description to all TTS providers.",
"href": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-generate-speech",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-17T23:20:42.873Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
