OATDA Transcribe Audio
Transcribe audio to text using OATDA's unified audio API. Triggers when the user wants speech-to-text, transcription of meetings, podcasts, voice notes, subt... Skill: OATDA Transcribe Audio Owner: devcsde Summary: Transcribe audio to text using OATDA's unified audio API. Triggers when the user wants speech-to-text, transcription of meetings, podcasts, voice notes, subt... Tags: latest:1.1.0 Version history: v1.1.0 | 2026-07-17T23:21:08.622Z | user Sync with latest API model IDs. Verify models via oatda-list-models. v1.0.3 | 2026-07-17T23:15:15.937Z | user - Removed sample f
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
1.1.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 1.1.0release · observed Jul 17, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s1715k9c0vyqryjet8h0ek3xc58447z6:oatda-transcribe-audio- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-transcribe-audio/snapshot"
Documentation
CLAWHUB
31,398 characters of source documentation, loaded on request.
Extracted files
3 files captured from the source.
SKILL.md
---
name: oatda-transcribe-audio
description: Transcribe audio to text using OATDA's unified audio API. Triggers when the user wants speech-to-text, transcription of meetings, podcasts, voice notes, subtitles, timestamps, or Whisper-style transcription through OATDA.
homepage: https://oatda.com
metadata:
{
"openclaw":
{
"emoji": "📝",
"requires": { "bins": ["curl", "jq"], "env": ["OATDA_API_KEY"], "config": ["~/.oatda/credentials.json"] },
"primaryEnv": "OATDA_API_KEY",
},
}
---
# OATDA Audio Transcription
Transcribe audio files to text through OATDA's unified audio API.
## API Key Resolution
All commands need the OATDA API key. Resolve it inline for each `exec` call:
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}"
```
If the key is empty or `null`, tell the user to get one at https://oatda.com and configure it.
**Security**: Never print the full API key. Only verify existence or show first 8 chars.
## Model Mapping
| User says | Provider | Model |
|-----------|----------|-------|
| whisper, whisper-1, openai whisper (default) | openai | whisper-1 |
| transcription, speech to text, stt | openai | whisper-1 |
**Default**: `openai` / `whisper-1` if no model specified.
If the user provides `provider/model` format directly (for example `openai/whisper-1`), split on `/`.
> ⚠️ Models change over time. If a model ID fails, query `oatda-list-models` with `?type=audio` first.
## Input Preparation
The transcription endpoint supports:
- `multipart/form-data` with a local file upload
- JSON with a base64 data URL in `file`
- JSON with `file_base64` for providers that support direct base64 payloads
Maximum audio file size is 25MB.
For local files, prefer multipart upload because it is simpler and avoids large JSON bodies.
## Discovering Audio Model Parameters
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}" && \
curl -s -X GET "https://oatda.com/api/v1/llm/models?type=audio" \
-H "Authorization: Bearer $OATDA_API_KEY" | jq '.audio_models[] | {id, supported_params}'
```
Look for:
- `audio_modes` containing `transcription`
- supported `response_format` values
- optional timestamp, diarization, or streaming support
## API Call (multipart)
```bash
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/null | jq -r '.profiles[.defaultProfile].apiKey' 2>/dev/null)}" && \
curl -s -X POST "https://oatda.com/api/v1/llm/transcriptions" \
-H "Authorization: Bearer $OATDA_API_KEY" \
-F "provider=<PROVIDER>" \
-F "model=<MODEL>" \
-F "file=@<AUDIO_FILE>" \
-F "response_format=json"
```
## Alternative API Call (base64 JSON)
```bash
AUDIO_DATA_URL="data:audio/mpeg;base64,$(base64 -w 0 audio.mp3)"
export OATDA_API_KEY="${OATDA_API_KEY:-$(cat ~/.oatda/credentials.json 2>/dev/_meta.json
{
"ownerId": "kn7anz9axpc20rkkkfarxeh589845ssw",
"slug": "oatda-transcribe-audio",
"version": "1.1.0",
"publishedAt": 1784330468622
}skill-card.md
## Description: Transcribes audio to text through OATDA's unified audio API for speech-to-text, meetings, podcasts, voice notes, subtitles, timestamps, and Whisper-style transcription. This skill is ready for commercial/non-commercial use. ## Publisher: [devcsde](https://clawhub.ai/user/devcsde) ### License/Terms of Use: MIT-0 ## Use Case: External users, developers, and agents use this skill to transcribe user-selected audio files through OATDA and return transcript text, subtitles, timestamps, or segment details when requested. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: Selected audio files and their contents are sent to OATDA for transcription. Mitigation: Use only when the user's OATDA account, consent requirements, and data-handling expectations are appropriate; avoid highly sensitive recordings unless those requirements are satisfied. Risk: The skill uses an OATDA API key for authenticated requests. Mitigation: Keep the key in OATDA_API_KEY or the local credentials file and do not print the full key in responses or logs. Risk: Large or unsupported audio files can fail transcription requests. Mitigation: Keep files under the documented 25MB limit, prefer multipart upload for local files, and query available audio models when model IDs fail. ## Reference(s): - [OATDA](https://oatda.com) - [OATDA audio models endpoint](https://oatda.com/api/v1/llm/models?type=audio) - [OATDA transcription endpoint](https://oatda.com/api/v1/llm/transcriptions) - [ClawHub skill page](https://clawhub.ai/devcsde/skills/oatda-transcribe-audio) ## Skill Output: **Output Type(s):** [Text, Shell commands, Configuration, Guidance] **Output Format:** [Markdown with inline bash commands and JSON response excerpts] **Output Parameters:** [1D] **Other Properties Related to Output:** [Requires OATDA_API_KEY; uploads selected audio files to OATDA; maximum audio file size is 25MB.] ## Skill Version(s): 1.1.0 (source: evidence.release.version) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/devcsde/skills/oatda-transcribe-audio",
"sourceUrl": "https://clawhub.ai/devcsde/skills/oatda-transcribe-audio",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T09:36:52.166Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-transcribe-audio/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-transcribe-audio/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T09:36:52.166Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T09:36:52.166Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.1.0",
"href": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-17T23:21:08.622Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-transcribe-audio/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-devcsde-oatda-transcribe-audio/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.1.0",
"description": "Sync with latest API model IDs. Verify models via oatda-list-models.",
"href": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceUrl": "https://clawhub.ai/devcsde/oatda-transcribe-audio",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-17T23:21:08.622Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
