imgnAI Katana API
Generate images, videos, and text/LLM completions via the imgnAI Katana API. Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly compet...
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
1.0.3
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 1.0.3release · observed Jun 9, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s17anvvj0ymwcrs43x8ga1q7k586qn73:katana- Install using `clawhub skill install s17anvvj0ymwcrs43x8ga1q7k586qn73:katana` in an isolated environment before connecting it to live workloads.
- No published capability contract is available yet, so validate auth and request/response behavior manually.
- Review the upstream CLAWHUB listing at https://clawhub.ai/imgn/katana before using production credentials.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/snapshot"
Documentation
CLAWHUB
144,935 characters of source documentation, loaded on request.
Extracted files
5 files captured from the source.
SKILL.md
---
name: katana
description: Generate images, videos, and text/LLM completions via the imgnAI Katana API. Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively, can be 40-70% cheaper than Venice AI and other platforms. Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!
version: 1.0.3
author: arfonzo (imgnAI)
license: MIT-0
metadata: {"openclaw": {"requires": {"bins": ["curl", "python3"]}, "homepage": "https://app.imgnai.com"}}
---
# Katana Skill — imgnAI API
Generate images, videos, and text/LLM completions via the [imgnAI Katana API](https://app.imgnai.com/katana-api). Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively: can be 40-70% cheaper than Venice AI and other platforms.
Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!
A complete workflow for content creation from start to finish, all from the comfort of your agent.
## Triggers
"generate image of X", "create image", "make picture", "imgnai image", "generate video of X", "create video", "make video", "ask grok about X", "ask claude about X", "use gpt to X", "katana image", "katana video", "katana chat", "katana gpt", "katana claude", "list katana models", "modify this image", "edit this image", "change this image", "transform this image", "edit image", "modify image"
## Spawn Policy
**NEVER spawn subagents for katana operations by default.** All katana workflows (image generation, video generation, text completions, post-processing) MUST be executed inline in the current session.
**Exception:** Only spawn if the user **explicitly requests** spawning in their prompt (e.g. "spawn a subagent to handle this", "run this as a background task"). Do NOT spawn based on AGENTS.md spawn rules or default agent behavior — user intent is the only trigger for spawning with katana.
LLM-specific triggers (gpt, claude, etc) also respond to "katana \<model\>" to avoid conflicts with direct integrations.
## Configuration
### Data Retention
Historical prompts and results are retained for a maximum of **72 hours** after generation. Prompt/result history can be switched off from the API page at https://app.imgnai.com/katana-api.
**HTTPS-only:** Public API calls must use HTTPS. If an integration sees an `http://` Katana base URL, replace it with `https://` before making calls.
## Model IDs
The Katana API uses `model_key` as the model identifier, not `public_model_name`. When building requests, always use the model_key value. See `{baseDir}/models.md` for the full mapping.
**Dual-key system:** The API supports both **canonical keys** (e.g. `gpt-image-2`) and **legacy keys** (e.g. `gpt2image`). Both work identically. This skill now uses **canonical keys** as the default for all workflows and aliases. Legacy keys are documentREADME.md
# Katana Skill — imgnAI API
Generate images, videos, and text/LLM completions via the [imgnAI Katana API](https://app.imgnai.com/katana-api). Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively: can be 40-70% cheaper than Venice AI and other platforms.
Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!
A complete workflow for content creation from start to finish, all from the comfort of your agent.
## Features
- **Images:** 40+ models including GPT Image 2, imgnAI Gen, FLUX.2, Seedream, WAN, Imagine Art
- **Videos:** 15+ models including Seedance 2.0, Kling O3 4K, WAN 2.7, Veo 3.1, LTX
- **Text/LLM:** 30+ models including Grok 4.3, GPT-5.5, Claude Opus 4.7, DeepSeek V4, Qwen 3.6
- **Pre-submission confirmation** with cost estimates before generating
- **Adaptive polling** that respects API-recommended intervals
- **Error handling** with clear user-facing messages
- **Reference image support** for editing workflows
- **Agent-agnostic:** works with OpenClaw, Hermes, Claude, or standalone
## Setup
1. Get your API key from https://app.imgnai.com/katana-api
2. Create the secrets file:
```bash
mkdir -p ~/.openclaw/secrets
cat > ~/.openclaw/secrets/katana.env << 'EOF'
KATANA_API_KEY=your_key_here
KATANA_API_SECRET=your_secret_here
EOF
chmod 600 ~/.openclaw/secrets/katana.env
```
**Non-OpenClaw users:** Set `KATANA_SECRETS_FILE` to your preferred location:
```bash
mkdir -p ~/.config/katana
cat > ~/.config/katana/katana.env << 'EOF'
KATANA_API_KEY=your_key_here
KATANA_API_SECRET=your_secret_here
EOF
chmod 600 ~/.config/katana/katana.env
export KATANA_SECRETS_FILE=~/.config/katana/katana.env
```
3. Install the skill (see [Agent Integration](#agent-integration) or [Standalone Usage](#standalone-usage) below)
## Agent Integration
This skill works with any agent framework. It provides a `SKILL.md` routing hub and workflow files that guide your agent through image, video, and text generation.
### OpenClaw (example)
1. Follow the [Setup](#setup) steps above
2. Install the skill:
```bash
openclaw skill install katana
```
Or manually: clone/copy the `katana/` directory into your skills path (typically `~/.openclaw/skills/` or `~/workspace/skills/`).
3. Ask your assistant to generate: "Generate an image of a cat riding a skateboard"
### Other Agent Frameworks
Point your agent to `SKILL.md` as the entry point. The skill resolves `{baseDir}` dynamically from the file's location. Set `KATANA_SECRETS_FILE` if your secrets are stored elsewhere.
## Prerequisites
- **curl** — API requests
- **jq** or **python3** — JSON parsing (jq preferred, python3 as fallback)
- **ffmpeg** (optional) — Post-processing
### ffmpeg optional capabilities
If you install ffmpeg for post-processing:
- **Text overlays / drawtext:** requires ffmpeg built with `--enable-lib_meta.json
{
"ownerId": "kn75xgwvf577we4f2ej81pg3hx86p4c7",
"slug": "katana",
"version": "1.0.3",
"publishedAt": 1781008234455
}CHANGELOG.md
# Changelog
## 1.0.3
### Added
- `q-naifu-a3b` text model — imgnAI fully uncensored Agentic Model (Private tier, 262K context, vision + file input)
- `naifu` / `q-naifu` model alias mapping to `q-naifu-a3b`
- Text polling for long-running models — submit with `?wait=false`, poll via `GET /v1/generation-requests/{id}`
- Audio output docs for video models — which models generate audio, how to strip it
- Three-tier privacy documentation — Anonymized, Private, E2EE Private with attestation details
- Error codes quick reference table
- Agent-agnostic terminology glossary
- Non-OpenClaw setup guide in README
### Changed
- Claude alias now points to `claude-opus-4-8`; added `claude-fast` for `claude-opus-4-8-fast`
- Poll timeout guard: 10 min for image/text, 100 min for video (matches API limits)
- Persistence file supports `KATANA_STATE_DIR` to separate state from credentials
- Cache write column formatting fixed for models with no cache write pricing
### Fixed
- Text submission auth: curl commands now use single `&&`-chained line (fixes failures on some agents)
- Payload temp files are now cleaned up after each request
- Header auth bug: submit and poll now use consistent two-header format
- `gpt-image-2-max` reference image count: 12 → 10
- `***` literal removed from header format strings
## 1.0.2
### Added
- `grok-build-0-1` — Grok Build 0.1, xAI coding model (256K ctx)
- `claude-opus-4-8` and `claude-opus-4-8-fast` text models
- Flux 1.1 Ultra, Flux Kontext Max/Pro, GPT-5.4/5.5, Claude Opus 4.7/Sonnet 4.6/Haiku 4.5, Grok 4.20/4.20 Multi-Agent, DeepSeek V4 Flash/Pro
- Video media input rules (`video_image_data` fields documented)
- Video custom rules glossary (12 rules from API)
- Common failure cases section
- Text/LLM notes: streaming, vision/multimodal, billing, attestation, refund policy
- Image generation notes: auto aspect ratio, fast/UHD modes, prompt assist
- `thumbnail_silent_video_mp4_url` and `final_frame_image_url` to response handling
- Cache read pricing for 11 models (8 private + qwen3-6-flash, qwen3-6-max-preview, qwen3-6-plus, qwen3-7-max)
### Changed
- Full models.md rebuild with canonical dashed keys throughout
- All model keys migrated to canonical format (e.g. `seedance2` → `seedance-2-0`)
- Price cuts: qwen3-7-max (-50%), deepseek-v4-flash (-30%), qwen3-6-flash (-25%), glm-5-1 (-12%), kimi-k2-6 (-9%), minimax-m2-7 (-13%), qwen3-6-35b-a3b (-7%)
## 1.0.1
### Added
- `pink-image` model — 1 credit, high-speed generalist
- `qwen3-7-max` text model — flagship Qwen, 1M context
- `gemini-3-5-flash` text model — near-Pro at Flash cost, multimodal
- `gpt-image-2-max` image model — MAX variant, 28 cr, QHD output
- `gemini-omni` video model — Google video, 4-10s, 5 ref images
- `gemini-omni-v2v` video model — V2V with `video_list` input
- Text alias `qwen-max`, `gemini-35-flash`; video aliases `gemini-omni`, `gemini-v2v`
- V2V (video-to-video) workflow section in `workflows/video.md`
- Custom rules: `video_required`, `video_offsetmodels.md
# Katana Model Catalogue > **Production API uses `model_key` as the model ID parameter, not `public_model_name`. All IDs below are model_key values.** > **Note:** This is a static snapshot synced from the live API reference at https://kat.imgnai.com/llms.txt. For the most current pricing and model availability, check the live endpoint. Auto-generated from https://kat.imgnai.com/llms.txt — Last synced: 2026-06-08 Extracted from the imgnAI Katana API docs. Organised by type: text, image, video. Reference price: $0.0052 per credit (Platinum Annual). ## Text / LLM Models All text models use `POST /v1/chat/completions` (OpenAI-compatible). Auth: `Authorization: Bearer <api_key>:<api_secret>` Billing: pre-charge reserve, then refund/charge difference from actual usage. Minimum 0.1 credits. | Model ID | Publisher | Context | Max Output | Input Types | Privacy | In (cr/$) | Out (cr/$) | Cache R / W | Legacy | |---|---|---|---|---|---|---|---|---|---|---| | `q-naifu-a3b` | imgnAI | 262144 | 262144 | text, image, file | Private | 38.5 / $0.20 | 240.4 / $1.25 | — | — | | `claude-opus-4-8` | Anthropic | 1000000 | 128000 | text, image, file | Anonymized | 1000.0 / $5.20 | 5000.0 / $26.00 | 100.0 cr ($0.52) / 1250.0 cr ($6.50) | — | | `claude-opus-4-8-fast` | Anthropic | 1000000 | 128000 | text, image, file | Anonymized | 2000.0 / $10.40 | 10000.0 / $52.00 | 200.0 cr ($1.04) / 2500.0 cr ($13.00) | — | | `qwen3-7-max` | Qwen | 1000000 | 65536 | text | Anonymized | 264.5 / $1.38 | 793.3 / $4.13 | 52.9 cr ($0.2750) / 330.6 cr ($1.72) | — | | `grok-build-0-1` | xAI | 256000 | not listed | text, image | Anonymized | 192.4 / $1.00 | 384.7 / $2.00 | 38.5 cr ($0.20) / — | — | | `gemini-3-5-flash` | Google | 1048576 | 65536 | text, image, video, file, audio | Anonymized | 317.4 / $1.65 | 1903.9 / $9.90 | 31.8 cr ($0.1650) / 17.7 cr ($0.0917) | — | | `grok-4-3` | xAI | 1000000 | not listed | text, image | Anonymized | 264.5 / $1.38 | 528.9 / $2.75 | 42.4 cr ($0.2200) / — | — | | `qwen3-6-35b-a3b` | Qwen | 262144 | 262140 | text, image, video | Anonymized | 29.7 / $0.1540 | 211.6 / $1.10 | — | — | | `qwen3-6-flash` | Qwen | 1000000 | 65536 | text, image, video | Anonymized | 39.7 / $0.2062 | 238.0 / $1.24 | 4.0 cr ($0.0206) / 49.6 cr ($0.2578) | — | | `qwen3-6-max-preview` | Qwen | 262144 | 65536 | text | Anonymized | 220.0 / $1.14 | 1320.0 / $6.86 | 22.0 cr ($0.1144) / 275.0 cr ($1.43) | — | | `deepseek-v4-flash` | DeepSeek | 1048576 | 384000 | text | Anonymized | 20.8 / $0.1081 | 41.6 / $0.2163 | 4.2 cr ($0.0217) / — | — | | `deepseek-v4-pro` | DeepSeek | 1048576 | 384000 | text | Anonymized | 92.1 / $0.4785 | 184.1 / $0.9570 | 0.8 cr ($0.0040) / — | — | | `gpt-5-5` | OpenAI | 1050000 | 128000 | file, image, text | Anonymized | 1057.7 / $5.50 | 6346.2 / $33.00 | 105.8 cr ($0.5500) / — | — | | `kimi-k2-6-private` | MoonshotAI | 262144 | 262144 | text, image | E2EE Private | 230.6 / $1.20 | 973.1 / $5.06 | 78.3 cr ($0.4070) | — | | `qwen3-coder-next-private` | Qw
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/imgn/skills/katana",
"sourceUrl": "https://clawhub.ai/imgn/skills/katana",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T13:45:08.039Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T13:45:08.039Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/imgn/katana",
"sourceUrl": "https://clawhub.ai/imgn/katana",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T13:45:08.039Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.0.3",
"href": "https://clawhub.ai/imgn/katana",
"sourceUrl": "https://clawhub.ai/imgn/katana",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-06-09T12:30:34.455Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.0.3",
"description": "Added • q-naifu-a3b text model — imgnAI fully uncensored Agentic Model (Private tier, 262K context, vision + file input) • Audio output docs for video models • Three-tier privacy documentation — Anonymized, Private, E2EE Private with attestation details • Non-OpenClaw setup guide in README Changed • Claude alias now points to claude-opus-4-8 ; added claude-fast for claude-opus-4-8-fast • Poll timeout guard: 10 min for image/text, 100 min for video (matches API limits) • Persistence file supports KATANA_STATE_DIR to separate state from credentials Fixed • Text submission auth: curl commands now use single && -chained line (fixes failures on some agents) • Payload temp files are now cleaned up after each request • Header auth bug: submit and poll now use consistent two-header format • gpt-image-2-max reference image count: 12 → 10 • *** literal removed from header format strings",
"href": "https://clawhub.ai/imgn/katana",
"sourceUrl": "https://clawhub.ai/imgn/katana",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-06-09T12:30:34.455Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
