agentCLAWHUBUnverified

imgnAI Katana API

Generate images, videos, and text/LLM completions via the imgnAI Katana API. Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly compet...

OpenClaw

Rank

62

Safety

84

Downloads

1.1k

Updated

Oct 11, 2026

Version

1.0.3

Source

CLAWHUB

About

What it does, and when to use it.

Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.

Avoid when

  • Contract metadata is missing or unavailable for deterministic execution.

Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing

Public facts

Every fact links back to the source it came from.

Vendor
Clawhubvendor · observed Oct 11, 2026
Protocol compatibility
OpenClawcompatibility · observed Oct 11, 2026
Adoption signal
1.1K downloadsadoption · observed Oct 11, 2026
Latest release
1.0.3release · observed Jun 9, 2026
Handshake status
UNKNOWNsecurity

Install and run

Setup complexity: low.

clawhub skill install s17anvvj0ymwcrs43x8ga1q7k586qn73:katana
  1. Install using `clawhub skill install s17anvvj0ymwcrs43x8ga1q7k586qn73:katana` in an isolated environment before connecting it to live workloads.
  2. No published capability contract is available yet, so validate auth and request/response behavior manually.
  3. Review the upstream CLAWHUB listing at https://clawhub.ai/imgn/katana before using production credentials.

Contract: missing

curl -s "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/snapshot"

Documentation

CLAWHUB

144,935 characters of source documentation, loaded on request.

Extracted files

5 files captured from the source.

SKILL.md

---
name: katana
description: Generate images, videos, and text/LLM completions via the imgnAI Katana API. Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively, can be 40-70% cheaper than Venice AI and other platforms. Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!
version: 1.0.3
author: arfonzo (imgnAI)
license: MIT-0
metadata: {"openclaw": {"requires": {"bins": ["curl", "python3"]}, "homepage": "https://app.imgnai.com"}}
---

# Katana Skill — imgnAI API

Generate images, videos, and text/LLM completions via the [imgnAI Katana API](https://app.imgnai.com/katana-api). Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively: can be 40-70% cheaper than Venice AI and other platforms.

Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!

A complete workflow for content creation from start to finish, all from the comfort of your agent.


## Triggers

"generate image of X", "create image", "make picture", "imgnai image", "generate video of X", "create video", "make video", "ask grok about X", "ask claude about X", "use gpt to X", "katana image", "katana video", "katana chat", "katana gpt", "katana claude", "list katana models", "modify this image", "edit this image", "change this image", "transform this image", "edit image", "modify image"

## Spawn Policy

**NEVER spawn subagents for katana operations by default.** All katana workflows (image generation, video generation, text completions, post-processing) MUST be executed inline in the current session.

**Exception:** Only spawn if the user **explicitly requests** spawning in their prompt (e.g. "spawn a subagent to handle this", "run this as a background task"). Do NOT spawn based on AGENTS.md spawn rules or default agent behavior — user intent is the only trigger for spawning with katana.

LLM-specific triggers (gpt, claude, etc) also respond to "katana \<model\>" to avoid conflicts with direct integrations.

## Configuration

### Data Retention
Historical prompts and results are retained for a maximum of **72 hours** after generation. Prompt/result history can be switched off from the API page at https://app.imgnai.com/katana-api.

**HTTPS-only:** Public API calls must use HTTPS. If an integration sees an `http://` Katana base URL, replace it with `https://` before making calls.

## Model IDs

The Katana API uses `model_key` as the model identifier, not `public_model_name`. When building requests, always use the model_key value. See `{baseDir}/models.md` for the full mapping.

**Dual-key system:** The API supports both **canonical keys** (e.g. `gpt-image-2`) and **legacy keys** (e.g. `gpt2image`). Both work identically. This skill now uses **canonical keys** as the default for all workflows and aliases. Legacy keys are document

README.md

# Katana Skill — imgnAI API

Generate images, videos, and text/LLM completions via the [imgnAI Katana API](https://app.imgnai.com/katana-api). Supports end-to-end-encrypted (E2EE) and anonymized models. Priced highly competitively: can be 40-70% cheaper than Venice AI and other platforms.

Includes post-processing such as combining videos and images, cutting, slicing, splicing, transitions, drawing text, re-encoding, resizing and much more!

A complete workflow for content creation from start to finish, all from the comfort of your agent.


## Features

- **Images:** 40+ models including GPT Image 2, imgnAI Gen, FLUX.2, Seedream, WAN, Imagine Art
- **Videos:** 15+ models including Seedance 2.0, Kling O3 4K, WAN 2.7, Veo 3.1, LTX
- **Text/LLM:** 30+ models including Grok 4.3, GPT-5.5, Claude Opus 4.7, DeepSeek V4, Qwen 3.6
- **Pre-submission confirmation** with cost estimates before generating
- **Adaptive polling** that respects API-recommended intervals
- **Error handling** with clear user-facing messages
- **Reference image support** for editing workflows
- **Agent-agnostic:** works with OpenClaw, Hermes, Claude, or standalone

## Setup

1. Get your API key from https://app.imgnai.com/katana-api
2. Create the secrets file:
   ```bash
   mkdir -p ~/.openclaw/secrets
   cat > ~/.openclaw/secrets/katana.env << 'EOF'
   KATANA_API_KEY=your_key_here
   KATANA_API_SECRET=your_secret_here
   EOF
   chmod 600 ~/.openclaw/secrets/katana.env
   ```

   **Non-OpenClaw users:** Set `KATANA_SECRETS_FILE` to your preferred location:
   ```bash
   mkdir -p ~/.config/katana
   cat > ~/.config/katana/katana.env << 'EOF'
   KATANA_API_KEY=your_key_here
   KATANA_API_SECRET=your_secret_here
   EOF
   chmod 600 ~/.config/katana/katana.env
   export KATANA_SECRETS_FILE=~/.config/katana/katana.env
   ```

3. Install the skill (see [Agent Integration](#agent-integration) or [Standalone Usage](#standalone-usage) below)

## Agent Integration

This skill works with any agent framework. It provides a `SKILL.md` routing hub and workflow files that guide your agent through image, video, and text generation.

### OpenClaw (example)

1. Follow the [Setup](#setup) steps above
2. Install the skill:
   ```bash
   openclaw skill install katana
   ```
   Or manually: clone/copy the `katana/` directory into your skills path (typically `~/.openclaw/skills/` or `~/workspace/skills/`).
3. Ask your assistant to generate: "Generate an image of a cat riding a skateboard"

### Other Agent Frameworks

Point your agent to `SKILL.md` as the entry point. The skill resolves `{baseDir}` dynamically from the file's location. Set `KATANA_SECRETS_FILE` if your secrets are stored elsewhere.

## Prerequisites

- **curl** — API requests
- **jq** or **python3** — JSON parsing (jq preferred, python3 as fallback)
- **ffmpeg** (optional) — Post-processing

### ffmpeg optional capabilities

If you install ffmpeg for post-processing:
- **Text overlays / drawtext:** requires ffmpeg built with `--enable-lib

_meta.json

{
  "ownerId": "kn75xgwvf577we4f2ej81pg3hx86p4c7",
  "slug": "katana",
  "version": "1.0.3",
  "publishedAt": 1781008234455
}

CHANGELOG.md

# Changelog

## 1.0.3

### Added
- `q-naifu-a3b` text model — imgnAI fully uncensored Agentic Model (Private tier, 262K context, vision + file input)
- `naifu` / `q-naifu` model alias mapping to `q-naifu-a3b`
- Text polling for long-running models — submit with `?wait=false`, poll via `GET /v1/generation-requests/{id}`
- Audio output docs for video models — which models generate audio, how to strip it
- Three-tier privacy documentation — Anonymized, Private, E2EE Private with attestation details
- Error codes quick reference table
- Agent-agnostic terminology glossary
- Non-OpenClaw setup guide in README

### Changed
- Claude alias now points to `claude-opus-4-8`; added `claude-fast` for `claude-opus-4-8-fast`
- Poll timeout guard: 10 min for image/text, 100 min for video (matches API limits)
- Persistence file supports `KATANA_STATE_DIR` to separate state from credentials
- Cache write column formatting fixed for models with no cache write pricing

### Fixed
- Text submission auth: curl commands now use single `&&`-chained line (fixes failures on some agents)
- Payload temp files are now cleaned up after each request
- Header auth bug: submit and poll now use consistent two-header format
- `gpt-image-2-max` reference image count: 12 → 10
- `***` literal removed from header format strings

## 1.0.2

### Added
- `grok-build-0-1` — Grok Build 0.1, xAI coding model (256K ctx)
- `claude-opus-4-8` and `claude-opus-4-8-fast` text models
- Flux 1.1 Ultra, Flux Kontext Max/Pro, GPT-5.4/5.5, Claude Opus 4.7/Sonnet 4.6/Haiku 4.5, Grok 4.20/4.20 Multi-Agent, DeepSeek V4 Flash/Pro
- Video media input rules (`video_image_data` fields documented)
- Video custom rules glossary (12 rules from API)
- Common failure cases section
- Text/LLM notes: streaming, vision/multimodal, billing, attestation, refund policy
- Image generation notes: auto aspect ratio, fast/UHD modes, prompt assist
- `thumbnail_silent_video_mp4_url` and `final_frame_image_url` to response handling
- Cache read pricing for 11 models (8 private + qwen3-6-flash, qwen3-6-max-preview, qwen3-6-plus, qwen3-7-max)

### Changed
- Full models.md rebuild with canonical dashed keys throughout
- All model keys migrated to canonical format (e.g. `seedance2` → `seedance-2-0`)
- Price cuts: qwen3-7-max (-50%), deepseek-v4-flash (-30%), qwen3-6-flash (-25%), glm-5-1 (-12%), kimi-k2-6 (-9%), minimax-m2-7 (-13%), qwen3-6-35b-a3b (-7%)

## 1.0.1

### Added
- `pink-image` model — 1 credit, high-speed generalist
- `qwen3-7-max` text model — flagship Qwen, 1M context
- `gemini-3-5-flash` text model — near-Pro at Flash cost, multimodal
- `gpt-image-2-max` image model — MAX variant, 28 cr, QHD output
- `gemini-omni` video model — Google video, 4-10s, 5 ref images
- `gemini-omni-v2v` video model — V2V with `video_list` input
- Text alias `qwen-max`, `gemini-35-flash`; video aliases `gemini-omni`, `gemini-v2v`
- V2V (video-to-video) workflow section in `workflows/video.md`
- Custom rules: `video_required`, `video_offset

models.md

# Katana Model Catalogue

> **Production API uses `model_key` as the model ID parameter, not `public_model_name`. All IDs below are model_key values.**

> **Note:** This is a static snapshot synced from the live API reference at https://kat.imgnai.com/llms.txt. For the most current pricing and model availability, check the live endpoint.

Auto-generated from https://kat.imgnai.com/llms.txt — Last synced: 2026-06-08

Extracted from the imgnAI Katana API docs. Organised by type: text, image, video.

Reference price: $0.0052 per credit (Platinum Annual).

## Text / LLM Models

All text models use `POST /v1/chat/completions` (OpenAI-compatible).
Auth: `Authorization: Bearer <api_key>:<api_secret>`
Billing: pre-charge reserve, then refund/charge difference from actual usage. Minimum 0.1 credits.

| Model ID | Publisher | Context | Max Output | Input Types | Privacy | In (cr/$) | Out (cr/$) | Cache R / W | Legacy |
|---|---|---|---|---|---|---|---|---|---|---|
| `q-naifu-a3b` | imgnAI | 262144 | 262144 | text, image, file | Private | 38.5 / $0.20 | 240.4 / $1.25 | — | — |
| `claude-opus-4-8` | Anthropic | 1000000 | 128000 | text, image, file | Anonymized | 1000.0 / $5.20 | 5000.0 / $26.00 | 100.0 cr ($0.52) / 1250.0 cr ($6.50) | — |
| `claude-opus-4-8-fast` | Anthropic | 1000000 | 128000 | text, image, file | Anonymized | 2000.0 / $10.40 | 10000.0 / $52.00 | 200.0 cr ($1.04) / 2500.0 cr ($13.00) | — |
| `qwen3-7-max` | Qwen | 1000000 | 65536 | text | Anonymized | 264.5 / $1.38 | 793.3 / $4.13 | 52.9 cr ($0.2750) / 330.6 cr ($1.72) | — |
| `grok-build-0-1` | xAI | 256000 | not listed | text, image | Anonymized | 192.4 / $1.00 | 384.7 / $2.00 | 38.5 cr ($0.20) / — | — |
| `gemini-3-5-flash` | Google | 1048576 | 65536 | text, image, video, file, audio | Anonymized | 317.4 / $1.65 | 1903.9 / $9.90 | 31.8 cr ($0.1650) / 17.7 cr ($0.0917) | — |
| `grok-4-3` | xAI | 1000000 | not listed | text, image | Anonymized | 264.5 / $1.38 | 528.9 / $2.75 | 42.4 cr ($0.2200) / — | — |
| `qwen3-6-35b-a3b` | Qwen | 262144 | 262140 | text, image, video | Anonymized | 29.7 / $0.1540 | 211.6 / $1.10 | — | — |
| `qwen3-6-flash` | Qwen | 1000000 | 65536 | text, image, video | Anonymized | 39.7 / $0.2062 | 238.0 / $1.24 | 4.0 cr ($0.0206) / 49.6 cr ($0.2578) | — |
| `qwen3-6-max-preview` | Qwen | 262144 | 65536 | text | Anonymized | 220.0 / $1.14 | 1320.0 / $6.86 | 22.0 cr ($0.1144) / 275.0 cr ($1.43) | — |
| `deepseek-v4-flash` | DeepSeek | 1048576 | 384000 | text | Anonymized | 20.8 / $0.1081 | 41.6 / $0.2163 | 4.2 cr ($0.0217) / — | — |
| `deepseek-v4-pro` | DeepSeek | 1048576 | 384000 | text | Anonymized | 92.1 / $0.4785 | 184.1 / $0.9570 | 0.8 cr ($0.0040) / — | — |
| `gpt-5-5` | OpenAI | 1050000 | 128000 | file, image, text | Anonymized | 1057.7 / $5.50 | 6346.2 / $33.00 | 105.8 cr ($0.5500) / — | — |
| `kimi-k2-6-private` | MoonshotAI | 262144 | 262144 | text, image | E2EE Private | 230.6 / $1.20 | 973.1 / $5.06 | 78.3 cr ($0.4070) | — |
| `qwen3-coder-next-private` | Qw
Github ReposUpdated 2d agoRank 70

AionUi

Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!

MCPOPENCLAW
Github ReposUpdated 6mo agoRank 70

activepieces

AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents

OPENCLAW
Github ReposUpdated 6mo agoRank 70

cherry-studio

AI productivity studio with smart chat, autonomous agents, and 300+ assistants.

MCPOPENCLAW
Github ReposUpdated 7mo agoRank 70

CopilotKit

The Frontend for Agents & Generative UI. React + Angular

OPENCLAW

Machine-readable data

The same record, as JSON, for agents and crawlers.

{
  "facts": [
    {
      "factKey": "vendor",
      "category": "vendor",
      "label": "Vendor",
      "value": "Clawhub",
      "href": "https://clawhub.ai/imgn/skills/katana",
      "sourceUrl": "https://clawhub.ai/imgn/skills/katana",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-11T13:45:08.039Z",
      "isPublic": true
    },
    {
      "factKey": "protocols",
      "category": "compatibility",
      "label": "Protocol compatibility",
      "value": "OpenClaw",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/contract",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/contract",
      "sourceType": "contract",
      "confidence": "medium",
      "observedAt": "2026-10-11T13:45:08.039Z",
      "isPublic": true
    },
    {
      "factKey": "traction",
      "category": "adoption",
      "label": "Adoption signal",
      "value": "1.1K downloads",
      "href": "https://clawhub.ai/imgn/katana",
      "sourceUrl": "https://clawhub.ai/imgn/katana",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-11T13:45:08.039Z",
      "isPublic": true
    },
    {
      "factKey": "latest_release",
      "category": "release",
      "label": "Latest release",
      "value": "1.0.3",
      "href": "https://clawhub.ai/imgn/katana",
      "sourceUrl": "https://clawhub.ai/imgn/katana",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-06-09T12:30:34.455Z",
      "isPublic": true
    },
    {
      "factKey": "handshake_status",
      "category": "security",
      "label": "Handshake status",
      "value": "UNKNOWN",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/trust",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-imgn-katana/trust",
      "sourceType": "trust",
      "confidence": "medium",
      "observedAt": null,
      "isPublic": true
    }
  ],
  "events": [
    {
      "eventType": "release",
      "title": "Release 1.0.3",
      "description": "Added • q-naifu-a3b text model — imgnAI fully uncensored Agentic Model (Private tier, 262K context, vision + file input) • Audio output docs for video models • Three-tier privacy documentation — Anonymized, Private, E2EE Private with attestation details • Non-OpenClaw setup guide in README Changed • Claude alias now points to claude-opus-4-8 ; added claude-fast for claude-opus-4-8-fast • Poll timeout guard: 10 min for image/text, 100 min for video (matches API limits) • Persistence file supports KATANA_STATE_DIR to separate state from credentials Fixed • Text submission auth: curl commands now use single && -chained line (fixes failures on some agents) • Payload temp files are now cleaned up after each request • Header auth bug: submit and poll now use consistent two-header format • gpt-image-2-max reference image count: 12 → 10 • *** literal removed from header format strings",
      "href": "https://clawhub.ai/imgn/katana",
      "sourceUrl": "https://clawhub.ai/imgn/katana",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-06-09T12:30:34.455Z",
      "isPublic": true
    }
  ]
}

Record generated Oct 11, 2026.

Sponsored

Ads related to imgnAI Katana API and adjacent AI workflows.