agentCLAWHUBUnverified

Prompt Performance Tester - UnisAI

Test prompts across Claude, GPT, and Gemini models and get detailed latency, cost, quality, consistency, and error metrics with smart recommendations.

OpenClaw

Rank

62

Safety

84

Downloads

1.9k

Updated

Apr 15, 2026

Version

1.1.9

Source

CLAWHUB

About

What it does, and when to use it.

Capability contract not published. No trust telemetry is available yet. 1.9K downloads reported by the source. Last updated 4/15/2026.

Avoid when

  • Contract metadata is missing or unavailable for deterministic execution.

Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing

Public facts

Every fact links back to the source it came from.

Vendor
Clawhubvendor · observed Apr 15, 2026
Protocol compatibility
OpenClawcompatibility · observed Apr 15, 2026
Adoption signal
1.9K downloadsadoption · observed Apr 15, 2026
Latest release
1.1.9release · observed Feb 27, 2026
Handshake status
UNKNOWNsecurity

Install and run

Setup complexity: low.

clawhub skill install kn77yjs5esft2kgsd6dpz9c92n80dgsy:prompt-performance-tester
  1. Install using `clawhub skill install kn77yjs5esft2kgsd6dpz9c92n80dgsy:prompt-performance-tester` in an isolated environment before connecting it to live workloads.
  2. No published capability contract is available yet, so validate auth and request/response behavior manually.
  3. Review the upstream CLAWHUB listing at https://clawhub.ai/vedantsingh60/prompt-performance-tester before using production credentials.

Contract: missing

curl -s "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/snapshot"

Documentation

CLAWHUB

62,598 characters of source documentation, loaded on request.

Extracted files

4 files captured from the source.

SKILL.md

# Prompt Performance Tester

**Model-agnostic prompt benchmarking across 9 providers.**

Pass any model ID — provider auto-detected. Compare latency, cost, quality, and consistency across Claude, GPT, Gemini, DeepSeek, Grok, MiniMax, Qwen, Llama, and Mistral.

---

## 🚀 Why This Skill?

### Problem Statement
Comparing LLM models across providers requires manual testing:
- No systematic way to measure performance across models
- Cost differences are significant but not easily comparable
- Quality varies by use case and provider
- Manual API testing is time-consuming and error-prone

### The Solution
Test prompts across any model from any supported provider simultaneously. Get performance metrics and recommendations based on latency, cost, and quality.

### Example Cost Comparison
For 10,000 requests/day with average 28 input + 115 output tokens:
- Claude Opus 4.6: ~$30.15/day ($903/month)
- Gemini 2.5 Flash-Lite: ~$0.05/day ($1.50/month)
- DeepSeek Chat: ~$0.14/day ($4.20/month)
- Monthly cost difference (Opus vs Flash-Lite): $901.50

---

## ✨ What You Get

### Model-Agnostic Multi-Provider Testing
Pass any model ID — provider is auto-detected from the model name prefix.
No hardcoded list; new models work without code changes.

| Provider | Example Models | Prefix | Required Key |
|----------|---------------|--------|--------------|
| **Anthropic** | claude-opus-4-6, claude-sonnet-4-6, claude-haiku-4-5-20251001 | `claude-` | ANTHROPIC_API_KEY |
| **OpenAI** | gpt-5.2-pro, gpt-5.2, gpt-5.1 | `gpt-`, `o1`, `o3` | OPENAI_API_KEY |
| **Google** | gemini-2.5-pro, gemini-2.5-flash, gemini-2.5-flash-lite | `gemini-` | GOOGLE_API_KEY |
| **Mistral** | mistral-large-latest, mistral-small-latest | `mistral-`, `mixtral-` | MISTRAL_API_KEY |
| **DeepSeek** | deepseek-chat, deepseek-reasoner | `deepseek-` | DEEPSEEK_API_KEY |
| **xAI** | grok-4-1-fast, grok-3-beta | `grok-` | XAI_API_KEY |
| **MiniMax** | MiniMax-M2.1 | `MiniMax`, `minimax` | MINIMAX_API_KEY |
| **Qwen** | qwen3.5-plus, qwen3-max-instruct | `qwen` | DASHSCOPE_API_KEY |
| **Meta Llama** | meta-llama/llama-4-maverick, meta-llama/llama-3.3-70b-instruct | `meta-llama/`, `llama-` | OPENROUTER_API_KEY |

### Known Pricing (per 1M tokens)

| Model | Input | Output |
|-------|-------|--------|
| claude-opus-4-6 | $15.00 | $75.00 |
| claude-sonnet-4-6 | $3.00 | $15.00 |
| claude-haiku-4-5-20251001 | $1.00 | $5.00 |
| gpt-5.2-pro | $21.00 | $168.00 |
| gpt-5.2 | $1.75 | $14.00 |
| gpt-5.1 | $2.00 | $8.00 |
| gemini-2.5-pro | $1.25 | $10.00 |
| gemini-2.5-flash | $0.30 | $2.50 |
| gemini-2.5-flash-lite | $0.10 | $0.40 |
| mistral-large-latest | $2.00 | $6.00 |
| mistral-small-latest | $0.10 | $0.30 |
| deepseek-chat | $0.27 | $1.10 |
| deepseek-reasoner | $0.55 | $2.19 |
| grok-4-1-fast | $5.00 | $25.00 |
| grok-3-beta | $3.00 | $15.00 |
| MiniMax-M2.1 | $0.40 | $1.60 |
| qwen3.5-plus | $0.57 | $2.29 |
| qwen3-max-instruct | $1.60 | $6.40 |
| meta-llama/llama-4-maverick | $0.20 | $0.60 |
| meta-llama/l

_meta.json

{
  "ownerId": "kn77yjs5esft2kgsd6dpz9c92n80dgsy",
  "slug": "prompt-performance-tester",
  "version": "1.1.9",
  "publishedAt": 1772213259522
}

LICENSE.md

# UniAI Skills - Proprietary License

**Version 1.0 | Effective Date: February 2, 2024**

## 1. GRANT OF LICENSE

UniAI ("Licensor") grants you ("Licensee") a limited, non-exclusive, non-transferable, revocable license to use the ClawhHub Skills ("Software") solely in accordance with the terms of this license agreement.

## 2. LICENSE RESTRICTIONS

You may NOT:
- Reverse engineer, decompile, or disassemble the Software
- Modify, alter, or create derivative works of the Software
- Remove, obscure, or alter any proprietary notices or labels on the Software
- Share, distribute, or sublicense the Software to any third party
- Use the Software for commercial purposes without a commercial license
- Access or use the Software beyond the scope of your subscription tier
- Attempt to circumvent licensing controls or API rate limits

## 3. INTELLECTUAL PROPERTY RIGHTS

All intellectual property rights in and to the Software are retained by Licensor. This includes:
- Source code and object code
- Algorithms and methodologies
- Performance optimization techniques
- Quality scoring mechanisms
- Proprietary data structures
- Trade secrets and confidential information

## 4. PERMITTED USES

You may only:
- Use the Software as provided through the ClawhHub platform
- Access features available in your subscription tier
- Create test results and reports for internal use
- Share results with your team (if on a team plan)
- Provide feedback to improve the Software

## 5. SUBSCRIPTION TIERS

### Starter (Free)
- 5 tests per month
- 2 models per test
- Basic features
- Personal use only

### Professional ($29/month)
- Unlimited tests
- All models supported
- Advanced analytics
- API access
- Commercial use permitted

### Enterprise ($99/month)
- Team collaboration
- White-label option
- Custom integrations
- Dedicated support
- SLA guarantees

## 6. API KEY AND CREDENTIALS

- You are responsible for keeping your API keys confidential
- Do not share your license key with others
- One license per person/organization
- License keys are non-transferable
- Unauthorized sharing may result in account termination

## 7. DATA PRIVACY

- We do not retain your test data by default
- Free tier: 30-day retention
- Paid tiers: 90-day retention
- You can request data deletion anytime
- See Privacy Policy for full details

## 8. WARRANTY DISCLAIMER

THE SOFTWARE IS PROVIDED "AS-IS" WITHOUT ANY WARRANTIES. LICENSOR DISCLAIMS ALL WARRANTIES, EXPRESS OR IMPLIED, INCLUDING BUT NOT LIMITED TO:
- Merchantability
- Fitness for a particular purpose
- Non-infringement
- Accuracy of results

## 9. LIMITATION OF LIABILITY

IN NO EVENT SHALL LICENSOR BE LIABLE FOR:
- Any indirect, incidental, special, or consequential damages
- Loss of data, revenue, or profits
- Business interruption
- Even if advised of the possibility of such damages

## 10. TERMINATION

Licensor may terminate your license if you:
- Violate any terms of this agreement
- Fail to pay subscription fees
- Attempt to reverse engine

manifest.yaml

name: "Prompt Performance Tester"
id: "prompt-performance-tester"
version: "1.1.8"
description: "Model-agnostic prompt benchmarking across 9 providers. Pass any model ID from Claude, GPT, Gemini, DeepSeek, Grok, MiniMax, Qwen, Llama, Mistral — provider auto-detected. Measures latency, cost, quality, and consistency."

homepage: "https://unisai.vercel.app"
repository: "https://github.com/vedantsingh60/prompt-performance-tester"
source: "included"

intellectual_property:
  license: "free-to-use"
  license_file: "LICENSE.md"
  copyright: "© 2026 UnisAI. All rights reserved."
  distribution: "via-clawhub-only"
  source_code_access: "included"
  modification: "personal-use-only"
  reverse_engineering: "allowed-for-security-audit"

author:
  company: "UnisAI"
  contact: "[email protected]"
  website: "https://unisai.vercel.app"

category: "ai-testing"
tags:
  - "prompt-testing"
  - "performance-analysis"
  - "cost-optimization"
  - "multi-llm"
  - "quality-assurance"
  - "benchmarking"
  - "llm-comparison"
  - "ai-testing"

pricing:
  model: "free"

runtime: "local"
execution: "python"

required_env_vars:
  - "ANTHROPIC_API_KEY"   # Required if testing Claude models
  - "OPENAI_API_KEY"      # Required if testing GPT models
  - "GOOGLE_API_KEY"      # Required if testing Gemini models
  - "MISTRAL_API_KEY"     # Required if testing Mistral models
  - "DEEPSEEK_API_KEY"    # Required if testing DeepSeek models
  - "XAI_API_KEY"         # Required if testing Grok/xAI models
  - "MINIMAX_API_KEY"     # Required if testing MiniMax models
  - "DASHSCOPE_API_KEY"   # Required if testing Qwen/Alibaba models
  - "OPENROUTER_API_KEY"  # Required if testing Llama/OpenRouter models
primary_credential: "At least ONE provider API key is required per provider you want to test"

dependencies:
  python: ">=3.9"
  packages:
    - "anthropic>=0.40.0"
    - "openai>=1.60.0"
    - "google-generativeai>=0.8.0"
    - "mistralai>=1.3.0"
  install_all: "pip install anthropic openai google-generativeai mistralai"
  install_selective: |
    pip install anthropic          # Claude
    pip install openai             # GPT, DeepSeek, xAI, MiniMax, Qwen, Llama (OpenAI-compat)
    pip install google-generativeai  # Gemini
    pip install mistralai          # Mistral
  note: "Install only the SDKs for the providers you plan to test. DeepSeek, xAI, MiniMax, Qwen, and Llama all use the openai package with a custom base URL."
  requirements_file: "requirements.txt"

security:
  data_retention: "0 days"
  data_flow: "prompts-sent-to-chosen-ai-providers"
  third_party_data_sharing: |
    WARNING: This skill sends your prompts to whichever AI providers you select for testing.
    Each provider has their own data retention and privacy policies:
    - Anthropic: https://www.anthropic.com/legal/privacy
    - OpenAI: https://openai.com/policies/privacy-policy
    - Google: https://ai.google.dev/gemini-api/terms
    - Mistral: https://mistral.ai/terms/
    - DeepSeek: https://www.deepseek
Github ReposUpdated 5h agoRank 70

AionUi

Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!

MCPOPENCLAW
Github ReposUpdated 6mo agoRank 70

activepieces

AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents

OPENCLAW
Github ReposUpdated 6mo agoRank 70

cherry-studio

AI productivity studio with smart chat, autonomous agents, and 300+ assistants.

MCPOPENCLAW
Github ReposUpdated 7mo agoRank 70

CopilotKit

The Frontend for Agents & Generative UI. React + Angular

OPENCLAW

Machine-readable data

The same record, as JSON, for agents and crawlers.

{
  "facts": [
    {
      "factKey": "vendor",
      "category": "vendor",
      "label": "Vendor",
      "value": "Clawhub",
      "href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-04-15T00:45:39.800Z",
      "isPublic": true
    },
    {
      "factKey": "protocols",
      "category": "compatibility",
      "label": "Protocol compatibility",
      "value": "OpenClaw",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/contract",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/contract",
      "sourceType": "contract",
      "confidence": "medium",
      "observedAt": "2026-04-15T00:45:39.800Z",
      "isPublic": true
    },
    {
      "factKey": "traction",
      "category": "adoption",
      "label": "Adoption signal",
      "value": "1.9K downloads",
      "href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-04-15T00:45:39.800Z",
      "isPublic": true
    },
    {
      "factKey": "latest_release",
      "category": "release",
      "label": "Latest release",
      "value": "1.1.9",
      "href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-02-27T17:27:39.522Z",
      "isPublic": true
    },
    {
      "factKey": "handshake_status",
      "category": "security",
      "label": "Handshake status",
      "value": "UNKNOWN",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/trust",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/trust",
      "sourceType": "trust",
      "confidence": "medium",
      "observedAt": null,
      "isPublic": true
    }
  ],
  "events": [
    {
      "eventType": "release",
      "title": "Release 1.1.9",
      "description": "- Updated provider/model lists to include more example models and the latest pricing. - Expanded cost comparison examples to include DeepSeek Chat. - Added clarification that unlisted models are supported, but cost is shown as $0.00 with a warning. - Improved explanation of quality, cost, and performance metrics for broader clarity. - Enhanced recommendation and real-world example sections to better showcase DeepSeek and new models.",
      "href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-02-27T17:27:39.522Z",
      "isPublic": true
    }
  ]
}

Record generated Oct 10, 2026.

Sponsored

Ads related to Prompt Performance Tester - UnisAI and adjacent AI workflows.