Prompt Performance Tester - UnisAI
Test prompts across Claude, GPT, and Gemini models and get detailed latency, cost, quality, consistency, and error metrics with smart recommendations.
Rank
62
Safety
84
Downloads
1.9k
Updated
Apr 15, 2026
Version
1.1.9
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.9K downloads reported by the source. Last updated 4/15/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Apr 15, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Apr 15, 2026
- Adoption signal
- 1.9K downloadsadoption · observed Apr 15, 2026
- Latest release
- 1.1.9release · observed Feb 27, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install kn77yjs5esft2kgsd6dpz9c92n80dgsy:prompt-performance-tester- Install using `clawhub skill install kn77yjs5esft2kgsd6dpz9c92n80dgsy:prompt-performance-tester` in an isolated environment before connecting it to live workloads.
- No published capability contract is available yet, so validate auth and request/response behavior manually.
- Review the upstream CLAWHUB listing at https://clawhub.ai/vedantsingh60/prompt-performance-tester before using production credentials.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/snapshot"
Documentation
CLAWHUB
62,598 characters of source documentation, loaded on request.
Extracted files
4 files captured from the source.
SKILL.md
# Prompt Performance Tester **Model-agnostic prompt benchmarking across 9 providers.** Pass any model ID — provider auto-detected. Compare latency, cost, quality, and consistency across Claude, GPT, Gemini, DeepSeek, Grok, MiniMax, Qwen, Llama, and Mistral. --- ## 🚀 Why This Skill? ### Problem Statement Comparing LLM models across providers requires manual testing: - No systematic way to measure performance across models - Cost differences are significant but not easily comparable - Quality varies by use case and provider - Manual API testing is time-consuming and error-prone ### The Solution Test prompts across any model from any supported provider simultaneously. Get performance metrics and recommendations based on latency, cost, and quality. ### Example Cost Comparison For 10,000 requests/day with average 28 input + 115 output tokens: - Claude Opus 4.6: ~$30.15/day ($903/month) - Gemini 2.5 Flash-Lite: ~$0.05/day ($1.50/month) - DeepSeek Chat: ~$0.14/day ($4.20/month) - Monthly cost difference (Opus vs Flash-Lite): $901.50 --- ## ✨ What You Get ### Model-Agnostic Multi-Provider Testing Pass any model ID — provider is auto-detected from the model name prefix. No hardcoded list; new models work without code changes. | Provider | Example Models | Prefix | Required Key | |----------|---------------|--------|--------------| | **Anthropic** | claude-opus-4-6, claude-sonnet-4-6, claude-haiku-4-5-20251001 | `claude-` | ANTHROPIC_API_KEY | | **OpenAI** | gpt-5.2-pro, gpt-5.2, gpt-5.1 | `gpt-`, `o1`, `o3` | OPENAI_API_KEY | | **Google** | gemini-2.5-pro, gemini-2.5-flash, gemini-2.5-flash-lite | `gemini-` | GOOGLE_API_KEY | | **Mistral** | mistral-large-latest, mistral-small-latest | `mistral-`, `mixtral-` | MISTRAL_API_KEY | | **DeepSeek** | deepseek-chat, deepseek-reasoner | `deepseek-` | DEEPSEEK_API_KEY | | **xAI** | grok-4-1-fast, grok-3-beta | `grok-` | XAI_API_KEY | | **MiniMax** | MiniMax-M2.1 | `MiniMax`, `minimax` | MINIMAX_API_KEY | | **Qwen** | qwen3.5-plus, qwen3-max-instruct | `qwen` | DASHSCOPE_API_KEY | | **Meta Llama** | meta-llama/llama-4-maverick, meta-llama/llama-3.3-70b-instruct | `meta-llama/`, `llama-` | OPENROUTER_API_KEY | ### Known Pricing (per 1M tokens) | Model | Input | Output | |-------|-------|--------| | claude-opus-4-6 | $15.00 | $75.00 | | claude-sonnet-4-6 | $3.00 | $15.00 | | claude-haiku-4-5-20251001 | $1.00 | $5.00 | | gpt-5.2-pro | $21.00 | $168.00 | | gpt-5.2 | $1.75 | $14.00 | | gpt-5.1 | $2.00 | $8.00 | | gemini-2.5-pro | $1.25 | $10.00 | | gemini-2.5-flash | $0.30 | $2.50 | | gemini-2.5-flash-lite | $0.10 | $0.40 | | mistral-large-latest | $2.00 | $6.00 | | mistral-small-latest | $0.10 | $0.30 | | deepseek-chat | $0.27 | $1.10 | | deepseek-reasoner | $0.55 | $2.19 | | grok-4-1-fast | $5.00 | $25.00 | | grok-3-beta | $3.00 | $15.00 | | MiniMax-M2.1 | $0.40 | $1.60 | | qwen3.5-plus | $0.57 | $2.29 | | qwen3-max-instruct | $1.60 | $6.40 | | meta-llama/llama-4-maverick | $0.20 | $0.60 | | meta-llama/l
_meta.json
{
"ownerId": "kn77yjs5esft2kgsd6dpz9c92n80dgsy",
"slug": "prompt-performance-tester",
"version": "1.1.9",
"publishedAt": 1772213259522
}LICENSE.md
# UniAI Skills - Proprietary License
**Version 1.0 | Effective Date: February 2, 2024**
## 1. GRANT OF LICENSE
UniAI ("Licensor") grants you ("Licensee") a limited, non-exclusive, non-transferable, revocable license to use the ClawhHub Skills ("Software") solely in accordance with the terms of this license agreement.
## 2. LICENSE RESTRICTIONS
You may NOT:
- Reverse engineer, decompile, or disassemble the Software
- Modify, alter, or create derivative works of the Software
- Remove, obscure, or alter any proprietary notices or labels on the Software
- Share, distribute, or sublicense the Software to any third party
- Use the Software for commercial purposes without a commercial license
- Access or use the Software beyond the scope of your subscription tier
- Attempt to circumvent licensing controls or API rate limits
## 3. INTELLECTUAL PROPERTY RIGHTS
All intellectual property rights in and to the Software are retained by Licensor. This includes:
- Source code and object code
- Algorithms and methodologies
- Performance optimization techniques
- Quality scoring mechanisms
- Proprietary data structures
- Trade secrets and confidential information
## 4. PERMITTED USES
You may only:
- Use the Software as provided through the ClawhHub platform
- Access features available in your subscription tier
- Create test results and reports for internal use
- Share results with your team (if on a team plan)
- Provide feedback to improve the Software
## 5. SUBSCRIPTION TIERS
### Starter (Free)
- 5 tests per month
- 2 models per test
- Basic features
- Personal use only
### Professional ($29/month)
- Unlimited tests
- All models supported
- Advanced analytics
- API access
- Commercial use permitted
### Enterprise ($99/month)
- Team collaboration
- White-label option
- Custom integrations
- Dedicated support
- SLA guarantees
## 6. API KEY AND CREDENTIALS
- You are responsible for keeping your API keys confidential
- Do not share your license key with others
- One license per person/organization
- License keys are non-transferable
- Unauthorized sharing may result in account termination
## 7. DATA PRIVACY
- We do not retain your test data by default
- Free tier: 30-day retention
- Paid tiers: 90-day retention
- You can request data deletion anytime
- See Privacy Policy for full details
## 8. WARRANTY DISCLAIMER
THE SOFTWARE IS PROVIDED "AS-IS" WITHOUT ANY WARRANTIES. LICENSOR DISCLAIMS ALL WARRANTIES, EXPRESS OR IMPLIED, INCLUDING BUT NOT LIMITED TO:
- Merchantability
- Fitness for a particular purpose
- Non-infringement
- Accuracy of results
## 9. LIMITATION OF LIABILITY
IN NO EVENT SHALL LICENSOR BE LIABLE FOR:
- Any indirect, incidental, special, or consequential damages
- Loss of data, revenue, or profits
- Business interruption
- Even if advised of the possibility of such damages
## 10. TERMINATION
Licensor may terminate your license if you:
- Violate any terms of this agreement
- Fail to pay subscription fees
- Attempt to reverse enginemanifest.yaml
name: "Prompt Performance Tester" id: "prompt-performance-tester" version: "1.1.8" description: "Model-agnostic prompt benchmarking across 9 providers. Pass any model ID from Claude, GPT, Gemini, DeepSeek, Grok, MiniMax, Qwen, Llama, Mistral — provider auto-detected. Measures latency, cost, quality, and consistency." homepage: "https://unisai.vercel.app" repository: "https://github.com/vedantsingh60/prompt-performance-tester" source: "included" intellectual_property: license: "free-to-use" license_file: "LICENSE.md" copyright: "© 2026 UnisAI. All rights reserved." distribution: "via-clawhub-only" source_code_access: "included" modification: "personal-use-only" reverse_engineering: "allowed-for-security-audit" author: company: "UnisAI" contact: "[email protected]" website: "https://unisai.vercel.app" category: "ai-testing" tags: - "prompt-testing" - "performance-analysis" - "cost-optimization" - "multi-llm" - "quality-assurance" - "benchmarking" - "llm-comparison" - "ai-testing" pricing: model: "free" runtime: "local" execution: "python" required_env_vars: - "ANTHROPIC_API_KEY" # Required if testing Claude models - "OPENAI_API_KEY" # Required if testing GPT models - "GOOGLE_API_KEY" # Required if testing Gemini models - "MISTRAL_API_KEY" # Required if testing Mistral models - "DEEPSEEK_API_KEY" # Required if testing DeepSeek models - "XAI_API_KEY" # Required if testing Grok/xAI models - "MINIMAX_API_KEY" # Required if testing MiniMax models - "DASHSCOPE_API_KEY" # Required if testing Qwen/Alibaba models - "OPENROUTER_API_KEY" # Required if testing Llama/OpenRouter models primary_credential: "At least ONE provider API key is required per provider you want to test" dependencies: python: ">=3.9" packages: - "anthropic>=0.40.0" - "openai>=1.60.0" - "google-generativeai>=0.8.0" - "mistralai>=1.3.0" install_all: "pip install anthropic openai google-generativeai mistralai" install_selective: | pip install anthropic # Claude pip install openai # GPT, DeepSeek, xAI, MiniMax, Qwen, Llama (OpenAI-compat) pip install google-generativeai # Gemini pip install mistralai # Mistral note: "Install only the SDKs for the providers you plan to test. DeepSeek, xAI, MiniMax, Qwen, and Llama all use the openai package with a custom base URL." requirements_file: "requirements.txt" security: data_retention: "0 days" data_flow: "prompts-sent-to-chosen-ai-providers" third_party_data_sharing: | WARNING: This skill sends your prompts to whichever AI providers you select for testing. Each provider has their own data retention and privacy policies: - Anthropic: https://www.anthropic.com/legal/privacy - OpenAI: https://openai.com/policies/privacy-policy - Google: https://ai.google.dev/gemini-api/terms - Mistral: https://mistral.ai/terms/ - DeepSeek: https://www.deepseek
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-04-15T00:45:39.800Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-04-15T00:45:39.800Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.9K downloads",
"href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-04-15T00:45:39.800Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.1.9",
"href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-02-27T17:27:39.522Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-vedantsingh60-prompt-performance-tester/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.1.9",
"description": "- Updated provider/model lists to include more example models and the latest pricing. - Expanded cost comparison examples to include DeepSeek Chat. - Added clarification that unlisted models are supported, but cost is shown as $0.00 with a warning. - Improved explanation of quality, cost, and performance metrics for broader clarity. - Enhanced recommendation and real-world example sections to better showcase DeepSeek and new models.",
"href": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceUrl": "https://clawhub.ai/vedantsingh60/prompt-performance-tester",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-02-27T17:27:39.522Z",
"isPublic": true
}
]
}Record generated Oct 10, 2026.
