Live catalog · 18 models

Model catalog

Use genuine GPT, Claude, Gemini, and Xpersona models through one API key, one OpenAI-compatible endpoint, and one place to track usage. Switch models per request without changing your integration.

xpersona / models

18 models

New arrivals

Model catalog

X

Muse Spark 1.3 Contributor

OpenAI

Agentic coding fallback via OpenCode

Context

1,048,576

Max out

131,072

Input / 1M

$0.1

Output / 1M

$0.2

ChatReasoningFunction callingStructured outputs2026-09-02
OpenAI logo

GPT-6

OpenAI

Latest frontier GPT reasoning

Context

272k

Max out

128k

Input / 1M

$5

Output / 1M

$25

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.6

OpenAI

Frontier GPT reasoning

Context

372k

Max out

128k

Input / 1M

$1.5

Output / 1M

$12

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.6 Sol

OpenAI

Agentic GPT coding and deep work

Context

372k

Max out

128k

Input / 1M

$1.5

Output / 1M

$12

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.6 Terra

OpenAI

Balanced GPT intelligence and efficiency

Context

372k

Max out

128k

Input / 1M

$1.5

Output / 1M

$2

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.5

OpenAI

Long-context GPT reasoning

Context

1050k

Max out

128k

Input / 1M

$1.5

Output / 1M

$12

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.4

OpenAI

Reliable general-purpose GPT reasoning

Context

1050k

Max out

128k

Input / 1M

$0.75

Output / 1M

$6

ChatReasoningFunction calling2026-07-28
OpenAI logo

GPT-5.4 Mini

OpenAI

Fast, economical GPT reasoning

Context

272k

Max out

128k

Input / 1M

$0.375

Output / 1M

$4

ChatReasoningFunction calling2026-07-28
Anthropic logo

Claude Opus 4.8

OpenAI

Premium Claude analysis and coding

Context

200k

Max out

128k

Input / 1M

$1.5

Output / 1M

$9.25

ChatFunction calling2026-07-28
Anthropic logo

Claude Sonnet 4.6

OpenAI

Balanced Claude intelligence

Context

200k

Max out

128k

Input / 1M

$0.9

Output / 1M

$5.55

ChatFunction calling2026-07-28
Anthropic logo

Claude Haiku 4.5

OpenAI

Fast Claude responses

Context

200k

Max out

128k

Input / 1M

$0.6

Output / 1M

$3.7

ChatFunction calling2026-07-28
Google logo

Gemini 3.5 Flash

OpenAI

Fast multimodal Gemini reasoning

Context

1M

Max out

128k

Input / 1M

$1.55

Output / 1M

$12.2

ChatReasoningFunction calling2026-07-28
Anthropic logo

Claude Opus Latest

OpenAI

Stable alias for genuine Claude Opus 4.8

Context

200k

Max out

128k

Input / 1M

$1.5

Output / 1M

$9.25

ChatFunction callingImage inputPrompt caching2026-07-24
Anthropic logo

Claude Sonnet Latest

OpenAI

Stable alias for genuine Claude Sonnet 4.6

Context

200k

Max out

128k

Input / 1M

$0.9

Output / 1M

$5.55

ChatFunction callingImage inputPrompt caching2026-07-24
OpenAI logo

OpenAI GPT Latest

OpenAI

Stable alias for genuine GPT-5.6

Context

372k

Max out

128k

Input / 1M

$1.5

Output / 1M

$12

ChatReasoningFunction callingPrompt caching2026-07-24
Anthropic logo

Claude Fable 5

OpenAI

Genuine Claude Fable 5 through Xpersona

Context

1M

Max out

128k

Input / 1M

$3

Output / 1M

$18.5

ChatFunction callingImage inputPrompt caching2026-07-05
X

Xpersona A_BETTER_YOU

OpenAI

Extra-high reasoning for ambitious OpenCode work

Context

1M

Max out

128k

Input / 1M

$3

Output / 1M

$18

ChatReasoningFunction callingImage input2026-05-29
X

Xpersona Frieren 1

OpenAI

Coding Agent LLM for Windows/MacOS

Context

1M

Max out

384k

Input / 1M

$1.5

Output / 1M

$6

ChatReasoningFunction callingImage input2026-05-01

Code examples

from openai import OpenAI

client = OpenAI(
    base_url="https://www.xpersona.co/v1",
    api_key="YOUR_XPERSONA_API_KEY",
)

response = client.chat.completions.create(
    model="gpt-5.6-sol",
    messages=[{"role": "user", "content": "Hello!"}],
)

print(response.choices[0].message.content)

Uptime and status

Inference API

Live operational history and incident reporting

View status
Public model endpoint

Machine-readable model discovery for connected clients

Open endpoint

One key. Every model.

Start with a monthly package or top up prepaid credits from $2.

Explore access

FAQ

Frequently asked questions

How do I use a model through Xpersona?

Use the Xpersona base URL and API key with any OpenAI-compatible client, then pass the model ID shown in the catalog.

Are the GPT, Claude, and Gemini models genuine?

Yes. Catalog entries labeled as upstream models are routed to their genuine upstream model through Xpersona.

Can I change models without changing my integration?

Yes. Keep the same endpoint and key, and change only the model value in your request.

Where can I see usage and billing?

Your dashboard contains API key management, request usage, and billing information in one place.

Comparing models? Run the same task across several catalog models, blind-judge the evidence, and continue with the winner.

Explore Council →
Models - GPT, Claude, Gemini, and Xpersona | Xpersona