Gwarden AI
28 models
28 models for code, agents and any task - via one simple API.
Get your key in Telegram, start with a free 7-day trial.
Model lineup

Claude Fable 5.1
AnthropicIncremental update over Fable 5: better instruction following, long-horizon coding and tool use.

Claude Opus 5
AnthropicThe most capable Opus generation. Built for hard long-running tasks and deep research.

Claude Fable 5
AnthropicAnthropic's 2026 flagship. The strongest coder in the Claude family, 200K context, vision.

GPT-5.6 Sol
OpenAIFrontier tier of the GPT-5.6 family. Top-tier reasoning for hard coding and agents.
Grok 4.6
xAIxAI's frontier model. Strong reasoning and coding with real-time knowledge.

Kimi K3
Moonshot AIMoonshot AI's flagship open-weight model. Top-tier coder and agent, 1M context.

GLM 5.3
Z.aiFlagship model by Z.ai. 200K context, always-on reasoning, top tier for code and agent tasks.

Gemini 3.8 Flash
GoogleGoogle's fast Gemini: low latency, multimodal, 200K context, great price/quality.

Qwen3.8 Max
AlibabaAlibaba's strongest Qwen. Excellent coding and multilingual coverage, 200K context.

Qwen3.8 2.4T A95B
AlibabaOpen-weight Qwen MoE: 2.4T total, 95B active. Frontier quality at open prices.

GPT-5.6 Terra
OpenAIBalanced mid tier of the GPT-5.6 family: frontier quality at a sane price.

Claude Sonnet 5
AnthropicThe balanced workhorse of the Claude family: fast, strong coding, 200K context.

GPT-5.5
OpenAIPrevious OpenAI flagship generation, still a reliable production workhorse.

Muse Spark 1.3
MetaMeta's fast contributor-tier model. 1M context for everyday chat, coding and creative writing.

DeepSeek V4 Pro
DeepSeekDeepSeek's flagship V4: reasoning-first coder with elite benchmark scores.

GLM-5.2
Z.aiPrevious generation of the GLM flagship. Still a strong coder for its price.

GPT-5.6 Luna
OpenAIFast lightweight tier of the GPT-5.6 family. Ultra-cheap everyday workhorse.

Qwen3.8 27B
AlibabaCompact open Qwen (27B): fast, cheap and surprisingly capable.

DeepSeek V4 Flash
DeepSeekDeepSeek's fast V4 variant: near-flagship quality, minimal latency and cost.

Gemini 3.1 Pro
GoogleGoogle's Pro-tier Gemini: 200K context, complex reasoning and multimodal work.

GPT-OSS 120B
OpenAIOpenAI's open-weight GPT-OSS (120B MoE, ~5B active). Frontier-adjacent quality, open price.

Agnes 2.5 Flash
Sapiens AISapiens AI's fast Flash-generation model. Low latency, 1M context, always-on reasoning.

MiniMax M3
MiniMaxMiniMax M3: fast long-context model tuned for agents and everyday coding.

MiMo V2.5 Pro
XiaomiXiaomi's MiMo Pro: budget-friendly reasoning model with solid coding chops.

DeepSeek R1 0528 Qwen3 8B
DeepSeekDeepSeek's R1 (0528) reasoning distill on Qwen3-8B. Compact, fast, strong at math and logic.

Claude Opus 4.8
AnthropicPrevious Opus generation, still a workhorse in production pipelines.
| Model | Input / 1M | Output / 1M | Cache / 1M | Quality |
|---|---|---|---|---|
| claude-fable-5-1200k / 128k out | $3.50 | $13.00 | $4.00 | 65.7% |
| claude-opus-5200k / 64k out | $2.00 | $10.00 | $2.00 | 63.1% |
| claude-fable-5200k / 128k out | $3.50 | $13.00 | $4.00 | 62.1% |
| gpt-5.6-sol200k / 128k out | $3.50 | $18.00 | $0.35 | 60.9% |
| grok-4.6200k / 64k out | $1.70 | $5.00 | $0.40 | 60.9% |
| kimi-k31M / 64k out | $2.60 | $13.00 | $0.55 | 59.7% |
| glm-5.3200k / 128k out | $0.15 | $0.50 | $0.03 | 59.5% |
| gemini-3.8-flash200k / 64k out | $1.00 | $4.50 | $0.25 | 58.7% |
| qwen3.8-max200k / 64k out | $1.70 | $5.20 | $0.40 | 58.1% |
| qwen3.8-2.4t-a95b200k / 64k out | $0.80 | $1.70 | $0.15 | 57.7% |
| gpt-5.6-terra200k / 128k out | $1.70 | $10.00 | $0.18 | 56.5% |
| claude-sonnet-5200k / 64k out | $1.70 | $8.50 | $0.17 | 56.4% |
| Qwen3.8-Flash-Next200k / 64k out | $0.15 | $0.50 | $0.03 | 56.0% |
| gpt-5.5200k / 128k out | $4.50 | $27.00 | $0.45 | 55.8% |
| muse-spark-1.3-contributor1M / 64k out | $0.40 | $1.60 | $0.08 | 55.1% |
| mercury-2.5200k / 32k out | $0.15 | $0.50 | $0.03 | 55.0% |
| deepseek-v4-pro200k / 64k out | $0.80 | $2.20 | $0.20 | 54.5% |
| glm-5.2200k / 128k out | $1.20 | $3.90 | $0.25 | 53.5% |
| gpt-5.6-luna200k / 128k out | $0.18 | $1.05 | $0.01 | 52.1% |
| deepseek-v4-flash200k / 32k out | $0.30 | $0.90 | $0.08 | 52.0% |
| qwen3.8-27b200k / 32k out | $0.22 | $0.80 | $0.05 | 52.0% |
| gemini-3.1-pro200k / 64k out | $1.75 | $10.50 | $0.45 | 51.2% |
| gpt-oss-120b1M / 64k out | $0.45 | $1.80 | $0.09 | 50.6% |
| agnes-2.5-flash1M / 64k out | $0.35 | $1.40 | $0.07 | 49.8% |
| mimo-v2.5-pro200k / 64k out | $0.60 | $2.20 | $0.15 | 47.5% |
| minimax-m3200k / 64k out | $0.27 | $1.05 | $0.08 | 47.5% |
| deepseek-r1-0528-qwen3-8b1M / 32k out | $0.12 | $0.45 | $0.03 | 44.2% |
| claude-opus-4-8200k / 64k out | $2.00 | $10.00 | $2.00 | - |
Faster API Endpoints
Dedicated high-speed URLs with priority routing — no rate limits, no auth hiccups, no delays. Buy once, use forever.
Dedicated high-speed path with priority routing. No rate limits, no auth hiccups, no waiting on key rotation.
https://boost.gwarden.su/v1Maximum throughput with redundant failover. Built for agents that cannot afford delays.
https://dev.gwarden.su/v1Subscriptions
Every 5 hours your balance refills to the plan limit. Start - free 7 days. Per 1M tokens: Input 0.15, Output 0.5, Cache 0.03. Top-up: 1$ = 10.0$.
I can't pay for the subscription - what do I do?
Open the @CryptoBot payment guideLaunch in a minute
Get a key
Open the bot, subscribe to the channel, press «Get API».
Add config
Add baseURL and apiKey to your tool.
Done
Works with opencode, Claude Code and any OpenAI client.
Launch Claude Code in one command
export ANTHROPIC_BASE_URL="https://gwarden.su" export ANTHROPIC_AUTH_TOKEN="GWAR-XXXX" export ANTHROPIC_MODEL="glm-5.3" claude
Your key is in the cabinet. Codex, opencode and aider setups - in the documentation.
Or the raw API
# opencode - check the model
curl https://gwarden.su/v1/chat/completions \
-H "Authorization: Bearer GWAR-XXXX" \
-H "Content-Type: application/json" \
-d '{"model": "MODEL", "messages": [{"role": "user", "content": "Hello"}]}'