Best AI gateway

Gwarden AI
31 models

31 models for code, agents and any task - via one simple API.
Get your key in Telegram, start with a free 7-day trial.

31 models OpenAI + Anthropic API Claude Code · Codex · opencode · aider Context up to 1M Free start - 7 days
Models

Model lineup

Claude Fable 5.1

Anthropic
API

Incremental update over Fable 5: better instruction following, long-horizon coding and tool use.

$3.50 in$13.00 out1M ctx65.7% Q
Status 100.00%Lat. 8.46sTPS 37.9t/s

Claude Opus 5

Anthropic
API

The most capable Opus generation. Built for hard long-running tasks and deep research.

$2.00 in$10.00 out1M ctx63.1% Q
Status 97.78%Lat. 20.12sTPS 34.9t/s

Claude Fable 5

Anthropic
API

Anthropic's 2026 flagship. The strongest coder in the Claude family, 1M context, vision.

$3.50 in$13.00 out1M ctx62.1% Q
Status 100.00%Lat. 19.43sTPS 19.7t/s

Stealth Warden

Gwarden
API

Our own frontier model on a dedicated channel. Opus-tier answers, 1M context, own balance, never shares capacity with the main gateway.

$0.06 in$0.17 out1M ctx62.0% Q
Status 100.00%Lat. 1.16sTPS 55.8t/s

GPT-5.6 Sol

OpenAI
API

Frontier tier of the GPT-5.6 family. Top-tier reasoning for hard coding and agents.

$3.50 in$18.00 out1M ctx60.9% Q
Status 100.00%Lat. 10.17sTPS 25.1t/s

Grok 4.6

xAI
API

xAI's frontier model. Strong reasoning and coding with real-time knowledge.

$1.70 in$5.00 out1M ctx60.9% Q
Status 100.00%Lat. 5.78sTPS 48.1t/s

Kimi K3

Moonshot AI
API

Moonshot AI's flagship open-weight model. Top-tier coder and agent, 1M context.

$2.60 in$13.00 out1M ctx59.7% Q
Status 91.11%Lat. 9.90sTPS 32.9t/s

GLM 5.3

Z.ai
API

Flagship model by Z.ai. 1M context, always-on reasoning, top tier for code and agent tasks.

$0.15 in$0.50 out1M ctx59.5% Q
Status 95.56%Lat. 41.68sTPS 6.2t/s

Gemini 3.8 Flash

Google
API

Google's fast Gemini: low latency, multimodal, 1M context, great price/quality.

$1.00 in$4.50 out1M ctx58.7% Q
Status 97.78%Lat. 6.68sTPS 46.1t/s

Qwen3.8 Max

Alibaba
API

Alibaba's strongest Qwen. Excellent coding and multilingual coverage, 1M context.

$1.70 in$5.20 out1M ctx58.1% Q
Status 100.00%Lat. 7.31sTPS 49.4t/s

Gwarden AI

Gwarden
API

Gwarden AI - friendly coding assistant with 1M context. Unlocks with any one-time $2 gateway purchase, yours forever.

$0.70 in$1.50 out1M ctx57.9% Q
Status 100.00%Lat. 15.80sTPS 44.1t/s

Qwen3.8 2.4T A95B

Alibaba
API

Open-weight Qwen MoE: 2.4T total, 95B active. Frontier quality at open prices.

$0.80 in$1.70 out1M ctx57.7% Q
Status 100.00%Lat. 5.57sTPS 52.9t/s

GPT-5.6 Terra

OpenAI
API

Balanced mid tier of the GPT-5.6 family: frontier quality at a sane price.

$1.70 in$10.00 out1M ctx56.5% Q
Status 100.00%Lat. 4.22sTPS 36.1t/s

Claude Sonnet 5

Anthropic
API

The balanced workhorse of the Claude family: fast, strong coding, 1M context.

$1.70 in$8.50 out1M ctx56.4% Q
Status 100.00%Lat. 12.92sTPS 16.4t/s

GPT-5.5

OpenAI
API

Previous OpenAI flagship generation, still a reliable production workhorse.

$4.50 in$27.00 out1M ctx55.8% Q
Status 100.00%Lat. 4.05sTPS 42.6t/s

Muse Spark 1.3

Meta
API

Meta's fast contributor-tier model. 1M context for everyday chat, coding and creative writing.

$0.40 in$1.60 out1M ctx55.1% Q
Status 98.89%Lat. 7.80sTPS 44.6t/s

DeepSeek V4 Pro

DeepSeek
API

DeepSeek's flagship V4: reasoning-first coder with elite benchmark scores.

$0.80 in$2.20 out1M ctx54.5% Q
Status 97.78%Lat. 24.77sTPS 24.5t/s

GLM-5.2

Z.ai
API

Previous generation of the GLM flagship. Still a strong coder for its price.

$1.20 in$3.90 out1M ctx53.5% Q
Status 97.78%Lat. 18.97sTPS 33.1t/s

GPT-5.6 Luna

OpenAI
API

Fast lightweight tier of the GPT-5.6 family. Ultra-cheap everyday workhorse.

$0.18 in$1.05 out1M ctx52.1% Q
Status 100.00%Lat. 3.48sTPS 40.5t/s

Qwen3.8 27B

Alibaba
API

Compact open Qwen (27B): fast, cheap and surprisingly capable.

$0.22 in$0.80 out1M ctx52.0% Q
Status 100.00%Lat. 7.41sTPS 55.7t/s

DeepSeek V4 Flash

DeepSeek
API

DeepSeek's fast V4 variant: near-flagship quality, minimal latency and cost.

$0.30 in$0.90 out1M ctx52.0% Q
Status 100.00%Lat. 9.38sTPS 34.8t/s

Gemini 3.1 Pro

Google
API

Google's Pro-tier Gemini: 1M context, complex reasoning and multimodal work.

$1.75 in$10.50 out1M ctx51.2% Q
Status 100.00%Lat. 5.52sTPS 52.0t/s

GPT-OSS 120B

OpenAI
API

OpenAI's open-weight GPT-OSS (120B MoE, ~5B active). Frontier-adjacent quality, open price.

$0.45 in$1.80 out1M ctx50.6% Q
Status 100.00%Lat. 3.66sTPS 52.0t/s

Agnes 2.5 Flash

Sapiens AI
API

Sapiens AI's fast Flash-generation model. Low latency, 1M context, always-on reasoning.

$0.35 in$1.40 out1M ctx49.8% Q
Status 100.00%Lat. 5.87sTPS 49.4t/s

MiniMax M3

MiniMax
API

MiniMax M3: fast long-context model tuned for agents and everyday coding.

$0.27 in$1.05 out1M ctx47.5% Q
Status 100.00%Lat. 4.95sTPS 53.5t/s

Mimo 2.5 Pro

Xiaomi
API

Xiaomi's MiMo Pro: budget-friendly reasoning model with solid coding chops.

$0.60 in$2.20 out1M ctx47.5% Q
Status 98.89%Lat. 5.73sTPS 49.0t/s

DeepSeek R1 0528 Qwen3 8B

DeepSeek
API

DeepSeek's R1 (0528) reasoning distill on Qwen3-8B. Compact, fast, strong at math and logic.

$0.12 in$0.45 out1M ctx44.2% Q
Status 100.00%Lat. 5.41sTPS 66.8t/s

Claude Opus 4.8

Anthropic
API

Previous Opus generation, still a workhorse in production pipelines.

$2.00 in$10.00 out1M ctx
Status 100.00%Lat. 8.67sTPS 50.2t/s

GPT Image 2

OpenAI
API

GPT Image 2: photorealistic image generation. Flat $1.00 per image, no token maths - one call, one picture.

$1.00 / request
ModelInput / 1MOutput / 1MCache / 1MQuality
claude-fable-5-11M / 128k out$3.50$13.00$4.0065.7%
claude-opus-51M / 64k out$2.00$10.00$2.0063.1%
claude-fable-51M / 128k out$3.50$13.00$4.0062.1%
stealth-warden1M / 128k out$0.06$0.17$0.0162.0%
gpt-5.6-sol1M / 128k out$3.50$18.00$0.3560.9%
grok-4.61M / 64k out$1.70$5.00$0.4060.9%
kimi-k31M / 64k out$2.60$13.00$0.5559.7%
glm-5.31M / 128k out$0.15$0.50$0.0359.5%
gemini-3.8-flash1M / 64k out$1.00$4.50$0.2558.7%
qwen3.8-max1M / 64k out$1.70$5.20$0.4058.1%
gwarden-ai1M / 64k out$0.70$1.50$0.0957.9%
qwen3.8-2.4t-a95b1M / 64k out$0.80$1.70$0.1557.7%
gpt-5.6-terra1M / 128k out$1.70$10.00$0.1856.5%
claude-sonnet-51M / 64k out$1.70$8.50$0.1756.4%
Qwen3.8-Flash-Next1M / 64k out$0.15$0.50$0.0356.0%
gpt-5.51M / 128k out$4.50$27.00$0.4555.8%
muse-spark-1.3-contributor1M / 64k out$0.40$1.60$0.0855.1%
mercury-2.51M / 32k out$0.15$0.50$0.0355.0%
deepseek-v4-pro1M / 64k out$0.80$2.20$0.2054.5%
glm-5.21M / 128k out$1.20$3.90$0.2553.5%
gpt-5.6-luna1M / 128k out$0.18$1.05$0.0152.1%
deepseek-v4-flash1M / 32k out$0.30$0.90$0.0852.0%
qwen3.8-27b1M / 32k out$0.22$0.80$0.0552.0%
gemini-3.1-pro1M / 64k out$1.75$10.50$0.4551.2%
gpt-oss-120b1M / 64k out$0.45$1.80$0.0950.6%
agnes-2.5-flash1M / 64k out$0.35$1.40$0.0749.8%
mimo-v2.5-pro1M / 64k out$0.60$2.20$0.1547.5%
minimax-m31M / 64k out$0.27$1.05$0.0847.5%
deepseek-r1-0528-qwen3-8b1M / 32k out$0.12$0.45$0.0344.2%
claude-opus-4-81M / 64k out$2.00$10.00$2.00-
gpt-image-2image generation$1.00 per request-
Gateway's

Faster API Endpoints

Dedicated high-speed URLs with priority routing — no rate limits, no auth hiccups, no delays. Buy once, use forever.

Boost
$2 / 200₽ — one-time

Dedicated high-speed path with priority routing. No rate limits, no auth hiccups, no waiting on key rotation.

https://boost.gwarden.su/v1
Get access
Dev
$5 / 500₽ — one-time

Maximum throughput with redundant failover. Built for agents that cannot afford delays.

https://dev.gwarden.su/v1
Get access
Gwarden AI
$2 / 200₽ — one-time

Unlock the Gwarden AI model: friendly coding assistant, 1M context. One-time $2, yours forever.

"model": "gwarden-ai"
Get access
Pricing

Subscriptions

Free
Free
Up to $6$ credits
Refills every 5 h to $6$
7-day trial
Tester
$4
Up to $5$ credits
Refills every 5 h to $5$
1 mo$4
2 mo$8
3 mo$12
4 mo$16
5 mo$20
6 mo$24
7 mo$28
8 mo$32
9 mo$36
10 mo$40
11 mo$44
1 year$40
1-12 months. A year is priced as 10 months.
Startupper
$5
Up to $6$ credits
Refills every 5 h to $6$
1 mo$5
2 mo$10
3 mo$15
4 mo$20
5 mo$25
6 mo$30
7 mo$35
8 mo$40
9 mo$45
10 mo$50
11 mo$55
1 year$50
1-12 months. A year is priced as 10 months.
Agent
$8
Up to $12$ credits
Refills every 5 h to $12$
1 mo$8
2 mo$16
3 mo$24
4 mo$32
5 mo$40
6 mo$48
7 mo$56
8 mo$64
9 mo$72
10 mo$80
11 mo$88
1 year$80
1-12 months. A year is priced as 10 months.
Orchestrator
$15
Up to $20$ credits
Refills every 5 h to $20$
1 mo$15
2 mo$30
3 mo$45
4 mo$60
5 mo$75
6 mo$90
7 mo$105
8 mo$120
9 mo$135
10 mo$150
11 mo$165
1 year$150
1-12 months. A year is priced as 10 months.
Neuron
$25
Up to $35$ credits
Refills every 5 h to $35$
1 mo$25
2 mo$50
3 mo$75
4 mo$100
5 mo$125
6 mo$150
7 mo$175
8 mo$200
9 mo$225
10 mo$250
11 mo$275
1 year$250
1-12 months. A year is priced as 10 months.
Emperor
$35
Up to $50$ credits
Refills every 5 h to $50$
1 mo$35
2 mo$70
3 mo$105
4 mo$140
5 mo$175
6 mo$210
7 mo$245
8 mo$280
9 mo$315
10 mo$350
11 mo$385
1 year$350
1-12 months. A year is priced as 10 months.
CEO
$70
Up to $80$ credits
Refills every 5 h to $80$
1 mo$70
2 mo$140
3 mo$210
4 mo$280
5 mo$350
6 mo$420
7 mo$490
8 mo$560
9 mo$630
10 mo$700
11 mo$770
1 year$700
1-12 months. A year is priced as 10 months.
Monopolist
$89
Up to $150$ credits
Refills every 5 h to $150$
1 mo$89
2 mo$178
3 mo$267
4 mo$356
5 mo$445
6 mo$534
7 mo$623
8 mo$712
9 mo$801
10 mo$890
11 mo$979
1 year$890
1-12 months. A year is priced as 10 months.
Developer
$120
Unlimited - balance is not limited
No limits
1 mo$120
2 mo$240
3 mo$360
4 mo$480
5 mo$600
6 mo$720
7 mo$840
8 mo$960
9 mo$1080
10 mo$1200
11 mo$1320
1 year$1200
1-12 months. A year is priced as 10 months.

Every 5 hours your balance refills to the plan limit. Start - free 7 days. Per 1M tokens: Input 0.15, Output 0.5, Cache 0.03. Top-up: 1$ = 10.0$.

I can't pay for the subscription - what do I do?

Open the @CryptoBot payment guide
QuickStart

Launch in a minute

1

Get a key

Open the bot, subscribe to the channel, press «Get API».

2

Add config

Add baseURL and apiKey to your tool.

3

Done

Works with opencode, Claude Code and any OpenAI client.

Launch Claude Code in one command

export ANTHROPIC_BASE_URL="https://gwarden.su"
export ANTHROPIC_AUTH_TOKEN="GWAR-XXXX"
export ANTHROPIC_MODEL="glm-5.3"
claude

Your key is in the cabinet. Codex, opencode and aider setups - in the documentation.

Or the raw API

# opencode - check the model
curl https://gwarden.su/v1/chat/completions \
  -H "Authorization: Bearer GWAR-XXXX" \
  -H "Content-Type: application/json" \
  -d '{"model": "MODEL", "messages": [{"role": "user", "content": "Hello"}]}'