AI Key Supported Models

Supported Models for AI Keys

AI Keys is OpenClacky's deeply optimized model service, exclusively available for OpenClacky client use. We provide a proxy layer with smart routing and automatic caching — no third-party service configuration required.

Overview

Series Models Highlights
Flagship Opus 4.7, Opus 4.6, Sonnet 4.6, Sonnet 4.5, Haiku 4.5 Top-tier reasoning, Prompt caching
Turbo V4 Pro, V4 Flash, V4 Flash Vision High cost-efficiency, Auto context caching, vision
Google Gemini 3.1 Pro, Gemini 3.5 Flash, Nano Banana Pro/2 1M long context, native multimodal, image generation

All prices include a 5% service fee — transparent billing, no hidden costs.


Flagship Series

Flagship models deliver the strongest reasoning capabilities, ideal for complex development tasks, deep research, and multi-step agent workflows.

Models & Pricing

Prices are in USD per 1,000 tokens (including 5% service fee).

Model Alias Input Output Cache Read Cache Write
abs-claude-opus-4-8 $0.00525 $0.02625 $0.000525 $0.0065625
abs-claude-opus-4-7 $0.00525 $0.02625 $0.000525 $0.0065625
abs-claude-opus-4-6 $0.00525 $0.02625 $0.000525 $0.0065625
abs-claude-sonnet-4-6 $0.00315 $0.01575 $0.000315 $0.0039375
abs-claude-sonnet-4-5 $0.00315 $0.01575 $0.000315 $0.0039375
abs-claude-haiku-4-5 $0.00105 $0.00525 $0.000105 $0.0013125

Choosing a Model

Use Case Recommended Model
Complex reasoning, research, multi-step agent tasks abs-claude-opus-4-8
General development, code generation, everyday tasks abs-claude-sonnet-4-6
Quick edits, simple queries, cost-sensitive workloads abs-claude-haiku-4-5

Turbo Series

Turbo models offer fast responses at extremely low cost, perfect for high-frequency calls and budget-sensitive workloads.

DeepSeek peak/off-peak pricing (in sync with DeepSeek's official 2026-08-17 adjustment)

Starting August 17, 2026 (00:00 Beijing time), DeepSeek uses peak/off-peak billing: peak hours (09:00-12:00, 14:00-18:00 Beijing time) at standard rates, off-peak at half price. From August 23, 2026, weekends (Saturday & Sunday, Beijing time) are billed entirely at off-peak rates.

AI Key requests are billed by the actual time slot. Combined with our 99% cache hit rate, real-world spend on long conversations and Agent workflows drops by ~50% on top of that.

Models & Pricing

Current Pricing (before Aug 17)

Prices are in USD per 1,000 tokens (including 5% service fee).

Model Alias Input Cache Read Output
dsk-deepseek-v4-pro $0.00045675 $0.000003806 $0.0009135
dsk-deepseek-v4-flash $0.000147 $0.00000294 $0.000294

Peak/Off-peak Pricing (from Aug 17)

Prices are in USD per 1,000 tokens (including 5% service fee). Peak: 01:00-04:00 & 06:00-10:00 UTC; all other hours are off-peak (half price).

Model Alias Tier Input Cache Read Output
dsk-deepseek-v4-pro Peak $0.001386 $0.0000462 $0.004158
dsk-deepseek-v4-pro Off-peak $0.000693 $0.0000231 $0.002079
dsk-deepseek-v4-flash Peak $0.000462 $0.0000147 $0.001386
dsk-deepseek-v4-flash Off-peak $0.000231 $0.00000735 $0.000693
dsk-deepseek-v4-flash-vision-exp Peak $0.000462 $0.0000147 $0.001386
dsk-deepseek-v4-flash-vision-exp Off-peak $0.000231 $0.00000735 $0.000693

dsk-deepseek-v4-flash-vision-exp supports image input; images are converted to tokens per DeepSeek's official rules and billed the same as V4 Flash.

Choosing a Model

Use Case Recommended Model
Complex tasks, deep reasoning dsk-deepseek-v4-pro
Fast, cost-effective responses dsk-deepseek-v4-flash
Image understanding, visual Q&A dsk-deepseek-v4-flash-vision-exp

Google Series

Google's Gemini family is natively multimodal (text / image / audio / video) with 1M-token long context. Text models excel at long-document processing and multimodal understanding; image models handle generation and editing.

🎉 Launch Promo — 20% off: or-gemini-3-5-flash, or-gemini-3-pro-image, and or-gemini-3-1-flash-image are now live. The whole Vertex Gemini lineup is 20% off for a limited time.

The tables below show standard prices (5% service fee included). Actual billing is table price × 0.8. End-of-promo will be announced separately.

Text Models & Pricing

Prices are in USD per 1,000 tokens (including 5% service fee, promo discount not included).

Model Alias Input Output Cache Read
or-gemini-3-1-pro $0.0021 $0.0126 $0.00021
or-gemini-3-5-flash 🔥 $0.001575 $0.00945 $0.0001575

Image Generation Models & Pricing

Prices are in USD per 1,000 tokens (including 5% service fee, promo discount not included).

Image generation uses the standard /v1/images/generations endpoint, but is billed by token — and image tokens are priced differently from text tokens:

  • Input — text and image-understanding tokens (prompt, reference images)
  • Image Output — pixel tokens of the generated image (high rate)
  • Text Output / Thinking — explanatory text and chain-of-thought tokens (low rate, same as text models)
  • Cache Read — input tokens served from context cache
Model Alias Input Image Output Text Output Cache Read
or-gemini-3-pro-image 🔥 $0.0021 $0.126 $0.0126 $0.00021
or-gemini-3-1-flash-image 🔥 $0.000525 $0.063 $0.00315 $0.0001575

One image consumes ~1024–1290 image tokens (depending on resolution). With the promo, Flash Image is ~$0.064 per image and Pro Image is ~$0.128 per image.

Choosing a Model

Use Case Recommended Model
Long-document analysis, complex multimodal reasoning or-gemini-3-1-pro
High-frequency calls, code generation, agent workflows or-gemini-3-5-flash
High-quality image generation and editing or-gemini-3-pro-image
Fast image generation (Nano Banana 2) or-gemini-3-1-flash-image

Why Choose AI Keys

Blazing Fast, Direct Connection

Direct connection to official APIs for fast, reliable responses. No multi-layer proxy forwarding — latency minimized to the absolute lowest.

Official Pricing, Transparent Billing

Same pricing as official APIs, no hidden fees — pay only for what you use. All prices include a 5% service fee, with clear and transparent billing.

Industry-Leading Caching, Massive Savings

Our proxy intelligently caches repeated prompt segments, achieving up to 99% cache hit rates and reducing overall costs by up to 50% compared to similar services.

When your request hits the cache:

  • Cache Read — pay only ~10% of the input price
  • Cache Write — tokens written to cache are billed at ~125% of input price, retained for 5 minutes

Turbo series models feature automatic context caching — cached tokens are automatically billed at a lower rate with no extra steps.


Usage

Create a Key

Generate an API key from your AI Keys dashboard. Keys use the format clacky-xxxx... and support quota limits, expiration dates, and usage tracking.

Configure Your Client

In the OpenClacky client, simply select OpenClacky as the Provider and enter your AI Key. That's it — no need to manually configure Base URL, API Type, or other parameters.

# Select OpenClacky in your client
# No additional configuration needed
provider: openclacky
api_key: clacky-your-key-here

Rate Limits & Quotas

  • Each API key can have optional usage quotas (daily, weekly, or monthly) set from your dashboard
  • The proxy enforces authentication and quota checks on every request
  • For high-volume needs, contact us about custom limits

Need help choosing a model? Visit the FAQ or contact support.