Inference API
Pay per token across all platform models. OpenAI-compatible endpoints, multi-region routing.
Language
| Model | Input | Output | Notes |
|---|---|---|---|
DeepSeek-V4-Flash | $0.16/1M | $0.33/1M | — |
GLM-5-Turbo | $0.83/1M | $3.67/1M | — |
GLM-5.2 | $1.00/1M | $4.00/1M | — |
gpt-oss-20b | $0.20/1M | $0.28/1M | Context: 125K |
Llama-3.1-8B-Instruct | $0.10/1M | $0.10/1M | Context: 125K |
MiniMax-M2.7 | $0.35/1M | $1.40/1M | — |
Qwen2.5-7B-Instruct | $0.20/1M | $0.20/1M | Context: 16K |
Qwen3-235B-A22B | $0.33/1M | $1.33/1M | — |
qwen3-coder-30b-a3b-instruct | $0.10/1M | $0.30/1M | Context: 32K |
Vision
| Model | Input | Output | Notes |
|---|---|---|---|
Gemma-4-31B-IT | $0.50/1M | $0.50/1M | Context: 32K |
Kimi-K2.6 | $1.08/1M | $4.50/1M | — |
qwen3-omni-30b-a3b-instruct | $0.40/1M | $0.80/1M | — |
qwen3-vl-8b-instruct | $0.15/1M | $0.50/1M | Context: 32K |
Qwen3.5-27B | $0.30/1M | $0.60/1M | Context: 32K |
Qwen3.5-35B-A3B | $0.40/1M | $0.40/1M | — |
Qwen3.6-27B | $0.30/1M | $0.60/1M | Context: 32K |
Qwen3.6-35B-A3B | $0.40/1M | $0.40/1M | Context: 32K |
Embedding
| Model | Input | Output | Notes |
|---|---|---|---|
Jina-Embeddings-V3 | $0.02/1M | — | — |
Jina-Embeddings-V4 | $0.12/1M | — | — |
qwen3-embedding-0.6b | $0.12/1M | — | — |
qwen3-embedding-4b | $0.12/1M | — | — |
qwen3-embedding-8b | $0.12/1M | — | — |
Reranker
| Model | Input | Output | Notes |
|---|---|---|---|
BGE-Reranker-V2-M3 | $0.03/1M | — | — |
Image
| Model | Input | Output | Notes |
|---|---|---|---|
FLUX.2 Klein flux2-klein | — | $20.00/1M out tok | Tokens / item: 1,000 |
Qwen-Image | — | $30.00/1M out tok | Tokens / item: 1,000 |
Z-Image-Turbo | — | $10.00/1M out tok | Tokens / item: 1,000 |
Speech
| Model | Input | Output | Notes |
|---|---|---|---|
Chatterbox | — | $2.00/1M out tok | Tokens / sec audio: 100 |
Fun-ASR-Nano | $0.05/1M in tok | — | Tokens / sec audio: 1,000 |
Kokoro-82M | — | $1.00/1M out tok | Tokens / sec audio: 100 |
Qwen3-ASR-1.7B qwen3-asr-1-7b | $0.05/1M in tok | — | Tokens / sec audio: 1,000 |
Qwen3-TTS | — | $2.00/1M out tok | Tokens / sec audio: 100 |
Whisper-Large-V3-Turbo | $0.10/1M in tok | — | Tokens / sec audio: 1,000 |
Video Generation
Billed on actual tokens used, at a rate set by output resolution and whether you supply a reference video. Higher resolutions bill at a lower per-token rate but consume substantially more tokens per second of video, so a 4K clip still costs more than the same clip at 720p.
EcoLink-Video-Gen-2.0
ecolink-video-gen-2.0
| Resolution | Text / image input | With reference video |
|---|---|---|
| 480p / 720p | $7.14/1M | $4.39/1M |
| 1080p | $7.85/1M | $4.79/1M |
| 4K | $4.08/1M | $2.45/1M |
GPU Compute
Pay per GPU-hour. Same rate for dedicated inference, workspace instances, and clusters. Billed per second of running time.
| GPU | VRAM | On-Demand |
|---|---|---|
| RTX Pro 6000 | 96 GB | $1.98/GPU-hr |
Storage
Persistent storage attached to your GPU instances and clusters. Cloud Drives are single-instance (ReadWriteOnce); Shared Filesystems mount across multiple instances (ReadWriteMany). Billed per second of attached time.
| Type | Size | Monthly |
|---|---|---|
| Cloud Drive | 50 GB | $5.00/mo |
| Cloud Drive | 100 GB | $10.00/mo |
| Cloud Drive | 200 GB | $20.00/mo |
| Cloud Drive | 300 GB | $30.00/mo |
| Cloud Drive | 400 GB | $40.00/mo |
| Cloud Drive | 500 GB | $50.00/mo |
| Shared Filesystem | 50 GB | $5.00/mo |
| Shared Filesystem | 100 GB | $10.00/mo |
| Shared Filesystem | 200 GB | $20.00/mo |
| Shared Filesystem | 300 GB | $30.00/mo |
| Shared Filesystem | 400 GB | $40.00/mo |
| Shared Filesystem | 500 GB | $50.00/mo |