Shared
Virtual machines on shared vCPUs. Dev, staging and small production.
- 1 to 32 vCPU, 1 GB to 192 GB RAM
- 25 GB to 3.8 TB NVMe storage
- Full root access, SSH keys
- Snapshots and backups
- Billed by the hour
33 models on one OpenAI-compatible endpoint, one key, up to 1M context.
H100, H200 and B200 accelerators on demand. Reserve a single GPU or a multi-node cluster, provisioned in under 90 seconds and billed by the second, with persistent volumes that follow your workload between sessions.
Multi-node H200, B200 and B300 clusters with NVIDIA NVLink 5.0 fabric, dedicated capacity and committed pricing. From a single 8-GPU node to thousand-GPU training runs; we handle the rest.
33 models, 8 providers, one OpenAI- and Anthropic-compatible endpoint. Switch providers with a string change.
claude-opus-5
Anthropic
Frontier · Tools · 1M context
gpt-6-astra
OpenAI
Frontier · Tools · 1M context
grok-4.6
xAI
Reasoning · Tools · 500k
kimi-k3
Moonshot
Open weights · Tools · 1M
glm-5.3
Zhipu
Open weights · Tools · 1M
deepseek-v4-pro
DeepSeek
Reasoning · Tools · 1M
claude-opus-5
Anthropic
Frontier · Tools · 1M context
gpt-6-astra
OpenAI
Frontier · Tools · 1M context
grok-4.6
xAI
Reasoning · Tools · 500k
kimi-k3
Moonshot
Open weights · Tools · 1M
glm-5.3
Zhipu
Open weights · Tools · 1M
deepseek-v4-pro
DeepSeek
Reasoning · Tools · 1M
claude-sonnet-5
Anthropic
Balanced · Tools · 1M context
gpt-5.6-terra
OpenAI
Balanced · Tools · 1M context
grok-4.5
xAI
Reasoning · Tools · 500k
kimi-k2.7-code
Moonshot
Code · Open weights · 256k
minimax-m3
MiniMax
Open weights · Tools · 1M
glm-5.3-flash-derisked
Zhipu
Derisked · Hosted · 1M
claude-sonnet-5
Anthropic
Balanced · Tools · 1M context
gpt-5.6-terra
OpenAI
Balanced · Tools · 1M context
grok-4.5
xAI
Reasoning · Tools · 500k
kimi-k2.7-code
Moonshot
Code · Open weights · 256k
minimax-m3
MiniMax
Open weights · Tools · 1M
glm-5.3-flash-derisked
Zhipu
Derisked · Hosted · 1M
claude-haiku-4.5
Anthropic
Fast · Tools · 200k
gpt-5.4-mini
OpenAI
Fast · Tools · 400k
gpt-5.3-codex
OpenAI
Code · Tools · 400k
deepseek-v4-flash
DeepSeek
Fast · Open weights · 1M
doubao-seed-2.1-turbo
ByteDance
Fast · Tools · 256k
claude-fable-5.1
Anthropic
Frontier · Tools · 1M context
claude-haiku-4.5
Anthropic
Fast · Tools · 200k
gpt-5.4-mini
OpenAI
Fast · Tools · 400k
gpt-5.3-codex
OpenAI
Code · Tools · 400k
deepseek-v4-flash
DeepSeek
Fast · Open weights · 1M
doubao-seed-2.1-turbo
ByteDance
Fast · Tools · 256k
claude-fable-5.1
Anthropic
Frontier · Tools · 1M context
Three tiers on one control plane. Move between them without rebuilding your stack: same API, same billing, same audit trail, whether you are running a staging box or a regulated workload on its own hardware.
Virtual machines on shared vCPUs. Dev, staging and small production.
Virtual dedicated servers on pinned physical cores. Production APIs and databases.
A whole physical server, no hypervisor. Your kernel, your hardware.