Model Library

Designed for the Age of AI

Search our library of open source models and deploy in seconds

Explore More

All Models

DeepSeek V4 Flash

DeepSeek V4 Flash is a MoE model (284B total, 13B active) with a 164K-token context, tuned for low-latency, cost-efficient inference.

DeepSeek V4 Pro

DeepSeek V4 Pro is a MoE model (1.6T total, 49B active) with a 164K-token context and hybrid compressed sparse attention.

Gemma 4 31B

Gemma 4 31B is a 31B dense multimodal model with a 262K-token context, configurable reasoning, and 140+ language support.

GLM 5.2

GLM 5.2 is an open-weight MoE model (753B total, 40B active) with a 1M-token context, built for coding and long-horizon agent workflows.

Kimi K2.6

Kimi K2.6 is a native multimodal MoE model (1T total, 32B active) with agent swarms scaling to 300 sub-agents and 4,000 coordinated steps.

Kimi K2.7 Code

Kimi K2.7 Code is a coding-focused MoE model (1T total, 32B active) that cuts reasoning tokens 30% while improving agentic coding accuracy.

Kimi K3

Kimi K3 is a MoE model (2.8T total, 104B active) with a 1M-token context, native visual understanding, and always-on reasoning.

MiniMax H3

MiniMax H3 is an omni-modal video model generating 15-second 2K clips with native stereo audio from text, image, video, or audio input.

MiniMax M3

MiniMax M3 is a natively multimodal MoE model (427B total, 26B active) with a 1M-token context and frontier coding performance.

Qwen3.5

Qwen3.5 is a multimodal MoE model (397B total, 17B active) with a 262K-token context and strong reasoning, coding, and vision-language performance.

Qwen3.8

Qwen3.8 is a sparse MoE model (2.4T total, 95B active) and the open-weight flagship of the Qwen family, with a 262K-token context.

Coming soon

Deepseek R1

Strong at multi-step problem solving, math, and structured analysis. Great when you want dependable “think it through” answers at a practical price.

GPT-OSS

Balanced quality across writing, coding, and everyday tasks with a smooth UX feel. Best when you need one model that handles most requests well.

Qwen

Good instruction-following, quick responses, and strong performance in multilingual scenarios. Solid pick for chat experiences and high-throughput workloads.

Deepseek R1

Strong at multi-step problem solving, math, and structured analysis. Great when you want dependable “think it through” answers at a practical price.
MiniMax  M2.5
Kimi K2.5
GLM 5
DeepSeek V3.2
gpt-oss-120b
gpt-oss-20b
Qwen3 Instruct
Qwen3 Thinking
Qwen3 Coder
Qwen3.5
Qwen3 VL Instruct
Qwen3 ASR
Qwen-Image
Qwen-Image-Edit
Flux2
Stable Diffusion 3.5
Hunyuan Image
Z-Image
Wan2.2-I2V
Wan2.2-T2V
Hunyuan Image
Z-Image