DeepSeek V4 Pro
A high-capacity reasoning route for agentic workflows, coding tasks, and production chat traffic through Jatevo.
Browse every playground-ready and request-access model route with provider logos, pricing markers, deployment type, and capability tags in one place.
Find model routes by provider, capability, or deployment style.
A high-capacity reasoning route for agentic workflows, coding tasks, and production chat traffic through Jatevo.
GLM 5.2 is Jatevo's dedicated Z.ai route for long-running software, agent, and tool-use sessions.
Nemotron 3 Ultra is a 550B hybrid MoE model from NVIDIA, optimized for demanding multi-agent AI and complex reasoning tasks.
Kimi K3 provides maximum-effort reasoning and a 1,048,576-token context window through Jatevo API Master.
Kimi K2.7 Code runs on Jatevo Inference for long-context software work, agent execution, and production chat workloads.
A Qwen Max route for text-only enterprise workflows, exposed through Jatevo with scoped key enforcement.
GLM 4.7 served on the existing Cerebras route for low-latency prompt iteration and streaming chat sessions.
Gemma 4 31B runs through the existing Cerebras playground route for fast streaming chat and long completion tests.
Spark Gemma 4 26B A4B is exposed through API Master as a serverless Gemma route for lightweight chat and generation workloads.
Qwen3.6 35B A3B NVFP4 runs on Jatevo's RTX PRO 6000 route for high-memory chat and long-context inference through API Master.