MoonshotAI: Kimi K2.7 Code
moonshotai/kimi-k2.7-codeMoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
input
$0.95/M
output
$4.00/M
context
262K
created
Aug 14, 2026
Supported API shape
input
text · image
output
text
tools
Supported
json mode
Supported
Verification
receipt
x-receipt-id
attestation
gateway report
session
attested upstream
provider
Phala
Provider
Phala
Intel TDX
input
$0.95/M
output
$4.00/M
context
262K
Performance comparison
Last 72h · UTCMore models
Other private inference routes.
Qwen: Qwen3 8B
Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between thinking mode for math, coding, and logical reasoning and non-thinking mode for efficient general-purpose dialogue. Served as a TEE deployment via Phala.
context
41K
input
$0.11/M
Z.ai: GLM 5.3
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window. Served as a text-only TEE deployment via Phala.
context
1.0M
input
$1.40/M
Z.ai: GLM 5.3 Flash
GLM-5.3-Flash is Z.ai's natively multimodal 320B MoE model with 18B active parameters, designed for efficient coding, long-horizon agent tasks, visual understanding, and long-context inference. Served as a TEE deployment via Phala.
context
1.0M
input
$0.15/M
Qwen: Qwen3.8 27B
Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be enabled or disabled. Served on Phala in a TDX-attested enclave.
context
262K
input
$0.30/M