Models
ReadyGPU TEE

Meta: Llama 3.3 70B Instruct

Model IDmeta-llama/llama-3.3-70b-instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model is optimized for multilingual dialogue use cases and outperforms many of the available open source and closed chat models on common industry benchmarks.

input

$2.00/M

output

$2.00/M

context

131K

created

Nov 28, 2025

Supported API shape

input

text

output

text

tools

Supported

json mode

Supported

Verification

receipt

x-receipt-id

attestation

gateway report

session

attested upstream

provider

Phala

Provider

Phala

GPU TEE

input

$2.00/M

output

$2.00/M

context

131K

Performance comparison

Last 72h · UTC