Models
ReadyIntel TDX

Google: Gemini 2.5 Flash Lite

Model IDgoogle/gemini-2.5-flash-lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.

input

$0.10/M

output

$0.40/M

context

1.0M

created

Jul 23, 2025

Supported API shape

input

file · image · text · audio

output

text

tools

Supported

json mode

Supported

Verification

receipt

x-receipt-id

attestation

gateway report

session

attested upstream

provider

Phala

Provider

Phala

Intel TDX

input

$0.10/M

output

$0.40/M

context

1.0M

Performance comparison

Last 72h · UTC