FreeInference Model Status

Each model is probed with a synthetic request every 20 minutes (Cloudflare cron) · monitoring staging.freeinference.org. Click a model to zoom in on its latency and throughput history.
9 up 1 down 10 models
bge-m3
StatusUP Latency394 ms TTFT Throughput Uptime92% Checked2026-09-12T14:00:58
Latency trend (92)394 ms · 255–893 ms
click to zoom ↗
deepseek-v4-flash
StatusUP Latency2104 ms TTFT566 ms Throughput434.98 tok/s Uptime90% Checked2026-09-12T14:00:56
TTFT trend (71)566 ms · 422–50654 ms
click to zoom ↗
diffusiongemma
StatusUP Latency4640 ms TTFT3652 ms Throughput205.39 tok/s Uptime92% Checked2026-09-12T14:00:56
TTFT trend (73)3652 ms · 3111–20485 ms
click to zoom ↗
glm-5.2
StatusUP Latency16352 ms TTFT3251 ms Throughput78.09 tok/s Uptime84.9% Checked2026-09-12T14:00:33
TTFT trend (62)3251 ms · 2153–22133 ms
click to zoom ↗
glm-5.3
StatusUP Latency16210 ms TTFT3131 ms Throughput78.22 tok/s Uptime84.9% Checked2026-09-12T14:00:33
TTFT trend (62)3131 ms · 2165–21974 ms
click to zoom ↗
glm-5.3-flash
StatusUP Latency21033 ms TTFT4521 ms Throughput61.95 tok/s Uptime81% Checked2026-09-12T14:00:33
TTFT trend (62)4521 ms · 2263–18250 ms
click to zoom ↗
kimi-k2.7-code
StatusUP Latency9898 ms TTFT1535 ms Throughput44.96 tok/s Uptime32.9% Checked2026-09-12T14:00:50
TTFT trend (24)1535 ms · 1515–8042 ms
click to zoom ↗
kimi-k3
StatusDOWN Latency1105 ms TTFT Throughput Uptime6.9% Checked2026-09-12T14:00:54
TTFT trend (5)11025 ms · 1603–11025 ms
click to zoom ↗
stream error: You've reached your concurrent request limit. Please wait for your ongoing requests to finish and try again. (request_id: req_d198822e21db4ccb81c5eb2e6debaa74)
minimax-m3
StatusUP Latency4936 ms TTFT1235 ms Throughput276.41 tok/s Uptime100% Checked2026-09-12T14:00:50
TTFT trend (73)1235 ms · 589–4325 ms
click to zoom ↗
qwen3.6-35b
StatusUP Latency1213 ms TTFT350 ms Throughput281.58 tok/s Uptime92% Checked2026-09-12T14:00:55
TTFT trend (73)350 ms · 330–3115 ms
click to zoom ↗