CI

CueInference

Live inference demo

GLM 5.2
Live display0.0 /s
Reply avg0.0 /s
CI

How can I help you today?

Pick a model and send a prompt. Watch tokens stream in real time with live throughput in the header.

CueInference can make mistakes. Live display / reply avg are demo paint speed — not billed upstream rate.