Skip to content

Cost to run

GLM-4.7-Flash

zai-org/glm-4.7-flash

Family
GLM-4
Context
202,752 tokens

GLM-4.7-Flash can be run self-hosted (rent a GPU + run vLLM/TGI) or through a serverless API (pay per token). Live pricing comparisons:

Need to benchmark this model against another? Try the calculator or see where it ranks on the InferenceScore leaderboard.