Skip to content

Cost to run

VILA 1.5 40B

nvidia/vila-1.5-40b

Family
VILA
Context
8,192 tokens

VILA 1.5 40B can be run self-hosted (rent a GPU + run vLLM/TGI) or through a serverless API (pay per token). Live pricing comparisons:

Need to benchmark this model against another? Try the calculator or see where it ranks on the InferenceScore leaderboard.