Explains how to self-host Tencent’s Hunyuan 3, a 295B MoE reasoning and agent model, using vLLM on Spheron’s GPU cloud.
Need help?
Contact usExplains how to self-host Tencent’s Hunyuan 3, a 295B MoE reasoning and agent model, using vLLM on Spheron’s GPU cloud.