🔥 noonghunna / club-3090 - Community recipes for serving LLMs on RTX 3090/4090/5090 CUD
GitHub热门项目 | Community recipes for serving LLMs on RTX 3090/4090/5090 CUDA gpus. Multi-engine (vLLM, llama.cpp, ik_llama) and model-agnostic. Currently shipping Qwen3.6-27B Qwen3.6 35B Gemma 4 26B Gemma 4 31B configs for 1× and 2× cards. | Stars: 2,048 | 102 stars this week | 语言: Python
本文内容来源于互联网,版权归原作者所有
查看原文