Popular repositories Loading
-
vLLM-2080Ti-Definitive
vLLM-2080Ti-Definitive PublicForked from weicj/vLLM-2080Ti-Definitive
The definitive vLLM runtime for dual RTX 2080 Ti 22GB + NVLink, delivering Qwen 27B local inference with maximum 100+ tok/s single-request decode with support of FP8 weight ( Join Discord :https://…
Python
-
FluxDown
FluxDown PublicForked from zerx-lab/FluxDown
Rust 驱动的多协议下载管理器,支持 HTTP/FTP/BitTorrent 磁力链接及 HLS/DASH 流媒体,智能多线程加速与浏览器无缝集成。精美界面,极致性能,永久免费,零广告。
Rust
-
FreeToken
FreeToken PublicForked from FlashML-org/FreeToken
FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently.
Python
If the problem persists, check the GitHub status page or contact support.