Pinned Loading
-
qwen3.8-27b-nvfp4-5090-laptop-24gb-optimal
qwen3.8-27b-nvfp4-5090-laptop-24gb-optimal PublicQwen3.8-27B TWIN-TURBO (NVFP4 GGUF, all-Q8_0 precision chain) tuned to its limits on RTX 5090 Laptop 24GB: 74-78 tok/s with built-in MTP + CUDA graphs, overthinking -93%, vision 3.9s, 160K stable c…
Python 13
-
qwen3.8-flash-next-strata-5090-laptop-24gb
qwen3.8-flash-next-strata-5090-laptop-24gb PublicQwen3.8-Flash-Next 177B MoE on RTX 5090 Laptop (24GB VRAM + 64GB RAM): llama.cpp 22-25 tok/s -> Strata 93.5 tok/s (3.7x), GPU+CPU both saturated. 7 quant tiers screened over 26 rounds, MTP corrupti…
Python 14
-
comfyui-5090-laptop-minimax-h3-wan2.2-nvfp4
comfyui-5090-laptop-minimax-h3-wan2.2-nvfp4 PublicRTX 5090 Laptop (24GB) + ComfyUI: 12 model lines / 21 measured configs — MiniMax H3, Wan 2.2, Qwen-Image 2512, HunyuanVideo 1.5, FLUX.2 Klein, Z-Image-Turbo, LTX-Video, ACE-Step 1.5, YuE2, Stable A…
HTML 10
If the problem persists, check the GitHub status page or contact support.