Local AI + vLLM / NVFP4 / Blackwell. Co-author on upstream fixes. Your Highness brand.
-
Your Highness
- Los Angeles
-
16:45
(UTC -12:00) - yourhighnessla.com
- @seanhighness
Pinned Loading
-
vllm-sm12x-nvfp4-dflash2
vllm-sm12x-nvfp4-dflash2 PublicAll-NVFP4 vLLM + DFlash2 K7 for Blackwell (SM120/SM121): NVFP4 target/draft/KV, 262K context, optional CPU vision sidecar.
-
vllm-sm120-nvfp4-mtp
vllm-sm120-nvfp4-mtp PublicTurnkey vLLM v0.27.1 NVFP4 weights + NVFP4 KV cache + MTP-3 for a single RTX 5090
Python 7
-
-
L0xRE-BeeLLama-Low
L0xRE-BeeLLama-Low PublicL0xRE BeeLLama runtime for L0xRE-27b-Low: universal Linux/WSL and Windows packages for RTX 30–50 series; 12 GB VRAM target.
C++ 3
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.



