rtx-5090
Here are 96 public repositories matching this topic...
Cross-platform installer for Triton and SageAttention on ComfyUI. Simplifies GPU-accelerated inference setup for Windows users with automated dependency management and RTX 5090 support.
-
Updated
Aug 11, 2026 - Python
Tiny local text-to-24x24 pixel art model, trained on roughly 200K samples in 30 minutes on an RTX 5090.
-
Updated
Jul 20, 2026 - JavaScript
Fastest MoE/LLM inference runtime for consumer and edge Blackwell GPUs. SN74 on Gittensor.
-
Updated
Sep 12, 2026 - C++
An LLM server for a single RTX 5090, built for agent workloads: tool calls, long conversations, reasoning, and many requests at once. One of the fastest engines on this card, at batch 1 and at dozens of concurrent streams, with the numbers in the repo.
-
Updated
Sep 13, 2026 - Cuda
RTX 5090 & RTX 5060 Docker container with PyTorch + TensorFlow. First fully-tested Blackwell GPU support for ML/AI. CUDA 12.8, Python 3.11, Ubuntu 24.04. Works with RTX 50-series (5090/5080/5070/5060) and RTX 40-series.
-
Updated
Jul 8, 2025 - Shell
异环(Neverness To Everness / Ananta)光线追踪一键部署面板,基于 OptiScaler winmm 方案,默认推荐 RTX 5090,并支持本机/RTX 4090/RTX 5080M 配置、备份、恢复和本地 WebUI。
-
Updated
May 18, 2026 - Python
Pixal3D ComfyUI integration for Windows (RTX 30/40/50) — single image to textured PBR mesh in 3-5 min
-
Updated
May 14, 2026 - Python
NVFP4 inference on Blackwell GeForce (RTX 5090/5080/5070 Ti/RTX PRO 6000) — SM120 patches for vLLM + FlashInfer + CUTLASS. 175 tok/s on Qwen3.6-35B MoE.
-
Updated
Apr 27, 2026 - Python
Research: vGPU unlock on consumer NVIDIA RTX 5090 (Blackwell/GB202). 19 binary patches, full CPU-side pipeline working, GSP firmware blocked by fused-off VF PRIV registers.
-
Updated
Apr 1, 2026 - C
Durable local inference for Oh My Pi: NInfer + Qwen3.8 27B on one RTX 3090/4090/5090, with restart-resumable OpenAI Responses state.
-
Updated
Sep 13, 2026 - Python
Windows prebuilt of llama.cpp combining Multi-Token Prediction (MTP) + TurboQuant KV cache compression + native sm_120 (Blackwell consumer GPU, FP4 tensor cores). For RTX 5060 Ti / 5070 / 5080 / 5090.
-
Updated
Jun 5, 2026
Dashboard for AI Studio, Open Source Continuous Inference | Deepseek-R1, Qwen2.5, Llama3.1 | 4xRTX-5090 inside PRU2500, 2xH100 inside PRU2500, 8xMI210 in SuperMicro
-
Updated
Aug 7, 2026 - TypeScript
Qwen3.8-27B on RTX 5090s — 262K ctx, 1.4M-token KV pool, ~220 t/s code decode. NVFP4 + vLLM + sm120 patches, reproducible.
-
Updated
Sep 11, 2026 - Shell
LLM benchmarking, GPU workload orchestration backend server | Deepseek-R1, Qwen2.5, Llama3.1 | 4xRTX-5090 inside PRU2500, 2xH100 inside PRU2500, 8xMI210 in SuperMicro
-
Updated
Aug 18, 2026 - Python
Per-pin 12VHPWR monitoring for ASUS ROG Astral cards as standard Linux hwmon sensors, plus astral-guard
-
Updated
Sep 3, 2026 - C
Fork of NVIDIA MinkowskiEngine modernized for CUDA 12.8+/Blackwell (RTX 50-series), NumPy 2.0, PyTorch 2.x prebuilt wheels, fp16/bf16 autocast, fused gather/scatter. Best-effort, no support commitment.
-
Updated
Jul 17, 2026 - Python
Fish Audio OpenAudio S2-Pro on vLLM-Omni. low-latency ~100ms TTFA, OpenAI-compatible, runs on NVIDIA Blackwell (RTX 5090 / RTX PRO 6000). Self-hosted streaming TTS & voice cloning.
-
Updated
Jun 27, 2026 - Python
Add this topic to your repo
To associate your repository with the rtx-5090 topic, visit your repo's landing page and select "manage topics."