gfx1100
Here are 18 public repositories matching this topic...
Native HIP/C++ Z-Image-Turbo inference optimized for AMD Radeon RX 7900 XTX gfx1100 WMMA.
-
Updated
Aug 5, 2026 - C++
Multi-GPU tensor-parallel vLLM on AMD Radeon RX 7900 XT / XTX / GRE (7900XT, 7900XTX, RX7900XT, gfx1100, RDNA3, ROCm): root cause and fix for the RCCL hostcall / PCIe atomics (AtomicOps) crash "NCCL error: unhandled cuda error" / "operation cannot be performed in the present state", Proxmox VFIO passthrough; LLM inference benchmarks on 13 machines
-
Updated
Sep 17, 2026 - Python
High-performance FP32 GEMV for AMD RDNA 3 (gfx1100), with reproducible performance and numerical analysis.
-
Updated
Aug 31, 2026 - C++
Early Compatibility for RDNA 3 cards for ML/HIP
-
Updated
Jul 26, 2024 - Dockerfile
Measurement notes and tools from a local LLM and diffusion stack on consumer AMD hardware. Pre-registered noise bands, deterministic evaluation, and null results reported as plainly as positive ones.
-
Updated
Sep 23, 2026 - Python
Run SenseNova-U1.5-8B-MoT (unified multimodal understanding + image generation) on AMD RDNA3 GPUs via ROCm — evidence-first, one command, receipts for every claim
-
Updated
Sep 1, 2026 - Python
⚡ Automated nightly builds & portable releases for lucebox (DFlash) inference server on NVIDIA CUDA (sm_75-120) and AMD ROCm 7 (Strix Halo gfx1151, RX 7900 XTX gfx1100, R9700 gfx1201) with zero-dependency $ORIGIN RPATH bundling.
-
Updated
Sep 20, 2026 - Python
ROCm/HIP port of cuPDLP-C for AMD Radeon 890M / gfx1150 with cross-device CUDA/ROCm benchmarks.
-
Updated
Sep 18, 2026 - Python
Accelerate Z-Image-Turbo inference on AMD Radeon RX 7900 XTX with custom HIP/C++ optimizations for maximum performance.
-
Updated
Sep 25, 2026 - C++
Add this topic to your repo
To associate your repository with the gfx1100 topic, visit your repo's landing page and select "manage topics."