Skip to content
#

rtx-4050

Here are 2 public repositories matching this topic...

Language: All
Filter by language

Chinmay's AI Assistant (edge-llm-runtime) is a 100% offline, high-efficiency local inference engine for NVIDIA consumer GPUs featuring group-wise INT4 weight quantization (3.39x memory reduction), register-level fused dequantization GEMM kernels, interactive persona modes, and a publication-grade PDF/charting tools suite.

  • Updated Sep 14, 2026
  • HTML

Add this topic to your repo

To associate your repository with the rtx-4050 topic, visit your repo's landing page and select "manage topics."

Learn more