-
AI ML Infrastrucutre Engineering
- New Delhi
- https://portfolio-webapp-roan.vercel.app
- in/nareshns2004
- @naresh_swe24
- @nareshns2004
- https://substack.com/@nareshns2004
Pinned Loading
-
ai-nic-performance-profiler
ai-nic-performance-profiler PublicAn observability primitive that closes the attribution gap between NIC hardware counters and distributed training throughput degradation enabling data-driven decisions on fabric topology, RDMA tuni…
Python
-
kernel-performance-toolkit
kernel-performance-toolkit PublicA Linux kernel performance analysis toolkit for profiling CPU scheduling, memory behavior, NUMA locality, cache efficiency, page faults and Huge Pages
Python
-
distributed-training-framework-nccl
distributed-training-framework-nccl PublicMini Distributed Training Framework using NCCL
C++
-
custom-cuda-fused-attention-triton
custom-cuda-fused-attention-triton PublicBuilding high-performance GPU kernels from first principles by progressively implementing and optimizing deep learning operators in CUDA and Triton
Python
-
-
fault-tolerant-training-orchestrator
fault-tolerant-training-orchestrator PublicA cross-layer reliability system for multi-node LLM training
Python
If the problem persists, check the GitHub status page or contact support.






