Phlip79/Megatron-LM
Ongoing research training transformer models at scale
Megatron and NeMo @NVIDIA
Ongoing research training transformer models at scale
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
AI agent skills published by NVIDIA
Jetson GPIO Python Library
KAI Scheduler is an open source Kubernetes Native scheduler for AI workloads at large scale
Final Project for CMU 15-618
Deep Learning (Python, C, C++, Java, Scala, Go)
Protocol Buffers - Google's data interchange format