whbldhwj/DeepSpeed
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Development repository for the Triton language and compiler
Large Language Model Text Generation Inference
A high-throughput and memory-efficient inference and serving engine for LLMs
A schedule language for large model training
Datasets, Transforms and Models specific to Computer Vision
TorchBench is a collection of open source benchmarks used to evaluate PyTorch performance.
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Open deep learning compiler stack for cpu, gpu and specialized accelerators
HeavyDB (formerly OmniSciDB)
🥶Vilio: State-of-the-art VL models in PyTorch & PaddlePaddle
Python wrapper for isl, an integer set library
Polyhedral Parallel Code Generation (source repository: http://repo.or.cz/ppcg.git)
Pluto: An automatic polyhedral parallelizer and locality optimizer
CS259 Project - GATK flow
Deep Pose Estimation implemented using Tensorflow with Custom Architectures for fast inference.
OpenPose: Real-time multi-person keypoint detection library for body, face, and hands estimation
Android version of MobileInsight app
code for SW and PairHMM using shared memory/shuffle instructions
presented by Yuze Chi, Jie Wang, Peipei Zhou
precision color scheme for multiple applications (terminal, vim, etc.) with both dark/light modes