Sai Enduri

@saienduri · User

GitHub profile ↗ · Compare

AMDSan Jose, California10 followers41 repositories

Repositories

saienduri/hipflex

Flexible GPU fractionalization — run more workloads per GPU.

★ 0RustForks 0

saienduri/pytorch

Tensors and Dynamic neural networks in Python with strong GPU acceleration

★ 0PythonForks 0

saienduri/vgpu.rs

vgpu.rs is the fractional GPU & vgpu-hypervisor implementation written in Rust

★ 0Forks 0

saienduri/tensor-fusion

Tensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.

★ 0GoForks 0

saienduri/hai-sglang

SGLang is a fast serving framework for large language models and vision language models.

★ 0Forks 0

saienduri/iree

A retargetable MLIR-based machine learning compiler and runtime toolkit.

★ 0C++Forks 0

saienduri/AKS-GitHubARC-Setup

Documentation for bringing up an Azure Kubernetes cluster integrated with GitHub Actions Runner Controller for IREE Project

★ 1Forks 3

saienduri/TheRock

The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm

★ 0Forks 0

saienduri/triton

Development repository for the Triton language and compiler

★ 0C++Forks 0

saienduri/ossci-fleet

The goal of the OSSCI Fleet is to provide a central mechanism to enable test automation, batch job scheduling, and developer access to a federated set of limited GPU resources

★ 0Forks 0

saienduri/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0PythonForks 0