Yancey0623/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
SGLang is a high-performance serving framework for large language models and multimodal models.
DeepEP: an efficient expert-parallel communication library
Development repository for the Triton language and compiler
Benchmarking Gotorch
atom 快捷键 shortcuts
Yancey's Website
Some algorihtms.
Triton DSL with BladeDISC backend
The Torch-MLIR project aims to provide first class support from the PyTorch ecosystem to the MLIR ecosystem.
🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
Tensors and Dynamic neural networks in Python with strong GPU acceleration
flannel is an etcd backed network fabric for containers
Unified Interface for Constructing and Managing Workflows on different workflow engines, such as Argo Workflows, Tekton Pipelines, and Apache Airflow.
A demo using PyTorch C++ functional API
Kubernetes-native Deep Learning Framework
Redis is an in-memory database that persists on disk. The data model is key-value, but many different kind of values are supported: Strings, Lists, Sets, Sorted Sets, Hashes, HyperLogLogs, Bitmaps.
Deploy SQLFlow service mesh on Windows, macOS, and Linux desktop computers
Elastic Deep Learning using PaddlePaddle and Kubernetes
Couler backend for Tekton. Couler for Argo is at https://github.com/sql-machine-learning/sqlflow/tree/develop/python/couler
hadoop on k8s
A Go package of Jupyter Notebook format
Go driver for Apache Hive
Apache Thrift