bigPYJ1151/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
Common recipes to run vLLM
This repo hosts code for vLLM CI & Performance Benchmark infrastructure.
TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
Gluten: Plugin to Double Trino's Performance
The official home of the Presto distributed SQL query engine for big data
Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
ClickHouse® is a free analytics DBMS for big data
Collaborative Collection of C++ Best Practices. This online resource is part of Jason Turner's collection of C++ Best Practices resources. See README.md for more information.
reimplementation of muduo for learning
Graduation project about VQA based on Pytorch
pymia - generic and modular code for medical image analysis