7 followers8 repositories
Repositories
TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
★ 0PythonForks 0
ONNX-TensorRT: TensorRT backend for ONNX
★ 1Forks 0
TensorRT is a C++ library for high performance inference on NVIDIA GPUs and deep learning accelerators.
★ 0C++Forks 0
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
★ 0Forks 0
TensorRT Plugin Autogen Tool
★ 0PythonForks 0
An Open Source Machine Learning Framework for Everyone
★ 0C++Forks 0
TensorFlow/TensorRT integration
★ 0Jupyter NotebookForks 0
PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
★ 0Jupyter NotebookForks 0