ynankani/llama.cpp
LLM inference in C/C++
LLM inference in C/C++
Create nvmo mixed precision recipes
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
Olive: Simplify ML Model Finetuning, Conversion, Quantization, and Optimization for CPUs, GPUs and NPUs.
This repo hosts the source for the DirectX Shader Compiler which is based on LLVM/Clang.