anthonyhchan/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
MLX: An array framework for Apple silicon
A framework for few-shot evaluation of language models.
Run LLMs with MLX
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning
Notebooks using the Hugging Face libraries 🤗
Standard Open Arm 100