Jackmin801/even-g2-fact-checker
Even Realities G2 fact-check card prototype
Machines teach me how to learn
Even Realities G2 fact-check card prototype
A high-throughput and memory-efficient inference and serving engine for LLMs
Public repo for HF blog posts
CUDA Templates and Python DSLs for High-Performance Linear Algebra
Official CLI and Python SDK for Prime Intellect - access GPU compute, remote sandboxes, RL environments, and distributed training infrastructure for AI development at scale.
NVIDIA Inference Xfer Library (NIXL)
A Datacenter Scale Distributed Inference Serving Framework
A framework for efficient model inference with omni-modality models
PyTorch memory allocation visualizer
Fast and memory-efficient exact attention
A Native-PyTorch Library for LLM Fine-tuning
PyTorch per step fault tolerance (actively under development)
A PyTorch native platform for training generative AI models
Environments to evaluate agentic capabilities in DeFi workflows
Democratizing Reinforcement Learning for LLMs
Code that implements Page Rank as a vaccination strategy and simulates the disease to measure effectiveness
SGLang is a fast serving framework for large language models and vision language models.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.