andrewwhitecdw/zerobrew
A drop-in, 5-20x faster, experimental Homebrew alternative
Andrew White is a Managing Technical Consultant on the AI Factory Team at CDW.
A drop-in, 5-20x faster, experimental Homebrew alternative
cuda-oxide is an experimental Rust-to-CUDA compiler that lets you write (SIMT) GPU kernels in safe(ish), idiomatic Rust. It compiles standard Rust code directly to PTX — no DSLs, no foreign language bindings, just Rust.
NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems.
Accelerated Computer Vision Lab (ACCV-Lab) is a systematic collection of packages with the common goal to facilitate end-to-end efficient training in the ADAS domain, each package offering tools & best practices for a specific aspect/task in this domain.
A toolkit for discovering cluster network topology.
Production-ready Generative AI for local, cloud native, airgap, and edge deployments.
Translation layer that maps any Kubernetes framework's Custom Resource Definitions (CRDs) into a standardized, generic structure.
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.
This is my professional portfolio demonstrating my interests in AI, DevOps, and software development.
NVIDIA Infra Controller - Hardware Lifecycle Management and multitenant networking
Collection of libraries for bare-metal telemetry collection
A Rust Crate for interacting with DTMF Redfish endpoints
NVIDIA's Redfish next generation redfish crate
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
C++ and Python support for the CUDA Quantum programming model for heterogeneous quantum-classical workflows
NVIDIA cuDF for Apache Spark plugin - accelerate Apache Spark with GPUs
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance with lower memory utilization in both training and inference.
CUDA Python: Performance meets Productivity
RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing high performance applications.
The NVIDIA GPU driver container allows the provisioning of the NVIDIA driver through the use of containers.
NVIDIA Fleet Intelligence Agent - Host agent for GPU telemetry collection and attestation
C++ SDK that provides resources for implementing and validating Trusted Computing Solutions on NVIDIA hardware
3rd party dependencies for DALI project
Helpful kernel tutorials, examples and SKILLs for tile-based GPU programming
NVIDIA Switch Infrastructure - Config Manager
Rust SDK and language plugin for the Pulumi Infrastructure as Code Platform (experimental)
Modular and Automated Scene Generation for Policy Learning and Evaluation