EthanCornell/Netgraph-Epochization-for-FreeBSD
Re-engineers the kernel’s Netgraph packet path to be lock-free with epoch-based reclamation, slashing contention and scaling cleanly across modern multi-core CPUs.
Ethan is a dedicated software engineer with a Master's in CS from Cornell University, specializing in algorithms, data structures, and systems optimization.
Re-engineers the kernel’s Netgraph packet path to be lock-free with epoch-based reclamation, slashing contention and scaling cleanly across modern multi-core CPUs.
Free one-click installer to run AiM RaceStudio 3 on Apple Silicon Macs: a notarized, drag-to-Applications DMG. No Windows, Parallels, or CrossOver. Community project, not affiliated with AiM.
A hands-on fork of NanoGPT with FlashAttention-2 CUDA kernels, INT8/INT4 GPTQ quantization, paged KV-cache reuse, and continuous batching, turning a tiny Shakespeare model into a full-speed GPU LLM inference demo.
A comprehensive memory allocation library implementation featuring multiple levels of sophistication, from basic first-fit allocation to security-enhanced allocators with extensive debugging capabilities.
claude-code full original source code from source maps
Gossip protocol implementation in C++
A beautiful, simple, clean, and responsive Jekyll theme for academics
Ultra-fast SIMD library delivering 18.7× speedups through AVX-512 optimization, supporting F32/F16/I8/BF16 data types with 100% accuracy across 175 test cases.
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Github Pages template based upon HTML and Markdown for personal, portfolio-based websites.
CY86 is a simplified, intermediate assembly language created to streamline the translation from high-level programming logic to optimized low-level machine code, enabling faster development, easier debugging, and more efficient performance tuning in compiler and systems projects.
Open deep learning compiler stack for cpu, gpu and specialized accelerators
A retargetable MLIR-based machine learning compiler and runtime toolkit.
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
ClickHouse® is a real-time analytics database management system
A family of header-only, very fast and memory-friendly hashmap and btree containers.
A Standards‑Compliant C/C++ Pre‑Processor Tokeniser with Full Phase 2 & 3 Translation
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
A modern replacement for Redis and Memcached
Header-only C++17 red-black tree with per-node locks—parallel look-ups, serialized writers, and a built-in stress test for heavy-load correctness.
Blazing-fast branch-and-bound TSP solver (≤ 18 cities) in single-file C using the MPI message-passing model; auto-detects triangular inputs and runs locally or via PBS with one command.
A minimal operating system (2K LOC) on QEMU and a RISC-V board
Cache replacement policies in C
Mini-Migration — Cross-platform resumable file-transfer tool. C++17 core, Objective-C++ macOS layer; built for Apple Backup & Migration workflows.
A header-only, hazard-pointer–protected, lock-free queue for C++20
micro-DFS is a minimalist, log-structured distributed file system designed for educational purposes and edge computing scenarios. Built as a reference implementation, it demonstrates core distributed systems concepts with a focus on simplicity, performance, and reliability.
LLM training in simple, raw C/CUDA
Branch-and-Bound Wandering Salesman solver in C with OpenMP, leveraging a shared-memory parallel model for fast, multi-core search.