fairydreaming

@fairydreaming · User

GitHub profile ↗ · Compare

34 followers12 repositories

Repositories

fairydreaming/distributed-llama

Tensor parallelism is all you need. Run LLMs on an AI cluster at home using any device. Distribute the workload, divide RAM usage, and increase inference speed.

★ 18C++Forks 0

fairydreaming/numaprof

NUMAPROF is a NUMA memory profliler based on Pintool to track your remote memory accesses.

★ 0Forks 0