fairydreaming/lineage-bench
Testing LLM reasoning abilities with lineage relationship quizzes.
Testing LLM reasoning abilities with lineage relationship quizzes.
LLM inference in C/C++
Testing LLM reasoning abilities with family relationship quizzes.
Results of lineage-bench benchmark
Tensor parallelism is all you need. Run LLMs on an AI cluster at home using any device. Distribute the workload, divide RAM usage, and increase inference speed.
Simple Tool Caller for llama.cpp
Python bindings for llama.cpp
NUMAPROF is a NUMA memory profliler based on Pintool to track your remote memory accesses.
A course on aligning smol models.
Performance monitoring and benchmarking suite
Can LLMs calculate XOR?
Simple script to export NeMo formatted models into safetensors.