openwall/john
John the Ripper jumbo - advanced offline password cracker, which supports hundreds of hash and cipher types, and runs on many operating systems, CPUs, GPUs, and even some FPGAs
2,758 repositories
John the Ripper jumbo - advanced offline password cracker, which supports hundreds of hash and cipher types, and runs on many operating systems, CPUs, GPUs, and even some FPGAs
Fast inference engine for Transformer models
oneAPI Deep Neural Network Library (oneDNN)
A fast, ergonomic and portable tensor library in Nim with a deep learning focus for CPU, GPU and embedded devices via OpenMP, Cuda and OpenCL backends
Kratos Multiphysics (A.K.A Kratos) is a framework for building parallel multi-disciplinary simulation software. Modularity, extensibility and HPC are the main objectives. Kratos has BSD license and is written in C++ with extensive Python interface.
stdgpu: Efficient STL-like Data Structures on the GPU
2021年最新总结,值得推荐的c/c++开源框架与库。持续更新中。
High-performance stateful serverless runtime based on WebAssembly
OptimLib: a lightweight C++ library of numerical optimization methods for nonlinear functions
C++ library for solving large sparse linear systems with algebraic multigrid method
Open Multiplayer, a multiplayer mod fully backwards compatible with SA-MP
Numerical linear algebra software package
Extended Memory Semantics - Persistent shared object memory and parallelism for Node.js and Python
A state-of-the-art multithreading runtime: message-passing based, fast, scalable, ultra-low overhead
A C++ header-only library of statistical distribution functions.
This repository consists for gpu bootcamp material for HPC and AI
muparser is a fast math parser library for C/C++ with (optional) OpenMP support.
Armadillo: fast C++ library for linear algebra (matrix maths) & scientific computing
This is a set of simple programs that can be used to explore the features of a parallel platform.
A GPU benchmark tool for evaluating GPUs and CPUs on mixed operational intensity kernels (CUDA, OpenCL, HIP, SYCL, OpenMP)
fast raw decoding library
Particle-in-cell code for plasma simulation
Portable and vendor neutral framework for parallel programming on heterogeneous platforms.
Abstraction Library for Parallel Kernel Acceleration :llama:
Exascale multiphase flow solver — 2025 Gordon Bell Prize Finalist | 200T grid points on 43K+ GPUs
Python extension language using accelerators
GAP Benchmark Suite
STREAM, for lots of devices written in many programming models
Multi-Threaded FP32 Matrix Multiplication on x86 CPUs
Low-latency NUMA-aware fork-join thread-pool with zero allocations, syscalls, CAS, or false-sharing on the hot path for C, C++, Rust, and Zig 🍴