Sultan

@SulRash · User

GitHub profile ↗ · Compare

Research engineer @AdaMLLab at KAUST

KAUST (King Abdullah University of Science & Technology)Saudi Arabia36 followers42 repositories

Repositories

SulRash/Stable-Retro

A fork of gym-retro with additional games, emulators and supported platforms

★ 0C++Forks 0

SulRash/EasyRogue

A simple rogue like game ripped out from my third year project titled "Perfect Information Versus Imperfect Information in Reinforcement Learning". I had made this simple game to benchmark performance, and I hope other people can get some use out of it too!

★ 1PythonForks 0

SulRash/nanotron

Minimalistic large language model 3D-parallelism training

★ 1Forks 0

SulRash/captum

fixing bug with some LLMs that don't generate bos

★ 1PythonForks 0

SulRash/pytorch

Tensors and Dynamic neural networks in Python with strong GPU acceleration

★ 0Forks 0

SulRash/lighteval

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

★ 0PythonForks 0

SulRash/pystk

For RL course students who want to run this on their macbook

★ 0CForks 0

SulRash/transformers

🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.

★ 0PythonForks 0

SulRash/Qwen2-VL

Qwen2-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

★ 1Forks 0

SulRash/envenc

Repository for environment encoder, an attempt at improving reinforcement learning agents' generalisability through learning how to act on universal multimodal embeddings generated by a vision-language model.

★ 3PythonForks 0

SulRash/Cheatsheet

An attempt at improving facial recognition performance through appending a 'cheatsheet' to an image with one positive sample and multiple negatives during training.

★ 6PythonForks 0

SulRash/MedievalSD

Just a quick little project I'm working on for my dungeons & dragons world, creating a stable diffusion model that can reliably create those medieval sketch drawings. Most probably will LoRA it.

★ 1Forks 0

SulRash/minLLMTrain

Minimal yet high performant code for pretraining llms. Attempts to implement some SOTA features. Implements training through: Deepspeed, Megatron-LM, and FSDP. WIP

★ 6PythonForks 0

SulRash/datatrove

Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.

★ 1Forks 0

SulRash/inseq

Interpretability for sequence generation models 🐛 🔍

★ 1PythonForks 0