VHellendoorn/Code-LMs
Guide to using pre-trained large language models of source code
AI4SE Researcher, Assistant Prof. at CMU
Guide to using pre-trained large language models of source code
Data and Code for Reproducing "Global Relational Models of Source Code"
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Unsupervised text tokenizer for Neural Network-based text generation.
[PNAS2021] The neural architecture of language: Integrative modeling converges on predictive processing
Dynamic detection of likely invariants
PLUR (Programming-Language Understanding and Repair) is a collection of source code datasets suitable for graph-based machine learning. We provide scripts for downloading, processing, and loading the datasets. This is done by offering a unified API and data structures for all datasets.
Implementation of Flash Attention in Jax
An implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
The simplest, fastest repository for training/finetuning medium-sized GPTs.
Base repo for homework 3