Research @mistralai, random art projects
Repositories
TevenLeScao/local_shadertoy_claude
TevenLeScao/local_shadertoy_codex
TevenLeScao/sentence-transformers
Multilingual Sentence & Image Embeddings with BERT
TevenLeScao/pet
This repository contains the code for "How many data points is a prompt worth?"
TevenLeScao/glucose
GLUCOSE: GeneraLized and COntextualized Story Explanations https://arxiv.org/abs/2009.07758
TevenLeScao/BasicSR
Basic Super-Resolution Toolbox, including SRResNet, SRGAN, ESRGAN, etc.
TevenLeScao/transformers
🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
TevenLeScao/Megatron-DeepSpeed
Ongoing research training transformer language models at scale, including: BERT & GPT-2
TevenLeScao/mup
maximal update parametrization (µP)
TevenLeScao/datasets
🤗 The largest hub of ready-to-use datasets for ML models with fast, easy-to-use and efficient data manipulation tools
TevenLeScao/text-dedup
All-in-one text de-duplication
TevenLeScao/deduplicate-text-datasets
TevenLeScao/epsilon
TevenLeScao/tetraencoder
TevenLeScao/mpww
Code to recreate the MPWW dataset
TevenLeScao/library_of_babel
experiments with deduplication on large datasets
TevenLeScao/gooaq
Question-answers, collected from Google
TevenLeScao/promptsource
Toolkit for collecting and applying templates of prompting instances
TevenLeScao/big-sleep
A simple command line tool for text to image generation, using OpenAI's CLIP and a BigGAN
TevenLeScao/what-time-is-it
Code and dataset for EMNLP 2020 submission
TevenLeScao/calculator
TevenLeScao/generalization
TevenLeScao/blog
fastpages-based website to publish random stuff
TevenLeScao/transformer-xl
TevenLeScao/awd-lstm-lm
LSTM and QRNN Language Model Toolkit for PyTorch
TevenLeScao/bloom
Visualizing deep learning with fractals
TevenLeScao/small_pretrained_lms
Efficient word embeddings for production transfer tasks
TevenLeScao/anode
TevenLeScao/super-resolution
collection of super-resolution models & algorithms