nikg4/relbench
RelBench: Relational Deep Learning Benchmark
RelBench: Relational Deep Learning Benchmark
Relational Transformer: Toward Zero-Shot Foundation Models for Relational Data
An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
GenAI on Google Cloud: Enterprise Generative AI Systems and Agents
Solving i18n for client-side and resource-constrained environments.
Ready-made tokenizer library for working with GPT and tiktoken
💥 Fast State-of-the-Art Tokenizers optimized for Research and Production
Split text into semantic chunks, up to a desired chunk size. Supports calculating length by characters and tokens, and is callable from Rust and Python.
An transformer based LLM. Written completely in Rust
Implementation of a memory efficient multi-head attention as proposed in the paper, "Self-attention Does Not Need O(n²) Memory"
Learn Kubernetes in a Month of Lunches
Helpful tools and examples for working with flex-attention
LLaVA-MORE: Enhancing Visual Instruction Tuning with LLaMA 3.1
Memory Efficient Attention (O(sqrt(n)) for Jax and PyTorch
Context Parallelism, support Blockwise Attention, Ring Attention and Tree Attention.
modified on origin repo
Transformers with Arbitrarily Large Context
Memory optimization and training recipes to extrapolate language models' context length to 1 million tokens, with minimal hardware.
Cambrian-1 is a family of multimodal LLMs with a vision-centric design.
Official repository for LightSeq: Sequence Level Parallelism for Distributed Training of Long Context Transformers
Recipes for shrinking, optimizing, customizing cutting edge vision models. 💜
Ring attention implementation with flash attention
The ALCF hosts a regular simulation, data, and learning workshop to help users scale their applications. This repository contains the examples used in the workshop.
Ongoing research training transformer language models at scale, including: BERT & GPT-2
Tree Attention: Topology-aware Decoding for Long-Context Attention on GPU clusters
Examples in the MLX framework
VILA - a multi-image visual language model with training, inference and evaluation recipe, deployable from cloud to edge (Jetson Orin and laptops)
My Rust deep dive down the rabbit-hole !
Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"