Mihaiii/semantic-autocomplete
A blazing-fast semantic search React component. Match by meaning, not just by letters. Search as you type without waiting (no debounce needed). Rank by cosine similarity.
A blazing-fast semantic search React component. Match by meaning, not just by letters. Search as you type without waiting (no debounce needed). Rank by cosine similarity.
LLaMA-Omni is a low-latency and high-quality end-to-end speech interaction model built upon Llama-3.1-8B-Instruct, aiming to achieve speech capabilities at the GPT-4o level.
Steer LLM outputs towards a certain topic/subject and enhance response capabilities using activation engineering by adding steering vectors
An easy-to-understand framework for LLM samplers that rewind and revise generated tokens
The best ChatGPT that $100 can buy.
Letting LLMs negotiate against each other
Train the smallest LM you can that fits in 16MB. Best model wins!
A live multiplayer trivia game where users can bid for the subject of the next question
Port of Facebook's LLaMA model in C/C++
A bot that provides Youtube vid chapters on Twitter (a.k.a. X )
A novel Multimodal Large Language Model (MLLM) architecture, designed to structurally align visual and textual embeddings.
AI management tool
The Truth Is In There: Improving Reasoning in Language Models with Layer-Selective Rank Reduction
A guidance language for controlling large language models.
The fastest way to create an HTML app
Efficient Retrieval Augmentation and Generation Framework
Evaluating LLMs with CommonGen-Lite
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference and quantization, ASP.NET core integration and UI.
Code for the paper "SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot" with LLaMA implementation.
Finetuning InstructLLaMA with portuguese data
Homework assignment made on 07.09.2021. Had 3h to complete.
Crossover Ethereum/Internet Computer dApp demo and proof of concept. Sign in using Metamask, link eth address to IC identity FTW! ∞