Haseebasif7/cir-siglip
Three complementary item retrieval heads compared at matched conditions over one frozen SigLIP backbone. Polyvore Outfits
CS Sopho
Three complementary item retrieval heads compared at matched conditions over one frozen SigLIP backbone. Polyvore Outfits
A world model in your agent's toolbox - see what happens before it does.
🎨 NeuroInk - Neural style transfer Algorithm using VGG19. Blends content from one image with the style of another to create AI-generated artwork Using Pytorch
Vector Space Model for Information Retrieval
BotX is an advanced Discord chatbot built with discord.py, featuring 24/7 hosting on Pylexnodes. It offers dynamic responses powered by the Mistral AI model, voice interaction with Google Text-to-Speech (gTTS) and FFmpeg, and an interactive help menu. Enhance your Discord server with BotX's sophisticated features.
🇵🇰 Province aware hierarchical image geolocation for Pakistan using dual vision encoders and geocell based inference
Train transformer language models with reinforcement learning.
Developed a fake news prediction model using Support Vector Machine (SVM) and achieved an impressive 99.8% accuracy on the training dataset. The text preprocessing included stemming to normalize words, followed by the use of TF-IDF vectorization to convert the text data into numerical form
UrduReason-Eval: A comprehensive evaluation dataset with 800 Urdu reasoning problems across 6 categories (arithmetic, logical deduction, temporal, comparative, and causal reasoning) for assessing reasoning capabilities in Urdu language models.
Config Files for Readme
A lightweight, local-first, and 🆓 experiment tracking library from Hugging Face 🤗
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
A TTS model capable of generating ultra-realistic dialogue in one pass.
SurfAgent is an AI agent built from scratch that performs web searches using Selenium and Brave Search, Langchain and supports GROQ or OLLAMA providing reliable sources for the results.
A PPO trained agent for VizDoom’s “Defend the Center” scenario, using a custom Gym wrapper and StableBaselines3
Collection of Notebooks from DeepLearning.AI GANs Specialization
This project enables the inversion of real facial images into the latent space of StyleGAN2-ADA, leveraging dlib based facial landmark detection for accurate alignment, and supports semantic attribute editing through controlled latent vector perturbations along interpretable directions
A PyTorch implementation of the Pix2Pix architecture for image-to-image translation, tailored for depth estimation from dashboard camera images
A PyTorch reimplementation of Neural Radiance Fields (NeRF) for 3D scene reconstruction from 2D images