HemanthSai7/transformers
๐ค Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
ML Engineer
๐ค Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Answers of all leetcode questions in Kunal Kushwaha's java bootcamp
Attempt to implement research papers
Revise your subjects with studybot using RAG
Infrastructure to enable deployment of ML models to low-power resource-constrained embedded targets (including microcontrollers and digital signal processors).
The best ChatGPT that $100 can buy.
Your hyper-personal, always-on, open-source AI companion.
Code for my pilot task
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
template for my all my TF projects
My GitHub Readme
Code Documentation, redefined: Where AI Meets Clarity!
Robust Speech Recognition via Large-Scale Weak Supervision
A tensorflow implementation of an HMM layer
Deploying your FastAPI Applications
Sign Language Prediction streamlit link
Internship IVIS Labs
Scrape images from google and create your own training dataset
Allosaurus is a pretrained universal phone recognizer for more than 2000 languages
Circuitbot
LLM101n: Let's build a Storyteller
Longformer: The Long-Document Transformer
https://legalease-capstone.streamlit.app/
[ASRU 2021] Efficient Conformer: Progressive Downsampling and Grouped Attention for Automatic Speech Recognition
Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC)
GSoC'22 @ TensorFlow Notebooks, Code and much more