Steven Zheng

@Deep-unlearning ยท User

GitHub profile โ†— ยท Compare

Open Source Audio MLE at ๐Ÿค—

@huggingfaceParis, France138 followers60 repositories

Repositories

Deep-unlearning/smol-audio

Practical, Colab-friendly notebooks for fine-tuning and running audio AI models

โ˜… 425Jupyter NotebookForks 29

Deep-unlearning/TTS-Audio-Suite

A ComfyUI custom node integration for multi-language High-quality Text-to-Speech and Voice Conversion nodes using multiple engines like RVC, ResembleAI's Chatterbox TTS, F5-TTS, Higgs Audio 2 and Microsoft VibeVoice with unlimited text length, SRT timing, Character support, Audio Analyzer, Silent Speech Analyzer, audio edit and more!!

โ˜… 2Forks 0

Deep-unlearning/transformers

๐Ÿค— Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

โ˜… 0PythonForks 0

Deep-unlearning/unsloth

Fine-tuning & Reinforcement Learning for LLMs. ๐Ÿฆฅ Train OpenAI gpt-oss, DeepSeek, Qwen, Llama, Gemma, TTS 2x faster with 70% less VRAM.

โ˜… 0Forks 0

Deep-unlearning/clawdbot

Your own personal AI assistant. Any OS. Any Platform. The lobster way. ๐Ÿฆž

โ˜… 1Forks 0

Deep-unlearning/Qwen3-TTS

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.

โ˜… 0Forks 0

Deep-unlearning/Fun-ASR

Fun-ASR is an end-to-end speech recognition large model launched by Tongyi Lab.

โ˜… 0Forks 0

Deep-unlearning/VoiceDiT

[ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis

โ˜… 0Forks 0

Deep-unlearning/trl

Train transformer language models with reinforcement learning.

โ˜… 0PythonForks 0