Satvik Dixit

@satvik-dixit · User

GitHub profile ↗ · Compare

Graduate student at Carnegie Mellon University

5 followers19 repositories

Repositories

satvik-dixit/CPP

Python implementation of Cepstral Peak Prominence (CPP)

★ 6Jupyter NotebookForks 1

satvik-dixit/MFCon

Code for the paper: Improving Speaker Representations Using Contrastive Losses on Multi-scale Features

★ 5PythonForks 1

satvik-dixit/mace

Code for the paper: MACE: Leveraging Audio for Evaluating Audio Captioning Systems

★ 13PythonForks 1

satvik-dixit/EzAudio

High-quality Text-to-Audio Generation with Efficient Diffusion Transformer

★ 0PythonForks 0

satvik-dixit/explainability_SER

Code for the paper: Explaining Deep Learning Embeddings for Speech Emotion Recognition by Predicting Interpretable Acoustic Features

★ 3Jupyter NotebookForks 1

satvik-dixit/fense

Fluency ENhanced Sentence-bert Evaluation (FENSE), metric for audio caption evaluation. And Benchmark dataset AudioCaps-Eval, Clotho-Eval.

★ 0PythonForks 0

satvik-dixit/T-FOLEY

Implementation of the paper, T-FOLEY: A Controllable Waveform-Domain Diffusion Model for Temporal-Event-Guided Foley Sound Synthesis, accepted in 2024 ICASSP

★ 0PythonForks 0

satvik-dixit/AudioLDM

AudioLDM: Generate speech, sound effects, music and beyond, with text.

★ 0PythonForks 0

satvik-dixit/SpeechTokenizer

This is the code for the SpeechTokenizer presented in the SpeechTokenizer: Unified Speech Tokenizer for Speech Language Models. Samples are presented on

★ 0PythonForks 0

satvik-dixit/pyroomacoustics

Pyroomacoustics is a package for audio signal processing for indoor applications. It was developed as a fast prototyping platform for beamforming algorithms in indoor scenarios.

★ 0PythonForks 0