Skip to content
View JinZr's full-sized avatar
🌖
𝚗𝚘𝚝 𝚊 𝚕𝚘𝚝 𝚐𝚘𝚒𝚗𝚐 𝚘𝚗 𝚊𝚝 𝚝𝚑𝚎 𝚖𝚘𝚖𝚎𝚗𝚝
🌖
𝚗𝚘𝚝 𝚊 𝚕𝚘𝚝 𝚐𝚘𝚒𝚗𝚐 𝚘𝚗 𝚊𝚝 𝚝𝚑𝚎 𝚖𝚘𝚖𝚎𝚗𝚝

Highlights

  • Pro

Block or report JinZr

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 44 3 Updated Jun 3, 2026

Modular self-supervised learning for sleep signals

Python 6 1 Updated Oct 7, 2026

CuBlaze is a CUDA implementation of the InfoNCE Loss (Information Noise-Contrastive Estimation) for self-supervised contrastive learning. This implementation provides full batch processing with nat…

Python 1 Updated Jul 21, 2025

PyTorch implementation of "Supervised Contrastive Learning" (and SimCLR incidentally)

Python 3,448 555 Updated Dec 26, 2023

PyTorch implementation of Contrastive Learning methods

Python 1,992 180 Updated Oct 4, 2023

Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deplo…

C 1,220 221 Updated May 8, 2026

A Deep Learning Python Toolkit for Healthcare Applications.

Python 1,669 805 Updated Oct 7, 2026

YASA (Yet Another Spindle Algorithm): a Python package to analyze polysomnographic sleep recordings.

Python 589 131 Updated Oct 6, 2026

PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html

Python 2,220 326 Updated Sep 10, 2025

State-of-the-art audio codec with 90x compression factor. Supports 44.1kHz, 24kHz, and 16kHz mono/stereo audio.

Python 1,871 193 Updated Jul 16, 2026

unofficial implementation of the High Fidelity Neural Audio Compression

Python 178 22 Updated Aug 15, 2024

A list of publicly available room impulse response datasets and scripts to download them.

Shell 616 54 Updated Aug 21, 2026

collection of awesome research in brain decoding, including interaction with multi-modalities, theories, and foundation models.

HTML 108 10 Updated Jun 12, 2025

An evolving, large-scale and multi-domain ASR corpus for low-resource languages with automated crawling, transcription and refinement

Python 206 14 Updated Apr 28, 2026

✨✨Latest Advances on Multimodal Large Language Models

18,046 1,135 Updated Oct 1, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 15,154 1,750 Updated Oct 5, 2026

Some fast-ish algorithms for batch text search in moderate-sized collections, intended for data cleanup

Python 79 15 Updated Jun 30, 2025

My Python scripts for crawling paper related on speech processing.

Python 7 Updated Mar 27, 2024

Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.

83,324 9,693 Updated Feb 5, 2026

Conv-TasNet: Surpassing Ideal Time-Frequency Magnitude Masking for Speech Separation Pytorch's Implement

Python 554 83 Updated May 26, 2023

Cantonese Linguistics and NLP

Python 420 41 Updated May 26, 2026

Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM

Python 17 Updated Nov 7, 2024

The RWTH ASR Toolkit.

C++ 59 17 Updated Oct 7, 2026

Kaldi-compatible online & offline feature extraction with PyTorch, supporting CUDA, batch processing, chunk processing, and autograd - Provide C++ & Python API

C++ 220 38 Updated Jul 10, 2026

Unofficial PyTorch implementation of Google AI's VoiceFilter system

Python 1,224 240 Updated Jul 25, 2024

This is the official repository for M2UGen

Jupyter Notebook 514 42 Updated Jan 2, 2025

Typing to Listen at the Cocktail Party: Text-Guided Target Speaker Extraction (LLM-TSE)

JavaScript 43 2 Updated Oct 13, 2023

A PyTorch implementation of "TasNet: Surpassing Ideal Time-Frequency Masking for Speech Separation" (see recipes in aps framework https://github.com/funcwj/aps)

Python 219 61 Updated Jul 6, 2023

Google Drive CLI Client

Rust 2,101 147 Updated Aug 3, 2024
Next