Skip to content
View Galileo-star's full-sized avatar

Block or report Galileo-star

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM

Python 1 Updated Aug 3, 2026

A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM

Python 823 217 Updated Sep 9, 2026

MiMo: Unlocking the Reasoning Potential of Language Model – From Pretraining to Posttraining

Python 2,309 106 Updated Jun 5, 2025

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice…

Python 13,328 1,727 Updated Mar 17, 2026
Python 1 Updated Jun 18, 2026

中国专利.skill:专利点挖掘与交底书(发明/实用/外观)编写,通俗解读专利,嗅探政策动向,辅助审查答复。

Python 8,946 922 Updated Sep 9, 2026

Automatically Update Text-to-speech (TTS) Papers Daily using Github Actions (Update Every 12th hours)

Python 668 41 Updated Sep 9, 2026

Zonos2 is a leading open-weight text-to-speech MoE.

Python 311 30 Updated Jul 6, 2026

QuantaAlpha transforms how you discover quantitative alpha factors by combining LLM intelligence with evolutionary strategies. Just describe your research direction, and watch as factors are automa…

Python 1,511 292 Updated Jun 29, 2026

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

Python 36,897 4,197 Updated Sep 2, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,199 108,622 Updated Aug 16, 2026

The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

TypeScript 389,299 81,815 Updated Sep 9, 2026
Python 9 Updated Dec 9, 2025

A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

Python 973 79 Updated Apr 9, 2026

[INTERSPEECH 2026 Oral]Official code for "Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis"

Python 125 8 Updated Jun 21, 2026

CosyVoice_DPO_NOTES: Supercharge Your Cosyvoice model with Cutting-Edge DPO Fine-Tuning!

Python 129 19 Updated Aug 8, 2025

Official Repository of Paper: "SynParaSpeech: Automated Synthesis of Paralinguistic Datasets for Speech Generation and Understanding" (ICASSP 2026)

JavaScript 72 4 Updated Apr 27, 2026

High-quality speech synthesis with LoRA fine-tuning on index-tts, enhancing prosody and naturalness for single and multi-speaker voices.

Python 337 31 Updated Mar 12, 2026

This repository presents an evaluation framework for speech-to-speech (S2S) models, following the methodology described in the EmphAsses paper (de Seyssel et al., 2023).

Python 25 1 Updated Jan 9, 2024

Official code for "EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting"

Python 125 13 Updated Oct 16, 2025

Wan: Open and Advanced Large-Scale Video Generative Models

Python 17,444 2,235 Updated Mar 17, 2026
HTML 5 Updated Aug 18, 2025

Open-Source Frontier Voice AI

Python 54,091 6,099 Updated Sep 3, 2026

Text-audio foundation model from Boson AI

Python 8,346 641 Updated Jun 5, 2026

Official code for "F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization"

Python 169 18 Updated Mar 3, 2026

Easily train a good VC model with voice data <= 10 mins!

Python 38,171 5,267 Updated Aug 4, 2026

Evaluation Metrics Used For The Performance Evaluation of Voice Conversion (VC) Models

Jupyter Notebook 19 2 Updated Jul 8, 2025

Multilingual Automatic Speech Recognition with word-level timestamps and confidence

Python 2,844 211 Updated Aug 17, 2026

SoTA open-source TTS

Python 26,332 3,539 Updated Jul 21, 2026

Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads

Jupyter Notebook 2,772 204 Updated Jun 25, 2024