Skip to content
View linan2's full-sized avatar

Organizations

@group122

Block or report linan2

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 1,374 140 Updated Sep 28, 2026

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

Python 54,664 6,133 Updated Oct 7, 2026

Google Scholar skills for Claude Code — search, citation tracking, full-text access, and Zotero export via Chrome DevTools MCP

Python 523 26 Updated Mar 13, 2026

「摘星」· 微信助手 — AI Agent 智能助手 / AI 群聊摘要 / 关键词即时提醒 / 公众号摘要与文章实时推送 / RAG 语义检索 / 聊天、朋友圈归档 / 收藏导出 。本地运行,数据不出本机。

Python 82 54 Updated Oct 7, 2026

AuK: An Open-Source Foundational Model for Speech Generation and Editing

Python 1,467 107 Updated Sep 25, 2026

Wespeaker implementations for speaker recognition and verification: U3-xi,Uncertainty aware AAM Softmax, Score normalization, calibration, and unofficial RecXi (NeurIPS 2023) source code.

Python 26 2 Updated Sep 19, 2026

A curated list of plugins for DeepSeek Harness (dsh) · DeepSeek Harness 插件精选列表

JavaScript 18,009 3,781 Updated Oct 7, 2026

FireRedTTS3: Multilingual and Multi-Dialect Voice Cloning with Instruction-Guided Voice Design and Speech Editing

Python 1,769 19 Updated Sep 8, 2026

DeepSeek Harness: Everything is a Plugin.

TypeScript 245,153 29,369 Updated Oct 3, 2026

Java 离线语音识别(ASR)、语音合成(TTS)、声纹识别、VAD、说话人分离、降噪、KWS 关键词唤醒 SDK。基于 sherpa-onnx,Spring Boot / Solon 自动装配,JDK 8+ 兼容。

Java 37 8 Updated Oct 1, 2026

Robust Speech Recognition via Large-Scale Weak Supervision

Python 110,111 13,345 Updated Aug 31, 2026

AI视频, AI动漫,AI 短剧,AI漫剧自动化生成工具

Python 1,736 350 Updated Sep 17, 2026

Open-Source Turn-Taking Detection Model and Dataset for Full-Duplex Spoken Dialogue Systems

Python 146 8 Updated Jan 25, 2026

Open Frontier Intelligence

8,900 755 Updated Aug 6, 2026

Reference implementation of an end-to-end voice agent built using the NVIDIA Nemotron models

Python 232 63 Updated Oct 7, 2026

Hy3 (295B A21B), a leading reasoning and agent model in its size, with great cost efficiency.

Python 664 210 Updated Jul 17, 2026

A Fully Self-Hosted Solution for Full-Duplex Voice Interaction

Python 589 48 Updated Sep 28, 2025

Adaptive Flow-Matching for Target Speaker Extraction

Python 44 5 Updated Jul 13, 2026

Plug-and-play streaming semantic VAD for real-time full-duplex spoken dialogue systems.

Python 322 46 Updated Jul 17, 2026

A curated list of full-duplex spoken dialogue models & benchmarks

260 15 Updated Oct 4, 2026

A curated list of models, benchmarks, tools and guides for audio editing

49 9 Updated Oct 4, 2026

Nano vLLM

Python 15,730 2,662 Updated Apr 26, 2026

猫世界

1 Updated Jun 11, 2026

Speed-optimized streaming neural speech enhancement network

Python 157 38 Updated Sep 28, 2026

Lightweight coding agent that runs in your terminal

Rust 128,187 20,085 Updated Oct 7, 2026

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.

Python 129,114 20,207 Updated Oct 7, 2026

Speaker-Reasoner: Scaling Interaction Turns and Reasoning Patterns for Timestamped Speaker-Attributed ASR

Python 96 2 Updated May 13, 2026

Robust Speech Recognition Across Languages, Dialects, and Complex Acoustic Scenarios

Python 340 35 Updated Apr 23, 2026
Next