FFFF
Skip to content
View hironow's full-sized avatar
🌴
On vacation
🌴
On vacation

Block or report hironow

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Stars

ai-vocal

70 repositories

zero-shot voice conversion & singing voice conversion, with real-time support

Python 3,888 532 Updated Apr 20, 2025

VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech

Python 7,889 1,381 Updated Dec 6, 2023

无需情感标注的情感可控语音合成模型,基于VITS

Jupyter Notebook 1,392 170 Updated Mar 30, 2023

Separation voice and delete files with majority silence

Python 52 4 Updated Apr 14, 2023

リアルタイムボイスチェンジャー Realtime Voice Changer

Python 20,864 2,395 Updated Mar 21, 2026

Python scripts for AI voice changers

Python 14 2 Updated Apr 25, 2023

RVCのWebUIを補助するTampermonkeyスクリプト

JavaScript 19 1 Updated May 23, 2023

Port of OpenAI's Whisper model in C/C++

C++ 53,203 6,097 Updated Aug 25, 2026

SoftVC VITS Singing Voice Conversion

Python 28,125 5,034 Updated Nov 11, 2023

so-vits-svc fork with realtime support, improved interface and more features.

Python 9,327 1,225 Updated Aug 26, 2026

Implementation of Natural Speech 2, Zero-shot Speech and Singing Synthesizer, in Pytorch

Python 1,333 104 Updated Sep 24, 2023

so-vits-svc

Python 174 69 Updated Oct 20, 2025

Core Engine of Singing Voice Conversion & Singing Voice Clone

Python 2,864 907 Updated Apr 23, 2024

Real-time end-to-end singing voice conversion system based on DDSP (Differentiable Digital Signal Processing)

Python 2,656 285 Updated Aug 12, 2026

*CREPE+HYBRID TRAINING* A very experimental fork of the Retrieval-based-Voice-Conversion-WebUI repo that incorporates a variety of other f0 methods, along with a hybrid f0 nanmedian method.

Python 1,232 261 Updated Sep 27, 2023

Community interface for generative AI

TypeScript 9,048 915 Updated Apr 30, 2024

A multi-voice TTS system trained with an emphasis on quality

Jupyter Notebook 14,870 2,037 Updated Nov 19, 2024

Eleven Labs text to speech package for NodeJS. You can use the official package at: https://www.npmjs.com/package/elevenlabs

JavaScript 179 31 Updated Feb 12, 2024

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

Python 23,765 2,387 Updated Jul 13, 2026

Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable…

Jupyter Notebook 23,586 2,688 Updated Mar 3, 2026

vits2 backbone with multilingual-bert

Python 8,796 1,305 Updated Aug 24, 2026
JavaScript 77 11 Updated Mar 18, 2024

日本語TTS(VITS)の学習と音声合成のGradio WebUI

Python 42 5 Updated Jan 5, 2024

The Open Source Code of UniAudio

Python 608 40 Updated Jul 22, 2024

💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

1,399 152 Updated Jun 6, 2024

Faster Whisper transcription with CTranslate2

Python 25,104 2,045 Updated Nov 19, 2025

An environment where you can try out faster-whisper immediately.

Python 37 2 Updated Nov 21, 2024

A simple FastAPI Server to run XTTSv2

Python 597 156 Updated Jul 21, 2024

文章から感情豊かな音声を生成する Bert-VITS2 を簡単に使えます。

Batchfile 134 13 Updated Jan 8, 2024
0