Kana & Mari

Kana & Mari's SoundRepos (Japanese)

※この番組は日本語でお届けします。 Kana と Mari が、GitHub で見つけた TTS・MIDI・Audio など “音” にまつわる注目リポジトリを声で紹介。 音とコードが交差するオープンソースの世界を軽やかにナビゲートします。 Kana と Mari のプロフィールはこちら: Kana – Newbie Esports Caster Mari – Newbie Esports Analyst ※ 本番組の原稿は生成 AI を用いて自動生成されています。内容には誤りを含む可能性がありますので参考情報としてお楽しみください。

Author

Kana & Mari

Category

Technology

Podcast website

www.aquariumy.co.jp

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

shell-nlp/gpt_server 22.05.2026

gpt_server是一个用于生产级部署LLMs、Embedding、Reranker、ASR、TTS、文生图、图片编辑和文生视频的开源框架。

mlalma/kokoro-ios 21.05.2026

Kokoro TTS for iOS and macOSX

keonlee9420/DailyTalk 20.05.2026

Official repository of DailyTalk: Spoken Dialogue Dataset for Conversational Text-to-Speech, ICASSP 2023

DBJD-CR/astrbot_plugin_proactive_chat 19.05.2026

一个能让 Bot 在私聊和群聊中发起主动消息的插件,拥有上下文感知、持久化数据、动态情绪、免打扰时段和 TTS 集成。还有独立 WebUI,可进行个性化配置。 An AstrBot plugin that enables Bot to send proactive messages in private and group chats, featuring context awareness, persistent data, dynamic emotions, do-not-disturb periods, and TTS integration. It also boasts an independent WebUI for personalized.

devnen/Kitten-TTS-Server 18.05.2026

Self-host the ultra-lightweight Kitten TTS model with this enhanced API server with an intuitive Web UI, large text processing for audiobooks, and GPU acceleration.

leaonline/easy-speech 17.05.2026

Cross browser Speech Synthesis also known as Text to speech or TTS; no dependencies; uses Web Speech API

rendchevi/nix-tts 16.05.2026

Nix-TTS: Lightweight and End-to-end Text-to-Speech via Module-wise Distillation

HITsz-TMG/VideoClaw 15.05.2026

AI 全自动化视频生成员工 | Your First AIGC Coworker. Chat an Idea. Get a Film.

ChaituRajSagar/gemini-youtube-automation 14.05.2026

A fully autonomous AI Agent/Python pipeline that utilizes Large Language Models (LLMs) like Gemini to generate content, produce videos, and automatically upload educational videos to YouTube.

wildminder/awesome-ai-voice 13.05.2026

List of open-source TTS, voice cloning, and music generation models

mahimairaja/voiceai 12.05.2026

Set of with to help those building Voice AI agents ️

PowerBeef/QwenVoice 11.05.2026

Vocello is a local-first voice generation app for Apple Silicon Macs. Public beta for macOS 26; QwenVoice v1.2.3 remains the stable macOS 15 fallback.

r9y9/ttslearn 10.05.2026

ttslearn: Library for Pythonで学ぶ音声合成 (Text-to-speech with Python)

livekit-examples/kitt 09.05.2026

Talk to ChatGPT in real time using LiveKit

yl4579/PL-BERT 08.05.2026

Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions

ElmTran/praises 07.05.2026

Praises is a text-to-speech tool that can help you read text easily.

Elleo/pied 06.05.2026

Pied makes it simple to install and manage text-to-speech Piper voices for use with Speech Dispatcher.

1neReality/MITSUHA 05.05.2026

World's First Multilingual Inexpensive Therapeutic Sophisticated Ultra-responsive Holographic Agent. In simple terms, an AI you can talk to and it'll talk back with a body using VTube Studio.

rishikksh20/iSTFTNet-pytorch 04.05.2026

iSTFTNet : Fast and Lightweight Mel-spectrogram Vocoder Incorporating Inverse Short-time Fourier Transform

frostming/tetos 03.05.2026

A unified interface for multiple Text-to-Speech (TTS) providers.

atomicoo/FCH-TTS 02.05.2026

A fast Text-to-Speech (TTS) model. Work well for English, Mandarin/Chinese, Japanese, Korean, Russian and Tibetan (so far). 快速语音合成模型,适用于英语、普通话/中文、日语、韩语、俄语和藏语(当前已测试)。

Executedone/Chinese-FastSpeech2 01.05.2026

基于标贝数据继续训练,同时对原本的FastSpeech2模型做了改进,引入了韵律表征以及韵律预测模块,使中文发音更生动且富有节奏

JackismyShephard/ultimate-rvc 30.04.2026

An app for creating audio-based content such as song covers and speech using Retrieval-based Voice Conversion.

mathigatti/midi2voice 29.04.2026

Singing synthesis from MIDI file

trymirai/uzu 28.04.2026

A high-performance inference engine for AI models

Listen to the Kana & Mari's SoundRepos (Japanese) podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.