Kana & Mari
Kana & Mari's SoundRepos (Japanese)
※この番組は日本語でお届けします。 Kana と Mari が、GitHub で見つけた TTS・MIDI・Audio など “音” にまつわる注目リポジトリを声で紹介。 音とコードが交差するオープンソースの世界を軽やかにナビゲートします。 Kana と Mari のプロフィールはこちら: Kana – Newbie Esports Caster Mari – Newbie Esports Analyst ※ 本番組の原稿は生成 AI を用いて自動生成されています。内容には誤りを含む可能性がありますので参考情報としてお楽しみください。
Author
Kana & Mari
Category
Podcast website
Latest episode
Jul 10, 2026
Where to listen?
Podcasts in the app Replaio Radio Coming soonPodcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts
Episodes
OpenMOSS/MOSS-Audio-Tokenizer 16.06.2026 1:51
MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, it supports streaming and variable bitrates, delivering SOTA reconstruction and strong performance in generation and understanding—serving as a unified interface for next-generation native audio language models.
worldwonderer/video-recap-skills 15.06.2026 1:52
Turn any video into a narration recap with claude code skill|用claude code skill把任何视频剪辑成中文解说视频,支持剪映导出
thuhcsi/Crystal 14.06.2026 1:38
Crystal - C++ implementation of a unified framework for multilingual TTS synthesis engine with SSML specification as interface.
dunky11/voicesmith 13.06.2026 1:46
[WIP] VoiceSmith makes training text to speech models easy.
ekwek1/soprano-factory 12.06.2026 1:43
Soprano-Factory: Train your own 2000x realtime text-to-speech model
small-cactus/M.I.L.E.S 11.06.2026 2:01
M.I.L.E.S, a GPT-4-Turbo voice assistant, self-adapts its prompts and AI model, can play any Spotify song, adjusts system and Spotify volume, performs calculations, browses the web and internet, searches global weather, delivers date and time, autonomously chooses and retains long-term memories. Available for macOS and Windows.
Yazdi9/Talking_Face_Avatar 10.06.2026 1:38
Avatar Generation For Characters and Game Assets Using Deep Fakes
XilinJia/Podcini 09.06.2026 1:51
Open source podcast instrument for Android supporting contents from YouTube and YT Music as well as normal podcasts.
LonePheasantWarrior/TalkifyTTS 08.06.2026 1:43
云端大模型驱动的 Android 语音合成应用(TTS引擎)。支持豆包、腾讯、微软、千问等模型。An Android text-to-speech (TTS) engine powered by cloud-based large language models. Supports models such as Doubao, Tencent, Microsoft, and Qwen.
rishikksh20/FastSpeech2 07.06.2026 1:55
PyTorch Implementation of FastSpeech 2 : Fast and High-Quality End-to-End Text to Speech
herimor/voxtream 06.06.2026 2:03
VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control
Hagsten/Talkify 05.06.2026 1:40
Javascript Text to speech library
robinhad/ukrainian-tts 04.06.2026 1:49
Ukrainian TTS (text-to-speech) using ESPNET
foyoux/pygtrans 03.06.2026 1:51
谷歌翻译, 支持 APIKEY 一口气翻译十万条
CMsmartvoice/One-Shot-Voice-Cloning 02.06.2026 1:42
:relaxed: One Shot Voice Cloning base on Unet-TTS
keonlee9420/DiffSinger 01.06.2026 1:45
PyTorch implementation of DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (focused on DiffSpeech)
AIFSH/ComfyUI-GPT_SoVITS 31.05.2026 1:43
a comfyui custom node for GPT-SoVITS! you can voice cloning and tts in comfyui now
yl4579/HiFTNet 30.05.2026 1:46
HiFTNet: A Fast High-Quality Neural Vocoder with Harmonic-plus-Noise Filter and Inverse Short Time Fourier Transform
asiff00/On-Device-Speech-to-Speech-Conversational-AI 29.05.2026 1:48
This is an on-CPU real-time conversational system for two-way speech communication with AI models, utilizing a continuous streaming architecture for fluid conversations with immediate responses and natural interruption handling.
BinWang28/audio-ai-hub 28.05.2026 1:46
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
Xerophayze/TTS-Story 27.05.2026 2:00
TTS-Story is a web-based multi‑voice TTS studio for turning tagged scripts into audiobooks—featuring full speaker management, chunk review/regeneration, a job queue and library system, and local GPU or API backends including Kokoro, Chatterbox, VOX CPM, Pocket-TTS, Kitten-TTS, IndexTTS-2, QWEN3 TTS and Omnivoice engines
AutoArk/GPA 26.05.2026 1:46
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
elevenlabs/skills 25.05.2026 1:56
Collections of skills for building with ElevenLabs
KevinMIN95/StyleSpeech 24.05.2026 1:47
Official implementation of Meta-StyleSpeech and StyleSpeech
ddxfish/sapphire 23.05.2026 1:46
She's the AI agent you come home to.
Similar podcasts
Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.