Kana & Mari

Kana & Mari's SoundRepos (Japanese)

※この番組は日本語でお届けします。 Kana と Mari が、GitHub で見つけた TTS・MIDI・Audio など “音” にまつわる注目リポジトリを声で紹介。 音とコードが交差するオープンソースの世界を軽やかにナビゲートします。 Kana と Mari のプロフィールはこちら: Kana – Newbie Esports Caster Mari – Newbie Esports Analyst ※ 本番組の原稿は生成 AI を用いて自動生成されています。内容には誤りを含む可能性がありますので参考情報としてお楽しみください。

Author

Kana & Mari

Category

Technology

Podcast website

www.aquariumy.co.jp

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

maum-ai/univnet 27.04.2026

Unofficial PyTorch Implementation of UnivNet Vocoder (https://arxiv.org/abs/2106.07889)

LSimon95/megatts2 26.04.2026

Unoffical implementation of Megatts2

makerjackie/MTTS 25.04.2026

A Demo of Mandarin/Chinese TTS frontend

haoheliu/voicefixer_main 24.04.2026

General Speech Restoration

ManimCommunity/manim-voiceover 23.04.2026

Manim plugin for all things voiceover

travisvn/obsidian-edge-tts 22.04.2026

Free, high quality text-to-speech for your Obsidian notes, leveraging Microsoft Edge's Read Aloud API.

zlargon/google-tts 21.04.2026

Google TTS (Text-To-Speech) for node.js

developersdigest/ai-devices 20.04.2026

AI Device Template Featuring Whisper, TTS, Groq, Llama3, OpenAI and more

zarazhangrui/personalized-podcast 19.04.2026

Turn any content into a personalized AI podcast. NotebookLM-style, except you control the script, voices, and hosts. Listen in Apple Podcasts, Spotify, or any podcast app.

debpalash/OmniVoice-Studio 18.04.2026

A Cinematic audio dubbing, Cloning and voice generation studio

akdeb/ElatoAI 17.04.2026

Realtime Voice AI with 100+ Models on Arduino ESP32 for AI Toys, Companions, and Devices

izwi-ai/izwi 16.04.2026

Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.

moonshine-ai/moonshine 15.04.2026

Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces

Saganaki22/ComfyUI-OmniVoice-TTS 14.04.2026

OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue

OpenMOSS/MOSS-TTS-Nano 13.04.2026

MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run directly on CPU without a GPU, and keeps the deployment stack simple enough for local demos, web serving, and lightweight product integration.

Aratako/T5Gemma-TTS 12.04.2026

Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM

lmnt-com/wavegrad 11.04.2026

A fast, high-quality neural vocoder.

mbzuai-oryx/LLMVoX 10.04.2026

LLMVoX: Autoregressive Streaming Text-to-Speech Model for Any LLM

Adri6336/gpt-voice-conversation-chatbot 09.04.2026

Allows you to have an engaging and safely emotive spoken / CLI conversation with the AI ChatGPT / GPT-4 while giving you the option to let it remember things discussed.

richardr1126/openreader 08.04.2026

An open-source read-along document reader server with high-quality TTS options, synchronized highlighting, and audiobook export for EPUB, PDF, DOCX, TXT, and MD.

Aratako/Irodori-TTS 07.04.2026

A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control

LlmKira/fast-langdetect 06.04.2026

⚡️ 80x faster Fasttext language detection out of the box | Split text by language

Sharrnah/whispering-ui 05.04.2026

Native UI for the Whispering Tiger project - https://github.com/Sharrnah/whispering (live transcription / translation)

keonlee9420/Expressive-FastSpeech2 04.04.2026

PyTorch Implementation of Non-autoregressive Expressive (emotional, conversational) TTS based on FastSpeech2, supporting English, Korean, and your own languages.

Agents365-ai/video-podcast-maker 03.04.2026

AI-powered video podcast creation skill for coding agents. Supports Bilibili & YouTube, multi-language (zh-CN/en-US), 6 TTS engines (Edge/Azure/ElevenLabs/OpenAI/Doubao/CosyVoice), 4K Remotion rendering.

Listen to the Kana & Mari's SoundRepos (Japanese) podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.