Kana & Mari

Kana & Mari's SoundRepos (Japanese)

※この番組は日本語でお届けします。 Kana と Mari が、GitHub で見つけた TTS・MIDI・Audio など “音” にまつわる注目リポジトリを声で紹介。 音とコードが交差するオープンソースの世界を軽やかにナビゲートします。 Kana と Mari のプロフィールはこちら: Kana – Newbie Esports Caster Mari – Newbie Esports Analyst ※ 本番組の原稿は生成 AI を用いて自動生成されています。内容には誤りを含む可能性がありますので参考情報としてお楽しみください。

Author

Kana & Mari

Category

Technology

Podcast website

www.aquariumy.co.jp

Latest episode

Jul 10, 2026

Where to listen?

Podcasts in the app Replaio Radio Coming soon

Podcasts are coming to the app soon. Install now and be the first to see a whole new take on podcasts

Get it on Google Play Install for free Android 5M+ downloads · 4.8 rating iOS soon

Episodes

OpenMOSS/MOSS-Audio-Tokenizer 16.06.2026

MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, it supports streaming and variable bitrates, delivering SOTA reconstruction and strong performance in generation and understanding—serving as a unified interface for next-generation native audio language models.

worldwonderer/video-recap-skills 15.06.2026

Turn any video into a narration recap with claude code skill|用claude code skill把任何视频剪辑成中文解说视频,支持剪映导出

thuhcsi/Crystal 14.06.2026

Crystal - C++ implementation of a unified framework for multilingual TTS synthesis engine with SSML specification as interface.

dunky11/voicesmith 13.06.2026

[WIP] VoiceSmith makes training text to speech models easy.

ekwek1/soprano-factory 12.06.2026

Soprano-Factory: Train your own 2000x realtime text-to-speech model

small-cactus/M.I.L.E.S 11.06.2026

M.I.L.E.S, a GPT-4-Turbo voice assistant, self-adapts its prompts and AI model, can play any Spotify song, adjusts system and Spotify volume, performs calculations, browses the web and internet, searches global weather, delivers date and time, autonomously chooses and retains long-term memories. Available for macOS and Windows.

Yazdi9/Talking_Face_Avatar 10.06.2026

Avatar Generation For Characters and Game Assets Using Deep Fakes

XilinJia/Podcini 09.06.2026

Open source podcast instrument for Android supporting contents from YouTube and YT Music as well as normal podcasts.

LonePheasantWarrior/TalkifyTTS 08.06.2026

云端大模型驱动的 Android 语音合成应用(TTS引擎)。支持豆包、腾讯、微软、千问等模型。An Android text-to-speech (TTS) engine powered by cloud-based large language models. Supports models such as Doubao, Tencent, Microsoft, and Qwen.

rishikksh20/FastSpeech2 07.06.2026

PyTorch Implementation of FastSpeech 2 : Fast and High-Quality End-to-End Text to Speech

herimor/voxtream 06.06.2026

VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control

Hagsten/Talkify 05.06.2026

Javascript Text to speech library

robinhad/ukrainian-tts 04.06.2026

Ukrainian TTS (text-to-speech) using ESPNET

foyoux/pygtrans 03.06.2026

谷歌翻译, 支持 APIKEY 一口气翻译十万条

CMsmartvoice/One-Shot-Voice-Cloning 02.06.2026

:relaxed: One Shot Voice Cloning base on Unet-TTS

keonlee9420/DiffSinger 01.06.2026

PyTorch implementation of DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (focused on DiffSpeech)

AIFSH/ComfyUI-GPT_SoVITS 31.05.2026

a comfyui custom node for GPT-SoVITS! you can voice cloning and tts in comfyui now

yl4579/HiFTNet 30.05.2026

HiFTNet: A Fast High-Quality Neural Vocoder with Harmonic-plus-Noise Filter and Inverse Short Time Fourier Transform

asiff00/On-Device-Speech-to-Speech-Conversational-AI 29.05.2026

This is an on-CPU real-time conversational system for two-way speech communication with AI models, utilizing a continuous streaming architecture for fluid conversations with immediate responses and natural interruption handling.

BinWang28/audio-ai-hub 28.05.2026

The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.

Xerophayze/TTS-Story 27.05.2026

TTS-Story is a web-based multi‑voice TTS studio for turning tagged scripts into audiobooks—featuring full speaker management, chunk review/regeneration, a job queue and library system, and local GPU or API backends including Kokoro, Chatterbox, VOX CPM, Pocket-TTS, Kitten-TTS, IndexTTS-2, QWEN3 TTS and Omnivoice engines

AutoArk/GPA 26.05.2026

[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!

elevenlabs/skills 25.05.2026

Collections of skills for building with ElevenLabs

KevinMIN95/StyleSpeech 24.05.2026

Official implementation of Meta-StyleSpeech and StyleSpeech

ddxfish/sapphire 23.05.2026

She's the AI agent you come home to.

Listen to the Kana & Mari's SoundRepos (Japanese) podcast in Replaio

Radio and podcasts in one app - free, with no sign-up. Install today and do not miss the launch

Get it on Google Play

Replaio is not a podcast publisher; show names, artwork and audio belong to their authors and are distributed through public RSS feeds.