Back to people
@ysaito_human
Y

Yuki Saito

音声
@ysaito_human

Lecturer (Sr. Assistant Professor) @ UTokyo-SaruLab, Japan, 特定フェロー@産総研 (JST BOOST 若手研究者支援, 2025 ~ 2030), 講談社「音声変換入門 Pythonで作って学ぶボイスチェンジャー」(🦜本)

832Followers593Following728PostsView on X

Recent posts

8月のMUS研究会で1件の発表(by D1 有田さん)がございます🎵 https://www.ipsj.or.jp/kenkyukai/event/mus147.html

Photo 1

電子情報通信学会誌 8月号の「知識の森」に「音声合成」の記事を寄稿しました🌳 https://www.journal.ieice.org/summary.php?id=k109_8_781&year=2026&lang=J

Photo 1

Three papers were accepted for APSIPA ASC 2026🇹🇭: - Y. Narahata+, Lead Vocal Separation - D. Yang+, CraBERT: Efficient Phoneme Encoder Pre-Training - R. Arita+, Acoustic Feature Analysis of Speech and Singing Voices for Multidimensional Evaluation Congrats!👏

今月号の日本音響学会誌「連載:初学者に薦める入門書 第2回」にて「Pythonで学ぶ音声合成(山本・高道)」を推薦させていただきました📖 Pythonで音声合成を学びたいという全ての人々にオススメの一冊です!🐍 https://book.impress.co.jp/books/1120101073

D2 中田さんがASJ 粟屋 潔学術奨励賞を,D1 有田さんが ASJ 学生優秀発表賞を受賞しました.おめでとうございます!👏👏👏👏👏

Photo 1Photo 2

ポスター聴きに来ていただいた皆様ありがとうございました!日本語TTSの評価は超難しいのでみんなで何とかしてできるようになっていきましょう💪

Photo 1

高道研の学生さん2名(岸さん&八木さん)による共著論文が INTERSPEECH2026 に採択されました!おめでとうございます🥳 金曜日からの音学シンポジウム2026 at 電通大 でも発表予定です🧐 https://x.com/ysaito_human/status/2047669071525879820

@ysaito_human
Y
Yuki Saito@ysaito_human

One paper by Kishi-san and Yagi-san (equally contributed under the supervision by Takamichi-sensei), "Do speech foundation models perceive speaker similarity as humans do?", has been accepted for #INTERSPEECH2026 👂 Kudos to the great collaborators👏👏👏

岸さんと八木さんによる論文(武満先生の指導下で同等に貢献)「Do speech foundation models perceive speaker similarity as humans do?」が #INTERSPEECH2026 に採択されました 👂 素晴らしい協力者たちへ敬意を 👏👏👏

原文を表示 (en)

One paper by Kishi-san and Yagi-san (equally contributed under the supervision by Takamichi-sensei), "Do speech foundation models perceive speaker similarity as humans do?", has been accepted for #INTERSPEECH2026 👂 Kudos to the great collaborators👏👏👏

Park-san(@nonmetal_)による我々の研究「Probing Token Spaces for Cross-Generator AI-Generated Music Detection」がICML 2026 Workshop on Machine Learning for Audio(ICMLWMLA)にアクセプトされました🎧 おめでとう!#ICML2026

原文を表示 (en)

Our work "Probing Token Spaces for Cross-Generator AI-Generated Music Detection" by Park-san (@nonmetal_ ) has been accepted for ICML 2026 Workshop on Machine Learning for Audio (ICMLWMLA) 🎧 Congurats! #ICML2026

私たちの論文「DialogueSidon: Recovering Full-Duplex Dialogue Tracks from In-the-Wild Dialogue Audio」が#SIGDIAL2026(ロングペーパー)に採択されました🙌 Nakata-san @wataru9871と素晴らしい共著者たち(Yamauchi-san、Tsunoo-san、Saruwatari-sensei)に大おめでとう👏👏👏

原文を表示 (en)

Our work "DialogueSidon: Recovering Full-Duplex Dialogue Tracks from In-the-Wild Dialogue Audio" has been accepted for #SIGDIAL2026 (Long paper) 🙌 Many congrats to Nakata-san @wataru9871 and great co-authors (Yamauchi-san, Tsunoo-san, and Saruwatari-sensei) 👏👏👏

@wataru9871
T
took@wataru9871

The paper is now available! Check it out https://arxiv.org/abs/2604.09344

🙇‍♂️🙇‍♂️🙇‍♂️🙇‍♂️🙇‍♂️🙇‍♂️🙇‍♂️

システム情報学専攻の説明会始まってます🤗 工学部2号館の212講義室です!

Photo 1