Publication

At the 155th (Spring 2026) Meeting of the Acoustical Society of Japan, Yuuki Yamakawa (B4) and Professor Sei Ueno gave presentations

At the 155th (Spring 2026) Meeting of the Acoustical Society of Japan, held from March 17 to 19, Yuuki Yamakawa (B4) and Professor Sei Ueno gave presentations on the following topics.

「LLM-based 音声合成における連続性を考慮したビームサーチ」 Sei Ueno, Akinobu Lee (Nagoya Institute of Technology)

「大規模モデルを用いた演じ分けに着目した落語音声合成」 Yuuki Yamakawa, Sei Ueno, Akinobu Lee (Nagoya Institute of Technology)

At the 32nd Annual Meeting of the Association for Natural Language Processing (NLP2026), Kaho Suzuki and Yu Kaneko (M2) gave presentations

At the 32nd Annual Meeting of the Association for Natural Language Processing (NLP2026), held from March 9 to 13, Kaho Suzuki and Yu Kaneko (M2) gave presentations on the following topics.

「発散・深掘り対話戦略に基づくLLM対話システムによる悩みの内省支援」 Kaho Suzuki, Sei Ueno, Akinobu Lee (Nagoya Institute of Technology)

「両価性の気づきの促進を重視する動機づけ面接システム」 Yu Kaneko, Sei Ueno, Akinobu Lee (Nagoya Institute of Technology)

At the Sixth Joint Meeting Acoustical Society of America and Acoustical Society of Japan, Keigo Ichikawa, Umi Okamoto, Junichi Shimazaki, Momone Suzuki gave presentations

At the The Sixth Joint Meeting Acoustical Society of America and Acoustical Society of Japan, held from December 1 to November 5, Keigo Ichikawa(D1), Umi Okamoto(M2), Junichi Shimazaki(M2), Momone Suzuki(M1) gave presentations on the following topics.

「Multi-talker conversational speech generation for training speaker diarization model via text-to-speech」 Keigo Ichikawa,Sei Ueno,Akinobu Lee

「Fine-Tuning Strategies for Large-Scale Face-Conditioned Text-to-Speech」 Umi Okamoto,Sei Ueno,Akinobu Lee

「Speech Synthesis with Diverse Laughter Types using Artificial Data」 Junichi Shimazaki,Sei Ueno,Akinobu Lee

At the 17th Asia Pacific Signal and Information Processing Association Annual Summit and Conference, Umi Okamoto(M2) gave presentations

At the 17th Asia Pacific Signal and Information Processing Association Annual Summit and Conference, held from October 22 to October 24, Umi Okamoto(M1) gave presentations on the following topics.

「Face-conditioned Large-scale Text-to-Speech via Speaker Embedding Prediction from Facial Images」 Umi Okamoto,Sei Ueno,Akinobu Lee

At the 153rd (Spring 2025) Research Presentation of the Acoustical Society of Japan, Keigo Ichikawa (M2), Momone Suzuki (B4), and Professor Sei Ueno gave presentations

At the 153rd (Spring 2025) Research Presentation of the Acoustical Society of Japan, held from March 17 to March 19, Keigo Ichikawa (M2), Momone Suzuki (B4), and Professor Sei Ueno gave presentations on the following topics.

「話者遷移確率に基づく話者ダイアライゼーションのためのデータ生成」 Keigo Ichikawa,Sei Ueno,Akinobu Lee 「音響情報を考慮した大規模言語モデルによる音声認識の誤り訂正」 Momone Suzuki,Sei Ueno,Akinobu Lee

「拡散モデルを用いた音声合成による音声認識のデータ拡張」 Sei Ueno,Akinobu Lee