発表文献

2026年度に発表された文献の一覧

学術論文誌

  1. J. Mi, X. Shi, D. Ma, J. He, T. Fujimura, T. Toda, "Robust speech emotion recognition under human speech noise," Computer Speech and Language, Vol. 100, Article 101987, pp. 1-16, Apr. 2026.
  2. T. Komatsu, H. Munakata, Y. Ishikawa, K. Takeda, T. Toda, "Semi-supervised text-audio contrastive learning method using pseudo-text input," APSIPA Transactions on Signal and Information Processing, Vol. 15, No. 1, pp. 183-198, Apr. 2026.
  3. Y. Hashizume, T. Toda, "Investigation of perceptual music similarity based on individual instrumental parts by large-scale listening test," APSIPA Transactions on Signal and Information Processing, Vol. 15, No. 1, pp. 249-269, Apr. 2026.
  4. J. Feng, Y. Yasuda, T. Toda, "An investigation of the robustness of flow- and diffusion-based speech generation models on noisy transcriptions," APSIPA Transactions on Signal and Information Processing, Vol. 15, No. 1, pp. 270-292, Apr. 2026.
  5. W.-C. Huang, E. Cooper, T. Toda, "MOS-Bench: benchmarking generalization abilities of subjective speech quality assessment models," IEEE Transactions on Audio, Speech and Language Processing, Vol. 34, pp. 2385-2397, Apr. 2026.
  6. X. Shi, X. Li, T. Toda, "Emotion similarity and shift: modeling temporal dynamic interactions for emotion prediction in conversation," IEEE Transactions on Audio, Speech and Language Processing, Vol. 34, pp. 2552-2567, Apr. 2026.
  7. R. Yoneyama, T. Toda, "SiFi-GAN: combining source-filter modeling and upsampling-based high-fidelity neural vocoder for fast and pitch-controllable speech synthesis," IEICE Trans. Inf. and Syst., Vol. E109-D, No. 6, pp. 945-956, June 2026.
 

国際会議

  1. T. Imamura, T. Komatsu, H. Munakata, T. Toda, "Audio-visual feature fusion for calibrating relevance scores of video moment retrieval," Proc. IEEE ICASSP, pp. 5551-5555, Barcelona, Spain, May 2026.
  2. L.P. Violeta, X. Zhang, J. Shi, Y. Yasuda, W.-C. Huang, Z. Wu, T. Toda, "The singing voice conversion challenge 2025: from singer identity conversion to singing style conversion," Proc. IEEE ICASSP, pp. 17707-17711, Barcelona, Spain, May 2026.
  3. J. Wang, T. Toda, "From fixed positions to free-form signals: Virtual Microphone signal estimation for general-purpose spatial audio processing," Proc. IEEE ICASSP, pp. 21011-21015, Barcelona, Spain, May 2026.
  4. H. Munakata, T. Imamura, T. Nishimura, T. Komatsu, "CASTELLA: long audio dataset with captions and temporal boundaries," Proc. IEEE ICASSP, pp. 15352-15356, Barcelona, Spain, May 2026.
  5. B. Xu, W. Zhang, D. Ma, J. Liang, Z. Sun, Z. Lei, "Modeling long-tail relations in the operating room via in-context multimodal learning," Proc. ICML, 10 pages, Seoul, South Korea, July 2026.
  6. H. Folkertsma, T. Tienkamp, S. de Visscher, M. Witjes, R. van Son, J. Guo, B.M. Halpern, "Improving automatic speech recognition for speakers treated for oral cancer using data augmentation and LLM error correction," Proc. EMBC, pp. 627-631, Tronto, Canada, July 2026.
 

研究会

  1. デ ポンテス ジェフェルソン マコト, ホワン ウェンチン, 戸田 智基, "モジュール分離型アーキテクチャによるオーディオエフェクト設定逆推定," 情報処理研報, Vol. 2026-MUS-146, No. 8, pp. 1-6, June 2026.【音学シンポジウム2026学生優秀発表賞(受賞者:デ ポンテス ジェフェルソン マコト)】
  2. 浪崎 恭佑, ホワン ウェンチン, 戸田 智基, "聴覚フィードバック音声制御に向けた体内伝導自己聴取音マスキングの調査," 情報処理研報, Vol. 2026-SLP-160, No. 44, pp. 1-6, June 2026.
  3. 齋藤 佑樹, HUANG Wen-Chin, 榎本 悠久, 今井 柊平, "Human Fooling Rateテストに基づく最先端の日本語テキスト音声合成モデルの評価および分析," 情報処理研報, Vol. 2026-MUS-160, No. 60, pp. 1-5, June 2026.
 

大会講演

  1. 中井 淳一, 藤村 拓弥, 浅野 憲司, 若松 智之, 戸田 智基, "説明性向上マルチモーダルAIによるMOCの異常予見~潜在的異常発見に向けたアテンションによる実験的分析~," 人工知能学会全国大会論文集, 1I3-GS-10f-02, 4 pages, June 2026.
 

その他発表

  1. S. Chen, T. Toda, "QHARMA-GAN: quasi-harmonic neural vocoder based on autoregressive moving average model," IEEE ICASSP, SPS journal paper presentation, Barcelona, Spain, May 2026.
  2. D. Ma, L.P. Violeta, K. Kobayashi, T. Toda, "Pretraining and fine-tuning techniques for electrolaryngeal speech enhancement based on sequence-to-sequence voice conversion," IEEE ICASSP, SPS journal paper presentation, Barcelona, Spain, May 2026.
  3. J. He, X. Shi, C.-H. Hu, J. Mi, X. Li, T. Toda, "M4SER: multimodal, multirepresentation, multitask, and multistrategy learning for speech emotion recognition," IEEE ICASSP, SPS journal paper presentation, Barcelona, Spain, May 2026.
  4. B.M. Halpern, T.B. Tienkamp, T. Rebernik, R.J.J.H. van Son, S.A.H.J. de Visscher, M.J.H. Witjes, D. Abur, T. Toda, "XPPG-PCA: reference-free automatic speech severity evaluation with principal components," IEEE ICASSP, SPS journal paper presentation, Barcelona, Spain, May 2026.
  5. T. Fujimura, G. Wichern, Y. Masuyama, C. Boeddeker, K. Saijo, J. Richter, T. Edo, and J. Le Roux, "The MERL systems for DCASE 2026 challenge task 2," Technical report, DCASE 2026 Challenge Task 2, 5 pages, July 2026.
  6. 戸田 智基, "名古屋大学キャリア講演会 情報学部コンピュータ科学科," 桑野高校, 三重, July 2026.
 


他の年度はこちら