Tomohiro Nakatani | NTT R&D Website
/center/dept Audio and speech signal processing for capturing and recognition of human conversations The
https://www.rd.ntt/e/organization/researcher/superior/s_012.html
Speech and Audio Signal Modeling | NTT Communication Science Laboratories | NTT R&D Website
Speech and Audio Signal Modeling | NTT Communication Science Laboratories | NTT R&D Website NTT R
https://www.rd.ntt/e/cs/team_project/media/recognition/research_media03.html
Computational Modeling Research Group | NTT Communication Science Laboratories | NTT R&D Website
/ACM Transactions on Audio, Speech, and Language Processing (IEEE/ACM TASLP), 32, 2213-2226. Daisuke
https://www.rd.ntt/e/cs/team_project/media/computational_modeling/
NTT Communication Science Laboratories Open House 2013
modeling approach to speech and audio signal processing - Hirokazu Kameoka, Media Information Laboratory
https://www.rd.ntt/cs/event/openhouse/2013/talk/research4/index_en.html
Hirokazu Kameoka | NTT R&D Website
signal processing, audio source separation, speech analysis, conversion, and synthesis, machine learning
https://www.rd.ntt/e/organization/researcher/superior/s_025.html
事象モデリング研究グループ|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Conversion with Annealed Langevin Dynamics. IEEE/ACM Transactions on Audio, Speech, and Language Processing
https://www.rd.ntt/cs/team_project/media/computational_modeling/
Media Information Laboratory | NTT Communication Science Laboratories | NTT R&D Website
Media Search Speech and Audio Signal Modeling High fidelity color reproduction and analysis Gallery of
https://www.rd.ntt/e/cs/team_project/media/
中谷 智広 | NTT R&D Website
~2014年12月 Member, Audio and Acoustics Technical Committee, IEEE Signal Processing Society 2011年1月~2012年
https://www.rd.ntt/organization/researcher/superior/s_012.html
メディア情報研究部|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Enhancement in Underdetermined Situations. EURASIP Journal on Audio, Speech, and Music Processing, 2024
https://www.rd.ntt/cs/team_project/media/
Kenta Niwa | NTT R&D Website
Splitting", IEEE/ACM Transactions on Audio, Speech, and Language Processing, 28, pp. 2036-2046, 2020. Yuma
https://www.rd.ntt/e/organization/researcher/special/s_061.html
Listening to human speech in noisy reverberant environments | NTT Communication Science Laboratories | NTT R&D Website
their outputs according to the statistical speech model. Unified framework of audio signal processing
https://www.rd.ntt/e/cs/team_project/media/signal/research_media06.html
Signal Processing Research Group | NTT Communication Science Laboratories | NTT R&D Website
Digital Signal Processing technology, we are researching speech recognition technology, acoustic signal
https://www.rd.ntt/e/cs/team_project/media/signal/
Recognition Research Group | NTT Communication Science Laboratories | NTT R&D Website
acoustic signals Media Search Speech and Audio Signal Modeling High fidelity color reproduction and
https://www.rd.ntt/e/cs/team_project/media/recognition/
WPE speech dereverberation
-normalized delayed linear prediction," IEEE Transactions on Audio, Speech, and Language Processing, vol. 18
https://www.rd.ntt/cs/team_project/media/signal/wpe/references.html
AI Hears Your Voice as if It Were Right Next to You-Audio Processing Framework for Separating Distant Sounds with Close-microphone Quality | NTT R&D Website
audio-signal processing As described above, our unified model provides theoretically and practically
https://www.rd.ntt/e/research/JN202208_19141.html
F04_leaf_e.pdf
audio signal mixture is called sound source separation, and it plays an important role in speech
https://www.rd.ntt/forum/2023/doc/F04_leaf_e.pdf
Hiroshi Sawada | NTT R&D Website
Processing(May 2013 - May 2014) Associate Editor, IEEE Transactions on Audio, Speech and Language Processing
https://www.rd.ntt/e/organization/researcher/superior/s_007.html
丹羽 健太 | NTT R&D Website
on Audio, Speech, and Language Processing, 28, pp. 2036-2046, 2020. Yuma Koizumi, Kenta Niwa, Yusuke
https://www.rd.ntt/organization/researcher/special/s_061.html
信号処理研究グループ|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Speech Enhancement in Underdetermined Situations. EURASIP Journal on Audio, Speech, and Music Processing
https://www.rd.ntt/cs/team_project/media/signal/
Speech Recognition for Computers | NTT Communication Science Laboratories | NTT R&D Website
Processing Research Group Signal Processing Research Group > Speech Recognition for Computers Speech
https://www.rd.ntt/e/cs/team_project/media/signal/research_media05.html
poster_en19.pdf
process,” in Proc. 38th International Conference on Acoustics, Speech, and Signal Processing (ICASSP
https://www.rd.ntt/cs/event/openhouse/2013/exhibition/media7/poster_en19.pdf
Media Information Laboratory Past news | NTT Communication Science Laboratories | NTT R&D Website
been accepted to IEEE Transactions on Audio, Speech and Language Processing. https://ieeexplore.ieee
https://www.rd.ntt/e/cs/team_project/media/past_news.html
Biomedical Informatics Research Group | NTT Communication Science Laboratories | NTT R&D Website
Audio, Speech, and Language Processing (IEEE/ACM TASLP), 32, 2391-2406. Peer-reviewed Conference Papers
https://www.rd.ntt/e/cs/team_project/media/biomedical_informatics/
スライド 1
Selected Topics in Signal Processing, 2019. Selective Hearing with Audio Speaker Clue Utilization of Audio
https://www.rd.ntt/cs/event/openhouse/2020/download/c_17_en.pdf
Developing AI that Pays Attention to Who You Want to Listen to: Deep-learning-based Selective Hearing with SpeakerBeam|NTT R&D Website
. It has been the goal of speech-processing researchers to reproduce a human’s selective hearing
https://www.rd.ntt/e/research/JN202107_14481.html
Dr.Takehiro Moriya | NTT R&D Website
Laboratories Research subject:Speech/audio signal processing and coding FellowMore Fellows Basic ResearchMore
https://www.rd.ntt/e/organization/researcher/fellow/f_002.html
program_for_web.pdf
WASPAA2007 2007 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics Mohonk
https://www.rd.ntt/cs/team_project/icl/signal/waspaa2007/program_for_web.pdf
paper_new4.dvi
new algorithm in the field of digital audio signal processing commonly con- sists of three development
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0015.pdf
スライド 1
-robust speech coding,” Proc. IEEE International Conference on Acoustics, Speech and Signal Processing
https://www.rd.ntt/cs/event/openhouse/2020/download/c_16_en.pdf
NTT Communication Science Laboratories Open House 2019
simulation of non-speech sounds, and (2) an sentence describing sounds, given an audio signal as an input
https://www.rd.ntt/cs/event/openhouse/2019/exhibition21/index_en.html
スライド 1
. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP2018), April 2018
https://www.rd.ntt/cs/event/openhouse/2019/download/20_c_en.pdf
メディア情報研究部 過去のニュース|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Recognition」が、IEEE Transactions on Audio, Speech and Language Processing に掲載されました。 https://ieeexplore.ieee.org
https://www.rd.ntt/cs/team_project/media/past_news.html
NTT Communication Science Laboratories Open House 2019
Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP2018), April 2018
https://www.rd.ntt/cs/event/openhouse/2019/exhibition20/index_en.html
澤田 宏 | NTT R&D Website
(2013年5月~2014年5月) Associate Editor, IEEE Transactions on Audio, Speech and Language Processing(2006年2月
https://www.rd.ntt/organization/researcher/superior/s_007.html
NTT Communication Science Laboratories Open House 2020
speaker extraction in speech mixtures,” IEEE Journal of Selected Topics in Signal Processing, 2019. Poster
https://www.rd.ntt/cs/event/openhouse/2020/exhibition17/index_en.html
生体情報処理研究グループ|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Framework. IEEE/ACM Transactions on Audio, Speech, and Language Processing (IEEE/ACM TASLP), 32, 2391-2406
https://www.rd.ntt/cs/team_project/media/biomedical_informatics/
NTT Communication Science Laboratories Open House 2020
. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2018. Poster Please
https://www.rd.ntt/cs/event/openhouse/2020/exhibition16/index_en.html
スライド 1
simulation of non-speech sounds, and (2) an sentence describing sounds, given an audio signal as an input
https://www.rd.ntt/cs/event/openhouse/2019/download/21_c_en.pdf
poster_en_18.pdf
speech and audio through Internet Protocol (IP) telephones, for example, in radio broadcasting. The near
https://www.rd.ntt/cs/event/openhouse/2017/exhibition/18/poster_en_18.pdf
Program of WASPAA2007
Makino 8:00-9:00 [Keynote1] Coherent ICA: Implications for Auditory Signal Processing Simon Haykin
https://www.rd.ntt/cs/team_project/icl/signal/waspaa2007/program.html
NTT Communication Science Laboratories Open House 2014
utilized in actual market, and how we expect them to open up new vistas for audio signal processing. Photos
https://www.rd.ntt/cs/event/openhouse/2014/talk/research2/index_en.html
雑音・残響の中で聞き取る―超高品質音声強調|NTTコミュニケーション科学基礎研究所|NTT R&D Website
separation," IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol. 28, pp. 2267-2282, 2020. N
https://www.rd.ntt/cs/team_project/media/signal/research_signal01.html
メディア認識研究グループ|NTTコミュニケーション科学基礎研究所|NTT R&D Website
Conference on Acoustics, Speech and Signal Processing (ICASSP). Seoul, Korea. Takuhiro kaneko (2024
https://www.rd.ntt/cs/team_project/media/recognition/
亀岡 弘和 | NTT R&D Website
Acoustics, Speech and Signal Processing (ICASSP2005) 2005 第20回 電気通信普及財団 テレコムシステム技術学生賞 2005 情報処理学会 平成17年度 山下
https://www.rd.ntt/organization/researcher/superior/s_025.html
Main Topics | NTT Communication Science Laboratories | NTT R&D Website
. Buru-Navi Sports Brain Science Media Information Science Media Search Speech and Audio Signal Modeling
https://www.rd.ntt/e/cs/research_topic/
Abstracts of plenary talks
dissemination of mobile communications, speech processing systems must be made robust with respect to
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/plenary_abst.html
Undertaking Research on Audio Source Separation and a Method for Training Hardware-oriented Neural Networks by Using Algorithms That Exploit Correlations and Complex Numbers | NTT R&D Website
Signal Processing (ICASSP) 2022, and the paper was published in IEEE/ACM Transactions on Audio, Speech
https://www.rd.ntt/e/research/JN202503_32652.html
articleICA03.dvi
used in [14] or the video of the lips corresponding to a noisy speech signal used in [8]. 3.2. Audio
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0085.pdf
NTT Communication Science Laboratories Open House 2013 Transcribing known and unknown sounds - Bayesian semi-supervised audio event recognition -
process,” in Proc. 38th International Conference on Acoustics, Speech, and Signal Processing (ICASSP) 2013
https://www.rd.ntt/cs/event/openhouse/2013/exhibition/media7/index_en.html
0085.pdf
,” IEEE Trans. Acoustics, Speech and Signal Processing, vol. 27, pp. 113–120, 1979. [2] M. Berouti, R
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0085.pdf
0128.pdf
Using LP Residual Signal” IEEE Transactions on Speech and Audio Processing, Vol. 8, No. 3, May 2000, pp
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0128.pdf
kalman.dvi
corrupted speech signal with an additive noise is the only information available for processing. Kalman
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0012.pdf
0055.pdf
Application of Signal Processing on Audio and Acoustics, New York, 2001. 4. Prasad.R.K, H.Saruwatari, A.Lee, K
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0055.pdf
poster_en_17.pdf
intelligible native-like speech. We take a signal processing-based approach using our recently developed model
https://www.rd.ntt/cs/event/openhouse/2017/exhibition/17/poster_en_17.pdf
Augmenting Communication Capabilities with Cutting-edge Voice-conversion Technology that Enables Users to Freely Customize the Impression of Their Voice | NTT R&D Website
University of Tokyo. His research interests include audio and speech processing and machine learning. He has
https://www.rd.ntt/e/research/JN202511_37040.html
Researchers Are the Source of Social Progress. Be Confident in Your Research Theme and Aim for the Next Big Topic | NTT R&D Website
research on speech- and audio-signal coding for about 40 years to improve and innovate communication and
https://www.rd.ntt/e/research/JN202302_20970.html
DemoSessionProgram.pdf
rendering. All the different features of digital audio signal processing will be introduced and practically
https://www.rd.ntt/cs/team_project/icl/signal/waspaa2007/DemoSessionProgram.pdf
WPE speech dereverberation
processing in utterance batch mode. When processing long audio files, you can use block batch processing by
https://www.rd.ntt/cs/team_project/media/signal/wpe/config.html
Science and Technology Are the Collective Wisdom of Our Predecessors. It Is Our Mission-the Researchers of Today-to Make Them Even Better | NTT R&D Website
international conferences, such as International Conference on Acoustics, Speech, and Signal Processing (ICASSP
https://www.rd.ntt/e/research/JN202305_21819.html
NTT Communication Science Laboratories Open House 2013 Schedule
- Generative modeling approach to speech and audio signal processing - Hirokazu Kameoka, Media Information
https://www.rd.ntt/cs/event/openhouse/2013/schedule_en.html
IWAENC2003 Program
-03] On the Application of the Unscented Kalman Filter to Speech Processing Sharon Gannot
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/preliminaryprogram.html
"Speech Enhancement based on Linear Prediction Error Signals and Spectral Subtraction"
., “Enhancement of Reverberant Speech Using LP Residual Signal”, IEEE Trans, on Speech and Audio Processing, Vol
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0077.pdf
main.dvi
Masahiro FURUKAWA, Yusuke HIOKA, Takuro EMA and Nozomu HAMADA Signal Processing Lab., School of Integrated
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0045.pdf
NTT Communication Science Laboratories Open House 2013 Program
information from speech and audio signals - Generative modeling approach to speech and audio signal processing
https://www.rd.ntt/cs/event/openhouse/2013/program_en.html
Speech information processing | NTT R&D Website
processing technology handles speech as signal data for analysis and processing, such as recognition and
https://www.rd.ntt/e/hil/category/voice/
IWAENC 2003 Online Proceedings
Rainer Martin, pp.1-6, [PDF] [T-2] The Automatic DJ: An Appealing and Instructive Signal Processing
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/index.htm
NTT Communication Science Laboratories Open House 2016 Program
deep learning and signal processing that are making speech recognition leap forward - Takuya Yoshioka
https://www.rd.ntt/cs/event/openhouse/2016/program_en.html
Noboru Harada | NTT R&D Website
▶ To the interview Speech/audio signal processing, coding, and standardization We aim to provide
https://www.rd.ntt/e/organization/researcher/superior/s_031.html
音声情報処理 | NTT R&D Website
-based ASR" , In Proc. International Conference on Acoustics, Speech, and Signal Processing (ICASSP
https://www.rd.ntt/hil/category/voice/
NTT Communication Science Laboratories Open House 2020
conversion,” IEEE/ACM Transactions on Audio, Speech, and Language Processing, under review. K. Tanaka, H
https://www.rd.ntt/cs/event/openhouse/2020/exhibition18/index_en.html
スライド 1
multichannel audio signals,” in Proc. International Conference on Acoustics, Speech, and Signal Processing
https://www.rd.ntt/cs/event/openhouse/2019/download/22_c_en.pdf
Microsoft Word - timetable.doc
Albert S. Bregman Conference House 9:00 10:00 Lecture ML1 Microphone Array Signal Processing Conference
https://www.rd.ntt/cs/team_project/icl/signal/waspaa2007/Timetable.pdf
hosica03.dvi
identification using gaussian mixture speaker model,” IEEE Trans. on Speech and Audio Processing, vol. 3, no. 1
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0213.pdf
NTT Communication Science Laboratories Open House 2019
segmentation results from multichannel audio signals,” in Proc. International Conference on Acoustics, Speech
https://www.rd.ntt/cs/event/openhouse/2019/exhibition22/index_en.html
NTT Communication Science Laboratories Open House 2016 Schedule
super-human speech recognizers - Advances in deep learning and signal processing that are making speech
https://www.rd.ntt/cs/event/openhouse/2016/schedule_en.html
0064.pdf
. Kellermann, “Adaptive beamforming for audio signal acquisition,” in Adaptive signal processing: Application
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0064.pdf
WPE speech dereverberation
," IEEE Transactions on Audio, Speech, and Language Processing, vol. 18, no. 7, pp. 1717-1731, Sep. 2010
https://www.rd.ntt/cs/team_project/media/signal/wpe/
proc_toc.pdf
by Shoji Makino and Masato Miyoshi IEEE Japan Council IEEE Kansai Section IEEE Signal Processing
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/proc_toc.pdf
0008.pdf
spectral amplitude estima- tor,” IEEE Trans. Acoustics, Speech and Signal Processing, vol. 32, pp. 1109
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0008.pdf
ica2003.dvi
mix- ing, e.g., in processing of audio signal [25]. This paper will not provide a detailed discussion
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0214.pdf
0082.pdf
,” IEEE Trans. Acous- tics, Speech and Signal Processing, vol. 28, pp. 137–145, December 1980. [2] Y
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0082.pdf
スライド 1
Transactions on Audio, Speech, and Language Processing, in Review. [2] K. Tanaka, H. Kameoka, T. Kaneko, N
https://www.rd.ntt/cs/event/openhouse/2020/download/c_18_en.pdf
0009.pdf
Advanced Brain Signal Processing, BSI RIKEN, 2-1, Hirosawa, Wakoh-City, Saitama 351-0198, Japan
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0009.pdf
ICA2003em.dvi
applications in speech processing, wireless communications, biomedical signal processing exist and re- cently
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0116.pdf
NTT R&D Website | NTT, Inc.
Basic research Fellow:Dr.Takehiro Moriya Research subject:Speech/audio signal processing and coding
https://www.rd.ntt/e/
Microsoft Word - 3E3495A9-54B4-18FBD3.doc
convolved audio source separation,” Proc. IEEE Workshop on Application of Signal Processing on Audio and
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0082.pdf
聴きたい音を聴く―選択的聴取|NTTコミュニケーション科学基礎研究所|NTT R&D Website
extraction in speech mixtures," IEEE Journal of Selected Topics in Signal Processing, vol. 13, no. 4, pp. 800
https://www.rd.ntt/cs/team_project/media/signal/research_signal02.html
2019_booklet_english.pdf
autoencoders,” in Proc. of 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing
https://www.rd.ntt/cs/event/openhouse/2019/download/2019_booklet_english.pdf
あなたの声を「すぐそば」品質で聴くAI ――遠くからでも近接マイク品質で混ざった音を聞き分ける革新的音響処理技術 | NTT R&D Website
prediction,” IEEE Trans. on Audio, Speech, and Language Processing, Vol. 18, No. 7, pp. 1717-1731, 2010. (2)N
https://www.rd.ntt/research/JN202208_19141.html
main.dvi
Acoustics, Speech and Signal Processing, Vol.33, No.4, pp.823-831, 1985. [3] M. Omologo and P. Svaizer, “Use
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0021.pdf
Pursuing Elegance and Creating a Research-goal Umbrella | NTT R&D Website
solving that problem based on a statistical signal-processing approach[2]. Fig. 1. Process of generating a
https://www.rd.ntt/e/research/JN202008_6056.html
ica_v11.dvi
in advanced statistics and signal processing, and it ap- plies to major fields such as audio
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0118.pdf
0067.pdf
, “Blind signal separation using overcomplete subband representation,” IEEE Trans. Speech Audio Process
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0067.pdf
0001.pdf
ON THE APPLICATION OF THE UNSCENTED KALMAN FILTER TO SPEECH PROCESSING Sharon Gannot Faculty of
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0001.pdf
0005.pdf
to speech,” IEEE Trans. Signal Processing, vol. 49, no. 8, pp. 1614–1626, August 2001. [3] B. Widrow
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0005.pdf
0037.pdf
ABSTRACT A speech enhancement scheme is presented integrating spatial and temporal signal processing
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0037.pdf
ica2003_revise.doc
(ICA) processing has become one of the hottest and emerging areas in signal processing with solid
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0023.pdf
NTT コミュニケーション科学基礎研究所 オープンハウス2020
on Acoustics, Speech and Signal Processing (ICASSP2019), pp. 6805-6809, 2019. ポスター アイコンをクリックすると、展示ポス
https://www.rd.ntt/cs/event/openhouse/2020/exhibition18/
Microsoft Word - ica_new.doc
on Oral Air Flow During Phonation”, IEEE Trans. on Acoustics, Speech, and Signal Processing, Vol.ASSP
https://www.rd.ntt/cs/team_project/icl/signal/ica2003/cdrom/data/0182.pdf
0074.pdf
Transactions on Speech and Audio Processing, vol. 9, no. 8, pp. 799–807, November 2001. [3] S. Gollamudi, S
https://www.rd.ntt/cs/team_project/icl/signal/iwaenc03/cdrom/data/0074.pdf