肖仲喆,女,博士,副教授,硕士生导师。
2008年毕业于法国里昂中央理工大学,获博士学位。2010年到苏州大学工作。
到苏州大学工作以来,承担国家自然科学基金青年基金项目、江苏省自然科学基金项目、江苏省高校自然科学基金等多项纵向科研课题,在国际学术期刊、学术会议等发表论文30余篇。
科研方向主要集中在音频信号处理方面,尤其是基于人工智能方法对语音/音乐进行情感识别,以及对其他通用音频信号的事件智能检测等。
1. 情感语音研究
研究语音与情感的关系,包括语音情感模型,语音参数,语音情感识别模型,以及语音情感合成/转换模型等。创新性的提出了结合发声生理过程挖掘语音情感本质特性的研究思想,通过对声道、声带特性的分析,体现语音情感细节本质。
2. 音乐计算
以音乐情绪表达为核心的音乐计算,通过旋律、节奏等音乐特征以及各种感知特征进行音乐情绪识别。具体研究对象主要为中国传统民族乐器。
3. 通用音频信号事件检测
对通过音频信号体现的事件/场景进行智能检测/分析。
主要主讲课程:
1. 信号与线性系统
2. 微弱信号检测
3. 专业英语
参与主讲课程:
数字信号处理
国际学术期刊论文:
· Ying Chen, Zhongzhe Xiao, Xiaojun Zhang, Zhi Tao, DSTL: Solution to Limitation of Small Corpus in Speech Emotion Recognition, accepted by Journal of Articial Intelligence Research, vol. 65, 2019.
· Zhongzhe Xiao, Emmanuel Dellandréa,Weibei Dou, Liming Chen, Classification of Emotional Speech Based on anAutomatically Elaborated Hierarchical Classifier, ISRN signal processing, vol.2011, Article ID 753819, 15 pages, 2011. doi:10.5402/2011/753819.
· Zhongzhe Xiao, Emmanuel Dellandréa,Weibei Dou, Liming Chen, Multi-stage classification of emotional speechmotivated by a dimensional emotion model, Multimedia Tools and Applications,Volume 46, Issue 1 (2010), Page 119-145.
国际会议论文:
zZhongzhe Xiao, Ying Chen, Zhi Tao, Proceedings of the 2018 IEEE International Conference on Progress in Informatics and Computing (PIC2018), Suzhou, China, 2018, Dec, pp.185-190.
B. Sun, X. Zhang, Y. Wang, D. Wu, Z. Xiao and Z. Tao, A Hybrid Coding Strategy to Improve Auditory Perception of Cochlear Implant, 2018 5th International Conference on Systems and Informatics (ICSAI), Nanjing, China, 2018, pp. 845-850.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, Speech Emotion Recognition Cross Language Families:Mandarin vs. Western Languages, in the Proceedings of the 2016 IEEEInternational Conference on Progress in Informatics and Computing (PIC2016),December 23-25, 2016, Shanghai, China, p253-257.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, A Cross-Corpus Recognition of Emotional Speech,Proceedings of 2016 9th International Symposium on Computational Intelligenceand Design (ISCID 2016), vol 2, p42-46, Dec. 10-11, 2016 in Hangzhou,China
· Di Wu,Zhi Tao, Yuanbo Wu, Cheng Shen, Zhongzhe Xiao, Xiaojun Zhang, Di Wu,Heming Zhao, Speech endpoint detection in noisy environment using SpectrogramBoundary Factor, 2016 9th International Congress on Image and SignalProcessing, BioMedical Engineering and Informatics (CISP-BMEI), p964-968,2016.10.15-17, Datong, P. R. China.
· Di Wu,Heming Zhao, Zhe Feng, Liyuan Chen, Zhongzhe Xiao, Xiaojun Zhang, ZhiTao, Perception auditory factor for speaker recognition in noisy environment,2016 12th International Conference on Natural Computation, Fuzzy Systems andKnowledge Discovery (ICNC-FSKD), p1916-1920, 2016.8.13-15, Changsha, P. R.China.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, Music Mood Tracking Based on HCS, Proceedings 2012 IEEE11th International Conference on Signal Processing, p. 1171-1175, October 21 -25, 2012, Beijing,
· Bingjie Li, Zhongzhe Xiao,Yan Shen, Qiang Zhou, Zhi Tao*, Emotional Speech Conversion Based onSpectrum-prosody Dual Transformation, Proceedings 2012 IEEE 11th InternationalConference on Signal Processing, p. 531-535, October 21 - 25, 2012, Beijing
· Huanzhang Fu, Zhongzhe Xiao,Emmanuel Dellandréa, Liming Chen, Image Categorization using ESFS: a newembedded Feature Selection Method Based on Evidence Theory, Intl. conf.Advanced Concepts for Intelligent Vision Systems (ACIVS 2009), Bordeaux,France, September 2009
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, What is the Best Segment Duration forMusic Mood Analysis? Sixth International Workshop on Content-Based MultimediaIndexing (CBMI 2008), June 2008, London, UK.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Ambiguous classification of emotionalspeech, International Workshop on EMOTION - satellite of InternationalConference on Language Resources and Evaluation (LREC), 2008.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Automatic Hierarchical Classificationof Emotional Speech,Ninth IEEE International Symposium on Multimedia Workshops(ISMW 2007), Taichung, Taiwan. pp. 291-296.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Two-stage Classification of EmotionalSpeech, International Conference on Digital Telecommunications (ICDT'06), p.32-37, August 29 - 31, 2006, Cap Esterel, C?te d’Azur, France.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Features extraction and selection foremotional speech classification, IEEE Conference on Advanced Video and SignalBased Surveillance, 2005. AVSS 2005, p411 - 416,Como, Italy
· Zhongzhe Xiao, ZaiwangDong, Improved GIB synchronization method for OFDM systems, 10th InternationalConference on Telecommunications, 2003. ICT 2003, p1417 - 1421,vol.2
国内期刊论文:
陈颖,肖仲喆,离散标签与维度空间结合的语音数据库设计,声学技术,2018,37(04):380-387.
肖仲喆,吴迪,张晓俊,陶智,面向情感分析的古筝乐曲声谱图信息主干提取,复旦学报(自然科学版),2017,(2):191-199+205.
沈燕,肖仲喆,李冰洁,周孝进,周强,陶智,采用GW-MFCC模型空间参数的语音情感识别. 计算机工程与应用,2015,51(10):219-222+226.
黄程韦,吴迪,张晓俊,肖仲喆,许宜申,季晶晶,陶智,赵力,基于级联投影高斯混合模型的语音与心电情绪识别. 东南大学学报:英文版,2015,(3): 320-326.
吴迪,赵鹤鸣,陶智,张晓俊,肖仲喆,许宜申,低信噪比下采用感知语谱结构边界参数的语音端点检测算法,声学学报,2014(3): 392-399.
张晓俊,陶智,吴迪,肖仲喆,赵鹤鸣,采用多特征组合优化的语音特征参数研究,通信技术,2012,45(12): 98-100.
授权专利:
一种基于声音信号情感识别的自动选台器,实用新型专利,专利号:ZL 2017 2 1194234.X
欢迎电子、测控、信号处理、计算机等背景的同学报考本组的硕士生(考研、保研都欢迎)。同时,欢迎本校本科生参与本组的本科生导师制科研活动。
联系方式:xiaozhongzhe@suda.edu.cn,13913520247。
肖仲喆,女,博士,副教授,硕士生导师。
2008年毕业于法国里昂中央理工大学,获博士学位。2010年到苏州大学工作。
到苏州大学工作以来,承担国家自然科学基金青年基金项目、江苏省自然科学基金项目、江苏省高校自然科学基金等多项纵向科研课题,在国际学术期刊、学术会议等发表论文30余篇。
科研方向主要集中在音频信号处理方面,尤其是基于人工智能方法对语音/音乐进行情感识别,以及对其他通用音频信号的事件智能检测等。
1. 情感语音研究
研究语音与情感的关系,包括语音情感模型,语音参数,语音情感识别模型,以及语音情感合成/转换模型等。创新性的提出了结合发声生理过程挖掘语音情感本质特性的研究思想,通过对声道、声带特性的分析,体现语音情感细节本质。
2. 音乐计算
以音乐情绪表达为核心的音乐计算,通过旋律、节奏等音乐特征以及各种感知特征进行音乐情绪识别。具体研究对象主要为中国传统民族乐器。
3. 通用音频信号事件检测
对通过音频信号体现的事件/场景进行智能检测/分析。
主要主讲课程:
1. 信号与线性系统
2. 微弱信号检测
3. 专业英语
参与主讲课程:
数字信号处理
国际学术期刊论文:
· Ying Chen, Zhongzhe Xiao, Xiaojun Zhang, Zhi Tao, DSTL: Solution to Limitation of Small Corpus in Speech Emotion Recognition, accepted by Journal of Articial Intelligence Research, vol. 65, 2019.
· Zhongzhe Xiao, Emmanuel Dellandréa,Weibei Dou, Liming Chen, Classification of Emotional Speech Based on anAutomatically Elaborated Hierarchical Classifier, ISRN signal processing, vol.2011, Article ID 753819, 15 pages, 2011. doi:10.5402/2011/753819.
· Zhongzhe Xiao, Emmanuel Dellandréa,Weibei Dou, Liming Chen, Multi-stage classification of emotional speechmotivated by a dimensional emotion model, Multimedia Tools and Applications,Volume 46, Issue 1 (2010), Page 119-145.
国际会议论文:
zZhongzhe Xiao, Ying Chen, Zhi Tao, Proceedings of the 2018 IEEE International Conference on Progress in Informatics and Computing (PIC2018), Suzhou, China, 2018, Dec, pp.185-190.
B. Sun, X. Zhang, Y. Wang, D. Wu, Z. Xiao and Z. Tao, A Hybrid Coding Strategy to Improve Auditory Perception of Cochlear Implant, 2018 5th International Conference on Systems and Informatics (ICSAI), Nanjing, China, 2018, pp. 845-850.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, Speech Emotion Recognition Cross Language Families:Mandarin vs. Western Languages, in the Proceedings of the 2016 IEEEInternational Conference on Progress in Informatics and Computing (PIC2016),December 23-25, 2016, Shanghai, China, p253-257.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, A Cross-Corpus Recognition of Emotional Speech,Proceedings of 2016 9th International Symposium on Computational Intelligenceand Design (ISCID 2016), vol 2, p42-46, Dec. 10-11, 2016 in Hangzhou,China
· Di Wu,Zhi Tao, Yuanbo Wu, Cheng Shen, Zhongzhe Xiao, Xiaojun Zhang, Di Wu,Heming Zhao, Speech endpoint detection in noisy environment using SpectrogramBoundary Factor, 2016 9th International Congress on Image and SignalProcessing, BioMedical Engineering and Informatics (CISP-BMEI), p964-968,2016.10.15-17, Datong, P. R. China.
· Di Wu,Heming Zhao, Zhe Feng, Liyuan Chen, Zhongzhe Xiao, Xiaojun Zhang, ZhiTao, Perception auditory factor for speaker recognition in noisy environment,2016 12th International Conference on Natural Computation, Fuzzy Systems andKnowledge Discovery (ICNC-FSKD), p1916-1920, 2016.8.13-15, Changsha, P. R.China.
· Zhongzhe Xiao, Di Wu,Xiaojun Zhang, Zhi Tao, Music Mood Tracking Based on HCS, Proceedings 2012 IEEE11th International Conference on Signal Processing, p. 1171-1175, October 21 -25, 2012, Beijing,
· Bingjie Li, Zhongzhe Xiao,Yan Shen, Qiang Zhou, Zhi Tao*, Emotional Speech Conversion Based onSpectrum-prosody Dual Transformation, Proceedings 2012 IEEE 11th InternationalConference on Signal Processing, p. 531-535, October 21 - 25, 2012, Beijing
· Huanzhang Fu, Zhongzhe Xiao,Emmanuel Dellandréa, Liming Chen, Image Categorization using ESFS: a newembedded Feature Selection Method Based on Evidence Theory, Intl. conf.Advanced Concepts for Intelligent Vision Systems (ACIVS 2009), Bordeaux,France, September 2009
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, What is the Best Segment Duration forMusic Mood Analysis? Sixth International Workshop on Content-Based MultimediaIndexing (CBMI 2008), June 2008, London, UK.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Ambiguous classification of emotionalspeech, International Workshop on EMOTION - satellite of InternationalConference on Language Resources and Evaluation (LREC), 2008.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Automatic Hierarchical Classificationof Emotional Speech,Ninth IEEE International Symposium on Multimedia Workshops(ISMW 2007), Taichung, Taiwan. pp. 291-296.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Two-stage Classification of EmotionalSpeech, International Conference on Digital Telecommunications (ICDT'06), p.32-37, August 29 - 31, 2006, Cap Esterel, C?te d’Azur, France.
· Zhongzhe Xiao, EmmanuelDellandréa, Weibei Dou, and Liming Chen, Features extraction and selection foremotional speech classification, IEEE Conference on Advanced Video and SignalBased Surveillance, 2005. AVSS 2005, p411 - 416,Como, Italy
· Zhongzhe Xiao, ZaiwangDong, Improved GIB synchronization method for OFDM systems, 10th InternationalConference on Telecommunications, 2003. ICT 2003, p1417 - 1421,vol.2
国内期刊论文:
陈颖,肖仲喆,离散标签与维度空间结合的语音数据库设计,声学技术,2018,37(04):380-387.
肖仲喆,吴迪,张晓俊,陶智,面向情感分析的古筝乐曲声谱图信息主干提取,复旦学报(自然科学版),2017,(2):191-199+205.
沈燕,肖仲喆,李冰洁,周孝进,周强,陶智,采用GW-MFCC模型空间参数的语音情感识别. 计算机工程与应用,2015,51(10):219-222+226.
黄程韦,吴迪,张晓俊,肖仲喆,许宜申,季晶晶,陶智,赵力,基于级联投影高斯混合模型的语音与心电情绪识别. 东南大学学报:英文版,2015,(3): 320-326.
吴迪,赵鹤鸣,陶智,张晓俊,肖仲喆,许宜申,低信噪比下采用感知语谱结构边界参数的语音端点检测算法,声学学报,2014(3): 392-399.
张晓俊,陶智,吴迪,肖仲喆,赵鹤鸣,采用多特征组合优化的语音特征参数研究,通信技术,2012,45(12): 98-100.
授权专利:
一种基于声音信号情感识别的自动选台器,实用新型专利,专利号:ZL 2017 2 1194234.X
欢迎电子、测控、信号处理、计算机等背景的同学报考本组的硕士生(考研、保研都欢迎)。同时,欢迎本校本科生参与本组的本科生导师制科研活动。
联系方式:xiaozhongzhe@suda.edu.cn,13913520247。
