F0 contour approximation model for a one-stream tonal word recognition system

บทความในวารสาร

ผู้เขียน/บรรณาธิการ

โกสินทร์ จำนงไทย

กลุ่มสาขาการวิจัยเชิงกลยุทธ์

ไม่พบข้อมูลที่เกี่ยวข้อง

รายละเอียดสำหรับงานพิมพ์

รายชื่อผู้แต่ง: Prukkanon N., Chamnongthai K., Miyanaga Y.

ผู้เผยแพร่: Elsevier

ปีที่เผยแพร่ (ค.ศ.): 2016

วารสาร: International Journal of Electronics and Communications (1434-8411)

Volume number: 70

Issue number: 5

หน้าแรก: 681

หน้าสุดท้าย: 688

จำนวนหน้า: 8

นอก: 1434-8411

URL: https://www.scopus.com/inward/record.uri?eid=2-s2.0-84959473571&doi=10.1016%2fj.aeue.2016.02.006&partnerID=40&md5=4f2814d372270b9e09d4aa1203d5fb79

ภาษา: English-Great Britain (EN-GB)

ดูในเว็บของวิทยาศาสตร์ | ดูบนเว็บไซต์ของสำนักพิมพ์ | บทความในเว็บของวิทยาศาสตร์

บทคัดย่อ

The performance of a non-tonal speech recognition system degrades when confronted with the task of recognizing tonal words. Several speech recognition applications require tonal word recognition. Therefore, this paper considers how to create a suitable tone model for a tonal syllable recognition system serving application devices based on a one-stream scheme. The fundamental frequency contour (F0 contour) approximation model is proposed here to estimate F0 continuity contours for all of a tonal word. The processes of approximation include voice detection, F0 smoothing, F0 forecasting, and F0 normalization. To model the F0 contours of unvoiced regions belonging to F0 forecasting, a linear regression function is used to create an approximate F0 contour. Experimental results indicate that the proposed model improves the accuracy of tonal word recognition by 8.6% and 12.2%, respectively, compared with conventional random and exponential approaches. ฉ 2016 Elsevier GmbH.

คำสำคัญ

F0 contour approximation, Fundamental frequency, Tonal syllable recognition, Tone model