การระบุและแก้ไขส่วนที่แตกต่างของบทถอดความระหว่างเสียงบันทึกการประชุมรัฐสภาไทยและรายงานการประชุม

ณัฐณรงค์ พ่วงศรี

Please use this identifier to cite or link to this item: https://cuir.car.chula.ac.th/handle/123456789/37618

Full metadata record

DC Field	Value	Language
dc.contributor.advisor	อติวงศ์ สุชาโต	-
dc.contributor.advisor	โปรดปราน บุณยพุกกณะ	-
dc.contributor.advisor	ชัย วุฒิวิวัฒน์ชัย	-
dc.contributor.author	ณัฐณรงค์ พ่วงศรี	-
dc.contributor.other	จุฬาลงกรณ์มหาวิทยาลัย. คณะวิศวกรรมศาสตร์	-
dc.coverage.spatial	ไทย	-
dc.date.accessioned	2013-12-31T14:39:51Z	-
dc.date.available	2013-12-31T14:39:51Z	-
dc.date.issued	2555	-
dc.identifier.uri	http://cuir.car.chula.ac.th/handle/123456789/37618	-
dc.description	วิทยานิพนธ์ (วศ.ม.)--จุฬาลงกรณ์มหาวิทยาลัย, 2555	en_US
dc.description.abstract	ข้อมูลเสียงพูด (Speech utterance) และคำบรรยายเสียง (Transcription) ที่มีความถูกต้องเป็นส่วนสำคัญที่ใช้ ในการพัฒนาระบบรู้จำเสียงพูดอัตโนมัติ (Automatic speech recognition) โดยเฉพาะอย่างยิ่งกับระบบที่นำไปใช้ในการถอดความการประชุมรัฐสภา สำหรับในประเทศไทยนั้น สำนักงานเลขาธิการสภาผู้แทนราษฏร ได้จัดทำรายงานการประชุมและจัดเก็บข้อมูลเสียงบันทึกระหว่างการประชุมไว้ตลอดช่วงสมัยประชุม ทำให้มีข้อมูลดังกล่าวเป็นจำนวนมากเพียงพอที่จะนำมาใช้ในการพัฒนาระบบรู้จำเสียงพูดอัตโนมัติอย่างไรก็ตามเนื่องจากข้อมูลทั้งสองส่วนยังมีความไม่สอดคล้องกันเกิดขึ้นในบางจุด ดังนั้น วิทยานิพนธ์นี้จึงนำเสนอวิธีในการระบุส่วนที่แตกต่างกันที่เกิดขึ้น กฎที่ได้จากการวิเคราะห์ หลักเกณฑ์การจัดทำรายงานการประชุมสภา และส่วนที่แตกต่างกันที่เกิดขึ้นจริงถูกนำมาใช้วิเคราะห์ประโยคจากรายงานการประชุม เพื่อสร้างประโยคสมมติฐานขึ้นมาเพิ่มเติม จากนั้น ประโยคจากรายงานการประชุมและประโยคสมมติฐานจะถูกนำไปผ่านกระบวนการปรับแนวเสียง (Force alignment) เพื่อประเมินความน่าจะเป็นของแต่ละประโยคซึ่งประโยคที่มีความน่าจะเป็นสูงที่สุด จะถูกเลือกเป็นคำบรรยายเสียงสำหรับข้อมูลเสียงพูดสำหรับใช้ใน กระบวนการระบุส่วนที่ไม่ตรงกัน จากการทดลองพบว่าระบบที่พัฒนาขึ้น มีค่าความแม่นยำในการระบุส่วนที่แตกต่างกัน 72.6% และคำบรรยายเสียงที่ได้จากประโยคที่มีความน่าจะเป็นสูงที่สุด มีความถูกต้องตรงกับข้อมูลเสียงพูดในระดับหน่วยเสียงย่อ 96.5% โดยเมื่อเปรียบเทียบกับคำบรรยายเสียงที่ได้จากรายงานการประชุมพบว่า สามารถลดความไม่ตรงกันได้ถึง 26.8%	en_US
dc.description.abstractalternative	Speech utterance and their accurate transcriptions are essential to train acoustic models of modern automatic speech recognition (ASR) especially for transcribing parliament meeting speech. In Thai, there are many speech data and their official meeting reports sufficient for developing good acoustic models. However, most of existing reports are not consistent with their corresponding utterances because of discrepancies. This article proposes a method for automatically detecting locations of the discrepancies. A process to generate alternative hypotheses supplied to a forced-alignment procedure can be done by applying rules derived from the standard transcript guidelines for Thai parliament stenographer and patterns of discrepancies to texts obtained from the reports. The forced-alignment procedure selects the best hypothesis to be the word-for-word transcription for each speech utterance. The accuracy to detect syllabic discrepancies is 72.6% while the accuracy to falsely detect correct syllables is kept minimal. With the proposed method, the word-for-word phonemic transcription accuracy of 96.5% is achieved due to the transcription error rate of word-for-word phonemic transcription from the best hypothesis is relatively reduced 26.8% compared to the transcription from official meeting report.	en_US
dc.language.iso	th	en_US
dc.publisher	จุฬาลงกรณ์มหาวิทยาลัย	en_US
dc.relation.uri	http://doi.org/10.14457/CU.the.2012.1171	-
dc.rights	จุฬาลงกรณ์มหาวิทยาลัย	en_US
dc.subject	การรู้จำเสียงพูดอัตโนมัติ	en_US
dc.subject	การประชุมรัฐสภา	en_US
dc.subject	Automatic speech recognition	en_US
dc.subject	Legislative bodies -- Thailand	en_US
dc.title	การระบุและแก้ไขส่วนที่แตกต่างของบทถอดความระหว่างเสียงบันทึกการประชุมรัฐสภาไทยและรายงานการประชุม	en_US
dc.title.alternative	Detecting and correcting transcription discrepancies between Thai parliament meeting speech utterances and their official meeting reports	en_US
dc.type	Thesis	en_US
dc.degree.name	วิศวกรรมศาสตรมหาบัณฑิต	en_US
dc.degree.level	ปริญญาโท	en_US
dc.degree.discipline	วิศวกรรมคอมพิวเตอร์	en_US
dc.degree.grantor	จุฬาลงกรณ์มหาวิทยาลัย	en_US
dc.email.advisor	Atiwong.S@Chula.ac.th	-
dc.email.advisor	Proadpran.Pu@Chula.ac.th	-
dc.email.advisor	ไม่มีข้อมูล	-
dc.identifier.DOI	10.14457/CU.the.2012.1171	-
Appears in Collections:	Eng - Theses

Files in This Item:

File	Description	Size	Format
natnarong_pu.pdf		2.26 MB	Adobe PDF	View/Open

Show simple item record