A language model adaptation using multiple varied corpora

IEEE Workshop on Automatic Speech Recognition and Understanding, 2001. ASRU '01. - Trang 389-392

H. Yamamoto¹, Y. Sagisaka¹

¹ATR Spoken Language Translation Research Laboratories, Soraku-gun, Kyoto, Japan

Tóm tắt

A new language model adaptation scheme is proposed to cope with multiple varied speech recognition tasks. Both topic difference and sentence style difference resulting from the speaker's role are reflected in the proposed language model adaptation. An adaptation is carried out using two different language corpora where only the topic or speaker's style is matched. New word clustering techniques are introduced to extract the topic or style dependency separately. Word neighboring characteristics in the two adaptation source data are regarded as different features in this clustering. All words are classified into commonly used word classes and topic or style dependent classes. Furthermore, target topic and sentence style dependent words and their neighboring characteristics are emphasized according to their frequency in the adaptation target data. In the evaluation experiment, the proposed method shows a 13% lower perplexity and a 9% lower word error rate in continuous speech recognition compared with the conventional adaptation method.

Từ khóa

#Adaptation model #Natural languages #Speech recognition #Data mining #Frequency #Error analysis #Vocabulary

Tài liệu tham khảo

10.1109/ICASSP.1997.596042 10.1109/ICASSP.1999.758180 10.1006/csla.1996.0021 takezawa, 1998, Speech and Language Databases for Speech Translation Research in ATR, Proc of the 1st International Workshop on East-Asian Language Resource and Evaluation shimizu, 1996, Spontaneous sialog speech recognition using cross-word context constrained word graphs, Proc ICASSP-96, 145 bai, 1998, Building Class-based Language Models with Contextual Statistics, Proc ICASSP-98, 173 moore, 2000, Class-based language model adaptation using mixture of word-class weight, Proc ICSLP 2000, 4, 512

Scholar Hub - Công cụ hỗ trợ trích dẫn và phân tích khoa học Việt Nam

Scholar Hub là công cụ hỗ trợ trích dẫn và phân tích ảnh hưởng của các bài báo, công bố khoa học Việt Nam và Quốc tế.
ScholarHub KHÔNG đăng thông tin tổng hợp, KHÔNG đăng lại nội dung từ các trang báo chí Việt Nam hoặc trang thông tin điện tử khác tại Việt Nam.

Thông tin, cập nhật

Đăng ký Tạp chí tham gia Scholar Hub

Phản hồi ý kiến về Scholar Hub

Bài viết, nội dung cập nhật

Chủ đề khoa học

Website liên kết

Hệ thống CSDL Khoa học & Công nghệ SciBase

Phần mềm kiểm tra trùng lặp Kiểm Tra Tài Liệu

Phần mềm xuất bản tạp chí điện tử VOJS

Hệ thống hội thảo khoa học Việt Nam

Nền tảng trắc nghiệm và đề thi đa lĩnh vực LetQA

Thông tin liên hệ & hỗ trợ

Đơn vị chủ quản, phát triển và vận hành: Công ty Cổ phần Metis

Địa chỉ liên hệ: 26A Lê Đức Thọ, Phường Từ Liêm, Thành phố Hà Nội

Số giấy chứng nhận ĐKKD: 0109293202 cấp ngày 03/08/2020 tại Sở Kế hoạch và Đầu tư thành phố Hà Nội

Người quản lý và chịu trách nhiệm nội dung: Nguyễn Ngọc Sơn

Hotline: 0566.685.688

Email: [email protected]