Liebe Besucherinnen und Besucher,
aufgrund unseres Sommerfestes sind wir am 03. September 2026 bis 14 Uhr erreichbar. Am 04. September 2026 sind wir wieder wie gewohnt für Sie da. Vielen Dank für Ihr Verständnis.
Ihr Team von Sack Fachmedien
Skadina / Gaizauskas / Babych Using Comparable Corpora for Under-Resourced Areas of Machine Translation
1. Auflage 2019
ISBN: 978-3-319-99004-0
Verlag: Springer International Publishing
Format: PDF
Kopierschutz: 1 - PDF Watermark
E-Book, Englisch, 323 Seiten
Reihe: Computer Science
ISBN: 978-3-319-99004-0
Verlag: Springer International Publishing
Format: PDF
Kopierschutz: 1 - PDF Watermark
This book provides an overview of how comparable corpora can be used to overcome the lack of parallel resources when building machine translation systems for under-resourced languages and domains. It presents a wealth of methods and open tools for building comparable corpora from the Web, evaluating comparability and extracting parallel data that can be used for the machine translation task. It is divided into several sections, each covering a specific task such as building, processing, and using comparable corpora, focusing particularly on under-resourced language pairs and domains.
The book is intended for anyone interested in data-driven machine translation for under-resourced languages and domains, especially for developers of machine translation systems, computational linguists and language workers. It offers a valuable resource for specialists and students in natural language processing, machine translation, corpus linguistics and computer-assisted translation, and promotes the broader use of comparable corpora in natural language processing and computational linguistics.
Zielgruppe
Research
Autoren/Hrsg.
Weitere Infos & Material
Introduction.- Cross-language comparability and its Applications for MT (Bogdan Babych, Fangzhong Su, Anthony Hartley, Ahmet Aker, Monica Lestari Paramita, Paul Clough, Robert Gaizauskas).- Collecting comparable corpora (Monica Lestari Paramita, Ahmet Aker, Paul Clough, Robert Gaizauskas, Nikos Glaros, Nikos Mastropavlos, Olga Yannoutsou, Radu Ion, Dan ?tefanescu, Alexandru Ceausu, Dan Tufi? and Judita Preiss).- Extracting data from comparable corpora (Marcis Pinnis, Nikola Ljubešic, Dan Stefanescu, Inguna Skadina, Marko Tadic, Tatjana Gornostaja, Špela Vintar, Darja Fišer).- Mapping and aligning units from comparable corpora (Ahmet Aker, Alexandru Ceau?u, Yang Feng, Robert Gaizauskas, Sabine Hunsicker, Radu Ion, Elena Irimia, Dan ?tefanescu, Dan Tufi?).- Training, enhancing, evaluating and using MT-Systems with comparable data (Bogdan Babych, Yu Chen, Andreas Eisele, Sabine Hunsicker, Marcis Pinnis, Inguna Skadina, Raivis Skadinš, Gregor Thurmair, Andrejs Vasiljevs, Mateja Verlic, Xiaojun Zhang).- New areas of application of comparable corpora (Reinhard Rapp, Vivian Xu, Michael Zock, Serge Sharoff, Richard Forsyth, Bogdan Babych, Chenhui Chu, Toshiaki Nakazawa, Sadao Kurohashi).- Appendices (Ahmet Aker, Radu Ion, Nikos Mastropavlos, Monica Paramita, Marcis Pinnis, Dan Stefanescu, Fangzhong Su, Gregor Thurmair,Elena Irimia, Nikola Ljubešic, Evangelos Kanoulas, Judita Preiss, Rob Gaizauskas, Paul Clough, Emma Barker, Nikos Glaros, Tiberiu Boro?, Inguna Skadina, Andrejs Vasiljevs).




