Nalazite se na CroRIS probnoj okolini. Ovdje evidentirani podaci neće biti pohranjeni u Informacijskom sustavu znanosti RH. Ako je ovo greška, CroRIS produkcijskoj okolini moguće je pristupi putem poveznice www.croris.hr
izvor podataka: crosbi

Extracting Data from Comparable Corpora (CROSBI ID 63155)

Prilog u knjizi | izvorni znanstveni rad | međunarodna recenzija

Pinnis, Mārcis ; Ljubešić, Nikola ; Ştefănescu, Dan ; Skadiņa, Inguna ; Tadić, Marko ; Gornostaja, Tatjana ; Vintar, Špela ; Fišer, Darja Extracting Data from Comparable Corpora // Using Comparable Corpora for Under-Resourced Areas of Machine Translation / Skadiņa, Inguna ; Gaizauskas, Robert ; Babych, Bogdan et al. (ur.). Berlin: Springer, 2019. str. 89-139 doi: 10.1007/978-3-319-99004-0_4

Podaci o odgovornosti

Pinnis, Mārcis ; Ljubešić, Nikola ; Ştefănescu, Dan ; Skadiņa, Inguna ; Tadić, Marko ; Gornostaja, Tatjana ; Vintar, Špela ; Fišer, Darja

engleski

Extracting Data from Comparable Corpora

Comparable corpora may comprise different types of single-word and multi-word phrases that can be considered as reciprocal translations, which may be beneficial for many different natural language processing tasks. This chapter describes methods and tools developed within the ACCURAT project that allow utilising comparable corpora in order to (1) identify terms, named entities (NEs), and other lexical units in comparable corpora, and (2) to cross- lingually map the identified single-word and multi-word phrases in order to create automatically extracted bilingual dictionaries that can be further utilised in machine translation, question answering, indexing, and other areas where bilingual dictionaries can be useful.

comparable corpora ; data extraction ; machine translation

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

Podaci o prilogu

89-139.

objavljeno

10.1007/978-3-319-99004-0_4

Podaci o knjizi

Using Comparable Corpora for Under-Resourced Areas of Machine Translation

Skadiņa, Inguna ; Gaizauskas, Robert ; Babych, Bogdan ; Ljubešić, Nikola ; Tufiş, Dan ; Vasiļjevs, Andrejs

Berlin: Springer

2019.

978-3-319-99004-0

Povezanost rada

Filologija, Informacijske i komunikacijske znanosti

Poveznice