![]() ![]() ![]() |
![]() |
|
![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]() ![]()
|
Return to Session IVb: Cross-language information retrieval It is crucial for cross-language information retrieval (CLIR) systems to deal with the translation of unknown queries1 due to that real queries might be short. The purpose of this paper is to investigate the feasibility of exploiting the Web as the corpus source to translate unknown queries for CLIR. We propose an online translation approach to determine effective translations for unknown query terms via mining of bilingual search-result pages obtained from Web search engines. This approach can alleviate the problem of the lack of large bilingual corpora, translate many unknown query terms, provide flexible query specifications, and extract semantically-close translations to benefit CLIR tasks – especially for cross-language Web search. @inproceedings{1009020, author = {Pu-Jen Cheng and Jei-Wen Teng and Ruei-Cheng Chen and Jenq-Haur Wang and Wen-Hsiang Lu and Lee-Feng Chien}, title = {Translating unknown queries with web corpora for cross-language information retrieval}, booktitle = {SIGIR '04: Proceedings of the 27th annual international conference on Research and development in information retrieval}, year = {2004}, isbn = {1-58113-881-4}, pages = {146--153}, location = {Sheffield, United Kingdom}, doi = {http://doi.acm.org/10.1145/1008992.1009020}, publisher = {ACM Press}, } ![]() ©2005 Association for Computing Machinery |