Cross-language search: The case of Google Language Tools

First Monday ◽

10.5210/fm.v14i3.2335 ◽

2009 ◽

Cited By ~ 7

Author(s):

Jiangping Chen ◽

Yu Bao

Keyword(s):

Information Access ◽

Query Translation ◽

Automatic Translation ◽

Search Terms ◽

Cross Language Information Retrieval ◽

Search Service ◽

Different Types ◽

Cross Language ◽

Access Services

This paper presents a case study of Google Language Tools, especially its cross-language search service. Cross-language search integrates machine translation (MT) and cross-language information retrieval (CLIR) technologies and allows Web users to search and read pages written in languages different from their search terms. In addition to cross-language search, Google Language Tools provides various language support services to multilingual information access. Our study examines the functions of Google Language Tools and the performance of its cross-language search. The results and analysis show that Google Language Tools are useful for Web users. Its cross-language search service provides quality query translation while the automatic translation of result pages needs further improvement. The paper suggests that cross-language search could be used by different types of Web users. The authors also discuss the strategies and important issues with regard to implementing multilingual information access services for information systems.

Download Full-text

International Students’ use of Multilingual Information Access (MLIA) tools: A Survey

Proceedings of the Annual Conference of CAIS / Actes du congrès annuel de l'ACSI ◽

10.29173/cais654 ◽

2013 ◽

Author(s):

Peggy Nzomo ◽

Victoria Rubin ◽

Isola Ajiferuke

Keyword(s):

International Students ◽

Language Proficiency ◽

Information Access ◽

Native English Speakers ◽

English Speakers ◽

Cross Language Information Retrieval ◽

The University ◽

Cross Language ◽

Linguistic Backgrounds

This research presents the results of a case study on potential users of Cross Language Information Retrieval (CLIR) systems –international students at the University of Western Ontario. The study is designed to test their awareness of Multi-Lingual Information Access (MLIA) tools on the internet and in select electronic databases. The study also investigates how non-native English speakers cope with language barriers while searching for information online. Based on the findings, we advocate for designing systems that incorporate CLIR options and other MLIA tools to support users from diverse linguistic backgrounds with varying language proficiency levels.Cette recherche présente les résultats d’une étude de cas auprès d’utilisateurs potentiels, des étudiants internationaux de l’University of Western Ontario, d’un système de repérage d’information par langue croisée (RILC). L’étude est conçue pour tester leur connaissance d’outils d’accès à l’information multilingues (AIM) sur Internet et dans certaines bases de données électroniques. L’étude s’intéresse également aux moyens que prennent les locuteurs non natifs de l’anglais pour palier aux barrières linguistiques lorsqu’ils cherchent de l’information en ligne. Selon les résultats, nous recommandons de concevoir des systèmes qui incorporent des options de RILC et d’autres outils d’AIM pour aider les utilisateurs d’origine linguistique diverse ayant des niveaux de maîtrise linguistique différents.

Download Full-text

English-Marathi Cross Language Information Retrieval System

International Journal of Advanced Research in Computer Science and Software Engineering ◽

10.23956/ijarcsse.v7i8.34 ◽

2017 ◽

Vol 7 (8) ◽

pp. 112

Author(s):

Kalyani Lokhande ◽

Dhanashree Tayade

Keyword(s):

Information Retrieval ◽

Retrieval System ◽

Information Retrieval System ◽

Indian Languages ◽

Query Translation ◽

Retrieval Systems ◽

Cross Language Information Retrieval ◽

Different Types ◽

Information Retrieval Systems ◽

Cross Language

Nowadays, diﬀerent types of content in diﬀerent languages are available on World Wide Web and their usage is increasing rapidly. Cross Language Information Retrieval (CLIR) deals with retrieval of documents in another language than the language of the requested query. Various researchers worked on Cross Language Information Retrieval systems for Indian languages using diﬀerent translation approaches. There is still CLIR system to be developed which allow user to retrieve Marathi documents when English query is given. In the proposed English to Marathi Cross Language Information Retrieval system, translation is based on query translation approach. The proposed system retrieves Marathi documents depending on matching terms in query. The performance of the proposed system is improved by query pre-processing and query expansion using WordNet.

Download Full-text

Comparing Different Units for Query Translation in Chinese Cross-Language Information Retrieval

Proceedings of the 2nd International ICST Conference on Scalable Information Systems ◽

10.4108/infoscale.2007.932 ◽

2007 ◽

Cited By ~ 3

Author(s):

Lixin Shi ◽

Jian-Yun Nie ◽

Jing Bai

Keyword(s):

Information Retrieval ◽

Query Translation ◽

Cross Language Information Retrieval ◽

Cross Language

Download Full-text

A comparison of query translation methods for English-Japanese cross-language information retrieval (poster abstract)

Proceedings of the 22nd annual international ACM SIGIR conference on Research and development in information retrieval - SIGIR '99 ◽

10.1145/312624.312690 ◽

1999 ◽

Cited By ~ 12

Author(s):

Gareth Jones ◽

Tetsuya Sakai ◽

Nigel Collier ◽

Akira Kumano ◽

Kazuo Sumita

Keyword(s):

Information Retrieval ◽

Query Translation ◽

Poster Abstract ◽

Cross Language Information Retrieval ◽

Translation Methods ◽

Cross Language

Download Full-text

Embedding Web-Based Statistical Translation Models in Cross-Language Information Retrieval

Computational Linguistics ◽

10.1162/089120103322711587 ◽

2003 ◽

Vol 29 (3) ◽

pp. 381-419 ◽

Cited By ~ 49

Author(s):

Wessel Kraaij ◽

Jian-Yun Nie ◽

Michel Simard

Keyword(s):

Information Retrieval ◽

Low Cost ◽

Standard Test ◽

Query Translation ◽

Web Based ◽

Parallel Corpora ◽

Cross Language Information Retrieval ◽

Cross Language ◽

Parallel Texts ◽

The Web

Although more and more language pairs are covered by machine translation (MT) services, there are still many pairs that lack translation resources. Cross-language information retrieval (CLIR) is an application that needs translation functionality of a relatively low level of sophistication, since current models for information retrieval (IR) are still based on a bag of words. The Web provides a vast resource for the automatic construction of parallel corpora that can be used to train statistical translation models automatically. The resulting translation models can be embedded in several ways in a retrieval model. In this article, we will investigate the problem of automatically mining parallel texts from the Web and different ways of integrating the translation models within the retrieval process. Our experiments on standard test collections for CLIR show that the Web-based translation models can surpass commercial MT systems in CLIR tasks. These results open the perspective of constructing a fully automatic query translation device for CLIR at a very low cost.

Download Full-text