Information Access Evaluation. Multilinguality, Multimodality, and Visual Analytics

Volume 7488 of the series Lecture Notes in Computer Science pp 54-66

Effects of Language and Topic Size in Patent IR: An Empirical Study

  • Florina PiroiAffiliated withVienna University of Technology
  • , Mihai LupuAffiliated withVienna University of Technology
  • , Allan HanburyAffiliated withVienna University of Technology

* Final gross prices may vary according to local VAT.

Get Access


We revisit the effects that various characteristics of the topic documents have on the effectiveness of the systems for the task of finding prior art in the patent domain. In doing so, we provide the reader interested in approaching the domain a guide of the issues that need to be addressed in this context.

For the current study, we select two patent based test collections with a common document representation schema and look at topic characteristics specific to the objectives of the collections. We look at the effect of languages on retrieval and at the length of the topic documents. We present the correlations between these topic facets and their retrieval results, as well as their relevant documents.