<oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
  <dc:contributor>Crestani, Fabio</dc:contributor>
  <dc:contributor>Landoni, Monica</dc:contributor>
  <dc:creator>Mahdabi, Parvaz</dc:creator>
  <dc:date>2014-06-23</dc:date>
  <dc:description xmlns:ns0="xml" ns0:lang="en">A patent is a contract between the inventor and the state, granting a limited time  period to the inventor to exploit his invention. In exchange, the inventor must put a  detailed description of his invention in the public domain. Patents can encourage  innovation and economic growth but at the time of economic crisis patents can  hamper such growth. The long duration of the application process is a big obstacle  that needs to be addressed to maximize the benefit of patents on innovation and  economy. This time can be significantly improved by changing the way we search the  patent and non-patent literature.Despite the recent advancement of general  information retrieval and the revolution of Web Search engines, there is still a huge  gap between the emerging technologies from the research labs and adapted by major  Internet search engines, and the systems which are in use by the patent search  communities.In this thesis we investigate the problem of patent prior art search in  patent retrieval with the goal of finding documents which describe the idea of a query  patent. A query patent is a full patent application composed of hundreds of terms  which does not represent a single focused information need. Other relevance  evidences (e.g. classification tags, and bibliographical data) provide additional details  about the underlying information need of the query patent. The first goal of this thesis  is to estimate a uni-gram query model from the textual fields of a query patent. We  then improve the initial query representation using noun phrases extracted from the  query patent. We show that expansion in a query-dependent manner is useful.The  second contribution of this thesis is to address the term mismatch problem from a  query formulation point of view by integrating multiple relevance evidences associated  with the query patent. To do this, we enhance the initial representation of the query  with the term distribution of the community of inventors related to the topic of the  query patent. We then build a lexicon using classification tags and show that query  expansion using this lexicon and considering proximity information (between query  and expansion terms) can improve the retrieval performance. We perform an empirical  evaluation of our proposed models on two patent datasets. The experimental results  show that our proposed models can achieve significantly better results than the  baseline and other enhanced models.</dc:description>
  <dc:format>application/pdf</dc:format>
  <dc:identifier>https://susi.usi.ch/global/documents/318502</dc:identifier>
  <dc:identifier>https://n2t.net/ark:/12658/srd1318502</dc:identifier>
  <dc:identifier>https://susi.usi.ch/documents/318502/files/2014INFO004.pdf</dc:identifier>
  <dc:language>eng</dc:language>
  <dc:relation>info:eu-repo/semantics/altIdentifier/urn/urn:nbn:ch:rero-006-113196</dc:relation>
  <dc:relation>info:eu-repo/semantics/altIdentifier/ark/12658/srd1318502</dc:relation>
  <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
  <dc:rights>License undefined</dc:rights>
  <dc:subject xmlns:ns1="xml" ns1:lang="en">Information retrieval</dc:subject>
  <dc:subject xmlns:ns2="xml" ns2:lang="en">Patent search</dc:subject>
  <dc:subject xmlns:ns3="xml" ns3:lang="en">Query expansion</dc:subject>
  <dc:subject xmlns:ns4="xml" ns4:lang="en">Query formulation</dc:subject>
  <dc:subject>info:eu-repo/classification/udc/004</dc:subject>
  <dc:title xmlns:ns5="xml" ns5:lang="en">Query refinement for patent prior art search</dc:title>
  <dc:type>http://purl.org/coar/resource_type/c_db06</dc:type>
</oai_dc:dc>
