The Use of Domain Modelling to Improve Performance Over a Query Session
|
|
- Beverley Harvey
- 5 years ago
- Views:
Transcription
1 The Use of Domain Modelling to Improve Performance Over a Query Session Deirdre Lungley Dyaa Albakour Udo Kruschwitz Apr 18, 2011
2 Table of contents 1 The AutoAdapt Project
3 The AutoAdapt Project Automatic Adaptation of Knowledge Structures for Assisted Information Seeking 3-year EPSRC project (November November 2011) Collaboration between: University of Essex Robert Gordon University Open University
4 - Aim Aim The Task Evaluation Metrics Evaluate the effectiveness of search engines in interpreting query reformulations [4] A good search engine should be able to utilise the previous queries in the sequence of a session to provide better results that reflect the user needs throughout the session Example: Britney Spears Paris Hilton France Hotels Paris Hilton The session track provides a framework to assess this particular issue in Information Retrieval systems
5 Aim The Task Evaluation Metrics - The Task Only sessions with two queries were considered in 2010 Participants were given a set of 150 query pairs, each query pair (original query, query reformulation) represented a user session. Roughly split between: 1 Generalisation: low carb high fat diet types of diets. 2 Specification: us map us map states and capitals 3 Drifting/Parallel Reformulation: music man performances music man script. The participants were asked to submit three ranked lists of documents from the ClueWeb09 dataset: One for the original query (RL1) One for the query reformulation ignoring the original query (RL2) One for the query reformulation taking the original query into consideration (RL3)
6 Aim The Task Evaluation Metrics - Evaluation Metrics 1 Can search engines improve their performance for a given query using previous queries (Task G1)? Compare nsdcg@10 for RL1 RL2 and RL1 RL3 2 How do they perform over an entire session (Task G2)? nsdcg@10 for RL1 RL3
7 General Methodology - essex3 FCA Lattice Approach - General Methodology RL1 - Indri - original query (e.g. hoboken ) RL2 - Indri - reformulated query (e.g. hoboken nightlife ) RL3 - Indri - supports query expansion - weighted belief operators. Allow us to combine original query, reformulated query and derived expansion terms, e.g., #weight( 0.7 #combine( hoboken nightlife ) 0.3 #combine( hoboken #1( hoboken bars ) #1( hoboken nightclubs ) #1( hoboken wine bar ) ) ) All lists were filtered for spam - Waterloo Spam Filter - 70% cutoff 1 1
8 General Methodology - essex3 FCA Lattice Approach - AutoAdapt@TREC10 - essex3 Anchor log as a simulated query log has been shown to be effective in query reformulation Dang and Croft, WSDM10[1] Anchor log for ClueWeb09 available from the University of Twente[3] Use Fonseca s Association rules [2] to extract suggestions for both constituents of the session pair Use the intersection of the suggestions extracted for both constituents as useful expansions Session gps devices garmin computer worms malware us geographic map us political map Expansion terms or phrases gps devices, wikipedia, usb, gps device, gps products, garmin nuvi880, garmin gps device, visit garmin computer worms, computer security, category, worm us political map, article
9 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach Anchor logs effective expensive Explore feasibility of generating expansion terms using Formal Concept Analysis (FCA) lattices Could we exploit the hierarchy of concepts of an FCA lattice created from documents and their index terms?
10 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach Figure: Sample Hasse Diagram for query volvo
11 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach
12 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach Build lattices for both queries in the session to generate expansion terms for RL3 Initial hypotheses: 1 Lattice concepts generated from documents returned by a search engine over ClueWeb09 would be more discriminate than those generated by documents returned by a WWW search engine 2 Using a combination of common and distinct lattice concepts depending on the query type would be more discriminate than solely distinct terms
13 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach Methods explored in creating document descriptors: Snippets returned by Microsoft s LiveSearch API Snippets returned by Indri over ClueWeb09 Noun phrases containing query from full Indri documents ClueWeb09 anchor logs for Indri result documents Methods to extract expansion terms Disjunction use concepts which appear in lattice generated by query 2 and not in lattice generated by query 1 Conjunction concepts common to both lattices Combination based on query type determined by term distribution, e.g., query 1 subsumed by query 2 (volvo : volvo semi trucks) specialisation
14 General Methodology - essex3 FCA Lattice Approach - FCA Lattice Approach Top Lattice Concepts Disjunction hoboken hoboken nightlife Derived Concepts new jersey hoboken bars hoboken bars free encyclopedia hoboken nightclubs hoboken nightclubs new york new jersey hoboken wine bar hudson county hoboken wine bar united states hoboken events Table: Example of expansion terms generated by disjunction method
15 % Average Increase % Average Increase TREC Metrics Exploratory Findings System All topics Spec. Gen. Drift. Essex Essex Lattice Table: % average increase from nsdcg@10.rl12 to nsdcg@10.rl13
16 TREC Metrics % Average Increase TREC Metrics Exploratory Findings Run nsdcg@10.rl13 (G2) all sessions Specification Generalisation Drifting Essex Essex Lattice CengageS10R Submitted nsdcg@10 (G1) ndcg@10 (G1) Run RL12 RL13 RL1 RL2 RL3 Essex Essex Lattice
17 Exploratory Findings % Average Increase TREC Metrics Exploratory Findings Search Descriptors Expansion Terms nsdcg@10.rl13 (G2) Engine Derived From Selection Method all sessions Spec. MSN Snippets Disjunction MSN Snippets Combination Indri Title, anchors Disjunction Indri Title, anchors Conjunction Indri Full Doc 1 Disjunction noun phrases containing a query term
18 The AutoAdapt Project TREC Session Track 2010 results demonstrated it is very hard to get any measureable improvement when utilising prior history FCA lattice - we believe it to be a promising structure not ready to dismiss Indexing the entire ClueWeb09 collection in house Category B Tier 1 web crawl Helps to explain success of the expansion term wikipedia Opportunity then to explore query expansion methods Effect of spam filtering
19 V. Dang and B. W. Croft. Query reformulation using anchor text. In WSDM 10: Proceedings of the third ACM international conference on Web search and data mining, pages 41 50, New York, NY, USA, ACM. B. M. Fonseca, P. B. Golgher, E. S. de Moura, and N. Ziviani. Using association rules to discover search engines related queries. In Proceedings of the First Latin American Web Congress, pages 66 71, D. Hiemstra and C. Hauff. Mirex: Mapreduce information retrieval experiments. Technical Report TR-CTIT-10-15, Centre for Telematics and Information Technology University of Twente, Enschede, April E. Kanoulas, B. Carterette, P. Clough, and M. Sanderson. Session track overview. In Proceedings of The Nineteenth Text REtrieval Conference Proceedings (TREC 2010), National Institute of Standards and Technology, To Appear.
Adaptive Search at Essex
Adaptive Search at Essex Udo Kruschwitz School of Computer Science and Electronic Engineering University of Essex udo@essex.ac.uk 7th October 2011 Adaptive Search - LAC Day - 7 October 2011 1 (Source:
More informationSIR 11: Information Retrieval Over Query Sessions
SIR 11: Information Retrieval Over Query Sessions Organizers Ben Cartere6e, University of Delaware Evangelos Kanoulas, University of Sheffield Paul Clough, University of Sheffield Mark Sanderson, RMIT
More informationOn Duplicate Results in a Search Session
On Duplicate Results in a Search Session Jiepu Jiang Daqing He Shuguang Han School of Information Sciences University of Pittsburgh jiepu.jiang@gmail.com dah44@pitt.edu shh69@pitt.edu ABSTRACT In this
More informationWebis at the TREC 2012 Session Track
Webis at the TREC 2012 Session Track Extended Abstract for the Conference Notebook Matthias Hagen, Martin Potthast, Matthias Busse, Jakob Gomoll, Jannis Harder, and Benno Stein Bauhaus-Universität Weimar
More informationCengage Learning at TREC 2010 Session Track Benjamin King University of Michigan 2260 Hayward Street Ann Arbor, MI USA
Cengage Learning at TREC 2010 Session Track Benjamin King University of Michigan 2260 Hayward Street Ann Arbor, MI 48109 USA benjaminking@umich.edu Ivan Provalov Cengage Learning 27500 Drake Road Farmington
More informationICTNET at Web Track 2010 Diversity Task
ICTNET at Web Track 2010 Diversity Task Yuanhai Xue 1,2, Zeying Peng 1,2, Xiaoming Yu 1, Yue Liu 1, Hongbo Xu 1, Xueqi Cheng 1 1. Institute of Computing Technology, Chinese Academy of Sciences, Beijing,
More informationReducing Redundancy with Anchor Text and Spam Priors
Reducing Redundancy with Anchor Text and Spam Priors Marijn Koolen 1 Jaap Kamps 1,2 1 Archives and Information Studies, Faculty of Humanities, University of Amsterdam 2 ISLA, Informatics Institute, University
More informationAutomatic Generation of Query Sessions using Text Segmentation
Automatic Generation of Query Sessions using Text Segmentation Debasis Ganguly, Johannes Leveling, and Gareth J.F. Jones CNGL, School of Computing, Dublin City University, Dublin-9, Ireland {dganguly,
More informationOn Duplicate Results in a Search Session
On Duplicate Results in a Search Session Jiepu Jiang Daqing He Shuguang Han School of Information Sciences University of Pittsburgh jiepu.jiang@gmail.com dah44@pitt.edu shh69@pitt.edu ABSTRACT In this
More informationTREC 2016 Dynamic Domain Track: Exploiting Passage Representation for Retrieval and Relevance Feedback
RMIT @ TREC 2016 Dynamic Domain Track: Exploiting Passage Representation for Retrieval and Relevance Feedback Ameer Albahem ameer.albahem@rmit.edu.au Lawrence Cavedon lawrence.cavedon@rmit.edu.au Damiano
More informationEffective Structured Query Formulation for Session Search
Effective Structured Query Formulation for Session Search Dongyi Guan Hui Yang Nazli Goharian Department of Computer Science Georgetown University 37 th and O Street, NW, Washington, DC, 20057 dg372@georgetown.edu,
More informationOverview of the TREC 2009 Web Track
Overview of the TREC 2009 Web Track Charles L. A. Clarke University of Waterloo Nick Craswell Microsoft Ian Soboroff NIST 1 Overview The TREC Web Track explores and evaluates Web retrieval technologies.
More informationBetter Contextual Suggestions in ClueWeb12 Using Domain Knowledge Inferred from The Open Web
Better Contextual Suggestions in ClueWeb12 Using Domain Knowledge Inferred from The Open Web Thaer Samar 1, Alejandro Bellogín 2, and Arjen P. de Vries 1 1 Centrum Wiskunde & Informatica, {samar,arjen}@cwi.nl
More informationUniversity of Delaware at Diversity Task of Web Track 2010
University of Delaware at Diversity Task of Web Track 2010 Wei Zheng 1, Xuanhui Wang 2, and Hui Fang 1 1 Department of ECE, University of Delaware 2 Yahoo! Abstract We report our systems and experiments
More informationClassification and retrieval of biomedical literatures: SNUMedinfo at CLEF QA track BioASQ 2014
Classification and retrieval of biomedical literatures: SNUMedinfo at CLEF QA track BioASQ 2014 Sungbin Choi, Jinwook Choi Medical Informatics Laboratory, Seoul National University, Seoul, Republic of
More informationAn Exploration of Query Term Deletion
An Exploration of Query Term Deletion Hao Wu and Hui Fang University of Delaware, Newark DE 19716, USA haowu@ece.udel.edu, hfang@ece.udel.edu Abstract. Many search users fail to formulate queries that
More informationNUSIS at TREC 2011 Microblog Track: Refining Query Results with Hashtags
NUSIS at TREC 2011 Microblog Track: Refining Query Results with Hashtags Hadi Amiri 1,, Yang Bao 2,, Anqi Cui 3,,*, Anindya Datta 2,, Fang Fang 2,, Xiaoying Xu 2, 1 Department of Computer Science, School
More informationNTU Approaches to Subtopic Mining and Document Ranking at NTCIR-9 Intent Task
NTU Approaches to Subtopic Mining and Document Ranking at NTCIR-9 Intent Task Chieh-Jen Wang, Yung-Wei Lin, *Ming-Feng Tsai and Hsin-Hsi Chen Department of Computer Science and Information Engineering,
More informationBetter Contextual Suggestions in ClueWeb12 Using Domain Knowledge Inferred from The Open Web
Better Contextual Suggestions in ClueWeb12 Using Domain Knowledge Inferred from The Open Web Thaer Samar 1, Alejandro Bellogín 2, and Arjen P. de Vries 1 1 Centrum Wiskunde & Informatica, {samar,arjen}@cwi.nl
More informationTREC 2017 Dynamic Domain Track Overview
TREC 2017 Dynamic Domain Track Overview Grace Hui Yang Zhiwen Tang Ian Soboroff Georgetown University Georgetown University NIST huiyang@cs.georgetown.edu zt79@georgetown.edu ian.soboroff@nist.gov 1. Introduction
More informationEfficient Diversification of Web Search Results
Efficient Diversification of Web Search Results G. Capannini, F. M. Nardini, R. Perego, and F. Silvestri ISTI-CNR, Pisa, Italy Laboratory Web Search Results Diversification Query: Vinci, what is the user
More informationNortheastern University in TREC 2009 Million Query Track
Northeastern University in TREC 2009 Million Query Track Evangelos Kanoulas, Keshi Dai, Virgil Pavlu, Stefan Savev, Javed Aslam Information Studies Department, University of Sheffield, Sheffield, UK College
More informationInformativeness for Adhoc IR Evaluation:
Informativeness for Adhoc IR Evaluation: A measure that prevents assessing individual documents Romain Deveaud 1, Véronique Moriceau 2, Josiane Mothe 3, and Eric SanJuan 1 1 LIA, Univ. Avignon, France,
More informationHeading-aware Snippet Generation for Web Search
Heading-aware Snippet Generation for Web Search Tomohiro Manabe and Keishi Tajima Graduate School of Informatics, Kyoto Univ. {manabe@dl.kuis, tajima@i}.kyoto-u.ac.jp Web Search Result Snippets Are short
More informationAn Improvement of Search Results Access by Designing a Search Engine Result Page with a Clustering Technique
An Improvement of Search Results Access by Designing a Search Engine Result Page with a Clustering Technique 60 2 Within-Subjects Design Counter Balancing Learning Effect 1 [1 [2www.worldwidewebsize.com
More informationOverview of the TREC 2010 Web Track
Overview of the TREC 2010 Web Track Charles L. A. Clarke University of Waterloo Nick Craswell Microsoft Ian Soboroff NIST Gordon V. Cormack University of Waterloo 1 Introduction The TREC Web Track explores
More informationVerbose Query Reduction by Learning to Rank for Social Book Search Track
Verbose Query Reduction by Learning to Rank for Social Book Search Track Messaoud CHAA 1,2, Omar NOUALI 1, Patrice BELLOT 3 1 Research Center on Scientific and Technical Information 05 rue des 03 frères
More informationUniversity of TREC 2009: Indexing half a billion web pages
University of Twente @ TREC 2009: Indexing half a billion web pages Claudia Hauff and Djoerd Hiemstra University of Twente, The Netherlands {c.hauff, hiemstra}@cs.utwente.nl draft 1 Introduction The University
More informationWashington, DC April 22, 2013
Structured Query Formulation and Result Organization for Session Search A Thesis submitted to the Faculty of the Graduate School of Arts and Sciences of Georgetown University in partial fulllment of the
More informationWeb document summarisation: a task-oriented evaluation
Web document summarisation: a task-oriented evaluation Ryen White whiter@dcs.gla.ac.uk Ian Ruthven igr@dcs.gla.ac.uk Joemon M. Jose jj@dcs.gla.ac.uk Abstract In this paper we present a query-biased summarisation
More informationA Comparative Analysis of Cascade Measures for Novelty and Diversity
A Comparative Analysis of Cascade Measures for Novelty and Diversity Charles Clarke, University of Waterloo Nick Craswell, Microsoft Ian Soboroff, NIST Azin Ashkan, University of Waterloo Background Measuring
More informationExploiting Global Impact Ordering for Higher Throughput in Selective Search
Exploiting Global Impact Ordering for Higher Throughput in Selective Search Michał Siedlaczek [0000-0002-9168-0851], Juan Rodriguez [0000-0001-6483-6956], and Torsten Suel [0000-0002-8324-980X] Computer
More informationNortheastern University in TREC 2009 Web Track
Northeastern University in TREC 2009 Web Track Shahzad Rajput, Evangelos Kanoulas, Virgil Pavlu, Javed Aslam College of Computer and Information Science, Northeastern University, Boston, MA, USA Information
More informationImproving Synoptic Querying for Source Retrieval
Improving Synoptic Querying for Source Retrieval Notebook for PAN at CLEF 2015 Šimon Suchomel and Michal Brandejs Faculty of Informatics, Masaryk University {suchomel,brandejs}@fi.muni.cz Abstract Source
More informationFinding Related Entities by Retrieving Relations: UIUC at TREC 2009 Entity Track
Finding Related Entities by Retrieving Relations: UIUC at TREC 2009 Entity Track V.G.Vinod Vydiswaran, Kavita Ganesan, Yuanhua Lv, Jing He, ChengXiang Zhai Department of Computer Science University of
More informationCS6200 Information Retrieval. Jesse Anderton College of Computer and Information Science Northeastern University
CS6200 Information Retrieval Jesse Anderton College of Computer and Information Science Northeastern University Major Contributors Gerard Salton! Vector Space Model Indexing Relevance Feedback SMART Karen
More informationMaster Project. Various Aspects of Recommender Systems. Prof. Dr. Georg Lausen Dr. Michael Färber Anas Alzoghbi Victor Anthony Arrascue Ayala
Master Project Various Aspects of Recommender Systems May 2nd, 2017 Master project SS17 Albert-Ludwigs-Universität Freiburg Prof. Dr. Georg Lausen Dr. Michael Färber Anas Alzoghbi Victor Anthony Arrascue
More informationDeriving Query Suggestions for Site Search
Deriving Query Suggestions for Site Search Udo Kruschwitz, Deirdre Lungley, M-Dyaa Albakour and Dawei Song November 20, 2012 Abstract Modern search engines have been moving away from very simplistic interfaces
More informationFocused Retrieval Using Topical Language and Structure
Focused Retrieval Using Topical Language and Structure A.M. Kaptein Archives and Information Studies, University of Amsterdam Turfdraagsterpad 9, 1012 XT Amsterdam, The Netherlands a.m.kaptein@uva.nl Abstract
More informationInformation Retrieval
Multimedia Computing: Algorithms, Systems, and Applications: Information Retrieval and Search Engine By Dr. Yu Cao Department of Computer Science The University of Massachusetts Lowell Lowell, MA 01854,
More informationAutomatic Query Type Identification Based on Click Through Information
Automatic Query Type Identification Based on Click Through Information Yiqun Liu 1,MinZhang 1,LiyunRu 2, and Shaoping Ma 1 1 State Key Lab of Intelligent Tech. & Sys., Tsinghua University, Beijing, China
More informationFormulating XML-IR Queries
Alan Woodley Faculty of Information Technology, Queensland University of Technology PO Box 2434. Brisbane Q 4001, Australia ap.woodley@student.qut.edu.au Abstract: XML information retrieval systems differ
More informationUniversity of Amsterdam at INEX 2010: Ad hoc and Book Tracks
University of Amsterdam at INEX 2010: Ad hoc and Book Tracks Jaap Kamps 1,2 and Marijn Koolen 1 1 Archives and Information Studies, Faculty of Humanities, University of Amsterdam 2 ISLA, Faculty of Science,
More informationPredicting Next Search Actions with Search Engine Query Logs
2011 IEEE/WIC/ACM International Conferences on Web Intelligence and Intelligent Agent Technology Predicting Next Search Actions with Search Engine Query Logs Kevin Hsin-Yih Lin Chieh-Jen Wang Hsin-Hsi
More informationIncreasing Stability of Result Organization for Session Search
Increasing Stability of Result Organization for Session Search Dongyi Guan and Hui Yang Department of Computer Science, Georgetown University 37th and O Street NW, Washington DC, 20057, USA dongyi.guan@gmail.com,
More informationOverview of the INEX 2009 Link the Wiki Track
Overview of the INEX 2009 Link the Wiki Track Wei Che (Darren) Huang 1, Shlomo Geva 2 and Andrew Trotman 3 Faculty of Science and Technology, Queensland University of Technology, Brisbane, Australia 1,
More informationAn Investigation of Basic Retrieval Models for the Dynamic Domain Task
An Investigation of Basic Retrieval Models for the Dynamic Domain Task Razieh Rahimi and Grace Hui Yang Department of Computer Science, Georgetown University rr1042@georgetown.edu, huiyang@cs.georgetown.edu
More informationAn adaptable search system for collection of partially structured documents
Samantha Riccadonna An adaptable search system for collection of partially structured documents by Udo Kruschwitz Web Information Retrieval Course A.Y. 2005-2006 Outline Search system overview Few concepts
More informationRetrieval and Feedback Models for Blog Distillation
Retrieval and Feedback Models for Blog Distillation Jonathan Elsas, Jaime Arguello, Jamie Callan, Jaime Carbonell Language Technologies Institute, School of Computer Science, Carnegie Mellon University
More informationAxiomatic Approaches to Information Retrieval - University of Delaware at TREC 2009 Million Query and Web Tracks
Axiomatic Approaches to Information Retrieval - University of Delaware at TREC 2009 Million Query and Web Tracks Wei Zheng Hui Fang Department of Electrical and Computer Engineering University of Delaware
More informationiarabicweb16: Making a Large Web Collection More Accessible for Research
iarabicweb16: Making a Large Web Collection More Accessible for Research Khaled Yasser, Reem Suwaileh, Abdelrahman Shouman, Yassmine Barkallah, Mucahid Kutlu, Tamer Elsayed Computer Science and Engineering
More informationOverview of the TREC 2013 Crowdsourcing Track
Overview of the TREC 2013 Crowdsourcing Track Mark D. Smucker 1, Gabriella Kazai 2, and Matthew Lease 3 1 Department of Management Sciences, University of Waterloo 2 Microsoft Research, Cambridge, UK 3
More informationUnderstanding the Query: THCIB and THUIS at NTCIR-10 Intent Task. Junjun Wang 2013/4/22
Understanding the Query: THCIB and THUIS at NTCIR-10 Intent Task Junjun Wang 2013/4/22 Outline Introduction Related Word System Overview Subtopic Candidate Mining Subtopic Ranking Results and Discussion
More informationChapter 6. Queries and Interfaces
Chapter 6 Queries and Interfaces Keyword Queries Simple, natural language queries were designed to enable everyone to search Current search engines do not perform well (in general) with natural language
More informationThis is an author-deposited version published in : Eprints ID : 15246
Open Archive TOULOUSE Archive Ouverte (OATAO) OATAO is an open access repository that collects the work of Toulouse researchers and makes it freely available over the web where possible. This is an author-deposited
More informationContextual Search Using Ontology-Based User Profiles Susan Gauch EECS Department University of Kansas Lawrence, KS
Vishnu Challam Microsoft Corporation One Microsoft Way Redmond, WA 9802 vishnuc@microsoft.com Contextual Search Using Ontology-Based User s Susan Gauch EECS Department University of Kansas Lawrence, KS
More informationThe University of Illinois Graduate School of Library and Information Science at TREC 2011
The University of Illinois Graduate School of Library and Information Science at TREC 2011 Miles Efron, Adam Kehoe, Peter Organisciak, Sunah Suh 501 E. Daniel St., Champaign, IL 61820 1 Introduction The
More informationUniversity of Glasgow at the NTCIR-9 Intent task
University of Glasgow at the NTCIR-9 Intent task Experiments with Terrier on Subtopic Mining and Document Ranking Rodrygo L. T. Santos rodrygo@dcs.gla.ac.uk Craig Macdonald craigm@dcs.gla.ac.uk School
More informationWrapper: An Application for Evaluating Exploratory Searching Outside of the Lab
Wrapper: An Application for Evaluating Exploratory Searching Outside of the Lab Bernard J Jansen College of Information Sciences and Technology The Pennsylvania State University University Park PA 16802
More informationA RECOMMENDER SYSTEM FOR SOCIAL BOOK SEARCH
A RECOMMENDER SYSTEM FOR SOCIAL BOOK SEARCH A thesis Submitted to the faculty of the graduate school of the University of Minnesota by Vamshi Krishna Thotempudi In partial fulfillment of the requirements
More informationAggregation for searching complex information spaces. Mounia Lalmas
Aggregation for searching complex information spaces Mounia Lalmas mounia@acm.org Outline Document Retrieval Focused Retrieval Aggregated Retrieval Complexity of the information space (s) INEX - INitiative
More informationInformation Retrieval
Introduction Information Retrieval Information retrieval is a field concerned with the structure, analysis, organization, storage, searching and retrieval of information Gerard Salton, 1968 J. Pei: Information
More informationText Mining. Munawar, PhD. Text Mining - Munawar, PhD
10 Text Mining Munawar, PhD Definition Text mining also is known as Text Data Mining (TDM) and Knowledge Discovery in Textual Database (KDT).[1] A process of identifying novel information from a collection
More informationUNIT-V WEB MINING. 3/18/2012 Prof. Asha Ambhaikar, RCET Bhilai.
UNIT-V WEB MINING 1 Mining the World-Wide Web 2 What is Web Mining? Discovering useful information from the World-Wide Web and its usage patterns. 3 Web search engines Index-based: search the Web, index
More informationEvaluating Learning-to-Rank Methods in the Web Track Adhoc Task.
Evaluating Learning-to-Rank Methods in the Web Track Adhoc Task. Leonid Boytsov and Anna Belova leo@boytsov.info, anna@belova.org TREC-20, November 2011, Gaithersburg, Maryland, USA Learning-to-rank methods
More informationExternal Query Reformulation for Text-based Image Retrieval
External Query Reformulation for Text-based Image Retrieval Jinming Min and Gareth J. F. Jones Centre for Next Generation Localisation School of Computing, Dublin City University Dublin 9, Ireland {jmin,gjones}@computing.dcu.ie
More informationEffective Tweet Contextualization with Hashtags Performance Prediction and Multi-Document Summarization
Effective Tweet Contextualization with Hashtags Performance Prediction and Multi-Document Summarization Romain Deveaud 1 and Florian Boudin 2 1 LIA - University of Avignon romain.deveaud@univ-avignon.fr
More informationTime-aware Approaches to Information Retrieval
Time-aware Approaches to Information Retrieval Nattiya Kanhabua Department of Computer and Information Science Norwegian University of Science and Technology 24 February 2012 Motivation Searching documents
More informationIntroduction to Information Retrieval. (COSC 488) Spring Nazli Goharian. Course Outline
Introduction to Information Retrieval (COSC 488) Spring 2012 Nazli Goharian nazli@cs.georgetown.edu Course Outline Introduction Retrieval Strategies (Models) Retrieval Utilities Evaluation Indexing Efficiency
More informationCross-Language Information Retrieval using Dutch Query Translation
Cross-Language Information Retrieval using Dutch Query Translation Anne R. Diekema and Wen-Yuan Hsiao Syracuse University School of Information Studies 4-206 Ctr. for Science and Technology Syracuse, NY
More informationA System for Query-Specific Document Summarization
A System for Query-Specific Document Summarization Ramakrishna Varadarajan, Vagelis Hristidis. FLORIDA INTERNATIONAL UNIVERSITY, School of Computing and Information Sciences, Miami. Roadmap Need for query-specific
More informationPredicting Query Performance on the Web
Predicting Query Performance on the Web No Author Given Abstract. Predicting performance of queries has many useful applications like automatic query reformulation and automatic spell correction. However,
More informationCIRGDISCO at RepLab2012 Filtering Task: A Two-Pass Approach for Company Name Disambiguation in Tweets
CIRGDISCO at RepLab2012 Filtering Task: A Two-Pass Approach for Company Name Disambiguation in Tweets Arjumand Younus 1,2, Colm O Riordan 1, and Gabriella Pasi 2 1 Computational Intelligence Research Group,
More informationModern Retrieval Evaluations. Hongning Wang
Modern Retrieval Evaluations Hongning Wang CS@UVa What we have known about IR evaluations Three key elements for IR evaluation A document collection A test suite of information needs A set of relevance
More informationThe Open University s repository of research publications and other research outputs. Search Personalization with Embeddings
Open Research Online The Open University s repository of research publications and other research outputs Search Personalization with Embeddings Conference Item How to cite: Vu, Thanh; Nguyen, Dat Quoc;
More informationBuilding Rich User Profiles for Personalized News Recommendation
Building Rich User Profiles for Personalized News Recommendation Youssef Meguebli 1, Mouna Kacimi 2, Bich-liên Doan 1, and Fabrice Popineau 1 1 SUPELEC Systems Sciences (E3S), Gif sur Yvette, France, {youssef.meguebli,bich-lien.doan,fabrice.popineau}@supelec.fr
More informationAutomatically Generating Queries for Prior Art Search
Automatically Generating Queries for Prior Art Search Erik Graf, Leif Azzopardi, Keith van Rijsbergen University of Glasgow {graf,leif,keith}@dcs.gla.ac.uk Abstract This report outlines our participation
More informationAPNIC Training Mini Survey
APNIC Training - 2007 Mini Survey Summary Report by John Earls February 2008 Introduction The following analysis summarises the results of the APNIC training survey conducted during the period November
More informationCLEF-IP 2009: Exploring Standard IR Techniques on Patent Retrieval
DCU @ CLEF-IP 2009: Exploring Standard IR Techniques on Patent Retrieval Walid Magdy, Johannes Leveling, Gareth J.F. Jones Centre for Next Generation Localization School of Computing Dublin City University,
More informationEntity and Knowledge Base-oriented Information Retrieval
Entity and Knowledge Base-oriented Information Retrieval Presenter: Liuqing Li liuqing@vt.edu Digital Library Research Laboratory Virginia Polytechnic Institute and State University Blacksburg, VA 24061
More informationIJREAT International Journal of Research in Engineering & Advanced Technology, Volume 1, Issue 5, Oct-Nov, ISSN:
IJREAT International Journal of Research in Engineering & Advanced Technology, Volume 1, Issue 5, Oct-Nov, 20131 Improve Search Engine Relevance with Filter session Addlin Shinney R 1, Saravana Kumar T
More informationA Model for Interactive Web Information Retrieval
A Model for Interactive Web Information Retrieval Orland Hoeber and Xue Dong Yang University of Regina, Regina, SK S4S 0A2, Canada {hoeber, yang}@uregina.ca Abstract. The interaction model supported by
More informationWelcome to the class of Web Information Retrieval!
Welcome to the class of Web Information Retrieval! Tee Time Topic Augmented Reality and Google Glass By Ali Abbasi Challenges in Web Search Engines Min ZHANG z-m@tsinghua.edu.cn April 13, 2012 Challenges
More informationSNUMedinfo at TREC CDS track 2014: Medical case-based retrieval task
SNUMedinfo at TREC CDS track 2014: Medical case-based retrieval task Sungbin Choi, Jinwook Choi Medical Informatics Laboratory, Seoul National University, Seoul, Republic of Korea wakeup06@empas.com, jinchoi@snu.ac.kr
More informationImproving Difficult Queries by Leveraging Clusters in Term Graph
Improving Difficult Queries by Leveraging Clusters in Term Graph Rajul Anand and Alexander Kotov Department of Computer Science, Wayne State University, Detroit MI 48226, USA {rajulanand,kotov}@wayne.edu
More informationSearching in All the Right Places. How Is Information Organized? Chapter 5: Searching for Truth: Locating Information on the WWW
Chapter 5: Searching for Truth: Locating Information on the WWW Fluency with Information Technology Third Edition by Lawrence Snyder Searching in All the Right Places The Obvious and Familiar To find tax
More informationIRCE at the NTCIR-12 IMine-2 Task
IRCE at the NTCIR-12 IMine-2 Task Ximei Song University of Tsukuba songximei@slis.tsukuba.ac.jp Yuka Egusa National Institute for Educational Policy Research yuka@nier.go.jp Masao Takaku University of
More informationPRIOR System: Results for OAEI 2006
PRIOR System: Results for OAEI 2006 Ming Mao, Yefei Peng University of Pittsburgh, Pittsburgh, PA, USA {mingmao,ypeng}@mail.sis.pitt.edu Abstract. This paper summarizes the results of PRIOR system, which
More informationRMIT University at TREC 2006: Terabyte Track
RMIT University at TREC 2006: Terabyte Track Steven Garcia Falk Scholer Nicholas Lester Milad Shokouhi School of Computer Science and IT RMIT University, GPO Box 2476V Melbourne 3001, Australia 1 Introduction
More informationQuery Likelihood with Negative Query Generation
Query Likelihood with Negative Query Generation Yuanhua Lv Department of Computer Science University of Illinois at Urbana-Champaign Urbana, IL 61801 ylv2@uiuc.edu ChengXiang Zhai Department of Computer
More informationSemantic Entity Retrieval using Web Queries over Structured RDF Data
Semantic Entity Retrieval using Web Queries over Structured RDF Data Jeff Dalton and Sam Huston CS645 Project Final Report May 11, 2010 Abstract We investigate the problem of performing entity retrieval
More informationBook Recommendation based on Social Information
Book Recommendation based on Social Information Chahinez Benkoussas and Patrice Bellot LSIS Aix-Marseille University chahinez.benkoussas@lsis.org patrice.bellot@lsis.org Abstract : In this paper, we present
More informationWeb Searcher Interactions with Multiple Federate Content Collections
Web Searcher Interactions with Multiple Federate Content Collections Amanda Spink Faculty of IT Queensland University of Technology QLD 4001 Australia ah.spink@qut.edu.au Bernard J. Jansen School of IST
More informationFrom Passages into Elements in XML Retrieval
From Passages into Elements in XML Retrieval Kelly Y. Itakura David R. Cheriton School of Computer Science, University of Waterloo 200 Univ. Ave. W. Waterloo, ON, Canada yitakura@cs.uwaterloo.ca Charles
More informationSE Workshop PLAN. What is a Search Engine? Components of a SE. Crawler-Based Search Engines. How Search Engines (SEs) Work?
PLAN SE Workshop Ellen Wilson Olena Zubaryeva Search Engines: How do they work? Search Engine Optimization (SEO) optimize your website How to search? Tricks Practice What is a Search Engine? A page on
More informationUnderstanding the Query: THCIB and THUIS at NTCIR-10 Intent Task
Understanding the Query: THCIB and THUIS at NTCIR-10 Intent Task Yunqing Xia 1 and Sen Na 2 1 Tsinghua University 2 Canon Information Technology (Beijing) Co. Ltd. Before we start Who are we? THUIS is
More informationConcept-Based Interactive Query Expansion
Concept-Based Interactive Query Expansion Bruno M. Fonseca 12 maciel@dcc.ufmg.br Paulo Golgher 2 golgher@akwan.com.br Bruno Pôssas 12 bavep@akwan.com.br Berthier Ribeiro-Neto 1 2 berthier@dcc.ufmg.br Nivio
More informationYork University at CLEF ehealth 2015: Medical Document Retrieval
York University at CLEF ehealth 2015: Medical Document Retrieval Andia Ghoddousi Jimmy Xiangji Huang Information Retrieval and Knowledge Management Research Lab Department of Computer Science and Engineering
More informationCS47300: Web Information Search and Management
CS47300: Web Information Search and Management Web Search Prof. Chris Clifton 17 September 2018 Some slides courtesy Manning, Raghavan, and Schütze Other characteristics Significant duplication Syntactic
More informationLearning to Rank Query Suggestions for Adhoc and Diversity Search
Information Retrieval Journal manuscript No. (will be inserted by the editor) Learning to Rank Query Suggestions for Adhoc and Diversity Search Rodrygo L. T. Santos Craig Macdonald Iadh Ounis Received:
More information