Open research questions in Information Retrieval and Search Behavior
58 unresolved questions extracted from the limitations and future-work sections of 2,260 Information Retrieval and Search Behavior papers in our library. Each links back to the study that raised it.
What the literature leaves open
Volume 4 | Issue 4 | April 2026 Discussion and Conclusion Most respondents reported that they have only limited knowledge of Open Access. These challenges point towards gaps in knowledge and skills rather than problems with the resources themselves.
Impact of Open Access Resources on Scholarly Communication: An Empirical Study of University of Jammu Research Scholars · 2026 · DOIOriginality/value While research on SEO has been conducted across various fields, no studies have yet explored the intersection of university Web visibility and searches for information on artificial intelligence in search engines.
Search engine visibility of university AI content: an analysis of thirty institutions from three countries · 2025 · DOIUse of LIS models, theories, and concepts was limited; while this reflects the multidisciplinary nature of TikTok research, it also meant that aspects of IB, such as re-finding, avoidance, and discovery, were underexplored and undertheorized.
We structure this paper around five key opportunities for AI assistance in scholarly reading —discovery, efficiency, comprehension, synthesis, and accessibility—and present an overview of our progress and discuss remaining open challenges.
The AI techniques used mostly in the scanned university libraries are self-directed learning and natural language processing techniques; while the challenges of using AI for reference services are the problem of quality intelligence, linguistic style, privacy, a threat to intellectual freedom, bias, and cost; inadequate experts, poor network, poor training and lack of innovation, and limited knowledge about the technology.
Application of Artificial Intelligence for Reference Services in Academic Libraries: A Global Overview through a Systematic Review of Literature · 2023 · DOISpecifically, because eye-tracking captures automatic as well as conscious processes, it is currently an open question how reliably and consistently eye-tracking captures the strategic sourcing processes that take place during multiple document reading, in particular compared to subjective methods that mainly target conscious processes, such as interviews.
Using eye-tracking to assess sourcing during multiple document reading: A critical analysis · 2018 · DOIOriginality/value – The general consensus is that uncertainty is a mental state of users reflecting a gap in knowledge which triggers an IS&R process, and that the gap is reduced as relevant information is found, and thus that the uncertainty disappears as the search process concludes.
Although highly recommended query clarification techniques, especially using the follow-up question before logging off, are generally prescribed to improve accuracy, only 50 percent of librarians used follow-up questions and 33 percent of all questions asked to users were open questions.
is extension of adaptive hermeneutic agent component, which was only partially implemented. Especially, to fully utilize it, the profile needs to be directly editable and al- low for more direct specification of preferences. Also meta- model, even though it is sufficient to describe a wide range of real world applications, needs to be extended, if it is to be used for construction of more comprehensive knowl- edge management applications. One simple, yet very power- ful, addition would be introduction of processes [9], which are widely used for description of sequences of actions and, thus, well suited to support many of management tasks. Other more long-term possibilities include addition of dif- ferent algorithms for constructive manipulation of data gathered in presented system. For example, network struc- tures could be analyzed, to compute impact factor of ob- jects [10]. Similar approach is used by some search en- gines [11], and would extend scope of potential applica- tions. Also interesting would be formalization of semantic structure of this system, built upon work done in fields of ontological engineering and semantic web [12], [13]. This would make the data amenable to more intricate automatic processing.
Abstract A large number of studies have investigated the transaction log of general‐purpose search engines such as Excite and AltaVista, but few studies have reported on the analysis of search logs for search engines that are limited to particular Web sites, namely, Web site search engines.
There are general methods and principles by which domains should be explored in IS (see, for instance, the UNISIST model presented in Fjordback Søndergaard, Andersen and Hjøorland) as displayed at a poster session at this conference.
Thus, indexing techniques that are more suited to the way people categorize and name things should be investigated for their potential application to IR system design.
Until the 1980s it had not been feasible to run large-scale IR tests because large full-text collections in machine-readable form did not exist, storage capacity was limited and computer processing speeds were insufficient.
An Overview of the <i>Annual Review of Information Science and Technology</i> and Summary of Volume 32 · 1998 · DOIThese issues fall into two broad categories: The Internet environment (its culture, virtual communities and netiquette; agency and authority; the nature of “publication;” the importance and lack of standards; and searching tools and processes) and the process of guide construction (the importance of people; the nature of “resources” on the Net; intellectual property; levels of connectivity; and time).
Our discipline is physical education. Those in the profession who pursue a historical focus are, of necessity, im pelled toward other disciplines if their scholarly labors are to reflect excellence in addition to enthusiasm. No accessible book or manuscript should be so far out of reach that we are unable to locate it. Our reach is certain to be longer and our grasp surer with the assistance provided by the knowledge and use of bibliographic aids and in formation storage and retrieval systems. Physical education in general and the area of history of physical education in particular will be better served if the mounting research we are accumulat ing is available in more readily acces D sible and useful form.
This should not, however, be construed as a test of the user, since he is limited by the degree to which shorter formats ac- the full text which curately represent the content of judgement of 3 Titles were submitted together with bibliographic citations ; titles contained on the average some 5 to 9 words.
Since few studies have methodically reviewed current publications, researchers and practitioners are unable to take full advantage of existing achievements, which, in turn, limits their progress in this field.
These and other findings discussed here suggest that retrospective application of the vocabulary using automated means should be investigated by catalogers and other technical services librarians.
On the State of Genre/Form Vocabulary: A Quantitative Analysis of LCGFT Data in WorldCat · 2021 · DOIAlthough related research exists on information manipulation and the importance of online communities, few studies have directly discussed the influence of key information on the fan pages of university brands.
information encountering, accidental information discovery, incidental information acquisition), the scope of the phenomenon has not been clearly defined and its nature was not fully understood or fleshed-out.
Research limitations/implications The method described in this paper does not account for direct access to Citable Content Downloads that originate outside Google Search properties.
Originality/value – This study provides important insights into the underinvestigated area of digitised newspaper collections, and shows the importance of webometric methods in analysing online user behaviour.
Exploring the information behaviour of users of Welsh Newspapers Online through web log analysis · 2016 · DOIAfter all, just 2 decades ago published media were limited to the tight controls of a few publishing companies that maintained a large staff of quality-control editors.
What are the basic approaches to search to ensure that end users will be able to find their data easily, quickly and accurately? The kinds of search algorithms used to build and implement search software systems vary widely.
Abstract When users have poorly defined or complex goals, search interfaces that offer only keyword‐searching facilities provide inadequate support to help them reach their information‐seeking objectives.
Most-cited papers in Information Retrieval and Search Behavior
- The concept of relevance in IR · Journal of the American Society for Information Science and Technology · 2003 · 316 citations
- Relevance: The whole history · Journal of the American Society for Information Science · 1997 · 203 citations
- The “so-called” UGC: an updated definition of user-generated content in the age of social media · Online Information Review · 2021 · 139 citations
- Mapping knowledge structure by keyword co-occurrence and social network analysis · Library Hi Tech · 2018 · 120 citations
- Scholarly communication and possible changes in the context of social media · The Electronic Library · 2011 · 93 citations
- Who will you ask? An empirical study of interpersonal task information seeking · Journal of the American Society for Information Science and Technology · 2006 · 90 citations
- Information encountering re-encountered · Journal of Documentation · 2020 · 82 citations
- Relevance reconsidered—Towards an agenda for the 21st century: Introduction to special topic issue on relevance research · Journal of the American Society for Information Science · 1994 · 75 citations
- Analysis of the query logs of a Web site search engine · Journal of the American Society for Information Science and Technology · 2005 · 71 citations
- Evolution of research topics in LIS between 1996 and 2019: an analysis based on latent Dirichlet allocation topic model · Scientometrics · 2020 · 70 citations
Most recent work
- Semantic Exhaustion: A Case Study in the Cost of Zero-Source Entity Substitution — Composition-Layer Substitution, the Scalar Transcription Homology, and the Recursive Atrophy Cost (v1.0) · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Patterns of Information-Seeking Behavior and Digital Access for Agrarian and Paramedical Community Users: A Study of Apex institution in Kota, India · International Journal of Environmental & Agriculture Research · 2026
- ReadSphere – Your Reading World in One Website · INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT · 2026
- CLEF HIPE-2026 - Shared Task Participation Guidelines · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Advancing information retrieval: a bibliometric and topic modeling perspective · Information Discovery and Delivery · 2026
- From the click race to the citation game: a conceptual exploration of the shift from search engine optimisation to generative engine optimisation · Information Research an international electronic journal · 2026
- Impact of Open Access Resources on Scholarly Communication: An Empirical Study of University of Jammu Research Scholars · Zenodo (CERN European Organization for Nuclear Research) · 2026
- Information behaviours and the inner life · Information Research an international electronic journal · 2026
- Who’s a heretic now? Conceptualising information seeking on controversial topics · Information Research an international electronic journal · 2026
- Optimal Web Page Retrieval Using Evolutionary Genetic Algorithm · International Scientific Journal of Engineering and Management · 2026
Find a gap in your own Information Retrieval and Search Behavior sub-topic
This page shows what the Information Retrieval and Search Behavior literature already flags as unresolved. To narrow it to your specific question, run the guided finder — it searches the gap library on demand and checks candidates against 250M+ OpenAlex works.
Open the Research Gap Finder →