Open research questions in Image Retrieval and Classification Techniques
53 unresolved questions extracted from the limitations and future-work sections of 322 Image Retrieval and Classification Techniques papers in our library. Each links back to the study that raised it.
What the literature leaves open
Large intra-class variations. Small inter-class differences in samples within the set captured in unconstrained environments. Limited performance of existing ISC methods in few-shot image set classification tasks.
DCSCR: a class-specific collaborative representation based network for few-shot image set classification · 2026 · DOIThe rapid growth of video content has made efficient video retrieval a critical challenge. Traditional approaches rely heavily on manual tagging and metadata generation, which are time-consuming, inconsistent, and not scalable. The lack of efficient video retrieval systems that can handle large-scale video data.
There is a need for a data model to add structured data to natural history specimen images on Wikimedia Commons. The existing data model does not support a wider range of Wikimedia Commons queries.
Limited generalization to unseen classes. Complex distributions and long-tailed, cross-domain differences. Heavy dependence on manual annotation.
Visual similarity among architectural styles. Varying illumination conditions. Lack of labelled datasets.
Deep Learning Framework for Indian Heritage Site Classification and Virtual Tour Generation using Transfer Learning · 2026 · DOIThe paper identifies a gap in the literature, which has traditionally focused on central place theory, neglecting the importance of gateways. The study highlights the need to consider both central place and gateway developments in understanding the rise of large centers.
The threshold θ for cosine similarity matching was set conservatively to prioritize precision over recall, but no systematic analysis of threshold selection or optimization methodology is provided.
Testing was limited to four events with maximum 5,000 photos and over 100 guests; performance at significantly larger scale (thousands of guests or tens of thousands of photos) is untested.
The study uses a fixed 70-30 train-test split without justification or ablation study. Cross-validation strategies, different data split ratios, and their impact on BGP-Model robustness across batik motif categories are not explored.
Batik Motif Recognition Using the BGP-Model: A Hybrid GLCM-PCA Approach with Machine Learning Classifiers · 2026 · DOIMisclassification patterns are noted (e.g., Parangsloboh misclassified as Parangklitih) but root cause analysis is absent. Systematic investigation of which GLCM texture features fail to discriminate between visually similar batik motifs would inform feature engineering improvements.
Batik Motif Recognition Using the BGP-Model: A Hybrid GLCM-PCA Approach with Machine Learning Classifiers · 2026 · DOIThere are many security, performance, scalability, and architectural design issues when it comes to implementing such APIs in web apps. Most of the applications designed for academic or research purposes are very much experimental minded.
The lack of paired data can make it challenging to ensure high-quality translations and maintain semantic coherence in the output images. Obtaining paired datasets can be impractical for many applications, as it requires manual annotation of images.
Future research could investigate the application of other quantitative metrics for image quality assessment. The development of more advanced AI image generation tools could lead to further research on the evaluation of their outputs.
Analysis of Digital Image Quality Between Conventional Design Outputs and Generative Artificial Intelligence Images Using Digital Image Processing Methods · 2026 · DOIThe increasing adoption of AI image generation tools has created a need for objective methods to evaluate the quality of their outputs. There is a gap in the development of quantitative tools for image quality assessment.
Analysis of Digital Image Quality Between Conventional Design Outputs and Generative Artificial Intelligence Images Using Digital Image Processing Methods · 2026 · DOIThe difficulties in obtaining power defect samples. The dominance of normal samples. The reliance on large-scale data of multi-modal large models.
Research on the Construction of Multimodal Large Models and Self-supervised Learning for Panoramic Perception of Power Equipment · 2026 · DOIThe need for intelligent automated methods to distinguish relevant posts from irrelevant ones. The challenge of filtering relevant urban improvement content from the overwhelming volume of social media data.
Multimodal detection of urban improvement indicators in public visual-text content using deep learning · 2026 · DOIThe requirement for extensive manual labeling and significant computing power is a major barrier to the implementation of deep learning approaches. The lack of user-friendly and accessible tools for non-AI specialists is a significant challenge.
Existing zero-shot semantic segmentation frameworks are limited by their reliance on manual annotation and closed-world assumptions. The large number of classes and complex distributions make comprehensive annotation virtually impossible.
There is a need for a framework that can model the formal visual language of contemporary Inner Mongolian painting. Earlier content-based retrieval systems usually described paintings through low-level visual cues, but there is a need for more advanced methods.
Multimodal Representation Learning and Visual Analytics for Modeling the Formal Visual Language of Contemporary Inner Mongolian Painting · 2026 · DOIThe underutilization of modern Multimodal Large Language Models in the Digital Humanities. The limited scalability and insufficient real-time adaptability of most existing intelligent multimodal systems.
Designing Intelligent Multimodal Assistants for Digital Humanities: A Comparative Study of Models, Modalities, and Domains · 2026 · DOIComplex backgrounds often led to incomplete or misleading text extraction. Norwegian characters such as Ø, AE, and Å were not consistently recognised, even after modifying the OCR configurations. Irrelevant text extraction remained a recurring challenge, particularly when multiple products appeared within a single cropped image.
Automated product and price information extraction from retail promotional flyers using YOLO and OCR · 2026 · DOITo improve the performance of the approach on complex backgrounds and Norwegian characters. To apply the approach to other retail analytics and automated document processing applications. To evaluate the performance of the approach on a larger dataset.
Automated product and price information extraction from retail promotional flyers using YOLO and OCR · 2026 · DOIExisting traditional ISC methods classify image sets based on raw pixel features. Deep ISC methods learn deep features but fail to adaptively adjust the features when measuring set distances.
DCSCR: a class-specific collaborative representation based network for few-shot image set classification · 2026 · DOIThe dataset was small and imbalanced. Some monuments had fewer than 10 samples. The system requires further development for multilingual support and AR/VR support.
Deep Learning Framework for Indian Heritage Site Classification and Virtual Tour Generation using Transfer Learning · 2026 · DOIThe need for a system that integrates image restoration and caption generation. The challenge of restoring degraded images and generating accurate captions.
Most-cited papers in Image Retrieval and Classification Techniques
- Recognize Anything: A Strong Image Tagging Model · 2024 · 154 citations
- EarthGPT: A Universal Multimodal Large Language Model for Multisensor Image Comprehension in Remote Sensing Domain · IEEE Transactions on Geoscience and Remote Sensing · 2024 · 146 citations
- Conv2Former: A Simple Transformer-Style ConvNet for Visual Recognition · IEEE Transactions on Pattern Analysis and Machine Intelligence · 2024 · 143 citations
- RSDiff: remote sensing image generation from text using diffusion model · Neural Computing and Applications · 2024 · 48 citations
- Folksonomies: Flickr image tagging: Patterns made visible · Bulletin of the American Society for Information Science and Technology · 2007 · 26 citations
- Mobile visual search model for Dunhuang murals in the smart library · Library Hi Tech · 2022 · 24 citations
- Image Clustering: An Unsupervised Approach to Categorize Visual Data in Social Science Research · Sociological Methods & Research · 2022 · 19 citations
- A novel label-based multimodal topic model for social media analysis · Decision Support Systems · 2022 · 18 citations
- Analysis of Clothing Image Classification Models: A Comparison Study between Traditional Machine Learning and Deep Learning Models · Fibres and Textiles in Eastern Europe · 2022 · 17 citations
- Envisaging a global infrastructure to exploit the potential of digitised collections · Biodiversity Data Journal · 2023 · 10 citations
Most recent work
- Unpaired image-to-image translation with content preserving perspective: a review · Multimedia Tools and Applications · 2026
- Noise-Robust Scale-Selective Local Binary Pattern For Texture Classification · Iconic Research and Engineering Journals · 2026
- AI-Driven Guest Identification and Photo Retrieval System · International Journal of Science, Strategic Management and Technology · 2026
- Batik Motif Recognition Using the BGP-Model: A Hybrid GLCM-PCA Approach with Machine Learning Classifiers · Engineering Technology & Applied Science Research · 2026
- Towards the design of hybrid transformer-driven deep representation learning for high-performance video retrieval systems · Discover Computing · 2026
- Unsplash FindAWall: Secure, Scalable and Responsive Image-Search Platform · International Journal of Technology & Emerging Research · 2026
- A Guide to Structureless Visual Localization · International Journal of Computer Vision · 2026
- PhotoDB: a customizable platform for image data management, processing, and annotation · Ecological Informatics · 2026
- Analysis of Digital Image Quality Between Conventional Design Outputs and Generative Artificial Intelligence Images Using Digital Image Processing Methods · Journal of Engineering Electrical and Informatics · 2026
- Research on the Construction of Multimodal Large Models and Self-supervised Learning for Panoramic Perception of Power Equipment · Tehnicki vjesnik - Technical Gazette · 2026
Find a gap in your own Image Retrieval and Classification Techniques sub-topic
This page shows what the Image Retrieval and Classification Techniques literature already flags as unresolved. To narrow it to your specific question, run the guided finder — it searches the gap library on demand and checks candidates against 250M+ OpenAlex works.
Open the Research Gap Finder →