Open research questions in Text Readability and Simplification
245 unresolved questions extracted from the limitations and future-work sections of 1,197 Text Readability and Simplification papers in our library. Each links back to the study that raised it.
What the literature leaves open
future research can use the proposed characterization approach to analyze LLM-generated scientific concept descriptions, - further validation of the methodology is needed, - exploration of other domains and prompt types is suggested
Characterizing LLM scientific concept generation: A multi-dimensional measurement study · 2026 · DOIThe assumption that semantic distance in embedding space serves as a reliable proxy for conceptual novelty is problematic. Embedding models are optimized for semantic similarity, not novelty detection. Different embedding models produce different distance distributions, making absolute thresholds arbitrary.
Characterizing LLM scientific concept generation: A multi-dimensional measurement study · 2026 · DOIThe current understanding of how productive and item-specific knowledge interact in language processing is limited. Prior work has not fully addressed how these knowledge types trade off as a function of frequency. The paper identifies a gap in the understanding of how language processing relies on both productive and item-specific knowledge.
Productive knowledge and item-specific knowledge trade off as a function of frequency in multiword expression processing · 2024 · DOIDespite this widely acknowledged significance, there remains a notable gap in research concerning the evaluation of key factors, such as syntactic complexity and text readability, which determine text difficulty.
Exploring Syntactic Complexity and Text Readability in an ELT Textbook Series for Chinese English Majors · 2025 · DOIThe analysis of participants’ responses to the questionnaire open questions revealed notable benefits and limitations associated with using ChatGPT in literary translation.
Empowering Student Translators: The Impact of ChatGPT Training on Self-Efficacy in Literary Translation · 2025 · DOIIn the Malaysian context, studies have shown that English proficiency is still lacking among undergraduate university students.
Readability analysis as a tool for evaluating English proficiency in first-year medical students · 2025 · DOIThere is a lack of comprehensive comparative analyses of AI-generated and human-written abstracts. The study aims to address this gap by comparing the readability and writing styles of AI-generated and human-written abstracts.
Comparative analysis of text readability and writing styles in AI-generated vs. Human-written academic abstracts · 2026 · DOIFuture research should focus on developing detection tools with improved linguistic capabilities. The study suggests the need for further research on cross-linguistic challenges in AI detection.
ZeroGPT, however, displayed inconsistent results for human-written Croatian texts: One sample was detected as ‘Mixed signals, with some parts generated by AI/GPT’ with a significant AI percentage, another as ‘Likely human-written, may include parts generated by AI/GPT’ with a moderate AI percentage, and a third as ‘Most of your text is AI/GPT-generated’ with a high AI Information Research, Vol.
The proposed framework may not generalize to other domains or languages. The evaluation is limited to the CNLAW dataset, which may not be representative of all legal texts.
Future research can focus on extending the proposed framework to other domains or languages. The use of Monte Carlo Dropout can be explored in other natural language processing tasks.
The study does not provide a comprehensive comparison with existing simplification models. The corpus size and diversity may be limited, potentially affecting the generalizability of the results.
The lack of aligned monolingual parallel datasets tailored to specific linguistic dialects, such as the Shahmukhi dialect. The need for simplified text, particularly for individuals with lower levels of literacy and language learners.
The field is heterogeneous, making it difficult to pool effect sizes or conduct a meta-analysis. There is a lack of studies on certain topics, such as reading and Chinese character acquisition. The review notes that AI-supported pedagogy for Chinese language teaching and learning poses unique challenges and requirements.
Artificial Intelligence in Teaching Chinese as A Target Second Foreign Language: A Scoping Review (2021–2026) · 2026 · DOIThe review is limited to studies published in JCR 2023-2024 SSCI/SCIE Q1 or Q2 venues, which may not capture all relevant research. The field is heterogeneous, making it difficult to pool effect sizes or conduct a meta-analysis. The review did not assign MMAT/JBI scores or present a formal risk-of-bias matrix due to the scoping review design.
Artificial Intelligence in Teaching Chinese as A Target Second Foreign Language: A Scoping Review (2021–2026) · 2026 · DOIThe development of a new research paradigm combining generative AI and linguistic proficiency levels.
Ensuring accessibility compliance in researcher-authored visual components. Addressing technical inaccuracies in LLM-generated informatics framing. Balancing the need for originality and pedagogical depth in contest-style computational thinking tasks.
Two significant gaps emerged in the broader authoring workflow: accessibility compliance in the researcher-authored visual components and technical inaccuracies in the LLM-generated informatics framing. The study used a single-case design, which may limit the generalizability of the findings.
Learners, in turn, become more independent and empowered, with AI granting them agency to explore, test and produce language in ways previously limited by the constraints of print media.
The paper suggests that future research should focus on empirical data collection and analyses to validate the theoretical framework. Future research should also investigate the impact of AI tools on critical thinking and originality in academic writing.
The paper identifies a gap in the use of AI tools in academic writing, particularly in terms of critical thinking and originality. The gap is related to the lack of awareness among students and educators about the limitations and risks of AI tools.
The study does not account for future capabilities of ChatGPT or any other alternative tools. The study used GPT-4o for analysis, which may not be representative of more advanced versions. The study recommends employing emerging AI tools or more advanced versions to examine potential differences and improvements.
ChatGPT-4o as an automated scoring tool for writing assessment: Strengths and weaknesses · 2026 · DOIThe study recommends employing emerging AI tools or more advanced versions to examine potential differences and improvements. The study suggests exploring whether any scoring and feedback differences exist across EFL and ESL contexts. The study recommends examining ChatGPT-generated scores and feedback for fairness and bias depending on students' demographic differences.
ChatGPT-4o as an automated scoring tool for writing assessment: Strengths and weaknesses · 2026 · DOITime constraints in providing feedback. Limited use of open-ended questions in language assessments. Need for efficient and effective writing assessments.
Statistical and qualitative analysis of ChatGPT and human raters in preservice teachers’ writing assessment · 2026 · DOIThe lack of research on the suitability of AI for writing skill assessments. The need to explore the strengths and weaknesses of experts and ChatGPT in providing feedback.
Statistical and qualitative analysis of ChatGPT and human raters in preservice teachers’ writing assessment · 2026 · DOI
Most-cited papers in Text Readability and Simplification
- A new readability yardstick. · Journal of Applied Psychology · 1948 · 3,972 citations
- How Does ChatGPT Perform on the United States Medical Licensing Examination (USMLE)? The Implications of Large Language Models for Medical Education and Knowledge Assessment · JMIR Medical Education · 2023 · 1,852 citations
- A SWOT analysis of ChatGPT: Implications for educational practice and research · Innovations in Education and Teaching International · 2023 · 863 citations
- A computer readability formula designed for machine scoring. · Journal of Applied Psychology · 1975 · 744 citations
- Inference during reading. · Psychological Review · 1992 · 731 citations
- From human writing to artificial intelligence generated text: examining the prospects and potential threats of ChatGPT in academic writing · Biology of Sport · 2023 · 666 citations
- Comparing scientific abstracts generated by ChatGPT to real abstracts with detectors and blinded human reviewers · npj Digital Medicine · 2023 · 595 citations
- Exploring the potential of using an AI language model for automated essay scoring · Research Methods in Applied Linguistics · 2023 · 576 citations
- AI-generated feedback on writing: insights into efficacy and ENL student preference · International Journal of Educational Technology in Higher Education · 2023 · 417 citations
- Comparing the quality of human and ChatGPT feedback of students’ writing · Learning and Instruction · 2024 · 371 citations
Most recent work
- The impact of generative AI on academic reading and writing: a synthesis of recent evidence (2023–2025) · Frontiers in Education · 2026
- Translating Culture-Specific Items in The True Story of Ah Q: A Comparative Evaluation of GPT-4o, KIMI, and Google Translate Under Multimodal Prompts · SAGE Open · 2026
- Generative Artificial Intelligence for Automated Qualitative Feedback: A Cross-Comparison of Prompting Strategies · RELC Journal · 2026
- AI writing detectors are ineffective, unreliable and harmful · English Teaching Practice & Critique · 2026
- Clue before correction: ChatGPT-enhanced strategy for promoting autonomous and reflective language learning · Innovation in Language Learning and Teaching · 2026
- Using AI to adapt reading input for mixed-proficiency EFL learners · Frontiers in Education · 2026
- ChatGPT-4o as an automated scoring tool for writing assessment: Strengths and weaknesses · International Journal of Assessment Tools in Education · 2026
- Statistical and qualitative analysis of ChatGPT and human raters in preservice teachers’ writing assessment · International Journal of Assessment Tools in Education · 2026
- Evaluating Rater Effects of Large Language Models in Automated Essay Scoring: GPT, Claude, Gemini, and DeepSeek · Educational Measurement Issues and Practice · 2026
- The use of Copilot, Gemini and ChatGPT in the context of foreign language learning and teaching: An academic technology review · Contemporary Educational Technology · 2026
Find a gap in your own Text Readability and Simplification sub-topic
This page shows what the Text Readability and Simplification literature already flags as unresolved. To narrow it to your specific question, run the guided finder — it searches the gap library on demand and checks candidates against 250M+ OpenAlex works.
Open the Research Gap Finder →