Between Methodological Expansion and Epistemic Boundary Dissolution: Generative AI, Distributed Epistemic Processes, and the Reorganization of Scientific Responsibility
DOI:
https://doi.org/10.54536/ajmri.v5i5.8515Keywords:
Educational Research, Epistemic Agency, Generative AI, Large Language Models, Mixed MethodsAbstract
Although individual studies have examined the capabilities of generative AI in annotation, coding, item construction, thematic analysis, and data simulation, there has so far been no cross-paradigm framework that systematically integrates technical performance, methodological validity, institutional embedding, and the attribution of scientific responsibility. Against this backdrop, the study examines under what conditions the delegation of knowledge-relevant research operations to large language models (LLMs) should be evaluated as a controlled methodological extension or as an epistemic dissolution of boundaries. Methodologically, it follows a critical integrative evidence synthesis. From 80 hits across the eight thematically differentiated search clusters and five additional articles identified through citation tracking, 58 publications remained in the core analytical corpus after removing duplicates, screening titles and abstracts, and reviewing full texts based on predefined criteria. The predefined criteria restricted the corpus to studies that empirically evaluated LLM-mediated research operations or examined their validity, bias, reproducibility, epistemic efficacy, or governance with direct or transferable relevance to social science and educational research. With regard to the synthesized findings, the evidence indicates that LLMs operationally scale quantitative methods but may in the process foster methodological pseudo-competence, non-classical measurement errors, and synthetic circularity. In qualitative designs, they expand the scope for contrastive interpretation but do not replace situated case understanding; in mixed-methods designs, they reduce translation costs without negating the independent logics of validity of the combined strands. Building on, but analytically distinct from, these synthesized findings, the study advances an exploratory conceptual framework in the form of an original five-level matrix of epistemic delegation, ranging from operational and infrastructural offloading to empirical surrogation. Each level is linked to specific epistemic risks and proportionate verification requirements. In addition, reflexive controllability is operationalized as a governance framework with eight dimensions. Both instruments can be directly applied to research design, quality assurance, institutional rule-making, and methodological training and are transferable across different empirical paradigms. Therefore, generative AI appears neither as a neutral tool nor as a responsible agent of knowledge but rather as a probabilistic epistemic actor within distributed knowledge arrangements, for which the ultimate scientific responsibility remains with identifiable individuals and institutions.
Downloads
References
Abdurahman, S., Atari, M., Karimi-Malekabadi, F., Xue, M. J., Trager, J., Park, P. S., Golazizian, P., Omrani, A., & Dehghani, M. (2024). Perils and opportunities in using large language models in psychological research. PNAS Nexus, 3(7), pgae245. https://doi.org/10.1093/pnasnexus/pgae245
Abdurahman, S., Salkhordeh Ziabari, A., Moore, A. K., Bartels, D. M., & Dehghani, M. (2025). A primer for evaluating large language models in social-science research. Advances in Methods and Practices in Psychological Science, 8(2). https://doi.org/10.1177/25152459251325174
Al-Abdullatif, A. M. (2024). Modeling teachers’ acceptance of generative artificial intelligence use in higher education: The role of AI literacy, intelligent TPACK, and perceived trust. Education Sciences, 14(11), 1209. https://doi.org/10.3390/educsci14111209
Andersen, J. P., Degn, L., Fishberg, R., Graversen, E. K., Horbach, S. P. J. M., Schmidt, E. K., Schneider, J. W., & Sørensen, M. P. (2025). Generative artificial intelligence (GenAI) in the research process—A survey of researchers’ practices and perceptions. Technology in Society, 81, 102813. https://doi.org/10.1016/j.techsoc.2025.102813
Argyle, L. P., Busby, E. C., Fulda, N., Gubler, J. R., Rytting, C., & Wingate, D. (2023). Out of one, many: Using language models to simulate human samples. Political Analysis, 31(3), 337–351. https://doi.org/10.1017/pan.2023.2
Ashokkumar, A., Hewitt, L., Ghezae, I., & Willer, R. (2026). Large language models can predict the results of social science experiments. Nature, 656(8126), 115–122. https://doi.org/10.1038/s41586-026-10742-x
Ashwin, J., Chhabra, A., & Rao, V. (2026). Using large language models for qualitative analysis can introduce serious bias. Sociological Methods & Research, 55(3), 795–839. https://doi.org/10.1177/00491241251338246
Balt, E., Salmi, S., Bhulai, S., Vrinzen, S., Eikelenboom, M., Gilissen, R., Creemers, D., Popma, A., & Mérelle, S. (2025). Deductively coding psychosocial autopsy interview data using a few-shot learning large language model. Frontiers in Public Health, 13, 1512537. https://doi.org/10.3389/fpubh.2025.1512537
Bender, E. M., Gebru, T., McMillan-Major, A., & Shmitchell, S. (2021). On the dangers of stochastic parrots: Can language models be too big? In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (pp. 610–623). Association for Computing Machinery. https://doi.org/10.1145/3442188.3445922
Binz, M., Alaniz, S., Roskies, A., Aczel, B., Bergstrom, C. T., Allen, C., Schad, D., Wulff, D., West, J. D., Zhang, Q., Shiffrin, R. M., Gershman, S. J., Popov, V., Bender, E. M., Marelli, M., Botvinick, M. M., Akata, Z., & Schulz, E. (2025). How should the advancement of large language models affect the practice of science? Proceedings of the National Academy of Sciences, 122(5), e2401227121. https://doi.org/10.1073/pnas.2401227121
Birhane, A., Kasirzadeh, A., Leslie, D., & Wachter, S. (2023). Science in the age of large language models. Nature Reviews Physics, 5, 277–280. https://doi.org/10.1038/s42254-023-00581-4
Bisbee, J., Clinton, J. D., Dorff, C., Kenkel, B., & Larson, J. M. (2024). Synthetic replacements for human survey data? The perils of large language models. Political Analysis, 32(4), 401–416. https://doi.org/10.1017/pan.2024.5
Blanchard, S. J., Duani, N., Garvey, A. M., Netzer, O., & Oh, T. T. (2025). New tools, new rules: A practical guide to effective and responsible generative AI use for surveys and experiments in research. Journal of Marketing, 89(6), 119–139. https://doi.org/10.1177/00222429251349882
Chang, Y.-C., Wang, X., Wang, J., Wu, Y., Yang, L., Zhu, K., Chen, H., Yi, X., Wang, C., Wang, Y., Ye, W., Zhang, Y., Chang, Y., Yu, P. S., Yang, Q., & Xie, X. (2024). A survey on evaluation of large language models. ACM Transactions on Intelligent Systems and Technology, 15(3), 1–45. https://doi.org/10.1145/3641289
Combrinck, C. (2024). A tutorial for integrating generative AI in mixed methods data analysis. Discover Education, 3, 116. https://doi.org/10.1007/s44217-024-00214-7
de Cassai, A., Dost, B., Augoustides, J. G. T., Azamfirei, L., Alanoğlu, Z., Azi, L. M., Calvache, J. A., Cerny, V., De Hert, S., Eldawlatly, A., Farber, M. K., Sobreira-Fernandes, D., Fettiplace, M. R., Galante, D., Garg, R., Goldstein, H. V., Abad-Gurumeta, A., Gupta, L., Hemmings, H. C., … Zdanowski, S. (2026). Responsible use of large language models in manuscript authorship, peer review, and editorial processes: A Delphi consensus among editors-in-chief of anaesthesia and pain medicine journals (RULE-AP). British Journal of Anaesthesia, 136(5), 1625–1633. https://doi.org/10.1016/j.bja.2026.01.029
Dengel, A., Gehrlein, R., Fernes, D., Görlich, S., Maurer, J., Pham, H. H., Großmann, G., & Eisermann, N. D. G. (2023). Qualitative research methods for large language models: Conducting semi-structured interviews with ChatGPT and BARD on computer science education. Informatics, 10(4), 78. https://doi.org/10.3390/informatics10040078
De Paoli, S. (2024). Performing an inductive thematic analysis of semi-structured interviews with a large language model: An exploration and provocation on the limits of the approach. Social Science Computer Review, 42(4), 997–1019. https://doi.org/10.1177/08944393231220483
Deutsche Forschungsgemeinschaft. (2025). Guidelines for safeguarding good research practice: Code of conduct (Version 3). https://doi.org/10.5281/zenodo.14281892
Dilek, M., Baran, E., & Aleman, E. (2025). AI literacy in teacher education: Empowering educators through critical co-discovery. Journal of Teacher Education, 76(3), 294–311. https://doi.org/10.1177/00224871251325083
Floridi, L. (2023). AI as agency without intelligence: On ChatGPT, large language models, and other generative models. Philosophy & Technology, 36, 15. https://doi.org/10.1007/s13347-023-00621-y
Halterman, A., & Keith, K. A. (2026). Codebook LLMs: Evaluating LLMs as measurement tools for political science concepts. Political Analysis, 34(2), 188–204. https://doi.org/10.1017/pan.2025.10017
Haltaufderheide, J., & Ranisch, R. (2024). The ethics of ChatGPT in medicine and healthcare: A systematic review on large language models (LLMs). npj Digital Medicine, 7, 183. https://doi.org/10.1038/s41746-024-01157-x
Haroud, S., & Saqri, N. (2025). Generative AI in higher education: Teachers’ and students’ perspectives on support, replacement, and digital literacy. Education Sciences, 15(4), 396. https://doi.org/10.3390/educsci15040396
Hayes, A. S. (2025). “Conversing” with qualitative data: Enhancing qualitative research through large language models (LLMs). International Journal of Qualitative Methods, 24, 1–19. https://doi.org/10.1177/16094069251322346
Hepp, A. (2010). Cultural Studies und Medienanalyse: Eine Einführung [Cultural studies and media analysis: An introduction] (3rd ed.). VS Verlag für Sozialwissenschaften. https://doi.org/10.1007/978-3-531-92190-7
Jose, B., Cleetus, A., Joseph, B., Joseph, L., Jose, B., & John, A. K. (2025). Epistemic authority and generative AI in learning spaces: Rethinking knowledge in the algorithmic age. Frontiers in Education, 10, 1647687. https://doi.org/10.3389/feduc.2025.1647687
Keane, D., & McNaughton, R. B. (2026). Using generative AI to enhance psychometric scale development in market research. International Journal of Market Research, 68(2), 194–218. https://doi.org/10.1177/14707853251384769
Leslie, D. (2025). Does the sun rise for ChatGPT? Scientific discovery in the age of generative AI. AI and Ethics, 5(4), 3439–3444. https://doi.org/10.1007/s43681-023-00315-3
Lin, H., & Zhang, Y. (2026). Navigating the risks of using large language models for text annotation in social science research. Social Science Computer Review, 44(3), 403–427. https://doi.org/10.1177/08944393251366243
Lin, Z. (2023). Why and how to embrace AI such as ChatGPT in your academic life. Royal Society Open Science, 10(8), 230658. https://doi.org/10.1098/rsos.230658
Lin, Z. (2026). A validity-guided workflow for robust large language model research in psychology. Behavior Research Methods, 58(8), Article 216. https://doi.org/10.3758/s13428-026-03073-2
Martin Kowal, J., Hurley Bryant, K., Segall, D., & Kantrowitz, T. (2025). Harnessing generative AI for assessment item development: Comparing AI-generated and human-authored items. International Journal of Selection and Assessment, 33(3), e70021. https://doi.org/10.1111/ijsa.70021
Mathis, W. S., Zhao, S., Pratt, N., Weleff, J., & De Paoli, S. (2024). Inductive thematic analysis of healthcare qualitative interviews using open-source large language models: How does it compare to traditional methods? Computer Methods and Programs in Biomedicine, 255, 108356. https://doi.org/10.1016/j.cmpb.2024.108356
Messeri, L., & Crockett, M. J. (2024). Artificial intelligence and illusions of understanding in scientific research. Nature, 627, 49–58. https://doi.org/10.1038/s41586-024-07146-0
Mou, X., Ding, X., He, Q., Wang, L., Liang, J., Zhang, X., Sun, L., Lin, J., Zhou, J., Huang, X., & Wei, Z. (2026). From individual to society: A survey on social simulation driven by large language model-based agents. ACM Computing Surveys, 58(11), 1–41. https://doi.org/10.1145/3800683
Olaniyan, Y. D., Martins, M. O., & Al Maqrashi, R. H. (2026). Generative artificial intelligence and epistemic (in)justice: Perspectives from higher education students in the Global North and South. Frontiers in Human Dynamics, 8, 1790324. https://doi.org/10.3389/fhumd.2026.1790324
Pangakis, N., Wolken, S., & Fasching, N. (2023). Automated annotation with generative AI requires validation [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2306.00176
Park, P. S., Schoenegger, P., & Zhu, C. (2024). Diminished diversity-of-thought in a standard large language model. Behavior Research Methods, 56(6), 5754–5770. https://doi.org/10.3758/s13428-023-02307-x
Peters, U., & Chin-Yee, B. (2025). Generalization bias in large language model summarization of scientific research. Royal Society Open Science, 12, 241776. https://doi.org/10.1098/rsos.241776
Prandner, D., Wetzelhütter, D., & Hese, S. (2025). ChatGPT as a data analyst: An exploratory study on AI-supported quantitative data analysis in empirical research. Frontiers in Education, 9, 1417900. https://doi.org/10.3389/feduc.2024.1417900
Qian, Y. (2025). Pedagogical applications of generative AI in higher education: A systematic review of the field. TechTrends, 69(5), 1105–1120. https://doi.org/10.1007/s11528-025-01100-1
Qiao, S., Fang, X., Wang, J., Zhang, R., Li, X., & Kang, Y. (2025). Generative AI for thematic analysis in a maternal health study: Coding semistructured interviews using large language models. Applied Psychology: Health and Well-Being, 17(3), e70038. https://doi.org/10.1111/aphw.70038
Saúde, S., Barros, J. P., & Almeida, I. (2024). Impacts of generative artificial intelligence in higher education: Research trends and students’ perceptions. Social Sciences, 13(8), 410. https://doi.org/10.3390/socsci13080410
Tai, R. H., Bentley, L. R., Xia, X., Sitt, J. M., Fankhauser, S. C., Chicas-Mosier, A. M., & Monteith, B. G. (2024). An examination of the use of large language models to aid analysis of textual data. International Journal of Qualitative Methods, 23, 1–14. https://doi.org/10.1177/16094069241231168
Terry, J., Strait, G., Alsarraf, S., Weinmann, E., & Waychoff, A. (2025). Artificial intelligence in scale development: Evaluating AI-generated survey items against gold standard measures. Current Psychology, 44(20), 16339–16350. https://doi.org/10.1007/s12144-025-08240-w
Törnberg, P. (2025). Large language models outperform expert coders and supervised classifiers at annotating political social media messages. Social Science Computer Review, 43(6), 1181–1195. https://doi.org/10.1177/08944393241286471
Tzirides, A. O., Zapata, G. C., Kastania, N. P., Saini, A. K., Castro, V., Ismael, S. A., You, Y.-L., Santos, T. A. D., Searsmith, D., O’Brien, C., Cope, B., & Kalantzis, M. (2024). Combining human and artificial intelligence for enhanced AI literacy in higher education. Computers and Education Open, 6, 100184. https://doi.org/10.1016/j.caeo.2024.100184
van Dis, E. A. M., Bollen, J., Zuidema, W., van Rooij, R., & Bockting, C. L. (2023). ChatGPT: Five priorities for research. Nature, 614, 224–226. https://doi.org/10.1038/d41586-023-00288-7
Verma, S., Kashive, N., & Gupta, A. (2026). Examining predictors of generative-AI acceptance and usage in academic research: A sequential mixed-methods approach. Benchmarking: An International Journal, 33(6), 1821–1849. https://doi.org/10.1108/BIJ-07-2024-0564
Vindigni, G. (2024a). Enhancing human-computer interaction in socially inclusive contexts: Flow heuristics and AI systems in compliance with DIN EN ISO 9241 standards. European Journal of Contemporary Education and E-Learning, 2(4), 115–139. https://doi.org/10.59324/ejceel.2024.2(4).10
Vindigni, G. (2024b). Optimierung der HCI in sozial inklusiven Kontexten: Flow-Heuristiken und KI-Systeme gemäß DIN EN ISO 9241. Zeitschrift für Sozialmanagement, 22(2), 111–123.
Vindigni, G. (2025a). Data-driven disparities: How AI applications in education may perpetuate or mitigate inequality. European Journal of Science and Modern Technologies, 1(4), 4–54. https://doi.org/10.59324/ejsmt.2025.1(4).02
Vindigni, G. (2025b). Gender bias and cultural misrepresentation in AI: A critical inquiry into cross-cultural communication and algorithmic design. European Journal of Applied Science, Engineering and Technology, 3(3), 51–72. https://doi.org/10.59324/ejaset.2025.3(3).06
Vindigni, G. (2026a). Platformized private higher education institutions in the EHEA: Instructional infrastructure, access conditionality, and platform governance. European Journal of Contemporary Education and E-Learning, 4(4), 275–356. https://doi.org/10.59324/ejceel.2026.4(4).13
Vindigni, G. (2026b). Beyond compliance: Future Skills 2030, digital accessibility and democratic participation in higher education—A scoping review. European Journal of Contemporary Education and E-Learning, 4(4), 56–137. https://doi.org/10.59324/ejceel.2026.4(4).05
Vindigni, G. (2026c). Digital exams as socio-technical interaction systems: An ISO 9241-based framework for validity, fairness and governance. American Journal of Education and Technology, 5(3), 61–93. https://doi.org/10.54536/ajet.v5i3.8137
Wachinger, J., Bärnighausen, K., Schäfer, L. N., Scott, K., & McMahon, S. A. (2025). Prompts, pearls, imperfections: Comparing ChatGPT and a human researcher in qualitative data analysis. Qualitative Health Research, 35(9), 951–966. https://doi.org/10.1177/10497323241244669
Wang, J., Bai, B., & An, Q. (2026). Enhancing college students’ AI literacy through generative AI use: A mixed-methods investigation. Frontiers in Psychology, 17, 1728785. https://doi.org/10.3389/fpsyg.2026.1728785
Weng, X., Xia, Q., Gu, M., Rajaram, K., & Chiu, T. K. F. (2024). Assessment and learning outcomes for generative AI in higher education: A scoping review on current research status and trends. Australasian Journal of Educational Technology, 40(6), 37–55. https://doi.org/10.14742/ajet.9540
Wu, J.-Y., Lee, Y.-H., Chai, C. S., & Tsai, C.-C. (2025). Strengthening human epistemic agency in the symbiotic learning partnership with generative artificial intelligence. Educational Researcher, 54(6), 358–368. https://doi.org/10.3102/0013189X251333628
Yan, L., Pammer-Schindler, V., Mills, C., Nguyen, A., & Gašević, D. (2025). Beyond efficiency: Empirical insights on generative AI’s impact on cognition, metacognition and epistemic agency in learning. British Journal of Educational Technology, 56(5), 1675–1685. https://doi.org/10.1111/bjet.70000
Yan, L., Sha, L., Zhao, L., Li, Y., Martinez-Maldonado, R., Chen, G., Li, X., Jin, Y., & Gašević, D. (2024). Practical and ethical challenges of large language models in education: A systematic scoping review. British Journal of Educational Technology, 55(1), 90–112. https://doi.org/10.1111/bjet.13370
Zhang, Y., & Dong, C. (2024). Exploring the digital transformation of generative AI-assisted foreign language education: A socio-technical systems perspective based on mixed methods. Systems, 12(11), 462. https://doi.org/10.3390/systems12110462
Ziems, C., Held, W., Shaikh, O., Chen, J., Zhang, Z., & Yang, D. (2024). Can large language models transform computational social science? Computational Linguistics, 50(1), 237–291. https://doi.org/10.1162/coli_a_00502
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Giovanni Vindigni

This work is licensed under a Creative Commons Attribution 4.0 International License.



