THE EPISTEMOLOGY OF ARTIFICIAL INTELLIGENCE AND THE CRISIS OF LINGUISTIC VALIDITY IN ARABIC LANGUAGE STUDIES
Keywords:
Epistemology of AI, Linguistic Validity, Large Language Models, Arabic Language, Epistemological HallucinationAbstract
Large Language Models (LLMs) have achieved demonstrable syntactic competence, yet this article argues that their outputs remain epistemologically deficient across three critical dimensions: semantic adequacy, pragmatic legitimacy, and epistemic accountability. Drawing on Wittgenstein, Chomsky, Searle, and Goldman, this study introduces a Four-Dimensional Linguistic Validity Model and the concept of epistemological hallucination to analyze AI outputs in Arabic language studies. The findings establish that AI’s formal-linguistic strength coexists with structural limitations irreducible by scale, particularly given the civilizational depth of Arabic as an epistemic language.
References
Abudalfa, S., Saad, M., & El-Beltagy, S. (2025). Editorial: Emerging techniques in Arabic natural language processing. Frontiers in Artificial Intelligence, 8. https://doi.org/10.3389/frai.2025.1715520
Ainirrohmah, N., Arridho, A. R., Makrom, C., Shofiyah, & Prasetya, B. (2026). Konsep Pendidikan Filsafat Menurut Al-Farabi dan Ibnu Sina. Jurnal Pendidikan Agama Islam, 2(2). https://doi.org/10.59829/1s5jkh80
Al-Dadi, M. bin A. (2025). The Linguistic Meaning Between Traditional Rhetorical Frameworks and Artificial Intelligence Models: A Post humanist Critique of Computational Semantics. Journal of Posthumanism, 5(7). https://doi.org/10.63332/joph.v5i7.2888
Al-Rawafi, A., Sudana, D., Lukmana, I., & Syihabuddin, S. (2021). Students’ apologizing in Arabic and English: An interlanguage pragmatic case study at an Islamic boarding school in Indonesia. Indonesian Journal of Applied Linguistics, 10(3). https://doi.org/10.17509/ijal.v10i3.31740
Alharbi, W. (2023). AI in the Foreign Language Classroom: A Pedagogical Overview of Automated Writing Assistance Tools. Education Research International, 2023, 1–15. https://doi.org/10.1155/2023/4253331
Andresen, J. T. (1990). Skinner and Chomsky thirty years later. Historiographia Linguistica, 17(1–2), 145–165. https://doi.org/10.1075/hl.17.1-2.12and
Arum, K. (2018). Pengembangan Pendidikan Agama Islam Berbasis Sosial Profetik (Analisis Terhadap Pemikiran Kuntowijoyo). Millah: Journal of Religious Studies, 177–196. https://doi.org/10.20885/millah.vol17.iss2.art2
Beuls, K., & Van Eecke, P. (2024). Humans Learn Language from Situated Communicative Interactions. What about Machines? Computational Linguistics, 50(4), 1277–1311. https://doi.org/10.1162/coli_a_00534
Bhojani, A.-R., & Schwarting, M. (2023). Truth and Regret: Large Language Models, the Quran, and Misinformation. Theology and Science, 21(4), 557–563. https://doi.org/10.1080/14746700.2023.2255944
Bishop, J. M. (2021). Artificial Intelligence Is Stupid and Causal Reasoning Will Not Fix It. Frontiers in Psychology, 11. https://doi.org/10.3389/fpsyg.2020.513474
Buhori, B., & Wahidah, B. (2017). Bahasa Arab dan Peradaban Islam: Telaah atas Sejarah Perkembangan Bahasa Arab dalam Lintas Sejarah Peradaban Islam. Al-Hikmah, 11(1). https://doi.org/10.24260/al-hikmah.v11i1.822
Campanini, M. (2020). The Evidence of Meaning ( bayān al-maʿnā ) in the Ẓāhirī Approach to the Qur’an. Journal of Qur’anic Studies, 22(1), 172–191. https://doi.org/10.3366/jqs.2020.0415
Cholily, N., Ghozali, M., Kholid, A., Wan Mokhtar, W. K. A., & Mahmut, R. İ. (2025). Bridging Fiqh and Religious Practice: Actualizing the Function of Ḥāshiyah as a Form of Worship in the Scribal Traditions of Madurese Pesantren Literature. Journal of Islamic Law, 6(1), 21–45. https://doi.org/10.24260/jil.v6i1.3749
Chomsky, N. (1959). Verbal behavior. By B. F. Skinner. (The Century Psychology Series.) Pp. viii, 478. New York: Appleton-Century-Crofts, Inc., 1957. Language, 35(1), 26–58. https://doi.org/10.2307/411334
Christof, M., & Armoundas, A. A. (2025). Implications of integrating large language models into clinical decision making. Communications Medicine, 5(1), 490. https://doi.org/10.1038/s43856-025-01216-8
Coeckelbergh, M. (2018). Technology Games: Using Wittgenstein for Understanding and Evaluating Technology. Science and Engineering Ethics, 24(5), 1503–1519. https://doi.org/10.1007/s11948-017-9953-8
Collins, H. (2021). The science of artificial intelligence and its critics. Interdisciplinary Science Reviews, 46(1–2), 53–70. https://doi.org/10.1080/03080188.2020.1840821
Contreras Kallens, P., Kristensen‐McLachlan, R. D., & Christiansen, M. H. (2023). Large Language Models Demonstrate the Potential of Statistical Learning in Language. Cognitive Science, 47(3). https://doi.org/10.1111/cogs.13256
Dentella, V., Günther, F., & Leivada, E. (2023). Systematic testing of three Language Models reveals low language accuracy, absence of response stability, and a yes-response bias. Proceedings of the National Academy of Sciences, 120(51). https://doi.org/10.1073/pnas.2309583120
Dentella, V., Günther, F., & Leivada, E. (2025). Language in vivo vs. in silico: Size matters but Larger Language Models still do not comprehend language on a par with humans due to impenetrable semantic reference. PLOS One, 20(7), e0327794. https://doi.org/10.1371/journal.pone.0327794
Dentella, V., Günther, F., Murphy, E., Marcus, G., & Leivada, E. (2024). Testing AI on language comprehension tasks reveals insensitivity to underlying meaning. Scientific Reports, 14(1), 1–11. https://doi.org/10.1038/s41598-024-79531-8
Dokic, K., Pisker, B., & Radisic, B. (2025). Mirroring Cultural Dominance: Disclosing Large Language Models Social Values, Attitudes and Stereotypes. Societies, 15(5), 142. https://doi.org/10.3390/soc15050142
Erdocia, I., Migge, B., & Schneider, B. (2024). Language is not a data set—Why overcoming ideologies of dataism is more important than ever in the age of AI. Journal of Sociolinguistics, 28(5), 20–25. https://doi.org/10.1111/josl.12680
Faisal, M. (2024). Dampak Kecerdasan Buatan (AI) terhadap Pola Pikir Cerdas Mahasiswa di Pontianak. NUCLEUS, 5(1), 60–66. https://doi.org/10.37010/nuc.v5i1.1684
Farquhar, S., Kossen, J., Kuhn, L., & Gal, Y. (2024). Detecting hallucinations in large language models using semantic entropy. Nature, 630(8017), 625–630. https://doi.org/10.1038/s41586-024-07421-0
Ferrario, A., Facchini, A., & Termine, A. (2024). Experts or Authorities? The Strange Case of the Presumed Epistemic Superiority of Artificial Intelligence Systems. Minds and Machines, 34(3), 1–27. https://doi.org/10.1007/s11023-024-09681-1
Fromkin, V. (1968). Speculations on performance models. Journal of Linguistics, 4(1), 47–68. https://doi.org/10.1017/S002222670000164X
Hajrah, Tang, R., Tahmir, S., & Daeng, K. (2019). Reconceptualization of local wisdom through kelong makassar: A semiotic review of michael riffaterre. Journal of Language Teaching and Research, 10(6), 1209–1216. https://doi.org/10.17507/jltr.1006.08
Heersmink, R., de Rooij, B., Clavel Vázquez, M. J., & Colombo, M. (2024). A phenomenology and epistemology of large language models: transparency, trust, and trustworthiness. Ethics and Information Technology, 26(3), 1–15. https://doi.org/10.1007/s10676-024-09777-3
Huttunen, R., & Kakkori, L. (2022). Heidegger’s critique of the technology and the educational ecological imperative. Educational Philosophy and Theory, 54(5), 630–642. https://doi.org/10.1080/00131857.2021.1903436
Irsyady, K. A., Qudsy, S. Z., Wijayati, M., & Muasssomah, M. (2023). The Authorship Of Shaykh Nawawi Al-Bantani In Arabic Linguistics Studies. Journal Of Indonesian Islam, 17(2), 259. https://doi.org/10.15642/JIIS.2023.17.2.259-282
Jidan, F. (2022). Perkembangan Ilmu Balaghah. Imtiyaz: Jurnal Ilmu Keislaman, 6(2), 142–150. https://doi.org/10.46773/imtiyaz.v6i2.355
Kreps, S., McCain, R. M., & Brundage, M. (2022). All the News That’s Fit to Fabricate: AI-Generated Text as a Tool of Media Misinformation. Journal of Experimental Political Science, 9(1), 104–117. https://doi.org/10.1017/XPS.2020.37
Kurnaz, S. (2017). Who is the Lawgiver? The Hermeneutical Grounds of the Methods of Interpreting Qur’an and Sunna (istinbāṭ al-aḥkām). Oxford Journal of Law and Religion, 6(2), 347–371. https://doi.org/10.1093/ojlr/rwx007
Lan, N., Chemla, E., & Katzir, R. (2024). Large Language Models and the Argument from the Poverty of the Stimulus. Linguistic Inquiry, 1–28. https://doi.org/10.1162/ling_a_00533
Lbs, M. (2020). Konsep Pendidikan Menurut Pemikiran Kh. Hasyim Asy’ari. Jurnal As-Salam, 4(1), 79–94. https://doi.org/10.37249/as-salam.v4i1.170
Lederman, H., & Mahowald, K. (2024). Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs. Transactions of the Association for Computational Linguistics, 12, 1087–1103. https://doi.org/10.1162/tacl_a_00690
Levy, N. (2003). Analytic and Continental Philosophy: Explaining the Differences. Metaphilosophy, 34(3), 284–304. https://doi.org/10.1111/1467-9973.00274
Loru, E., Nudo, J., Di Marco, N., Santirocchi, A., Atzeni, R., Cinelli, M., Cestari, V., Rossi-Arnaud, C., & Quattrociocchi, W. (2025). The simulation of judgment in LLMs. Proceedings of the National Academy of Sciences, 122(42). https://doi.org/10.1073/pnas.2518443122
Lumbard, J. E. B. (2024). Islam and the Challenge of Epistemic Sovereignty. Religions, 15(4), 406. https://doi.org/10.3390/rel15040406
Mahfud, C., Astari, R., Kasdi, A., Mu’ammar, M. A., Muyasaroh, M., & Wajdi, F. (2021). Islamic cultural and Arabic linguistic influence on the languages of Nusantara; From lexical borrowing to localized Islamic lifestyles. Wacana, 22(1), 224. https://doi.org/10.17510/wacana.v22i1.914
Masnun, M. (2019). Teori Linguistik dan Psikologi dalam Pengajaran Bahasa Arab di Lembaga Pendidikan Islam. Jurnal Pendidikan Islam, 8(1), 172–204. https://doi.org/10.38073/jpi.v8i1.107
Meng, L., Li, Y., Wei, W., & Yang, C. (2025). Resolving Linguistic Asymmetry: Forging Symmetric Multilingual Embeddings Through Asymmetric Contrastive and Curriculum Learning. Symmetry, 17(9), 1386. https://doi.org/10.3390/sym17091386
Merrill, W., Goldberg, Y., Schwartz, R., & Smith, N. A. (2021). Provable Limitations of Acquiring Meaning from Ungrounded Form: What Will Future Language Models Understand? Transactions of the Association for Computational Linguistics, 9, 1047–1060. https://doi.org/10.1162/tacl_a_00412
Mitchell, M., & Krakauer, D. C. (2023). The debate over understanding in AI’s large language models. Proceedings of the National Academy of Sciences, 120(13). https://doi.org/10.1073/pnas.2215907120
Musfardi, M., Nelma, U., Nathania, N., Anof, I. M., Astuti, D., & Widodo, L. S. (2025). Studi Kasus Tinjauan Aspek Ontologi, Epsitemologi dan Aksiologi Efektivitas Terapi Musik Klasik terhadap Penurunan Tingkat Halusinasi Pendengaran pada Pasien Skizoafektif. Jurnal Ners, 10(1), 127–136. https://doi.org/10.31004/jn.v10i1.52443
Nasir, M. R., & Subet, M. F. (2024). Sorotan Literatur Bersistematik: Semantik Inkuisitif sebagai Kaedah Pembelajaran Peribahasa Melayu Sekolah Menengah Systematic Literature Review: Inquisitive Semantics as a Method for Learning Malay Proverbs in Secondary School. 39(1), 171–190.
O’Regan, J. P., & Ferri, G. (2025). Artificial intelligence and depth ontology: implications for intercultural ethics. Applied Linguistics Review, 16(2), 797–807. https://doi.org/10.1515/applirev-2024-0189
Pasquale, F. (2019). Professional Judgment in an Era of Artificial Intelligence and Machine Learning. Boundary 2, 46(1), 73–101. https://doi.org/10.1215/01903659-7271351
Ramadhan, F. H., & Munawaroh, S. (2025). Development National Ethical Framework for AI Use in Indonesia: Perspectives Regulation and Socio-Cultural Values. Journal of Artificial Intelligence Research, 1(2), 70–79. https://doi.org/10.64910/jouair.v1i2.14
Ramoglou, S., & McMullen, J. S. (2024). “What Is an Opportunity?”: From Theoretical Mystification to Everyday Understanding. Academy of Management Review, 49(2), 273–298. https://doi.org/10.5465/amr.2020.0335
Rega, M. L., Telaretti, F., Alvaro, R., & Kangasniemi, M. (2017). Philosophical and theoretical content of the nursing discipline in academic education: A critical interpretive synthesis. Nurse Education Today, 57, 74–81. https://doi.org/10.1016/j.nedt.2017.07.001
Rolin, K. H. (2021). Objectivity, trust and social responsibility. Synthese, 199(1–2), 513–533. https://doi.org/10.1007/s11229-020-02669-1
Rosyad, A. M., & Maarif, M. A. (2020). Paradigma Pendidikan Demokrasi Dan Pendidikan Islam Dalam Menghadapi Tantangan Globalisasi Di Indonesia. Nazhruna: Jurnal Pendidikan Islam, 3(1), 75–99. https://doi.org/10.31538/nzh.v3i1.491
Sari, D. N., & Utomo, A. P. Y. (2020). Directive speech act in President Joko Widodo’s speech related to handling coronavirus (Covid-19) in Indonesia (Pragmatic review). Journal of Social Studies (JSS), 16(1), 35–50. https://doi.org/10.21831/jss.v16i1.32072
Schmitt, F. F. (2000). Veritistic value. Social Epistemology, 14(4), 259–280. https://doi.org/10.1080/713865188
Searle, J. R. (1980). Minds, brains, and programs. Behavioral and Brain Sciences, 3(3), 417–424. https://doi.org/10.1017/S0140525X00005756
Šekrst, K. (2024). Chinese Chat Room: AI Hallucinations, Epistemology and Cognition. Studies in Logic, Grammar and Rhetoric, 69(1), 365–381. https://doi.org/10.2478/slgr-2024-0029
Sun, Y., Sheng, D., Zhou, Z., & Wu, Y. (2024). AI hallucination: towards a comprehensive classification of distorted information in artificial intelligence-generated content. Humanities and Social Sciences Communications, 11(1), 1278. https://doi.org/10.1057/s41599-024-03811-x
Supena, I. (2024). Epistemology of Tafsīr, Ta’wīl, and Hermeneutics: Towards an Integrative Approach. Journal of Islamic Thought and Civilization, 14(1), 121–136. https://doi.org/10.32350/jitc.141.08
Taherkhani, H., Shin, J., Tahir, M. A., Misu, M. R. H., Gattani, V. S., & Hemmati, H. (2026). Toward Automated Validation of Language Model Synthesized Test Cases Using Semantic Entropy. IEEE Transactions on Software Engineering, 52(4), 1426–1445. https://doi.org/10.1109/TSE.2026.3664287
Tahiri, H. (2018). When The Present Misunderstands The Past How A Modern Arab Intellectual Reclaimed His Own Heritage. Arabic Sciences and Philosophy, 28(1), 133–158. https://doi.org/10.1017/S0957423917000108
Titus, L. M. (2024). Does ChatGPT have semantic understanding? A problem with the statistics-of-occurrence strategy. Cognitive Systems Research, 83, 101174. https://doi.org/10.1016/j.cogsys.2023.101174
Tonta, R. F. (2025). Relaks Dengan Fear Of Missing Out (Fomo) Dalam Perspektif Gelassenheit Martin Heidegger. Sanjiwani: Jurnal Filsafat, 16(1), 1–10. https://doi.org/10.25078/sjf.v16i1.2886
Utami, S. P. T., Andayani, A., Winarni, R., & Sumarwati, S. (2023). Utilization of artificial intelligence technology in an academic writing class: How do Indonesian students perceive? Contemporary Educational Technology, 15(4), ep450. https://doi.org/10.30935/cedtech/13419
Wulandari, S., & Muhid. (2022). Pemahaman Terhadap Hadis Dengan Pendekatan Linguistik. Universum, 16(2), 1–23. https://doi.org/10.30762/universum.v16i2.285
Zainal, M. Z. (2017). Lakuan Ilokusi Guru dalam Pengajaran Bahasa Melayu (Teacher’s Illocutionary Act in Teaching Malay Language). GEMA Online® Journal of Language Studies, 17(4), 191–208. https://doi.org/10.17576/gema-2017-1704-13
Zainodin, U. Z., Omar, N., & Saif, A. (2017). Semantic Measure Based on Features in Lexical Knowledge Sources. Asia-Pacific Journal of Information Technology & Multimedia, 06(01), 39–55. https://doi.org/10.17576/apjitm-2017-0601-04
Zawacki-Richter, O., Marín, V. I., Bond, M., & Gouverneur, F. (2019). Systematic review of research on artificial intelligence applications in higher education – where are the educators? International Journal of Educational Technology in Higher Education, 16(1). https://doi.org/10.1186/s41239-019-0171-0
Zönnchen, B., Dzhimova, M., & Socher, G. (2025). From intelligence to autopoiesis: rethinking artificial intelligence through systems theory. Frontiers in Communication, 10. https://doi.org/10.3389/fcomm.2025.1585321