THE EPISTEMOLOGY OF ARTIFICIAL INTELLIGENCE AND THE CRISIS OF LINGUISTIC VALIDITY IN ARABIC LANGUAGE STUDIES

Authors

  • Muhammad Zahid ‘Afifarrasyihab Rahimadinullah Author
  • Imam Asrori Author

Keywords:

Epistemology of AI, Linguistic Validity, Large Language Models, Arabic Language, Epistemological Hallucination

Abstract

Large Language Models (LLMs) have achieved demonstrable syntactic competence, yet this article argues that their outputs remain epistemologically deficient across three critical dimensions: semantic adequacy, pragmatic legitimacy, and epistemic accountability. Drawing on Wittgenstein, Chomsky, Searle, and Goldman, this study introduces a Four-Dimensional Linguistic Validity Model and the concept of epistemological hallucination to analyze AI outputs in Arabic language studies. The findings establish that AI’s formal-linguistic strength coexists with structural limitations irreducible by scale, particularly given the civilizational depth of Arabic as an epistemic language.

 

References

Abudalfa, S., Saad, M., & El-Beltagy, S. (2025). Editorial: Emerging techniques in Arabic natural language processing. Frontiers in Artificial Intelligence, 8. https://doi.org/10.3389/frai.2025.1715520

Ainirrohmah, N., Arridho, A. R., Makrom, C., Shofiyah, & Prasetya, B. (2026). Konsep Pendidikan Filsafat Menurut Al-Farabi dan Ibnu Sina. Jurnal Pendidikan Agama Islam, 2(2). https://doi.org/10.59829/1s5jkh80

Al-Dadi, M. bin A. (2025). The Linguistic Meaning Between Traditional Rhetorical Frameworks and Artificial Intelligence Models: A Post humanist Critique of Computational Semantics. Journal of Posthumanism, 5(7). https://doi.org/10.63332/joph.v5i7.2888

Al-Rawafi, A., Sudana, D., Lukmana, I., & Syihabuddin, S. (2021). Students’ apologizing in Arabic and English: An interlanguage pragmatic case study at an Islamic boarding school in Indonesia. Indonesian Journal of Applied Linguistics, 10(3). https://doi.org/10.17509/ijal.v10i3.31740

Alharbi, W. (2023). AI in the Foreign Language Classroom: A Pedagogical Overview of Automated Writing Assistance Tools. Education Research International, 2023, 1–15. https://doi.org/10.1155/2023/4253331

Andresen, J. T. (1990). Skinner and Chomsky thirty years later. Historiographia Linguistica, 17(1–2), 145–165. https://doi.org/10.1075/hl.17.1-2.12and

Arum, K. (2018). Pengembangan Pendidikan Agama Islam Berbasis Sosial Profetik (Analisis Terhadap Pemikiran Kuntowijoyo). Millah: Journal of Religious Studies, 177–196. https://doi.org/10.20885/millah.vol17.iss2.art2

Beuls, K., & Van Eecke, P. (2024). Humans Learn Language from Situated Communicative Interactions. What about Machines? Computational Linguistics, 50(4), 1277–1311. https://doi.org/10.1162/coli_a_00534

Bhojani, A.-R., & Schwarting, M. (2023). Truth and Regret: Large Language Models, the Quran, and Misinformation. Theology and Science, 21(4), 557–563. https://doi.org/10.1080/14746700.2023.2255944

Bishop, J. M. (2021). Artificial Intelligence Is Stupid and Causal Reasoning Will Not Fix It. Frontiers in Psychology, 11. https://doi.org/10.3389/fpsyg.2020.513474

Buhori, B., & Wahidah, B. (2017). Bahasa Arab dan Peradaban Islam: Telaah atas Sejarah Perkembangan Bahasa Arab dalam Lintas Sejarah Peradaban Islam. Al-Hikmah, 11(1). https://doi.org/10.24260/al-hikmah.v11i1.822

Campanini, M. (2020). The Evidence of Meaning ( bayān al-maʿnā ) in the Ẓāhirī Approach to the Qur’an. Journal of Qur’anic Studies, 22(1), 172–191. https://doi.org/10.3366/jqs.2020.0415

Cholily, N., Ghozali, M., Kholid, A., Wan Mokhtar, W. K. A., & Mahmut, R. İ. (2025). Bridging Fiqh and Religious Practice: Actualizing the Function of Ḥāshiyah as a Form of Worship in the Scribal Traditions of Madurese Pesantren Literature. Journal of Islamic Law, 6(1), 21–45. https://doi.org/10.24260/jil.v6i1.3749

Chomsky, N. (1959). Verbal behavior. By B. F. Skinner. (The Century Psychology Series.) Pp. viii, 478. New York: Appleton-Century-Crofts, Inc., 1957. Language, 35(1), 26–58. https://doi.org/10.2307/411334

Christof, M., & Armoundas, A. A. (2025). Implications of integrating large language models into clinical decision making. Communications Medicine, 5(1), 490. https://doi.org/10.1038/s43856-025-01216-8

Coeckelbergh, M. (2018). Technology Games: Using Wittgenstein for Understanding and Evaluating Technology. Science and Engineering Ethics, 24(5), 1503–1519. https://doi.org/10.1007/s11948-017-9953-8

Collins, H. (2021). The science of artificial intelligence and its critics. Interdisciplinary Science Reviews, 46(1–2), 53–70. https://doi.org/10.1080/03080188.2020.1840821

Contreras Kallens, P., Kristensen‐McLachlan, R. D., & Christiansen, M. H. (2023). Large Language Models Demonstrate the Potential of Statistical Learning in Language. Cognitive Science, 47(3). https://doi.org/10.1111/cogs.13256

Dentella, V., Günther, F., & Leivada, E. (2023). Systematic testing of three Language Models reveals low language accuracy, absence of response stability, and a yes-response bias. Proceedings of the National Academy of Sciences, 120(51). https://doi.org/10.1073/pnas.2309583120

Dentella, V., Günther, F., & Leivada, E. (2025). Language in vivo vs. in silico: Size matters but Larger Language Models still do not comprehend language on a par with humans due to impenetrable semantic reference. PLOS One, 20(7), e0327794. https://doi.org/10.1371/journal.pone.0327794

Dentella, V., Günther, F., Murphy, E., Marcus, G., & Leivada, E. (2024). Testing AI on language comprehension tasks reveals insensitivity to underlying meaning. Scientific Reports, 14(1), 1–11. https://doi.org/10.1038/s41598-024-79531-8

Dokic, K., Pisker, B., & Radisic, B. (2025). Mirroring Cultural Dominance: Disclosing Large Language Models Social Values, Attitudes and Stereotypes. Societies, 15(5), 142. https://doi.org/10.3390/soc15050142

Erdocia, I., Migge, B., & Schneider, B. (2024). Language is not a data set—Why overcoming ideologies of dataism is more important than ever in the age of AI. Journal of Sociolinguistics, 28(5), 20–25. https://doi.org/10.1111/josl.12680

Faisal, M. (2024). Dampak Kecerdasan Buatan (AI) terhadap Pola Pikir Cerdas Mahasiswa di Pontianak. NUCLEUS, 5(1), 60–66. https://doi.org/10.37010/nuc.v5i1.1684

Farquhar, S., Kossen, J., Kuhn, L., & Gal, Y. (2024). Detecting hallucinations in large language models using semantic entropy. Nature, 630(8017), 625–630. https://doi.org/10.1038/s41586-024-07421-0

Ferrario, A., Facchini, A., & Termine, A. (2024). Experts or Authorities? The Strange Case of the Presumed Epistemic Superiority of Artificial Intelligence Systems. Minds and Machines, 34(3), 1–27. https://doi.org/10.1007/s11023-024-09681-1

Fromkin, V. (1968). Speculations on performance models. Journal of Linguistics, 4(1), 47–68. https://doi.org/10.1017/S002222670000164X

Hajrah, Tang, R., Tahmir, S., & Daeng, K. (2019). Reconceptualization of local wisdom through kelong makassar: A semiotic review of michael riffaterre. Journal of Language Teaching and Research, 10(6), 1209–1216. https://doi.org/10.17507/jltr.1006.08

Heersmink, R., de Rooij, B., Clavel Vázquez, M. J., & Colombo, M. (2024). A phenomenology and epistemology of large language models: transparency, trust, and trustworthiness. Ethics and Information Technology, 26(3), 1–15. https://doi.org/10.1007/s10676-024-09777-3

Huttunen, R., & Kakkori, L. (2022). Heidegger’s critique of the technology and the educational ecological imperative. Educational Philosophy and Theory, 54(5), 630–642. https://doi.org/10.1080/00131857.2021.1903436

Irsyady, K. A., Qudsy, S. Z., Wijayati, M., & Muasssomah, M. (2023). The Authorship Of Shaykh Nawawi Al-Bantani In Arabic Linguistics Studies. Journal Of Indonesian Islam, 17(2), 259. https://doi.org/10.15642/JIIS.2023.17.2.259-282

Jidan, F. (2022). Perkembangan Ilmu Balaghah. Imtiyaz: Jurnal Ilmu Keislaman, 6(2), 142–150. https://doi.org/10.46773/imtiyaz.v6i2.355

Kreps, S., McCain, R. M., & Brundage, M. (2022). All the News That’s Fit to Fabricate: AI-Generated Text as a Tool of Media Misinformation. Journal of Experimental Political Science, 9(1), 104–117. https://doi.org/10.1017/XPS.2020.37

Kurnaz, S. (2017). Who is the Lawgiver? The Hermeneutical Grounds of the Methods of Interpreting Qur’an and Sunna (istinbāṭ al-aḥkām). Oxford Journal of Law and Religion, 6(2), 347–371. https://doi.org/10.1093/ojlr/rwx007

Lan, N., Chemla, E., & Katzir, R. (2024). Large Language Models and the Argument from the Poverty of the Stimulus. Linguistic Inquiry, 1–28. https://doi.org/10.1162/ling_a_00533

Lbs, M. (2020). Konsep Pendidikan Menurut Pemikiran Kh. Hasyim Asy’ari. Jurnal As-Salam, 4(1), 79–94. https://doi.org/10.37249/as-salam.v4i1.170

Lederman, H., & Mahowald, K. (2024). Are Language Models More Like Libraries or Like Librarians? Bibliotechnism, the Novel Reference Problem, and the Attitudes of LLMs. Transactions of the Association for Computational Linguistics, 12, 1087–1103. https://doi.org/10.1162/tacl_a_00690

Levy, N. (2003). Analytic and Continental Philosophy: Explaining the Differences. Metaphilosophy, 34(3), 284–304. https://doi.org/10.1111/1467-9973.00274

Loru, E., Nudo, J., Di Marco, N., Santirocchi, A., Atzeni, R., Cinelli, M., Cestari, V., Rossi-Arnaud, C., & Quattrociocchi, W. (2025). The simulation of judgment in LLMs. Proceedings of the National Academy of Sciences, 122(42). https://doi.org/10.1073/pnas.2518443122

Lumbard, J. E. B. (2024). Islam and the Challenge of Epistemic Sovereignty. Religions, 15(4), 406. https://doi.org/10.3390/rel15040406

Mahfud, C., Astari, R., Kasdi, A., Mu’ammar, M. A., Muyasaroh, M., & Wajdi, F. (2021). Islamic cultural and Arabic linguistic influence on the languages of Nusantara; From lexical borrowing to localized Islamic lifestyles. Wacana, 22(1), 224. https://doi.org/10.17510/wacana.v22i1.914

Masnun, M. (2019). Teori Linguistik dan Psikologi dalam Pengajaran Bahasa Arab di Lembaga Pendidikan Islam. Jurnal Pendidikan Islam, 8(1), 172–204. https://doi.org/10.38073/jpi.v8i1.107

Meng, L., Li, Y., Wei, W., & Yang, C. (2025). Resolving Linguistic Asymmetry: Forging Symmetric Multilingual Embeddings Through Asymmetric Contrastive and Curriculum Learning. Symmetry, 17(9), 1386. https://doi.org/10.3390/sym17091386

Merrill, W., Goldberg, Y., Schwartz, R., & Smith, N. A. (2021). Provable Limitations of Acquiring Meaning from Ungrounded Form: What Will Future Language Models Understand? Transactions of the Association for Computational Linguistics, 9, 1047–1060. https://doi.org/10.1162/tacl_a_00412

Mitchell, M., & Krakauer, D. C. (2023). The debate over understanding in AI’s large language models. Proceedings of the National Academy of Sciences, 120(13). https://doi.org/10.1073/pnas.2215907120

Musfardi, M., Nelma, U., Nathania, N., Anof, I. M., Astuti, D., & Widodo, L. S. (2025). Studi Kasus Tinjauan Aspek Ontologi, Epsitemologi dan Aksiologi Efektivitas Terapi Musik Klasik terhadap Penurunan Tingkat Halusinasi Pendengaran pada Pasien Skizoafektif. Jurnal Ners, 10(1), 127–136. https://doi.org/10.31004/jn.v10i1.52443

Nasir, M. R., & Subet, M. F. (2024). Sorotan Literatur Bersistematik: Semantik Inkuisitif sebagai Kaedah Pembelajaran Peribahasa Melayu Sekolah Menengah Systematic Literature Review: Inquisitive Semantics as a Method for Learning Malay Proverbs in Secondary School. 39(1), 171–190.

O’Regan, J. P., & Ferri, G. (2025). Artificial intelligence and depth ontology: implications for intercultural ethics. Applied Linguistics Review, 16(2), 797–807. https://doi.org/10.1515/applirev-2024-0189

Pasquale, F. (2019). Professional Judgment in an Era of Artificial Intelligence and Machine Learning. Boundary 2, 46(1), 73–101. https://doi.org/10.1215/01903659-7271351

Ramadhan, F. H., & Munawaroh, S. (2025). Development National Ethical Framework for AI Use in Indonesia: Perspectives Regulation and Socio-Cultural Values. Journal of Artificial Intelligence Research, 1(2), 70–79. https://doi.org/10.64910/jouair.v1i2.14

Ramoglou, S., & McMullen, J. S. (2024). “What Is an Opportunity?”: From Theoretical Mystification to Everyday Understanding. Academy of Management Review, 49(2), 273–298. https://doi.org/10.5465/amr.2020.0335

Rega, M. L., Telaretti, F., Alvaro, R., & Kangasniemi, M. (2017). Philosophical and theoretical content of the nursing discipline in academic education: A critical interpretive synthesis. Nurse Education Today, 57, 74–81. https://doi.org/10.1016/j.nedt.2017.07.001

Rolin, K. H. (2021). Objectivity, trust and social responsibility. Synthese, 199(1–2), 513–533. https://doi.org/10.1007/s11229-020-02669-1

Rosyad, A. M., & Maarif, M. A. (2020). Paradigma Pendidikan Demokrasi Dan Pendidikan Islam Dalam Menghadapi Tantangan Globalisasi Di Indonesia. Nazhruna: Jurnal Pendidikan Islam, 3(1), 75–99. https://doi.org/10.31538/nzh.v3i1.491

Sari, D. N., & Utomo, A. P. Y. (2020). Directive speech act in President Joko Widodo’s speech related to handling coronavirus (Covid-19) in Indonesia (Pragmatic review). Journal of Social Studies (JSS), 16(1), 35–50. https://doi.org/10.21831/jss.v16i1.32072

Schmitt, F. F. (2000). Veritistic value. Social Epistemology, 14(4), 259–280. https://doi.org/10.1080/713865188

Searle, J. R. (1980). Minds, brains, and programs. Behavioral and Brain Sciences, 3(3), 417–424. https://doi.org/10.1017/S0140525X00005756

Šekrst, K. (2024). Chinese Chat Room: AI Hallucinations, Epistemology and Cognition. Studies in Logic, Grammar and Rhetoric, 69(1), 365–381. https://doi.org/10.2478/slgr-2024-0029

Sun, Y., Sheng, D., Zhou, Z., & Wu, Y. (2024). AI hallucination: towards a comprehensive classification of distorted information in artificial intelligence-generated content. Humanities and Social Sciences Communications, 11(1), 1278. https://doi.org/10.1057/s41599-024-03811-x

Supena, I. (2024). Epistemology of Tafsīr, Ta’wīl, and Hermeneutics: Towards an Integrative Approach. Journal of Islamic Thought and Civilization, 14(1), 121–136. https://doi.org/10.32350/jitc.141.08

Taherkhani, H., Shin, J., Tahir, M. A., Misu, M. R. H., Gattani, V. S., & Hemmati, H. (2026). Toward Automated Validation of Language Model Synthesized Test Cases Using Semantic Entropy. IEEE Transactions on Software Engineering, 52(4), 1426–1445. https://doi.org/10.1109/TSE.2026.3664287

Tahiri, H. (2018). When The Present Misunderstands The Past How A Modern Arab Intellectual Reclaimed His Own Heritage. Arabic Sciences and Philosophy, 28(1), 133–158. https://doi.org/10.1017/S0957423917000108

Titus, L. M. (2024). Does ChatGPT have semantic understanding? A problem with the statistics-of-occurrence strategy. Cognitive Systems Research, 83, 101174. https://doi.org/10.1016/j.cogsys.2023.101174

Tonta, R. F. (2025). Relaks Dengan Fear Of Missing Out (Fomo) Dalam Perspektif Gelassenheit Martin Heidegger. Sanjiwani: Jurnal Filsafat, 16(1), 1–10. https://doi.org/10.25078/sjf.v16i1.2886

Utami, S. P. T., Andayani, A., Winarni, R., & Sumarwati, S. (2023). Utilization of artificial intelligence technology in an academic writing class: How do Indonesian students perceive? Contemporary Educational Technology, 15(4), ep450. https://doi.org/10.30935/cedtech/13419

Wulandari, S., & Muhid. (2022). Pemahaman Terhadap Hadis Dengan Pendekatan Linguistik. Universum, 16(2), 1–23. https://doi.org/10.30762/universum.v16i2.285

Zainal, M. Z. (2017). Lakuan Ilokusi Guru dalam Pengajaran Bahasa Melayu (Teacher’s Illocutionary Act in Teaching Malay Language). GEMA Online® Journal of Language Studies, 17(4), 191–208. https://doi.org/10.17576/gema-2017-1704-13

Zainodin, U. Z., Omar, N., & Saif, A. (2017). Semantic Measure Based on Features in Lexical Knowledge Sources. Asia-Pacific Journal of Information Technology & Multimedia, 06(01), 39–55. https://doi.org/10.17576/apjitm-2017-0601-04

Zawacki-Richter, O., Marín, V. I., Bond, M., & Gouverneur, F. (2019). Systematic review of research on artificial intelligence applications in higher education – where are the educators? International Journal of Educational Technology in Higher Education, 16(1). https://doi.org/10.1186/s41239-019-0171-0

Zönnchen, B., Dzhimova, M., & Socher, G. (2025). From intelligence to autopoiesis: rethinking artificial intelligence through systems theory. Frontiers in Communication, 10. https://doi.org/10.3389/fcomm.2025.1585321

Downloads

Published

2026-08-03