روش‌های یادگیری خودکار هستی‌نگاشت‌ها در حوزۀ مفاهیم قرآنی: مطالعۀ مروری دامنه‌ای مقاله

نویسنده: میرعرب، علی ؛ طباطبایی امیری، فائزه سادات ؛ دهقانی سانیج، سمیه ؛

پژوهشنامه کتابداری و اطلاع رسانی پاییز و زمستان 1402، سال سیزدهم - شماره 2 رتبه ب (وزارت علوم/ISC (‎21 صفحه - از 29 تا 49 )

کلیدواژه ها: استخراج دانش فناوری معنایی یادگیری خودکار هستی‌نگاشت هستی‌نگاشت قرآن داده‌کاوی Data mining Ontology automatic learning Quran ontologies Semantic technology Knowledge Extraction

fa en

چکیده:

مقدمه: امروزه فناوری‌های معنایی رویکرد جدیدی را در پردازش و بازنمون معارف قرآنی با هدف ارائۀ اطلاعات معنادار ارائه می‌دهند. هستی‌نگاشت‌ها به‌عنوان یکی از فناوری‌های معنایی، ابزاری جهت بیان رسمی مفاهیم و روابط موجود در حوزۀ خاصی بوده که توسعه و کاربرد آن جهت استخراج معارف و علوم قرآنی مورد توجه قرار گرفته است. یادگیری هستی‌نگاشت‌ها و روش‌های آن به‌صورت خودکار جهت استخراج مفاهیم از مباحث مهم در حوزۀ وب معنایی و فناوری‌های آن است. به‌تازگی توسعه و کاربرد یادگیری هستی‌نگاشت‌ها جهت استخراج مفاهیم قرآنی مورد توجه قرار گرفته است. ازاین‌رو، هدف پژوهش حاضر، بررسی جامع یادگیری خودکار هستی‌نگاشت‌ها در حوزۀ استخراج مفاهیم قرآنی به‌منظور شفاف‌سازی وضعیت فعلی و آینده است. معیارهای مورد بررسی مجموعه داده‌ها، روش‌های یادگیری، روش‌های ارزیابی، نتایج و پیشنهاد‌های آتی پژوهش‌ها در حوزۀ یادگیری خودکار هستی‌نگاشت‌های قرآنی بود. روش‌شناسی: روش بررسی پژوهش حاضر، مرور دامنه‌ای بر اساس دستورالعمل‌های پریزما و بر اساس رویۀ استفاده‌شده توسط آرکسی و امالی (2005) است. این فرآیند پروتکلی را به‌منظور تطبیق نتایج پژوهش موجود با سؤالات و معیارهای تحقیق توصیف می‌کند. پنج مرحلۀ پیشنهادی آرکسی و امالی عبارت‌اند از: 1. شناسایی و طراحی سؤال(ها) پژوهش، 2. انجام استراتژی‌های جستجو برای استخراج مطالعات مرتبط از طریق انتخاب واژه‌های کلیدی مناسب و عملگرهای بولی، 3. انتخاب نهایی پژوهش‌های مرتبط با تعیین معیارهای ورود و خروج، 4. خلاصه‌سازی و گزارش یافته‌ها و درنهایت، 5. گزارش و بحث پیرامون نتایج حاصل. جستجوی منابع در هفت پایگاه دادۀ علمی مشتمل برEmerald, Science Direct, IEEE Xplore Digital Library, Google Scholar, Web of Science, Scopus انجام شد. فرایند جستجو در فروردین 1402 صورت گرفت. تعداد 811 مقاله، بدون توجه به محدودۀ زمانی، مورد ارزیابی و انتخاب قرار گرفت. به‌منظور سازماندهی مقالات بازیابی‌شده، از نرم‌افزار مدیریت منابع اطلاعاتی اندنوت استفاده شد و پس از تطبیق عناوین در پایگاه‌های اطلاعاتی مختلف، تعداد 317 مقاله تکراری حذف گردید. پس از بررسی چکیده‌ها، معیارهای ورود و خروج و کیفیت مقالات اعمال گردید. همچنین به‌منظور جلوگیری از سوگیری در انتخاب مقالات، طی بررسی تصادفی مجددی، توسط دو پژوهشگر مستقل در حوزۀ یادگیری خودکار هستی‌نگاشت نیز ارزیابی صورت گرفت و درنهایت تعداد 25 اثر به‌عنوان ملاک مرور انتخاب گردید. یافته‌ها: یافته‌ها نشان داد اغلب پژوهش‌ها در حوزۀ مجموعۀ داده‌های قرآنی به زبان‌های انگلیسی و عربی بودند و بخش عمده آن‌ها نیز از ترجمۀ انگلیسی قرآن الهلالی و خان استفاده کرده‌اند. استفاده از مجموعه داده‌های بسیار محدود، مهم‌ترین محدودیت پژوهش‌های انجام شده بود. بخش عمدۀ پژوهش‌ها از روش‌های نرمال‌سازی، خوشه‌بندی و دسته‌بندی متن، خلاصه‌سازی متن، استخراج اطلاعات، تشابه و یافتن موجودیت‌های نامدار استفاده کرده‌اند. البته در برخی پژوهش‌ها، روش‌های هوش مصنوعی نظیر شبکۀ عصبی نیز به کار گرفته شده است. علاوه بر این، یافته‌ها نشان داد که الگوریتم‌های داده‌کاوی مبتنی بر روش‌های آمار و احتمال برای یادگیری و ساخت هستی‌نگاشت‌های خودکار در میان محققان با محبوبیت روبرو شده است. همچنین از روش‌های محاسبۀ دقت، فراخوانی و معیار F برای ارزیابی نتایج کاربرد الگوریتم‌های یادگیری خودکار در هستی‌نگاشت‌های قرآنی استفاده کرده‌اند. پژوهش‌هایی که از روش‌های هوش مصنوعی بهره‌برداری کرده‌اند، با تحلیل معنایی، استنتاج، مدل‌سازی و تأیید اعتبار داده‌های استنتاج‌شده به نتایجی مانند تشخیص صوت برای آموزش قرائت قرآن، تشخیص آرایه‌های ادبی و ایجاد ارتباط‌های موضوعی در مفاهیم قرآنی و همچنین ایجاد ارتباط بین این مفاهیم با مفاهیم سایر ادیان نائل شده‌اند. ارزیابی‌ روش‌های ارائه‌شده برای یادگیری خودکار هستی‌نگاشت‌های قرآنی نشان می‌دهد استفاده توأمان از روش‌های داده‌کاوی و هوش مصنوعی نتایج بهتری را به‌همراه دارد. بخش عمدۀ نتایج این حوزه در دو دسته کلی قرار دارد. دستۀ اول مبتنی بر به‌کارگیری روش‌های داده‌کاوی، متن‌کاوی و یادگیری ماشین جهت استخراج خودکار مفاهیم و ابعاد سه‌گانه (فعل، فاعل، مفعول) به‌همراه روابط معنایی از متن قرآن بود. دستۀ دیگر به مقایسه عملکرد روش‌ها و الگوریتم‌های مبتنی بر آمار و مشابهت‌یابی نظیر TF، TF-IDF، AVE-TF، Ridf، TIM، N-gram، FREyA، Pos Taggin، Levenshtein، Log Likelihod، هِرسِت، و جز این‌ها در استخراج مفاهیم خودکار جهت ساخت هستی‌نگاشت قرآنی پرداخته‌اند. یافته‌های حاصل از بررسی کارهای آینده نشان از علاقۀ محققان به الگوریتم‌های هوش مصنوعی و استفاده در یادگیری هستی‌نگاشت و توسعۀ خودکار و نیمه‌خودکار هستی‌نگاشت‌های قرآنی دارد. فقدان مجموعه داده‌های صحیح، دلیل عجز سامانه‌های هوش مصنوعی پیشرفتۀ دنیا مانند جی‌پی‌تی 4 است که در آینده باید به این مهم پرداخته شود. نتیجه‌گیری: نتایج این مطالعه می‌تواند به جهت‌دهی پژوهش‌های آتی درباره بهترین روش‌ها در توسعۀ خودکار هستی‌نگاشت‌های قرآنی کمک کند. این مسئله می‌تواند با طراحی هستی‌نگاشت جامع قرآنی که تمام موضوعات و مفاهیم را با توجه به بافت قرآن، پوشش دهد، مدنظر قرار گرفته و با ایجاد هستی‌نگاشتی جامع از مفاهیم قرآن، کاربران را به‌سمت بازیابی دانش قرآنی رهنمون سازد. همچنین بهره‌برداری بیشتر از روش‌های هوش مصنوعی و پردازش زبان طبیعی نظیر جی.پی.تی. به‌عنوان مدل یادگیری ماشینی برای تولید متن به زبان طبیعی با استفاده از شبکۀ عصبی عمیق، در توسعۀ خودکار هستی‌نگاشت‌های قرآنی ضروری به نظر می‌رسد. با توجه به اینکه یادگیری ماشین مستلزم وجود داده‌های کلان در حوزۀ قرآن است، ساخت مجموعه داده‌های استاندارد ازجمله کارهای آتی محققان است.

Objective: Today, semantic technology offers a new approach in organizing Quranic knowledge with the aim of providing meaningful information and representing Quranic teachings. Ontologies are a tool to formally express concepts and relationships in a specific domain. In the same way, the development of ontology as a tool for representing the effulgence and extracting the knowledge of the Quran is not only valuable, but also necessary. Ontology learning and its methods automatically to extract concepts are important topics in the field of Semantic Web and its technologies. Recently, the development and application of ontologies learning for the extraction of Quranic concepts has been considered. Therefore, the aim of the current research is to comprehensively investigate the ontologies automatic learning in the field of extracting knowledge and Quranic concepts in order to clarify the current and future situation. The investigated criteria were data set, learning methods, evaluation methods, results and future suggestions of studies in the field of ontologies automatic learning of the Quran. Methodology: The research was conducted by the scoping review method in accordance with PRISMA guidelines and based on Arksey & O’Malley procedure. This process describes a protocol for matching the results of existing studies with research questions and criteria. The five steps suggested by Arksey & O’Malley are as follows: 1. Identify and design the research question (s) , 2. Conduct search strategies advocate for relevant studies through the selection of appropriate keywords and Boolean operators, 3. Final selection of relevant studies, considering the inclusion and exclusion criteria, 4. Tabulating the data, and finally, 5. Reporting its results. Sources were searched in seven scientific databases including Emerald, Science Direct, IEEE Xplore Digital Library, Google Scholar, Web of Science, and Scopus. The search process has been done in April 2023. A number of 811 articles, regardless of the time limit, were evaluated and selected. In order to organize the retrieved articles, EndNote resource management software was used and after matching the titles in different databases, 317 duplicate articles were removed. After reviewing the abstracts, the entry and exit criteria and the quality of the articles were applied. Also, in order to avoid bias in the selection of articles, during a random review, two independent researchers in the field of ontology automatic learning were evaluated and finally 25 articles were selected as review criteria. Findings: Most of the study in the field of Quranic data set were in English and Arabic languages, and most of them used the English translation of Al-Hilali and Khan's Quran. The use of a limited data set was the most important limitation of the research conducted in the field of automatic learning of Quranic ontologies. Most of the studies have used normalization methods, text clustering and categorization, text summarization, information extraction, similarity and finding famous entities. Of course, in some studies, artificial intelligence methods such as neural network have also been used. In addition, the findings showed that data mining algorithms based on statistics and probability methods for learning and constructing automatic ontologies was apparently surging in popularity among researchers. Evaluation methods includes calculating accuracy, recall and F criteria in the application of automatic learning algorithms in Quranic ontologies. The studies that have used artificial intelligence techniques, by Semantic analysis, inference, modeling and validation of inferred data have achieved results such as sound recognition for teaching Quran reading, recognition of literary arrays and creating thematic connections in Quranic concepts as well as creating connections between these concepts and concepts in other religions. The evaluation of the presented methods for ontology automatic learning shows that the combined use of data mining methods and artificial intelligence brings better results. Most of the results of this field are in two general categories. The first category was based on the use of data mining, text mining and machine learning methods to automatically extract three concepts and dimensions (subject-predicate-object) along with Semantic relationships from the text of the Quran. The other category compares the performance of methods and algorithms based on statistics and similarity, such as TF, TF-IDF, AVE-TF, Ridf, TIM, N-gram, FREyA, Pos Taggin, Levenshtein, Log Likelihod, Herset, etc. in extracting concepts for the construction of the Quranic ontologies. The findings of the future studies review show the researchers' interest in artificial intelligence algorithms and their use in ontology learning and the automatic and semi-automatic development of Quranic ontologies. The lack of correct data sets is the reason for the inability of the world's advanced artificial intelligence systems such as GPT 4, which must be addressed in the future. Discussion and conclusion: The results of this study can help to direct future research about the best practices in the automatic development of Quranic ontologies. This issue can be taken into consideration by designing a comprehensive Quranic ontology that covers all topics and concepts according to the context of the Quran, and by creating a comprehensive ontology of the Quranic concepts, it will guide users towards the retrieval of Quranic knowledge. Also, more use of artificial intelligence and natural language processing methods, such as GPT as a machine learning model for natural language text generation by deep neural network, it seems essential in the development of automatic learning of Quranic ontologies. Machine learning requires the existence of big data in the field of the Qur'an, hence the creation of standard data sets is one of the future studies.

دریافت فایل ارجاع :
(پژوهیار, , , )

دانلود HTML
دانلود PDF

ورود / عضویت

برای مشاهده محتوای مقاله لازم است وارد پایگاه شوید. در صورتی که عضو نیستید از قسمت عضویت اقدام فرمایید.

ورود

عضویت

تحتاج دخول لعرض محتوى المقالة. إذا لم تكن عضوًا ، فتابع من الجزء الاشتراک.
إن كنت لا تقدر علی شراء الاشتراك عبرPayPal أو بطاقة VISA، الرجاء ارسال رقم هاتفك المحمول إلی مدير الموقع عبر webmaster@noormags.com .

You need Sign in to view the content of the article. If you are not a member, proceed from part Sign up.
If you fail to purchase subscription via PayPal or VISA Card, please send your mobile number to the Website Administrator via webmaster@noormags.com .

1402

1401

1400

1399

1398

1397

1396

1395

1394

1393

1392

1391

1390

روش‌های یادگیری خودکار هستی‌نگاشت‌ها در حوزۀ مفاهیم قرآنی: مطالعۀ مروری دامنه‌ای مقاله