Λεπτομέρειες βιβλιογραφικής εγγραφής
| Τίτλος: |
THE APPROACH DEVELOPMENT OF DATA EXTRACTION FROM LAMBDA TERMS. |
| Alternate Title: |
РОЗРОБКА ПІДХОДУ ДО ВИЛУЧЕННЯ ДАНИХ ЛЯМБДА-ТЕРМІВ. (Ukrainian) |
| Συγγραφείς: |
Deineha, Oleksandar, Donets, Volodymyr, Zholtkevych, Grygoriy |
| Πηγή: |
Eastern-European Journal of Enterprise Technologies; 2024, Vol. 129 Issue 2, p42-54, 13p |
| Θεματικοί όροι: |
Compilers (Computer programs), Language models, Data extraction, Hierarchical clustering (Cluster analysis), Machine learning |
| Περίληψη: |
The study's object is the process of extracting the characteristics of lambda terms, which indicate the optimality of the reduction strategy and increase the productivity of compilers and interpreters. The solution to the problem of extracting specific strategy priority data from lambda terms using Machine Learning methods was considered. Such data was extracted using the large language model Microsoft CodeBERT, which was trained to solve the problem of summarizing the software code. The resulting matrices of embeddings were used to obtain vectors of average embeddings of size 768 and a latent space of size 8 thousand. Further, vectors of average embeddings were used for cluster analysis using the DBSCAN and Hierarchical Agglomerative clustering methods. The most informative variables affecting clustering were determined. Next, the clustering results were compared with the priorities of reduction strategies, which showed the impossibility of separating terms with RI priority. A feature of the obtained results is using machine learning methods to obtain knowledge. The clustering results showed many of the same informative variables, which is explained by the similar shape of the obtained clusters. The results of comparing the clustering values with the real priority are explained by the impossibility of clearly determining the priority and the use of the Microsoft CodeBERT model, which was not trained for the analysis of lambda terms. The proposed approach can find application in the development of compilers and interpreters of functional programming languages, allowing to analyze the code and extract important data to optimize the execution of programs. The obtained data can be used to develop rules aimed at improving the efficiency of compilation and interpretation. [ABSTRACT FROM AUTHOR] |
|
Copyright of Eastern-European Journal of Enterprise Technologies is the property of PC TECHNOLOGY CENTER and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) |
| Βάση Δεδομένων: |
Complementary Index |