{"id":959,"date":"2020-03-31T22:48:33","date_gmt":"2020-03-31T20:48:33","guid":{"rendered":"https:\/\/www.cjvt.starkmat.si\/template-projekt\/work-packages\/work-package-2\/"},"modified":"2024-01-25T11:37:24","modified_gmt":"2024-01-25T10:37:24","slug":"work-package-2","status":"publish","type":"page","link":"https:\/\/www.cjvt.si\/povejmo\/en\/work-packages\/slollamai\/","title":{"rendered":"Work Package 2: SloLLaMai"},"content":{"rendered":"
Extremely large language models (ChatGPT and GPT-4) have recently shown remarkable progress in certain tasks, but they also face numerous practical challenges in their utilization, such as closedness and lack of transparency, high computational requirements, and the high cost of customization and broader usage, which is unattainable for most research organizations and companies. Further scaling up these models is no longer practical. Smaller, open-access models like LLaMA, Alpaca, GPT4All, and Koala have also emerged, which can be trained or adapted for specific tasks on regular GPU computers and achieve similar or nearly equal performance to the largest models.<\/p>\n
In the SloLLaMAi project, we will develop an open-access, computationally efficient generative language model for Slovenian. This model will be the first of its kind for a morphologically rich language with limited resources, presenting a significant research challenge. The development of this new large general model will serve as fundamental infrastructure for industrial projects and all new products requiring natural language processing. Previously developed models (e.g., SloT5, SloBERTa, CroSloEn BERT) have already enabled the creation of technologies that, just a few years ago, could not be developed for Slovenian with comparable accuracy to larger languages (e.g., machine translation, summarization, question answering). The development of such technologies would not have been possible if models for Slovenian did not exist.<\/p>\n
The results of the first project (RRP1) are necessary for the development of the next generation of large language models. This will make it possible to prepare general language models that can be specialized for specific natural language processing tasks.<\/p>\n<\/div><\/section><\/div>\n