Please use this identifier to cite or link to this item: https://er.chdtu.edu.ua/handle/ChSTU/9062
Full metadata record
DC FieldValueLanguage
dc.contributor.authorShvets, Sofiia-
dc.contributor.authorШвець, Софія-
dc.date.accessioned2026-03-24T13:50:16Z-
dc.date.available2026-03-24T13:50:16Z-
dc.date.issued2025-
dc.identifier.issn2306-4412 (print)-
dc.identifier.issn2708-6070 (online)-
dc.identifier.urihttps://er.chdtu.edu.ua/handle/ChSTU/9062-
dc.description.abstractThe intensive growth in the use of real-time language models requires mechanisms for their dynamic adaptation to changes in queries, terminology, and user expectations. The study aimed to investigate approaches to continuous feedback-based retraining of large language models. To achieve this goal, the theoretical and structural-functional modelling of the adaptation architecture, experimental implementation of the language model retraining cycle with processing and classification of different types of feedback, and quantitative evaluation of the results using automatic and user metrics were applied. The results of the study showed the effectiveness of the architecture of continuous online learning, which ensures the relevance and stability of the language model in real time. The study determined that implicit feedback is 4-10 times more common than explicit feedback, but explicit feedback gives a higher increase in the accuracy of answers. The proposed system successfully integrated different types of user signals, providing dynamic generation of training examples and hybrid relearning while maintaining the quality and consistency of the results. The Python software cycle for adaptive retraining of the language model involved processing and filtering user signals to form a high-quality buffer of training pairs. After 500 retraining steps on 52,912 query-response pairs, a significant improvement of the model was observed, which was confirmed by a decrease in the loss function from 3.82 to 3.15 and stability of the fine-tuning process without signs of overtraining. The results of the pre-training showed a moderate improvement in the quality of answers after adaptation: lexical similarity according to the Recall-Oriented Understudy for Gisting Evaluation was 0.102, accuracy according to the Bilingual Evaluation Understudy was 0.006, and subjective user satisfaction increased to 0.24, while maintaining the stability of the model with an average cosine similarity value of 0.396. The approach proposed in this study improves the quality and relevance of real-time responses of language models while maintaining their stability and can be used in productive systems to improve user experience.uk_UA
dc.description.abstractІнтенсивне зростання використання мовних моделей реального часу вимагає механізмів їх динамічної адаптації до змін у запитах, термінології та очікуваннях користувачів. Метою дослідження було вивчення підходів до перенавчання великих мовних моделей на основі безперервного зворотного зв’язку. Для досягнення цієї мети було застосовано теоретичне та структурно-функціональне моделювання архітектури адаптації, експериментальну реалізацію циклу перенавчання мовної моделі з обробкою та класифікацією різних типів зворотного зв’язку, а також кількісну оцінку результатів за допомогою автоматичних та користувацьких метрик. Результати дослідження показали ефективність архітектури безперервного онлайннавчання, яка забезпечує актуальність та стабільність мовної моделі в реальному часі. У дослідженні визначено, що неявний зворотний зв’язок зустрічається в 4-10 разів частіше, ніж явний зворотний зв’язок, але явний зворотний зв’язок дає вищий приріст точності відповідей. Запропонована система успішно інтегрує різні типи користувацьких сигналів, забезпечуючи динамічну генерацію навчальних прикладів та гібридне перенавчання, зберігаючи при цьому якість та узгодженість результатів. Програмний цикл Python для адаптивного перенавчання мовної моделі включав обробку та фільтрацію користувацьких сигналів для формування високоякісного буфера навчальних пар. Після 500 кроків перенавчання на 52 912 парах запит-відповідь спостерігалося значне покращення моделі, що підтверджувалося зменшенням функції втрат з 3,82 до 3,15 та стабільністю процесу точного налаштування без ознак перенавчання. Результати попереднього навчання показали помірне покращення якості відповідей після адаптації: лексична подібність за даними Recall-Oriented Understudy for Gisting Evaluation становила 0,102, точність за даними Bilingual Evaluation Understudy – 0,006, а суб’єктивна задоволеність користувачів зросла до 0,24, зберігаючи при цьому стабільність моделі із середнім значенням косинусної подібності 0,396. Підхід, запропонований у цьому дослідженні, покращує якість та релевантність відповідей мовних моделей у реальному часі, зберігаючи їх стабільність, і може бути використаний у продуктивних системах для покращення користувацького досвіду.uk_UA
dc.language.isoenuk_UA
dc.publisherВісник Черкаського державного технологічного університетуuk_UA
dc.subjectadaptive relearninguk_UA
dc.subjectgenerative transformersuk_UA
dc.subjectdynamic model adaptationuk_UA
dc.subjectPython implementationuk_UA
dc.subjecthybrid learninguk_UA
dc.subjectquality assessment metricsuk_UA
dc.subjectlanguage model stabilityuk_UA
dc.subjectадаптивне перенавчанняuk_UA
dc.subjectгенеративні трансформаториuk_UA
dc.subjectадаптація динамічної моделіuk_UA
dc.subjectреалізація на Pythonuk_UA
dc.subjectгібридне навчанняuk_UA
dc.subjectметрики оцінки якостіuk_UA
dc.subjectстабільність мовної моделіuk_UA
dc.titleContinuous feedback loops: Online fine-tuning of LLMs with user signalsuk_UA
dc.title.alternativeБезперервні цикли зворотного зв’язку: онлайн-налаштування LLM за допомогою сигналів користувачаuk_UA
dc.typeArticleuk_UA
dc.citation.volume30uk_UA
dc.citation.issue3uk_UA
dc.citation.spage106uk_UA
dc.citation.epage120uk_UA
dc.identifier.doihttps://doi.org/10.62660/bcstu/3.2025.106-
Appears in Collections:том 30, №3/2025

Files in This Item:
File Description SizeFormat 
зміст.pdf161.04 kBAdobe PDFThumbnail
View/Open
титул.pdf234.55 kBAdobe PDFThumbnail
View/Open
11.pdf3.65 MBAdobe PDFThumbnail
View/Open


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.