Please use this identifier to cite or link to this item:
https://scidar.kg.ac.rs/handle/123456789/23296Full metadata record
| DC Field | Value | Language |
|---|---|---|
| dc.contributor.author | Svičević, Marina | - |
| dc.contributor.author | Pavković, Miloš | - |
| dc.contributor.author | Vučićević, Nemanja | - |
| dc.contributor.author | Milenković, Aleksandar | - |
| dc.contributor.author | Milutinović, Aleksandar | - |
| dc.contributor.editor | Marković, Goran | - |
| dc.date.accessioned | 2026-09-29T10:28:43Z | - |
| dc.date.available | 2026-09-29T10:28:43Z | - |
| dc.date.issued | 2026 | - |
| dc.identifier.isbn | 978-86-82434-15-3 | en_US |
| dc.identifier.uri | https://scidar.kg.ac.rs/handle/123456789/23296 | - |
| dc.description.abstract | This paper examines the effect of QLoRA fine-tuning of locally executable large language models on mathematics competition tasks for high school students in Serbia. The main focus is on comparing the general-purpose model Mistral-7BInstruct and the mathematically specialized model Mathstral-7B, in order to determine whether prior specialization for mathematical reasoning affects the success of fine-tuning on a small, domain-specific corpus of tasks in Serbian. For the purposes of the study, an automatic evaluation framework was applied, in which generated solutions were assessed according to several criteria, including final answer accuracy, logical coherence of the solution, and explanation quality. The results indicate different effects of fine-tuning for the observed models: the general-purpose model showed a slight performance decline, while the mathematically specialized model showed an improvement. This outcome suggests that models previously adapted to mathematical reasoning may benefit more from domain-specific fine-tuning, while also confirming that Serbian mathematics competition tasks remain a highly demanding test for locally executable large language models. | en_US |
| dc.language.iso | en | en_US |
| dc.subject | Large language models | en_US |
| dc.subject | QLoRA fine-tuning | en_US |
| dc.subject | Mistral-7B-Instruct | en_US |
| dc.subject | Mathstral-7B | en_US |
| dc.subject | Mathematics competitions | en_US |
| dc.subject | Automatic evaluation | en_US |
| dc.subject | Mathematical reasoning | en_US |
| dc.title | Comparing QLoRA fine-tuning of a general-purpose and a mathematically specialized LLM on Serbian mathematics competition tasks | en_US |
| dc.type | conferenceObject | en_US |
| dc.description.version | Published | en_US |
| dc.identifier.doi | https://doi.org/10.46793/ET26.D12S | en_US |
| dc.type.version | PublishedVersion | en_US |
| dc.source.conference | XII Triennial International Scientific Conference Engineering TODAY | en_US |
| Appears in Collections: | Faculty of Science, Kragujevac | |
Files in This Item:
| File | Size | Format | |
|---|---|---|---|
| Comparing QLoRA fine-tuning of a general-purpose and a mathematically.pdf | 1.18 MB | Adobe PDF | View/Open |
Items in SCIDAR are protected by copyright, with all rights reserved, unless otherwise indicated.
