Fine-tuning ruGPT-3 XL 1.3B to a Modern LLM Level and Comparison with Gemma3 1B
The ancient ruGPT-3 XL 1.3B model from 2020 was fine-tuned using SFT and DPO methods, turning it into a modern instruct-following LLM. The result was compared with the 2025 Gemma3 1B model, showing that the fine-tuned ruGPT often outperforms Gemma3 in logic and creativity, though it struggles with negative instructions and role-playing.
Сбербанк
Google/DeepMind
