---
title: "Syayvo: Ukraine to Launch National LLM Based on Gemma 3 by 2027"
description: "🇺🇦 Ukraine is creating its own LLM \"Syayvo\" based on Google Gemma 3. The project is supervised by the Ministry of Digital Transformation, with infrastructure provided by \"Kyivstar\". The model will be trained on data from 90 state institutions and archives. Launch for users — January 2027. #IT #Ukraine #AI"
date: 2026-08-10T14:22:03.000Z
lang: en
url: https://xab.info/en/posts/ukraine-creates-national-llm-syayvo-based-on-gemma-3
tags: [ukraine, artificial-intelligence, llm, ministry-of-digital-development, kyivstar, gemma-3, tech-sovereignty]
publisher: "XAB.info"
---

# Syayvo: Ukraine to Launch National LLM Based on Gemma 3 by 2027

![Olga Boyko, founder of Swayvo, speaks at a conference about creating a national LLM based on Gemma 3 by 2027](https://xab.info/media/2026/08/10/ukraina-sozdaet-natsionalnuyu-llm-syayvo-na-baze-gemma-3/ukraina-sozdaet-natsionalnuyu-llm-syayvo-na-baze-gemma-3-1.webp)

## 🎯 Key Points

- Ukraine is developing a national LLM "Syayvo" based on Google Gemma 3.
- The project is supervised by the Ministry of Digital Transformation, with infrastructure and financing provided by "Kyivstar".
- The model is trained on data from 90 state institutions and archives, including 10 TB of historical documents.
- The test version launch is scheduled for the end of 2026, with a public release in January 2027.
- After transfer to the state, the model and datasets will be open source.

Amid the global race for artificial intelligence, Ukraine is taking a strategic step towards digital sovereignty. Head of Ukraine's Ministry of Digital Transformation, Oksana Ferchuk, confirmed the creation of a national large language model (LLM) named "Syayvo". The project, which will become a key element of the country's digital infrastructure, is based on Google's cutting-edge Gemma 3 architecture and is being adapted to the unique linguistic and cultural features of the Ukrainian language.

### Technological Foundation and Infrastructure

The development of "Syayvo" is a large-scale public-private partnership. The telecommunications giant "Kyivstar" is taking on the technical infrastructure, financing, and organizational support. The choice of the Gemma 3 base model from Google is due to its high efficiency and openness, allowing engineers to focus on fine-tuning for specific tasks of the Ukrainian market. The key task at the current stage is the adaptation of the tokenizer—the component that breaks text into semantic units—to ensure the fastest and most accurate processing of Ukrainian vocabulary and grammar.

### Creation of a National Data Corpus

The main challenge for developers remains the formation of a high-quality national data corpus. To train the model, the team is collecting and verifying information from more than 90 state institutions, scientific institutions, universities, and media resources. Special attention is paid to historical accuracy: the State Archive of Ukraine (Ukrhosarkhiv) has already provided more than 10 terabytes of materials, including historical documents, state acts, and scientific works. A significant part of this data requires complex digitization from paper media, which is a laborious but necessary process for preserving historical memory in digital format.

### Expert Control and Security

To guarantee the quality and safety of AI operations, a special expert committee of more than 70 specialists has been formed. They evaluate the model in four critical areas: language accuracy and depth of context understanding, security (resilience to hallucinations, disinformation, and bias), absence of discriminatory responses, and comparative analysis with global counterparts. In June 2025, a small-scale prototype of "Syayvo" successfully passed the pre-training and supervised fine-tuning stages, confirming the viability of the chosen architecture.

### Open Source and Integration Plans

After testing is completed and the project is handed over to the state, the base version of "Syayvo" and the Ukrainian datasets created during development will be published in the open source domain. This will allow developers, businesses, and scientists to obtain a technical pipeline for training the model for their own products. In the public sector, "Syayvo" will become the basis for future services, such as "Diia AI", an educational assistant in the "Mriya" app, and tools for legislative analysis. The test version of the model is scheduled to be completed by the end of 2026, and it will be opened to a wide range of users in January 2027.

## ❓ FAQ

### Q: When will the "Syayvo" model be launched?
**A:** The test version is planned for the end of 2026, and the public launch for users and developers — in January 2027.

### Q: What is "Syayvo" based on?
**A:** The model is built on the Gemma 3 base from Google, adapted for the Ukrainian language and context.

### Q: Who is financing the project?
**A:** The company "Kyivstar" provides the technical infrastructure, financing, and organizational processes.

### Q: Will the model be open source?
**A:** Yes, after the project is transferred to the state, the base version and datasets will be published in the open source domain.