Sber GigaChat News

This feed includes only recent releases, roughly from the last 12 months, not the full historical archive. That is why a long-established product may have only a few news items.

3 news

Sber GigaChat
GigaChat 3 Ultra is now available to individuals in the API on Freemium

GigaChat 3 Ultra became available to individuals through the public API in a free Freemium mode (the limit was raised to 365 million tokens by 22.07.2026). Cloud access only; the announcement does not mention published weights for this specific build.

Sber GigaChat
GigaChat 3.5 Ultra: hybrid linear attention, open weights

A 432B MoE model (28B active) using a hybrid MLA + GatedDeltaNet architecture (linear attention): long-text generation is 4× faster, while the model is almost twice as compact as GigaChat 3.1 Ultra with comparable code and math quality. Weights are available on HuggingFace (ai-sage/GigaChat3.5-432B-A28B), but the model was not yet available in the official cloud API at developers.sber.ru on the release date.

Hybrid linear architectures for speeding up long-context are already being tested by competitors in the same MoE class. Qwen →
Sber GigaChat
GigaChat-3.1 Ultra and Lightning: open MoE weights under the MIT license

Sber released GigaChat-3.1-Ultra (702B, 36B active) and GigaChat-3.1-Lightning (10B, 1.8B active) with a full DPO stage in FP8; the weights are available on HuggingFace and GitVerse under the MIT license. At the same time, the GigaChat assistant was moved to the flagship model with long-term memory across sessions.

Llama weights are released under Meta's restrictive community license, while GigaChat-3.1 uses a standard MIT license without commercial restrictions. Meta Llama →

Discuss Sber GigaChat News

Enter your email or phone number so we can get back to you.

Send via: