30 repos
Fine-tuned and adapted GPT-style transformer models for text generation across low-resource and non-English languages, including Kalmyk, Tuvan, Bashkir, Turkmen, Mongolian, Tatar, and other language variants. This cluster contains implementations built on PyTorch exploring how large language models can be scaled and adapted for linguistic diversity, with particular focus on minority and endangered language communities.