alenisaw/turkicocr-svtrv2-b

Model

TurkicOCR-SVTRv2-B (Official PyTorch Weights)

0

7 commits

2 linked in READMEs

updated Aug 15, 2026

See the code

README

TurkicOCR-SVTRv2-B (Official PyTorch Weights)

TurkicOCR-SVTRv2-B is a lightweight (~35M parameters) line-grounded visual text recognizer tailored for Kazakh and Kyrgyz Cyrillic scripts, trained on the large-scale synthetic TurkicOCR-Cyrillic benchmark.

Model Details

  • Architecture: SVTRv2-B (Single Visual Model Text Recognizer v2, Base configuration).
  • Supervision: Connectionist Temporal Classification (CTC) sequence-to-sequence loss.
  • Parameters: ~35M (~84.4 MB FP32 weights).
  • Input: Bounded document image crops resized to $48 imes 640$ px.
  • Output: UTF-8 encoded Cyrillic text string.
  • Supported Scripts: Kazakh Cyrillic, Kyrgyz Cyrillic, Russian Cyrillic (including mixed/bilingual text).

Performance Benchmark

ModelPrecisionSizeKazakh CERKyrgyz CERRussian CEROverall CERThroughput
TurkicOCR-SVTRv2-BFP3284.4 MB1.71%1.89%1.68%1.76%~142 lines/s
TurkicOCR-SVTRv2-BINT827.1 MB1.88%2.05%1.85%1.93%~385 lines/s
PaddleOCR ServerFP32~115 MB4.82%5.12%3.20%4.38%~95 lines/s
EasyOCRFP32~95 MB6.94%7.31%4.55%6.27%~60 lines/s
Tesseract 5 (rus+kaz)--14.2%16.8%9.4%13.5%~25 lines/s

Quick Usage

git clone https://github.com/alenisaw/turkicocr-svtrv2-b.git
cd turkicocr
pip install -e .
import cv2
from turkicocr.recognition.predictor import TurkicOCRPredictor

predictor = TurkicOCRPredictor(model_path="epoch_9.pth")
image = cv2.imread("document_line.jpg")
result = predictor.predict(image)
print(result["text"])

Citation

@inproceedings{issayev2026turkicocr,
  title={TurkicOCR-SVTRv2-B: Lightweight Line-Grounded Recognizer for Kazakh and Kyrgyz Optical Character Recognition},
  author={Issayev, Alen and Zhalgas, Aidana},
  booktitle={Analysis of Images, Social Networks and Texts (AIST 2026)},
  series={Lecture Notes in Computer Science (LNCS)},
  publisher={Springer},
  year={2026},
  doi={10.1007/978-3-031-XXXXX-X_XX}
}

@misc{issayev_2026_turkicocr_cyrillic,
  author       = {Issayev, Alen},
  title        = {TurkicOCR-Cyrillic},
  year         = {2026},
  publisher    = {Hugging Face},
  doi          = {10.57967/hf/9255},
  url          = {https://huggingface.co/datasets/alenisaw/turkicocr-cyrillic}
}

License

Apache 2.0. Full code and model checkpoints available at https://github.com/alenisaw/turkicocr.

cyrillic
image-to-text
openocr
svtrv2
text-recognition
turkicocr

alenisaw/turkicocr-svtrv2-b

Model

TurkicOCR-SVTRv2-B (Official PyTorch Weights)

0

7 commits

2 linked in READMEs

updated Aug 15, 2026

See the code

README

TurkicOCR-SVTRv2-B (Official PyTorch Weights)

TurkicOCR-SVTRv2-B is a lightweight (~35M parameters) line-grounded visual text recognizer tailored for Kazakh and Kyrgyz Cyrillic scripts, trained on the large-scale synthetic TurkicOCR-Cyrillic benchmark.

Model Details

  • Architecture: SVTRv2-B (Single Visual Model Text Recognizer v2, Base configuration).
  • Supervision: Connectionist Temporal Classification (CTC) sequence-to-sequence loss.
  • Parameters: ~35M (~84.4 MB FP32 weights).
  • Input: Bounded document image crops resized to $48 imes 640$ px.
  • Output: UTF-8 encoded Cyrillic text string.
  • Supported Scripts: Kazakh Cyrillic, Kyrgyz Cyrillic, Russian Cyrillic (including mixed/bilingual text).

Performance Benchmark

ModelPrecisionSizeKazakh CERKyrgyz CERRussian CEROverall CERThroughput
TurkicOCR-SVTRv2-BFP3284.4 MB1.71%1.89%1.68%1.76%~142 lines/s
TurkicOCR-SVTRv2-BINT827.1 MB1.88%2.05%1.85%1.93%~385 lines/s
PaddleOCR ServerFP32~115 MB4.82%5.12%3.20%4.38%~95 lines/s
EasyOCRFP32~95 MB6.94%7.31%4.55%6.27%~60 lines/s
Tesseract 5 (rus+kaz)--14.2%16.8%9.4%13.5%~25 lines/s

Quick Usage

git clone https://github.com/alenisaw/turkicocr-svtrv2-b.git
cd turkicocr
pip install -e .
import cv2
from turkicocr.recognition.predictor import TurkicOCRPredictor

predictor = TurkicOCRPredictor(model_path="epoch_9.pth")
image = cv2.imread("document_line.jpg")
result = predictor.predict(image)
print(result["text"])

Citation

@inproceedings{issayev2026turkicocr,
  title={TurkicOCR-SVTRv2-B: Lightweight Line-Grounded Recognizer for Kazakh and Kyrgyz Optical Character Recognition},
  author={Issayev, Alen and Zhalgas, Aidana},
  booktitle={Analysis of Images, Social Networks and Texts (AIST 2026)},
  series={Lecture Notes in Computer Science (LNCS)},
  publisher={Springer},
  year={2026},
  doi={10.1007/978-3-031-XXXXX-X_XX}
}

@misc{issayev_2026_turkicocr_cyrillic,
  author       = {Issayev, Alen},
  title        = {TurkicOCR-Cyrillic},
  year         = {2026},
  publisher    = {Hugging Face},
  doi          = {10.57967/hf/9255},
  url          = {https://huggingface.co/datasets/alenisaw/turkicocr-cyrillic}
}

License

Apache 2.0. Full code and model checkpoints available at https://github.com/alenisaw/turkicocr.

cyrillic
image-to-text
openocr
svtrv2
text-recognition
turkicocr