24 repos
Models and tools for generating vector embeddings from text across multiple languages, enabling semantic search, similarity matching, and downstream NLP tasks in language-agnostic ways. The cluster centers on pretrained embedding models like Snowflake Arctic Embed, multilingual E5, and XLM-RoBERTa variants, along with frameworks and utilities for deploying and using these embeddings in production systems. Developers working on cross-lingual search, recommendation systems, or multilingual machine learning applications would find both the foundational models and integration tooling here.