13 repos
Pre-trained transformer models and tools for natural language processing across multiple languages (French, English, German, Spanish) with a focus on semantic similarity, sentence representation, and legal domain applications. The cluster centers on PyTorch-based models like legal-xlm variants and multilingual sentence embeddings, designed for cross-lingual understanding, paraphrase detection, and specialized tasks like sentence boundary detection. This is a resource collection for practitioners building multilingual NLP systems and legal tech applications.