Text Embeddings & Semantic Search

17 repos

Libraries and models for generating dense vector embeddings from text, enabling semantic similarity search and retrieval-augmented applications. The cluster centers on sentence-transformers, PyTorch-based embedding models (particularly multilingual variants like BGE for Chinese), and inference optimization tools like text-embeddings-inference and SPLADE. These components support building retrieval systems, vector databases, and text classification pipelines where semantic understanding matters more than keyword matching.

sentence-transformers ·2,450
safetensors ·2,350
endpoints_compatible ·2,034
feature-extraction ·2,034
text-embeddings-inference ·2,034
sentence-similarity ·1,904
gemma3_text ·1,898
eval-results ·1,898
custom_code ·443
model-index ·429